NV

    Nvidia Voice Clone

    2

    Clone any voice or generate professional text-to-speech using NVIDIA's zero-shot Magpie NIM technology.

    Secure checkout via Stripe

    0 installsSecurity scanned

    Works with the AI tools you already use

    CClaude CodeCCursorCCodex CLIGGitHub CopilotGGemini CLI+17 more

    See it in action

    You say

    Clone the voice from ./samples/founder_voice.wav and say: Welcome to the new platform interface.

    Your agent does

    Cloning voice and generating speech... Text: Welcome to the new platform interface. Voice sample: ./samples/founder_voice.wav VOICE: /home/user/.claude-voice-clone/voice_1715234892.wav (452 KB) Done! Play the output file to hear the cloned voice.

    What you get

    Create personalized voiceovers for demos using a short audio referenceGenerate multilingual narration for documentation and tutorialsPrototype voice-enabled applications without local GPU resourcesAutomate the production of audio assets for developer presentations

    About this skill

    What it does

    This skill enables high-fidelity voice cloning and text-to-speech (TTS) generation directly through your AI agent. By leveraging the NVIDIA Magpie TTS NIM, it can replicate any voice from a brief 10-30 second audio sample or generate professional narration using high-quality preset voices.

    Why use this skill

    Integrating professional-grade voice synthesis into a developer workflow usually requires complex SDKs or expensive subscriptions. This skill streamlines the process by using NVIDIA's zero-shot cloning technology, allowing your agent to produce localized audio assets, narration, or personalized voice feedback without leaving the terminal. It is significantly faster than manual audio processing and utilizes a powerful cloud infrastructure for low-latency synthesis.

    Supported tools & features

    • NVIDIA Magpie TTS Zeroshot: Clone voices from WAV/MP3 files with minimal data.
    • NVIDIA Magpie Multilingual: Support for diverse accents and languages including Spanish, French, and German.
    • Local File Management: Automatically manages audio output and storage in a dedicated local directory.
    • Bypass Setup: Works with a simple API key, removing the need for local GPU-heavy TTS models.

    Output format

    The skill produces high-fidelity WAV audio files stored locally, providing clear, natural-sounding speech that is ready for use in applications, videos, or testing.

    How to install

    Drop the file into your AI Agent. Works with Claude, Cursor, ChatGPT, and 20+ more.

    Reviews

    No reviews yet

    Be one of the first to try it. Every listed skill passes our trust checks below.

    Security scanned

    Passed our 8-point scan before listing

    Fresh listing

    Recently published to Agensi

    30-day refund

    Not a fit? Get your money back

    Trust & safety

    Security scanned

    Verified clean 4 months ago

    Listed4 months ago

    Frequently Asked Questions