Suno Style Songsmith
Produce full songs from lyrics and style tags using the open-source HeartMuLa model stack. Use when a user wants an AI-generated track, a Suno-style demo, or lo
Secure checkout via Stripe
Works with the AI tools you already use
See it in action
You say
Generate a 30-second acoustic folk song about a traveler returning home. Use a female vocal and 90bpm. Lyrics include [verse] and [chorus] sections.
Your agent does
Generation complete using seed 4029. Track rendered to outputs/folk_traveler.wav. Style: folk, acoustic, female vocal, 90bpm. Model: HeartMuLa-Base. Note: Model license requires attribution for non-commercial use.
What you get
About this skill
The problem
Generating high-quality music demos usually requires expensive cloud subscriptions or complex DAW setups. Cloud-based AI music services often lack the fine-grained control and privacy needed for local development and rapid iteration.
What it does
- Converts raw lyrics and style tags into full-length audio tracks locally.
- Orchestrates the HeartMuLa model stack, including language models and audio codecs.
- Optimizes VRAM usage through lazy loading to support GPUs with as little as 8GB.
- Automates environment patching for RoPE caches and transformer incompatibilities.
- Supports multi-GPU setups by splitting model and codec processing across devices.
Frameworks & tools
Python 3.10, HeartMuLa, NVIDIA CUDA, Transformers, and Datasets. Uses uv for dependency management and nvidia-smi for resource monitoring.
Why this beats prompting it yourself
Manually configuring the HeartMuLa stack is error-prone due to specific dependency versioning and hardware requirements. This skill handles the technical setup, applies necessary code patches for modern library compatibility, and manages the iterative generation workflow for you.
Use cases
- Generate scratch vocals and instrumental demos for songwriting sessions.
- Create unique background tracks for video content or podcasts without cloud fees.
- Prototype musical ideas across different genres like metal, k-pop, or lo-fi.
- Produce localized audio assets in environments without reliable internet access.
Known limitations
Requires an NVIDIA GPU with at least 8GB VRAM; CPU-only generation is not supported. Performance is limited to the underlying HeartMuLa model capabilities and requires local Python environment setup.
How to install
Drop the file into your AI Agent. Works with Claude, Cursor, ChatGPT, and 20+ more.
Reviews
No reviews yet
Be one of the first to try it. Every listed skill passes our trust checks below.
Security scanned
Passed our 8-point scan before listing
Fresh listing
Recently published to Agensi
30-day refund
Not a fit? Get your money back
Trust & safety
Security scanned
Verified clean today
- 30-day refund guarantee
- One-time purchase, yours forever
- Secure checkout via Stripe