Zero-Shot Voice Cloning
Clone any voice from a few seconds of audio — no fine-tuning, no cloud. Works offline with Metal, CUDA, and ROCm acceleration.
Voice, music & more — free, local-first AI software.
Aispanvok builds free desktop AI tools that run entirely on your own hardware — a free, local alternative to cloud voice services like ElevenLabs and WisprFlow. Start with VoxLoom, our flagship AI voice studio: clone any voice, generate speech in 23 languages, and dictate anywhere.
Every category is a family of products that ship independently — download what you need, add more as we launch.
Local-first software runs every computation on your own hardware — the same models, without the meter.
Local-first means every model, voice sample, and transcript stays on your device. Cloud voice services process your most personal asset — your voice — on someone else's servers.
Local generation is unlimited generation. No per-character billing, no API keys, no quotas, and no rate limits — cloud services charge you for exactly what you could compute for free.
Aispanvok tools are free to download and use for a limited time — not a trial that converts to a subscription. Install it, own it, run it forever on your own hardware.
Aispanvok builds voice software that runs on your machine. Each product ships independently — download what you need, add more as we launch.
More voice products are on the way. See the full lineup →
One studio for cloning, generation, transcription, and effects.
Clone any voice from a few seconds of audio — no fine-tuning, no cloud. Works offline with Metal, CUDA, and ROCm acceleration.
Qwen3-TTS, Chatterbox, Kokoro, LuxTTS, and more — switch engines with one click and generate speech from English to Arabic, Japanese, and Hindi.
Hold a hotkey, talk, release: Whisper-based dictation lands in any text field, with raw and LLM-refined transcripts kept side by side.
One tool call and Claude Code, Cursor, or Cline speaks in a voice you cloned — each agent bound to its own voice profile.
A multi-track timeline for podcasts and narratives, plus pitch shift, reverb, delay, chorus, and compression built in.
No character limits, no API keys, no subscriptions. Auto-chunking with crossfade handles books and full chapters.
Pick a preset from VoxLoom — Free AI Voice Studio. Full cloning and unlimited generation run in the desktop app — on your machine, not in the cloud.
"Welcome to VoxLoom — your voice, your machine, no cloud required."
Audio samples play here when available. Clone any voice in seconds inside the app.
VoxLoom is not a cloud freemium voice changer. It is a free, local-first workstation — cloning, dictation, TTS, and agent voices in one app.
| Feature | VoxLoom | ElevenLabs | WisprFlow | Dubbing AI |
|---|---|---|---|---|
| Pricing | Free (limited time) | $5–330/mo | $144/yr | Freemium + sub |
| Runs locally | ✓ | — | — | — |
| Voice cloning | Zero-shot, offline | Cloud | — | Cloud, real-time |
| Text-to-speech | 7 engines · 23 langs | Cloud | — | 500+ presets |
| System dictation | ✓ | — | Cloud | — |
| AI agent voices (MCP) | ✓ | — | — | — |
| Local REST API | No keys · no limits | Metered cloud | — | SDK only |
| Privacy | 100% on-device | Cloud processing | Cloud processing | Cloud processing |
| Best for | Creators · devs · privacy | Studio dubbing | Writing flow | Gaming · memes |
Six small steps from download to your first locally generated voice with VoxLoom — Free AI Voice Studio.
Grab the installer for macOS, Windows, or Linux. No account required.
Drop a short audio clip, record 10 seconds, or pick a preset voice.
Type any text, pick an engine, and hit Generate — all on your machine.
Assign a global hotkey. Hold, speak, release — text lands in any app.
Optional: point Discord, OBS, or your IDE MCP client at VoxLoom.
Clone, dictate, compose, and build — unlimited, offline, on your hardware.
Upload audio files, record with your microphone, or create a voice from a sample text. Every voice you import is yours — stored locally on your machine.
The voices included with Aispanvok were created with the same technology you get to use. Import your own and they are added to your library instantly.
No uploads. No cloud. Just your files, on your machine.
Click to record from your microphone. Maximum duration: 30 seconds.
Press a shortcut to start dictating into any text field — notes, editors, chats. Aispanvok transcribes with Whisper and types for you.
Hold, speak, release — from anywhere on your machine, into any app.
Raw and refined transcripts kept side by side, with the original audio forever.
Agents talk back through the same pill, in any cloned voice.
Base, Small, Medium, Large, and Turbo. Pick the size that fits your hardware — 99 languages across every tier, all running locally.
A local LLM cleans ums, self-corrections, and punctuation without rephrasing. Optional, toggleable, and never leaves your machine.
Any MCP-aware agent gets a voice with one tool call. The pill surfaces when an agent is speaking, so you always see what's coming out of your machine.
Give your AI agents a voice with the Aispanvok MCP server. One tool call and any MCP-aware agent — Claude Code, Cursor, Cline — speaks in a voice you cloned.
Transparent — every generation is logged locally.
Credible — clone your own voice for your assistant.
Customizable — pick any personality for your agent.
$ claude run ✓ Tests passing (42 files) ✓ Build succeeded in 12.4s → aispanvok.speak({ profile: "Morgan" })
Bind each MCP client to a voice profile. Claude Code in Morgan, Cursor in Scarlett — you know which agent is talking without looking.
Every agent-initiated speech surfaces the pill. No silent background TTS — you always see what's coming out of your machine.
MCP ships day one. ACP, A2A, and anything else built on a tool-call primitive slots into the same endpoint.
Give any voice profile a free-form personality. Then Rewrite your text in their voice, or let them Compose a fresh line of their own — your cloned voice, in full character.
🎭 Compose voices from scratch
✍️ Rewrite existing voices into new characters
✨ Apply voice effects on top
“1940s noir detective. World-weary, cynical, every situation a metaphor for the city's underbelly. Talks like he's seen one stack trace too many.”
“Build's wrapped, ship's left the dock. Another stack of code makes its way into prod, another row of green checks lining the wall.”
“Some days the city hums clean and the tests all pass. Not tonight. Tonight a flaky build walks into my CI pipeline, and nothing green comes out the other side.”
Restate your text in their voice while preserving every idea. Same content, their delivery — for scripts, dubs, and consistent character voice across long-form work.
No input needed — hit the button and the character improvises a fresh line of their own. Roll again for another take. Useful for game dialogue, narration cues, or character barks.
Every engine you download becomes a REST endpoint on your machine. Build apps, games, and voice tools with full programmatic control — no API keys, no rate limits, no per-character fees.
🛠️ Use it with any language or framework
⚙️ Automate your voice workflows
📴 Fully offline-capable
http://127.0.0.1:17493 /generateGenerate speech/generate/:id/cancelCancel a generation/profilesList voice profiles/profilesCreate a new profile/models/statusModel catalog & state/historyPast generations$ curl -X POST http://127.0.0.1:17493/generate \
-H "Content-Type: application/json" \
-d '{"text": "Welcome to the game, player one.", "profile": "morgan"}' \
--output line.wav Generate NPC dialogue on the fly, localize characters into new languages, or ship expressive voice lines without a studio.
Give your app or AI agent a voice. Real-time narration, accessibility readouts, voice replies — all running on the user's machine.
Batch-generate audiobook chapters, automate podcast intros, or wire Aispanvok into your Stream Deck. It's just a localhost URL.
VoxLoom bundles 7 TTS engines, Whisper, and a local LLM into one app — all interchangeable from the same interface, no separate installs.
From podcast production to AI agent speech — local, free, and private. Not a gaming voice changer; a studio for people who own their audio.
Clone voices, generate narration, and edit multi-track stories — all locally. VoxLoom is the free voice studio for YouTubers, podcasters, and audiobook makers.
Learn more →Give Claude Code, Cursor, and Cline a voice you own. VoxLoom ships MCP and a local REST API — no API keys, no rate limits.
Learn more →Your voice is personal data. VoxLoom processes everything on-device — no accounts, no cloud uploads, no third-party training.
Learn more →Dictate into any app with a global hotkey. Whisper transcription plus optional LLM cleanup — all local, all private.
Learn more →Multi-track Stories editor, built-in effects, and unlimited local generation. Produce podcast episodes without a monthly voice bill.
Learn more →Generate speech in 23 languages with seven TTS engines. Localize characters and narration without re-recording or cloud API fees.
Learn more →Free for a limited time, runs locally. No account, no API keys, no per-character fees.
Feedback from early adopters. Join us on GitHub and help shape what ships next.
"Finally a voice studio that runs locally. I replaced two subscriptions with one free app on my own machine."
"The MCP integration is exactly what I wanted — Cursor speaks in a voice I cloned, and I see every generation."
"Dictation plus TTS in one app. I dictate blog posts in the morning and generate narration in the afternoon."
"Privacy was non-negotiable. VoxLoom keeps everything on-device — that is the whole reason I switched."
"The local REST API means I can batch-generate podcast intros without hitting a cloud rate limit."
"Seven TTS engines in one UI. I pick the engine that fits my hardware and switch when I need quality."
Aispanvok builds free, local-first AI tools. Our products — starting with the VoxLoom AI voice studio — run entirely on your own machine: voice cloning, text-to-speech, dictation, and AI agent voices with no cloud, no accounts, and no subscriptions.
Yes — our products are free to download and use for a limited time. No accounts, no subscriptions, no per-character fees, and no rate limits. Unlimited local generation on your own hardware.
VoxLoom is a free, local-first AI voice studio that runs on your machine. It combines zero-shot voice cloning, text-to-speech across 7 engines and 23 languages, system-wide dictation, and AI agent voices via MCP — a local alternative to both ElevenLabs and WisprFlow.
Seven engines ship in one app: Qwen3-TTS, Qwen CustomVoice, LuxTTS, Chatterbox Multilingual, Chatterbox Turbo, HumeAI TADA, and Kokoro — plus Whisper for speech-to-text and a bundled local LLM for transcript refinement.
Locally on your machine. Voices, transcripts, and generated audio never leave your device — privacy is the architecture, not a setting.
Yes. VoxLoom ships with a built-in MCP server — a single voxloom.speak tool call lets any MCP-aware agent (Claude Code, Cursor, Cline) speak in a voice you own, with each agent bound to its own voice profile.
VoxLoom runs natively on macOS (Apple Silicon and Intel), Windows (64-bit), and Linux. Built with Tauri (Rust) for native performance with Metal, CUDA, and ROCm acceleration.
Aispanvok builds free, local-first AI tools. VoxLoom is our flagship studio — cloning, dictation, TTS, and agent voices in one native app. More products are launching soon, each shipping independently on your machine. If our tools save you a subscription bill or just made your day, a coffee helps us keep shipping.