Free for a limited time · local-first

The local alternative to
cloud AI tools

Voice, music & more — free, local-first AI software.

Aispanvok builds free desktop AI tools that run entirely on your own hardware — a free, local alternative to cloud voice services like ElevenLabs and WisprFlow. Start with VoxLoom, our flagship AI voice studio: clone any voice, generate speech in 23 languages, and dictate anywhere.

7 TTS engines23 languages0 API keys100% on-device
macOS Windows Linux
Product categories

Free AI tools, by category

Every category is a family of products that ship independently — download what you need, add more as we launch.

Why local-first

Cloud AI charges you for what your machine can do

Local-first software runs every computation on your own hardware — the same models, without the meter.

Complete privacy

Local-first means every model, voice sample, and transcript stays on your device. Cloud voice services process your most personal asset — your voice — on someone else's servers.

No meter running

Local generation is unlimited generation. No per-character billing, no API keys, no quotas, and no rate limits — cloud services charge you for exactly what you could compute for free.

Free, not freemium

Aispanvok tools are free to download and use for a limited time — not a trial that converts to a subscription. Install it, own it, run it forever on your own hardware.

Learn more about our local-first AI voice tools →

Product family

Free, local-first voice tools

Aispanvok builds voice software that runs on your machine. Each product ships independently — download what you need, add more as we launch.

More voice products are on the way. See the full lineup →

Why VoxLoom

Everything you need to create with voice

One studio for cloning, generation, transcription, and effects.

Zero-Shot Voice Cloning

Clone any voice from a few seconds of audio — no fine-tuning, no cloud. Works offline with Metal, CUDA, and ROCm acceleration.

7 TTS Engines, 23 Languages

Qwen3-TTS, Chatterbox, Kokoro, LuxTTS, and more — switch engines with one click and generate speech from English to Arabic, Japanese, and Hindi.

System-Wide Dictation

Hold a hotkey, talk, release: Whisper-based dictation lands in any text field, with raw and LLM-refined transcripts kept side by side.

AI Agent Voices via MCP

One tool call and Claude Code, Cursor, or Cline speaks in a voice you cloned — each agent bound to its own voice profile.

Stories Editor & Effects

A multi-track timeline for podcasts and narratives, plus pitch shift, reverb, delay, chorus, and compression built in.

Unlimited Local Generation

No character limits, no API keys, no subscriptions. Auto-chunking with crossfade handles books and full chapters.

Explore all VoxLoom features →

Try a voice

Hear the local difference

Pick a preset from VoxLoom — Free AI Voice Studio. Full cloning and unlimited generation run in the desktop app — on your machine, not in the cloud.

"Welcome to VoxLoom — your voice, your machine, no cloud required."

Audio samples play here when available. Clone any voice in seconds inside the app.

Compare

The local alternative

VoxLoom is not a cloud freemium voice changer. It is a free, local-first workstation — cloning, dictation, TTS, and agent voices in one app.

Feature VoxLoomElevenLabsWisprFlowDubbing AI
Pricing Free (limited time) $5–330/mo $144/yr Freemium + sub
Runs locally
Voice cloning Zero-shot, offline Cloud Cloud, real-time
Text-to-speech 7 engines · 23 langs Cloud 500+ presets
System dictation Cloud
AI agent voices (MCP)
Local REST API No keys · no limits Metered cloud SDK only
Privacy 100% on-device Cloud processing Cloud processing Cloud processing
Best for Creators · devs · privacy Studio dubbing Writing flow Gaming · memes
Quick start

Up and running in five minutes

Six small steps from download to your first locally generated voice with VoxLoom — Free AI Voice Studio.

1

Download & install

Grab the installer for macOS, Windows, or Linux. No account required.

~1 min
2

Import or clone a voice

Drop a short audio clip, record 10 seconds, or pick a preset voice.

~30 sec
3

Generate your first line

Type any text, pick an engine, and hit Generate — all on your machine.

~10 sec
4

Set up dictation

Assign a global hotkey. Hold, speak, release — text lands in any app.

~1 min
5

Route to your apps

Optional: point Discord, OBS, or your IDE MCP client at VoxLoom.

~30 sec

You're creating

Clone, dictate, compose, and build — unlimited, offline, on your hardware.

Done
Built for real work

Local power, zero compromise

7 TTS engines One-click switch
23 Languages Multilingual cloning
0 API keys Local REST endpoint
100% On-device Privacy by design
Import Voices

Import a voice in 3 ways

Upload audio files, record with your microphone, or create a voice from a sample text. Every voice you import is yours — stored locally on your machine.

The voices included with Aispanvok were created with the same technology you get to use. Import your own and they are added to your library instantly.

No uploads. No cloud. Just your files, on your machine.

Upload a clip Microphone System Audio
🎵 Drag & drop any audio file — WAV, MP3, FLAC, or WebM
my_voice_sample.mp3 00:32 · 44.1kHz
narration.wav 01:15 · 48kHz
Import voice
Dictation

Dictate anywhere in your system

Press a shortcut to start dictating into any text field — notes, editors, chats. Aispanvok transcribes with Whisper and types for you.

Hold, speak, release — from anywhere on your machine, into any app.

Raw and refined transcripts kept side by side, with the original audio forever.

Agents talk back through the same pill, in any cloned voice.

Hold on macOS, CtrlAlt on Windows
Recording 0:00
Whisper Base 74M Small 244M Medium 769M Large 1.5B Turbo 809M 99 langs

Whisper, sized for every machine

Base, Small, Medium, Large, and Turbo. Pick the size that fits your hardware — 99 languages across every tier, all running locally.

Refined transcripts

A local LLM cleans ums, self-corrections, and punctuation without rephrasing. Optional, toggleable, and never leaves your machine.

Agents speak in voices you own

Any MCP-aware agent gets a voice with one tool call. The pill surfaces when an agent is speaking, so you always see what's coming out of your machine.

Agent Integration

Bring your agent to life

Give your AI agents a voice with the Aispanvok MCP server. One tool call and any MCP-aware agent — Claude Code, Cursor, Cline — speaks in a voice you cloned.

Transparent — every generation is logged locally.

Credible — clone your own voice for your assistant.

Customizable — pick any personality for your agent.

Terminal
$ claude run
✓ Tests passing (42 files)
✓ Build succeeded in 12.4s
→ aispanvok.speak({ profile: "Morgan" })

Per-agent voice

Bind each MCP client to a voice profile. Claude Code in Morgan, Cursor in Scarlett — you know which agent is talking without looking.

Always visible

Every agent-initiated speech surfaces the pill. No silent background TTS — you always see what's coming out of your machine.

Open protocols

MCP ships day one. ACP, A2A, and anything else built on a tool-call primitive slots into the same endpoint.

Personalities

Voices with a personality

Give any voice profile a free-form personality. Then Rewrite your text in their voice, or let them Compose a fresh line of their own — your cloned voice, in full character.

🎭  Compose voices from scratch

✍️  Rewrite existing voices into new characters

✨  Apply voice effects on top

🕵️
Marlowe Voice profile · cloned from a 12s sample
Personality

“1940s noir detective. World-weary, cynical, every situation a metaphor for the city's underbelly. Talks like he's seen one stack trace too many.”

Rewrite Compose
In character · Marlowe

“Build's wrapped, ship's left the dock. Another stack of code makes its way into prod, another row of green checks lining the wall.”

Rewrite

Restate your text in their voice while preserving every idea. Same content, their delivery — for scripts, dubs, and consistent character voice across long-form work.

Compose

No input needed — hit the button and the character improvises a fresh line of their own. Roll again for another take. Useful for game dialogue, narration cues, or character barks.

Built-in REST API

Your local voice API

Every engine you download becomes a REST endpoint on your machine. Build apps, games, and voice tools with full programmatic control — no API keys, no rate limits, no per-character fees.

🛠️  Use it with any language or framework

⚙️  Automate your voice workflows

📴  Fully offline-capable

API Reference http://127.0.0.1:17493
POST/generateGenerate speech
POST/generate/:id/cancelCancel a generation
GET/profilesList voice profiles
POST/profilesCreate a new profile
GET/models/statusModel catalog & state
GET/historyPast generations
$ curl -X POST http://127.0.0.1:17493/generate \
  -H "Content-Type: application/json" \
  -d '{"text": "Welcome to the game, player one.", "profile": "morgan"}' \
  --output line.wav
No API keys No rate limits No per-character fees Works offline Your audio, your machine

Games

Generate NPC dialogue on the fly, localize characters into new languages, or ship expressive voice lines without a studio.

Apps & agents

Give your app or AI agent a voice. Real-time narration, accessibility readouts, voice replies — all running on the user's machine.

Scripts & tools

Batch-generate audiobook chapters, automate podcast intros, or wire Aispanvok into your Stream Deck. It's just a localhost URL.

Supported Models

Your favorite models, out of the box

VoxLoom bundles 7 TTS engines, Whisper, and a local LLM into one app — all interchangeable from the same interface, no separate installs.

🔊 Text-to-Speech
Qwen3-TTS Multilingual cloning, voice instructions
Qwen CustomVoice 50+ preset voices, delivery control
Chatterbox Multilingual Conversational multilingual speech
Chatterbox Turbo Expressive tags — [laugh], [sigh]
LuxTTS Ultra-light, 48kHz, CPU-friendly
HumeAI TADA Emotionally expressive delivery
Kokoro Compact and expressive voices
🎙️ Speech-to-Text
Whisper Base → Large, Turbo — 99 languages
🧠 LLM
Qwen3 1.7B Bundled local LLM for refinement & personas

Install VoxLoom — Free AI Voice Studio, start dictating.

Free for a limited time, runs locally. No account, no API keys, no per-character fees.

View all releases on GitHub

← See everything VoxLoom — Free AI Voice Studio can do

Community

Built for people who own their voice

Feedback from early adopters. Join us on GitHub and help shape what ships next.

"Finally a voice studio that runs locally. I replaced two subscriptions with one free app on my own machine."
E
Early adopter
Developer
"The MCP integration is exactly what I wanted — Cursor speaks in a voice I cloned, and I see every generation."
E
Early adopter
AI builder
"Dictation plus TTS in one app. I dictate blog posts in the morning and generate narration in the afternoon."
E
Early adopter
Creator
"Privacy was non-negotiable. VoxLoom keeps everything on-device — that is the whole reason I switched."
E
Early adopter
Journalist
"The local REST API means I can batch-generate podcast intros without hitting a cloud rate limit."
E
Early adopter
Podcaster
"Seven TTS engines in one UI. I pick the engine that fits my hardware and switch when I need quality."
E
Early adopter
Engineer
FAQ

Frequently asked questions

What is Aispanvok?

Aispanvok builds free, local-first AI tools. Our products — starting with the VoxLoom AI voice studio — run entirely on your own machine: voice cloning, text-to-speech, dictation, and AI agent voices with no cloud, no accounts, and no subscriptions.

Is Aispanvok free?

Yes — our products are free to download and use for a limited time. No accounts, no subscriptions, no per-character fees, and no rate limits. Unlimited local generation on your own hardware.

What is VoxLoom?

VoxLoom is a free, local-first AI voice studio that runs on your machine. It combines zero-shot voice cloning, text-to-speech across 7 engines and 23 languages, system-wide dictation, and AI agent voices via MCP — a local alternative to both ElevenLabs and WisprFlow.

Which TTS engines does VoxLoom support?

Seven engines ship in one app: Qwen3-TTS, Qwen CustomVoice, LuxTTS, Chatterbox Multilingual, Chatterbox Turbo, HumeAI TADA, and Kokoro — plus Whisper for speech-to-text and a bundled local LLM for transcript refinement.

Where is my voice data stored?

Locally on your machine. Voices, transcripts, and generated audio never leave your device — privacy is the architecture, not a setting.

Can AI agents like Claude Code speak with my cloned voice?

Yes. VoxLoom ships with a built-in MCP server — a single voxloom.speak tool call lets any MCP-aware agent (Claude Code, Cursor, Cline) speak in a voice you own, with each agent bound to its own voice profile.

Which platforms are supported?

VoxLoom runs natively on macOS (Apple Silicon and Intel), Windows (64-bit), and Linux. Built with Tauri (Rust) for native performance with Metal, CUDA, and ROCm acceleration.

About

Free for a limited time, local-first

Aispanvok builds free, local-first AI tools. VoxLoom is our flagship studio — cloning, dictation, TTS, and agent voices in one native app. More products are launching soon, each shipping independently on your machine. If our tools save you a subscription bill or just made your day, a coffee helps us keep shipping.

100% local processing No subscriptions No account required Free for a limited time Open on GitHub