Local Voice (FluidAudio TTS/STT)
What this skill does
Local text-to-speech (TTS) and speech-to-text (STT) using FluidAudio on Apple Silicon. Sub-second voice synthesis and transcription running entirely on-device via the Apple Neural Engine. Use when setting up local voice capabilities, voice assistant integration, or replacing cloud TTS/STT services.
Local Voice (FluidAudio TTS/STT) is part of the Personal Assistant category — personal-assistant skills for everyday life, finance, and planning. You can install it on its own or alongside other personal assistant skills from the OpenClaw catalog.
How to install Local Voice (FluidAudio TTS/STT)
The easiest path is via the OpenClaw Easy desktop app — one click, no terminal required:
- Download OpenClaw Easy for macOS or Windows (free, one-click installer, ~30 seconds).
- Open the in-app Skills panel.
- Search for
local-voiceand click Install. - The skill activates automatically when an incoming message matches its description.
Install from the command line
If you already run the OpenClaw CLI, add Local Voice (FluidAudio TTS/STT) with a single command:
openclaw skills add local-voice
This pulls local-voice from ClawHub and installs it into ~/.openclaw/skills/local-voice/. Restart the OpenClaw gateway afterwards so the new skill is discovered.
How to use Local Voice (FluidAudio TTS/STT)
Once installed, Local Voice (FluidAudio TTS/STT) activates on its own: when an incoming message on WhatsApp, Telegram, Slack, Discord, Feishu or Line matches the skill's description, your OpenClaw agent loads it and runs the workflow. You can also trigger it explicitly by describing the task in chat. No extra configuration is required after install.
Putting audio in front of an audience
Local Voice (FluidAudio TTS/STT) works with audio. Getting from there to something postable is a separate job: ViralMint — an open-source video pipeline — pairs audio with footage or generated visuals, burns synced captions, and exports a finished short.
It runs as an MCP server, so an OpenClaw agent can drive it from the same chat you already use: ask for a short, and the render comes back finished. See how to connect a video pipeline to your agent.
Manual install (advanced)
If you prefer manual installation:
- Click the Download .zip button above to grab
local-voice-1.0.1.zipdirectly from our S3 mirror. - Unzip into
~/.openclaw/skills/local-voice/(create the directory if it does not exist). - Restart OpenClaw Easy (or the OpenClaw CLI gateway) so the new skill is discovered.
Frequently asked questions
How do I install Local Voice (FluidAudio TTS/STT)?
Install Local Voice (FluidAudio TTS/STT) in the OpenClaw Easy desktop app by opening the Skills panel, searching for local-voice, and clicking Install. From a terminal you can run: openclaw skills add local-voice. Either way the skill is placed in ~/.openclaw/skills/local-voice/.
Is Local Voice (FluidAudio TTS/STT) free?
Yes. Local Voice (FluidAudio TTS/STT) is free and open-source, distributed under the Apache-2.0 license through ClawHub. No account or payment is required to download or run it.
What does Local Voice (FluidAudio TTS/STT) do?
Local text-to-speech (TTS) and speech-to-text (STT) using FluidAudio on Apple Silicon. Sub-second voice synthesis and transcription running entirely on-device via the Apple Neural Engine. Use when setting up local voice capabilities, voice assistant integration, or replacing cloud TTS/STT services.
Related: more personal assistant skills
If Local Voice (FluidAudio TTS/STT) looks useful, you may also want to check out other personal assistant skills in the OpenClaw catalog:
Browse the full OpenClaw skill catalog
This page covers just one skill. The OpenClaw skill hub has 10,000+ more — search, sort by downloads or stars, and install any of them in one click. There is also a curated awesome-openclaw-skills list grouped by use case.
Get OpenClaw Easy — Free
Install Local Voice (FluidAudio TTS/STT) and 10,000+ other OpenClaw skills in one click. Free, open-source, runs locally on macOS & Windows.
Free, open-source · Apache-2.0 · Works with Claude, ChatGPT, Gemini, or local Ollama models