Home › Skills › speech-recognition
Data & APIs

speech-recognition skill for OpenClaw

⬇ 6K downloads ★ 3 stars Version 1.0.1 Rank #1000 of 10,000+

What this skill does

通用语音识别 Skill。支持多种音频格式(ogg/mp3/wav/m4a),使用硅基流动 SenseVoice API 进行语音转文字。当用户发送语音消息、音频文件,或需要转录音频时触发。

The speech-recognition skill is part of the Data & APIs category — data and API skills that fetch, query, and work with external data sources. You can install it on its own or alongside other data & apis skills from the OpenClaw catalog.

The speech-recognition skill is ranked #1000 by downloads in the OpenClaw skill catalog (6K total downloads, 3 stars). It belongs to the Data & APIs category alongside 945 other top-10000 skills.

How to install the speech-recognition skill

The easiest path is via the OpenClaw Easy desktop app — one click, no terminal required:

  1. Download OpenClaw Easy for macOS or Windows (free, one-click installer, ~30 seconds).
  2. Open the in-app Skills panel.
  3. Search for speech-recognition and click Install.
  4. The skill activates automatically when an incoming message matches its description.

Install from the command line

If you already run the OpenClaw CLI, add the speech-recognition skill with a single command:

openclaw skills add speech-recognition

This pulls speech-recognition from ClawHub and installs it into ~/.openclaw/skills/speech-recognition/. Restart the OpenClaw gateway afterwards so the new skill is discovered.

How to use the speech-recognition skill

Once installed, the speech-recognition skill activates on its own: when an incoming message on WhatsApp, Telegram, Slack, Discord, Feishu or Line matches the skill's description, your OpenClaw agent loads it and runs the workflow. You can also trigger it explicitly by describing the task in chat. No extra configuration is required after install.

Putting audio in front of an audience

speech-recognition works with audio. Getting from there to something postable is a separate job: ViralMint — an open-source video pipeline — pairs audio with footage or generated visuals, burns synced captions, and exports a finished short.

It runs as an MCP server, so an OpenClaw agent can drive it from the same chat you already use: ask for a short, and the render comes back finished. See how to connect a video pipeline to your agent.

Manual install (advanced)

If you prefer manual installation:

  1. Click the Download skill .zip button above to grab speech-recognition-1.0.1.zip directly from our S3 mirror.
  2. Unzip into ~/.openclaw/skills/speech-recognition/ (create the directory if it does not exist).
  3. Restart OpenClaw Easy (or the OpenClaw CLI gateway) so the new skill is discovered.

Frequently asked questions

How do I install speech-recognition?

Install speech-recognition in the OpenClaw Easy desktop app by opening the Skills panel, searching for speech-recognition, and clicking Install. From a terminal you can run: openclaw skills add speech-recognition. Either way the skill is placed in ~/.openclaw/skills/speech-recognition/.

Is the speech-recognition skill free?

Yes. It is free to download and run through ClawHub, with no account or payment required. Each skill is published by its own author under its own licence — see its ClawHub page for the licence and source.

What does speech-recognition do?

通用语音识别 Skill。支持多种音频格式(ogg/mp3/wav/m4a),使用硅基流动 SenseVoice API 进行语音转文字。当用户发送语音消息、音频文件,或需要转录音频时触发。

Related: more data & apis skills

If the speech-recognition skill looks useful, you may also want to check out other data & apis skills in the OpenClaw catalog:

Skills in this catalog are community-contributed integrations published on ClawHub and distributed under their own open-source licences. Product and company names, and any third-party service a skill connects to, are trademarks of their respective owners; a listing here does not imply affiliation with, sponsorship by, or endorsement from them. This page does not distribute any third-party application. Rights holders can reach us at hello@openclaw-easy.com.

Browse the full OpenClaw skill catalog

This page covers just one skill. The OpenClaw skill hub has 10,000+ more — search, sort by downloads or stars, and install any of them in one click. There is also a curated awesome-openclaw-skills list grouped by use case.

Get OpenClaw Easy — Free

Install speech-recognition and 10,000+ other OpenClaw skills in one click. Free, open-source, runs locally on macOS & Windows.

Free, open-source · Apache-2.0 · Works with Claude, ChatGPT, Gemini, or local Ollama models