LLM Skills
~/catalog/Media creation/Audio & voice

Skills, plugins and agents for Audio & voice

23 results

Sherpa ONNX text to speech (TTS)

386.5k

1. Download the runtime for your OS (extracts into `$OPENCLAW_STATE_DIR/tools/sherpa-onnx-tts/runtime`, default `~/.openclaw/tools/sherpa-onnx-tts/runtime`)

Claude CodeCodexAntigravity
Audio & voice
openclaw

OpenAI transcriptions API

386.5k

Transcribe audio through `/v1/audio/transcriptions`. Set `OPENAI_BASE_URL` for an OpenAI-compatible proxy or local gateway.

Claude CodeCodexAntigravity
Audio & voice
openclaw

Whisper (CLI)

386.5k

Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞

Claude CodeCodexAntigravity
Audio & voice
openclaw

Text to speech (ElevenLabs TTS)

386.5k

Use `sag` for ElevenLabs TTS with local playback.

Claude CodeCodexAntigravity
Audio & voice
openclaw

songsee

386.5k

Generate spectrograms + feature panels from audio.

Claude CodeCodexAntigravity
Audio & voice
openclaw

Albums de musique

87.8k

Full-lifecycle AI music album production : concept, lyric drafting, track sequencing, and export. Useful for indie album experiments and brand soundtr

Claude CodeCodexAntigravity
Audio & voice
nexu-io

Audio lecture on Venice

87.8k

Text-to-speech models, voices, formats, and streaming via Venice.ai. Useful for narration, voiceover, and conversational agent voices.

Claude CodeCodexAntigravity
Audio & voice
nexu-io

Musique de venise

87.8k

Music generation queueing, retrieval, and completion endpoints via Venice.ai. Suited for jingles, background loops, and prototype scoring.

Claude CodeCodexAntigravity
Audio & voice
nexu-io

Click anywhere on MuseScore

47.7k

Command-line interface for music notation: transposition, export to PDF/audio/MIDI, part extraction, and management of instruments

Claude CodeCodexAntigravity
Audio & voice
HKUDS

Click anywhere in Audacity

47.7k

Command-line interface for Audacity - A stateful command-line interface for audio editing, following the same patterns as the GIMP and Ble...

Claude CodeCodexAntigravity
Audio & voice
HKUDS

Azure ai transcription py

45.0k

SDK Azure AI Transcription for Python. Use it for real-time and batch speech-to-text transcription, with timestamping and diarization.

Claude CodeCodexAntigravity
Audio & voice
sickn33

Azure ai voicelive ts

45.0k

SDK Azure AI Voice Live for JavaScript/TypeScript. Build real-time voice AI applications using bidirectional WebSocket communication

Claude CodeCodexAntigravity
Audio & voice
sickn33

Azure for Voicelive Java

45.0k

SDK Azure AI VoiceLive for Java. Real-time, two-way voice conversations with AI assistants via WebSocket.

Claude CodeCodexAntigravity
Audio & voice
sickn33

Azure voice live .net

45.0k

SDK Azure AI Voice Live for .NET. Build real-time voice AI applications using bidirectional WebSocket communication.

Claude CodeCodexAntigravity
Audio & voice
sickn33

Azure, I'm going live on py.

45.0k

Build real-time voice AI applications using bidirectional WebSocket communication.

Claude CodeCodexAntigravity
Audio & voice
sickn33

Transcripteur audio

45.0k

Turn your audio recordings into professional-quality Markdown documentation with smart summaries, thanks to the integration of a

Claude CodeCodexAntigravity
Audio & voice
sickn33

Audio transcription (ElevenLabs Scribe)

1.3k

Transcribe audio or video files using the ElevenLabs Speech-to-Text API (Scribe v2). Accepts a file path and optional parameters, reads the API key from the project's .env file, and returns a formatte

Claude CodeCodexAntigravity
Audio & voice
qdhenry

Openai Whisper

325

Local speech-to-text with the Whisper CLI (no API key).

Claude CodeCodexAntigravity
Audio & voice
ClawHub community

Video Subtitles

17

Generate SRT subtitles from video/audio with translation support. Transcribes Hebrew (ivrit.ai) and English (whisper), translates between languages, burns subtitles into video. Use for creating captions, transcripts, or

Claude CodeCodexAntigravity
Audio & voice
ClawHub community

Local Whisper

12

Local speech-to-text using OpenAI Whisper. Runs fully offline after model download. High quality transcription with multiple model sizes.

Claude CodeCodexAntigravity
Audio & voice
ClawHub community

Songsee

10

Generate spectrograms and feature-panel visualizations from audio with the songsee CLI.

Claude CodeCodexAntigravity
Audio & voice
ClawHub community

Voice Wake Say

5

Speak responses aloud on macOS using the built-in `say` command when user input indicates Voice Wake/voice recognition (for example, messages starting with "User talked via voice recognition on <device>").

Claude CodeCodexAntigravity
Audio & voice
ClawHub community

Suno Music

OpenClaw plugin for creating Suno songs, lyrics, personas, credit checks, and local Suno task history.

OpenClaw
Audio & voice
Plugin@talkl3ss