LLM Skills
~/catalog/audio & voice//SKILL
Audio & voiceGitHub source

OpenAI transcriptions API

/SKILL

Transcribe audio through `/v1/audio/transcriptions`. Set `OPENAI_BASE_URL` for an OpenAI-compatible proxy or local gateway.

openclawopenclaw
386.5k
June 5, 2026
NOASSERTION
// skill content

--- name: openai-whisper-api description: "OpenAI Audio TranscriptionsAPI via curl; gpt-4o-transcribe, mini, diarize, or whisper-1." homepage:https://platform.openai.com/docs/guides/speech-to-text metadata: { "openclaw": { "emoji": "🌐", "requires": { "bins": ["curl", "node"], "env": ["OPENAIAPIKEY "] }, "primaryEnv ": "OPENAIAPIKEY ", "install": [ { "id": "brew", "kind": "brew", "formula": "curl", "bins": ["curl"], "label": "Install curl (brew)", }, ], }, } --- #OpenAI transcriptionsAPI Transcribe audio through/v1/audio/transcriptions . SetOPENAI_BASE_URL to anOpenAI -compatible proxy or local gateway. ## Quick start ``bash {baseDir}/scripts/transcribe.sh /path/to/audio.m4a ` Defaults: - Model: gpt-4o-transcribe - Output: <input>.txt ## Useful flags `bash {baseDir}/scripts/transcribe.sh /path/to/audio.ogg --model gpt-4o-transcribe --out /tmp/transcript.txt {baseDir}/scripts/transcribe.sh /path/to/audio.ogg --model gpt-4o-mini-transcribe {baseDir}/scripts/transcribe.sh /path/to/audio.ogg --model gpt-4o-transcribe-diarize --json {baseDir}/scripts/transcribe.sh /path/to/audio.ogg --model whisper-1 {baseDir}/scripts/transcribe.sh /path/to/audio.m4a --language en {baseDir}/scripts/transcribe.sh /path/to/audio.m4a --prompt "Speaker names: Peter, Daniel" {baseDir}/scripts/transcribe.sh /path/to/audio.m4a --json --out /tmp/transcript.json ` Notes: - Supported upload formats include mp3, mp4, mpeg, mpga, m4a, wav, webm. - 25 MB upload limit on the hosted API. - Use diarize for speaker labels; script sends chunking_strategy =auto and rejects --prompt. ## API key Set OPENAIAPIKEY, or configure it in the active OpenClaw config file ( $OPENCLAWCONFIGPATH, default ~/.openclaw/openclaw.json). Optionally set OPENAIBASEURL: `json5 { skills: { "openai-whisper-api": { apiKey: "OPENAI_KEY_HERE", }, }, } ``

// original public source
openclaw/openclaw
/skills/openai-whisper-api/SKILL.md
License: NOASSERTION. Review the repository before reusing it.
Independent project, not affiliated with Anthropic. This skill remains the property of its original author.
// install this skill
Paste this command in your terminal at the root of your project:
mkdir -p .claude/commands && curl -o ".claude/commands/SKILL.md" "https://raw.githubusercontent.com/openclaw/openclaw/main/skills/openai-whisper-api/SKILL.md"
Then in Claude Code, type /SKILL to activate it.
open_in_newOpen original source
// save
Save available after sign in.
loginSign in to save
// information
Creatoropenclaw
Stars 386.5k
LicenseNOASSERTION
UpdatedJune 5, 2026
Format.md
AccessFree
// similar

Skills Audio & voice

View allarrow_forward