OpenAI transcriptions API
/SKILLTranscribe audio through `/v1/audio/transcriptions`. Set `OPENAI_BASE_URL` for an OpenAI-compatible proxy or local gateway.
--- name: openai-whisper-api description: "OpenAI Audio TranscriptionsAPI via curl; gpt-4o-transcribe, mini, diarize, or whisper-1." homepage:https://platform.openai.com/docs/guides/speech-to-text metadata: { "openclaw": { "emoji": "🌐", "requires": { "bins": ["curl", "node"], "env": ["OPENAIAPIKEY "] }, "primaryEnv ": "OPENAIAPIKEY ", "install": [ { "id": "brew", "kind": "brew", "formula": "curl", "bins": ["curl"], "label": "Install curl (brew)", }, ], }, } --- #OpenAI transcriptionsAPI Transcribe audio through/v1/audio/transcriptions . SetOPENAI_BASE_URL to anOpenAI -compatible proxy or local gateway. ## Quick start ``bash {baseDir}/scripts/transcribe.sh /path/to/audio.m4a ` Defaults: - Model: gpt-4o-transcribe - Output: <input>.txt ## Useful flags `bash {baseDir}/scripts/transcribe.sh /path/to/audio.ogg --model gpt-4o-transcribe --out /tmp/transcript.txt {baseDir}/scripts/transcribe.sh /path/to/audio.ogg --model gpt-4o-mini-transcribe {baseDir}/scripts/transcribe.sh /path/to/audio.ogg --model gpt-4o-transcribe-diarize --json {baseDir}/scripts/transcribe.sh /path/to/audio.ogg --model whisper-1 {baseDir}/scripts/transcribe.sh /path/to/audio.m4a --language en {baseDir}/scripts/transcribe.sh /path/to/audio.m4a --prompt "Speaker names: Peter, Daniel" {baseDir}/scripts/transcribe.sh /path/to/audio.m4a --json --out /tmp/transcript.json ` Notes: - Supported upload formats include mp3, mp4, mpeg, mpga, m4a, wav, webm. - 25 MB upload limit on the hosted API. - Use diarize for speaker labels; script sends chunking_strategy =auto and rejects --prompt. ## API key Set OPENAIAPIKEY, or configure it in the active OpenClaw config file ( $OPENCLAWCONFIGPATH, default ~/.openclaw/openclaw.json). Optionally set OPENAIBASEURL: `json5 { skills: { "openai-whisper-api": { apiKey: "OPENAI_KEY_HERE", }, }, } ``