OpenAI Whisper API (curl)

SkillMedia

Transcribe audio via OpenAI Audio Transcriptions API (Whisper).

Available today. Use it from your connected AI after setup.

Connect ahel once, and every AI you use reads what you have installed.

Then ask your AI: use the OpenAI Whisper API (curl) skill

What this skill tells your AI

The instructions your AI receives, as published by understudy-ai/understudy in skills/openai-whisper-api/SKILL.md and read by ahel’s review.

Transcribe an audio file via OpenAI’s /v1/audio/transcriptions endpoint.

Quick start

{baseDir}/scripts/transcribe.sh /path/to/audio.m4a

Defaults:

  • Model: whisper-1
  • Output: <input>.txt

Useful flags

{baseDir}/scripts/transcribe.sh /path/to/audio.ogg --model whisper-1 --out /tmp/transcript.txt
{baseDir}/scripts/transcribe.sh /path/to/audio.m4a --language en
{baseDir}/scripts/transcribe.sh /path/to/audio.m4a --prompt "Speaker names: Peter, Daniel"
{baseDir}/scripts/transcribe.sh /path/to/audio.m4a --json --out /tmp/transcript.json

API key

Set OPENAI_API_KEY, or configure it in ~/.understudy/config.json5:

{
  skills: {
    "openai-whisper-api": {
      apiKey: "OPENAI_KEY_HERE",
    },
  },
}

Signals

GitHub stars
456
Forks
34
Last commit
Jun 2026

ahel review

  • S4info
    community integration — published by understudy-ai, not openai

Automated review, not a security audit. Ruleset v1+k2.

Advanced
Catalog kind
skill
Gateway key
openai-whisper-api-understudy-ai
Source
github.com/understudy-ai/understudy