Quick Answer
Last verified:
High confidence

Whisper (OpenAI) costs $0.00 to $0.01 per minute as of July 2026, with 3 plans available including a free tier. Plans: GPT-4o Mini Transcribe (free), Whisper / GPT-4o Transcribe (free), and Enterprise (via ChatGPT Enterprise / API) (free). Enterprise pricing is available on request. Pricing depends on your chosen tier, contract length, and negotiated discounts.

Use the interactive pricing calculator to estimate your exact cost based on team size and requirements.

  • Free tier: Yes

Whisper (OpenAI) offers 3 pricing tiers: GPT-4o Mini Transcribe, Whisper / GPT-4o Transcribe, Enterprise (via ChatGPT Enterprise / API). The Whisper / GPT-4o Transcribe plan is developers needing accurate, affordable transcription with the simplest possible integration and no add-on fees.

Whisper (OpenAI) lists $0.003-$0.006/minute, but hidden costs like implementation and support add to the total as of July 2026. Key hidden costs: actual costs may exceed the $0.006/min headline rate: developer reports indicate real-world costs averaging $0.010/min due to billing rounding, retries on failed requests, and processing overhead -- across 648 hours one developer reported spending $397 vs an estimated $233 (70% over budget), no built-in speaker diarization on legacy whisper model: while gpt-4o transcribe now includes diarization at $0.006/min, the legacy whisper model requires a separate post-processing step using gpt-4o or a third-party service, adding $0.002-$0.01/min in additional costs, 25 mb file size limit forces chunking overhead: audio files over 25 mb must be split into smaller segments before upload, requiring engineering effort for chunk management, overlap handling, and transcript reassembly -- budget $500-$1,500 for initial chunking pipeline development. Verified from 5 sources by CostBench.

Hidden Costs Breakdown

1

Actual costs may exceed the $0.006/min headline rate: Developer reports indicate real-world costs averaging $0.010/min due to billing rounding, retries on failed requests, and processing overhead -- across 648 hours one developer reported spending $397 vs an estimated $233 (70% over budget)

2

No built-in speaker diarization on legacy Whisper model: While GPT-4o Transcribe now includes diarization at $0.006/min, the legacy Whisper model requires a separate post-processing step using GPT-4o or a third-party service, adding $0.002-$0.01/min in additional costs

3

25 MB file size limit forces chunking overhead: Audio files over 25 MB must be split into smaller segments before upload, requiring engineering effort for chunk management, overlap handling, and transcript reassembly -- budget $500-$1,500 for initial chunking pipeline development

4

No HIPAA BAA available: OpenAI does not offer a Business Associate Agreement, making the Whisper API unusable for Protected Health Information (PHI) -- organizations with healthcare data must self-host Whisper on HIPAA-compliant infrastructure at $1,400+/month

5

Self-hosting break-even at 500+ hours/month: At $0.006/min, 500 hours costs $180/month via API vs ~$276/month for self-hosted GPU infrastructure -- above 500 hours self-hosting becomes cheaper but requires DevOps expertise and GPU management overhead

6

Rate limits throttle high-volume processing: Default tier allows only 50 requests per minute -- processing 10,000+ files requires careful queue management, retry logic, and potentially upgrading to higher API tiers which require spending history with OpenAI

Check current Whisper (OpenAI) pricing

Prices and terms change; verify against the live pricing page.

See Whisper (OpenAI) Pricing