Gifts

Culture

Reviews

Local Spots

Whisper API API

★★★★★ 4.4 (9,500 reviews)
AI Powered API Access Cloud Based
Pay-per-use $

Overview

OpenAI's speech recognition API based on the Whisper model, offering accurate transcription and translation across 57 languages.

API Details

TypeREST
AuthenticationAPI Key
Rate Limits50 RPM

Whisper API Pricing Plans

Pay-as-you-go

usage-based

  • ✓ $0.006/min transcription
  • ✓ $0.006/min translation

Own Whisper API?

Claim this profile to add your website link, custom description, and get a verified badge.

$50/month or $350/year (save $250)

Whisper API Alternatives & Competitors

Looking for something different? Here are the top alternatives to Whisper API:

Google Cloud Vision

★★★★☆ 4.3 (5,200)
Ai Api Computer Vision Api

Google's computer vision API for image analysis including label detection, OCR, face detection, and explicit content detection.

Baichuan API

★★★★☆ 3.7 (320)
Ai Api Llm Api

Chinese AI company offering large language models optimized for Chinese language understanding and generation tasks.

Together AI

★★★★☆ 4.3 (3,200)
Ai Api Ai Infrastructure

Fast inference and fine-tuning platform for open-source models with competitive pricing and OpenAI-compatible endpoints.

Qdrant API

★★★★☆ 4.3 (2,400)
Ai Api Embedding Api

High-performance open-source vector search engine with filtering, payload indexing, and distributed deployment support.

AI21 Labs API

★★★★☆ 3.9 (900)
Ai Api Llm Api

AI21's Jamba model family API providing efficient text generation with hybrid Mamba-Transformer architecture.

AssemblyAI

★★★★★ 4.4 (3,200)
Ai Api Speech Api

Accurate speech-to-text API with built-in audio intelligence features like summarization, sentiment analysis, and topic detection.

DeepSeek API

★★★★☆ 4.3 (8,500)
Ai Api Llm Api

Chinese AI lab offering high-performance reasoning and coding models at very competitive prices through an OpenAI-compatible API.

Azure Speech

★★★★☆ 4.2 (3,600)
Ai Api Speech Api

Microsoft's comprehensive speech service offering text-to-speech, speech-to-text, translation, and speaker recognition.