OpenAI/GPTActive

GPT-4o Transcribe

GPT-4o Transcribe is OpenAI's high-quality speech-to-text model built on GPT-4o audio capabilities. It's priced per token (input and output), making it suitable for workflows that benefit from token-level billing transparency.

Context
128K
Max output
Unknown
Input
$2.5
Output
$10
DECISION SUMMARY

Recommended use cases

Strengths in this dataset

  • 128,000-token context window
  • audio input
  • 12 supported API parameters listed

Limits and caveats

  • Provider behavior and pricing can change; verify the linked sources before production use.
CAPABILITIES

Capability

Model-native facts

Model
Reasoning
Unknown
Open weights
Unknown

Provider endpoint facts

Provider endpoint
Tool calling
Not supported
Structured output
Supported
Streaming
Unknown
Prompt cache
Unknown
Batch
Unknown
Fine-tuning
Unknown
PROVIDER PRICING

GPT-4o Transcribe Provider pricing

Provider endpoint: openai/gpt-4o-transcribe

Input
$2.5
per 1M tokens
Output
$10
per 1M tokens
Cached input
Unknown
per 1M tokens
Image output
Unknown
per 1M tokens
SOURCE RECORDS

Sources and verification

OpenRouter Models API

Fields: identity, description, modalities, context window, maximum output, pricing, supported parameters

MODEL FAQ

Frequently asked questions

Answers are generated from the same sourced model and provider facts shown above.

GPT-4o Transcribe is OpenAI's high-quality speech-to-text model built on GPT-4o audio capabilities. It's priced per token (input and output), making it suitable for workflows that benefit from token-level billing transparency.

Model specifications and prices may vary by provider and change over time. AIToolly displays sources and verification dates so users can confirm critical details before production use.