Transcribe 1

Transcribe 1 is a speech-to-text model from Fish Audio. It is suited for audio transcription with automatic language detection and can return timestamped word-level segments when alignment details are requested.

Context
0
Max output
Unknown
Input
$100
Output
$0.00
DECISION SUMMARY

Recommended use cases

Strengths in this dataset

  • 0-token context window
  • audio input
  • 0 supported API parameters listed

Limits and caveats

  • Provider behavior and pricing can change; verify the linked sources before production use.
CAPABILITIES

Capability

Model-native facts

Model
Reasoning
Unknown
Open weights
Unknown

Provider endpoint facts

Provider endpoint
Tool calling
Not supported
Structured output
Not supported
Streaming
Unknown
Prompt cache
Unknown
Batch
Unknown
Fine-tuning
Unknown
PROVIDER PRICING

Transcribe 1 Provider pricing

Provider endpoint: fish-audio/transcribe-1

Input
$100
per 1M tokens
Output
$0.00
per 1M tokens
Cached input
Unknown
per 1M tokens
Image output
Unknown
per 1M tokens
SOURCE RECORDS

Sources and verification

OpenRouter Models API

Fields: identity, description, modalities, context window, maximum output, pricing, supported parameters

MODEL FAQ

Frequently asked questions

Answers are generated from the same sourced model and provider facts shown above.

Transcribe 1 is a speech-to-text model from Fish Audio. It is suited for audio transcription with automatic language detection and can return timestamped word-level segments when alignment details are requested.

Model specifications and prices may vary by provider and change over time. AIToolly displays sources and verification dates so users can confirm critical details before production use.