CATALOG

Models

Find the right model by capability, context and price. Compare options or try one in Chat.

Output / endpointExact metadata
1 of 647 modelsReference token prices are per million tokens.
Audio input

Universal-3.5 Pro is AssemblyAI's speech-to-text model served through its Sync API, returning a complete transcript with word-level timestamps in a single synchronous response for audio clips up to 120 seconds....

by assemblyaiSep 22, 2026N/A contextSpecialized media pricing · open model detailsAudio → Transcription