CATALOG

Models

Find the right model by capability, context and price. Compare options or try one in Playground.

Output / endpointExact metadata
19 of 599 modelsReference token prices are per million tokens.
Audio input

Muse Voice Transcribe 1.0 is a synchronous speech-to-text model from Meta. It is suited for push-to-talk, endpointing, and speaker-aware transcription, with keyword biasing for domain terms and language biasing through...

by metaSep 11, 2026N/A contextSpecialized media pricing · open model detailsAudioTranscription
Image outputVision

Muse Image is an agentic image generation model from Meta that generates and edits images from text and reference images. Unlike single-pass image models, it reasons before it renders, breaking...

by metaAug 26, 202665.54K contextSpecialized media pricing · open model detailsText, ImageImage
VisionToolsReasoning

Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon...

by metaAug 9, 2026131.07K context$0.35/M input$1.5/M outputText, ImageText
VisionTools

Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...

by meta-llamaApr 5, 20251.05M context$0.1875/M input$0.6525/M outputText, ImageText
VisionTools

Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It supports native multimodal input...

by meta-llamaApr 5, 20251.31M context$0.1/M input$0.3/M outputText, ImageText
Tools

The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...

by meta-llamaDec 6, 2024131.07K context$0.1/M input$0.32/M outputTextText
Text

Llama 3.2 1B is a 1-billion-parameter language model focused on efficiently performing natural language tasks, such as summarization, dialogue, and multilingual text analysis. Its smaller size allows it to operate...

by meta-llamaSep 25, 202460K context$0.027/M input$0.201/M outputTextText
Text

Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...

by meta-llamaSep 25, 2024131.07K context$0.05/M input$0.33/M outputTextText
Tools

Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 70B instruct-tuned version is optimized for high quality dialogue usecases. It has demonstrated strong...

by meta-llamaJul 23, 2024131.07K context$0.4/M input$0.4/M outputTextText
Tools

Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 8B instruct-tuned version is fast and efficient. It has demonstrated strong performance compared to...

by meta-llamaJul 23, 2024131.07K context$0.05/M input$0.08/M outputTextText
Reasoning

Hermes 4 is a large-scale reasoning model built on Meta-Llama-3.1-405B and released by Nous Research. It introduces a hybrid reasoning mode, where the model can choose to deliberate internally with...

by nousresearchAug 26, 2025131.07K context$1/M input$3/M outputTextText
Reasoning

DeepSeek R1 Distill Llama 70B is a distilled large language model based on [Llama-3.3-70B-Instruct](/meta-llama/llama-3.3-70b-instruct), using outputs from [DeepSeek R1](/deepseek/deepseek-r1). The model combines advanced distillation techniques to achieve high performance across...

by deepseekJan 23, 20258.19K context$0.8/M input$0.8/M outputTextText