CATALOG

Models

Find the right model by capability, context and price. Compare options or try one in Chat.

Output / endpointExact metadata
9 of 647 modelsReference token prices are per million tokens.
VisionTools

Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...

by meta-llamaApr 5, 20251.05M context$0.1875/M input$0.6525/M outputText, Image → Text
VisionTools

Llama 4 Scout 17B Instruct (16E) is a mixture-of-experts (MoE) language model developed by Meta, activating 17 billion parameters out of a total of 109B. It supports native multimodal input...

by meta-llamaApr 5, 20251.31M context$0.1/M input$0.3/M outputText, Image → Text
Tools

The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...

by meta-llamaDec 6, 2024131.07K context$0.1/M input$0.32/M outputText → Text
Text

Llama 3.2 1B is a 1-billion-parameter language model focused on efficiently performing natural language tasks, such as summarization, dialogue, and multilingual text analysis. Its smaller size allows it to operate...

by meta-llamaSep 25, 202460K context$0.027/M input$0.201/M outputText → Text
Text

Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...

by meta-llamaSep 25, 2024131.07K context$0.05/M input$0.33/M outputText → Text
Tools

Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 70B instruct-tuned version is optimized for high quality dialogue usecases. It has demonstrated strong...

by meta-llamaJul 23, 2024131.07K context$0.4/M input$0.4/M outputText → Text
Tools

Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 8B instruct-tuned version is fast and efficient. It has demonstrated strong performance compared to...

by meta-llamaJul 23, 2024131.07K context$0.05/M input$0.08/M outputText → Text
Reasoning

Hermes 4 is a large-scale reasoning model built on Meta-Llama-3.1-405B and released by Nous Research. It introduces a hybrid reasoning mode, where the model can choose to deliberate internally with...

by nousresearchAug 26, 2025131.07K context$1/M input$3/M outputText → Text