CATALOG

Models

Find the right model by capability, context and price. Compare options or try one in Playground.

Output / endpointExact metadata
24 of 599 modelsReference token prices are per million tokens.
VisionToolsReasoning

DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of [DeepSeek V4 Flash 0731](https://openrouter.ai/deepseek/deepseek-v4-flash-0731) from DeepSeek, adding image understanding while matching the base model on text capabilities including agents,...

by deepseekAug 21, 20261.05M context$0.11/M input$0.33/M outputText, ImageText
ToolsReasoning

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....

by deepseekJul 31, 20261.31M context$0.06/M input$0.12/M outputTextText
ToolsReasoning

DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding,...

by deepseekApr 24, 20261.05M context$1.6/M input$3.2/M outputTextText
ToolsReasoning

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...

by deepseekApr 24, 20261.05M context$0.0809/M input$0.1618/M outputTextText
ToolsReasoning

DeepSeek-V3.2 is a large language model designed to harmonize high computational efficiency with strong reasoning and agentic tool-use performance. It introduces DeepSeek Sparse Attention (DSA), a fine-grained sparse attention mechanism...

by deepseekDec 1, 2025163.84K context$0.269/M input$0.4/M outputTextText
ToolsReasoning

DeepSeek-V3.2-Exp is an experimental large language model released by DeepSeek as an intermediate step between V3.1 and future architectures. It introduces DeepSeek Sparse Attention (DSA), a fine-grained sparse attention mechanism...

by deepseekSep 29, 2025163.84K context$0.27/M input$0.41/M outputTextText
ToolsReasoning

DeepSeek-V3.1 Terminus is an update to [DeepSeek V3.1](/deepseek/deepseek-chat-v3.1) that maintains the model's original capabilities while addressing issues reported by users, including language consistency and agent capabilities, further optimizing the model's...

by deepseekSep 22, 2025163.84K context$0.27/M input$1/M outputTextText
ToolsReasoning

DeepSeek-V3.1 is a large hybrid reasoning model (671B parameters, 37B active) that supports both thinking and non-thinking modes via prompt templates. It extends the DeepSeek-V3 base with a two-phase long-context...

by deepseekAug 21, 2025163.84K context$0.25/M input$0.95/M outputTextText
ToolsReasoning

May 28th update to the [original DeepSeek R1](/deepseek/deepseek-r1) Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active...

by deepseekMay 28, 2025163.84K context$0.5/M input$2.15/M outputTextText
Tools

DeepSeek V3, a 685B-parameter, mixture-of-experts model, is the latest iteration of the flagship chat model family from the DeepSeek team. It succeeds the [DeepSeek V3](/deepseek/deepseek-chat-v3) model and performs really well...

by deepseekMar 24, 2025163.84K context$0.25/M input$1/M outputTextText
Reasoning

DeepSeek R1 Distill Llama 70B is a distilled large language model based on [Llama-3.3-70B-Instruct](/meta-llama/llama-3.3-70b-instruct), using outputs from [DeepSeek R1](/deepseek/deepseek-r1). The model combines advanced distillation techniques to achieve high performance across...

by deepseekJan 23, 20258.19K context$0.8/M input$0.8/M outputTextText
ToolsReasoning

DeepSeek R1 is here: Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active in an inference pass....

by deepseekJan 20, 202564K context$0.7/M input$2.5/M outputTextText
Tools

DeepSeek-V3 is the latest model from the DeepSeek team, building upon the instruction following and coding abilities of the previous versions. Pre-trained on nearly 15 trillion tokens, the reported evaluations...

by deepseekDec 26, 2024163.84K context$0.2574/M input$1.03/M outputTextText
ToolsReasoning

Aion-3.0 Mini is a multi-model roleplaying and storytelling system from AionLabs, built on the DeepSeek family of models. It uses a collaborative generation process in which multiple specialized models each...

by aion-labsJul 7, 2026131.07K context$0.7/M input$1.4/M outputTextText
ToolsReasoning

Aion-2.0 is a variant of DeepSeek V3.2 optimized for immersive roleplaying and storytelling. It is particularly strong at introducing tension, crises, and conflict into stories, making narratives feel more engaging....

by aion-labsFeb 23, 2026131.07K context$0.8/M input$1.6/M outputTextText
VisionReasoning

Note: Sonar Pro pricing includes Perplexity search pricing. See [details here](https://docs.perplexity.ai/guides/pricing#detailed-pricing-breakdown-for-sonar-reasoning-pro-and-sonar-pro) Sonar Reasoning Pro is a premier reasoning model powered by DeepSeek R1 with Chain of Thought (CoT). Designed for...

by perplexityMar 7, 2025128K context$2/M input$8/M outputText, ImageText