CLSSAI 独立产品正在逐项复刻与接入真实能力。查看进度

CATALOG

Models

按能力、模态、上下文、价格与数据策略寻找模型。默认展示官方标价,实际路由费用会在调用前透明显示。

24 of 456 modelsOutput and specialty tabs may overlap when a model supports multiple paths.
FFish Audio: S2.1 Pro Free (free)fish-audio/s2.1-pro-free:freeOpenRouter preview · CLSSAI route pending
Free

S2.1 Pro Free is the no-cost variant of Fish Audio S2.1 Pro, intended for testing, prototyping, and low-volume applications. It provides the same synthesis capabilities without production latency or availability...

by fish-audioJul 29, 2026N/A contextSpecialized media pricing · open model detailsTextSpeech
VVoyageAI by MongoDB: rerank-2.5-litevoyageai/rerank-2.5-liteOpenRouter preview · CLSSAI route pending
Free

rerank-2.5-lite is a reranker optimized for both latency and quality, delivering a 7.16% improvement in retrieval accuracy over Cohere Rerank v3.5 across 93 datasets. It also outperformed Cohere Rerank v3.5...

by voyageaiJul 27, 202632K context$0/M input$0/M outputTextRerank
VVoyageAI by MongoDB: rerank-2.5voyageai/rerank-2.5OpenRouter preview · CLSSAI route pending
Free

rerank-2.5 is a cutting-edge reranker optimized for quality, delivering a 7.94% improvement in retrieval accuracy over Cohere Rerank v3.5 across 93 datasets. It also outperformed Cohere Rerank v3.5 by 12.70%...

by voyageaiJul 27, 202632K context$0/M input$0/M outputTextRerank
FreeToolsReasoning

*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...

by inclusionaiJul 23, 2026262.14K context$0/M input$0/M outputTextText
FreeToolsReasoning

Laguna S 2.1 is the latest coding agent model from [Poolside](<https://poolside.ai/>). Laguna S 2.1 is a 118B total parameter model with 8B active parameters, scoring 70.2% on Terminal-Bench 2.1 and...

by poolsideJul 21, 2026262.14K context$0/M input$0/M outputTextText
Free

NVIDIA Nemotron 3 Embed 1B is an open text embedding model from NVIDIA, optimized for high-throughput, low-latency retrieval. It is suited for enterprise search, RAG, code retrieval, and agentic retrieval...

by nvidiaJul 16, 202632.77K context$0/M input$0/M outputTextEmbeddings
FreeToolsReasoning

Laguna XS 2.1 is the latest coding agent model in the 33B-A3B category from [Poolside](https://poolside.ai/) and a step forward from their Laguna XS.2 model (released in April 2026). It combines...

by poolsideJul 2, 2026262.14K context$0/M input$0/M outputTextText
FreeToolsReasoning

North Mini Code is Cohere's first agentic coding model and the debut of its North family. A sparse mixture-of-experts model with 30B total parameters and 3B active, it is optimized...

by cohereJun 17, 2026256K context$0/M input$0/M outputTextText
FreeVisionReasoning

NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...

by nvidiaJun 4, 2026128K context$0/M input$0/M outputText, ImageText
FreeToolsReasoning

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...

by nvidiaJun 4, 20261M context$0/M input$0/M outputTextText
FreeVisionAudio inputVideo inputTools

NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and...

by nvidiaApr 28, 2026256K context$0/M input$0/M outputText, Audio, Image, VideoText
CCohere: Rerank 4 Procohere/rerank-4-proOpenRouter preview · CLSSAI route pending
Free

Cohere's AI search foundation model for enhancing the relevance of information surfaced within search and RAG systems. Features a 32K context window, multilingual support across 100+ languages, no data pre-processing...

by cohereApr 6, 202632.77K context$0/M input$0/M outputTextRerank
CCohere: Rerank 4 Fastcohere/rerank-4-fastOpenRouter preview · CLSSAI route pending
Free

Cohere's AI search foundation model for enhancing the relevance of information surfaced within search and RAG systems. Features a 32K context window, multilingual support across 100+ languages, no data pre-processing...

by cohereApr 6, 202632.77K context$0/M input$0/M outputTextRerank
CCohere: Rerank v3.5cohere/rerank-v3.5OpenRouter preview · CLSSAI route pending
Free

Rerank v3.5 is designed to reorder search results for improved relevance. It supports multi-aspect and semi-structured data reranking over 100+ languages. Ideal for refining results from semantic or keyword search...

by cohereApr 5, 20264.1K context$0/M input$0/M outputTextRerank
FreeVisionVideo inputToolsReasoning

Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...

by googleApr 3, 2026262.14K context$0/M input$0/M outputImage, Text, VideoText
FreeVisionVideo inputToolsReasoning

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

by googleApr 2, 2026262.14K context$0/M input$0/M outputImage, Text, VideoText
FreeToolsReasoning

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...

by nvidiaMar 11, 2026262.14K context$0/M input$0/M outputTextText
OFree Models Routeropenrouter/freeCLSSAI live · temporary metadata
FreeVisionToolsReasoning

The simplest way to get free inference. openrouter/free is a router that selects free models at random from the models available on OpenRouter. The router smartly filters for models that...

by openrouterFeb 1, 2026200K context$0/M input$0/M outputText, ImageText
FreeToolsReasoning

NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully...

by nvidiaDec 14, 2025256K context$0/M input$0/M outputTextText
FreeVisionVideo inputToolsReasoning

NVIDIA Nemotron Nano 2 VL is a 12-billion-parameter open multimodal reasoning model designed for video understanding and document intelligence. It introduces a hybrid Transformer-Mamba architecture, combining transformer-level accuracy with Mamba’s...

by nvidiaOct 28, 2025128K context$0/M input$0/M outputImage, Text, VideoText
FreeToolsReasoning

NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and...

by nvidiaSep 5, 2025128K context$0/M input$0/M outputTextText
FreeToolsReasoning

gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...

by openaiAug 5, 2025131.07K context$0/M input$0/M outputTextText