CATALOG

Models

Find the right model by capability, context and price. Compare options or try one in Chat.

Output / endpointExact metadata
2 of 647 modelsReference token prices are per million tokens.
VisionVideo inputReasoning

Perceptron Mk1 (Mark One) is Perceptron's highest-quality vision-language model for video and embodied reasoning.** It accepts image and video inputs paired with natural language queries, and produces detailed visual understanding...

by perceptronMay 12, 202632.77K context$0.15/M input$1.5/M outputText, Image, Video → Text