Perceptron Mk1 is Perceptron's vision-language model for image and video understanding, OCR, document parsing, object detection, counting, spatial localization, and embodied visual reasoning.
Added May 12, 2026
Context Window
32.8K
Max Output
8.2K
Input Price (Auto)
$0.15/1M
Output Price (Auto)
$1.50/1M
Cache Read (Auto)
$0.075/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
No benchmark data is available yet for this model.
Providers
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare Perceptron Mk1 with similar models from the same provider or model family.
Ornith 1.5 35B
ornith-ai/ornith-1.5-35b-a3bOrnith 1.5 35B A3B is an open-weight mixture-of-experts model for agentic coding, tool use, image understanding, and long-context work. This variant disables thinking for faster direct responses.
Ornith 1.5 35B Thinking
ornith-ai/ornith-1.5-35b-a3b:thinkingOrnith 1.5 35B A3B is an open-weight mixture-of-experts model for agentic coding, reasoning, tool use, image understanding, and long-context work. This variant enables thinking by default.
Qwen3.5 0.8B
qwen3.5-0.8bQwen3.5 0.8B is a lightweight open-weight multimodal model from Alibaba for fast reasoning, visual understanding, tool use, and JSON output.
Qwen3.5 2B
qwen3.5-2bQwen3.5 2B is a small open-weight multimodal model from Alibaba for efficient reasoning, coding, visual understanding, tool use, and JSON output.
Qwen3.5 4B
qwen3.5-4bQwen3.5 4B is a compact open-weight multimodal model from Alibaba for reasoning, coding, visual understanding, tool use, and structured output.
Qwen3.8 27B
qwen3.8-27bQwen3.8 27B is an open-weight multimodal model from Alibaba for coding, visual understanding, tool use, and structured output. This variant keeps thinking disabled for faster direct responses.