K2-Think is a 32B open-weights general reasoning model with strong competitive math performance. Benchmarks: AIME 2024 90.83, AIME 2025 81.24, GPQA-Diamond 71.08, LiveCodeBench v5 63.97.
Added Jul 26, 2025
Context Window
128.0K
Max Output
32.8K
Input Price (Auto)
$0.17/1M
Output Price (Auto)
$0.68/1M
Cache Read (Auto)
$0.085/1M
Benchmarks
Benchmarks
Performance metrics and benchmarks
No benchmark data is available yet for this model.
Providers
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare K2-Think with similar models from the same provider or model family.
Ornith 1.5 35B
ornith-ai/ornith-1.5-35b-a3bOrnith 1.5 35B A3B is an open-weight mixture-of-experts model for agentic coding, tool use, image understanding, and long-context work. This variant disables thinking for faster direct responses.
Ornith 1.5 35B Thinking
ornith-ai/ornith-1.5-35b-a3b:thinkingOrnith 1.5 35B A3B is an open-weight mixture-of-experts model for agentic coding, reasoning, tool use, image understanding, and long-context work. This variant enables thinking by default.
Qwen3.5 0.8B
qwen3.5-0.8bQwen3.5 0.8B is a lightweight open-weight multimodal model from Alibaba for fast reasoning, visual understanding, tool use, and JSON output.
Qwen3.5 2B
qwen3.5-2bQwen3.5 2B is a small open-weight multimodal model from Alibaba for efficient reasoning, coding, visual understanding, tool use, and JSON output.
Qwen3.5 4B
qwen3.5-4bQwen3.5 4B is a compact open-weight multimodal model from Alibaba for reasoning, coding, visual understanding, tool use, and structured output.
Qwen3.8 27B
qwen3.8-27bQwen3.8 27B is an open-weight multimodal model from Alibaba for coding, visual understanding, tool use, and structured output. This variant keeps thinking disabled for faster direct responses.