Ling-2.6-1T is an inclusionAI instruction model optimized for large-scale agentic and coding workloads with long-context support and structured output capabilities.
Added Apr 23, 2026
Context Window
262.1K
Max Output
32.8K
Input Price (Auto)
$0.30/1M
Output Price (Auto)
$2.50/1M
Cache Read (Auto)
$0.060/1M
Capabilities
Benchmarks
Benchmarks
Performance metrics and benchmarks
Sourced from Artificial Analysis.
Intelligence Index
26.1
Reasoning
GPQA Diamond
Graduate-level scientific reasoning
75.2%
Better than 65% of models compared
HLE
Humanity's Last Exam
8.2%
Better than 59% of models compared
IFBench
Instruction-following benchmark
56.9%
Better than 69% of models compared
T²-Bench Telecom
Conversational AI agents in dual-control scenarios
89.8%
Better than 84% of models compared
AA-LCR
Long context reasoning evaluation
34.7%
Better than 49% of models compared
CritPt
Research-level physics reasoning
0.3%
Coding
SciCode
Python programming for scientific computing
37.0%
Better than 65% of models compared
Terminal-Bench Hard
Agentic coding and terminal use
31.1%
Better than 75% of models compared
Knowledge
AA-Omniscience Accuracy
Proportion of correctly answered questions
21.9%
AA-Omniscience Hallucination Rate
Rate of incorrect answers among non-correct responses
92.8%
Last updated Jun 28, 2026
Artificial AnalysisProviders
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare Ling 2.6 1T with similar models from the same provider or model family.
Ling 3.0 Flash
inclusionai/ling-3.0-flashLing-3.0-flash is a 124B-parameter Mixture-of-Experts model with approximately 5.1B parameters active per token. It prioritizes token efficiency and production-scale agentic inference, helping coding and tool-using agents complete more work within constrained latency and serving budgets.
Ling 3.0 Flash Thinking
inclusionai/ling-3.0-flash:thinkingLing-3.0-flash Thinking enables visible reasoning on inclusionAI's token-efficient 124B-parameter Mixture-of-Experts model for harder coding, tool use, planning, and production-scale agent workflows.
Ling 2.6 Flash
inclusionai/ling-2.6-flashLing-2.6-flash is an instant (instruct) model from inclusionAI with 104B total parameters and 7.4B active parameters, designed for real-world agents that require fast responses, strong execution, and high token efficiency. It delivers performance comparable to state-of-the-art models at a similar scale while significantly reducing token usage across coding, document processing, and lightweight agent workflows.
Ring 2.6 1T
inclusionai/ring-2.6-1tRing-2.6-1T is an inclusionAI thinking model for real-world agent workflows, coding agents, tool use, and long-horizon task execution.
Ornith 1.5 35B
ornith-ai/ornith-1.5-35b-a3bOrnith 1.5 35B A3B is an open-weight mixture-of-experts model for agentic coding, tool use, image understanding, and long-context work. This variant disables thinking for faster direct responses.
Ornith 1.5 35B Thinking
ornith-ai/ornith-1.5-35b-a3b:thinkingOrnith 1.5 35B A3B is an open-weight mixture-of-experts model for agentic coding, reasoning, tool use, image understanding, and long-context work. This variant enables thinking by default.