Ling 2.6 Flash API Benchmarks, Pricing & Provider Data
Compare Ling 2.6 Flash with another model
Choose a model to open its comparison page.
Ling-2.6-flash is an instant (instruct) model from inclusionAI with 104B total parameters and 7.4B active parameters, designed for real-world agents that require fast responses, strong execution, and ...
Category Performance
Observed-capability estimates with uncertainty reported separately from independent benchmark families.
- Coverage
- 5 / 8
- Methodology
- V3.0
#1Instruction following48.5Provisional1/4 Measured dimensions
#2Agents44.3Estimated2/4 Measured dimensions
#3Coding42.7Provisional1/4 Measured dimensions
#4Reasoning39Estimated2/4 Measured dimensions
#5Knowledge34.1Provisional1/4 Measured dimensions
Specifications
Input and output token limits for this model, plus how it ranks on long-context understanding.
Features
Technical Details
- Input
- Output
- Released
- Apr 2026
- Tokenizer
- Other
- Architecture
- text->text
- Moderated
- No
- Expiration date
- 2026-08-24
- Supported parameters
- frequency_penaltylogprobsmax_tokenspresence_penaltyrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_logprobstop_p
Rankings
Excels at
Falls behind in
Detailed scores
Updated: Aug 25, 2026Overall
undefined metric} other undefined metrics}}
Overall score42.0#96 / 100
Overall
undefined metric} other undefined metrics}}
Agents
V3.0undefined metric} other undefined metrics}} · Estimated
Score44.380% interval32.4–56.12/4 Measured dimensionsΤ²-bench results86.0#42 / 82GDPval-AA2.2#60 / 61
Agents
V3.0undefined metric} other undefined metrics}} · Estimated
Coding
V3.0undefined metric} other undefined metrics}} · Provisional
Score42.780% interval26.8–58.71/4 Measured dimensionsSciCode27.1%#169 / 204SciCode27.0#12 / 12
Coding
V3.0undefined metric} other undefined metrics}} · Provisional
Reasoning
V3.0undefined metric} other undefined metrics}} · Estimated
Score3980% interval28.2–49.82/4 Measured dimensionsGPQA59.3%#164 / 211HLE6.3%#145 / 208
Reasoning
V3.0undefined metric} other undefined metrics}} · Estimated
Knowledge
V3.0undefined metric} other undefined metrics}} · Provisional
Score34.180% interval20.1–48.11/4 Measured dimensionsKnowledge score46.5#64 / 69Artificial Analysis Intelligence Index14.1#98 / 111
Knowledge
V3.0undefined metric} other undefined metrics}} · Provisional
Math
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Math
V3.0undefined metric} other undefined metrics}} · No data
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Multilingual
V3.0undefined metric} other undefined metrics}} · No data
Multimodal
V3.0undefined metric} other undefined metrics}} · No data
No data80% interval30.8–69.20/4 Measured dimensions
Multimodal
V3.0undefined metric} other undefined metrics}} · No data
Instruction following
V3.0undefined metric} other undefined metrics}} · Provisional
Score48.580% interval32.5–64.51/4 Measured dimensionsInstruction Following score39.0#25 / 25IFBench57.0#14 / 14
Instruction following
V3.0undefined metric} other undefined metrics}} · Provisional
Alternatives & Similar Models
GLM-4.7 Flash
glm-4-7-flash
Zhipu AI GLM-4.7 Flash is a speed-focused GLM variant for real-time chat, function calling, and bilingual enterprise copilots with competitive token economics.
DeepSeek V4 Flash
deepseek-v4-flash
DeepSeek V4 Flash is a fast, cost-efficient language model in the DeepSeek V4 family, optimized for low-latency chat, coding assistance, and high-throughput API workloads while retaining strong reasoning quality.
Kimi K2.5
kimi-k2-5
Moonshot Kimi K2.5 is an open-weight multimodal agent model with native vision and text input, strong coding performance, and a 256K context window.
GLM-5
glm-5
Zhipu GLM-5 is Zhipu flagship GLM series model with enhanced reasoning, agent capabilities, and strong performance on Chinese enterprise and coding scenarios.
DeepSeek V3.2
deepseek-v3-2
DeepSeek V3.2 is an upgraded V3-series MoE model with stronger reasoning, coding, and math performance, widely available through OpenAI-compatible API relays.
MiniMax M2.7
minimax-m2-7
MiniMax M2.7 is a high-tier M2-series model tuned for complex reasoning, long-context dialogue, and production-grade API workloads.
