all the models — AI benchmark observatory
← Models

Ling 3.1 Flash

InclusionAI · proprietary · ling-3.1-flash

Ling 3.1 Flash (Ling-3.1-flash) is InclusionAI / Ant Group's hybrid-reasoning MoE flash model: ~560B total parameters with ~25B active per token. API-only at launch (weights not yet on Hugging Face). Hosted context is commonly 262K (256K-class); 1M context is the announced design target, not what all gateways serve today. Text→text; hybrid reasoning / tool-using agent positioning. Vendor launch benches (self-reported): GDPval-AA v2.1 1673 Elo, SkillsBench 68.7, AutomationBench 52.5, Terminal-Bench 4.0 40.4, CyberGym 87.9, FrontierSWE 75.16, SWE Atlas Codebase QnA 55.92, Finance Agent v2 57.87, HealthBench Professional 65.35, DRACO 85.49, MultiChallenge 69.78. No stable paid price row in an existing first-party/gateway provider dir in this repo (Vercel AI Gateway / Novita free trial noted but not stored) — available_in_zeroeval false.

GDPval-AA 2.CyberGymDRACOFrontierSWE

Benchmark scores

BenchmarkScore
GDPval-AA 2.11673.00
CyberGym0.88
DRACO0.85
FrontierSWE0.75

Pricing

  • No provider pricing.

AA metrics

Intelligence
—
Speed tok/s
211.39
Blended $/M
—

Arena Elo

  • No Arena snapshot linked.