← Models
LongCat-Flash-Chat
Meituan · open weight · longcat-flash-chat
LongCat-Flash-Chat is Meituan's first open-source foundation model, a 560B parameter Mixture-of-Experts (MoE) model that dynamically activates 18.6B-31.3B parameters (~27B average) based on contextual demands. It features Zero-Computation Experts for efficient routing and supports 128K context. Optimized for conversational and agentic tasks, it shows competitive performance across reasoning, coding, instruction following, and domain benchmarks with particular strengths in tool use and complex multi-step interactions. Achieves over 100 tokens per second on H800 GPUs.
Benchmark scores
| Benchmark | Score |
|---|---|
| MATH-500 | 0.96 |
| MMLU | 0.90 |
| IFEval | 0.90 |
| HumanEval | 0.88 |
| MMLU-Pro | 0.83 |
| GPQA | 0.73 |
| AIME 2025 | 0.61 |
| SWE-Bench Verified | 0.60 |
| LiveCodeBench | 0.48 |
Pricing
- No provider pricing.
AA metrics
No Artificial Analysis link yet.
Arena Elo
- No Arena snapshot linked.
