specs at a glance
| Leaderboard status | Verifying — no confirmed SWE-bench Verified score |
|---|---|
| SWE-bench Verified | — |
| SWE-bench Pro | 59.5% |
| Terminal-Bench | 70.8% |
| Input price / 1M | $0.30 |
| Output price / 1M | — |
| Context window | 1M |
| Open weights | Yes |
| Access | Open weights (MIT) · API |
| Maker | Meituan |
how good is LongCat-2.0 at coding?
LongCat-2.0 is on our board but deliberately unranked. We rank by SWE-bench Verified, Meituan has not published that number, and no independent evaluation has run it, so there is nothing here we could stand behind. What the maker does publish is 59.5% on SWE-bench Pro and 70.8% on Terminal-Bench, measured on its own scaffold and not comparable like-for-like with the ranked rows. It moves into the ranking the moment a confirmed score exists. On the harder SWE-bench Pro it scores 59.5%. Terminal-Bench (agentic terminal work): 70.8%.
Score provenance: Vendor-reported (Meituan, Jun 30 2026): SWE-bench Pro 59.5%, Terminal-Bench 70.8%. No SWE-bench Verified score or independent eval published yet, so unranked pending confirmation. vals.ai has not evaluated it and does not cover Meituan at all, and llm-stats does not list the model.
Re-checked 2026-09-09: vals.ai's SWE-bench Verified board is unchanged since the September 1 archival and still shows no evaluation for this model; all previously-tracked ranked scores were also re-verified this run and hold within normal variance (largest drift under 0.4pp).
what does LongCat-2.0 cost?
$0.30 per 1M input tokens — as listed by the maker. Coding workloads are output-heavy, so weight the output rate when budgeting. Run your own volume through the AI API cost calculator for a monthly estimate.
where can you use it?
Available via Open weights (MIT) · API. Because it ships open weights, you can also self-host it on your own hardware or any inference provider — with the version pinned so the model can't change under you.
Full storyMeituan's LongCat-2.0 Is a 1.6T Coder on Chinese Chips
Ranked on our AI Coding Leaderboard — scores confirmed against primary sources only, updated 2026-09-03.