跳轉到主要內容

LongCat-Flash-Chat

Meituan開源權重MIT · 商用許可

描述

LongCat-Flash-Chat is Meituan's first open-source foundation model, a 560B parameter Mixture-of-Experts (MoE) model that dynamically activates 18.6B-31.3B parameters (~27B average) based on contextual demands. It features Zero-Computation Experts for efficient routing and supports 128K context. Optimized for conversational and agentic tasks, it shows competitive performance across reasoning, coding, instruction following, and domain benchmarks with particular strengths in tool use and complex multi-step interactions. Achieves over 100 tokens per second on H800 GPUs.

發布日期
2025-08-29
參數規模
560.0B
上下文長度
支援模態
text

能力雷達圖

70
general
60
coding
80
reasoning
60
science估算
70
agents
0
multimodal

缺少專門科學評測時,Science 由 LLM Stats 科學得分或推理能力估算。

排行榜排名

領域#排名分數來源
智慧體能力模型榜30
55.0
LS

基準測試分數 (LLM Stats)

(LLM Stats (zeroeval))

Agents

Terminal-Bench39.5%自報

Biology

GPQANYU + Cohere + Anthropic (2023)73.2%自報

Code

HumanEvalOpenAI (2021)88.4%自報
SWE-Bench Verified60.4%自報
LiveCodeBench48.0%自報

Communication

Tau2 Telecom73.7%自報
Tau2 Retail71.3%自報
Tau2 Airline58.0%自報

Finance

MMLU89.7%自報
MMLU-Pro82.7%自報

General

IFEvalGoogle Research (2023)89.6%自報
CMMLU84.3%自報

Math

MATH-50096.4%自報
DROP79.1%自報
AIME 202561.3%自報

Reasoning

ZebraLogic89.3%自報

AA 評測指數

(Artificial Analysis)

暫無 AA 評測資料

LLM Stats 分類評分

(LLM Stats (zeroeval))
Legal
90
Structured Output
90
Instruction Following
90
Language
90
Finance
90
Healthcare
90
Math
80
Physics
70
Reasoning
70
General
70
Biology
70
Chemistry
70
Communication
70
Tool Calling
70
Frontend Development
60
Code
60
Agents
40

定價

暫無定價資料

速度

暫無速度資料

供應商價格排行

暫無提供商資料

外部連結