Atlas / Models / Meituan / Meituan: LongCat Flash Chat

Meituan: LongCat Flash Chat✓ Catalog verified

meituan/longcat-flash-chat

LongCat-Flash-Chat is a large-scale Mixture-of-Experts (MoE) model with 560B total parameters, of which 18.6B–31.3B (≈27B on average) are dynamically activated per input. It introduces a shortcut-connected MoE design to reduce communication overhead and achieve high throughput while maintaining training stability through advanced scaling strategies such as hyperparameter transfer, deterministic computation, and multi-stage optimization. This release, LongCat-Flash-Chat, is a non-thinking foundation model optimized for conversational and agentic tasks. It supports long context windows up to 128K tokens and shows competitive performance across reasoning, coding, instruction following, and domain benchmarks, with particular strengths in tool use and complex multi-step interactions.

Input price
$0.20 /1M
Output price
$0.80 /1M
Context
131K
Modalities
text
Released
Sep 9, 2025
Tool calling
✓ Yes
Atlas signal
83/100
01

Overview

LongCat-Flash-Chat is a large-scale Mixture-of-Experts (MoE) model with 560B total parameters, of which 18.6B–31.3B (≈27B on average) are dynamically activated per input. It introduces a shortcut-connected MoE design to reduce communication overhead and achieve high throughput while maintaining training stability through advanced scaling strategies such as hyperparameter transfer, deterministic computation, and multi-stage optimization. This release, LongCat-Flash-Chat, is a non-thinking foundation model optimized for conversational and agentic tasks. It supports long context windows up to 128K tokens and shows competitive performance across reasoning, coding, instruction following, and domain benchmarks, with particular strengths in tool use and complex multi-step interactions.

Access: available through the official Meituan API. Context window: 131K.

Pricing and metadata from the ModelsAtlas catalog.Last refreshed Aug 5, 2026
02

Specifications

API identifiermeituan/longcat-flash-chat
ProviderMeituan
Model typeText LLM
Context window131K catalog
Input modalitiestext
Output modalitiestext
ReleasedSep 9, 2025
ModeratedNo
Architecture modalitytext->text
Supported parameters
frequency_penaltylogit_biasmax_tokensmin_ppresence_penaltyrepetition_penaltyresponse_formatseedstopstructured_outputstemperaturetool_choicetoolstop_ktop_p
Source: provider documentation + catalog feedMethodology →
03

Pricing

Live pricing components for meituan/longcat-flash-chat as published in the catalog.

$0.20
Input /1M
$0.80
Output /1M
Per request
Per image
Pricing componentRaw unit priceNormalizedUnit
Prompt tokens$2e-7$0.20 / 1MPer input token
Completion tokens$8e-7$0.80 / 1MPer output token
Request feeN/AN/APer request
Image feeN/AN/APer image unit
Web search feeN/AN/APer search request
Source: catalog pricing feedCompare all pricing →Cheapest models →
04

Cost calculator

Estimate monthly spend from your own token volumes.

Scale / Volume

Estimated Total Cost

$0.26Calculated from current list pricing in the ModelsAtlas catalog.
05

Capabilities

CapabilityStatusWhat it means
Visual UnderstandingNot advertisedImage and document analysis support
Audio ProcessingNot advertisedSpeech and voice aligned flows
Tool Calling✓ SupportedSupports tools / function calling
Self-HostingNot advertisedDeploy outside managed APIs
Derived from catalog capability tags and supported parameters
06

Capability signals

Directional signals derived from model metadata and capability tags — not official benchmark submissions.

Meituan: LongCat Flash Chat
MMLU Signal (Reasoning)
85signal
Coding Signal (HumanEval proxy)
88signal
Math Signal (GSM8K proxy)
79signal
Science Signal (GPQA proxy)
79signal
Estimated from metadata — not an official benchmark runAll benchmarks →Methodology →
07

Quick start

Call Meituan: LongCat Flash Chat through an OpenAI-compatible client.

from openai import OpenAI

client = OpenAI(
    base_url="https://openrouter.ai/api/v1",
    api_key="YOUR_API_KEY"
)

response = client.chat.completions.create(
    model="meituan/longcat-flash-chat",
    messages=[{"role": "user", "content": "Explain quantum physics."}]
)

print(response.choices[0].message.content)
08

Alternatives

Closest models by context window from a different provider.

09

Sources & attribution

Pricing and metadata are maintained in the ModelsAtlas catalog.

Capability bars use metadata tags — directional estimates onlyHow we source and verify data →
10

Frequently asked questions

How much does Meituan: LongCat Flash Chat cost?

$0.20 per 1M input tokens and $0.80 per 1M output tokens on the official Meituan API.

What is the context window of Meituan: LongCat Flash Chat?

Meituan: LongCat Flash Chat supports up to 131K tokens of context.

Does Meituan: LongCat Flash Chat support tool / function calling?

Yes, Meituan: LongCat Flash Chat supports tool / function calling — you can register tools and the model will emit structured tool calls.

How do I access Meituan: LongCat Flash Chat?

Use the official Meituan API with the model id `meituan/longcat-flash-chat`. See the Quick start section above for code examples.