Meituan Longcat Flash Chat features 560B parameters, MoE architecture and 131k-token context window. LongCat-Flash-Chat is a large-scale Mixture-of-Experts (MoE) model with 560B total parameters, of which 18.6B–31.3B (≈27B on average) are dynamically activated per input. It introduces a shortcut-connected MoE design to reduce communication overhead and achieve high throughput while maintaining training stability through advanced scaling strategies such as hyperparameter transfer, deterministic computation, and multi-stage optimization. This release, LongCat-Flash-Chat, is a non-thinking foundation model optimised for conversational and agentic tasks. It supports long context windows up to 128K tokens and shows competitive performance across reasoning, coding, instruction following, and domain benchmarks, with particular strengths in tool use and complex multi-step interactions. Available at $0.2/1M tokens.
Visit Meituan: LongCat Flash ChatAI-Powered
Leverages advanced AI technology to deliver cutting-edge capabilities and results.
Fast & Efficient
Optimized performance ensures quick results without compromising on quality.
Purpose-Built
Specifically designed for llms tasks and workflows.
Meituan Model Timeline
131k tokens context
Specifications
AI Evaluation
Optimized for programming tasks, this model excels at code generation, debugging, and software engineering workflows with solid benchmark performance.
Pros
- Competitive pricing ($0.2/1M)
- 131k token context window
- Large-scale 560B architecture
- Strong code generation and debugging
Cons
- Requires substantial compute
- May lack creative flair
- Speed/quality trade-off
Related Tools
FLUX
FLUX, from Black Forest Labs, is a family of high-quality open and commercial image generation models prized for photorealism and prompt adherence. Widely integrated across third-party tools and APIs, it has become a default backbone for image generation.
Claude Opus 5.5
Claude Opus 5.5, launched 22 September 2026, is Anthropic's current Opus-line flagship, replacing Opus 5. It performs at roughly the level of Claude Fable 5.1 on most work while costing 40% less to run than Opus 5, at $4 input / $20 output per million tokens (both 20% down, cache reads cut 60% to $0.20). It leads on Terminal-Bench 4.0, GDPval-AA v2.1 and Humanity's Last Exam, and is the first Opus model to ship with Fable-5.1-class cybersecurity, biology and distillation safeguards.
Claude Sonnet 5.5
Claude Sonnet 5.5, released 28 September 2026, is the second model in Anthropic's Claude 5.5 family and the faster, lower-cost complement to Claude Opus 5.5. It keeps Sonnet 5's $2 input / $10 output per million tokens ($0.20 cache reads) while generating output over 30% faster, and scores 70.6% on Terminal-Bench 4.0, above Opus 5.5's 66.4%. It has a 1M-token context window, 128K max output, adaptive thinking, and is the first Sonnet with cyber and anti-distillation classifiers.
