LongCat AI Models — Independent Resources, Guides & Comparisons

LongCat is Meituan’s open-source AI model family — language models, image generation, video, and text-to-speech.

Not affiliated with, endorsed by, or sponsored by Meituan or LongCat.

Explore LongCat models

Each overview separates verified source material from future independent testing.

Open the LongCat-2.0 model overview

LongCat-2.0

A model documented in the meituan-longcat GitHub organization and exposed through the LongCat publisher API platform.

  • Documented in the meituan-longcat GitHub repository.
  • Listed in the publisher API documentation as a supported model.
  • Exposed through the publisher API platform in both OpenAI-compatible and Anthropic-compatible formats.
Last verified: 2026-07-17
Open the LongCat-Flash-Chat model overview

LongCat-Flash-Chat

A 560-billion-parameter Mixture-of-Experts language model with dynamic computation, averaging 27B activated parameters per token. Built for high-throughput chat and agentic tasks.

  • General conversational AI with 560B total parameters (MoE).
  • Per-token dynamic activation of 18.6B–31.3B parameters (averaging ~27B).
  • 256K context window (upgraded December 2025 from 128K).
Last verified: 2026-07-26
Open the LongCat-Image model overview

LongCat-Image

A 6-billion-parameter bilingual image generation and editing model covering all 8,105 standard Chinese characters, with open-source SOTA on image editing benchmarks.

  • Bilingual (Chinese-English) text-to-image generation.
  • Text-driven image editing with open-source SOTA (ImgEdit-Bench 4.50, GEdit-Bench 7.60/7.64).
  • Chinese text rendering covering all 8,105 standard characters (ChineseWord 90.7).
Last verified: 2026-07-26
Open the LongCat-Flash-Omni model overview

LongCat-Flash-Omni

A 560B-parameter open-source omni-modal model with 27B activated, excelling at real-time audio-visual interaction — text, image, audio, and video in a single end-to-end framework.

  • Real-time audio-visual interaction — text, image, audio, and video within a single framework.
  • 560B total parameters (MoE), 27B activated per token on average.
  • 128K token context window with multi-turn dialogue support.
Last verified: 2026-07-26

View all 10 models →

Resource areas

Browse developer guides

Developer guides

Practical setup and usage guides, published with source references.

View benchmark records

Benchmarks

Independent test records with methodology and limitations disclosed.

Read model comparisons

Model comparisons

Task-focused comparisons designed to explain trade-offs.

How this site handles information

  • Official announcements, repositories, and project pages are linked as sources.
  • Independent tests are labeled separately with dates, methodology, and limitations.
  • Editorial analysis is presented as analysis, not as an official statement.
Read our source policy

10 model profiles · 2 news articles · 3 benchmark records · every claim source-attributed

View all models