LongCat-Image

A 6-billion-parameter bilingual image generation and editing model covering all 8,105 standard Chinese characters, with open-source SOTA on image editing benchmarks.

Independent third-party resource. Not affiliated with or endorsed by LongCat or Meituan.

Overview

LongCat-Image is documented in the meituan-longcat GitHub repository and is listed in the publisher API documentation. The model is the same entity referenced in both sources.

Documented capabilities

The statements below are described in the cited primary documentation. They describe what the cited documentation says the model is, not an independent performance evaluation.

  • Bilingual (Chinese-English) text-to-image generation.
  • Text-driven image editing with open-source SOTA (ImgEdit-Bench 4.50, GEdit-Bench 7.60/7.64).
  • Chinese text rendering covering all 8,105 standard characters (ChineseWord 90.7).
  • Photorealistic image generation with enhanced aesthetic quality through curated reward models.
  • Compact 6B parameter design — significantly smaller than ~20B MoE alternatives — for minimal VRAM usage.
  • Text-to-image generation quality: GenEval 0.87, DPG-Bench 86.8.
  • Available on LongCat Web and LongCat App with 24 editing templates.

Access and license

Open source. Review the repository license file before commercial use. Multiple model versions released including mid-training and post-training checkpoints.

Independent-site notes

This profile is a source-based summary. It does not contain any benchmark, evaluation, or quality claim produced by this site.

No independent testing has been published by this site.

Any performance figures attributed to the model in third-party materials are vendor-reported and have not been independently verified on this site.

Sources

  • LongCat-Image Technical Report (arXiv:2512.07584)

    Primary sourcePublished 2025-12-08Accessed 2026-07-26

    Technical report describing the MM-DiT + Single-DiT hybrid architecture, data curation pipeline, RL fine-tuning with reward models, and benchmark results including ChineseWord 90.7 and GenEval 0.87.

  • LongCat-Image on HuggingFace

    Primary sourceAccessed 2026-07-26

    Model weights available under the meituan-longcat organization. Includes mid-training and post-training checkpoints, training toolchain, and multiple model versions for text-to-image and image editing.

  • LongCat Image on LongCat Official Website

    Publisher documentationAccessed 2026-07-26

    LongCat-Image is accessible through the LongCat Web interface and LongCat App, with 24 image editing templates and image-to-image capabilities.

FAQ

What is LongCat-2.0?
LongCat-2.0 is a model documented in the meituan-longcat GitHub organization and listed in the publisher API documentation.
Where can the primary source for LongCat-2.0 be found?
The meituan-longcat GitHub repository is the primary source. The repository contains the model code, configuration, and license file.
Does LongCat-2.0 have a documented API?
Yes. The publisher API documentation lists LongCat-2.0 as a supported model on both OpenAI-compatible and Anthropic-compatible endpoints.
Is this page an official LongCat or Meituan page?
No. This page is published by an independent third-party site that summarizes publicly-available primary documentation. It is not affiliated with, endorsed by, or sponsored by LongCat or Meituan.

Independent third-party disclosure

This page is published by an independent third-party site. It is not affiliated with, endorsed by, sponsored by, or operated by Meituan, LongCat, or any of their affiliates. The content summarizes publicly-available primary documentation and does not represent the views of any referenced organization.

Last reviewed: 2026-07-26