All models

LongCat-Next

byMeituan LongCatMeituan LongCat· 25 Mar 2026
MultimodalImage generationAudio generation

A multimodal model which can process text, vision, and audio as both inputs and outputs.

Specs
Params73B total, 3B active
LicenseMIT
Adoption · Hugging Face
RAM score
Relative Adoption Metric not scored because required parameter, download, or API metadata is not cataloged. This is not a zero score.
Hugging Face Downloads
200
last 30d
31K
all time
HF Likes
211

Relative Adoption Metric not scored because required parameter, download, or API metadata is not cataloged. This is not a zero score.

Related Models