All models

LongCat-Next

byMeituan LongCatMeituan LongCat· 25 Mar 2026
MultimodalImage generationAudio generation

A multimodal model which can process text, vision, and audio as both inputs and outputs.

Specs
Params73B total, 3B active
LicenseMIT
Adoption · Hugging Face
RAM score
Relative Adoption Metric not scored because required parameter, download, or API metadata is not cataloged. This is not a zero score.
Hugging Face Downloads
216
last 30d
30.9K
all time
HF Likes
207

Relative Adoption Metric not scored because required parameter, download, or API metadata is not cataloged. This is not a zero score.

Related Models