ToolNavs Find Useful AI Tools
Submit Sign in

Meituan LongCat

560B total parameters with only 27B active, plus 100,000 free tokens a day: LongCat pushes the barrier down to cost. ScMoE and asynchronous RL are its two main threads, and the models now reach from text to image, speech and video.

Hitting 100 tokens per second with 27B active parameters, LongCat makes its trade-off on cost: Flash-Chat pairs 560B total with 27B active, Flash-Omni folds text, image, audio and video into one discrete token space, and Flash-Thinking trades asynchronous RL for reasoning depth. The API grants 100,000 tokens a day and speaks both OpenAI and Anthropic formats.