Meituan LongCat 2.0 (Tested): China's 1.6T OPEN MODEL looks CRAZY!
In this video, I'll be telling you about Meituan’s new LongCat-2.0 model, a massive 1.6 trillion parameter Mixture-of-Experts model that could become one of the biggest open-weight AI models once the weights are fully uploaded. -- Key Takeaways: 🚀 Meituan has released LongCat-2.0, a huge Mixture-of-Experts model with 1.6 trillion total parameters. 📦 The Hugging Face page is live, but the actual model weights are still being uploaded. 🧠 LongCat-2.0 activates around 48 billion parameters per token, making it much larger than the first LongCat model. 📚 The model includes LongCat Sparse Attention and was trained for long-context, coding, research, and agentic tasks. ⚙️ Meituan says the model was trained and deployed on AI ASIC superpods instead of the usual Nvidia GPU setup. 📊 Official benchmarks look strong, especially for agentic coding, SWE-bench, reasoning, and long-context tasks. 🧪 In my one-shot testing through the free chat site, LongCat-2.0 did not perform very well on coding tasks. 🤖 The model may perform better in a real agent setup, but the API and coding plan are currently only available in China. 🌍 Even with mixed test results, a 1.6 trillion parameter open-weight model is still a major release for the open-source AI ecosystem. 👍 Overall, LongCat-2.0 is an exciting open model release, but I need proper agent testing before judging its real coding ability.