DeepSeek Releases V4-Flash-0731, Outperforming Its Bigger Model
DeepSeek's official V4-Flash-0731 release outperforms its larger V4-Pro model on coding and agent benchmarks, despite activating far fewer parameters per token.
Read more →This project is suspended: no new articles or editions will be published. The archive stays available.
Every AI News story tagged with both AI Agents and Mixture of Experts (MoE) — the two topics side by side, updated as new articles publish.
2 articles
DeepSeek's official V4-Flash-0731 release outperforms its larger V4-Pro model on coding and agent benchmarks, despite activating far fewer parameters per token.
Read more →Ant Group's Ling-3.0-Flash matches models several times its size on reasoning and coding benchmarks using just 5.1 billion active parameters, the company says.
Read more →