Model
Ling-3.0-flash: 124B Total Parameters, Only 5.1B Activated, Agent Execution Cost Approaches Zero
Ling-3.0-flash model is now live on OpenRouter, with 124B total parameters but only 5.1B activated. Its output quality is close to some flagship models, and token cost is about half of Claude. It excels at execution tasks like cross-file bug fixing and converting long documents into structured tables, with 256K context and full tool calling enabled. It is currently in a free period. The community has developed a layered approach of "flagship model planning + Ling-3.0-flash execution," improving overall efficiency several times while reducing costs to a fraction.
Read the original (opens in a new tab)
News stream data aggregated by AI HOT