@10xmylife: Do large models really have a moat? As everyone gradually masters the training methods, even if a new model is impressive at launch, the novelty wears off after about six months, and latecomers slowly catch up. Will the winner of this race be a single super-powerful model, or will multiple models carve out their own domains?
Summary
Discussion on whether large models have a moat. As training methods become widespread, new models are impressive but easily caught up. Explores whether the final landscape will be dominated by a single super model or multiple models in their respective fields.
Similar Articles
@RookieRicardoR: Domestic models break through again, matching top models like Claude 4.6 and Gemini 3.1 Pro. Just tested Qwen3.7-Max, sharing some real thoughts. Last night I topped up as soon as the API went live and chose three tasks (see video) to test Qwen3.7-Max's frontend capabilities…
The user tested Qwen3.7-Max and believes it matches top models like Claude 4.6 and Gemini 3.1 Pro in frontend, computing power, and Agent capabilities. Its reasoning ability has significantly improved, and with monthly iteration speed, it has become a first-tier domestic model.
@seclink: Meituan has also started recruiting large model talents. kimi and stepfun are in danger if they don't go public soon. Xiaomi mimo is also strong, Meituan longcat is also strong, deepseek is also strong... https://zhaopin.meituan.com/we…
Meituan has started recruiting large model talents. The tweet mentions models such as Kimi, Stepfun, Xiaomi Mimo, Meituan Longcat, DeepSeek, implying intensified competition for large model talent.
@Xudong07452910: Let the model set its own problems, solve its own problems, and train itself — the biggest fear is learning incorrect problems along with the correct ones. This paper by the Qwen team proposes Skill Self-Play, adding a continuously updated skill library to the model's self-evolution. There are three roles in training: The Proposer generates tasks that are just challenging enough based on the skills...
The Qwen team proposes the Skill Self-Play framework, which significantly improves model capabilities on tool-calling and reasoning tasks through the collaboration of Proposer, Solver, and a dynamic skill controller in self-play.
@10xmylife: I try every domestic model update, and one trend is becoming clearer: each version update of Kimi, GLM, and DeepSeek is better than the last. I think this indicates that the domestic model R&D system is entering a virtuous cycle: more users bring richer real data; engineering infrastructure is maturing; compute power is still tight but forces optimization; pricing advantages are clear.
The author observes that domestic AI models (Kimi, GLM, DeepSeek) improve with each update, believing that domestic model R&D has entered a virtuous cycle: user growth brings abundant data, mature engineering infrastructure, tight compute power forcing efficiency optimization, and clear pricing advantages.
@jakevin7: I have deleted Superpower. Now models are getting stronger, Agent capabilities are getting stronger, there is no need for such complex skills anymore, they only hinder the model's performance.
User @jakevin7 states they have deleted Superpower, believing that as model and Agent capabilities improve, complex skills are no longer needed and may even impair model performance.