@0xLogicrw: Zhipu AI founder and chief scientist Tang Jie predicts that the biggest breakthrough in large models this year will be long-horizon tasks, where AI can continuously operate in real environments and solve complex problems. Once long-horizon tasks are achieved, today's 'one-person companies' will rapidly become 'no-employee companies...
Summary
Zhipu AI founder Tang Jie predicts that the biggest breakthrough in large models this year will be long-horizon tasks, where AI can continuously solve complex problems in real environments, and mentions three technical pillars and Anthropic's progress in autonomous training.
Similar Articles
@jietang: Recent thoughts: The Shift to Long-Horizon Tasks The most likely breakthrough this year will be in long-horizon tasks. …
The article discusses the anticipated breakthrough in long-horizon AI tasks and autonomous agents, suggesting a shift from 'one-person' to 'none-person' companies. It highlights technical pillars like memory, continual learning, and self-judging as key to realizing fully self-evolving AI systems that could redefine AGI and operating systems.
@jakevin7: Let me make a prediction: The next phase of the AI era will become "Infra is all you need". AI-generated code is already very powerful, but it's still far from adequate in terms of usability and stability. Recently, OpenAI's subscription system had a huge bug, and the membership system completely broke down. The system…
The author predicts that the next phase of the AI era will shift from model capabilities to infrastructure capabilities, emphasizing infra abilities such as reproducibility, observability, recoverability, and security isolation, believing that stably carrying AI behavior will be the key to competition.
@MaxForAI: A brutal fact: AI's mathematical abilities have already surpassed 99% of humans on this planet. @MenloVentures partner Deedy @deedydas, who invested in Anthropic and OpenRouter, conducted a test. He used the just-concluded 2026 International Mathematical Olympiad...
AI models Fable, Sol, K3, and Axiom all achieved a perfect score of 42/42 in the 2026 International Mathematical Olympiad, solving the competition completely for the first time at low cost. Among them, Claude Fable 5 was the fastest, while GPT 5.6 Sol had the lowest cost.
@seclink: Zhipu's latest flagship model GLM-5.3 has been officially unveiled, achieving major breakthroughs in programming capabilities and cybersecurity vulnerability detection, launching a new challenge to AI leaders like Anthropic and OpenAI. Indeed, large models are now applied in the cybersecurity field, especially replacing the blue team in previous red-blue team confrontations, those who did daily vulnerability hunting…
Zhipu releases flagship AI model GLM-5.3, achieving major breakthroughs in programming and cybersecurity vulnerability detection, aiming to challenge leading enterprises like Anthropic and OpenAI.
@VincentLogic: If Ilya Is Right, the Three Strongest Consensuses in AI Over the Past Few Years Might All Be Wrong: Scaling Is No Longer the Universal Answer. High Benchmark Scores Don't Equal True Intelligence. RL Might Even Be Making Models 'Dumber'. This Interview, Called 'the Last Interview Before Ilya Disappeared'...
Ilya Sutskever suggested in an in-depth interview that the three core consensuses of the AI industry over the past few years could all be mistaken: Scaling is no longer a silver bullet, high benchmark scores do not equate to real intelligence, and RL is instead making models 'dumber'. He believes the dividends from pre-training and RL are nearly exhausted, AI has re-entered the era of research, and true superintelligence should possess a strong learning capability like a gifted teenager, not a static repository of knowledge.