Tag
TechCrunch Disrupt 2026 will feature a session with Nvidia executives Nader Khalil and Sydney Sykes debating the trade-offs between open and closed AI models and their impact on startup decisions.
Cognition's SWE-2 model, post-trained from Kimi K3, achieves a score of 92.8 on Terminal-Bench 2.1, offering competitive performance with lower cost compared to frontier models like Fable 5.1 and GPT-6 Astra.
A sarcastic reaction to Pokee AI's Pokee-Isaac 28B, a claimed 10M-token context agentic model that fits on a single RTX 4090 but is not open-weights and only available via API.
This paper benchmarks agentic review systems for peer review, evaluating open-source and proprietary systems on research papers. The best configuration achieves 83.0% pairwise accuracy and catches 71.6% of injected errors, but user feedback highlights issues with false positives and nitpicks.
Based on OpenRouter data, open-source LLMs have overtaken proprietary models in token market share, shifting from a 60-40 split favoring proprietary to 60-40 favoring OSS in three months.
GLM-5.2 (max) is currently ranked as the third best AI model overall according to Artificial Analysis' Intelligence Index, with detailed analysis of intelligence, openness, cost, and token usage.
Apple has developed its own foundation models for AI, signaling its entry into the large language model space with proprietary technology.
Meta has abandoned its open-weight Llama model family in favor of a fully proprietary model called Muse Spark, developed by Alexandr Wang's team, marking the end of Meta's role as a champion of open-source AI.
Epoch AI Research analyzed the capability gap between open-weight and proprietary AI models, finding that open-weight models have been trailing the state of the art by approximately four months since the start of the year.