标签
文章批评了一个病毒式传播的AI基准测试,该测试声称Grok 4.6得分为1753,而人类专家为1000。文章指出,该测试基于AI输出之间的偏好比较,而非客观正确性,因此外观精美的工作可能会获胜,但未必真正更好。
埃隆·马斯克宣布,Grok 4.6 针对 Grok Build 工具链进行了优化,没有它性能会显著下降。
Aravind Srinivas 祝贺 SpaceXAI 推出 Grok 4.6,指出它在使用 Perplexity Computer 工具链的 Wide-And-Deep-Research 基准测试中表现出色,现已向 Pro 和 Max 用户开放。
Grok 4.6 正式发布,定价与 4.5 相同,Agent 能力和编程能力增强,复杂任务跑得更久,AA Intelligence Index 追平 GPT-5.6 Sol,输入每百万 Token 2 美元、输出每百万 Token 6 美元。
A hands-on field guide to Grok 4.6, highlighting its speed, dense communication style, and effective prompting patterns for coding and knowledge work.
Ray Fernando 主持一场使用 Cursor 中的 Grok 4.6 进行的直播设计挑战,展示 AI 辅助设计工作流程。
xAI releases Grok 4.6, a frontier model focused on long-running agents and ambitious interactive/visual work, matching GPT-5.6 Sol on the Artificial Analysis Intelligence Index and available in Cursor and Grok Build.