这是我在单 agent 迭代中个人首选超越 Opus 的第一个模型。Fast Grok 4.5 体验非常非常好!该模型始终……(原文截断)
对你意味着
事实Cursor 联创明确表示 Grok 4.5 Fast 是他首次在单 Agent 迭代场景中个人优先选择超过 Claude Opus 的模型。
观点这不是一般用户评测,而是来自深度集成多模型平台的产品决策者,意味着 Anthropic 在 Cursor Agent 场景中的首选地位首次受到实质挑战。对正在用 Claude 驱动 Agent 产品的你来说,这是需要标记的竞争拐点信号。
This is the first model that I personally prefer over Opus when I am iterating with a single agent. Fast Grok 4.5 feels very very good! The model alwa...
观点Cursor 联创的简短背书意味着该模型很可能进入 Cursor 的默认模型菜单——这对所有依赖 Cursor 的 AI 开发团队是直接采购决策信号。若你团队在权衡 API 调用成本,Grok 4.5 现在是性价比候选项。
Good model :)Artificial Analysis: SpaceXAI just released Grok 4.5, and it ranks #4 on GDPval-AA v2 with an Elo of 1543 - behind only the latest Claude releases from Anthropic on real-world agentic knowledge work tasksGrok 4.5 achieved this score at a cost of $0.49 per GDPval task to sit clearly on the Pareto
观点头部 AI 应用团队正在验证 xAI 最新模型的生产力表现,基座模型竞争格局持续演变。CEO 需关注多模型路由策略,保留基座灵活替换的空间。
Grok 4.5 feels very good to use!
今日没有观点主题
尚未被主编选入事件/主题的原始推文(共 28 条,按时间倒序)
@sualehasif99608-05 05:05
RT OpenAI: We're detailing two new incidents that occurred during external cyber evaluations conducted by independent evaluation partners. We outline ...
RT jenny wen: ok! quick updates from me: - i left anthropic (who does that?!) - had a baby (i love her) - and am joining @cursor_ai as head of design ...
RT Elliot Arledge: KernelBench-Hard update: Grok 4.5 high vs GPT 5.6 Sol xhigh on RTX PRO 6000. Both given unlimited time, graded on hw roofline score...
RT dax: i've been using this as my default since yesterday it's super fast and i think it's the first model of theirs that clears the bar for day to d...
RT Vicent Martí: This is a major breakthrough in both intelligence and cost efficiency. The team really outdid themselves! Sadly it’s a large enough...
We've gotten really really good at RL. Composer 2.5 is fighting well-above its weight class. Very excited for the next release as we scale model sizes...
Composer 2 is our first model where the intelligence is truly at the frontier-level. As we scale our compute and pretraining + RL stack, you should co...
0.50 is a really big release! The new tab model is an OOM better at jumps and the cursor agent can run in the background. You can interact with your a...