Karpathy has Opus 5 build a playable Lord of the Rings game for $10
curated from 7 items across 42 tracked sources
Two rival frontier models turned up over the long weekend, and the yardstick shifted with them: less benchmark table, more what a model builds unattended for pocket money — while OpenAI answered on a different front entirely, pure math.
🧭 Recent trends
Deployment and price did the talking: insurer Univé put ChatGPT Enterprise across its staff, and a Japanese electronics chain's shopping agent served 30,000 customers in two weeks.
Agent tooling settled into polish: Grok Build shipped two patches in three days, mostly permissions and navigation, while Naval Ravikant reread the API as an agent-facing interface.
Open weights kept arriving with infrastructure attached: Thinking Machines Lab published Inkling's weights under a staged safety framework, and MiniMax's open video model got day-one serving support.
🔥 Top signals
- Karpathy has Claude Opus 5 build a playable Lord of the Rings game for $10 One paragraph of instructions, a **1M-token** budget, about **$10** spent. A whole finished build, not a chat reply, is now how a top researcher sizes up a new model. · Andrej Karpathy
- Grok 4.5 ships, beats OpenAI's GPT-5.6 Terra on most shared benchmarks xAI also claims the best speed-and-cost tradeoff of any frontier model, and Musk singles out its C code. The pitch is price per answer, not a new capability. · Elon Musk
- OpenAI publishes ten results on open problems in math and computer science The **ten** items span geometry, cryptography and complexity theory — a lab pushing its output into fields where every claim is checkable line by line, rather than judged by vibes. · OpenAI
Karpathy 让 Opus 5 造出可玩的《魔戒》游戏,只花 10 美元
从 42 个追踪信源的 7 条动态中精选
长周末里两家对手的前沿模型同时到场,衡量标准也跟着变:不再只看基准表,而看模型自己能造出什么;OpenAI 则在纯数学这条完全不同的战线上回应。
🧭 最近趋势
落地在说话:保险公司 Univé 全员部署 ChatGPT Enterprise,日本山田电机购物助手两周服务3万顾客。
工具进入打磨期:Grok Build 三天发两版补丁多是权限与导航,Naval 把 API 重新解读为写给 Agent 的接口。
Thinking Machines 分阶段放出 Inkling 权重,MiniMax 开源视频模型上线当天即获推理支持。
🔥 今日要点
- Karpathy 让 Claude Opus 5 造出可玩的魔戒游戏 只给一段提示、**100万** token 预算,花费约 **10 美元**;顶级研究者现在看模型能不能独立做出完整作品。 · Andrej Karpathy
- Grok 4.5 发布,多数共享基准超过 GPT-5.6 Terra xAI 还称它在速度与成本上位居前沿模型最优,Musk 另夸其 C 代码;卖点是每次回答的性价比,而非新能力。 · Elon Musk
- OpenAI 公布数学与计算机十项未解问题的进展 这 **十** 项覆盖几何、密码学与复杂性理论;实验室把成果推向对错可逐行核验的领域,而不是靠感觉评判。 · OpenAI