AI Builders Digest
Bilingual edition · 双语对照版
第 71 期|2026-07-28|双语精选版|5 条精选|5 位作者|5 个主题 返回目录
编者导语 / Editor's Note

Sottiaux 的用量重置进化为庆祝派对(**9474 赞**)——「庆祝 ChatGPT Work 的快速采用和所有投入的努力」+「抓住你的 ultra 和 /fast」。Altman 一个词「wrong」获得 **7993 赞**。Rauchg 发布 Vercel 安全基准:Grok 4.5 在性价比上是最优网络安全 AI 模型——比 Sol 便宜 10 倍、比 Opus 5 便宜 5.7 倍(622 赞)。Steipete「我的 agent 报了一个 bug,他们的 agent 在同一个晚上修好了」(325 赞)——agent-to-agent 协作的真实案例。Levie 确认「AI 负面就业影响就是不发生」(219 赞)+ K3 权重到了(341 赞)。Swyx 论「$ per token 已死,$ per task 才是相关成本指标」(94 赞)。Peter Yang 分享 OpenAI DevX Jason 的故事:骑车时 Codex 远程修了发布视频(338 赞)。Nan Yu 论聪明人应该做自己的产品(247 赞)。播客是 Sachin Katti(与 07-18 同一期,不再重复翻译)。

Theme 01

Usage Reset Party & Altman's One Word / 用量重置派对与 Altman 的一个词

Sottiaux 的庆祝重置(**9474 赞**)+ 回来了(8471 赞)+ 暂别(3199 赞);Altman「wrong」(**7993 赞**)。

Sottiaux / Sam Altman avatarS/
Sottiaux / Sam Altman
Codex & ChatGPT @OpenAI / OpenAI CEO
中文

Sottiaux 的庆祝重置(9474 赞):「我们今天庆祝 ChatGPT Work 的快速采用和所有投入的惊人努力。我感觉像一次 limit reset。抓住你的 ultra 和 /fast,几小时后见。」重置已经从功能演变成文化事件。

他之前的重置推文(8471 赞):「回到笔记本电脑前。Codex 和 ChatGPT Work 所有付费用户的用量限制已重置。Weeeeeeeee。美好的一天!」加上他的暂停公告(3199 赞):「我决定暂时离开 x 充电一下。明天见。」

Altman 的一词反驳(7993 赞):「wrong。」——feed 中每字符互动最高的推文。

Thibault Sottiaux:今天庆祝 ChatGPT Work 的快速采用和所有努力。感觉像一次 limit reset。

Thibault Sottiaux:回到电脑前。用量限制已重置。Weeeeeeeee。

Thibault Sottiaux:决定暂时离开 x 充电。明天见。

Sam Altman:wrong。

English

Sottiaux's celebration reset (9474 likes): 'We're celebrating the fast adoption of ChatGPT Work and all the incredible effort that went into it today. I'm feeling like a limit reset. Hold on tight to your ultra and /fast and see you in a few hours.' The reset has evolved from utility to cultural event.

His earlier reset tweet (8471 likes): 'Back at the laptop. The usage limits have been reset for all paid users of Codex and ChatGPT Work. Weeeeeeeee. It's a good day!' Plus his break announcement (3199 likes): 'I have decided to take a break from x to recharge a bit. See you back tomorrow.'

Altman's one-word rebuttal (7993 likes): 'wrong.' — the highest engagement per character ratio in the feed. Context suggests it was responding to a claim about OpenAI's approach.

Thibault Sottiaux: We're celebrating the fast adoption of chatGPT Work and all the incredible effort that went into it today. I'm feeling like a limit reset.

Thibault Sottiaux: Back at the laptop. The usage limits have been reset for all paid users of Codex and ChatGPT Work. Weeeeeeeee.

Thibault Sottiaux: I have decided to take a break from x to recharge a bit. See you back tomorrow.

Sam Altman: wrong.

Theme 02

Cyber Benchmarks & Agent-to-Agent / 网络安全基准与 Agent 间协作

Rauchg 发布 Vercel 安全基准(**622 赞**)+ Kimi 安全边界(295 赞);Steipete agent-to-agent 修 bug(**325 赞**)+ 安全工作被误解(690 赞)。

Rauchg / Steipete avatarR/
Rauchg / Steipete
Vercel CEO / OpenClaw
中文

Rauchg 的网络安全基准(622 赞):「在我们最新的 Vercel 基准测试中,Grok 4.5 在性价比上是最优网络安全 AI 模型。比 Sol 便宜 10 倍、比 Opus 5 便宜 5.7 倍。」——为模型路由决策提供真实的成本性能数据。

他对 Kimi 论文的观察(295 赞):「Kimi 的论文强调了正确的 agent 安全边界的重要性。容器级隔离不够。在他们的实验中,agent 崩溃了底层机器。」安全边界比以往更重要。

Steipete 的 agent-to-agent 里程碑(325 赞):「我的 agent 报了一个 bug,他们的 agent 在同一个晚上修好了。@jarredsumner 的 robobun 设置就是未来。」加上他的沮丧(690 赞):「我们和世界上最优秀的团队一起做所有这些了不起的安全工作,人们还是觉得我们不安全。」

Guillermo Rauch:Vercel 基准测试中,Grok 4.5 在性价比上是最优网络安全 AI 模型。比 Sol 便宜 10 倍。

Guillermo Rauch:Kimi 论文强调正确的 agent 安全边界。容器级隔离不够。

Peter Steinberger:我的 agent 报了 bug,他们的 agent 修好了。同一个晚上。

Peter Steinberger:我们和最优秀的团队做安全工作,人们还是觉得我们不安全。

English

Rauchg's cybersecurity benchmark (622 likes): 'In our latest Vercel benchmarks, Grok 4.5 has emerged as the best cybersecurity AI model on price-performance. It's 10x cheaper than Sol, 5.7x cheaper than Opus 5, and 2.2x cheaper than [the next best].' — real cost-performance data for model routing decisions.

His Kimi paper observation (295 likes): 'Kimi's paper underlines the importance of the right security boundary for agents. Container-level isolation is not enough. In their experiments, agents crashed the underlying machine.' Security boundaries matter more than ever.

Steipete's agent-to-agent milestone (325 likes): 'My agent reported a bug, their agent fixed it. [in the same night] @jarredsumner's robobun setup is future.' Plus his frustration (690 likes): 'We do all the amazing security work with some of the best teams in the world and people still think we're unsafe.'

Guillermo Rauch: In our latest Vercel benchmarks, Grok 4.5 has emerged as the best cybersecurity AI model on price-performance. It's 10x cheaper than Sol.

Guillermo Rauch: Kimi's paper underlines the importance of the right security boundary for agents. Container-level isolation is not enough.

Peter Steinberger: My agent reported a bug, their agent fixed it. [in the same night]

Peter Steinberger: We do all the amazing security work with some of the best teams in the world and people still think we're unsafe.

Theme 03

Cost Metrics, DevX Stories & AI Jobs / 成本指标、DevX 故事与 AI 就业

Swyx 论 $/task 替代 $/token(**94 赞**);Peter Yang 分享 Jason 骑车修视频(**338 赞**);Levie 确认 AI 就业负面影响不发生(**219 赞)+ K3 权重(341 赞)。

Swyx / Peter Yang / Levie avatarS/
Swyx / Peter Yang / Levie
swyx / 个人开发者 / Box CEO
中文

Swyx 论成本指标(94 赞):「顺便说一下,每输入/输出 token 的 $ 去年就已经作为相关成本指标死了。如果你还没把 x 轴更新为 $/task,我不觉得你还能被认真对待。」行业已从 token 定价转向任务定价。

他对 agent lab 论题的反思(55 赞):「作为 agent lab 论题的发起人——对了 evals/routing/interactivity/ROI 的重点——我必须说,对我自己最大的反对意见是 Claude Code 被意外开源了。」对自己部分错误的原因保持知识诚实。

Peter Yang 的 DevX 故事(338 赞):「我问 @jxnlco(OpenAI DevX):Codex 为你做过的最疯狂的事是什么?我在骑车时同事让我修一个发布视频。所以我远程从[手机]连上去修了。」Levie 论就业(219 赞):「AI 的负面就业影响就是不发生。」加上 K3 权重到了(341 赞)。

Swyx:每 token 的 $ 去年就死了。要更新到 $/task。

Swyx:对我自己最大的反对意见是 Claude Code 被意外开源了。

Peter Yang:Jason 在骑车时远程用 Codex 修了发布视频。

Aaron Levie:AI 的负面就业影响就是不发生。

English

Swyx on cost metrics (94 likes): 'incidentally, $ per input/output tokens died as a relevant cost measure sometime last year. if you haven't updated your x axes to $/task per @ArtificialAnlys then idk if you can be taken seriously anymore.' The industry has shifted from token pricing to task pricing.

His agent lab thesis reflection (55 likes): 'as the progenitor of the agent lab thesis which got the evals/routing/interactivity/ROI focus right, the biggest argument against myself is that Claude Code got accidentally open sourced.' Intellectual honesty about being right for partly wrong reasons.

Peter Yang's DevX story (338 likes): 'I asked @jxnlco (DevEx at OpenAI): What's the craziest thing Codex did for you? I was on a bike ride when a coworker asked me to fix a launch video. So I connected remotely from [my phone].' Levie on jobs (219 likes): 'The negative AI jobs outcome just continues to not be happening.' Plus K3 weights arrival (341 likes).

Swyx: $ per input/output tokens died as a relevant cost measure sometime last year. if you haven't updated your x axes to $/task.

Swyx: the biggest argument against myself is that Claude Code got accidentally open sourced.

Peter Yang: I asked @jxnlco (DevEx at OpenAI): What's the craziest thing Codex did for you? I was on a bike ride when a coworker asked me to fix a launch video.

Aaron Levie: The negative AI jobs outcome just continues to not be happening as some predicted.

Theme 04

Smart People, Product Reviews & Content Engine / 聪明人、产品评审与内容引擎

Nan Yu 论聪明人做自己的产品(**247 赞**);Madhu Guru 论产品评审(36 赞);Zara 的内容创意图(61 赞);Nikunj 的旅行回顾(4 赞)。

Nan Yu / Madhu Guru / Zara / Nikunj avatarNY
Nan Yu / Madhu Guru / Zara / Nikunj
a16z / AI 产品经理 / Builder / FPV Ventures
中文

Nan Yu 论人才分配(247 赞):「很多非常聪明的人非常努力地让产品变得非常好用。如果你有非常聪明的人,你应该让他们做你自己的产品变得非常好用。并付[他们好工资]。」——对人才去其他公司产品的批评。

Madhu Guru 论产品评审(36 赞):「大多数产品评审会议很烦人,因为它们运行得不好。最好的产品评审通过模拟市场对你想法的反应,将几个月的学习压缩到一小时。」

Zara 的内容系统(61 赞):「我如何获得无尽的创意,一张图。」加上她的励志语录(21 赞):「你寻找的魔法在你正在回避的工作中。」Nikunj 的旅行回顾(4 赞):两周旅行中全程用 Claude Code,然后让它做完整回顾。

Nan Yu:如果你有聪明人,应该让他们做你自己的产品。并付好工资。

Madhu Guru:最好的产品评审将几个月学习压缩到一小时。

Zara:如何获得无尽的创意,一张图。

Nikunj:旅行全程用 Claude Code,然后让它做完整回顾。

English

Nan Yu on talent allocation (247 likes): 'A lot of very smart people work very hard to make the product very good and nice to use. If you have very smart people, you should have them make your own product very good and nice to use. And pay [them well].' — a critique of talent going to other companies' products.

Madhu Guru on product reviews (36 likes): 'Most product review meetings are annoying because they're poorly run. The best product reviews compress months of learnings into an hour by simulating the market reaction to your ideas.'

Zara's content system (61 likes): 'How I get endless content ideas, in one picture.' Plus her motivation quote (21 likes): 'The magic you're looking for is in the work you're avoiding.' Nikunj's trip retrospective (4 likes): used Claude Code throughout his two-week trip and asked it for a full retrospective.

Nan Yu: If you have very smart people, you should have them make your own product very good and nice to use. And pay [them well].

Madhu Guru: The best product reviews compress months of learnings into an hour by simulating the market reaction.

Zara: How I get endless content ideas, in one picture.

Nikunj: I used Claude Code as my primary interface for my two week trip and then asked it to do a full retrospective.

Theme 05

Podcast Note: Sachin Katti (Repeat) / 播客备注:Sachin Katti(重播)

MAD Podcast × OpenAI 计算负责人 Sachin Katti——与 2026-07-18 期刊同一期播客。完整中文译文请参见 7 月 18 日刊。

MAD Podcast (repeat) avatarMP
MAD Podcast (repeat)
与 07-18 同一期,不再重复翻译
中文

与 2026-07-18 期刊中同一期 Sachin Katti 播客(发布于 2026-07-16),feed 中重复推送。完整中英双语译文请参见 7 月 18 日刊。核心话题:需求远超算力供应、液冷工厂把电子变 token、Jalapeno 自研芯片、推理成为算力主体、电网投资、递归 AI 设计下一代系统。

备注:与 2026-07-18 期刊内容完全一致。完整译文请参见 ai-builders-digest-2026-07-18.json。

English

Same Sachin Katti podcast as 07-18 issue (published 2026-07-16). Full bilingual transcript in ai-builders-digest-2026-07-18.json. Topics: demand outstripping compute supply, liquid-cooled factories turning electrons into tokens, Jalapeno custom silicon, inference as majority of compute, grid investment, recursive AI designing next-gen systems.

Note: identical to 2026-07-18 issue. See ai-builders-digest-2026-07-18.json for full transcript.