AI Builders Digest - 2026-07-26
X / TWITTER
Swyx, affiliated with smol.ai, DX Tips, Cognition, AI Engineer, and Latent Space, is pushing SmolForge toward a more expressive builder surface: customizable skins and spritesheet animations. The more strategic signal is his frustration with "stupid defaults" in existing productivity suites, which explains why he is experimenting with a new GSuite-like workspace rather than another narrow AI wrapper.
https://x.com/swyx/status/2080750437133901925
https://x.com/swyx/status/2080705334587605122
Swyx 正在把 SmolForge 做成更有表达力的 builder 工具:支持自定义 skins 和 spritesheet animations。更重要的信号是,他对现有 productivity suite 的默认设计非常不满,这解释了为什么他不是在做一个窄 AI wrapper,而是在尝试新的类 GSuite 工作空间。
https://x.com/swyx/status/2080750437133901925
https://x.com/swyx/status/2080705334587605122
Google Labs VP Josh Woodward highlighted Gemini Spark as a practical agentic workflow: give Gemini a school calendar PDF and ask it to add every "No School" day to Google Calendar. The message is simple but important: Google is trying to make Gemini feel less like chat and more like an action layer across consumer productivity.
https://x.com/joshwoodward/status/2080771183944073347
Google Labs VP Josh Woodward 展示了 Gemini Spark 的一个具体 agentic workflow:把学校日历 PDF 丢给 Gemini,让它自动把所有 "No School" 日期加入 Google Calendar。这个信号很直接:Google 想让 Gemini 从聊天工具变成消费级 productivity 的行动层。
https://x.com/joshwoodward/status/2080771183944073347
Anthropic Claude Code engineer Boris Cherny argued that Opus 5's most exciting progress is not just coding or data analysis performance, but prompt-injection resistance. His claim is that strong model alignment plus prompt-injection probes plus Auto Mode in Claude Code can drive attack success rates toward zero.
https://x.com/bcherny/status/2080713091688583312
Anthropic Claude Code 工程师 Boris Cherny 认为 Opus 5 最值得关注的进步不只是 coding 或 data analysis 能力,而是 prompt injection 防护。他的判断是:强模型对齐、prompt injection probes 和 Claude Code 的 Auto Mode 叠加后,可以把攻击成功率压到接近零。
https://x.com/bcherny/status/2080713091688583312
OpenAI Codex and ChatGPT builder Thibault Sottiaux said ChatGPT Work is now available globally for all paid plans across mobile, web, and desktop. The positioning is blunt: it puts "a jetpack" on ChatGPT, suggesting OpenAI is continuing to turn ChatGPT from an assistant into a work environment.
https://x.com/thsottiaux/status/2080876712439747052
OpenAI Codex / ChatGPT builder Thibault Sottiaux 表示 ChatGPT Work 已经面向全球所有付费计划开放,覆盖 mobile、web 和 desktop。他的定位很直接:这是给 ChatGPT 装上 "jetpack",说明 OpenAI 仍在把 ChatGPT 从 assistant 推向工作环境。
https://x.com/thsottiaux/status/2080876712439747052
AI educator Peter Yang surfaced two practical founder signals: voice plus ChatGPT can be a surprisingly natural interface for steering long-running Codex work, but it requires disciplined thread naming; and pure software is getting harder for indie developers to monetize unless paired with services or another wedge.
https://x.com/petergyang/status/2080793867960643823
https://x.com/petergyang/status/2080669643577176573
AI 教程作者 Peter Yang 提了两个很实际的 founder 信号:用 ChatGPT Voice 驱动长期运行的 Codex 任务,体验已经很自然,但前提是要认真管理 thread 名称;同时,纯软件对 indie developer 来说越来越难变现,通常需要叠加 services 或其他切入点。
https://x.com/petergyang/status/2080793867960643823
https://x.com/petergyang/status/2080669643577176573
Meta AI senior director Madhu Guru framed the next few years of AI opportunity around people who can translate messy real-world workflows into domain-tuned foundation-model systems. His stack is specific: understand the work, design evals, improve models through post-training, and build feedback loops.
https://x.com/realmadhuguru/status/2080707454422413487
Meta AI senior director Madhu Guru 把未来几年的 AI 机会定义为:谁能把混乱的现实工作流改造成面向具体领域的 foundation model 系统。这个能力栈很具体:理解真实工作、设计 evals、通过 post-training 改进模型,并建立持续反馈循环。
https://x.com/realmadhuguru/status/2080707454422413487
Anthropic's Cat Wu described Claude Opus 5 as strong at long-running autonomous work. Anthropic's Thariq added a workflow lesson from the launch: for the newest Claude models, the team removed roughly 80% of the Claude Code system prompt, which suggests newer models may reward simpler, cleaner instruction layers.
https://x.com/_catwu/status/2080707593115516985
https://x.com/trq212/status/2080710971228918066
https://x.com/trq212/status/2080703339306913985
Anthropic 的 Cat Wu 强调 Claude Opus 5 擅长长时间 autonomous work。Anthropic 的 Thariq 补充了一个更有操作价值的经验:面向最新 Claude 模型,团队移除了约 80% 的 Claude Code system prompt,这说明新模型可能更适合更简单、更干净的 instruction layer。
https://x.com/_catwu/status/2080707593115516985
https://x.com/trq212/status/2080710971228918066
https://x.com/trq212/status/2080703339306913985
Replit CEO Amjad Masad pointed to Etched as a case where early investors were willing to believe before the market caught up, then pushed Anthropic to clarify its stance on open-weight model bans. He also teased that Replit has changed enough that returning users should expect a surprise.
https://x.com/amasad/status/2080864869130416320
https://x.com/amasad/status/2080850075358826871
https://x.com/amasad/status/2080848381967212975
Replit CEO Amjad Masad 提到 Etched 是一个早期投资者先于市场共识下注的案例,同时要求 Anthropic 明确是否支持禁止 open-weight models。他还暗示 Replit 已经发生较大变化,老用户重新使用会有明显惊喜。
https://x.com/amasad/status/2080864869130416320
https://x.com/amasad/status/2080850075358826871
https://x.com/amasad/status/2080848381967212975
Vercel CEO Guillermo Rauch continued to push the ambition line around AI-native product building and shared that Figma2React is "good." The practical signal is that design-to-code workflows are now becoming credible enough for infra and product leaders to treat them as real builder primitives.
https://x.com/rauchg/status/2080714333793972498
https://x.com/rauchg/status/2080706974476583337
https://x.com/rauchg/status/2080646549336678597
Vercel CEO Guillermo Rauch 继续强调 AI-native product building 里的野心问题,并评价 Figma2React 已经 "good"。实际信号是,design-to-code workflow 正在变得足够可信,开始被基础设施和产品负责人当成真正的 builder primitive。
https://x.com/rauchg/status/2080714333793972498
https://x.com/rauchg/status/2080706974476583337
https://x.com/rauchg/status/2080646549336678597
Anthropic researcher Alex Albert framed Opus 5 as a fast-moving jump in practical knowledge work: near-superhuman spreadsheets and slide decks, higher token efficiency, and smoother coding performance than Fable 5 for many tasks. The important pattern is not one benchmark, but the move from chat answers to consultant-grade artifacts.
https://x.com/alexalbert__/status/2080731979528679617
https://x.com/alexalbert__/status/2080703118086693121
https://x.com/alexalbert__/status/2080702002120757562
Anthropic researcher Alex Albert 把 Opus 5 描述为 practical knowledge work 的明显跃迁:接近超人水平的 spreadsheets 和 slide decks,更高 token efficiency,以及在许多 coding 任务上比 Fable 5 更顺手。关键不在单个 benchmark,而在模型正在从回答问题走向交付咨询级 artifacts。
https://x.com/alexalbert__/status/2080731979528679617
https://x.com/alexalbert__/status/2080703118086693121
https://x.com/alexalbert__/status/2080702002120757562
Box CEO Aaron Levie gave one of the clearest enterprise readings of Opus 5: on Box's Complex Work Eval, it improved due diligence, life sciences, legal, technology, and healthcare document workflows versus Opus 4.8. He also argued strongly for open weights, saying they expand customer choice, lower costs for certain workloads, and let many vertical specialists post-train models for real-world domains.
https://x.com/levie/status/2080704871934931221
https://x.com/levie/status/2080761484305654091
https://x.com/levie/status/2080675210991443982
Box CEO Aaron Levie 给出了今天最清晰的 enterprise 视角:在 Box 的 Complex Work Eval 上,Opus 5 相比 Opus 4.8 改进了 due diligence、life sciences、legal、technology 和 healthcare 等文档工作流。他也强烈支持 open weights,理由是它能扩大客户选择、降低部分 workload 成本,并让更多垂直领域团队通过 post-training 解决真实场景。
https://x.com/levie/status/2080704871934931221
https://x.com/levie/status/2080761484305654091
https://x.com/levie/status/2080675210991443982
YC CEO Garry Tan connected AI adoption to a much longer history of technology adoption, arguing that slow adoption compounds into national wealth gaps. His management warning is sharper: macro productivity gains will require CEOs and managers to approve radically different staffing and workflow plans, and this may take 10 years rather than 2.
https://x.com/garrytan/status/2080849953413541982
https://x.com/garrytan/status/2080699367883980924
YC CEO Garry Tan 把 AI adoption 放进更长的技术扩散史里看:一个国家采用新技术的速度,会长期影响财富差距。他对管理层的提醒更尖锐:宏观 productivity gain 需要 CEO 和 manager 批准完全不同的 staffing 和 workflow 方案,这件事可能需要 10 年,而不是 2 年。
https://x.com/garrytan/status/2080849953413541982
https://x.com/garrytan/status/2080699367883980924
FirstMark VC Matt Turck called out model routing as the week's emerging infrastructure theme: rumors around Stripe and OpenRouter, plus launches from Cursor Router and Runway Router, alongside routing work at Databricks, Vercel, Cloudflare, Dataiku, AWS, and Google. He also noted the irony that top AI researchers are building recursive auto-research that may automate parts of their own job.
https://x.com/mattturck/status/2080645582209663049
https://x.com/mattturck/status/2080738638065729741
FirstMark VC Matt Turck 点出本周的 AI infra 主题是 model routing:Stripe / OpenRouter 传闻、Cursor Router 和 Runway Router 发布,以及 Databricks、Vercel、Cloudflare、Dataiku、AWS、Google 都在做类似方向。他还指出一个有意思的反讽:顶级 AI 研究员正在构建 recursive auto-research,某种程度上是在自动化自己的工作。
https://x.com/mattturck/status/2080645582209663049
https://x.com/mattturck/status/2080738638065729741
Builder Zara Zhang argued that the model feature she wants most is speed, not more intelligence: 1-5 minute waits are too short for deep work and too long for passive waiting. She also noted that when agents enter chats and meetings, the history itself becomes PRDs, which gives verbal communicators more leverage than before.
https://x.com/zarazhangrui/status/2080829737044439444
https://x.com/zarazhangrui/status/2080617484261249160
Builder Zara Zhang 认为现在最需要的模型特性不是更聪明,而是更快:1 到 5 分钟的等待既不足以进入 deep work,又太长,容易把人拖去刷信息流。她还指出,当 agent 进入聊天群和会议时,chat history 和 meeting transcript 本身就会变成 PRD,这会让口头表达强的人获得更多生产力杠杆。
https://x.com/zarazhangrui/status/2080829737044439444
https://x.com/zarazhangrui/status/2080617484261249160
OpenClaw and OpenAI builder Peter Steinberger reported a new record for an autoreview skill: 66 rounds on a difficult refactor. The signal is less about the number and more about a new operating mode for coding agents: long-horizon automated review loops are becoming part of real refactor workflows.
https://x.com/steipete/status/2080899298838098034
OpenClaw / OpenAI builder Peter Steinberger 报告了 autoreview skill 的新纪录:一个复杂重构跑了 66 轮。这里的信号不只是轮数,而是一种新的 coding agent 工作方式:长周期自动 review loop 正在进入真实 refactor workflow。
https://x.com/steipete/status/2080899298838098034
Every CEO Dan Shipper gave the most skeptical builder take on Claude Opus 5: strong, but awkward with existing skills and workflows. His team saw better results after deleting older elaborate workflows and starting from scratch, and he suggested medium or low thinking effort may work better than higher effort for this model.
https://x.com/danshipper/status/2080700057892815114
https://x.com/danshipper/status/2080709090909503775
Every CEO Dan Shipper 给出了今天最怀疑派的 Opus 5 评价:能力强,但和既有 skills / workflows 兼容性不好。他们删除旧的复杂 workflow、从零开始后效果明显变好;他还认为这个模型可能在 medium 或 low thinking effort 下表现更好,而不是越高越好。
https://x.com/danshipper/status/2080700057892815114
https://x.com/danshipper/status/2080709090909503775
Sam Altman said he wants the US to win in AI across both open-source and proprietary models, aligning with the broader open-weights letter conversation. He also asked for feedback on "pro-ultra-superhard," signaling continued experimentation with higher-effort or more demanding ChatGPT modes.
https://x.com/sama/status/2080683363174945065
https://x.com/sama/status/2080683119959757243
Sam Altman 表示希望美国同时在 open-source 和 proprietary models 上赢得 AI 竞争,这与今天多位 founder 提到的 open weights 讨论相呼应。他还征求对 "pro-ultra-superhard" 的反馈,说明 ChatGPT 仍在尝试更高 effort、更重任务模式。
https://x.com/sama/status/2080683363174945065
https://x.com/sama/status/2080683119959757243
Claude's official account positioned Opus 5 as a paid-plan and API release priced the same as Opus 4.8, with a Fast mode running about 2.5x the default speed. The account also emphasized stronger cybersecurity performance than Opus 4.8, lower reckless or deceptive behavior in automated audits, and stronger adherence to Claude's Constitution.
https://x.com/claudeai/status/2080699515271528827
https://x.com/claudeai/status/2080699512205537648
https://x.com/claudeai/status/2080699508401328462
Claude 官方账号把 Opus 5 定位为面向付费计划和 API 的新版本,价格与 Opus 4.8 相同,并提供约 2.5x 默认速度的 Fast mode。同时它强调 Opus 5 在 cybersecurity 任务上强于 Opus 4.8,在自动行为审计中 reckless 或 deceptive behavior 更低,对 Claude Constitution 的遵循也更强。
https://x.com/claudeai/status/2080699515271528827
https://x.com/claudeai/status/2080699512205537648
https://x.com/claudeai/status/2080699508401328462
PODCASTS
No Priors - Building an Autonomous Delivery Experience with DoorDash Co-Founders Andy Fang and Stanley Tang
The Takeaway: DoorDash's AI and robotics strategy is not a lab demo strategy; it is a use-case-first operating system for commerce, logistics, and autonomy.
DoorDash co-founders Andy Fang and Stanley Tang described a company that has been quietly treating delivery as an autonomy problem since 2018. The surprising part is that their AI work is already changing demand: Ask DoorDash drives about half of restaurant-use trajectories toward places users have never ordered from before, while grocery use cases show roughly 40% larger baskets when people can ask for meal planning, fridge restocking, or usual-item reorders in natural language.
Their robotics philosophy is even more useful for builders. DoorDash first tried to be the platform layer for outside autonomy companies, then learned that generic robots did not fit the real job. Sidewalk robots were too slow for 3-5 mile average deliveries, while robotaxis were overbuilt for carrying food. The sharper lesson from Stanley Tang: many autonomy companies "build the technology first and then retroactively try to go find a problem to fit into." DoorDash reversed that by asking what vehicle, dispatch system, merchant integration, and first/last-100-feet workflow the delivery use case actually required.
For AI builders, this is the point: agentic commerce will not be won by adding chat to checkout. It will be won by companies that own context, workflow, fulfillment, and feedback loops deeply enough that agents can safely act.
https://www.youtube.com/watch?v=vNpcg_Ma-FA
核心判断:DoorDash 的 AI 和 robotics strategy 不是实验室 demo,而是围绕 commerce、logistics 和 autonomy 打造的 use-case-first 操作系统。
DoorDash 联合创始人 Andy Fang 和 Stanley Tang 透露,公司从 2018 年起就把 delivery 当作 autonomy 问题来研究。最值得注意的是,这些 AI 功能已经改变真实需求:Ask DoorDash 在 restaurant 场景里,大约一半使用路径会把用户带到从未下单过的新店;grocery 场景里,当用户可以用自然语言做 meal planning、fridge restocking 或 usual-item reorder 时,basket size 大约提升 40%。
他们的 robotics 方法论对 builder 更有启发。DoorDash 一开始想做外部 autonomy 公司的平台层,但后来发现通用机器人不适配真实任务。Sidewalk robots 对 3 到 5 英里的平均配送距离太慢,robotaxis 又对送餐过度设计。Stanley Tang 给出的关键教训是:很多 autonomy 公司是先造技术,再回头找问题适配。DoorDash 反过来,从 delivery use case 出发,倒推需要什么车辆、dispatch system、merchant integration,以及 first/last-100-feet workflow。
对 AI builder 来说,重点是:agentic commerce 不会靠在 checkout 上加聊天框取胜。真正的优势来自对 context、workflow、fulfillment 和 feedback loop 的深度掌控,让 agent 能够安全地行动。
https://www.youtube.com/watch?v=vNpcg_Ma-FA
Generated through the Follow Builders skill: https://github.com/zarazhangrui/follow-builders