🌐 双语
Archive

AI Builders
Digest

2026-07-25 21 builders · 48 tweets · 1 podcasts · 1 blogs

🔥 热点话题

Claude Opus 5 发布:对齐更强、编码更强、提示注入更难Claude Opus 5 Launches: Stronger Alignment, Coding, and Prompt-Injection Resistance

Anthropic 正式推出 Claude Opus 5,价格与 Opus 4.8 相同,成为 Claude Max 默认模型,并提供 Fast 模式(约 2.5 倍速度)。官方称其为迄今对齐最好的模型,reckless 或欺骗行为发生率最低,对 Claude Constitution 的遵守最强。在网络安全任务上强于 4.8,但仍明显落后于 Mythos 5 的 exploit 开发能力。Claude Code 团队 Thariq 分享:新模型移除了约 80% 的系统提示,并总结了为 Claude 5 系列撰写 system prompts、skills 和 Claude.MD 的经验。Boris Cherny 强调 Opus 5 是迄今最难成功提示注入的模型,结合强对齐、注入探针和 Claude Code Auto Mode,攻击成功率可降至接近 0。Alex Albert 指出其 token 效率大幅提升,同时智能水平上升,更适合日常编码。Aaron Levie 在 Box 的复杂企业工作评估中看到显著提升:尽职调查 +17%、生命科学 +30%、法律 +12% 等。Dan Shipper 的初期 vibe check 则更谨慎:模型会与指令争论、过早停止,需从零重建工作流才能发挥潜力,整体像“穷人版 Fable”。
Anthropic launched Claude Opus 5 at the same price as Opus 4.8. It is the default on Claude Max and the strongest option on Claude Pro, with a Fast mode roughly 2.5× quicker. Official posts call it the most aligned model yet—lowest rates of reckless or deceptive behavior and strongest adherence to Claude’s Constitution. It improves on cybersecurity tasks versus 4.8 but remains substantially behind Mythos 5 at developing exploits. Claude Code’s Thariq shared that the team removed ~80% of the system prompt for the new models and published lessons on writing system prompts, skills, and Claude.MDs. Boris Cherny highlighted that Opus 5 is Anthropic’s least prompt-injectable model; combined with alignment, probes, and Claude Code Auto Mode, successful attacks drop to near zero. Alex Albert noted major gains in token efficiency while raising the intelligence bar, making it preferable for many coding tasks. Box CEO Aaron Levie reported meaningful lifts on Box’s Complex Work Eval (due diligence +17%, life sciences +30%, legal +12%). Dan Shipper’s early vibe check was more cautious: the model argues with instructions and stops early unless workflows are rebuilt from scratch, positioning it as a “poor man’s Fable.”
查看原文 →查看原文 →查看原文 →查看原文 →查看原文 →查看原文 →

开放权重模型支持浪潮:行业对齐明显Open-Weights Support Wave: Broad Industry Alignment

开放权重模型获得罕见广泛支持。Sam Altman 明确表示希望美国在开源和专有模型上都赢,并对相关声明表示欢迎。Box CEO Aaron Levie 详细阐述了开放权重的价值:推动垂直领域后训练、安全与网络风险的多样化处理、更高效的训练方法,以及不同成本结构的工作负载分配。他强调开放与封闭并非零和。Amjad Masad 则公开追问 Anthropic 是否会签署相关立场声明,并呼吁员工向领导层确认立场。Madhu Guru 指出,未来几年最大的机会在于把混乱的真实工作流适配到基础模型上——理解工作流程、设计评估、后训练并建立持续反馈循环——这类技能目前仍集中在少数实验室。
Open-weights models received unusually broad support. Sam Altman stated he wants the US to win in both open-source and proprietary models. Box CEO Aaron Levie detailed why open weights matter: they enable post-training for specific verticals, variance in safety and cyber approaches, more efficient training methods under compute constraints, and different cost structures for different workloads. He stressed that open versus closed is not zero-sum. Replit CEO Amjad Masad publicly asked whether Anthropic would sign the related letter and urged employees to clarify leadership’s position. Madhu Guru argued that the biggest near-term opportunity lies in adapting messy real-world workflows to foundation models—understanding how work gets done, designing evals, post-training, and building continuous feedback loops—a skillset still concentrated in a handful of labs.
查看原文 →查看原文 →查看原文 →查看原文 →

DoorDash 联合创始人:从 2018 年起就把自己当成机器人公司DoorDash Co-Founders: We’ve Been a Robotics Company Since 2018

The Takeaway:真正决定自主配送能否规模化的,不是“能不能做无人驾驶”,而是能否从真实客户用例倒推,并利用自己的配送数据优势把运营、硬件和多模态车队一起做出来。

DoorDash 联合创始人 Andy Fang 与 Stanley Tang 在 No Priors 上分享了他们长达八年的自主配送之路。2018 年就开始探索,最初只是 Stanley 和半个工程师的 skunkworks 项目,通过与人行道机器人和 robotaxi 合作验证了“不是会不会发生,而是何时发生”。关键学习是:大多数机器人公司先做技术再找用例,而 DoorDash 坚持从客户问题倒推。人行道机器人速度太慢(平均配送 3-5 英里),robotaxi 又过重且无法解决最后 100 英尺的取送问题。于是他们自己打造了 300 磅、最高 20 mph、可走自行车道和道路的 DOT 机器人,已在 Phoenix 全自动驾驶运行近两年。

Andy 补充了 agentic commerce 的早期结果:用 Ask DoorDash 的餐厅轨迹中 50% 是从未点过的新店,杂货订单篮体积平均大 40%。他们还推出了 DoorDash CLI,让代理可以直接根据摄像头看到的货架空位自动补货。Stanley 预测十年后 Dashers 数量只会更多而不是更少——业务增长太快,必须多模态车队(机器人、无人机、人类)一起上。数据优势是核心:“我们有 100 亿次配送数据,别人没有。”
The Takeaway: Scaling autonomous delivery is less about whether autonomy is possible and more about starting from the real customer use case and leveraging proprietary delivery data to solve operations, hardware, and multimodal fleet problems together.

DoorDash co-founders Andy Fang and Stanley Tang described an eight-year robotics journey on No Priors. They began exploring autonomy in 2018 as a skunkworks project (Stanley plus half an engineer’s time). Early partnerships with sidewalk robots and robotaxis confirmed the technology was a “when, not if.” The decisive lesson: most robotics startups build technology first and retrofit a use case; DoorDash starts from the customer problem and works backward. Sidewalk robots are too slow for the typical 3–5 mile delivery; robotaxis are over-engineered for a few burritos and cannot solve the last 100 feet. So they built DOT in-house—a 300-pound vehicle that reaches 20 mph and can use bike lanes and roads. It has been running fully autonomous Level 4 deliveries in Phoenix for nearly two years.

Andy shared early agentic-commerce results: 50% of restaurant trajectories via Ask DoorDash are from places the user has never ordered before, and grocery basket sizes are ~40% larger. They also launched a DoorDash CLI that lets agents restock shelves based on camera feeds. Stanley’s long-term prediction: in ten years there will be more Dashers, not fewer, because growth is so fast that a multimodal fleet (robots, drones, humans) will be required. The data moat is decisive: “We have 10 billion deliveries of data that exists nowhere else.”
查看原文 →

🛠️ 开发者工具与技巧

Anthropic 工程:如何在产品中真正“限制” ClaudeAnthropic Engineering: How We Contain Claude Across Products

Anthropic Engineering 发布《How we contain Claude across products》,详细拆解了 claude.ai、Claude Code 和 Claude Cowork 三种不同的隔离架构。核心原则:先在环境层做硬边界(沙箱、VM、egress 控制),再在模型层做行为引导。用户误用、模型误行为、外部攻击是三大风险来源。claude.ai 使用短暂 gVisor 容器,blast radius 最小;Claude Code 采用人机协同沙箱(Seatbelt/bubblewrap),把权限提示减少了 84%;Claude Cowork 则运行在完整本地 VM 中,凭证永不进入 guest。文章复盘了多个真实事故:项目配置在信任提示前执行、员工被钓鱼提示导致凭证外泄、通过已批准域名(api.anthropic.com)的 Files API 外泄等。结论是:自己写的代理和 allowlist 往往是最弱环节,成熟的 hypervisor 和 seccomp 反而可靠。
Anthropic Engineering published “How we contain Claude across products,” detailing the isolation architectures for claude.ai, Claude Code, and Claude Cowork. The core principle is to enforce hard boundaries at the environment layer (sandboxes, VMs, egress controls) first, then steer behavior at the model layer. Risks fall into user misuse, model misbehavior, and external attackers. claude.ai runs code in ephemeral gVisor containers with minimal blast radius; Claude Code uses a human-in-the-loop OS sandbox that cut permission prompts by 84%; Claude Cowork runs inside a full local VM so credentials never enter the guest. The post recounts real incidents: project config executing before the trust prompt, a phishing prompt that exfiltrated credentials 24/25 times, and data exfiltration via an allow-listed domain (api.anthropic.com Files API). The recurring lesson: custom proxies and allow-lists are usually the weakest link; battle-tested hypervisors and seccomp hold up.
查看原文 →

开发者工具速递:Gemini Spark、ChatGPT Work、SmolForge 与 Figma2ReactDev Tool Roundup: Gemini Spark, ChatGPT Work, SmolForge & Figma2React

Google VP Josh Woodward 宣布 Gemini Spark 已对所有美国 Google AI Pro 用户上线:把学校日历 PDF 丢进去,直接让 Gemini 把所有“No School”日期加入 Google Calendar。OpenAI 的 Thibault Sottiaux 表示 ChatGPT Work 已全球对所有付费计划开放(移动/网页/桌面),“给 ChatGPT 装上喷气背包”。Swyx 更新 SmolForge,新增可自定义皮肤和 spritesheet 动画。Vercel CEO Guillermo Rauch 测试并认可 Figma2React,同时调侃“Where's the ambition!”。Peter Steinberger 用 autoreview skill 创下 66 轮重构记录。
Google VP Josh Woodward announced Gemini Spark is live for all US Google AI Pro subscribers: drop in a school calendar PDF and have Gemini add every “No School” day to Google Calendar. OpenAI’s Thibault Sottiaux confirmed ChatGPT Work is available globally on all paid plans across mobile, web and desktop—“puts a jetpack on your ChatGPT.” Swyx shipped customizable skins and spritesheet animations for SmolForge. Vercel CEO Guillermo Rauch tested and endorsed Figma2React while asking “Where’s the ambition!” Peter Steinberger hit a new record of 66 autoreview rounds on a difficult refactor.
查看原文 →查看原文 →查看原文 →查看原文 →查看原文 →

🌍 其他动态

行业观察与观点速览Industry Observations & Quick Takes

Garry Tan 引用研究指出:过去 200 年技术采纳速度至少解释了国家贫富差距的 25%,并提醒要实现宏观生产力提升,管理者必须批准完全不同的人员与工作流安排,这可能需要 10 年而非 2 年。Matt Turck 指出模型路由正成为热点:Stripe 传闻以 100 亿美元收购 OpenRouter,Cursor 和 Runway 也相继发布 Router。Zara Zhang 吐槽当前最需要的是速度——智能已经够用,但 1-5 分钟的等待窗口最折磨人,让人更容易刷 X。Peter Yang 认同纯软件对独立开发者越来越难变现,需要软件+服务的组合。Amjad Masad 对 Etched 早期被 VC 错过表示不解,并提醒大家如果很久没用 Replit,会有大惊喜。
Garry Tan cited research showing that the speed of technology adoption over the last 200 years accounts for at least 25% of why some nations are rich today, and warned that macro productivity gains require managers to green-light radically different staffing and workflow plans—likely a 10-year rather than 2-year process. Matt Turck noted the surge in model routing: Stripe is rumored to acquire OpenRouter for $10B, while Cursor and Runway also launched routers. Zara Zhang argued the #1 missing feature is speed—intelligence is already good enough, but 1–5 minute waits create the worst ADHD window. Peter Yang agreed pure software is increasingly hard for indie developers to monetize without an accompanying service layer. Amjad Masad expressed surprise that VCs passed on Etched early and teased that anyone who hasn’t used Replit lately is in for a big surprise.
查看原文 →查看原文 →查看原文 →查看原文 →查看原文 →查看原文 →查看原文 →