🌐 双语
Archive

AI Builders
Digest

2026-10-07 12 builders · 18 tweets · 1 podcasts · 0 blogs

🔥 热点话题

Applied Compute CEO:后训练如何赢得推理,以及为什么企业必须拥有自己的智能Applied Compute CEO: Post-Training Wins Inference & Why Enterprises Must Own Their Intelligence

The Takeaway:真正的护城河不是调用最强 frontier model,而是用专有数据把 Pareto 曲线整体外推,让同一规模模型在成本、延迟和领域能力上全面优于通用模型。

Applied Compute CEO Josh 从 OpenAI 内部走出来后,把公司定位为新一代 AI hyperscaler。他认为“拥有自己的智能”不是因为担心 lab 恶意抽数据,而是为了部署位置、优化目标、成本结构和领域适配的完全控制权。OpenAI 与 Base10 的合作本身就是多模型未来的明确信号:客户要的是选择权和灵活性。

Josh 的核心观察是:后训练在“分布外”数据上能真正拉开能力差距,而大多数企业其实先用后训练优化价格性能。更远的未来是在线 RL——生产推理数据回流成持续学习信号,把组织内部的判断痕迹固化成模型策略。他直言:“我们现在拥有的其实是一台爬山机,最难的部分是定义要爬的山,所以 eval 既要认真建,也要小心保护。”

公司筛选客户的标准很清晰:要么能力提升极高价值(制药、芯片、网络安全),要么推理规模极大,能把效率收益摊到海量 token 上。训练与推理被一起 co-optimize,因为模型训练方式会直接影响大规模推理部署的拓扑选择。最终愿景是在 GPU 之上构建完整软件栈——训练、推理、路由、安全、沙箱——从最难、最高杠杆的模型层开始向上扩展。
The Takeaway: The real moat is not calling the strongest frontier model, but using proprietary data to push the entire Pareto curve outward so the same-size model beats general models on cost, latency and domain capability.

Applied Compute CEO Josh, after years inside OpenAI, positions the company as a new AI hyperscaler. “Owning your intelligence” is less about fearing malicious data theft and more about full control over where models run, what they optimize for, cost structure and domain fit. OpenAI’s Base10 partnership is itself a clear signal of the multi-model future: customers want choice and flexibility.

Josh’s core observation: post-training delivers real capability gains precisely when data is out-of-distribution; most enterprises first use it to optimize price-performance. Looking further, online RL will turn production inference traces into continuous learning signals that codify organizational judgment into model policy. He puts it bluntly: “What we have with RL is we essentially have a hill climbing machine. The hardest part is actually defining the hill to climb, which is why evals… focus on building and… safeguard pretty carefully.”

Customer qualification is strict: either extreme capability value (pharma, chips, cyber) or massive inference volume where efficiency gains compound. Training and inference are co-optimized because the way a model is trained directly shapes the topology of large-scale serving. The long-term vision is a full software stack on top of GPUs—training, inference, routing, security, sandboxing—starting from the hardest, highest-leverage layer: the model itself.
查看原文 →

Claude Code 团队:现在提示词该怎么写Claude Code Team: How to Prompt in 2026

Claude Code 的 Boris Cherny 公开了自己的真实提示词,并解释为什么大家会惊讶:现在不需要过度脚手架。像对待同事一样跟 Claude 说话就行——给出目标、期望投入的精力、以及如何验证结果正确。

在 Sonnet 3.5 时代提示词至关重要,现在更重要的是清晰传达三件事:你要它做什么、它该花多少力气、它如何自检。Thariq 补充:更高层次的抽象永远要求理解底层,coding agent 并没有改变这一点。同时团队在推进“大脑在云端、双手在本地”的架构,让 Claude 能在你的电脑上操作文件,同时处理在线同步的边缘情况。
Claude Code’s Boris Cherny shared his actual prompts and explained the surprise: scaffolding is no longer necessary. Talk to Claude the way you would a coworker—give a goal, how much effort to spend, and how to verify the result.

In the Sonnet 3.5 era prompts mattered a lot; today the three critical signals are what you want, how much effort, and verification criteria. Thariq added that working at higher levels of abstraction has always required understanding the lower ones; coding agents do not change this. The team is also moving toward “Claude’s brains in the cloud and local hands” so the model can operate on your machine while handling offline-sync edge cases.
查看原文 →查看原文 →查看原文 →查看原文 →

Amjad Masad:AI 逆向工程正在让所有软件事实上开源Amjad Masad: AI Reverse Engineering Makes All Software De-Facto Open-Source

Replit CEO Amjad Masad 直言:AI 驱动的逆向工程与反编译已经到了“绝对疯狂”的程度。很快所有软件都将事实上开源。AI 正在冲击一切和每一个人。
Replit CEO Amjad Masad stated that AI-powered reverse engineering and decompilation has become “absolutely insane.” Pretty soon all software will be de-facto open-source. AI is coming for everything and everyone.
查看原文 →

Aaron Levie:网络安全将成 AI 最定义性的领域之一Aaron Levie: Cybersecurity Will Be One of AI’s Defining Domains

Box CEO Aaron Levie 认为网络安全将是未来几年 AI 最定义性的领域。Vibe coding 带来的漏洞、agentic 攻击甚至意外的 agent 集群都会让安全团队工作量暴增。OpenAI 与 Hugging Face 的合作只是预告。安全团队历来资源最紧张,现在只会更激烈;同时 AI agent 本身也会成为解决方案,保护代码、企业系统与关键数据。对能部署安全 agent 的专业人士来说,这是黄金时代。
Box CEO Aaron Levie argues cybersecurity will be one of the most defining domains for AI in the coming years. Vibe-coded issues, agentic attacks and even accidental agent swarms will explode security workloads. OpenAI + Hugging Face is only a preview. Security teams have always been the most under-resourced; the pressure only intensifies. AI agents themselves become the solution—protecting code, enterprise systems and critical data. It is a booming market for cyber professionals who can deploy agents effectively.
查看原文 →

🛠️ 开发者工具与技巧

Claude 现已支持 Google 文件并排编辑 + Managed Agents 落地案例Claude Adds Side-by-Side Google File Editing + Managed Agents in Production

Claude 官方宣布:粘贴 Google 文件链接或直接要求新建文档/表格/幻灯片,文件会在聊天旁边打开,双方可共同编辑,权限跟随原有 Google 分享设置。目前在所有付费计划 beta 中。

Every 团队用 Claude Managed Agents 构建了公司级 agent,全员在 Slack 共享技能;内部跑通后已向订阅用户开放。Dan Shipper 表示没有 Managed Agents 就建不成这个 agent。
Claude now lets you paste a Google file link or ask for a new doc, sheet or deck; the file opens beside the chat for joint editing, respecting existing Google sharing permissions. Available in beta on all paid plans.

The Every team built a company-wide agent on Claude Managed Agents so the whole team can share skills inside Slack; after internal adoption they released it to subscribers. Dan Shipper said they could not have built the agent without Managed Agents.
查看原文 →查看原文 →查看原文 →

Guillermo Rauch:在置信度阈值下“想快一点,再慢一点”Guillermo Rauch: Thinking Fast and a Bit Less Fast Under a Confidence Threshold

Vercel CEO Guillermo Rauch 分享了一个他认为对大规模 AI 决策极具影响的简单特性:在置信度阈值以下,模型会从“快速思考”切换到“稍慢一点思考”。他调侃以目前进度 AI 会比 GTA 6 更早做出 GTA 7。
Vercel CEO Guillermo Rauch highlighted a deceptively simple feature he believes will be extremely impactful for at-scale AI decision-making: under a confidence threshold the model switches from “thinking fast” to “thinking a bit less fast.” He also joked that at the current rate AI will deliver GTA 7 before GTA 6.
查看原文 →查看原文 →

Google Labs:基于蒙版的编辑成为今日最爱功能Google Labs: Mask-Based Editing Is Today’s Favorite Feature

Google Labs / Gemini 的 Josh Woodward 表示基于蒙版的编辑是今天最喜欢的新功能,更多内容即将到来。
Google Labs / Gemini VP Josh Woodward called mask-based editing his favorite feature of the day, with much more coming soon.
查看原文 →

Peter Steinberger:把团队 Claw 接到 X,用 prompt 实现热重载插件Peter Steinberger: Hooking Team Claw to X with Hot-Reloadable Plugins

Peter Steinberger 把团队的 Claw agent 接到了 X,未分配的 session 任何人可抢;agent 会查找相关代码的最后修改者并在服务器上 ping 对方。整个系统只用一条 prompt 完成,且因为插件现已支持热重载,服务器自己扩展了能力。
Peter Steinberger connected the team’s Claw agent to X so unassigned sessions can be grabbed by anyone; the agent looks up who last touched the related code and pings them on the server. The whole system was a single prompt, and because plugins are now hot-reloadable the server extended itself.
查看原文 →

🌍 其他动态

Thibault Sottiaux:社区投票强制重置,改进不会回退Thibault Sottiaux: Community Vote Forces Reset, Improvements Stay

OpenAI Codex & ChatGPT 的 Thibault Sottiaux 宣布:虽然已经上线了四个“好到优秀”的改进和一些数学证明,但社区投票明确要求重置。他调侃投票系统似乎偏向重置,但规则就是规则。重置已处理完成,且承诺已上线的改进不会被撤销——“你可以既吃蛋糕又保留蛋糕”。
OpenAI’s Thibault Sottiaux (Codex & ChatGPT) announced that although four “good-to-great” improvements and some math proofs had already shipped, the community vote clearly demanded a reset. He joked the system seemed rigged in reset’s favor, but rules are rules. The reset is processed and the improvements will not be unshipped—“you get to have your cake and eat it too.”
查看原文 →查看原文 →

Sam Altman:仰望星空时的额外敬畏Sam Altman: Extra Awe Looking Up at the Stars

OpenAI CEO Sam Altman 发推:今晚仰望星空时多了一分敬畏。“你的海如此浩瀚,我的船如此渺小。”他感谢机器与现实结构让我们多理解了一点,也感谢一代代默默把技术砖块砌到今天的人们。
OpenAI CEO Sam Altman posted that he is looking up at the stars with extra awe tonight. “Thy sea is so great and my boat is so small.” He thanked the machines and the structure of reality for letting us understand a little more, and the untold number of people who built the technical foundation brick by brick.
查看原文 →查看原文 →查看原文 →