🌐 双语
Archive

AI Builders
Digest

2026-08-22 16 builders · 34 tweets · 1 podcasts · 0 blogs

🔥 热点话题

Chess.com CEO:超级智能时代人类技能仍有价值Chess.com CEO: Human Skills Still Matter in the Age of Superintelligence

The Takeaway:即使机器在棋类上早已超越人类,棋类反而比30年前更受欢迎,因为人类本质上想做人类的事——与真人竞争、创造、解决问题。

Chess.com CEO Erik Allebest 从2005年以5.5万美元拍下域名起步,拒绝几乎所有VC,纯靠自有资金和盈利滚雪球,今天拥有超过2.5亿注册用户、约1000万日活、年收入超2亿美元、650人全远程团队。他指出,Deep Blue击败卡斯帕罗夫后很多人预言棋类消亡,但早期引擎让棋变得无聊,神经网络引擎(如Leela Chess Zero)却让棋重新变得激进、非常规且有趣,AI反而帮助人类更快进步:游戏复盘教练、个性化谜题、口袋教练都在加速学习曲线。

他强调没有捷径,只有重复:每天做5道谜题乘以1000天就是5000次练习。关于作弊,Chess.com投入大量资源用统计与机器学习模型检测,保护从休闲到职业赛事的诚信。对AGI/ASI,Allebest偏乐观,认为技术本身无关好坏,关键是文化与分配:希望它帮助解决气候、教育与公平问题,但也担心财富进一步集中。他总结:“人类想做人类的事。”棋类证明了即使机器更强,人们仍渴望习得并欣赏人类技能。

The Takeaway: Even after machines surpassed humans at chess decades ago, the game is more popular than ever because humans fundamentally want to do human things—compete against other humans, create, and solve problems.

Chess.com CEO Erik Allebest bought the domain out of bankruptcy for $55,000 in 2005, ignored nearly every VC who called it uninvestable, and bootstrapped to 250M+ members, ~10M DAU, $200M+ revenue and a 650-person fully remote team. Early engines made chess boring by pushing perfect play; neural-net engines like Leela Chess Zero made it aggressive and exciting again. AI now accelerates human improvement through game review coaches, personalized puzzles and on-demand pocket coaches. There are no shortcuts—only repetitions: five puzzles a day times a thousand days equals real progress. On AGI he is cautiously optimistic: the technology is neutral; the real risk is cultural and distributional. As he puts it, “Humans want to do human stuff.”
The Takeaway: Even after machines surpassed humans at chess decades ago, the game is more popular than ever because humans fundamentally want to do human things—compete against other humans, create, and solve problems.

Chess.com CEO Erik Allebest bought the domain out of bankruptcy for $55,000 in 2005, ignored nearly every VC who called it uninvestable, and bootstrapped to 250M+ members, ~10M DAU, $200M+ revenue and a 650-person fully remote team. Early engines made chess boring by pushing perfect play; neural-net engines like Leela Chess Zero made it aggressive and exciting again. AI now accelerates human improvement through game review coaches, personalized puzzles and on-demand pocket coaches. There are no shortcuts—only repetitions: five puzzles a day times a thousand days equals real progress. On AGI he is cautiously optimistic: the technology is neutral; the real risk is cultural and distributional. As he puts it, “Humans want to do human stuff.”
查看原文 →

OpenAI Codex 限速与 banked reset 更新OpenAI Codex Rate Limits and Banked Reset Update

OpenAI Codex & ChatGPT 团队的 Thibault Sottiaux 更新:部分用户本周缓存命中率下降,导致用量消耗更快;团队正在调查,次日会有进一步说明。同时宣布 banked reset 已对所有 ChatGPT Work 和 Codex 付费用户上线,方便周末使用。
OpenAI Codex & ChatGPT’s Thibault Sottiaux reported that some users saw worse cache hit rates this week, draining usage faster; the team is investigating with an update due the next day. Separately he confirmed the banked reset had landed for all paid ChatGPT Work and Codex users.
查看原文 →查看原文 →查看原文 →

Aaron Levie:AI 进步速度前所未有,应用层机会巨大Aaron Levie: AI Progress Is Unprecedented, Applied Layer Opportunity Is Huge

Box CEO Aaron Levie 指出,当前 AI 在成本、通用能力、速度和深度上的进步速度是科技史上独有的。当智能变得“便宜到无法计量”时,最大机会是推动 AI 在经济中的扩散。这对能抓住模型与竞争尾风的应用型 AI 创业公司是巨大时刻。
Box CEO Aaron Levie notes that the current rate of AI progress—cheaper on like-for-like tasks, more capable, faster, and deeper across domains—is unlike any other period in tech history. As intelligence becomes too cheap to meter, the big opportunity is driving diffusion into the economy. This is a huge moment for applied AI startups that can ride the innovation and competition tailwind.
查看原文 →

💰 创业成功案例

Erik Allebest:20年自举打造2.5亿用户棋类帝国Erik Allebest: 20-Year Bootstrap to a 250M-User Chess Empire

Chess.com CEO Erik Allebest 在斯坦福商学院期间买下域名,几乎所有VC都说“不可投资”,同学也不愿加入。他与朋友两人从零开始,靠前期出售业务的资金和少量借款,2007年上线后很快靠会员收费实现盈利,之后完全按现金速度扩张,从未再考虑融资。直到规模足够大才引入二级市场PE(General Atlantic、CVC),但始终没有需要外部资本来增长。他的建议:听内心想要创造什么,用最少资源证明,找到对的人一起执行,其余都是噪音。
Chess.com CEO Erik Allebest acquired the domain while at Stanford GSB; nearly every VC called it uninvestable and classmates declined to join. He and a friend bootstrapped, funded by prior exits and a small loan, launched in 2007, quickly turned profitable on memberships, and grew purely at the speed of cash—never raising again. Later secondary PE (General Atlantic, CVC) came in only after scale. His advice to founders: listen to what you want to exist, use the minimum resources to prove it, work with great people, and ignore the rest of the noise.
查看原文 →

🛠️ 开发者工具与技巧

Anthropic 内部热用 ELI5 技能:用大图少字解释复杂问题Anthropic’s Popular ELI5 Skill: Explain Complex Topics with Big Pictures and Few Words

Claude Code 的 Thariq 分享 Anthropic 内部大量使用的 ELI5 技能:/eli5 <问题>,要求“像对完全不懂的人解释,用带大图和少字的 HTML artifact”。常用于解释模块、权衡取舍或事故原因。可通过社区 marketplace 安装试用,团队还在讨论是否做成官方插件。
Claude Code’s Thariq shared a skill heavily used inside Anthropic: /eli5 <topic>, which asks the model to “explain like I’m someone who knows nothing… using a HTML artifact with big pictures and few words.” Common for diving into modules, tradeoffs or incidents. Install via the community marketplace; the team is debating making it an official plugin.
查看原文 →查看原文 →查看原文 →

Claude Security:用 Mythos 扫描 GitHub 仓库漏洞并给出修复建议Claude Security: Mythos Scans GitHub Repos for Vulnerabilities with Suggested Fixes

Claude 官方宣布 Claude Security:指向 GitHub 仓库后,Mythos 会跨文件追踪数据流并推理组件交互,返回带 CWE 分类、置信度与严重性评级的发现,以及建议补丁(可直接在 Claude Code 网页端打开)。扫描按标准 token 计费,模型不直接暴露给用户。同时推出 3500 万美元 Defender Advantage Fund 支持开源安全,并扩展 Cyber Verification Program。
Claude announced Claude Security: point it at a GitHub repo and Mythos scans for vulnerabilities by tracing data across files and reasoning about component interactions. Each finding includes CWE category, confidence/severity ratings and a suggested fix that opens in Claude Code on the web. Scans are billed as normal token usage; the model stays behind the scan. Also launching a $35M Defender Advantage Fund for open-source security and expanding the Cyber Verification Program.
查看原文 →查看原文 →查看原文 →

Guillermo Rauch:Vercel 用 is-agentic 循环把评分刷到 100/100Guillermo Rauch: Vercel Ran is-agentic in a Loop Until 100/100

Vercel CEO Guillermo Rauch 分享团队用 is-agentic 对目标站点循环运行直到拿到 100/100 分,过程中暴露并修复了多项差距。同时宣布相关工具现已支持 Grok 与 Codex 订阅,可在沙箱中即时安装测试。Python 团队也在快速推进。
Vercel CEO Guillermo Rauch described running is-agentic in a loop against a target until it scored 100/100, forcing the team to close multiple gaps while keeping criteria high-quality. The same tooling now supports Grok and Codex subscriptions and installs instantly on a sandbox. The Python team is also making rapid progress.
查看原文 →查看原文 →查看原文 →

Madhu Guru:停止把复杂 eval 压成单一分数Madhu Guru: Stop Collapsing Complex Evals into a Single Score

Meta AI 高级总监 Madhu Guru(前 Google Gemini/Veo 负责人)在 eval 系列第5篇中强调:不要把多维度的复杂评估结果压成一个平均分,这会掩盖模型在 frontier 用例上的真实退步(例如简单摘要提升但复杂金融分析下降)。加权分数只是给主观判断披上数学外衣。正确做法是保留按优先级排列的 eval 列表,让能看细节的人深度理解每项结果的成败,再决定系统对用户是更好还是更差。
Meta AI Sr Director Madhu Guru (ex-Google Gemini/Veo) argues in part 5 of his eval series against collapsing multi-dimensional results into a single average score—the “tyranny of the average.” A model can improve on simple summarization while regressing on complex financial analysis; a weighted score merely dresses a judgment call in false mathematical precision. Keep a prioritized ladder of evals, have people who can read the details, and decide based on full understanding of where the system shines or fails for users.
查看原文 →

Swyx:Simulation 是新的 scaling law,Simile 已找到 PMFSwyx: Simulation Is a New Scaling Law — Simile Already Finding PMF

smol.ai 等创始人 Swyx 从最初的半玩笑转向严肃:如果认真对待递归自我改进(RSI),模型自动化越来越大块的 ML 研究与 AI 工程后,最后一道障碍就是模拟人类与人类反馈。Smallville 当年零商业应用,但 Karpathy、李飞飞等 backing 的团队方向正确;Simile 如今已在 Fortune 100 客户中找到 PMF。他表示“从未如此高兴自己错得这么离谱”。
smol.ai founder Swyx moved from half-shitposting to very serious: if you take recursive self-improvement seriously, after models automate ever-larger parts of ML research and AI engineering the last barrier is simulating humans and human feedback. Smallville had zero commercial applications at the time, yet the direction Karpathy, Fei-Fei Li and others backed was correct; Simile is already finding PMF with Fortune 100s. “I’ve never been so happy to be so wrong.”
查看原文 →

Nikunj:用 Claude Code 解析幼儿园餐单 API 并接入家庭机器人Nikunj: Claude Code Reverse-Engineered Kindergarten Meal API for Home Bot

FPV Ventures 合伙人 Nikunj Kothari 分享实用案例:女儿幼儿园餐单藏在结构混乱的网页里。他让 Claude Code 通过网络请求找到未认证的 API、解析正确格式,再交给现有 Hermes 家庭机器人。现在每天早晨机器人会自动告知早餐和午餐内容,方便准备便当。
FPV Ventures partner Nikunj Kothari used Claude Code to reverse-engineer an unauthenticated API behind his daughter’s kindergarten meal site (messy unstructured data), then fed the structured output to an existing Hermes home bot. Every morning the bot now announces breakfast and lunch so the family can pack accordingly.
查看原文 →

🌍 其他动态

Peter Yang 试用 Instinct:onboarding 优秀但多线程工作流仍不足Peter Yang on Instinct: Smooth Onboarding, Limited Multi-Thread Work

Peter Yang 称赞 Instinct 连接 iMessage、Google Workspace 和 MCP 的 onboarding 非常流畅,且比许多助手更主动建议可立即执行的任务。但所有交互挤在一个线程里,难以并行处理真实工作,因此他仍把主力留给 ChatGPT Work 与 Codex,Instinct 更适合零散杂事。同时他公开批评 Instinct 在未获许可情况下索引并保留邮件且无法删除。
Peter Yang praised Instinct’s smooth onboarding for iMessages, Google Workspace and MCPs, plus its proactive suggestions after connecting tools. However the single-thread UX limits real parallel work, so he continues using ChatGPT Work and Codex for serious tasks and Instinct mainly for chores. He also publicly flagged that Instinct indexes and retains emails without clear deletion rights.
查看原文 →查看原文 →

Peter Steinberger:Agentic AI 峰会演讲“No Doors for Agents”并发布新 skillPeter Steinberger: “No Doors for Agents” Talk + New Skill Drop

OpenClaw 相关的 Peter Steinberger 在伯克利 Agentic AI Summit 发表题为“No Doors for Agents”的演讲,并随后宣布即将发布新 skill。
Peter Steinberger (OpenClaw) spoke at the Agentic AI Summit in Berkeley on “No Doors for Agents” and later teased a new skill drop.
查看原文 →查看原文 →

Zara Zhang:每天出现、持续发布、不怕重复Zara Zhang: Show Up Every Day, Always Be Launching, Repeat Yourself

Builder Zara Zhang 强调三个习惯:每天出现、永远在发布、不要害怕一遍又一遍地重复同一信息。
Builder Zara Zhang distilled her operating system: show up every day, always be launching, and do not be afraid of repeating yourself over and over.
查看原文 →

其他简讯Other Quick Notes

Amjad Masad(Replit CEO)转发并互动多条关于手机赚钱与相关进展的帖子。Garry Tan(YC CEO)转发并评论多条涉及“奢侈信念”与亚裔美国人被抹除的讨论。Aditya Agarwal 分享与 Ramaswamy Sridhar 的完整访谈,并指出 frontier 模型竞赛才刚开始。Dan Shipper、Matt Turck 等有轻量互动帖,无更多实质技术或产品更新。
Replit CEO Amjad Masad amplified posts on making money from a phone. YC CEO Garry Tan quote-tweeted discussions around luxury beliefs and Asian-American erasure. Aditya Agarwal shared a full episode with Ramaswamy Sridhar and noted the frontier model race is only getting started. Dan Shipper and Matt Turck posted lighter engagement content with no additional technical substance.
查看原文 →查看原文 →查看原文 →