← 2026-10-08

Daily Edition

2026-10-09

2026-10-10 →

AI Builders 日报 — 10月8日

追踪 AI 领域真正在做事的人,而不是空谈者。

今日思考

今天最值得关注的信号有两个。第一,OpenAI 上周发布的数学证明文档正在持续发酵——不仅 Thomas Breeden 在转发,Yann LeCun 也在用法语回应说"自动化证明时代已开启,人类数学家的工作重心将转向新概念和猜想",而 Garry Tan 也引用了 Ethan Mollick 的判断:"数学家们并没有要求这项工作被做出来"。这个争议的实质不是技术,而是在 AI 能证明一切之后,人类数学家的存在意义是什么。第二,Paul Graham 关于"亚马逊封禁 AI 购买代理是 17 年来最大的创业机会"这一判断值得深思——当平台开始替用户做选择时,绕过它的需求就会催生新物种。


产品与发布

Every 推出个人 AI 评测平台 Checks

Every(Dan Shipper 创立的 AI 产品公司)推出了一款名为 Checks 的个人评测平台,帮助任何人衡量新模型在真实工作中的表现。Dan 在推特上密集发布招聘启事,团队包括负责 Evals 的 Mike Taylor、负责 Platform 的 Big Willie Style 和设计师 Tyler Nishida,后者的工作还曾被 Anthropic 在新模型发布时点名推荐。Checks 的核心理念是:与其依赖通用 benchmark,不如让每个人建自己的评测集,量化自己的 vibe check。faviconx.com

OpenAI 面向 GPT-6.1 Sol 推出 Ultrafast 模式

OpenAI 宣布 GPT-6.1 Sol 在 API、Codex 和 ChatGPT Work 中全面推出 Ultrafast 模式,速度最高达 Sol Standard 的 8 倍,同时保持 Near-Astra 的智能水平。这是继 GPT-5 发布之后又一次在推理速度上的重大突破。faviconx.com

Garry Tan 开源完整 Claude Code 配置 gstack

YC CEO Garry Tan 将自己整套 Claude Code 配置完整开源,发布在 GitHub(garrytan/gstack)。被誉为"最好的 GitHub 仓库之一",展示了如何系统化配置 AI 编程环境。faviconx.com


观点与判断

Garry Tan(Y Combinator CEO)

  • 亚马逊封禁 AI 代理是 17 年来最大创业机会 Paul Graham 指出,亚马逊禁止 AI 代理替用户购物,是"自亚马逊成立以来创业公司创建竞争对手的第一个机会"。Garry 转发并表示认同:人们会想要 AI 代理帮他们买东西,而这不可能用亚马逊自己的代理来实现。faviconx.com

  • Agent 集群将带来人类认知的全面突破 Garry 在转发 Deedy 关于 OpenAI 数学发布的推文时评论:Agent 集群是真实的,它们将把人类对世界的理解提升到前所未有的高度。他引用 Opus 的判断,称 OpenAI 的数学发布是"有史以来最具影响力的数学发布"。faviconx.com

  • "数学家们并没有要求这项工作被做出来" Garry 引用 Ethan Mollick 的观察,称这份 OpenAI 的数学文档将成为大学课堂的必读材料,因为它在几段话里浓缩了大量的历史意义。faviconx.com

Peter Yang(AI 投资人/创业者)

  • AI 已经解决图像、音乐和视频,游戏即将被解决 Peter 认为 AI 已经分别在图像、音乐和视频领域取得了实质性突破,游戏是下一个被攻克的领域。faviconx.com

  • Suno 将超越 Spotify Peter 断言 AI 音乐生成工具 Suno 的质量已经极高,很可能在未来某个时刻超越 Spotify 的市场份额。faviconx.com

Dan Shipper(Every 联合创始人)

  • 软件即内容 Dan 认为,当你能用写一篇博客文章的时间发布一个新的实验产品时,写作和编程之间的界限就开始模糊了。对于有分发渠道的作家来说,这是一个难以置信的世界。faviconx.com

  • Albert Wenger:证明、药物、新材料都解决之后,真正的难题是政治、哲学和心理的 Dan 转发 Albert Wenger 的观点:当所有形式的技术问题都被解决之后,人类最深层的挑战依然是政治、哲学和心理层面的。faviconx.com

Yann LeCun(Meta 首席 AI 科学家)

  • AI 数学自动化将开启一个证明形式被大量自动化、人类聚焦于新概念和猜想的新时代 LeCun 反驳"AI 是数学之害"的论调,认为自动化证明的普及类似于"船的发明削弱了游泳的重要性,但让人类发现了新大陆"。faviconx.com

  • Mistral 面对批评应该被支持而非嘲讽 LeCun 对 Mistral AI 表示支持,认为在资源远不及超大规模云服务商的情况下,能以这样的信念建造前沿模型值得敬佩,"Arena 里建造是艰难的"。faviconx.com

  • 大型 AI 公司不应该对科学家施加 6-12 个月的"花园假"并拒绝资助学术教育项目 LeCun 转发了 Nando de Freitas 的批评,称年收入十亿和万亿的 AI 公司在道德上有义务支持学术研究,而不是限制人才流动或忽视教育投资。faviconx.com

Allie Miller(AI 领域增长专家)

  • 判断 AI 使用价值的真正标准:当所有人都用 AI 时,你的工作是否依然有价值 Allie 提出了一个关键框架:如果你的 AI 使用价值建立在"别人不用 AI"的前提上,那么你的 leverage 很有限。更好的测试是:当所有人都用 AI 时,你的用法是否依然创造价值。这引出了她对"AI 使用场景的纳什均衡"的思考——哪些用例在所有人适应之后仍然值得。faviconx.com

  • Meta 的 GTM 策略:在个人主页上加一个图标 Meta 正在重复其最成功的增长策略——在 Instagram 个人主页上添加新功能图标。Reels 和 Threads 都验证了这条路,预计下周会有数百万新用户。faviconx.com

  • 别忽视 EEO(邮件引擎优化) Allie 指出,太多人担心基于网页的 GEO(生成引擎优化),却忘了利用 EEO(Email Engine Optimization)——她在持续给 AI 发送附带笔记的邮件。faviconx.com

swyx(Latent Space 联合创始人)

  • 11月3日 SF Forge 大会:领域专属的everything swyx 将在 Forge 舞台上深入解析数据、模型、评测和芯片的领域专属趋势。这场活动汇聚了最认真的 AI 团队,探讨为什么未来属于 domain-specific everything。faviconx.com

技术动态

Yann LeCun(Meta 首席 AI 科学家)

  • RoboJEPA:机器人世界模型首次证明存在 scaling laws Meta 和 Mila 联合发布了 RoboJEPA,一个 8B 参数的 JEPA 模型,在 15,000 小时的机器人视频数据上训练,首次证明机器人世界模型存在 scaling laws。模型在 V-JEPA 2.1 的特征空间中想象未来,然后规划到单一目标图像。在真实 Franka 机械臂上无需任务微调,8B 模型抓取成功率达 67%,远超 π0.5 的 5%。更重要的是,研究者能够用 22M 到 2B 的模型拟合 scaling law 并准确预测 4B 和 8B 的结果——这意味着可以通过计算预测来指导研发投入,而不是盲目试错。faviconx.com

X / Twitter

38
garrytan
garrytan @garrytan
Retweeted
Garry's List Garry's List
California is defined by the families who give up everything to come and build a better life here.
Because here the best minds are funded, not hunted. Families like Forrest’s have seen what the alternative is.
We’re building a list of people fighting to keep it that way. You can show your support with a public donation, no matter how small, and by adding your name to our growing community.
🔗 https://secure.actblue.com/donate/californiadream
steipete
steipete @steipete
Retweeted
Joma Tech Joma Tech
Introducing ChatGPT for Dishwashing
GPT-6 achieves state-of-the-art results on dishwashing benchmarks for long-running jobs and difficult stains.
The Navier–Stokes Millennium Prize Problem is a $1 million challenge about whether the equations describing fluid motion can break down. In this video, I accidentally found a counter-example while trying to optimize my dishwasher. But in reality, OpenAI used 10,000 AI agents to find a proof in 88 hours: https://openai.com/index/navier-stokes-solution/
Parody of the original GPT-6 Astra ad
@OpenAI #ChatGPT_Partner #chatgpt
Credits:
Written and Directed by Joma
Co-Writer - Henry Connor Coan IV @coan_henry
Production by Dauntlus Studios @dauntlus.studios
Producer/Co-Director - Marcus Liew @dauntlus
Camera Operator - Amos Lee @amoselijahlee
Art Director - Alethea Soo @aletheasoo
Art Assistant - Alicia Lim @aliciatkl_
Assistant Editor - GPT-6 Astra
Music: Tensions Run High by Soundridemusic
https://www.youtube.com/watch?v=Ly9H63SLJJo
steipete
steipete @steipete
Retweeted
McBain McBain
Over the last few months, we've added a ton of new features and improved a lot of quality-of-life issues
This video covers a few of those and shares a little bit about what we're going to be doing over the coming months
0:11 New models
0:17 Multiplayer
0:44 Multiplayer in action
1:20 Interactive dashboards
1:57 Memory & skills
2:08 Meetings & voice
2:20 Easier Mac setup
2:32 What's next
jeremyphoward
jeremyphoward @jeremyphoward
Jasmine says that she accidentally accessed an executive's email that OpenAI had explicitly given her access to, even although she had explicitly requested that OpenAI remove her access, and IT had failed to do so.

Which implies Jasmine was fired for IT's failure?
Jasmine Wang: OpenAI fired me last week, along with two of my safety colleagues. I was given one reason: that I accessed an executive's email. I want to say this plainly, because too many of OpenAI’s history is smoke and mirrors when people disappear:
steipete
steipete @steipete
WE GOT IT! .claw incoming!
David Rodecker: .Claw cometh!

The winning applicant for .claw top level domain in ICANN’s 2026 round: the OpenClaw Foundation. Huge step in giving our agents a natural home presence.

Congratulations to @davemorin @steipete and the @openclaw team. Looking forward to great things crawling 🦞
swyx
swyx @swyx
going rate for this role is between 5-50m comp package btw at the frontier agent labs
adel 🌟: literal hardest role to hire for rn and every ai startup wants someone

> chronically online (knows trends)
> has taste & can create (new media)
> understands ai & technical
> can execute with quality (operator)

such a small niche of folks who can do this
ylecun
ylecun @ylecun
Retweeted
Morgan J. Freeman Morgan J. Freeman
Goodnight 🌙
swyx
swyx @swyx
Re recap of last time this discourse happened for those who are insufficiently online https://dx.tips/cant-hire
danshipper
danshipper @danshipper
Retweeted
Mike Taylor Mike Taylor
DSPyUI took me a whole month to build in the summer 2024. I just rebuilt and tested it in the latest DSPy/Gradio versions with one prompt with $50 of Fable credits. This is a project that went viral and got me 73k views two years ago and now it's trivial, anyone can do it.
Mike Taylor: Made myself a no code interface for optimizing prompts with DSPy. Thought I'd open source it.
rauchg
rauchg @rauchg
Few things more enjoyable in work and life than discovering new talent. Seeing greatness in people, sometimes even before they fully see it themselves.
garrytan
garrytan @garrytan
I refuse to take advice about software from a guy who doesn’t believe in mustache grooming to this extent
Louis Anslow: TikTok blowing up telling people not to use vibe coded photoshop alternative because it is AI slop

jeremyphoward
jeremyphoward @jeremyphoward
Narrator: he was not, in fact, allowed to say “I’m quite unhappy with much of what OpenAI does.”
Tomek Korbak: I’m quite unhappy with much of what OpenAI does.

I am very happy that I’m allowed to say “I’m quite unhappy with much of what OpenAI does.”
petergyang
petergyang @petergyang
I don't understand how Grok @bot's new email feature works. I just tagged it on a thread and it doesn't seem to reply? What are ppl using it for?
ylecun
ylecun @ylecun
Retweeted
Ian L Richardson Ian L Richardson
Let's be clear, the free world has nothing to learn about western values & civilisation from MAGA USA. Most of us can see what Vance & Rubio really are, as they strut around Europe telling us where we went wrong. They should clean up their own filthy & stinking yard, a country of botched executions & now a return to public executions. A cesspit where milliions of its citizens can't afford decent health care & thousands die each year due to archaic gun laws. It's a country that routinely breaks international law. Worst of all, it's a nation in thrall to a corrupt & criminal authoritarian leader who is all that Christianity, liberal democacy, the Renaissance & the Enlightenment were not & are not. MAGA has only one thing to teach us; the urgent need to avoid sliding back into barbarism. Thank you for your attention to this matter.
ylecun
ylecun @ylecun
Retweeted
Yann LeCun Yann LeCun
Re June 1986. It was my first conference in the US and my first time visiting MIT.
Had a poster on multilayer nets.
I met Marvin Minsky for the first time at the reception hosted by Thinking Machines Inc (the original one that built the Connection Machine). He was surprised that you could train a neural net to compute the product of two binary numbers (I had done the experiment).
I was on my way to the first Connectionist Summer School at CMU.
The Bell Labs folks, who I met the year before, heard I was in the US and invited me for a talk in Holmdel on my way back to Paris.
garrytan
garrytan @garrytan
Logical: in the future ICs with agents will be more productive and create better outcomes than equivalent people managers from prior eras
Lenny Rachitsky: “If we truly believe that AI is changing how we work, then ICs should get paid more than managers, because an IC can drive more impact than a manager could in the pre-AI system.” — @ElenaVerna

garrytan
garrytan @garrytan
Retweeted
Suhail Suhail
It is time to prepare for a world where there are tasks so complex and difficult to solve, humans will no longer know how they are solved.
drfeifei
drfeifei @drfeifei
Every one of my mathematician friend sees beauty in the numbers, shapes and the natural phenomena they work with. Who defines beauty in the age of machines?🤔
garrytan
garrytan @garrytan
Retweeted
Rui Ma Rui Ma
I was talking with a friend today about how tough the job market is for tech, and how even if you do have a job, it's so unstable because things are changing so fast.
I told them that even though I technically have my own business and I'm supposed to have "visibility" into how I'm doing, do I really? No. I just assume the business will change every few months and I don't get too fussed when it does, and trust me it has been changing quite consistently
This is why it's obvious to me that, more than a set of skills or a knowledge base, you really have to teach your kids how to adapt to new situations, quickly figure out opportunities, and take risks
garrytan
garrytan @garrytan
Retweeted
geoff geoff
🗞️ if your team is too busy doing their 'normal job' to experiment with AI, you're preparing them to be replaced
https://ghuntley.com/replaced/
In this post, I guess this's the last time I'll say it. AI use is no longer optional if you wish to be employed. What follows is advice for current employees and hiring managers on how to hire in 2026 to source AI-first candidates.
ylecun
ylecun @ylecun
Retweeted
ELLIS ELLIS
Sessions for the ELLIS UnConference are online! Kicking off with two sessions on World Models 🌍, including a 30-min tutorial by @ylecun!
🔎 Sessions: https://ellis-unconference.github.io/#sessions
❗ Register: https://ellis-unconference.github.io/#registration
Heading to @EurIPSConf in Paris? Join us a day early on Dec 8!
danshipper
danshipper @danshipper
Retweeted
Paridhi Agarwal Paridhi Agarwal
Over the past few months i've been having way too much fun building the @every agent - a coworker that lives in your slack and your whole team can delegate complex tasks to it.
wrote a piece on how it works and why we decided to build it using @claudeai managed agents.
https://every.to/source-code/why-we-handed-our-agent-s-infrastructure-to-anthropic?utm_cta_source=home_main_a_3
garrytan
garrytan @garrytan
The rate of progress in the labs is accelerating

The rate of adoption of that progress in society is not

It’s going to go both very fast on one end and it will feel slower than ever to those who understand AI because society has natural brakes

Fast takeoff with slow uptake
carried_no_interest: @AlexFreitasAI Far far far beyond that

Most of them are fully convinced the labs are on a dominant 15-20 year cycle

And that every single thing that can be farmed for tokens (or is a sensor as they call it) will be

And that AI inevitably will run entire companies (ostensibly) in the near
garrytan
garrytan @garrytan
Kind of an interesting idea: explicitly hire the best AI-enabled workers at full salary but only 20 hours a week.
Forbes: Jeff Bezos Claims AI Could Allow For 3-Day Workweeks And One-Income Families
https://go.forbes.com/bKpnur

garrytan
garrytan @garrytan
Education bureaucrats who say crazy things like this fool need to be removed from office
Dale Chu: Imagine running for California's top K–12 education job and arguing that teaching kids elementary school math is the responsibility of college professors.

That's quite the campaign pitch.

danshipper
danshipper @danshipper
Retweeted
Letterland LIVE Letterland LIVE
Is @every a media company or a software company?
Yes.
Their latest drop is the Every Agent, an AI coworker in Slack, bundled with the whole Every archive.
Shout out @danshipper and team for shipping cool stuff all the time. We're excited to try it out and report back...
danshipper
danshipper @danshipper
Working With Agents in Slack https://x.com/i/broadcasts/1qGoNYyeYkbKv
danshipper
danshipper @danshipper
Retweeted
Mike Taylor Mike Taylor
I finally have a writing skill I'm happy with that writes like me, and I was able to write 4 posts this week (normally I would average 2 per month). The problem is our editors can only really handle publishing 2 per month from me and AI editing isn't working yet so we're stuck.
garrytan
garrytan @garrytan
Retweeted
Elena Elena
pat gelsinger (@PGelsinger) designed chips at intel in a period when doing so could involve writing your own hardware description language, which he did for the 486. his team also built a compiler and tools for automating placement and routing. “I wrote my own language,” he says in this pod with @RaghuRaghuram and @appenz. it's a good reminder of how many things had to be invented alongside the processors themselves, and why the prospect of AI taking over more of that work is so exciting.
pat then asks us to imagine designing an excellent AI accelerator in three months. getting it fabricated, packaged, and into a working rack takes roughly nine more in his example (the relevant unit is now the rack: “nothing's a chip anymore, it's a rack”). so there's a period in which your very clever design has to sit and wait for its physical existence, while the models whose workloads you've designed it for can keep evolving. his description of graphcore is that it “wasn't a bad design, but the world moved on.” i think about how often people announce a new model or a new way of using one, and nine months seems like a very long commitment to your current understanding of what the hardware ought to do.
guido's argument is that agents should make it easier to support a greater variety of architectures, because writing all the software to use an unfamiliar chip has historically been one of the barriers. pat's answer is that somebody still has to pay for manufacturing and for a place to put all these chips. he expects consolidation, even with better software tools. if you're a customer committing to a data center and its electricity supply, you're making a rather different bet than the engineer who's just got a new design to work.
which leaves a lot for hardware people to do. pat is especially dissatisfied with hbm (high-bandwidth memory), pointing to problems with density, bandwidth, and heat; he calls it “a hideous memory,” and says he's just funded a new memory company that's still in stealth. he wants better memory closer to compute, optical connections between systems, and less electricity lost in conversion on its way to the chip. he also seems very happy to talk about cooling. “engineers are becoming plumbers,” is how he puts it.
if we're excited about being able to design more things with AI, we should be excited about the work that allows us to manufacture and operate them, too. pat has already lived through a period of having to invent quite a lot of the surrounding machinery. it makes sense that he'd see another one as an opportunity.
a16z: Former Intel CEO Pat Gelsinger with a16z's Raghu Raghuram and Guido Appenzeller on the next wave of semiconductor innovation and the physical constraints shaping the AI buildout:
Drawing on his experience designing Intel's 386 and 486 processors, Pat explains how AI could
swyx
swyx @swyx
Retweeted
Diogo Almeida Diogo Almeida
launched a model
accidentally served trillions of tokens a day
29.4% of the fortune 500 showed up
raised a really big series A from @a16z
it's been 3 weeks
we would like to sleep now
danshipper
danshipper @danshipper
Retweeted
Mike Taylor Mike Taylor
We're experiencing Accelerando at @every – the tempo is getting progressively faster as we go deeper on adopting AI.
I finally dialed things in with Claude and ran 8 new experiments and wrote 4 posts in one week, which would normally be a whole month's work. None of it is live yet because realistically we only have space for 2-4 things from me per month in the newsletter.
@kieranklaassen had to go around the normal clogged up release process and put things live on a new frontier page. He asked in today's meeting "what happens when we can publish 100 apps in one week?".
Nobody has a good answer for that yet. We're not prepared for what happens when our output is 100x.
amasad
amasad @amasad
Some communities are excited by AI’s impact on their field. Others are petrified. What’s the deciding factor(s)?
alliekmiller
alliekmiller @alliekmiller
His Dot caught a mistake before he did.

My dot has already caught scheduling conflicts, missing meeting information, and important billing notices for me.

As we move more into high-agency proactive AI, business leaders need to make sure their teams are not solely relying on AI to catch conflicts or errors. Honestly, I can already see myself falling into the trap a bit.
danshipper
danshipper @danshipper
Retweeted
Every 📧 Every 📧
We don’t condone stealing, but you should definitely steal your colleague’s skills.
@bran_don_gell turned a motion-design process into a skill. @beckyisj thought it was cool and used it to make this video.
Install Every in @SlackHQ: https://every.to/agent?utm_source=x&utm_campaign=every-agent-launch&utm_content=every-261009-skill-steal
petergyang
petergyang @petergyang
Pro @bot tip:

If you want your Grok Bot as a chief of staff it doesn't actually make sense to register its email as your (your-name) because it's weird to copy in yourself.

e.g., "Let me copy in Peter to find a time for us" doesn't make alot of sense.

So I suggest picking another name for your Grok Bot email - whatever you want your chief of staff to be called.
Grok Bot: Grok Bot now has its own email.

Bot can use it to sign up for services, contact businesses for you, or schedule time with someone.

danshipper
danshipper @danshipper
Retweeted
Mike Taylor Mike Taylor
"Don't freak out, I cloned you" - the story of how I cloned @NataliaZarina, @kdaigle, and @danshipper.
https://every.to/also-true-for-humans/you-already-signed-this-off-in-the-simulation?gift=hLQGJ0b4jVRfTxQ7Sn1JMAe5jDxPVEOp
openai
openai @openai
Retweeted
ChatGPT ChatGPT
Fresh updates for dots.
First up: you can now create your dot straight from your phone in the ChatGPT app on iOS and Android.
danshipper
danshipper @danshipper
Retweeted
Lenny Rachitsky Lenny Rachitsky
Tired: Two-pizza teams
Wired: Two-slice teams
@danshipper on the best team structure in the AI-era

YouTube

3
No Priors: AI, Machine Learning, Tech, & Startups

Beam: The Great American Open Model with ReflectionAI Co-Founder and CEO Misha Laskin

Watch Video
The MAD Podcast with Matt Turck

AI writes 60% of open-source databases? #ai #podcast

Watch Video
No Priors: AI, Machine Learning, Tech, & Startups

Beam: The Great American Open Model with ReflectionAI Co-Founder and CEO Misha Laskin

Watch Video