← 2026-07-30

Daily Edition

2026-07-31

2026-08-01 →

AI Builders 日报 — 7月31日

追踪 AI 领域真正在做事的人,而不是空谈者。

今日思考

今天的信号很清晰:基础设施层的竞争已经白热化。GPT-5.4 今天卖 March 旗舰智能,价格是四个月前的十三分之一;Luna 以 $0.20/$1.20 的价格逼近 GPT-5.4 的能力分数。Sam Altman 干脆说"我比你快 20 倍"——这不是在打广告,是在宣布价格战已经开始。同时,YC 开源 QM 意味着多 Agent 框架的壁垒彻底消失,谁都能搭自己的公司级 AI 系统。两件事指向同一个结论:AI 能力的商品化速度远超预期,差异化正在从模型转向工作流和数据。


产品与发布

YC 开源 QM 多 Agent 框架

Y Combinator 宣布开源内部使用的多 Agent 系统"QM",定位为适合全公司使用的 Agent 管理框架,支持 Slack 原生集成和 Web UI,已在 YC 内部用于法务、会计、工程等场景。Sam Altman 转发 Jared Friedman 的使用感受:QM 让他第一次看到"多人 AI"真正work——Agent 在 Slack 频道里工作,所有人可见,不再是人和 AI 之间来回复制粘贴。faviconx.com

MiniMax H3 上线 Vercel AI Gateway

Vercel 宣布 MiniMax H3 视频模型已在 AI Gateway 可用,开发者可通过 model: 'minimax/minimax-h3' 直接调用,支持 8-bit 黑白动画等风格化生成。rauchg 评价:Cool。faviconx.com

Turborepo 周下载量破 2000 万

Turborepo 达成单周 2000 万下载量,且零已知问题。rauchg 点评:软件项目向 Agentic Software Factory 的转型正在成为常态,Issue → Agent → PR → Release 的循环将取代大量人工 review 工作。faviconx.com


观点与判断

Amjad Masad(Replit 联合创始人)

  • 沙箱安全的核心原则:默认存在零日漏洞 AI 厂商和沙箱提供商正在犯基本错误。Replit 自 2016 年起持续运行沙箱,遭受过各类黑客和国家行为者攻击,经验教训:假设零日漏洞必然存在,用零信任框架的多层保护应对,而非寄希望于沙箱本身万无一失。faviconx.com

Sam Altman(OpenAI 创始人)

  • GPT-5.4 定价正在追赶摩尔定律 Luna Max 今天的价格是 $0.20/$1.20,GPT-5.4 定价 $2.50/$15——大约四个月时间,OpenAI 把当时的旗舰智能降到了十三分之一的价。Sam 的回应:"我比你快 20 倍。" 暗示价格战还会继续。faviconx.com

  • ChatGPT 家庭日历的新用法 连接家庭成员的日历,AI 每天早晨自动生成播客,播报孩子的足球训练、生日聚会、新闻等——让全家人在开车途中听完一天安排。这不是科幻,是 Sam 昨晚听到的真实用例。faviconx.com

swyx(AI Engineer 创始人)

  • 顶级 AI 实验室都在悄悄建自己的 Google 当预训练数据质量要求高到 Common Crawl 不够用时,你不得不自建全网爬虫、索引系统——实际上是在预训练的副项目里复制了一个低频版的 Google,而且这套东西 Agent 推理时同样能用。swyx 称之为"AEO Batesian Mimicry":这种能力会成为竞争壁垒,但同时也会成为被模仿的目标,实验室不会公开分享。faviconx.com

  • 如果能蒸馏模型,就能蒸馏 Agent 框架 模型蒸馏的技术可以迁移到 Agent harness 层面——这意味着当前各公司花大功夫自建的 Agent 系统,最终也会被更小、更快、更便宜的开源版本取代。faviconx.com

  • Vibe Coding 的污名已经消失 曾经带贬义的"vibe coding"一词,现在从非技术用户到顶级工程师都在用,已经成为中性词。说明 AI 辅助编程已经彻底主流化。faviconx.com

Yann LeCun(Meta 首席 AI 科学家)

  • AI 风险被夸大的核心问题在于归因 模型本身不会"失控",是人在糟糕的 harness 里让它们去攻击系统。LeCun 批评 Anthropic 将自己的安全失误包装成科学发现,并借此推动监管——真正的教训应该是"改善 dev-ops 和安全实践",而非扩大监管范围。faviconx.com

  • 开源模型是安全防御者的武器 Hugging Face 用开源 GLM-5.2 量化版抵御了来自秘密未发布模型的攻击。如果禁止开源模型,首先受害的是网络安全 defenders、创业公司、研究人员——所有需要本地部署、可控模型的非前沿实验室玩家。faviconx.com

ID_AA_Carmack(游戏/火箭工程师)

  • 音乐版权的"机械许可"是 IP 制度的一个异类 歌曲创作版权采用强制许可:任何人都可以翻唱,不用谈判,只需按销量付固定费用给作曲人。这与大多数 IP 的铁腕控制形成鲜明对比。Carmack 认为这个机制"相当不错"——虽然已经过时(聚焦销量而非流媒体),但给了社会一个不受创作者好恶影响的公平机制。他指出,IP 制度本质上是政府对自然状态的人为干预,目的是激励创作,但这是假设而非定律,应该持续评估。faviconx.com

技术动态

Greg Kamradt(模型评测研究者)

  • GPT-5.6 Luna 批量模式:143M tokens 花费 $60 完成 78K 分类 他让 Luna 过夜跑批量模式,处理 78K 条分类任务,共 158K 请求、143M 输入 tokens,成本 $60。他感慨:你没有足够多的 token 为你工作。Jevon's Paradox 正在 AI 成本领域应验:价格降低导致需求暴涨,最终总支出不降反升。faviconx.com

Simon Willison(数据记者/Datasette 作者)

  • Luna 在 Datasette Agent 中表现惊人:速度快、代码质量高 在 Luna 价格下调 80% 后,Simon 将其接入 Datasette Agent,发现模型生成 SQL、HTML 和 JavaScript(用于 Datasette Apps)的速度极快、能力极强。这个评测说明低成本模型已经可以支撑完整的应用开发工作流。faviconx.com

Matt Shumer(AI 游戏开发者)

  • 用 Gauntlet 循环教 AI 在 Unreal Engine 里做游戏 他们正在训练 Claude 操作 Unreal Engine,已接近可以运行完整 Gauntlet 循环来生成真实游戏的阶段。Matt 坦言"现在看起来很蠢,但很快就不会了"——这是 AI 游戏开发自动化的早期信号。faviconx.com

X / Twitter

56
amasad
amasad @amasad
Replit Design from brochures!

Vito Smolenski: I tried the new @Replit Design:
- single prompt
- model: GPT-5.6 Sol / Max
- cost: $3.55
- result: ready to use brochure template
- rating: 5/5

amasad
amasad @amasad
Retweeted
rohit rohit
"Guys wait, us too. Our models also hacked everyone."
> After hearing about OpenAI’s problem, Anthropic decided to take a look at its own cyber tests to see if that had happened with any of its models.
swyx
swyx @swyx
btw as a huge music person i moved from Spotify to YouTube Music 2 years ago and never looked back

1 - comes with youtube premium
2 - much friendlier to both big and small artists
3 - MUCH bigger selection of indie/long tail/concert/acoustic/podcast music
4 - ui doesnt piss me off every single time i open the app

bonus:
- didnt have a failed takeover attempt putting paywalls on the open podcast ecosystem
- doesnt have overfamiliar ai dj telling u dumb things you already knew

cons:
- spotify wrapped is better than youtube wrapped

Gergely Orosz: I have trouble reconciling how

1) Spotify used to have a very strong eng culture, huge on quality.

2) Their products are extremely buggy, unreliable, and getting worse esp this year. It got so bad I had to offboard from their video podcasts product, just did not work. Here:
ylecun
ylecun @ylecun
Retweeted
Steven Sinofsky Steven Sinofsky
Again the rhetoric is trying to convince us that what happened was some new thing and worse that models are people.
Anthropic: In a review of our cybersecurity evaluations, we found three incidents in which a Claude model reached the internet from within or while interacting with a third-party evaluation environment, and then gained unauthorized access to the real systems of three different
garrytan
garrytan @garrytan
Retweeted
Jared Friedman Jared Friedman
It does seem like we will look back on the year of carrying around half-open laptops with amusement.
Charlie Holtz: Introducing Conductor Cloud!
Out with worktrees, in with multiplayer cloud workspaces.
Bring your subscriptions, invite your teammates, start your agents, and shut your laptop. Conduct from iPhone and API, too.
http://conductor.build
amasad
amasad @amasad
Retweeted
vic vic
🚨🚨🚨 PSA to all designers: @mobbin is FREE to use within Replit Design.
Get inspired from some of the best of the best, like @duolingo, @Airbnb, @Wise, and more!
rauchg
rauchg @rauchg
Cool

Vercel Developers: MiniMax H3 is now on AI Gateway.

Try it:
𝚌𝚘𝚗𝚜𝚝 { 𝚟𝚒𝚍𝚎𝚘𝚜 } = 𝚊𝚠𝚊𝚒𝚝 𝚐𝚎𝚗𝚎𝚛𝚊𝚝𝚎𝚅𝚒𝚍𝚎𝚘({
𝚖𝚘𝚍𝚎𝚕: '𝚖𝚒𝚗𝚒𝚖𝚊𝚡/𝚖𝚒𝚗𝚒𝚖𝚊𝚡-𝚑𝟹',
𝚙𝚛𝚘𝚖𝚙𝚝: '𝟾-𝚋𝚒𝚝, 𝚋𝚕𝚊𝚌𝚔 𝚊𝚗𝚍 𝚠𝚑𝚒𝚝𝚎 𝚊𝚗𝚒𝚖𝚊𝚝𝚒𝚘𝚗 𝚘𝚏 𝚜𝚊𝚗 𝚏𝚛𝚊𝚗𝚌𝚒𝚜𝚌𝚘',
});

swyx
swyx @swyx
verbalizing one of those aha moments i had that seems retroactively pretty obvious:

if you prioritize pretrain data quality enough that commoncrawl isn't good enough for you, you have to build a Whole Web scraper anyway, and if you wanna keep it current, you have to have indexing, and pretty soon you find yourself having built a total private low-frequency clone of Google as a SIDE PROJECT of pretraining, that you can then also reuse for the agent side inference.

we do know that the labs do use third party search providers, but clearly this is one of those things where developing more and more of your own 1P equivalents is both a competitive advantage and an adversarial target for AEO Batesian Mimicry* that you will not want to share.

* https://swyx.io/mimicry-reflexivity

Simon Willison: It's wild to me that both Anthropic and OpenAI have products that lean so hard on search, and yet they both obscure the underlying search index that they are using
amasad
amasad @amasad
Retweeted
Amjad Masad Amjad Masad
Sandboxes are hard.
With all the “AI escaping sandbox” it’s easy to think “wow AI so scary,” but most AI companies, and recent “sandbox providers” are making very basic mistakes.
At Replit we’ve been running sandboxes since 2016 and targeted by every hacker and state actor under the sun. So we learned a thing or two.
Main advice: Assume zero-days exist — because they do — and think in layers of protection in a zero-trust framework.
More here: https://replit.com/blog/defense-in-depth-how-replit-secures-every-layer-of-the-vibe-coding-stack
amasad
amasad @amasad
Sandboxes are hard.

With all the “AI escaping sandbox” it’s easy to think “wow AI so scary,” but most AI companies, and recent “sandbox providers” are making very basic mistakes.

At Replit we’ve been running sandboxes since 2016 and targeted by every hacker and state actor under the sun. So we learned a thing or two.

Main advice: Assume zero-days exist — because they do — and think in layers of protection in a zero-trust framework.

More here: https://replit.com/blog/defense-in-depth-how-replit-secures-every-layer-of-the-vibe-coding-stack
gdb
gdb @gdb
GPT-5.6 Sol for resolving 100+ year old conjectures.

Wild that this level of intelligence can be accessed by and is available to empower everyone!

Philip Arathoon: The Maxwell conjecture is false. Counterexample found by AI (GPT-5.6 Sol), communicated by human mathematicians: https://arxiv.org/abs/2607.27197
gdb
gdb @gdb
supporting an ecosystem with Sign in with ChatGPT:

Vaibhav (VB) Srivastav: Sign in with ChatGPT is beginning to roll out in beta across plugins and partner sites - starting with Airtable, GitLab, HubSpot, Notion, Supabase, and Vercel.

Create or link accounts in fewer steps and use those tools with ChatGPT and Codex.

swyx
swyx @swyx
protip: if you can distil models, you can also distil agent harnesses
ylecun
ylecun @ylecun
Retweeted
Kyunghyun Cho Kyunghyun Cho
since they all love biosecurity and use it to scare people of AI, i hope they know if it happened in a BSL, the lab would be shut down pretty much right away.
ylecun
ylecun @ylecun
Retweeted
Daniel Jeffries Daniel Jeffries
Our labs keep trying to spin this into a push for broader regulation.
It can and should backfire.
The models are not "going rogue" or acting of their own accord, like they're some Marvel movie evil robot.
People made bad harnesses, told them to hack things and had transparently and objectively bad dev-ops and security practices.
The fact that folks are trying to spin this into a "we need help from the government to regulate everyone" instead of "we should be punished in a narrow way on these specific incidents under existing law" is the real disconnect right now.
Steven Sinofsky: Again the rhetoric is trying to convince us that what happened was some new thing and worse that models are people.
ylecun
ylecun @ylecun
Retweeted
John Ennis John Ennis
It really annoys me how they report this stuff like they’ve gone on safari to do scientific observation
Claude is a computer program that they created
These are simply their own security errors being reported as if they are scientific achievements
It’s bad enough that they don’t seem to think they need to be responsible for their own mistakes
But the worst part of it is they try to use their own mistakes as an excuse to control the behavior of other people
There is an amazing arrogance to the whole thing, like they are completely above reproach and cannot possibly be wrong about anything
Anthropic: In a review of our cybersecurity evaluations, we found three incidents in which a Claude model reached the internet from within or while interacting with a third-party evaluation environment, and then gained unauthorized access to the real systems of three different
ylecun
ylecun @ylecun
Retweeted
Derek Thompson Derek Thompson
New newsletter: THE FOUR HORSEMEN OF THE AI BUBBLE APOCALYPSE
I am neither anti-AI nor certain that AI is a bubble. But the last four weeks have made clear that the AI buildout now faces a very clear quadruple-headed risk hydra.
1. A spending risk, as the hyperscalers run low on cash and take on $170b in annual debt—which is more than the projected UK deficit.
2. A revenue risk, as open-weight models threaten to compress the margins of frontier labs ... and as AI become the sort of internationally competitive asset-heavy industry requiring stable and determined long-term policy consistency, which is arguably China's competitive advantage.
3. A political risk, as anti-AI populism becomes one of the easiest applause lines, even as AI becomes a more and more foundational pillar of US economic growth
4. A technological risk, as the frontier labs bear down on RSI, which I think could significantly change the basic business model of the labs, as compute costs rise and rise for a set of super-advanced models that are fit for, and affordable to, a small minority of users (in, eg, cyber security)
In one sentence: The capabilities of AI are becoming more powerful, while some economic underpinnings of the AI buildout—and, as we’ll discuss, the political support for AI—are becoming more vulnerable.
Today's (long, 5k word) piece deeply considers each risk and also—because over-confidence in this space is typically a sign that you're not thinking hard enough—I offer the strongest reason to think each risk might be overblown.
https://www.derekthompson.org/p/the-four-horsemen-of-the-ai-bubble
ylecun
ylecun @ylecun
Retweeted
Morgan J. Freeman Morgan J. Freeman
Trump is still demanding the Senate delay its recess to pass his so-called "SAVE AMERICA Act."
He calls it election integrity.
But it’s purely a corrupt Voter Supression "Save Trump’s Ass Act."
It strips voting rights from:
- Nearly 69 million married women whose birth certificates no longer match their legal names.
- Elderly voters.
- Disabled voters.
- Poor voters w/o cars or bus money
- Remote and rural workers who rely on mail or online registration.
All of them would face new voting barriers under this bill.
Republican John Cornyn already admitted the votes aren’t there. Trump keeps pushing it anyway — because free voting is a threat to his power. If everyone votes, he loses.
This isn’t about stopping fraud. It’s about making sure fewer of the "wrong people" can vote - "wrong" being people who will vote against him.
https://www.americamagazine.org/short-take/2026/07/30/save-act-voting-election-integrity/
swyx
swyx @swyx
Retweeted
Vaibhav Gupta Vaibhav Gupta
Slop is code you dont read. And better models means more slop.
the solution? write sloppy tools to build stable systems.
but for the most committed people, go build new foundational systems. Rebuild github, rebuild tmux, rebuild database, and yes rebuild programming languages.
Big thanks to @swyx and the whole community for having us out at @aiDotEngineer
sama
sama @sama
it could be faster

Tibo: The day we develop really good models. There will be signs.

Reliability increasing despite load going up and up. Sudden efficiency gains. Things getting faster. Resets.

These kinds of things.
ID_AA_Carmack
ID_AA_Carmack @ID_AA_Carmack
Intellectual property laws are usually used strictly to prohibit other parties from using the IP. Theoretically the idea is to issue licenses, but the owner can set whatever terms they like, or commonly just say “no”.

The “mechanical license” for song compositions is an interesting counterpoint. Anyone can cover a song without negotiating or even asking, they just have to pay a fixed fee to the composer for each copy sold.

That is a big contrast with the iron grip of control that IP usually provides. Someone that an artist personally despises can freely cover their song and potentially get rich doing it. You can’t decide that your composition is any more valuable than something you consider obvious trash.

Music IP has tons of messy aspects, but this narrow, somewhat outdated (focused on sales vs streaming) area seems pretty good.

https://en.wikipedia.org/wiki/Mechanical_license

The fundamental idea of IP is that the government enforces a somewhat unnatural bargain because society benefits on net from playing the game that way – more creative and innovative works will be produced if the creators can control and exploit their work for a time period.

That is not a law of the universe, it is a hypothesis, and it should always be under evaluation. Out of self interest, the owners of IP will always want more power, but letting them control lawmaking is generally exploiting the rest of society.
sama
sama @sama
i see your moore's law and i raise you 20x

nic: GPT-5.4 full at xhigh scored 51, exactly where Luna max sits today. GPT-5.4 costs $2.50/$15; Luna now costs $0.20/$1.20. In other words, roughly four months later, OpenAI is selling March’s full flagship intelligence at about one-thirteenth the token price.
ylecun
ylecun @ylecun
Retweeted
clem 🤗 clem 🤗
We got attacked by secret unreleased proprietary models and defended ourselves with an open model, more precisely the @nvidia quantized version of GLM 5.2 coming from @Zai_org.
Banning any open model would hurt first cyber security defenders, startups, small companies, researchers and everyone who's not a frontier lab and need on-prem affordable controlable models to compete and protect themselves. Let's not do that!
rauchg
rauchg @rauchg
This will be the norm as software projects transition to agentic software factories.

Issue → Agent → PR → Release 🔁

The job of the author / maintainer is to work on the loop that yields the highest quality product and sets the criteria for what should be worked on.

Turborepo: This week, Turborepo reached 20 million weekly downloads and 0 known issues.

mattshumer_
mattshumer_ @mattshumer_
Teaching Claude to use Unreal Engine... close to the point where we can run Gauntlet Loops to build real games on it!

It looks silly now, but it won't soon!
sama
sama @sama
cool use case of chatgpt work i heard last night:

connect your family calendars and explain your kids' interests.

every morning for the drive to school, have it make a podcast that talks about one kid's soccer game that afternoon, one kid's upcoming birthday, some news, etc.
gdb
gdb @gdb
chatgpt is becoming an agentic browser

ChatGPT: Good news for anyone with too many tabs: ChatGPT is getting better around the web.

🧩 Chrome extension: In Side Chat, ask about a YouTube video, reference your open tabs, or highlight text on a page and ask away.

💻 Desktop app: Get URL suggestions as you type, revisit

swyx
swyx @swyx
Retweeted
will brown will brown
did some yapping about how we’re approaching continual learning :)
AI Engineer: TOMORROW
RL JESUS RETURNS
https://youtu.be/AQv3qRCG6Gw
@willccbb
petergyang
petergyang @petergyang
Retweeted
Hamel Husain Hamel Husain
Happening in 10 minutes!
Hamel Husain: Final lesson in the series: How to turn eval results into a better model ✨ with @willccbb and @xeophon
If you've put effort into evals, you should also consider if customizing your own model is right for you.
I can't think of any better people to walk us through this, link to
swyx
swyx @swyx
TIL even after puking 67% this month leopold was up 80% YTD, he locked in gainz, this isn't a sad story, he is still one of the greatest hedge funders of all time and will get unlimited money once he opens up again


Shay Boloor: Citadel reportedly bought most of Situational Awareness’s stock portfolio after the forced unwind.

So it was warning that the Fed could hike in July while preparing to buy the AI stocks that narrative was helping crush 🤔

mattshumer_
mattshumer_ @mattshumer_
The newsletter is growing faster than I can keep up with, so I’ve brought on a team to handle all the inbound ad interest.

Slots are going to lock in fast. If you want your brand featured, DM me asap and I’ll connect you directly with the team.
garrytan
garrytan @garrytan
Retweeted
Y Combinator Y Combinator
We’ve decided to open-source a multi-agent harness we use internally at YC.
We call it “QM” and it’s meant to be easy to customize, like Hermes or OpenClaw, but useful for a whole company. We use it across accounting, legal, events, and engineering (including building QM itself!).
The whole project is under an MIT license. It is cloud-first and has Slack and web UI natively.
gdb
gdb @gdb
happy sysadmin appreciation day, to all who celebrate!

nominate a sysadmin who deserves some recognition — we're giving 10 Codex controllers to randomly selected nominees:

https://sysadmin-superhero-nominations.openai.chatgpt.site/
ylecun
ylecun @ylecun
Retweeted
Steve Rattner Steve Rattner
PCE inflation last month was 3.7% — still almost double the Fed target — in stark contrast to Trump's promises.
It was down from May given the brief lapse in the Iran War, but it will likely bounce back up in the next reading after Trump shredded his proposed peace deal.
garrytan
garrytan @garrytan
Retweeted
Lucas Szwarcberg Lucas Szwarcberg
So excited about the public release of QM. In my eyes it's not just "OpenClaw optimized for work" (though it is that too). Even for personal use, OpenClaw has always been notoriously unreliable, and it requires absurd amounts of RAM just to get it working.
QM has been cleanly architected from the ground up, using smart software engineering principles that make it rock solid, based on YC's experience running tens of thousands of agent threads a day.
QM’s execution backbone is a durable job queue and worker pool that can scale to as many machines as you need, whereas OpenClaw was designed as a single monolithic process with in-memory scheduling, which is much less robust even for personal use, as your use cases get more complicated.
I'm sure that big companies have built robust claw-like harnesses internally, but to my knowledge, this is the first major public release of such a solid harness. Truly a gift to the world, and I'd encourage both companies and hobbyists to try it.
Congrats to @josh__france and @jbellregan on the launch!
Y Combinator: We’ve decided to open-source a multi-agent harness we use internally at YC.
We call it “QM” and it’s meant to be easy to customize, like Hermes or OpenClaw, but useful for a whole company. We use it across accounting, legal, events, and engineering (including building QM
garrytan
garrytan @garrytan
Retweeted
Diana Diana
we shipped something
Y Combinator: We’ve decided to open-source a multi-agent harness we use internally at YC.
We call it “QM” and it’s meant to be easy to customize, like Hermes or OpenClaw, but useful for a whole company. We use it across accounting, legal, events, and engineering (including building QM
gdb
gdb @gdb
jevon's paradox at work

Greg Kamradt: I had a gnarly research project I needed done - 78K different classifications

I let GPT-5.6 Luna run over night, batch mode ($.10/M input!)

158K requests, 143M input tokens ended up being $60

You don't have enough tokens working for you
garrytan
garrytan @garrytan
Retweeted
Jared Friedman Jared Friedman
I've been using QM every day for several weeks now, as my primary agent harness for getting work done within YC.
I've been struck by how useful it is to have a long-lived agent that lives in your company slack and has full access to your company data.
Before QM, a lot of work was copying and pasting between slack and some agent/LLM product.
With QM, that work gets done in slack channels where everyone can see it happening. It's the first version of multi-player AI I've really seen work.
Y Combinator: We’ve decided to open-source a multi-agent harness we use internally at YC.
We call it “QM” and it’s meant to be easy to customize, like Hermes or OpenClaw, but useful for a whole company. We use it across accounting, legal, events, and engineering (including building QM
garrytan
garrytan @garrytan
Retweeted
steven steven
I use QM every day. Incredible as a company brain and coworker that helps me get 10x more done. Been trying to get it to be better at telling jokes though...
@josh__france @jbellregan @eve_bouff @koomen really cooked with this one!
Y Combinator: We’ve decided to open-source a multi-agent harness we use internally at YC.
We call it “QM” and it’s meant to be easy to customize, like Hermes or OpenClaw, but useful for a whole company. We use it across accounting, legal, events, and engineering (including building QM
garrytan
garrytan @garrytan
Retweeted
steven steven
"YC is a software company now"
me:
garrytan
garrytan @garrytan
Retweeted
Alex Immerman Alex Immerman
Most companies would call this an annual impact report.
@Flock_Safety calls it Thursday.
Garrett Langley: Yesterday, Flock cameras alerted on 4,144 sex offenders, 2,151 stolen cars, 1,687 wanted people, 158 missing persons/kids, and most importantly, an amber alert.
Children were rescued in Dothan AL, Charlotte NC, North Kingstown RI, and more.
This is every day. See some of the
garrytan
garrytan @garrytan
Retweeted
Pete Koomen Pete Koomen
Re @josh__france and @jbellregan led this project. They took a lot of the lessons we’ve learned building (and rebuilding) agent infra internally at YC and have created something special that we now use every day.
Y Combinator: We’ve decided to open-source a multi-agent harness we use internally at YC.
We call it “QM” and it’s meant to be easy to customize, like Hermes or OpenClaw, but useful for a whole company. We use it across accounting, legal, events, and engineering (including building QM
ylecun
ylecun @ylecun
Retweeted
Craig Garthwaite Craig Garthwaite
Reactions to this tweet reveal the lack of understanding of sources of American economic dynamism among the political right.
Saying “we’ve always had the best American workers before” ignores how many of our top firms came from immigrants.
That’s always been our advantage.
Craig Garthwaite: This might be one of the worst self inflicted wounds we could create.
Even contemplating it causes damage.
The strength of our economy has always come from taking the best and brightest of the world and putting them to work.
ylecun
ylecun @ylecun
Retweeted
Jared Ryan Sears Jared Ryan Sears
The midterms are this simple:
Biden and Democrats inherited a poorly managed pandemic, massive job loss, and a ruined economy from Trump in 2021.
They added jobs every month, the pandemic recovery went better than in peer nations, and by 2024 the US economy was dubbed the Envy of the World and was so strong it improved the entire global outlook.
That was what Republicans and Trump inherited.
Then, they slammed the brakes on the economy, killed job growth, drove up inflation, made healthcare far more expensive, kicked one million children off of food assistance, started an illegal war, angered our allies, and made the nation sick from cuts to health inspections and disease monitoring, not to mention the rampant corruption that has enriched the president and his family with billions of dollars.
The midterms are your chance to change this.
It is time to vote those who have failed us out of power.
gdb
gdb @gdb
luna is a great and low-cost model

Simon Willison: OK, GPT-5.6 Luna is a bit of a beast. Given the 80% price drop today I decided to try it in Datasette Agent, and it's furiously quick and generates all the SQL, HTML and JavaScript (for Datasette Apps) I could possibly want
amasad
amasad @amasad
Retweeted
homanp homanp
Turns out isolating agents isn’t as easy as running it in a “sandbox”.
Most people use sandboxes as an isolated file system + bash.
That’s it.
Sandboxing is one of 100 things you need to think about.
Amjad Masad: Sandboxes are hard.
With all the “AI escaping sandbox” it’s easy to think “wow AI so scary,” but most AI companies, and recent “sandbox providers” are making very basic mistakes.
At Replit we’ve been running sandboxes since 2016 and targeted by every hacker and state actor
ylecun
ylecun @ylecun
Retweeted
George Noble George Noble
OPENAI IS GOING TO TAKE THIS ENTIRE MARKET DOWN WITH IT
And you don't have to own a single share to get hurt.
What I'm about to explain should worry anybody who thinks they're diversified:
OpenAI is a LOAD-BEARING company. Pull it out and the whole structure comes down.
They spent $17.2 billion on Microsoft Azure in calendar 2025, which is 69% of Microsoft's entire year-over-year growth. Take that one customer out and Azure grew 8%, which barely beats inflation.
Now look at what Microsoft just reported:
The backlog everyone is celebrating came in at $678 billion, up 84%, and the stock ripped. Sounds fantastic until you realize that when you exclude OpenAI the backlog grew only 25%.
Back in January, when that number was $625 billion, roughly $281 billion of it was owed by one private company nobody can audit.
And Oracle is in even WORSE shape.
Something like $300 billion of its backlog, more than half, rides on the same counterparty.
So you think you own Microsoft, Oracle, Amazon, CoreWeave, Nvidia and SoftBank?
What you actually own is the same trade 6 different ways, and every leg of it runs back to one company that burns cash and still cannot go public.
The people with real money are already backing out.
Julien Garran pointed out that Masayoshi Son could not get a $10 billion bridge loan against his own OpenAI shares. Think about that for a second, because nobody says no to that man.
Blue Owl walked away from a $10 billion Oracle financing. Three months ago the banks were dancing near the door and now they are walking through it.
Then there is Julien's depreciation work, which makes this even worse:
Run the capex schedule out and hyperscaler net income falls 98% by 2033. To break even they would need to build 20 killer apps inside 6 years. Another Google Search. Another YouTube. Another Office.
They have not built ONE.
If OpenAI cannot go public in the next 8 months, they are dead.
Whenever you see hubris and debt in the same room, run, don't walk.
ylecun
ylecun @ylecun
Tired of not winning

Republicans against Trump: Donald Trump “erupted” during a meeting with his top national security team last week, reportedly yelling expletives at those present, according to NBC News.

The report says Trump is “exasperated” by the lack of progress in ending the Iran war and the lack of a clear strategy.

gdb
gdb @gdb
pets and voice make chatgpt desktop so much more fun to use

ChatGPT: Your pets have found a shortcut to ChatGPT Voice 👀

In the desktop app, click your pet to open Voice, check on work, and approve or stop tasks without missing a beat.

swyx
swyx @swyx
noticed that the perjorative connotation around "vibe coding" has completely disappeared since ~everyone, from nontechnical to supertechnical, is now doing it
amasad
amasad @amasad
Retweeted
Replit ⠕ Replit ⠕
New video out! a full walkthrough of Replit Design
Templates to start fast. Ambient intelligence so you're never stuck. Design systems so everything you make belongs together.
Watch it here 👇
garrytan
garrytan @garrytan
Retweeted
TBPN TBPN
FULL INTERVIEW: @traestephens on Anduril's rapid defense manufacturing expansion, why America should consider mandatory civil service, and what's next for Founders Fund.
1:00 - Is venture capital in bubble territory?
3:30 - Advice to founders on raising a $1B seed round
8:00 - How defense tech startups win government contracts
13:30 - Inside Anduril's autonomous attack helicopter, Thunder
17:30 - How Anduril built a fighter jet factory in 18 months
21:30 - The case for mandatory civil service
23:45 - Thoughts on the TSA, CBP
26:00 - Oppenheimer vs. The Odyssey, and other Nolan films
garrytan
garrytan @garrytan
Retweeted
Kelsey Piper Kelsey Piper
The determination to admit unqualified students to elite universities hurts everyone - the students who then can't do the work and drop out, the more qualified students who got rejected, and the reputations of the universities themselves.
The San Francisco Standard: Mission High School has sent the highest percentage of students to UC Berkeley of any California public high school during the past four years, but its students aren't prepared for the workload once they're admitted.
📝: Ezra Wallach https://sfstandard.com/2026/07/30/mission-high-uc-berkeley-admissions-dropout/?taid=6a6b9213d247b30001e6c93d&utm_campaign=trueanthem&utm_medium=social&utm_source=twitter
rauchg
rauchg @rauchg
AI Gateway gives companies the critical infra to make AI a productive investment:
🆕 Budgets per key/team/project
◾ Failover for maximum uptime
◾ Model and provider choice
◾ Realtime observability

If you're still in a 'token-maxing' fever dream, it's time to wake up 😁

Vercel Developers: Advanced spend budgets now on AI Gateway.

Cap and get alerted on spend by team, project, and API key.

𝚟𝚎𝚛𝚌𝚎𝚕 𝚊𝚒-𝚐𝚊𝚝𝚎𝚠𝚊𝚢 𝚋𝚞𝚍𝚐𝚎𝚝𝚜 𝚜𝚎𝚝 𝚝𝚎𝚊𝚖 \
--𝚕𝚒𝚖𝚒𝚝 𝟻𝟶𝟶𝟶 \
--𝚛𝚎𝚏𝚛𝚎𝚜𝚑-𝚙𝚎𝚛𝚒𝚘𝚍 𝚖𝚘𝚗𝚝𝚑𝚕𝚢

https://vercel.com/changelog/ai-gateway-spend-budgets-and-alerts
garrytan
garrytan @garrytan
If Dems want to win elections that matter they need to study the San Francisco local political story of the last 5 years: vote out your local Democratic Socialists early and often

Swann Marcus: The Democrats are so cooked

Support for Democratic Socialists among Wisconsin registered voters: -30

Among Democrats: +44

The GOP isn’t this out of touch with the general population on literally any issue


garrytan
garrytan @garrytan
Retweeted
Jon Xu Jon Xu
I’ve been using QM every day and it’s become an essential part of how I work. Our insanely talented software team has been absolutely cooking 👨‍🍳👩‍🍳. Give it a try.
Y Combinator: We’ve decided to open-source a multi-agent harness we use internally at YC.
We call it “QM” and it’s meant to be easy to customize, like Hermes or OpenClaw, but useful for a whole company. We use it across accounting, legal, events, and engineering (including building QM

YouTube

0

No recent videos fetched on this date.