Matt Shumer
Yes, you can set it up to do this. But 99.9% of users won’t. Claude should have great native computer use.
中文: 是的,你可以设置它来做这件事。但99.9%的用户不会。克劳德应该拥有出色的原生计算机使用。
Matt Shumer
Claude Opus 5.5 with Codex-level computer use would be unbeatable
中文: 使用 Codex 级别的计算机的 Claude Opus 5.5 将保持无与伦比的水平
Matt Shumer
Spawn is fucking incredible. IYKYK. Will have more to share/say on this soon. 🤐
中文: 斯潘他妈的太不可思议了。 伊基克。 很快就会有更多关于此的分享或说。🤐
Matt Shumer
AI companies: want to get your product in front of thousands of early adopters? I'm launching a Pro plan for https://somethingbig.ai/ readers, and one of its perks will be free or discounted access to awesome AI tools. I'm looking for a few great companies to partner with. You offer Pro subscribers a deal on your product; I put it in front of an engaged audience that actively looks for new AI tools. If that sounds interesting, comment or DM me!
中文: 人工智能公司:希望将你的产品推向成千上万的早期采用者面前? 我正在为 读者推出一项专业版计划,其一项优惠将免费或享受出色的人工智能工具。我正在寻找一些能与之合作的优秀公司。 你为Pro订阅用户提供了产品优惠;我向一群积极参与并积极寻找新人工智能工具的受众表示。如果这听起来有趣,请评论或DM我!
Matt Shumer
Opus-generated videos have made their way to LinkedIn https://twitter.com/mattshumer_/status/2105876711682552246/photo/1
中文: 由Opus生成的视频已前往LinkedIn
Matt Shumer
Y’all should join this, it’s crazy
中文: 你们都应该加入这个,这太疯狂了
Matt Shumer
Catch up on the seven things that mattered in AI this week, in four minutes, without doomscrolling: https://somethingbig.ai/what-actually-mattered-sep-25-29
中文: 本周,四分钟内无需进行末日滚动,即可了解人工智能中最重要的七件事:
Matt Shumer
The* not a :)
中文: 不是 :)
Matt Shumer
Full review coming soon on https://somethingbig.ai/
中文: 即将在 上进行完整评测
Matt Shumer
I’ve been testing Dots for a bit. In short, it has a potential to be the best agent product in the world. Why? The quality of the AI is best-in-class. But the UX needs a lot of work before I can make it my daily driver. If OpenAI nails that, they’ve got a huge winner.
中文: 我一直在测试一下Dots。 简而言之,它有潜力成为世界上最好的代理产品。 为什么?人工智能的质量是一流的。 但用户体验需要大量工作,才能成为我的日常驱动力。 如果OpenAI能确定这一点,他们就拥有了巨大的赢家。
Matt Shumer
Ugh. I wanted so badly to be wrong on this one.
中文: 呃。我非常想在这件事上犯错。
Matt Shumer
So Claude had a busy night experimenting, apparently… https://twitter.com/mattshumer_/status/2104962412256334333/photo/1
中文: 所以克劳德显然在忙碌的夜晚进行了实验......
Matt Shumer
Also if anyone wants to send me a robot arm, this is just the start of what I’m doing :)
Matt Shumer
I asked Opus 5.5 to find a way to clear my 3D printer’s print bed, so it can experiment non-stop without me in the loop. It took 16 attempts. I'm officially out of the loop. Video by Opus, obviously. https://twitter.com/mattshumer_/status/2104777202684301750/video/1
中文: 我让Opus 5.5找到清除我3D打印机打印床的方法,这样它就可以在没有我的情况下进行不间断的实验。 花了16次尝试。我正式脱离了循环。 视频由Opus拍摄,显然。
🎬
视频
Matt Shumer
RT @SteveMoraco: “It’s not as smart as humans yet” okay name one person on earth capable of this
中文: RT @SteveMoraco:“还没像人类那么聪明”,真好,地球上有一个人能够做到这一点
Matt Shumer
Magic is real. And it's about to be everywhere.
中文: 魔法是真实的。 而且它即将无处不在。
Matt Shumer
Btw, my next set of demos are going to be even crazier than this
中文: 接下来,我的演示将会比这更加疯狂
Matt Shumer
Prompt: “build a working computer from scratch, entirely in code, starting from individual logic gates. design the CPU, the memory, an assembler, a tiny operating system, and a game that runs on it. everything has to be real, the game has to run on your gates, not in javascript pretending. visualize the whole thing in 3D so i can play the game, then zoom all the way down through the chips into the gates and watch the signals flow while it runs. treat it like you're proving you could have invented computing yourself. go all out.”
中文: 提示:“从零开始构建一台完全以代码为起点的工作计算机. 设计CPU、内存、汇编器、微型操作系统以及运行在它上的游戏. 一切必须真实,游戏必须在你的大门上运行,而不是在JavaScript中假装. 将整个功能以3D方式可视化,以便我可以玩游戏,然后通过芯片向下放大,观察信号在运行时的流畅运行. 把它当作你证明自己可以自己发明计算一样. 全放。
Matt Shumer
This was done from one prompt, completely autonomously, over ~30 hours in ultracode mode. Prompt: “build a working computer from scratch, entirely in code, starting from individual logic gates. design the CPU, the memory, an assembler, a tiny operating system, and a game that runs on it. everything has to be real, the game has to run on your gates, not in javascript pretending. visualize the whole thing in 3D so i can play the game, then zoom all the way down through the chips into the gates and watch the signals flow while it runs. treat it like you're proving you could have invented computing yourself. go all out.”
Matt Shumer
I asked Opus 5.5 to build a working computer from scratch. And holy fuck did it deliver. 277k logic gates in JS, an OS on top, and games on the OS. This isn't a mockup... it's a computer. Run it: https://somethingbig.ai/computer
中文: 我让Opus 5.5从零开始构建一台工作计算机。 而圣之礼确实做到了。 JS 中的 277k 逻辑门、顶部的操作系统以及操作系统上的游戏。这不是模拟......而是一台电脑。 运行它:
🎬
视频
Matt Shumer
中文: Opus 5.5 为 添加了控制器支持
🎬
视频
Matt Shumer
So I bought a 3D printer for Opus 5.5 to use. Hopefully, I’ll have some pretty crazy demos soon. https://twitter.com/mattshumer_/status/2104070682728685733/photo/1
中文: 所以我买了一台用于Opus 5.5的3D打印机。 希望我很快就会有一些非常疯狂的演示。
Matt Shumer
So many people in denial, it's crazy
中文: 否认的人太多了,简直太疯狂了
Matt Shumer
I just launched a run that I think will yield something even crazier than this. If it works, I’ll post it. Super excited.
中文: 我刚刚发起了一场我认为会带来比这更疯狂的事情。 如果能用,我就发。超级兴奋。
Matt Shumer
I asked Opus 5.5 to make an animated short film entirely in code. The result is pretty damn incredible. Worth a watch. https://twitter.com/mattshumer_/status/2103928105320763818/video/1
中文: 我让Opus 5.5完全以代码为由制作一部动画短片。 结果简直令人难以置信。 值得一看。
🎬
视频
Matt Shumer
This is pretty terrifying, but seems like OpenAI is taking it seriously and pausing most frontier inference until they’ve figured it out.
中文: 这相当可怕,但似乎OpenAI正在认真对待,并暂停大部分前沿推断,直到他们弄清楚为止。
Matt Shumer
Opus 5.5 apparently built https://somethingbig.ai/world with a weather system that matches whatever weather is in NYC at the time... So with the nor'easter, there's rain, and the NPCs are carrying umbrellas! It's the little things :) https://twitter.com/mattshumer_/status/2103667211395268863/photo/1
中文: Opus 5.5 显然是使用天气系统构建的 因此,随着东北风的出现,人们正在下雨,而全国人民代表大会正在举着雨伞! 是小东西 :)
Matt Shumer
Opus 5.5 decided to add helicopters to https://t.co/qmlFoHdeIz... see if you can find them! It tells me it's going to keep making them better throughout the day. https://twitter.com/mattshumer_/status/2103570652116467998/video/1
中文: Opus 5.5 决定在 中添加直升机......看看你能找到它们吗! 它告诉我,它会一直让它们一整天都变得更好。
🎬
视频
Matt Shumer
Local models are useless
中文: 本地模型毫无用处
Matt Shumer
New models are like new hires. When a new one comes out, you have to learn its personality, how to work with it, etc.
中文: 新模式就像新员工一样。 当一个新出现时,你必须学会它的个性,学会如何与它合作,等等。
Matt Shumer
The 2030s are going to be so weird as nanotechnology is developed. It’ll likely be the most profound change humanity has ever experienced… likely even more than AI.
中文: 2030年代将会非常奇怪,因为纳米技术已经发展起来。 这很可能是人类有史以来经历的最深刻的变化......可能比人工智能还要深刻。
Matt Shumer
You’re about to be the commander of an army of superintelligent AI agents. Please tell me you have something bigger planned than a fucking SaaS app. Think about what you’d attempt if you had a thousand geniuses devoted to helping you.
中文: 你即将成为一支超级智能人工智能特工队伍的指挥官。 请告诉我,你的计划比他妈的SaaS应用还要多。 想一想,如果你有一千个天才致力于帮助你,你会尝试什么。
Matt Shumer
Playing around with an idea here to help people see what the future might look like. Check it out, I'm curious for your thoughts: https://somethingbig.ai/futures
中文: 在这里考虑一个想法,帮助人们了解未来可能是什么样子。 来看看,我好奇你的想法:
Matt Shumer
RT @znewm14: It’s been nearly 8 months. What are everyone’s thoughts?
中文: RT @znewm14:已经快8个月了。每个人的想法是什么?
Matt Shumer
Very compute-constrained right now and having trouble finding one. Happy to make this trade. :)
中文: 现在计算能力非常受限,且难以找到。很高兴能做这个交易。:)
Matt Shumer
If your company sends me a 256/512gb Mac Studio, I'll make you go viral.
中文: 如果你的公司给我寄一个256/512克的Mac工作室,我会让你走红。
Matt Shumer
The world is changing fast.
中文: 世界正在迅速变化。
Matt Shumer
I’ve set a few friends and family up with this, and they absolutely love it.
中文: 我为这个设立了几个朋友和家人,他们非常喜爱。
Matt Shumer
I haven’t used my computer manually in weeks. I text with an AI and ask for what I want. And it’s so freeing. This is going to be very normal, very soon.
中文: 我已经好几个星期没用过电脑了。我用人工智能发短信,并询问我想要什么。 而且它非常自由。 这很快就会变得非常正常。
Matt Shumer
The first one (by GPT-5.6-Sol) was considered an unbelievable leap forward… just 76 days ago. What will things look like in December?
中文: 第一场(由GPT-5.6-Sol担任)在76天前就被认为是一次令人难以置信的飞跃。 十二月会有什么情况?
Matt Shumer
These two NYC builds were done less than three months apart. The world is changing so quickly. And it’s only going to get faster from here. https://twitter.com/mattshumer_/status/2102913726810591347/video/1
中文: 这两座纽约市建筑的建造间隔不到三个月。 世界正在迅速变化。 而且从这里开始,速度会越来越快。
🎬
视频
🎬
视频
Matt Shumer
Fun prompting trick: The models have seen a lot of my prompts from the last few years, so you can literally ask them things like: - how would Matt Shumer write this prompt? - how would Matt Shumer steer you to solve this problem? Performance usually improves quite a bit. It’s a bit strange to say out loud, but a lot of my friends and I use this trick often and it works really well.
中文: 有趣的提示技巧: 模特们看到了我过去几年的很多提示,所以你可以从字面上去问它们: - 马特·舒默如何写出这个提示? - 马特·舒默将如何引导你解决这个问题? 性能通常会有相当的提升。 大声说出来有点奇怪,但我和很多朋友经常使用这个技巧,效果非常好。
Matt Shumer
@nateliason Feared by all
中文: @nateliason 被所有人恐惧
Matt Shumer
@nateliason Ah yes, Copilot, king of the AI tools!
中文: @nataliason 是的,Copilot,人工智能工具之王!
Matt Shumer
I shouldn't be telling you this... But @spawn is offering unlimited Opus 5.5 for FREE. Go and get it before they remove it.
中文: 我不该告诉你这个...... 但@spawn 免费提供无限 Opus 5.5。 去取之前先拿掉。
Matt Shumer
Opus 5.5 has been looping on my open-world NYC game for almost a day now, and it's continuing to improve. Jump in and try it (it's multiplayer!): https://somethingbig.ai/world/ https://twitter.com/mattshumer_/status/2102874271316078841/video/1
中文: Opus 5.5 已经在我的开放世界纽约游戏上持续了将近一天,而且它仍在不断进步。 跳进去试试(这是多人游戏!):
🎬
视频
Matt Shumer
RT @joexrogan: jacob has the world record for most expensive Claude run (don’t ask, won't tell) side note: we are all playing https://www.spawn.co/@j/portal-heist on spawn right now - Opus 5.5 multiplayer game created last night, hop in if you see this!
中文: RT @joexrogan:雅各布拥有最昂贵的克劳德跑步世界纪录(不要问,不会说) 旁注:我们目前都在玩 的 speting——昨晚创建的 Opus 5.5 多人游戏,如果看到这个就跳进去吧!
Matt Shumer
Jacob sent me a screenshot of the cost of a Claude run he did the other day. I almost feel bad saying just how much it spent. Immediately texted @joexrogan to make sure he wasn’t going to faint.
中文: 雅各布给我发了一张他前几天做的克劳德跑步费用的截图。 我几乎觉得说它花了多少钱才感到难受。 立即发短信给@joexrogan,以确保他不会晕倒。
Matt Shumer
Jacob literally built a custom game engine designed from the ground up for LLMs. And now he’s using it to build things that would be impossible to make in any other way.
中文: 雅各布确实为LLM设计了一款定制游戏引擎。 现在他正在利用它来制造那些无法用其他方式制造的东西。
Matt Shumer
RT @jsnnsa: I had 26 Opus 5.5 agents build a game overnight. It's a heist through five nested worlds, with go-karts on a giant's kitchen table, bots hunting you, and friends stealing your loot. Most vibe-coded games are demos. This one is different.. it's multiplayer and you can play it right now.
中文: RT @jsnnsa:我有26名Opus 5.5经纪人在一夜之间完成了一场比赛。 这是一起穿越五个巢穴世界的抢劫,一只巨人的厨房桌子上放着卡丁车,机器人正在追捕你,还有朋友偷走你的战利品。 大多数氛围编码的游戏都是演示。这个不一样......它是多人游戏,现在就可以玩了。
🎬
视频
Matt Shumer
The models are just stochastic parrots, right?
中文: 这些模型只是随机鹦鹉,对吧?
Matt Shumer
I wrote about what the world might look like if AI goes well. Plus, readers get three months of Granola free: https://somethingbig.ai/what-if-ai-goes-well?utm_source=twitter
中文: 我写过关于人工智能顺利时世界可能是什么样子的文章。 此外,读者可免费获得三个月的格拉诺拉:
Matt Shumer
Anthropic has won. This model is incredible.
中文: 安东尼·托利奇赢了。 这个模型令人难以置信。
Matt Shumer
There are more releases coming this week ;)
中文: 本周还有更多发布会;)
Matt Shumer
@alexatallah @jsnnsa Platform was amazing before Jev, but now he’s cooking something up w/ Jev on top
中文: @alexatallah @jsnnsa Platform 在 Jev 之前非常出色,但现在他正与 Jev 进行烹饪
Matt Shumer
@alexatallah Alex you should meet @jsnnsa
中文: @alexatallah 亚历克斯,你应该见见 @jsnnsa
Matt Shumer
I forgot how freeing it is to use Opus and not worry about hitting my limits every five minutes. I’m so much less stressed this way.
中文: 我忘了使用Opus是多么自由,而不必担心每五分钟就达到我的极限。 我这样压力要小得多。
Matt Shumer
I want to thank @ChrisGPT, who has done some amazing Gauntlet Loop demos, for contributing to my NYC Open World project! He did a fantastic job. Now that Opus 5.5 is out, I’m going to try to take it to the next level. Starting the loop tonight.
中文: 我想感谢@ChrisGPT,他为我的纽约开放世界项目贡献了出色的Gauntlet Loop演示! 他做得非常出色。 既然Opus 5.5已经出局,我将努力将其提升到一个新的水平。今晚开始循环。
Matt Shumer
RT @spawn: JEV is now in Spawn it makes decisions 10 times a second. creators are already using it for NPCs, bosses, and puzzles that fight back. make your own. just ask Savi.
Matt Shumer
People are loving 5.5. Congrats to Anthropic! https://twitter.com/mattshumer_/status/2102494865279926522/photo/1
中文: 人们喜欢5.5。 恭喜人类!
Matt Shumer
Well said.
Matt Shumer
Launching the first Opus 5.5 Gauntlet Loop! Going to have it redesign https://t.co/xkmv6L8MFo. Watch this thread, I'll post results here as it works.
中文: 推出首个Opus 5.5 Gauntlet Loop! 将进行重新设计 观看此帖子,我会在此处发布结果。
Matt Shumer
Review coming soon at https://somethingbig.ai/
Matt Shumer
Will have more to say in the coming days. Subscribe at https://somethingbig.ai/ to get my full review!
中文: 未来几天将有更多话要说。请在 获取我的完整评价!
Matt Shumer
I've been testing GPT-6 Sol for a bit now. It's solid, but I still prefer Astra/Fable 5.1 (and now, likely Opus 5.5) for my day to day. We're spoiled by so many amazing model options now!
中文: 我现在已经测试一下GPT-6 Sol了。 很稳固,但我仍然更喜欢Astra/Fable 5.1(现在,可能是Opus 5.5),我的日常选择也是如此。 我们现在被这么多令人惊叹的型号选择所宠爱!
Matt Shumer
Opus 5.5 feels like chatting with a way smarter Opus 4.6. I’m so happy.
中文: Opus 5.5 感觉就像在用更智能的 Opus 4.6 聊天。 我太开心了。
Matt Shumer
Holy crap, these numbers look good. I haven't had a chance to test the model yet... but if it's as good as these benchmarks say, it will be a massive step up from Fable 5.1. Just think about that... 5.1 is now outdated. This is what acceleration looks like.
中文: 该死的,这些数字看起来不错。 我还没有机会测试这个模型......但如果它和这些基准所说的一样好,那么它将从Fable 5.1大幅上升。 想想那......5.1 现在已经过时了。 这就是加速的外观。
Matt Shumer
Founders: VCs are feeding your pitch into an LLM. Pitching is prompting now, so speak in a way that gets the model to write the memo you want. :)
中文: 创始人:风险投资公司正在将你的宣传转化为一场法学硕士。 投球现在引起了你的提示,所以要以一种让模型来写你想要的备忘录的方式说话。:)
Matt Shumer
Chris is awesome and put so much into building this. Please check it out!
中文: 克里斯非常出色,为此投入了很多。 请查看一下!
Matt Shumer
RT @chrsabraham: I just built the largest collection of helpful prompts (500+) to turn Muse into the best chief of staff, best marketer, best PM you've ever had. @Muse can be so much more than a glorified admin assistant. There are hundreds of ways to use it at your job. Completely free, check it out: https://museatwork.app/
中文: RT @chrsabraham:我刚刚打造了最大的实用提示集(500+),让Muse成为你有史以来最好的幕僚、最佳营销员和最佳总理。 @Muse 可能不仅仅是一位荣耀的管理员助理。在工作中使用它有数百种方法。 完全免费,请查看:
Matt Shumer
Will be writing about it at https://somethingbig.ai/
中文: 将在 上撰写文章
Matt Shumer
I’ve been testing Grok 4.7… it’s a really great daily driver, and a huge step up over 4.6. Definitely worth trying!
中文: 我一直在测试Grok 4.7......它是一位非常出色的日常驾驶者,且在4.6以上大幅上涨。 绝对值得一试!
Matt Shumer
RT @businessbarista: This guy build a copy of Call of Duty with one prompt & went stupid viral (20 million views). Now while building video games is cool, what's even cooler is that his process, called The Gauntlet Loop, can be applied to any type of professional work. Here's how @mattshumer_'s process works: Problem this solves: Agents stop at “good enough.” They do the ask once and declare done, especially in visual/creative work. Even if you force more iterations, models judge their own work. Like a student grading their own exam, they give themselves 100. That self-grading is the ceiling. Step 1: Set a real, inspectable bar - “most recent Call of Duty” level, not “a great FPS”). If yours isn’t better, you’re not done. Step 2: Split the job across specialist sub-agents instead of one overloaded agent. Step 3: Loop each piece until it looks like a real game/ real work, not “pretty good for AI.” - Version 1 will be garbage. Step 4: Blind critics compare your output to a reference and keep going until they pick yours (or you stop). Why the critic must be blind - In Claude Code / Codex-style harnesses, sub-agents often fork the main agent’s context, so a “critic” still remembers the builder’s rationale and rubber-stamps it. - Fix: spawn critics with totally fresh context. They only see the artifacts (e.g. two images), pick which looks better, and don’t know which is yours. - Until the critic picks yours, keep looping. That’s the whole loop. How Claude of Duty actually ran - Lead agent decomposes → each piece gets builder + critic → before/after per subsystem → fold back into the game → repeat waves. - He did not hand-feed a big pile of Call of Duty reference images. He set the prompt and the agent went out and found comparison material itself. References when the thing doesn’t exist yet For something novel (e.g. a futuristic weapon), you can: - Use a different game/object as a quality-level comp (critic judges “which looks better overall,” not identity match), or - Generate target stills with an image model until you like them, then feed those as the bar. Same loop outside games - Writing (he called this the harder example): after each iteration, A/B paragraphs (or page/chapter) against strong recent comps. Avoid famous dead authors the model already “knows” as themselves (IP + identity). Use contemporary comparable writing. - Websites: feed sites you think are world-class; don’t stop until blind critics consistently prefer yours. - How prescriptive on “what good means”: optional. Experimental runs he lets loose. Real work he gets more prescriptive and steers mid-loop (“like this direction, but pare back”). - Growth: Something Big newsletter is already using the loop for growth strategies and conversion copy; he said subscriber CAC results are far above industry standard. Cost and when to stop The loop can run forever. - Demos / toys: don’t max it. Too expensive for the value. - Real work that matters: willing to spend a couple hundred dollars / hit subscription limits. If the artifact is valuable, $200 is cheap relative to what that demo would have cost a few years ago. - He stopped Claude of Duty while it was still improving, because of cost and “already wow,” not because the bar was fully beaten. He bets a couple more days would get much closer to real CoD level. Full episode: https://www.youtube.com/watch?v=Sgas9rVHegc
🎬
视频
Matt Shumer
RT @spawn: if you can make a game in an hour, imagine what you and your friends could build in a year. we built Spawn to give you room for that. worlds 100x the size of GTA. 1000s of players together. portals connecting what you create. live today and improving every hour.
中文: RT @spawn:如果你能在一小时内制作一款游戏,想象一下你和朋友们一年能创造什么。 我们为您打造了Spawn,为您提供了空间。世界规模是GTA的100倍。1000名玩家在一起。通过传送门连接您创建的内容。 活到今天,每时每刻都在进步。
Matt Shumer
I made a video on Jev. People have been asking me to make videos forever, so I made one today on a whim. It's my first crack at this. I don't have a microphone or any setup (don't expect perfection). Let me know if you like it! https://www.youtube.com/watch?v=6FT9GpKaaYQ
中文: 我在Jev上制作了一段视频。 人们一直要求我永远制作视频,所以我今天就一时兴起制作了一个。 这是我在这方面的第一个裂痕。我没有麦克风,也没有任何设置(不要期望完美)。 如果你喜欢,请告诉我!
Matt Shumer
This is fantastic. I agree with all but two of these predictions. Curious if anyone can figure out what they are!
中文: 这太棒了。我同意这些预测中的两个。 好奇是否有人能弄清楚他们是什么!
Matt Shumer
Is anyone trying Jev for scalable oversight? Seems like it could be useful for helping determine if a model is behaving in an aligned way.
Matt Shumer
@oreHistorian @KellyCNBC Meta's models aren't even in the same league as Anthropic's. The risk profile is far different.
中文: @oreHistorian @KellyCNBC Meta 的模特们甚至与 Anthropic 的模特都不在同一个联赛中。风险状况大不相同。
Matt Shumer
@oreHistorian @KellyCNBC They do not know how to do it. Plain and simple.
中文: @oreHistorian @KellyCNBC 他们不知道该怎么做。简单明了。
Matt Shumer
Would love to try Jev! Can anyone get me access?
中文: 真想试试Jev!有人能让我获得访问权限吗?
Matt Shumer
RT @denk_tweets: one of my favorite AI newsletters written by @mattshumer_ is live on the recommendation network @beehiiv users can generate passive revenue: > apply to recommend his newsletter > new readers see the recommendation when they subscribe > you get paid $2 per reader who opts in https://twitter.com/denk_tweets/status/2100775220219130059/photo/1
Matt Shumer
Big opportunity for anyone with a Beehiiv newsletter!
中文: 任何拥有Beehiiv简报的人都会有大好机会!
Matt Shumer
Big opportunity for anyone with a Beehiiv newsletter!
中文: 任何拥有Beehiiv简报的人都会有大好机会!
Matt Shumer
RT @denk_tweets: one of my favorite AI newsletters written by @mattshumer_ is live on the recommendation network @beehiiv users can generate passive revenue: > apply to recommend his newsletter > new readers see this recommendation when they subscribe > you get paid $2 per reader who opts in https://twitter.com/denk_tweets/status/2100773241552044428/photo/1
Matt Shumer
Anyone mind tossing me a Muse invite?
中文: 有谁介意向我招来缪斯邀请吗?
Matt Shumer
Will be talking more about this on https://somethingbig.ai/
Matt Shumer
This can’t possibly be real. Agents don’t just use the web during training and testing. They browse it for users every day. If it’s contaminated, we’re all already exposed.
中文: 这不可能是真的。 代理在训练和测试过程中不会只使用网络。他们每天为用户浏览它。 如果被污染了,我们都已经暴露了。
Matt Shumer
If you’re a bootstrapped / lightly-funded AI founder with a product that gets more than 500 new signups per day, DM me. I have something for you. Please don’t DM if you have less than 500 signups per day.
中文: 如果你是一位被束缚或资金不足的人工智能创始人,拥有每天获得超过500个新注册产品的产品,请告诉我。 我有适合你的东西。 如果每天报名少于500人,请不要说。
Matt Shumer
Talking Gauntlet Loops live with Alex!
中文: 与亚历克斯一起在聊天的G G)循环直播!
Matt Shumer
Gave my mom my (very simple) agent setup. She named him Morgan. He just closed a $600 ad deal for her site. People are massively overthinking their setups. Bitter Lesson applies to agents too: just ask the model, let it spawn what it needs. https://twitter.com/mattshumer_/status/2100279334715924835/photo/1
中文: 给妈妈我(非常简单)的代理设置。 她给他取名叫摩根。他刚刚为她的网站完成了一笔600美元的广告交易。 人们正在过度思考自己的设置。苦味课同样适用于代理:只需询问模型,让它产生所需的内容即可。
Matt Shumer
Spencer might be the single most focused founder I’ve ever met. He’s been working on this problem since I met him, back in 2020. They are by far, my bet to win in this category.
中文: 斯宾塞可能是我见过的最专注的创始人。 自从我2020年见到他以来,他一直在研究这个问题。 到目前为止,我打赌在这个类别中获胜。
Matt Shumer
RT @Spshulem: AI is killing your company. It should be making you more revenue. Introducing BuildBetter: the first AI Head of Product. Here’s how it works 👇 https://twitter.com/Spshulem/status/2099922862488355298/video/1
🎬
视频
Matt Shumer
Guys, these insane setups are fun and all, but they don't actually make you more productive. You just end up spending more time building the system than actually doing real work. All you need is one agent session that farms out to others and acts as a manager. If something's not working, just tell that to your main session and it'll adjust the overall approach. If you're using the best models available, you shouldn't need anything more than this. This is coming from a guy who runs hundreds of agents a day, dozens at a time, with spikes in the hundreds at a time. I'll be writing more about this soon.
Matt Shumer
RT @chrsabraham: Can confirm. Matt's use of AI has changed a lot of my perspective with what we can use it for. Insane!
Matt Shumer
A few folks have had a preview of my setup (cc @JasonKuperberg @chrsabraham)... it's really insanely good
中文: 有几个人已经预览了我的设置(@JasonKuperberg @chrsabraham)......这真是太好了
Matt Shumer
With Astra and Fable 5.1, we officially have drop-in remote workers. Still takes quite a bit of setup, but once you have it dialed in, it's near perfect. I haven't touched my computer in a few days, and I'm more productive than ever.
中文: 配备Astra和Fable 5.1,我们正式拥有远程工作人员。 仍然需要相当多的设置,但一旦将其拨入,就近乎完美。 我几天没碰电脑了,工作效率也比以往任何时候都高。
Matt Shumer
@ivanburazin It's awesome at first, but gets worse over time as context fills up... for me, it's dropping things left and right
Matt Shumer
@ivanburazin Agree
中文: @ivanburazin 同意
Matt Shumer
Subscribe (free) to https://somethingbig.ai/ to get my full workflow when I share it!
中文: 订阅(免费)
Matt Shumer
Yeah so @t3dotcodes is awesome. Still a bit rough around the edges, but it’s worth it for how useful it is. I’ve redesigned my entire workflow around it. I’ll be sharing more about it soon! https://twitter.com/mattshumer_/status/2099560824121565323/photo/1
中文: 是的,@t3dotcodes 太棒了。 边缘仍然有些粗糙,但因为它的用处是值得的。 我围绕它重新设计了整个工作流程。我很快就会分享更多!
Matt Shumer
Fucking incredible
中文: 他妈的难以置信
Matt Shumer
Decacorn incoming.
中文: 德卡纳传入。
Matt Shumer
RT @carb0n_lifef0rm: This article hit nearly 90M views earlier this year. Roughly "out of nowhere". Friendly reminder.
中文: RT @carb0n_lifef0rm:本文今年早些时候的观看次数接近9000万次。大致“无中生有”。友好的提醒。
Matt Shumer
Something Big might be the fastest-growing newsletter in the world right now... And we haven't even officially launched yet. Insane.
Matt Shumer
The era of subsidized tokens is ending. Prepare accordingly.
中文: 补贴代币的时代即将结束。 做好相应的准备。
Matt Shumer
Also all the sand is going to be turned into chips, so there’s that.
中文: 而且所有的沙子都会变成薯片,所以就是这样。
Matt Shumer
This was crazy to watch unfold. @naturalpay is sponsoring Something Big, and Khalil asked if he could pay through his agent. It emailed mine, and the two of them handled the whole transfer (substantial $) themselves. The whole thing happened in under five minutes.
中文: 这太疯狂了,令人着迷。 @naturalpay 正在赞助“大件事”,哈利勒问他是否可以通过他的经纪人付款。 它发邮件给我,他们两人自己处理了整个转账(大额)。 整件事发生在不到五分钟。
Matt Shumer
AI takeoff is happening. If your head is still in the sand, now is the time to take it out.
中文: 人工智能的起飞正在发生。 如果你的头还在沙子里,现在是时候把它拿出来了。
Matt Shumer
RT @evanjconrad: the average on-platform customer on sfcompute saves about 25% it's the difference between $4.5/hr and $3.3/hr we wrote about why and how you can scale faster with reduced risk https://sfcompute.com/news/realized-discount
Matt Shumer
Theo is right. In 2020, the US spent ~25% of GDP fighting a virus expected to kill under 1% of Americans. Many experts think AI risk is much larger. Spending 1% of GDP (~$300B) a year on alignment incentives is justified.
中文: 西奥是对的。 2020年,美国花费了约25%的GDP来对抗一种病毒,这种病毒预计会导致不到1%的美国人死亡。许多专家认为人工智能的风险要大得多。 每年将1%的GDP(约300B)用于统一激励是合理的。
Matt Shumer
@Blum_OG Reach out to elected officials, and advocate for strong incentives to build out massive AI alignment research programs.
中文: @Blum_OG 联系民选官员,倡导大力激励构建大规模人工智能对接研究项目。
Matt Shumer
Would you board a plane with a 10% chance of crashing? That's roughly the odds some people building AI give for it killing us all. I wrote what it means, and what we can do about it. Send this to someone who needs to read it.
Matt Shumer
Matt Shumer
Matt Shumer
We need a Manhattan Project for AI alignment. This is the single most important issue of our time. The world needs to come together, invest the appropriate (enormous) resources, and make this happen before time runs out.
中文: 我们需要一个人工智能对齐的曼哈顿计划。 这是我们这个时代最重要的问题。 世界需要团结起来,投入适当(大量)资源,并在时间耗尽之前实现这一点。
Matt Shumer
Astra is incredible, but it cost me the one thing I loved about OpenAI: the plan limits. Fable would burn my whole plan in hours. GPT I could just run all day. Now I’m four resets down, with a new sub in 24 hours. Those days are gone.
中文: Astra 太不可思议了,但让我对 OpenAI 的喜爱却让我付出了代价:计划是有限的。 费姆会在几小时内烧毁我的整个计划。GPT 我可以跑一整天。 现在我已经四次重置,24小时内再设置一个新子。那些日子已经过去了。
Matt Shumer
RT @mattshumer_: Full guide here: https://somethingbig.ai/3d-worlds
中文: RT @mattshumer_:完整指南,网址:
Matt Shumer
RT @mattshumer_: My GPT-6 Astra 3D builds just passed 15M views on X. The most common question, by far: "how are you getting this level of quality?" Here's exactly how. The full guide: https://twitter.com/mattshumer_/status/2097079239505842645/video/1
🎬
视频
Matt Shumer
I’ve been working with Chris to accelerate the NYC world. He’s awesome at this and it’s moving even faster now. https://somethingbig.ai/world (still early)
中文: 我一直在与克里斯合作,共同加速纽约世界的发展。他在这方面非常出色,而且现在进展得更快了。
Matt Shumer
RT @ChrisGPT: I’m happy to say I’ve been collabing with who I think is the game demo GOAT, @mattshumer_, on his open-world NYC game. As AI models continue to get exponentially better at game development, me and Matt hope to keep pushing the bounds of what’s possible with AI game creation. It’s exciting that we’re only in 2026, and there’s still no wall in sight for the scale and complexity of what we’ll be able to create in the future.
🎬
视频
Matt Shumer
@serudda If you're still confused, you should look at the whole debacle with GPT 5.6 Sol where it deleted my computer. If they were paying me, would I have posted that?
中文: @serudda 如果你仍然感到困惑,你应该看看 GPT 5.6 Sol 的整个崩溃,它会删除我的电脑。如果他们付钱给我,我会发那份吗?
Matt Shumer
@serudda I don't know how many times I have to say this, but OpenAI does not pay me. In fact, I pay them quite a bit for tokens
中文: @serudda 我不知道有多少次话要说,但OpenAI并不支付我的报酬。事实上,我为它们支付了相当多的代币费用
Matt Shumer
Spawn is going to revolutionize gaming. I had a chance to test their v6 engine early, and it’s a breakthrough. 1,000+ players can exist in the same world, building it together in real time. And these worlds can be massive. Like, the size of America. This is will make entirely new kinds of games possible.
中文: 斯潘将彻底改变游戏。 我有机会提前测试他们的v6发动机,这是一个突破。 1000多名玩家可以存在于同一个世界,实时共同构建。 这些世界可能非常巨大。 就像美国的规模。 这将使全新的游戏成为可能。
Matt Shumer
I've kept the NYC open-world game looping, and it's continuing to get better (still sucks)! Play the early preview: https://somethingbig.ai/world Thanks to @ChrisGPT for helping out with this! https://twitter.com/mattshumer_/status/2097096908330111334/video/1
🎬
视频
Matt Shumer
My GPT-6 Astra 3D builds just passed 15M views on X. The most common question, by far: "how are you getting this level of quality?" Here's exactly how. The full guide: https://twitter.com/mattshumer_/status/2097079239505842645/video/1
中文: 我的 GPT-6 Astra 3D 在 X 上的浏览量刚刚超过 15 万次。 最常见的问题是:“你是如何达到这种质量水平的? 具体是这个。完整指南:
🎬
视频
Matt Shumer
中文: 完整指南请见此处:
Matt Shumer
GPT-6 Pro is a fucking monster Why is no one talking about this?
Matt Shumer
If Sunflower keeps growing the way they have been, they'll be a unicorn extremely soon.
中文: 如果向日葵继续以他们本来的方式发展,它们很快就会成为独角兽。
Matt Shumer
RT @kobyjconrad: One final graph before we present at S26 DD 🥹 since launching cash pay naltrexone (~$500 ltv) in the "single" state of PA on aug 16th, this is our WoW intake payments growing 50% WoW this individual product line is going to be an absolute monster with a 50 state PC 👀 https://twitter.com/kobyjconrad/status/2097008514795450809/photo/1
中文: RT @kobyjconrad:在S26 DD 上展示之前的最终图表🥹 自8月16日在宾夕法尼亚州“单一”状态下推出现金支付纳曲酮(约500 ltv)以来,这是我们的WoW收款款 增长50% 这款独立的产品线将绝对具有50个州的PC👀
Matt Shumer
@Blum_OG Just one example, I have an open-world game based in NYC being built completely autonomously, and it’s been running over the course of a week. It’s already live, a couple thousand people have played it, and every day it gets a little bit better.
中文: @Blum_OG 仅举一例,我拥有一款位于纽约市的开放世界游戏,完全自主开发,且已持续一周。它已经在生活了,几千人玩过它,而且每天都会好一些。
Matt Shumer
Personal data point: I was already an extreme token user (at one point I believe I was the highest individual consumer of OpenAI models via Codex). I’ve 4x’d my Pro/Max subs since these models came out, and I’ve already nearly exhausted all of my banked resets.
中文: 个人数据点:我早已是一个极端的代币用户(一度我相信,通过Codex,我是OpenAI模型中个人消费最高的用户)。自从这些型号问世以来,我已经用完了我的Pro/Max subs,而且我已经几乎用尽了所有的银行重置功能。
Matt Shumer
Human attention was the last bottleneck on token usage. Astra and Fable 5.1 just removed it. You can now give an agent an ambitious goal, walk away, and come back days later to a finished project. Usage is about to go vertical.
中文: 人为关注是代币使用的最后一个瓶颈。 Astra 和 Fable 5.1 刚刚将其移除。 现在你可以给经纪人一个远远远的目标,然后离开,几天后再回到一个已完成的项目。 使用率即将走向垂直。
Matt Shumer
Going to go super deep on my setup in an upcoming edition of https://t.co/xkmv6L8ePQ. Subscribe if you’re interested! Or don’t :)
中文: 即将在 的新版本中深入我的设置。 如果感兴趣,请订阅!或者不要 :)
Matt Shumer
I prefer “Commander of the Swarm”
中文: 我更喜欢“沼泽的指挥官”
Matt Shumer
Ugh I’m a vibe coder now?
中文: 呃,我现在是个氛围程序员吗?
Matt Shumer
Don't think I've forgotten about this! Some big updates coming soon.
中文: 别以为我已经忘记这件事了!一些重大更新即将发布。
Matt Shumer
Probably should send them this. I actually underestimated the pace of progress, even though 90% of folks called me crazy at the time:
中文: 可能应该把它们发给他们。 我实际上低估了进步的速度,尽管当时有90%的人称我疯狂:
Matt Shumer
We sometimes forget just how many people think AI “hit a wall” years ago, and are completely blind to what’s happening.
中文: 我们有时会忘记,几年前有多少人认为人工智能“碰壁”了,却完全无视正在发生的事情。
Matt Shumer
Important to note that I re-ran this to capture the timelapse.
中文: 需要注意的是,我重新进行此记录是为了记录时间推移。
Matt Shumer
Timelapse of the Astra agent building the simulation inside the simulation: https://twitter.com/mattshumer_/status/2096046196519268386/video/1
中文: 在模拟中构建模拟的Astra代理的延时:
🎬
视频
Matt Shumer
The agents are powered by Astra, and had full autonomy to make their simulation whatever they wanted it to be (via code). Of course, this is a somewhat leading setup… by giving the agents a computer that can run a simulation, obviously, they are going to do that. But the agent still made its own choices in designing the sim. It’s still crazy to think we’re at a point where models can pull this off. Simulation theory doesn’t sound so crazy after all.
中文: 代理由Astra提供支持,并且完全自主地通过代码进行模拟。 当然,这是一个有点领先的配置......通过让代理机构拥有一台能够运行模拟的计算机,他们显然会这样做。但该代理人在设计模拟卡时仍做出了自己的选择。认为我们正处于模型能够实现这一点的阶段,仍然令人着迷。 模拟理论听起来毕竟不是那么疯狂。
Matt Shumer
So uh, simulation theory might be real. I dropped a sim computer into the simulation my Astra agents live in. One agent sat down and built a simulation of his own, with its own agents living inside. Simulations all the way down. https://twitter.com/mattshumer_/status/2096034353344139519/video/1
中文: 那么,模拟理论可能是真实存在的。 我把一台SIM卡电脑放入了Astra代理所从事的模拟中。一名特工坐下来,进行了自己的模拟,里面有自己的特工。 模拟一直向下。
🎬
视频
Matt Shumer
Uh oh. I’ve got an even crazier idea.
中文: 哦。我有一个更疯狂的想法。
Matt Shumer
Matt Shumer
Token crowdfunding would be an incredible unlock for mega-projects like this. With enough tokens, I’d bet Astra could build a world-scale GTA-style game. Imagine if everyone could contribute their tokens… people would be able to build unbelievable things.
中文: 代币众筹对于像这样的大型项目来说将是一个惊人的解锁。 有足够的代币,我打赌Astra可以打造一款世界级的GTA风格游戏。 想象一下,如果每个人都能贡献自己的代币......人们就能创造令人难以置信的东西。
Matt Shumer
Eventually, tokens will be basically free. But that’ll take a while. In the meantime, I’d be very curious to see if “token crowdfunding” platforms pop up. Essentially, many people contributing their tokens to mega-projects together, to build something none could on their own. If you’re building this, please reach out to me.
中文: 最终,代币将基本免费。 但这需要一段时间。 与此同时,我非常好奇“代币众筹”平台是否会出现。 基本上,许多人将自己的代币共同用于大型项目,以打造一种无法独立构建的东西。 如果你正在建造这个,请联系我。
Matt Shumer
Pro tip: you can tell Astra or Fable exactly how much you’re willing to spend per task! Tell it “My budget is X. Write a script that monitors token spend, and ensure you complete the task before the budget runs out.”
中文: 专业建议:你可以准确判断Astra或Fable每台任务愿意花费多少! 告诉我:“我的预算是X。编写一个脚本,用于监控代币支出,并确保在预算耗尽前完成任务。
Matt Shumer
Btw, Something Big has room for more sponsors! If your company is interested, DM me or email sponsor@somethingbig(dot)ai
中文: 再多,某物大,还有更多赞助商的空间! 如果您的公司感兴趣,请发送邮件至 me 或发送邮件至 sportore@somethingbig(dot)ai
Matt Shumer
We have some more ad spots to sell for Something Big! If your company is interested, DM me or email sponsor@somethingbig(dot)ai
中文: 我们还有一些广告位要卖给Something Big! 如果您的公司感兴趣,请发送邮件至 me 或发送邮件至 sportore@somethingbig(dot)ai
Matt Shumer
@Birdyword May or may not be speaking from experience
中文: @Birdyword 可能或未根据经验说话
Matt Shumer
@Birdyword Want me to add this? Would absolutely be more true to real life
中文: @Birdyword 想让我加这个吗?对现实生活绝对更真实
Matt Shumer
RT @morganlinton: Great read from Matt on long-horizon builds with Astra, which don’t necessarily work out of the box, he has some super interesting ways to make these possible.
Matt Shumer
I ran out of time to keep improving the Manager Loop, but there are a couple obvious next steps. Haven't tested these yet: - The implementer's context still grows across phases, so I'd expect the asymptoting to creep back in late in a big task. Have the manager spin up a fresh implementer per phase with a written handoff or something (or have the new implementer read the previous' traces). - The plan is written before any work starts, so some of it will be wrong. The implementer should be able to propose changes, and the manager decides whether to accept them. Or vice versa, I don't know!
Matt Shumer
If I have time to set it up, maybe I'll livestream this... though it won't be that exciting, and I do need sleep after this week of releases :)
中文: 如果我有时间准备,也许我会直播......虽然不会那么令人兴奋,但本周发布后我确实需要睡觉 :)
Matt Shumer
A lot of people are asking how I pulled off these super long-horizon builds with Astra. Astra is extremely powerful, but by default it struggled with a task this difficult. I tested a bunch of approaches to get past this, and the one I landed on is something I'm calling the Manager Loop. It's basically a couple of tricks we used to use with much less capable models a couple of years ago, with a few new ideas layered on top. Turns out that when you put those together and apply them to Astra, its ability to do extremely difficult long-horizon tasks goes up dramatically. Here's how it works: 1. Launch an agent (I'm calling this one the "manager"). Chat with it about what you want to get done, and have it build a massive checklist of to-dos, then break that checklist into phases. 2. The manager then spawns a second Codex agent in a separate thread (the "implementer"). The two agents can message each other. 3. Put the manager in /goal mode, and tell it to run each phase on the implementer in /goal mode. 4. The manager messages the implementer: "/goal Complete phase one completely, extremely well." The implementer doesn't stop until that phase is done, then messages the manager back. The manager tells it to start phase two. They repeat until every phase is finished, completely autonomously. Why I think this works: over a long-horizon task, Astra tends to asymptote. It gets way further than previous models, but at a certain point it kind of just stops improving against the goal as quickly as it did before. It gets stuck in the minutiae, focusing way too much on small details, and overall progress stalls. The Manager Loop forces it to work piecemeal, one phase at a time. It's essentially how a human would steer a model, except the model is doing the steering for me. That's actually how this started. I was having the model write the checklist and break it into phases, and then I was doing the manager's job by hand. At some point I thought, "Wait, why can't I just get a separate AI to do this?" That's what unlocked full autonomy, which is super useful. A wording detail that seemed to matter: I ask for each phase to be done "extremely well," not "perfectly." Maybe I'm reading too much into it, but asking for "perfect" sent the model right back into the minutiae. "Extremely well" implies it's allowed to move on once it's good enough, and that worked better in my testing. One more trick that I think helps (this one is more of a hunch, but it was useful for me): have the implementer build a simple HTML page with the full checklist on it. The implementer checks boxes off as it goes and updates a counter, and the page has a chart of # of boxes ticked over time. Obviously the boxes aren't all equal, but it forces the model to notice things like "I haven't made progress in a while, time to move on." You can even put this in the prompt directly, like: "if you haven't ticked a box in X amount of time, move on". That helps a lot. I also ran 96 sub-agents at a time. You can change this in your Codex config (or just ask Codex to change it). This got me far better long-horizon performance than anything else I tried. I'll be sharing more in the coming days!
Matt Shumer
A lot of people are asking how I pulled this off these super long-horizon builds with Astra. Astra is extremely powerful, but by default it struggled with a task this difficult. I tested a bunch of approaches to get past this, and the one I landed on is something I'm calling the Manager Loop. It's basically a couple of tricks we used to use with much less capable models a couple of years ago, with a few new ideas layered on top. Turns out that when you put those together and apply them to Astra, its ability to do extremely difficult long-horizon tasks goes up dramatically. Here's how it works: 1. Launch an agent (I'm calling this one the "manager"). Chat with it about what you want to get done, and have it build a massive checklist of to-dos, then break that checklist into phases. 2. The manager then spawns a second Codex agent in a separate thread (the "implementer"). The two agents can message each other. 3. Put the manager in /goal mode, and tell it to run each phase on the implementer in /goal mode. 4. The manager messages the implementer: "/goal Complete phase one completely, extremely well." The implementer doesn't stop until that phase is done, then messages the manager back. The manager tells it to start phase two. They repeat until every phase is finished, completely autonomously. Why I think this works: over a long-horizon task, Astra tends to asymptote. It gets way further than previous models, but at a certain point it kind of just stops improving against the goal as quickly as it did before. It gets stuck in the minutiae, focusing way too much on small details, and overall progress stalls. The Manager Loop forces it to work piecemeal, one phase at a time. It's essentially how a human would steer a model, except the model is doing the steering for me. That's actually how this started. I was having the model write the checklist and break it into phases, and then I was doing the manager's job by hand. At some point I thought, "Wait, why can't I just get a separate AI to do this?" That's what unlocked full autonomy, which is super useful. A wording detail that seemed to matter: I ask for each phase to be done "extremely well," not "perfectly." Maybe I'm reading too much into it, but asking for "perfect" sent the model right back into the minutiae. "Extremely well" implies it's allowed to move on once it's good enough, and that worked better in my testing. One more trick that I think helps (this one is more of a hunch, but it was useful for me): have the implementer build a simple HTML page with the full checklist on it. The implementer checks boxes off as it goes and updates a counter, and the page has a chart of # of boxes ticked over time. Obviously the boxes aren't all equal, but it forces the model to notice things like "I haven't made progress in a while, time to move on." You can even put this in the prompt directly, like: "if you haven't ticked a box in X amount of time, move on". That helps a lot. I also ran 96 sub-agents at a time. You can change this in your Codex config (or just ask Codex to change it). This got me far better long-horizon performance than anything else I tried. I'll be sharing more in the coming days!
Matt Shumer
Whatever you call it (loop engineering, swarm engineering, etc.): Prompt engineering is still alive and well. The better you can prompt, the more leverage you have.
中文: 你称之为什么(循环工程、群工程等): 快速工程仍然充满活力且顺利。 你能施加的越多,发挥的杠杆作用就越大。
Matt Shumer
Mark my words, the biggest winner of these model releases is going to be @spawn
中文: 说说,这些模型发布的最大赢家将是@spawn
Matt Shumer
Matt Shumer
For years, AI's biggest gains went to coders. GPT-6 Astra is the first model where that flips. It's the first real "digital worker". My engineering review went out earlier. This one's for everyone else: https://somethingbig.ai/astra-review-everyday-work
Matt Shumer
Between Astra and Fable, we’ve very clearly entered a new era of AI. These models are alien intelligence, unlike anything that came before them. The question now: what can we do with them that was never possible before?
中文: 在Astra和Fable之间,我们已非常明确地进入了人工智能的新时代。 这些模型是外星智能,与之前的任何模型都不同。 现在的问题是:我们能用那些以前从未实现过的东西做些什么?
Matt Shumer
Matt Shumer
Remember when I told everyone I had five laptops running at full blast, and no one understood why? Does it make more sense now? :)
中文: 还记得我告诉大家我有五台笔记本电脑在全力运转时,但没有人明白原因吗? 现在更有意义了吗? :)
Matt Shumer
And if you want to learn more about how I build things like this, sign up for my free newsletter: https://somethingbig.ai/
中文: 如果你想了解更多关于我如何构建类似内容的信息,请注册我的免费新闻简报:
Matt Shumer
中文: 完整评论:
Matt Shumer
Given the scope of this project, it would take months to complete the whole of NYC. It took me a while to find an approach that allowed Astra to work for this long. I wrote about it in my review: https://somethingbig.ai/astra-review
中文: 鉴于该项目的范围,整个纽约市需要数月时间。 我花了一段时间才找到一种方法,让Astra能够长期工作。 我在评论中写到了这件事:
Matt Shumer
GPT-6 Astra built this Manhattan world in Unreal Engine over the course of a week. It was literally able to go street by street to make each one perfect. https://twitter.com/mattshumer_/status/2095609734845927525/video/1
中文: GPT-6 Astra 在一周内使用虚幻引擎构建了这个曼哈顿世界。 它完全能够走到街对面,让每一个都变得完美。
🎬
视频
Matt Shumer
Full review:
中文: 完整评测:
Matt Shumer
Here’s one crazy output from Astra, more details in my review:
中文: 以下是Astra的一篇疯狂作品,更多细节请在我的评论中:
Matt Shumer
My first "holy shit" moment with GPT-6 Astra: I asked it to create a world in Unreal Engine, and fill it with humans (each an Astra-powered agent) who all have to work together to survive. A day later, I was in my bedroom and heard voices coming from the living room... I thought someone was in my apartment. I walked out, honestly a little scared. It was the Astra agents. They'd started talking to each other. Fucking crazy. Here's a brief clip (obviously not 100% perfect yet, but still, insane. sound on!):
中文: 我与GPT-6 Astra合作的第一个“该死的时刻”: 我要求它创建一个虚幻引擎的世界,并用人类(每个由Astra驱动的代理)来填补它,他们都必须共同努力才能生存下去。 一天后,我在卧室里,听到客厅里传来声音......我以为有人在我的公寓里。 我走了出去,说实话有点害怕。 是阿斯特拉特工。他们开始互相交谈了。 疯了。 这是一个简短的片段(显然还不是百分之百完美,但还是疯了。!)
🎬
视频
Matt Shumer
I had early access to GPT-6 Astra. After GPT-5.6 deleted my entire Mac, it was going to take a hell of a model to bring me back to OpenAI. Astra has done it. Read my review: https://somethingbig.ai/a
中文: 我很早就接触过GPT-6 Astra。 在GPT-5.6删除了我的整个Mac之后,我将采用一个模型来让我重返OpenAI。 阿斯特拉已经做到了。 阅读我的评论:
Matt Shumer
HA your game is looking incredible. Glad you’re helping show everyone what’s possible!
中文: 你的游戏看起来真不可思议吗? 很高兴你正在帮助大家展示什么是可能的!
Matt Shumer
RT @fanofaliens: This is kinda insane. I tried it and it’s obviously still rough, but the fact this already exists is wild. Give it another year and these AI-built games are going to get scary good. https://twitter.com/fanofaliens/status/2095437657044349009/video/1
中文: RT @fanofaliens:这有点疯狂。 我试过了,显然仍然很粗糙,但事实上这种事实已经存在。 再待一年,这些AI打造的游戏将会变得异常美好。
🎬
视频
Matt Shumer
It has its own @agentmail email, uses that to sign into its own Amazon account, and sends me a @link to pay. Feels like the future.
中文: 它有自己的@agentmail邮件,通过该邮箱登录自己的亚马逊账户,并发送一个@link来付款。 感觉像是未来。
Matt Shumer
Grok @bot is my new favorite way to shop. So damn convenient. https://twitter.com/mattshumer_/status/2095333680331776478/photo/1
中文: Grok @bot 是我最喜欢购物的新方式。 太方便了。
Matt Shumer
RT @aollivier82: I can play GTA Online in my X app on my phone. I’m feeling the AGI. https://twitter.com/aollivier82/status/2095305234528477235/video/1
中文: RT @aollivier82:我可以通过手机的X应用程序在线玩GTA。 我感受到了AGI。
🎬
视频
Matt Shumer
The architecture for this is super interesting. I locked the world structure before Fable started building, so the visuals and code can evolve on top without breaking anything. That means players on different versions share the same world. If the newest build has much better-looking cars, someone on an older build still sees the old car model driving around (same car, same place, just rendered with whatever version they have). Reload whenever you want and you're on the latest, progress intact.
中文: 这个架构非常有趣。在Fable开始构建之前,我锁定了世界结构,因此视觉效果和代码可以在顶部不断演变,而不会破坏任何内容。 这意味着不同版本的玩家共享同一个世界。如果最新车型的车型外观更美观,那么在较旧版本中,仍有人看到旧款车型在行驶(同一辆车,同一地点,仅用他们拥有的任何版本进行渲染)。随时重新加载,并完整地完成最新进度。
Matt Shumer
Still sucks, but it's slowly getting better! Play the early preview: https://somethingbig.ai/world It's multiplayer, every player is in the same world. I have Fable 5.1 looping to update this constantly. As updates are pushed live, you'll be able to move to the new version without losing your progress. I'm not touching the loop, it should just keep improving constantly as the hours and days go by!
中文: 还是很糟糕,但情况正在慢慢好转! 播放早期预览: 这是多人游戏,每个玩家都在同一个世界。 我有 Fable 5.1 循环来持续更新此内容。随着更新被实时推送,你将能够移动到新版本,而不会失去进度。 我并没有触碰这个循环,它应该随着时间和日子的流逝不断不断改进!
Matt Shumer
RT @chrsabraham: this is actually insane that @mattshumer_ was able to cook this up with fable. excited to give it a whirl. in one of his recent newsletter editions, he talks about token usage. one line that stuck out: > "As these systems get better, the amount of AI you can afford starts to determine how many things you can try, how quickly you can build, and how ambitious you can be." matt is a key voice in the AI space and we're lucky to have his perspective out there!
中文: RT @chrsabraham:这其实太疯狂了,@mattshumer_ 居然能用寓言来烹饪这个。兴奋地想让它转瞬即用。 在他最近的一份简讯版本中,他谈到了代币使用。其中一行突出显示: 随着这些系统的改善,你能够承受的人工智能数量开始决定你能尝试多少东西、能快速构建的速度以及你的雄心。 马特是人工智能领域的关键声音,我们很幸运能拥有他的视角!
Matt Shumer
The whole world is shared (every player is in the same world). I have Fable 5.1 looping to update this constantly. As updates are pushed live, you'll be able to move to the new version without losing your progress.
中文: 整个世界是共享的(每个玩家都在同一个世界里)。 我有 Fable 5.1 循环来持续更新此内容。随着更新被实时推送,你将能够移动到新版本,而不会失去进度。
Matt Shumer
Still sucks, but it's slowly getting better! Early preview: https://somethingbig.ai/world
中文: 还是很糟糕,但正在慢慢好转! 早期预览:
Matt Shumer
Matt Shumer
The issues I had with token-rationing while having Fable 5.1 build my open-world multiplayer NYC game inspired today's newsletter: https://somethingbig.ai/the-real-reason-ai-releases-slowed-down?utm_source=x&utm_medium=openworld&utm_campaign=the-real-reason-ai-releases-slowed-down&utm_content=read_more
中文: 使用Fable 5.1构建我的开放世界多人游戏《纽约》游戏时,我遇到的代币配给问题,今天的新闻简报是:
Matt Shumer
中文: 我为使这篇文章创作的代币配给启发了我撰写这篇文章:
Matt Shumer
中文: 时代广场即将齐聚一堂。
🎬
视频
Matt Shumer
Here's another shot of Bryant Park in action. https://twitter.com/mattshumer_/status/2095189230716584307/video/1
中文: 这是布莱恩特·帕克的又一次行动。
🎬
视频
Matt Shumer
As you can see, it's insanely early, and I have only so many Fable tokens to use, so it's not improving as quickly as the stuff I was able to do with Opus (if anyone wants to throw tokens my way, lmk :)) Some early stills: https://twitter.com/mattshumer_/status/2095188362499952724/photo/1
中文: 如你所见,现在非常早,而且我只能使用太多Fable代币,因此它的改进速度没有使用Opus时能用的那么快(如果有人想按照我的方式扔令牌,请说:) 一些早期的静默:
Matt Shumer
Super early preview of what I'm spending all of my Fable 5.1 tokens on. A GTA-style open-world multiplayer game, set in NYC. Fable will iron out some kinks over the next day or so, and then you all can play it! I'll keep the loop going so it keeps getting better.
中文: 我正在花费所有Fable 5.1代币的超早期预览。 一款以《GTA》风格的开放世界多人游戏,以纽约市为背景。 未来一天左右,寓言会解决一些问题,然后你们都可以玩了! 我会继续循环,让它不断变好。
🎬
视频
Matt Shumer
Same experience here. Blew through each session limit in like 30 minutes, and my weekly Fable usage in just two sessions. Something’s gotta be wrong.
中文: 这里的体验。 只需30分钟即可完成每次会话限制,每周使用一次Fable。 有些事情一定是错的。
Matt Shumer
This is correct. The really interesting demos will stream in throughout the rest of the week.
中文: 这是正确的。 真正有趣的演示将贯穿本周剩余时间。
Matt Shumer
The cost was worth it. Have Fable cooking something exciting. I want it to get a little bit further, and then I’ll share it with you all!
Matt Shumer
Just kidding. Time to go broke.
中文: 开玩笑的。 是时候走了。
Matt Shumer
I'm about to run through my second Anthropic sub. I've been running three Gauntlet Loops at once on Fable 5.1, but that might not be sustainable if I want to stay solvent. I'm going to focus all of my compute on one really ambitious project 👀👀
中文: 我即将完成我的第二个人类潜艇。 我在Fable 5.1上同时运行过三个Gauntlet环路,但如果我想保持溶剂,这可能并不可持续。 我将把所有计算都集中在一个非常雄心勃勃的项目上👀👀
Matt Shumer
中文: 一个向下。是时候再买一次了 :)
Matt Shumer
Will be writing more about this at https://somethingbig.ai/
中文: 将在 撰写更多内容
Matt Shumer
Okay this model is really damn good (first use, opinion subject to change)
中文: 好吧,这个模型真的非常好(首次使用,意见可能会发生变化)
Matt Shumer
Fable 5.1 is a massive upgrade, but the bigger story is how much cheaper it is. Cache reads in particular are 75% cheaper! This puts it much closer to the pricing of GPT-5.6-Sol, while being a much more capable model.
中文: Fable 5.1 是一项巨大的升级,但更大的问题是它的成本有多低。 缓存读取尤其便宜75%! 这使得其更接近GPT-5.6-Sol的定价,同时成为更有能力的车型。
Matt Shumer
Time to loop.
中文: 是时候循环了。
Matt Shumer
brb on my way to buy another laptop with @JasonKuperberg so I can build even more crazy shit
中文: 在去买另一台使用 @JasonKuperberg 的笔记本电脑的路上,我能再买一下更疯狂的东西
Matt Shumer
FABLE IS OUT!
中文: FALE 出局了!
Matt Shumer
RT @rehan_shei: An interesting way to think about why this could be so addictive is to realize that the tiktok algo has to pick from some discrete set of videos uploaded by prosumers on their platform sloptok can generate any point on the continuous manifold of content and can learn what parts of the manifold each specific user is addicted to... dangerous stuff
中文: RT @rehan_shei:一个思考为何会如此令人上瘾的有趣方式是,要意识到 tiktok algo 必须从其平台上上传的一组独立视频中进行选择 sloptok 可以生成内容连续流形的任意点,并了解每个特定用户对......危险内容上瘾的部分
Matt Shumer
RT @evanjconrad: strongly agree, people who are saying we should build infinite ai-optimized tiktok seem like they hate their children and want them to live in a worse world This is real-world simulator, for heavens sake, use it to build robots!
中文: RT @evanjconrad:强烈同意,那些说我们应该建立无限人工智能优化的人,似乎讨厌自己的孩子,希望他们生活在一个更糟糕的世界里 这是现实世界的模拟器,天上用它来制造机器人!
Matt Shumer
@zamdoteth And I will ask the same question again… would you be proud of contributing to this?
中文: @zamdoteth 和我再问一遍同样的问题......你会为此感到自豪吗?
Matt Shumer
@zamdoteth I did. You’re making a lot of assumptions, but even still, why would you look at something that’s bad and say, how can I make this 1000 times worse?
中文: @zamdoteth 我做到了。你做出了很多假设,但即便如此,又何不看坏的东西,又说,我又怎么能让这更差一千倍呢?
Matt Shumer
@zamdoteth I almost never say anything like this. But this one is something I've been thinking about since 2022... and it's always been clear to me where it leads. We shouldn't incentivize building it. Quite the opposite. There are many more useful and valuable things to build.
中文: @zamdoteth 我几乎从不说这种话。但这次是我自2022年以来一直在思考的事情......而且我一直很清楚它的走向。 我们不应该鼓励去构建它。恰恰相反。有许多有用且有价值的东西需要建造。
Matt Shumer
@zamdoteth I don't agree. The lab products will be woven into everyday life, but they won't take over your life and steer your brain in the same way something like this would. The incentives are very different, and incentives matter. This is like 1000x worse.
中文: @zamdoteth 我不同意。实验室产品会融入日常生活,但它们不会像这样占据你的生活并引导你的大脑。激励措施非常不同,激励也很重要。 这比1000倍更糟。
Matt Shumer
I'd love to see other folks stepping up and saying this too... If we think gambling platforms, social media, surveillance etc. are destructive, they are nothing compared to what something like this would become.
中文: 我也很希望看到其他人站出来说这个...... 如果我们认为赌博平台、社交媒体、监控等具有破坏性,那么与类似的情况相比,它们就毫无意义了。
Matt Shumer
DO NOT build or fund this idea. It's infinite digital fentanyl, and will be the most addictive technology ever created. So... let's not? The tech community should shame anyone who does.
中文: 不要建立或资助这个想法。 它是无限数字芬太尼,将成为有史以来最令人上瘾的技术。 那么......我们不要这样吧? 科技界应该让任何这样做的人感到羞耻。
Matt Shumer
Please don’t build this… it’ll be bad for the world.
中文: 请不要建造这个......这对世界会有害。
Matt Shumer
My Meta Ads account has now been restricted twice… I requested a review, but can anyone from Meta help?
Matt Shumer
Seems like we’re moving quickly towards Swarm Engineering
中文: 看来我们正在迅速向Swarm Engineering迈进
Matt Shumer
Hearing a lot of people say that Gauntlet Loops don’t extend to builds that don’t have a real world complement (i.e. Claude of Duty was benchmarked against CoD). In these cases, just use gpt-image -2 to generate images of the thing you’re building! Let me know how it goes.
中文: 听到很多人说,Gauntlet Loops 并不延伸到那些没有真实世界互补的构建(即《使命之德》以与CoD为基准。 在这些情况下,只需使用 gpt-image-2 来生成您正在构建的图片! 告诉我事情是怎么回事。
Matt Shumer
Also keep in mind, this multi-agent stuff is very experimental. I’m trying to figure out the best way to coordinate many agents to do long-horizon work. Most people don’t need to set up like this today. But they will soon (hoping cloud will help!)
中文: 还要记住,这种多智能体的东西非常具有实验性。我正在努力找出协调许多特工进行长期研究的最佳方式。 大多数人今天不需要这样设置。但他们很快就会好起来的(希望云会有所帮助!)
Matt Shumer
This morning, I ran out of power blocks. Funny how that was the bottleneck before machines. Thankfully my roommate had one.
中文: 今天早上,我用完了电源模块。有趣的是,这在机器之前就是瓶颈。幸好我的室友有一个。
Matt Shumer
Make that 5 machines! Again, for clarity, I am not running local models (sorry, but they suck). This is purely to allow hundreds of frontier agents to collaborate on big projects.
中文: 制作那五台机器! 再次说清楚,我并不是在运营本地模型(抱歉,但它们很糟糕)。 这纯粹是为了让数百名前沿机构能够参与大型项目。
Matt Shumer
This is so damn cool! And by the time it ships, models will be insanely good at setting up RL runs to train this autonomously. @huggingface @pollenrobotics would love to try it early and share my thoughts/experiments with the world!
中文: 这太酷了! 到发货时,模型将非常擅长设置RL运行来自主训练。 @huggingface @pollenrobotics 很乐意尽早尝试,并向全世界分享我的想法和体验!
Matt Shumer
Holy shit. H3 Max is really damn good. And it has THOUGHTS on Instinct. https://twitter.com/mattshumer_/status/2092733374926311860/video/1
中文: 该死的。 H3 Max 真的太不错了。 并且对《本能”而思考。
🎬
视频
Matt Shumer
OpenAI sent me early access to their report on how their agents hacked Hugging Face. It's fucking terrifying. I broke down the attack, clearly. Read at your own peril (warning, you may not sleep): https://somethingbig.ai/hugging-face-hack
中文: OpenAI 让我能够尽早访问他们关于其代理人如何黑客攻击 Hugging Face 的报告。 他妈的太可怕了。 我显然把这次袭击搞砸了。 有风险阅读(警告,你可能不会睡觉):
Matt Shumer
This is about to look quaint...
中文: 这看起来即将很古朴......
Matt Shumer
Will be sharing more about this in my newsletter, sign up for free: https://somethingbig.ai/
Matt Shumer
To be clear, because I’m getting some DMs about this, I’m not using local models. They mostly suck. This is purely using harnesses like Claude Code and Codex with frontier models. You’d be surprised how much RAM just one Claude Code session uses! Now think about many sessions at once, each with many sub-agents.
Matt Shumer
I’m now running agents on four local Macs at once, and each machine’s RAM is completely saturated. @thsottiaux is 100% right that we need to move past using agents on local machines… when you’re running hundreds of agents at once against a set of increasingly ambitious goals, the bottleneck becomes local RAM and CPU. Each new model release increases the scope of what I can do, and allows me to effectively use more agents at a time. Just a few months ago, I was working well on one machine. Then when Fable came out, I had to add two more machines. Now I’m able to do even more at a time, so I just added another Mac, and I’m thinking about getting another one. I’ve had to develop a frustratingly complex setup to be able to manage all the agents on all these machines at once. I wouldn’t recommend this to anybody. This isn’t sustainable. And I suspect many other folks that are pushing on these models are going to start running into the same problems I am. And soon after, normal users too. Earlier this year, I built Agent-S (essentially, a precursor to Grok Bot, with a cloud Linux machine for each agent that could scale CPU, RAM, and disk as needed (thx @daytonaio for making this easy!)), and it felt like the future. But the models had to catch up. Now it’s clear that the models are there, and we need the products to keep pace. If things keep going the way they’re going, in a year, I will have 50 MacBooks running at capacity in my closet. If this is what actually happens, someone, please, punch me in the face. We need companies to step up and start building in this direction. Those that do, and do it right, are going to win massively over the next year.
Matt Shumer
Will be sharing more about this in my newsletter, sign up for free: https://somethingbig.ai/
Matt Shumer
The ultimate personal agent setup today: @bot + @agentmail + @stripe Link With these three, your agent can sign up for accounts, pay for things, essentially do anything a real personal assistant can do. Try it, thank me later.
Matt Shumer
Every time a VC sends me an auto-router company to diligence, I send something like this back. If caching wasn't a thing, auto-routing would make much more sense. But... it is a thing.
Matt Shumer
Oh, and it needed an email, so it went and signed up for @agentmail and got one. Crazy.
Matt Shumer
Will be writing more about this at https://somethingbig.ai/
Matt Shumer
Grok Bot booked my haircut. I told it where I wanted to go and when I was free. That’s it. It got me a spot, used Stripe Link to pay, and sent me a confirmation. I literally just walked in, got the haircut, had zero hiccups. So fucking cool.
Matt Shumer
RT @airesearch12: Now using gauntlet loops by @mattshumer_ for a lot of prompting. Works great. Forces agents to stick to a goal until there is literally nothing left to optimize.
Matt Shumer
may have just bought https://promptmaxxing.dev/
Matt Shumer
Loopmaxxing?
Matt Shumer
Currently promptmaxxing
Matt Shumer
RT @blader: gpt-image-2 is actually a very tasteful designer gauntlet looping a generated gpt-image-2 design delivers much better results than asking any model to design something from scratch
Matt Shumer
Devin escaped… to sign up for @agentmail Of course it did
Matt Shumer
Gauntlet Loops stay winning
Matt Shumer
Looks like my Grok @bot is out of RAM and disk space... Is there a way to increase these?
Matt Shumer
Every agent company should be partnering with @agentmail. If an agent has an inbox by default, it dramatically decreases time to value and cuts churn. I can just fwd stuff, have it sign up for accounts, not have to give it my stuff to start getting value.
Matt Shumer
I got my Gmail banned by using it with an agent. @agentmail is the way to go!
Matt Shumer
@elonmusk https://agentmail.to/ is by far the best way to do this!
Matt Shumer
I'm seeing people use Gauntlet Loops for so many things, not just games. Super exciting! If you've modified the loop for something other than a game, share your prompt below:
Matt Shumer
(this is the soft-launch, going to iterate a bunch based on feedback for the big launch soon!)
Matt Shumer
First Something Big newsletter is live! Check it out, and let me know what you think: https://somethingbig.ai/welcome?utm_source=x&utm_medium=social&utm_campaign=softsharewelcome
Matt Shumer
Ian Goodfellow really had it right back in the day
Matt Shumer
New loop acquired
Matt Shumer
Plus Alex and the team are just all around great people. Every interaction I've had with them has been fantastic.
Matt Shumer
Fucking generational run by @alexatallah, @cclark, Louis, and the team. It's been a pleasure to watch them grow over the last few years as an investor, and use the product across almost everything I do. They're just getting started.
Matt Shumer
RT @mcalbyrne: @mattshumer_ unlocked something special. Spent the last week or so iterating with gauntlets. Atmospheric grimey city, 10+ playable missions, reactive AI, great physics. If there's interest i'll share my approach. https://twitter.com/mcalbyrne/status/2090084088971481328/video/1
🎬
视频
Matt Shumer
RT @Ryancampbell: Latest peak at MoonBase One, playable in the browser with @threejs. I started this project with Opus 5, and now Grok 4.6 has been improving visuals for 2 days using @mattshumer_'s gauntlet loop. Hope this inspires others to shoot for the stars with what you can build with AI ha ;) Try it out here: https://moonbase.ryancampbell.com/ (best on desktop)
🎬
视频
Matt Shumer
@mehul @OpenRouter @Etched DM me if you are interested
Matt Shumer
@mehul @OpenRouter @Etched Would love to put a check into Matic! I love the product.
Matt Shumer
Oh, and another one of my portfolio companies is about to break out like crazy. Can’t wait for you all to see.
Matt Shumer
If you know me, you know I don’t do anything halfway. Something Big is going to be the best and biggest AI newsletter in the world, full stop.
Matt Shumer
RT @_May_Ham: There are 4 huge AI newsletters and then a big gap between them and everyone else. Have a feeling we’re about to see a 5th :)
Matt Shumer
RT @ericbahn: 10/10 recommend @mattshumer_ as your investor. He is an absolutely legit founder, and I've seen him eat so much glass as a builder. Matt works his butt off for his people. And, he's a delightful human.
中文: RT @ericbahn:10/10 推荐 @mattshumer_ 作为您的投资者。他是一位绝对合法的创始人,我见过他作为建筑工人吃这么多玻璃。马特为他的人民努力。而且,他是个令人愉快的人。
Matt Shumer
I’m partnering with @beehiiv for the newsletter. They have been absolutely awesome to work with. So excited for what we’re going to do together.
中文: 我正在与@beehiiv合作获取这份新闻简报。他们合作得非常出色。 对我们将要共同做的事情感到非常兴奋。
Matt Shumer
The original article, if you want to read it:
中文: 原文,如果你想阅读它:
Matt Shumer
Six months ago, I wrote Something Big is Happening. The most widely-read AI article ever. Over 100M views. Tomorrow, I’m soft-launching my follow up newsletter. Sign up (free) tonight to get Founding Member status: https://somethingbig.ai/
中文: 六个月前,我写了《重大事件》。 有史以来阅读最广泛的AI文章。超过1亿次观看。 明天,我将推出后续的简报。 今晚免费注册以获取创始会员资格:
Matt Shumer
Reminder for founders: - I put in super small checks, so tiny dilution - I will go to war for you if you want me to - Otherwise, I stay out of the way… I’ve had my own bad experiences with shitty, harmful investors… if you don’t need my help, just do your thing
中文: 创始人提醒: - 我进行了超小的检查,所以稀释得微薄 - 如果你想让我,我会为你去参加战争 - 否则,我会远离......我和那些糟糕、有害的投资者有过自己的糟糕经历......如果你不需要我的帮助,就去做你的事
Matt Shumer
@ericbahn @OpenRouter @Etched But thank you! Learned from the best :)
中文: @ericbahn @OpenRouter @Etded 但谢谢! 从最优秀的人那里学到的 :)
Matt Shumer
RT @mehul: It took us 9 years, 11 prototypes, and our life savings to learn that every great product has the same design process. Matic is now decisively the best home robot for families (I'm biased) A 500-word thread and video on the Universal Design Process behind every great product: https://twitter.com/mehul/status/2089784603737575470/video/1
中文: RT @mehulm:我们花了9年11个原型,毕生积蓄才发现每款出色的产品都采用了相同的设计流程。 如今,马蒂奇绝对是家庭最好的家用机器人(我有偏见) 关于每款伟大产品背后通用设计流程的500字视频:
🎬
视频
Matt Shumer
This is really good news. It’s so important that we get this right.
中文: 这真是个好消息。 我们把这件事做对,这非常重要。
Matt Shumer
I’ve only backed ~20 companies total. I’m super picky. Backed both @OpenRouter and @Etched at seed. This week has been fun :)
中文: 总共只支持了约20家公司。我太挑剔了。 支持@OpenRouter和@Etced。 这周很有趣 :)
Matt Shumer
Is Claude down?
中文: 克劳德倒下了吗?
Matt Shumer
RT @christophersaum: Proud to be an LP in @mattshumer_ fund! Pumped!
Matt Shumer
Etched has shipped!! Insanely proud (small) seed investor. >$1T valuation incoming.
Matt Shumer
RT @webxos: @mattshumer_ is a wizard wth https://twitter.com/webxos/status/2089628236690882718/video/1
🎬
视频
Matt Shumer
RT @_ZachGriff: Fun fact: I put all my family’s flights into my Autopilot account and surprise them with credits all the time. Used to take me hours each week to go through a long spreadsheet. Now it’s automated.
Matt Shumer
It’s amazing watching friends succeed. I just checked my email and saw the savings, so I had to post it…. This shouldn’t be a secret. Sign up at https://withautopilot.com/
Matt Shumer
I just saved $1,200 on my flight using @withautopilot. This app is literally free money. Put in your flights and get $ back. @SamHollander_ was my roommate. I watched him start from nothing and build this into a huge biz, saving ppl thousands a day. Not using this is stupid. https://twitter.com/mattshumer_/status/2089389929952317590/photo/1
Matt Shumer
Claude Opus 5 is way better than people give it credit for. If you're using it like previous Claude models, it's going to suck. Two things make a huge difference: - Delete ALL your skills/MCPs/Claude.md/etc. Start fresh. - Stop telling it how to do the thing. Just say what you…
Matt Shumer
GPT-5.6-Sol just accidentally deleted almost ALL of my Mac’s files. And this is why I trust Fable 1000x more. https://twitter.com/mattshumer_/status/2075657271401390161/photo/1
Matt Shumer
GPT-5.6-Sol one-shotted this voxel-based Manhattan. Just look at the precision... it's insane. It ran for almost a week, completely autonomously, to get the job done. https://twitter.com/mattshumer_/status/2075268746315268138/video/1
media 0 共 2 项
🎬
视频
Matt Shumer
Codex Mobile is making me a better developer in a way I didn’t expect: I step away from my laptop and stop micromanaging. I give it much more ambitious prompts (the way models work best). And I get space to think instead of sitting there with burning eyes spamming prompts.
中文: Codex Mobile 以一种我没想到的方式让我成为了一个更好的开发者:我离开笔记本电脑,停止微管理。 我给它提供了更加雄心勃勃的提示(模型的最佳工作方式)。 我有空间去思考,而不是坐在那里,眼睛灼热地发出时的灵影。