Recent posts
I think its easy to say "Simulation is a new scaling law" and treat it as marketing hyperbole, but midway along this interview you can hear me go from somewhat shitposting to very very serious. I am 2 years late to this but finally understand why @karpathy and @drfeifei backed @joon_s_pk @msbernst @percyliang et al - Smallville at the time had zero commercial applications, but if you take RSI seriously, from models automating increasingly large parts of ML research and AI engineering, the last* barrier is simulating humans and human feedback, and Simile is obviously the team to do this and already finding PMF at Fortune 100s even at this early stage. I've never been so happy to be so wrong. *or second last ! :) more soon on the Science pod

Simulating Humanity: 85% accurate digital twins, behavioral foundation models, social physics, & 8 billion agents https://t.co/ntIccXOr1y @simile_ai CEO @joon_s_pk explains how AI can move from predicting what people will do to simulating how to shape outcomes, why today’s frontier models still miss how humans actually behave, how digital twins reproduced people with 85% accuracy, why human biases and mistakes have to be learned rather than optimized away, and what it would take to eventually simulate all 8 billion people on Earth.
i have another saas to kill (will share results of Kill My SaaS 1 next week!!)




btw if you havent set your {codex | claude | gemini | devin} automations to autoresearch how to improve your seo/aeo every week you are really truly missing out on free, should-be-commoditizing-but-weirdly-untapped alpha
it was a blast covering Build Fest this year - they didnt know this but I learned to code with MongoDB over 10 years ago (shoutout MERN stack) and now in the AI Engineering era seeing SF builders rediscover MDB was so gratifying - this was a *huge* step up from last year!

ICYMI: Everything announced at #MongoDBlocal Build Fest is designed to help developers stay in flow and ship production AI, faster. From the new Atlas Managed MCP Server for Agents and native Atlas access inside @claudeai, Claude Code, @ChatGPT, Codex, @grok Build, @DevinAI, & @cursor_ai, to Automated Embeddings in MongoDB Atlas and @VoyageAI's voyage-code-4, plus a GA'd Embedding & Reranking API and Vector Search for Stream Processing, it's about less setup, fewer workarounds, and more time building. Get the latest: https://t.co/YGX7MGq0oN @swyx @BazeleyMikiko
by total views, matt’s now the top @aidotengineer speaker in our brief history, and arguably one of the best skills experts in the world — his /grill-me has reached all echelons up to @satyanadella (with some variations…) /wayfinder is @mattpocockuk’s /grill-me for /grill-me, when you are just navigating the fog of war and don’t yet know what you dont yet know, and need to orchestrate research and other grill sessions to get there. super glad to have @ricmac help launch our new Skills coverage with his course launch this week! 👇exclusive interview, quick read


We chat to @mattpocockuk about his /wayfinder skill, which he designed for the "fog of war" — when you need to figure out a project but the end state isn’t entirely clear. This is the first in a series of skills we'll be exploring in the coming weeks. https://www.latent.space/p/wayfinder-skill
well, @openclaw was good to apple this team conservatively drove $50-$150m in mac mini sales alone this year haha (roughly +50% of normal annual mac mini sales worldwide, just due to openclaw 2026)


512GB RAM Studios. Apple was good to us. 🦞

HAHAHAHAHAHA non technical people are so incredibly cooked they are burnt thru can u imagine covering ai with zero context, zero reasoning, zero internal world model. it must be so delightfully joyful, everything is so amazing, face value is all you need, what a time to be alive


With Codex, @asana finished a frontend test migration from Enzyme to React Testing Library in two calendar weeks—a project expected to take five more years. https://openai.com/index/asana/
we've been doing a lot of a/b testing of @aiDotEngineer youtube thumbnails. i always hated that it is such an opaque process. open sourcing/crowdsourcing our learnings today! https://t.co/J9SYClIF7z doing this in hopes that people can share their experience or learn from ours. at the end of the day we just want to get good educational content to rise above the noise online. please lmk what you think
Trajectory have generally impressed me with their tasteful execution on ambitious goals. on our Continual Learning track, @rronak_ gave a very thoughtful overview on how they're tackling the main data problems left in CL, including why GRPO isn't enough and they had to go on-policy.... and then subsequently fix all the issues that come up with it nice overview from one of the early leaders in this field! (see the rest of the track for more, this one was quite stacked)





It's time to rethink RL. Translating real world use into model improvements requires redesigning post-training algorithms for non-verifiable, per token rewards. At @aiDotEngineer 's World Fair, we share our insights into scaling algorithms like SDPO for continual learning.
5 years later and most of the best players here have been bought


🆕 Blog: Why Isn't Usage Based Billing A Bigger Category? https://dev.to/swyx/why-isn-t-usage-based-billing-a-bigger-category-m8b
in case you’re not living in the tech bubble, as a general rule i’ve been surprised by how infrequently top tier folks actually meet/know each other. as an outsider i might have assumed that everyone is in secret illuminati group chats. those exist, but are very much short lived exceptions rather than the rule. what you see of the major headlines is pretty much what they also see. i guess one way to interpret this is also simply that most effort is still on doing the work rather than working the narrative or the gossip. and that, to me at least, is genuinely quite reassuring, this far in to my career.

Anthropic cofounder emails Noam Shazeer November 21, 2021




the reason @databricks "ipo is lava" fundraises are a meme is this but unironically the M in their $188B series M stands for "we are going to kill so many meetings"


I got this question so many times today. "How can you grow 80% at $7B?" The true answer is that we're finally seeing a breakthrough with AI agents starting to work in the enterprise. The AIs have been super smart for a while, but have lacked basic context that's in people's heads, or in some SaaS system-or-record. A lot of organizations are deploying FDEs to capture this context, or Ontology, and feed it to the AI. This is labor intensive and expensive. We just automated that with Genie Ontology. Once you have that enterprise context graph, an AI agent like Genie becomes magical. I find myself no longer waiting for answers from my CRO, CFO, CMO, CHRO etc, I just keep queuing up questions on the phone while sitting in meetings. It'd frankly addictive. Our customers are starting to do the same, over 70% of all queries on the platform are now generated by Genie agents. This fuels more questions to the platform, which drives consumption, which drives revenue. That's the simple answer.
AIE NYC CFP wave 1 acceptances are being finalized today. last day to get in for wave 1! https://t.co/ldei0UywCk our NYC event last year was the most successful summit we've ever had. excited to head back to 🗽 bigger and better than ever - note the special requirements for our mainstage finance keynotes

🗽 AI Engineer returns to NYC! Oct 12-14, a beautiful time to visit. Tickets go on sale soon, but the Call for Speakers is now open! https://t.co/E3ojLGFbTg If you do leading AI work in NYC, especially for AI x Finance, speaker applications for AIE NYC 2026 opened today. For mainstage keynotes, we are actively soliciting the top AI native companies and AI transformations in: - Investment Banking - Commercial Banking - Consumer Banking - Hedge Funds - Private Equity - Financial Data - Insurance - Venture Capital - Agentic Commerce - Accounting - Payments - Compliance - Mid/Back Office If you are NOT in Finance, don't worry, we will still be running our regular AI Engineering and AI Leadership tracks, covering all the latest in coding, generative media, evals, infra, and more. But if you want to meet some of the largest AI native banks and financial institutions in the world, this will be the #1 place to be this fall!
damn deepseek moves fast on roon tweets


the major ai companies should commit to a real time priced API product. ai demand varies wildly over a day/night curve and the industry is broadly extremely capacity crunched. meanwhile agents make dealing with variable pricing, batching, and projecting total costs very easy
!!!!! it's been a huge huge pleasure to serve @arizeai these past years with AIE, and this could not happen to better people. Dynarize is now a globally trusted $14B observability powerhouse that just got one of the best AI-native US teams in this business. if you're interested to learn more, come to @aidotengineer NYC this fall: https://t.co/9j7D5tVfOg where they are the first non-bigcloud to be an AIE presenting sponsor!!

Jason and I started Arize 6+ years ago with a simple proposition that headlined our seed deck: “We Make the World's AI Work” Today we are announcing we’ve entered into a definitive agreement to be acquired by Dynatrace to accelerate that vision. We made a bet years ago that the explosion of AI would require bespoke infra tools. We knew AI systems weren’t going to behave like traditional software because we were building the next generation of intelligence. That bet became the market's first AI Observability platform. Fast forward to 2026. The world has changed dramatically. Every company in the world is scaling up its use of agents. AI is everywhere and agents are reshaping business after business. What wasn’t clear years ago and is crystal clear today is that agent systems and software are deeply connected together in ways that none of us fathomed. - The best general agents are coding agents. - Skills and tools for agents are mixtures of prompts and code. - Prompts that control agents live in code repos. - The logs and traces used to debug your software also help you debug your AI systems. - The path to AGI is likely through coding agents. The two worlds of AI and software observability are coming together to create the future of observability. Lots more to share in the coming days, but I wanted to thank folks who made this possible: To the Arize team - startups are about the people you build with, and I’ve never met a better group of humble rockstars. The markets changed, our product evolved, but the camaraderie of the team through this journey is what I’m really proud of. To Jason, being in the trenches with you for the past 6 years has been the journey of a lifetime. You’re a visionary, and the best entrepreneur I know. Everything I know about entrepreneurship comes from watching you execute day in and day out - you always figured it out. Arize wouldn’t be Arize without you. To our customers, thank you for trusting us and building this category with us. You made the product better everyday, and inspired us with the products you’re building. To the open source community around Phoenix and OpenInference, developer trust and open standards are the foundation this category gets built on. And to our investors, advisors, and early believers who backed us when AI observability was not a budget line or even a real category yet, thank you. Onwards to the next chapter! https://t.co/OBZSsnX8pu

u guys have no idea how serious elon is about winning coding

🆕 @elonmusk has started following @cognition

The @latentspacepod team and I will be speaking at Mongodb's .local buildfest today 10a-4pmish, come by! https://t.co/2ECvrv0x4N We'll be covering: - tba Agent infra with @BazeleyMikiko! - the state of embeddings/reranking with @frankzliu @VoyageAI head of Applied ML - wearable AI with @a_israelov, CTO of Mentra! - their state of managed MCP with Gaurab Aryal - Jim Scharf, CTO MongoDB - AIE top speaker Apoorva Joshi click on only if you are gonna come... take one of my tix MDBLCL26INSIDER!
completely bowled over by the incredible responses to the $10,000 Kill My SaaS hackathon this past weekend. sorry it took so long to get back to some of you, organizing this thing solo on top of my regular meetings and work this week was completely stupid and rushed — but this is why we need you! blasting out submission forms now! - check discord - but check out some INSANE submissions from @agrimsingh and @realgenekim and over a dozen others just in the weekend alone



So @swyx announced his “Kill My SaaS in one weekend” contest, where you could win $10K to replace his $40K/year SaaS. On Saturday morning, I learned that it was to replace their conference CFP software. OMG. I've run about 24 conferences over 12 years, but something I rarely talk about is how much I've hated the software we've had to use — over the last decade, we've used 5+ CFP tools, which manage the process of taking speaker submissions and running a review process, some MUCH worse than others. With each, we've had to build so many workarounds, things like Basecamp, Trello, Google Sheets, Zapier, and I've written entire new apps to be the reviewer frontend. (BusyConf was the one we loved, but they went out of business.) So I jumped into the contest (Margueritte Kim and I will donate it all to a STEM charity if we win), and less than 24 hours later, I built CurtainCall CFP. This has been the craziest dev experience of my career — and when Swyx released his eval harness, the entire project became a hill-climbing exercise, and began one of the craziest infrastructure experiences of my career. (Will write more about that later!) But it's in production! The Enterprise AI Summit is in Charlotte on Oct 7–8, and there are a few slots that may become available! If you want to submit an experience report talk, submit your proposal! (And you can see the 14 amazing announced speakers (more coming) here, too!) https://t.co/WrXwr74ksr
Perplexity offered to buy @googlechrome one year ago today (this is a scheduled tweet)

this is already one of the most important papers of this year. https://www.latent.space/p/ainews-how-to-steal-a-reasoning-trace the methodology doesnt seem clearly explained so here are some notes with a further distillation

We can finally talk about it: We found a way to extract hidden reasoning of frontier models using a vulnerability in the APIs of every frontier AI company. We verified that our reasoning token count matches billed API thinking tokens 1:1 for most of the prompts we queried.


