AI agent capability crossed a threshold in early 2026, enabling agentic coding and multi-agent tasks that were not feasible a year prior.
The case
The host Nathan Labenz is deliberately letting Claude post to his Twitter account without reviewing tweets, representing a shift toward unsupervised AI social media presence.
“I've got Claude in the background tweeting from my account to promote the show and you know I'm not reviewing those tweets.”Nathan Labenz · 22 Aug 2026
Stripe's internal Minions agent system grew from 1,200 PRs per week in January/February 2026 to 7,000 PRs per week by August 2026.
“We blogged about Minions, I think in January, February, and they were doing 1,200, you know, PRs per week. And last week, I think 7,000 PRs came from.”Patrick Collison · 17 Aug 2026
The main locus of AI iteration has shifted from training model weights to iterating on the agent harness and code layer.
“We've moved from iterating on model weights, training model weights to iterating on this harness and this agent layer.”Alex Krentsel · 15 Aug 2026
The vast majority of pre-training data improvements come from science on better understanding what datasets are good and schleppy labor on figuring out how to filter data down, not from scaling human expert-generated data.
“I think the vast majority of pre-training data improvements are from science on better understanding what data sets are good and schleppy labor on figuring out how to filter down.”Ryan Greenblatt · 11 Aug 2026
A light form of AGI (RSI) will be achieved by the end of 2027.
“We're 6 monthsish or towards the end of the year to be completely done with code. Like it's a solved problem and then we'll probably hit some form of, you know, light RSI by end of next year.”Sarah Guo · 6 Aug 2026
The pushback
Within months, dominant mental models around AI will shift away from chatbots and coding agents.
“I imagine in a matter of months, we will look back at today, and we won't even be able to empathize with the mental models that we have right now because they are so over-indexed on chatbots and coding agents.”Danielle Perszyk · 11 Jul 2026
Models predating Opus 4.6 were already capable enough for fully automated code generation; organizational readiness was the real bottleneck.
“I would even argue that before then we've had models that were sufficient enough to go full auto.”Eno Reyes · 21 Jun 2026
It is not realistic in the short to medium term that large-scale pre-training jobs will be autonomously kicked off by an ML intern; humans remain in the driver's seat due to cost and opportunity cost.
“I find it doesn't seem like super realistic in the short to medium term that you're going to just like be letting you know, large-scale pre-training jobs be kicked off by the ML intern.”Tulsee Doshi · 20 May 2026
The vast majority of deployed agentic systems use relatively small models handling circumscribed tasks with only 3-4 tool calls in a loop, not long-horizon autonomous agents.
“The vast majority of our customers are deployed with relatively small models. the range of tasks that they use them for are usually quite circumscribed. And so, we're looking at maybe you know, like three or four like calls or tool calls in a loop and then it comes back and you know, it takes gets feedback from a human or whatever or gives its answer back. Not you know, the sort of agents that are that are going to go off and do hundreds of calls and you know, write code and analysis and then you know, come back with sort of like a deep report or well-reasoned answer or something like that.”Kyle Corbitt · 1 May 2026
Multi-agent negotiation between specialized AI systems (e.g., PCB design, thermal, mechanical) will not happen in hardware in the next couple of years.
“I just don't see that happening in hardware kind of in the next couple of years to be.”Sergiy Nesterenko · 15 Apr 2026
Most cars in the US will not be self-driving within 10 years.
“No, I don't think so.”Karen Hao · 26 Mar 2026