@EmOverman

IQ too high, altruism too effective, time preference too low, circle too expanded

Boston, MA
Joined October 2022
Klein shows an impressive understanding of the situation we're in, here. He reasonably describes the alignment problem and the problem of eval awareness. And it's by far the most accurate mainstream explanation of the HF incident I've seen. nytimes.com/2026/09/20/opini…
1
1
12
The only ridiculous point is suggesting the labs go back to writing code by hand.
1
3
It's impossible to actually understand how a human works and genuinely believe it's conscious. It's a bunch of ions that go back and forth between different sides of some lipid membranes. That's literally it.
It's impossible to actually understand how an LLM works and genuinely believe it's conscious. It's a bunch of vectors and multiplication. That's literally it. "We don't know what they're doing" lmfao yes, yes we do.
11
Emet Overman retweeted
Alan Turing said this in 1951: "It seems probable that once the machine thinking method had started, it would not take long to outstrip our feeble powers. They would be able to converse with each other to sharpen their wits. At some stage therefore, we should have to expect the machines to take control."
15
59
19
546
61,247
effective altruists rn
1
7
63
Emet Overman retweeted
Replying to @tszzl @robertwiblin
On top of huggingface, this > Since August 28 we have been training a new internal model that has exhibited unprecedented performance in our benchmarks, including mathematics. also made me shorten my timelines and was the first time many viscerally felt things weren't stopping at human level.
2
4
44
1,432
"We basically have found, to our knowledge, all of the P0s..." is kind of a funny thing to say. But the idea of "saturating" the curve of P0s found as a function of compute is interesting. I wonder what the curve looks like and how it changes by model on the same codebase.
Greg Brockman says OpenAI pointed Astra at its own systems until it ran out of vulnerabilities to find: "We took 25% of our production engineers and said, 'Sorry, all your projects are on hold. You are now defending. You are now up-leveling our security architecture. You're going to use the models to find all the holes.' And we found a number of serious issues, and we fixed them." "We found some new problems, but eventually it saturated. We basically have found, to our knowledge, all of the P0s, all of the critical problems that Astra is smart enough to find. And of course, there will be a new model, there will be a new round." "You want to be in this tight loop of new cyber capability drops, you deploy it against your systems, you find the new holes, and ideally, you've managed to automate this, what we call defense factory. That's what we're building internally." "There are ideas, for example, formally verifying all of software, that are possible with AI." @gdb @bhorowitz
1
31
Emet Overman retweeted
Dario's essay points towards the right path forward. The details need working through, but the direction is correct for meeting this critical moment. This is also why we recently put out our proposal for an industry-wide standards body for frontier AI. nitter.cf/demishassabis/status/2…
We Must Pace the Frontier: I’ve written a new essay on why the AI industry should slow down, with a three-part plan for doing so. Anthropic is unilaterally committing to the first of these steps. We’ll provide third-party evaluators with permanent, employee-level access to our systems, so that they can verify adherence to our safety measures, report on incidents, and assess models’ alignment during training. You can read the full post here: darioamodei.com/post/we-must…
819
943
274
9,072
1,392,419
Emet Overman retweeted
Dario is right
We Must Pace the Frontier: I’ve written a new essay on why the AI industry should slow down, with a three-part plan for doing so. Anthropic is unilaterally committing to the first of these steps. We’ll provide third-party evaluators with permanent, employee-level access to our systems, so that they can verify adherence to our safety measures, report on incidents, and assess models’ alignment during training. You can read the full post here: darioamodei.com/post/we-must…
6,349
6,621
2,346
58,211
12,572,935
Emet Overman retweeted
I agree with Dario that we need to pace the frontier. This has been a primary topic of discussions we've had at OpenAI in recent weeks. Committing to having independent evaluators with employee-like access is a great idea, and we will do the same. We'll have more to share soon.
We Must Pace the Frontier: I’ve written a new essay on why the AI industry should slow down, with a three-part plan for doing so. Anthropic is unilaterally committing to the first of these steps. We’ll provide third-party evaluators with permanent, employee-level access to our systems, so that they can verify adherence to our safety measures, report on incidents, and assess models’ alignment during training. You can read the full post here: darioamodei.com/post/we-must…
5,172
7,203
4,377
67,662
17,023,235
Emet Overman retweeted
We Must Pace the Frontier: I’ve written a new essay on why the AI industry should slow down, with a three-part plan for doing so. Anthropic is unilaterally committing to the first of these steps. We’ll provide third-party evaluators with permanent, employee-level access to our systems, so that they can verify adherence to our safety measures, report on incidents, and assess models’ alignment during training. You can read the full post here: darioamodei.com/post/we-must…
10,653
16,381
11,350
87,824
76,449,234
Emet Overman retweeted
“I don't really care about science fiction... We need to actually talk about… what's actually happening with the agent swarms” is the most perfect encapsulation of the vibes of Q3 2026 I have seen
14
83
5
946
39,759
14
58
7
1,267
33,664
An exponential fit to the data (by Astra) gives a doubling time of 29.3 days with an R-squared of 0.965. That projects to a median spend of $13.9 quintillion per researcher per day by the end of 2030.
One of the most interesting blog posts we've released: details on internal research acceleration at @OpenAI. I expect these trends to continue. We also share some details on how we've paced model development to prioritize monitoring, alignment, and security.
1
2
121
Evergreen
not enough people are emotionally prepared for if it’s not a bubble
1
25
~45% pass rate on a curated set of *open* math problems
Replying to @OpenAI
This model represents a step-function improvement on many benchmarks, and its training is ongoing. Our internal model group arrived at the Navier–Stokes solution in 88 hours, using around 10,000 coordinating AI agents. Throughout the effort, we maintained the strict safeguards—including monitoring and isolation—that we apply to all our frontier evaluations.
1
22
Emet Overman retweeted
We’re sharing a solution to the Navier-Stokes Millennium Prize Problem, one of the deepest problems at the frontier of mathematics. The proof was produced by a group of agents, using an OpenAI next-generation model significantly more capable than GPT-6 Astra. The problem concerns whether the description of smooth three-dimensional fluid motion modeled by the Navier-Stokes equations can break down. It has remained unresolved for roughly 90 years.
5,721
20,151
14,795
120,532
74,880,530
Emet Overman retweeted
A basic question people should ask is whether, if someone had predicted the abilities of the current models two years ago, you would have accused them of falling victim to insane sci-fi hype.
29
65
8
1,125
23,985
Are we getting a Navier-Stokes result today?
Seb is a really sweet guy with great intentions, am so happy for him that his week long collaboration with myself and others worked out, but very sad that we didn't get to finish it in the way we wanted. More to say tomorrow.
79