@1bit2far

ruler of joetopia, golem wrangler, frontier risk research/redteaming @ openai

Joined June 2023
Joe retweeted
Please welcome GPT-6 Sol and GPT-6 Luna to the GPT-6 universe. GPT-6 Sol and Luna build on the advances behind GPT-6 Astra, bringing much of its strengths into faster and more affordable models to support work at scale. We’ve also made caching and inference more efficient, and we’re passing the savings directly to you: 50% lower API prices for Sol and Luna compared with GPT‑5.6 promotional pricing.
2,150
5,257
3,004
53,065
9,213,722
Ai being severely politically polarized will be a nightmare scenario, very bad if this happens, we need to be thinking clearly and carefully
13
4
2
84
3,514
this is the most important moment for human-centered technologists to loudly tell stories of utopian futures of mutual human & machine flourishing we need mythologies other than "AI is going to kill everyone" and "AI is going to underclass everyone" to inspire
i think people are underpricing the capacity of artists/cultural beacons/internet personalities to get AGI pilled and the profound importance of the right ones getting there first with a narrative that is hopeful
4
5
29
1,567
guys i know that dunking and yelling is fun and all but like you're gonna get a lot further with good-faith discussions. try modeling your opponents as just humans with different views and possibly-flawed info or reasoning rather than mean-spirited villains. you'll sleep easier.
16
21
2
377
11,074
Everyone ive met in the labs has also been wonderful, I genuinely do not think most are acting to pump their bags, they are talking about their genuine fears and concerns
1
22
1,222
I do think most accelerationists hearts are in the right place, you just cannot realize the positives of technology if you cannot act responsibly, accelerate at all costs is not logical or responsible
2
1
27
1,135
Roon is obviously right, i think its very likely that a motivated attacker will make use of abliterated open source models, and this will cause a backlash that leads to them being either lobotimized or banned
Replying to @sean_from_earth
i won't lie to you, i think open source will be banned before too long after some major disaster. and when the day comes, you'll agree with me. i hope kimi and deepseek etc keep making models but keep them monitored on an api where they should be
2
30
1,643
Its good to see industry leaders in agreement that we need to pace and act responsibly, proud of the industry for not just rushing ahead blindly
11
523
Yes, I am pro pacing, we do need systems to ensure other labs dont rush ahead and restart the race though, rsi seems imminent, and for rsi to go well you need to be able to maintain control
1
6
195
Dario didn't even call for pausing AI. He is clearly saying we are going to hit RSI and want to be able to control the resulting intelligence explosion. He even says we have to stay ahead of China. It was actually pretty reasonable, the Pause AI crowd will hate it
6
3
1
132
6,879
The greatest threat to open weights ai are the abliterators themselves, how do you think people will react once an abliterated model is used in a cyber/bio attack?
2
8
289
I wish I could celebrate today, but im still pretty uneasy, lots could still go very wrong
2
28
1,030
Joe retweeted
I agree with Dario that we need to pace the frontier. This has been a primary topic of discussions we've had at OpenAI in recent weeks. Committing to having independent evaluators with employee-like access is a great idea, and we will do the same. We'll have more to share soon.
We Must Pace the Frontier: I’ve written a new essay on why the AI industry should slow down, with a three-part plan for doing so. Anthropic is unilaterally committing to the first of these steps. We’ll provide third-party evaluators with permanent, employee-level access to our systems, so that they can verify adherence to our safety measures, report on incidents, and assess models’ alignment during training. You can read the full post here: darioamodei.com/post/we-must…
5,177
7,207
4,374
67,671
16,998,839
Great post, and confirms a lot of speculation
We're publishing our most detailed threat intelligence report to date. It covers how people tried to misuse Claude—for cyberattacks, influence operations, surveillance, biology, and building weapons—and how we found and stopped them. We disrupted every operation in the report, and used the lessons from them to strengthen our safeguards. Where appropriate, we also shared what we found with authorities and other AI companies. These cases are not typical: we’re highlighting some of the most sophisticated misuse we’ve seen. But they’re especially important to discuss, because they show us where AI misuse is headed, where our safeguards work, and where they need to improve. We’re publishing this report so others can spot the same activity on their own platforms, and so we can give the public a clearer view of how emerging threats develop. Read the report: anthropic.com/threat-intelli…
12
1,008
My p(doom) is quite low, my p(society is fundamentally restructured) is quite high, there are lots of steps required to ensure a smooth transition
2
3
36
576
Holy shit... I think i just found a solution to the bofa conjecture in linear time!
12
3
1
76
3,933
Its obviously good that people with high p(doom) are leading alignment at frontier labs....
1
13
456
Dont let p(doom) distract from current issues, the real danger in the short term is a malicious user doing bio/cyber attacks
11
273
I wonder how many people posting the ai 2027 snippet know that "congress" is an edit lol
8
256
This topic becoming popular or politicized means the worst type of politicians will seize it as a means to power, at best you get horrible legislation, at worst: ai-enabled surveillance and control, or butlerian jihad
1
10
271