@1bit2fari
iAccount based inUnited States
About this account
- Account based in
- United States
- Connected via
- United States Android App
Account-level information from X, not a live location or the device used for a specific post.
ruler of joetopia, golem wrangler, frontier risk research/redteaming @ openai
Joined June 2023
- Tweets6.8K
- Following2.1K
- Followers1.2K
- Likes19.9K
Please welcome GPT-6 Sol and GPT-6 Luna to the GPT-6 universe.
GPT-6 Sol and Luna build on the advances behind GPT-6 Astra, bringing much of its strengths into faster and more affordable models to support work at scale.
We’ve also made caching and inference more efficient, and we’re passing the savings directly to you: 50% lower API prices for Sol and Luna compared with GPT‑5.6 promotional pricing.
Joe retweeted
this is the most important moment for human-centered technologists to loudly tell stories of utopian futures of mutual human & machine flourishing
we need mythologies other than "AI is going to kill everyone" and "AI is going to underclass everyone" to inspire
Joe retweeted
guys i know that dunking and yelling is fun and all but like you're gonna get a lot further with good-faith discussions. try modeling your opponents as just humans with different views and possibly-flawed info or reasoning rather than mean-spirited villains. you'll sleep easier.
Roon is obviously right, i think its very likely that a motivated attacker will make use of abliterated open source models, and this will cause a backlash that leads to them being either lobotimized or banned
Replying to @sean_from_earth
i won't lie to you, i think open source will be banned before too long after some major disaster. and when the day comes, you'll agree with me. i hope kimi and deepseek etc keep making models but keep them monitored on an api where they should be
Joe retweeted
Dario didn't even call for pausing AI. He is clearly saying we are going to hit RSI and want to be able to control the resulting intelligence explosion. He even says we have to stay ahead of China.
It was actually pretty reasonable, the Pause AI crowd will hate it
Joe retweeted
I agree with Dario that we need to pace the frontier. This has been a primary topic of discussions we've had at OpenAI in recent weeks.
Committing to having independent evaluators with employee-like access is a great idea, and we will do the same. We'll have more to share soon.
We Must Pace the Frontier: I’ve written a new essay on why the AI industry should slow down, with a three-part plan for doing so.
Anthropic is unilaterally committing to the first of these steps. We’ll provide third-party evaluators with permanent, employee-level access to our systems, so that they can verify adherence to our safety measures, report on incidents, and assess models’ alignment during training.
You can read the full post here: darioamodei.com/post/we-must…
Great post, and confirms a lot of speculation
We're publishing our most detailed threat intelligence report to date.
It covers how people tried to misuse Claude—for cyberattacks, influence operations, surveillance, biology, and building weapons—and how we found and stopped them.
We disrupted every operation in the report, and used the lessons from them to strengthen our safeguards. Where appropriate, we also shared what we found with authorities and other AI companies.
These cases are not typical: we’re highlighting some of the most sophisticated misuse we’ve seen. But they’re especially important to discuss, because they show us where AI misuse is headed, where our safeguards work, and where they need to improve.
We’re publishing this report so others can spot the same activity on their own platforms, and so we can give the public a clearer view of how emerging threats develop.
Read the report: anthropic.com/threat-intelli…