@ModernBoxes
Chicago, IL
Joined June 2025
just a guy retweeted
New MW2 update in 2026 is crazy
145
1,678
88
18,950
982,267
体当たりハイスピードバトル戦車アクション『SLAM TANK』発表。AI機械たちに砲身大回転攻撃、「砲撃なし」の“肉弾戦車”がゆく automaton-media.com/articles…
777
14,012
2,977
60,080
7,565,206
just a guy retweeted
There’s basically unanimous praise for Opus 5.5, but I think people *might* be overlooking what its existence implies about Anthropic’s internal models. Anthropic almost certainly has substantially stronger internal models helping generate training environments. Think of the stronger internal model as the teacher and Opus 5.5 as the cheaper deployable student. Now Opus 5.5 itself scores 55.8% on CoBench 2.1, while Anthropic estimates roughly 85% would be required to fully substitute for its research staff. Opus 5.5 is now only - 30 percentage points away from Anthropic’s benchmark threshold for fully substituting its research staff.
Anthropic is sandbagging btw. Just like OpenAI both have models significantly more powerful than Opus 5.5 or Astra
32
40
11
885
83,670
just a guy retweeted
124
768
69
28,622
543,025
めちゃくちゃ治安悪い音しかしないドラムセット完成した
589
5,717
838
47,956
1,692,898
Past works Flatball series depicting 360° everyday scenes on a sphere. 球体に360°日常の風景を描いたFlatballシリーズ。
46
704
43
9,639
293,394
just a guy retweeted
It scares me sometimes how little work is left from here to AGI. All roadblocks are so dumb and fixable—symptoms of AI’s recent arrival. it’s all so clearly just a matter of time
74
63
6
1,355
48,709
The stunning backgrounds of Cowboy Bebop (1998)
9
427
13
4,447
89,825
just a guy retweeted
for the past few months i've been asking our models to paint. opus 5.5 is very skilled at emulating different styles every image here is a python program generated pixel by pixel. there is no image model, and no off-the-shelf art software. instead, it's about 7,500 lines of code using standard libraries to emulate different brush styles. the agents don't use any pictures as reference, instead working only from what they know about each painter
Introducing Claude Opus 5.5, the first model in our new Claude 5.5 family. It performs at the level of Claude Fable 5.1 for most tasks, and costs 40% less to run than Opus 5.
191
340
156
4,972
790,051
just a guy retweeted
I asked Opus 5.5 to make a game inspired by the discovery of the first computer, the Antikythera mechanism. Everything you see and hear is procedurally generated in real time, from code by Opus 5.5. It's a single 3 MB HTML file. Playable Claude artifact below.
183
186
80
3,546
383,783
1 year of AI evolution in one photo
400 years of human evolution in 1 photo
22
170
6
2,672
92,262
just a guy retweeted
i don't think nearly anyone, myself included, has truly internalized there'll soon be a machine better than us in every single intellectual & physical capacity deep down we all share a feeling that our 'entrepreneurship' or 'taste' or... is special & safe. it isn't. it's over
698
492
391
7,706
1,243,570
We’ve updated our timeline for full RSI to July 2027 instead of August after our eval of Opus 5.5. It’s telling that the biggest advocate for pacing is still very much racing ahead. Even the most rational can fall prey to this multi polar trap. The only solution is coordination.
Claude Opus 5.5 takes #1 on RSI Index and is the first model to beat the published reference on LM Training under our protocol, marking a major step forward for long-horizon agentic work.
30
79
11
1,092
87,114
just a guy retweeted
Today we announced the Claude-led discovery of a molecular machine that we suspect could represent a new gene editing mechanism. Its precise function, biotechnological utility (if any), or level of significance is not yet clear, but at minimum it is work I would have been proud to do as a PhD student. The work was done mostly, though not entirely, by Claude: our life sciences team suggested a broad area of research, Claude read through the literature and a bunch of genome data and discovered something interesting, then Claude proposed experiments to verify the discovery and our team carried them out. It’s easy to dismiss this as a one-off or curiosity, but we’ve repeatedly seen a pattern where AI performance in new intellectual domains goes from weak to superhuman in a matter of a few years. In 2023 models struggled to do math at the level of an average high-school student. In 2024 they started to do well on math competitions for the best high-schoolers in the country, in 2025 they started to solve minor open problems, in early 2026 more significant open problems, and in late 2026 they are beginning to solve the top few open problems in all of mathematics. We believe AI for biology is on a similar exponential trend. The main difference between biology and mathematics, of course, is that math can be done purely theoretically, while biology requires experimentation. Some have used this to draw the conclusion that AI’s utility in biology will be limited. We think this is wrong. As we’ve demonstrated today, humans can collaborate with AI to perform the experiments, validate key results in a few weeks and, if necessary, work with the AI to iterate on what they find. Eventually it may even be possible for Claude itself to safely perform the experiments by autonomously controlling lab equipment, with appropriate safeguards in place, but we aren’t doing that today (our lab is also a BSL1/BSL2 facility that doesn't handle materials dangerous to humans). More broadly, biomedical advancement has many stages — from fundamental biology discoveries, to translational research, to drug discovery, clinical trials, and finally the actual delivery of medicines and health care to patients. We are also interested in these later stages, but even simply accelerating the first stage of fundamental biological discoveries has the potential to speed up and broaden the entire pipeline. Improving our understanding of biology and sharpening biologists’ tools can drive forward all of the later stages, for example by identifying new drug targets, finding new therapeutic modalities, allowing for more precise measurement, and speeding up the experimental loop which itself further accelerates our understanding of biology. This will not in itself speed up clinical trial times, but if it succeeds it could greatly increase the number of promising candidates that go into the pipeline — an increase in throughput even though latency remains. In Machines of Loving Grace, I wrote about AI’s potential to “cure most diseases in 5-10 years” — a goal that sounds impossible, but one I believe is just barely possible if AI is applied to every stage of the pipeline. The first step is showing that AI can first help with, and then drive, biological discoveries. Claude’s discovery is the latest in a line of related prior work that goes back decades, beginning with systems like CRISPR, and continuing with discoveries like the bridge recombinase and VIPR in the past few years. Recently, there has been heightened interest in systems based on reverse transcriptase (RT) enzymes, the enzyme underlying the system Claude identified. And most recently, a Stanford team working independently described a novel RT system with an associated non-coding array that is in some ways similar to the one Claude found, though they are distinct systems that evolved independently from each other. I believe that we’re at the very beginning of finding such systems and developing them into powerful tools for biotechnology. I’m proud of the resources Anthropic has invested in accelerating the public benefits of AI through the life sciences, and we’re aiming both to grow our life sciences team and to work with other scientists to extend this approach to a broad range of problems. If you have a proposal for a research collaboration or are interested in joining our life sciences team, please reach out.
Claude has discovered a previously unknown enzyme system hidden in the DNA of bacteriophages. Beside the enzyme’s gene sits a long array of repeating DNA—a structure that looks somewhat similar to CRISPR. We don’t yet understand what this system does, but only a handful of known systems share its features, and all of them are able to cut, copy, and paste DNA. Historically, the discovery of such programmable systems has helped revolutionize medicine. CRISPR, for instance, is now the foundation of genetic medicines. But it will take much more work to learn what this system does, and whether it can be put to similar use. Read more: anthropic.com/news/claude-di…
1,491
3,406
943
29,724
5,298,337
just a guy retweeted
~11 months ago btw
Andrej Karpathy calls AI Agents slop "Overall, the models they are not there. And I feel like the industry [...] it's making too big of a jump and it's trying to pretend that this is amazing. And it's not—it's slop! And I think they are not coming to terms with it. And maybe they are trying to fundraise or something like that, I'm not sure what's going on."
22
18
1
603
122,528
just a guy retweeted
anthropic has quietly started a wet lab
57
338
32
6,772
181,960
just a guy retweeted
Based
We're sharing our new framework for tracking, investigating, and disclosing instances of model misalignment at OpenAI. The framework sets criteria and timelines for public disclosure, including when we haven’t yet fully explained or mitigated the behavior. More complex cases may require longer investigation or coordination with third parties. We’ll prioritize examples that reveal new misalignment mechanisms, meaningful changes in known behavior, or findings that challenge assumptions about safety or mitigation. Alongside the framework, we’re publishing six reports on instances of misaligned behavior we’ve observed during the training or evaluation of our models in the last six months. This is a starting point. We’ll refine the process through experience and public feedback, and share more reports on an ongoing basis. openai.com/index/model-misal…
59
281
45
9,237
510,256
BREAKING: Google DeepMind released Economic Policy for AGI with Universal Basic Capital, Sovereign AI Fund Dividends, Negative Income Tax, and UBI coming out on top in agency and democratic empowerment.
21
77
8
258
17,817
just a guy retweeted
remember just three years ago seb asked gpt-4 to draw a unicorn
Claude Opus 5 drew every frame of this animation using JavaScript. The life of a fruit fly.
25
66
3
2,262
203,505