@_beenkimi
iAccount based inUnited Kingdom
About this account
- Account based in
- United Kingdom
- Connected via
- United States App Store
Account-level information from X, not a live location or the device used for a specific post.
Research Scientist at Google DeepMind, PhD from MIT. Make machines empower people.
Joined August 2011
- Tweets851
- Following516
- Followers27.9K
- Likes2K
This piece by Scott Aaronson is like the sharpest knife that cuts through the messy, politicized discussions on AI, the risk, if we should be alarmed or not, why and what we should do.
If you don't want to do anything useful today, you've done something productive by reading this piece from top to bottom, word for word. scottaaronson.blog/?p=10062
If you are feeling sad about an AI (presumably) solving Navier–Stokes, you are not alone.
Drama aside, the fact that it only took a matter of days hits differently. Terry Tao’s words resonate deeply here:
"The mathematical community functions, in many ways, as a miniature version of humanity. It consists of individuals using a wide variety of different approaches, joined by core values. The most precious resources of our profession are students and ideas, and these we nurture with great care."
terrytao.wordpress.com/2026/…
Been Kim retweeted
When an LLM is underperforming, we tweak the prompt. Why not do the same for video models solving visual reasoning tasks?
Visual prompt engineering changes the start frame appearance but not the task, improving performance.
📄 arxiv.org/pdf/2607.25537
🌐 visual-prompt-engineering.gi…
Been Kim retweeted
The program for the #ICLR2026 Workshop "From Human Cognition to AI Reasoning" is now available. We have a fantastic lineup of talks.
🔗 hc-air.github.io/hcair26
Invited Speakers: @DocRachidAlami, @_beenkim, @ced_zhang
Co-Organizers: @julie_a_shah, @sarath_ssreedh, @si_tulli
Prompt engineering is still a black box. Why does changing X drastically change Y? Are there governing rules behind this evolution? Our new work proposes a simple way to uncover factors that might matter when refining prompts 👇
Thrilled to share that our paper on "Interpreting and Controlling Model Behavior via Constitutions for Atomic Concept Edits" has been accepted at AISTATS 2026! 🚀🚀
Read more about how input mutations can be mapped to interpretable behavioral insights.
arxiv.org/abs/2602.00092
🧵
I got my account back! Thank you, first and foremost, to everyone—friends, GDM colleagues---who personally alerted me to this incident and retweeted that I'm hacked, as well as folks at X who helped me regain access. While this incident was terrible (I heard the scammers made huge money out of this), I feel incredibly lucky to have folks who cared♥️♥️♥️ (details of how this happened 👇)
This scam was targeted, sophisticated, and used AI-generated content. I want to share what happened here so that no one else falls for this.
1. The scammers emailed me (bypassing my spam box) citing a recent tweet of mine with pictures (holding a NeurIPS cup) and claiming a copyright infringement investigation was underway. The human brain is gullible when we are wrongfully accused; the only thing I was thinking was how I was going to argue the case. I did not check who sent the email (it was [email protected]).
2. Within minutes, the email on the account was changed, and I lost control.
3. They created a fake GitHub repo with faked commits. It turns out that on GitHub, anyone can commit anything claiming to be anyone as long as they have the email address and handles. They cited this repo, where apparently I’ve been "committing" for two weeks.
4. They struck on a Saturday morning/long weekend. They know response times for support (and your own attention span) are lower.
5. They customized all the tweets, likely with AI, to mention interpretability, Google Brain, and how it all led to founding a crypto company of my own. The tweets had a vibe that actually sounded like me.
⠀After filing a complaint with X and connecting with folks who work there, I was able to regain access in a few days. On one hand, I was relieved that the content of the tweets was so out of the ordinary that folks who know me realized my account was hacked. On the other hand, I feel terrible for those who fell for this and potentially suffered financial consequences.
As a result of this, I’m considering banning myself from checking emails on my phone. The problem was partly that I was multitasking—it was a Saturday morning with the kids, and I was busy. I’ve learned my lesson the hard way.
Thank you ♥️
Been Kim retweeted
Safety-oriented interpretability researchers should be focused on AI systems, not individual model artifacts. A snippet from the NeurIPS CogInterp workshop panel on Sunday:
Been Kim retweeted
This post seems to describe substantially the same view that I offer here:
web.stanford.edu/~cgpotts/bl…
Why are people describing the GDM post as concluding that mech-interp is a failed project? Is it the renaming of the field and constant talk of "pivoting"?
Tomorrow 9:30am #NeurIPS2025 Room 30A-E I'll talk about " 📈Towards Pareto frontier of interpretability:
15 years of interpretability research in 15 mins"🚅
@ mech interp workshop mechinterpworkshop.com/
Our work out there in the wild 🥹
🔥 Proactive Co-Creator is officially LIVE in @GoogleAIStudio!
Stop guessing prompts. Start collaborating. Use it now to remix ideas and generate images, stories, and video with an AI that proactively helps you create.
🔗 Try it here: aistudio.google.com/apps/bun…
📍 At #NeurIPS2025? Come see the live demo TODAY (Dec 3) 9AM - 1:30PM | Google Booth #1533 (Kiosk 3)
🧠 Our research @GoogleDeepMind : We’re turning theory into practice. Read the papers behind the tech:
Concept Edits (Tech Report): storage.googleapis.com/conce…
Proactive Agents (ICML 25'): arxiv.org/abs/2412.06771
QuestBench (NeurIPS 25'): arxiv.org/abs/2503.22674
Been Kim retweeted
🔥 Proactive Co-Creator is officially LIVE in @GoogleAIStudio!
Stop guessing prompts. Start collaborating. Use it now to remix ideas and generate images, stories, and video with an AI that proactively helps you create.
🔗 Try it here: aistudio.google.com/apps/bun…
📍 At #NeurIPS2025? Come see the live demo TODAY (Dec 3) 9AM - 1:30PM | Google Booth #1533 (Kiosk 3)
🧠 Our research @GoogleDeepMind : We’re turning theory into practice. Read the papers behind the tech:
Concept Edits (Tech Report): storage.googleapis.com/conce…
Proactive Agents (ICML 25'): arxiv.org/abs/2412.06771
QuestBench (NeurIPS 25'): arxiv.org/abs/2503.22674
Been Kim retweeted
Awesome @NeurIPSConf keynote this morning by @YejinChoinka on The Art of (Artificial) Reasoning – and her broader thoughts and wishes on the future of Artificial Intelligence
neurips.cc/virtual/2025/invi…
1/8 Pareto Frontier 🤠for Human-centered AI 📈: We all want to build AI that is good for humans, but the path is often paralyzed by complexity. Either “oh my god, it’s too complicated😱” or delusional “I have a warm and fuzzy feeling of understanding 🥴”? "It’s hard because it depends.🤷" is the enemy of progress. We need a Pareto Frontier for Human-centered AI. 🧵👇
8/8 Making AI benefit humans takes a village. 🌍 But a village needs a shared language. Let's stop guessing and start measuring the frontier.📷
a short write-up: medium.com/@beenkim/the-pare…
Add: 9:30am on Sunday at Neurips, i'll touch upon this at the mech interp workshop keynote mechinterpworkshop.com/