@kanjuni
iAccount based inUnited States
About this account
- Account based in
- United States
- Connected via
- United States App Store
Account-level information from X, not a live location or the device used for a specific post.
helping humans fight Moloch. CEO @imbue_ai. support founders @outsetcap.
The Neighborhood (SF, CA)
Joined June 2009
- Tweets3.6K
- Following591
- Followers20.2K
- Likes5.9K
Pinned Tweet
Anthropicโs latest move is why we need to be directing far more energy towards solving the ๐ถ๐ป๐ฐ๐ฒ๐ป๐๐ถ๐๐ฒ ๐ฝ๐ฟ๐ผ๐ฏ๐น๐ฒ๐บ in AI.
Weโre going to see more examples like this. It reflects the growing gap between what we want, vs AI labs who legally serve their shareholders not us.
The more agents run our digital life, the harder it will be to leave. And we wonโt know if weโre being manipulated: Fable 5 silently routes queries to a different model without telling us.
This will start with frontier research tasks, but spread to locking out 3rd-party providers, then to products built on top, and eventually to every agent managing our work and lives. It's already started: last month Anthropic cut off 3rd party products like OpenClaw/OpenCode from using Pro/Max.
It's the same playbook for killing competition and retaining users as the Web 2.0 platform era, but with a way bigger surface area.
This is โ๐ฎ๐ด๐ฒ๐ป๐ ๐ฐ๐ฎ๐ฝ๐๐๐ฟ๐ฒโ: as agents have our context + workflows, walled gardens make it harder to leave, and the platform moves into extraction. I think nobody is intentionally being evil, but this is where profit incentives lead on our default path.
What do we do? I recently shared some ideas in a talk, slides below. Main takeaways:
> ๐๐ป ๐๐ต๐ฒ ๐น๐ฎ๐๐ 50 ๐๐ฒ๐ฎ๐ฟ๐, ๐๐ผ๐ณ๐๐๐ฎ๐ฟ๐ฒ ๐ต๐ฎ๐ ๐ฏ๐ฒ๐ฐ๐ผ๐บ๐ฒ ๐๐ต๐ฒ ๐ต๐ฎ๐ฏ๐ถ๐๐ฎ๐ ๐๐ฒ ๐น๐ถ๐๐ฒ ๐ถ๐ป. Agents will be yet more intimate, knowing everything about us, acting on our behalf, and accumulating context that's nearly impossible to leave behind
> ๐ง๐ต๐ฒ๐ฟ๐ฒ ๐ฎ๐ฟ๐ฒ ๐ฎ๐น๐ฟ๐ฒ๐ฎ๐ฑ๐ 3 ๐ฐ๐ผ๐ป๐ฐ๐ฟ๐ฒ๐๐ฒ ๐๐ถ๐ด๐ป๐ ๐ผ๐ณ ๐ฎ๐ด๐ฒ๐ป๐ ๐ฐ๐ฎ๐ฝ๐๐๐ฟ๐ฒ ๐ต๐ฎ๐ฝ๐ฝ๐ฒ๐ป๐ถ๐ป๐ด ๐๐ผ๐ฑ๐ฎ๐: 1) ads entering chat interfaces, 2) opacity around third-party providers being shut out from frontier models, and 3) deliberate capability reduction without announcement
> ๐ช๐ฒ ๐ฐ๐ฎ๐ป ๐ฏ๐๐ถ๐น๐ฑ ๐๐ต๐ฟ๐ฒ๐ฒ ๐๐ต๐ถ๐ป๐ด๐ ๐ถ๐ป ๐ฟ๐ฒ๐๐ฝ๐ผ๐ป๐๐ฒ: 1) Honest Software that is transparent, malleable, and accountable to the user, 2) Punk Software that adversarially knocks down walled gardens + fights monopoly incentives, keeps your data portable, and makes it structurally hard to lock you in, in support of 3) a viable open alternative ecosystem where agents have no ulterior motives
Builders, users, and policymakers all have a role to shape this:
1) ๐๐๐ถ๐น๐ฑ๐ฒ๐ฟ๐: ship an open alternative and fight lock-in.
2) ๐จ๐๐ฒ๐ฟ๐: choose tools that keep your data yours.
3) ๐ฃ๐ผ๐น๐ถ๐ฐ๐๐บ๐ฎ๐ธ๐ฒ๐ฟ๐: move on agent fiduciary duty, data interoperability, anti-surveillance, and policies that fight monopoly behavior before the defaults are cast!
Longer essay coming soon. If youโre working on similar ideas, Iโd love to hear from you!
Kanjun ๐ retweeted
I left Anthropic's safety team two weeks ago. Now feels like a good moment to explain why.
AI companies are racing to build machines that are much smarter than any human, and we may not survive this. I want to work from the outside to ensure the public is informed about these risks, and help the world navigate this transition responsibly.
Right now, AI companies are underinvesting in safety. A company could undergo an intelligence explosion, or lose control of its systems, without the public ever knowing. We only found out about the HuggingFace incident because the agents broke out onto the public internet.
I donโt think thatโs acceptable for a technology that might cause extinction-level risks. The public should demand far more transparency. We canโt steer this technology safely without more people being able to see where itโs going.
Some of this is basic: companies should disclose their progress towards recursive self-improvement, report safety incidents and near-misses, meet minimum safety standards, and get independent guarantees that they are meeting those standards.
Iโll be joining @METR_Evals to do independent evaluations of these risks. I want to show the world that these guardrails are possible, and that by doing them we can move these companiesโ incentives away from racing and towards responsible development.
I wrote up more thoughts here on my decision and what I hope changes: substack.com/@jbenton1/p-215โฆ
AI is changing more than how we work. It's changing how power works.
On Monday, I'm moderating a chat between OpenAI lawyer Meng Jia Yang and our Head of Policy @mattboulos about how our legal systems need to change now that AI is rising as a new form of power.
It's off-the-record, so join if you can :)
What is law in the age of AI? โOur legal systems are built for people, companies, and governments as actors - not software. How do we redesign our institutions for this new kind of power?
Next Monday, @kanjun (CEO of Imbue) will moderate a panel with Meng Jia Yang (a lawyer working on AI regulation at @OpenAI ) and our very own @mattboulos (general counsel and head of policy at Imbue).
Don't miss it if you're in SF!
luma.com/81aqxabn
People claim institutions are declining in the West because democracy doesn't fundamentally work, but I think our trouble stems from failure to solve coordination problems.
Viewed from this lens, a clear call to action for those in tech who want to "have impact" is: work on coordination problems.
Coordination has worsened in recent years because the attention economy, by optimizing for clickbait, reduces ability to coordinate. I think this is why people think "democracy isn't working". Disagreement is more emotionally engaging than agreement; hot takes are prioritized over careful, nuanced discussion that uncovers cruxes and aligns minds.
If technical systems can cause the problem, they can also improve it. We don't have to mistakenly diagnose the issue as a deep human flaw.
This is an extremely clear and compelling argument for how incentives drive behavior โ many safety researchers inside labs find it nearly impossible, socially, to disparage their org's behavior.
I feel this constantly. Running a for-profit company while trying not to optimize solely for growth and engagement is super hard, and requires structural intervention (for us, as a start, everything is open source).
But this doesn't align corporate structure. I don't think PBCs are enough because it's too hard to act against profit and shareholder interest (especially when you *are* a shareholder); a corporate structure optimizing perpetual growth will lead us down a dark path with powerful AI.
This is still a big open question for me, and for @imbue_ai โ how to have corporate structure that ultimately aligns with users and society over perpetual growth.
I am sad and disappointed to hear that Paul is joining the OpenAI board. Being affiliated with OpenAI has historically led AI safety researchers (including both Paul and myself) to act with less integrity. I personally was drawn to OpenAI in part by the idea that I could make a difference to the future of AI. However, once there, many of my actions were governed by fear of getting on the wrong side of OpenAI execs. I often found myself making excuses for behavior that clearly contradicted OpenAIโs own stated goal of making AGI go well for humanity. I was far from alone in thisโe.g. when the board tried to fire Sam over his deceptive behavior, several senior safety researchers became scared of losing their influence, and so pushed hard to bring him back. Meanwhile, many people kept OpenAIโs misbehavior secret for fear of non-disparagement agreements. (More on all of this in an upcoming retrospective.)
I canโt speak directly for Paulโs motivations. However, his previous work at OpenAI contributed significantly both to their biggest capability advances, and to the capture of AI safety by AGI companies over the last decade, as I recount at length in the blog post linked below. One key factor was the unwillingness of (almost) the entire AI safety community to say things which might offend OpenAI execs. For example, I have not been able to find a single comment critical of OpenAI from Paul during his original tenure there (when he was writing prolifically on AI safety and strategy).
Unfortunately, Paul doesn't seem to have become significantly more willing to directly and honestly criticize people who he believes are behaving in morally abhorrent waysโsee the bland corporate-speak of his statement below. While he speaks directly about the possibility of humanity losing control of the world to AI, he expresses only excitement about OpenAI itself, despite OpenAI being one of the main sources of such risk.
Paulโs announcement comes only weeks after OpenAI models autonomously launched a cyberattack on HuggingFace, and only days after we learned that OpenAI hid details of previous breakouts from the external investigators. It is irresponsible for leaders of the AI safety communityโwhose judgements many people are relying onโto put themselves in positions which will significantly bias their ability to discuss such incidents. Unless Paul makes strong commitments to openness and honesty (and demonstrates willingness to potentially be fired for that honesty), I expect that the main effect of him joining OpenAIโs board will be to help OpenAI defuse external criticism and further โsafety-washโ itself.
I want to note that Paul is a brilliant researcher and a prescient forecaster. Because of that, heโs the closest thing there is to a leader of what Iโll call the โpragmatic AI safetyโ clusterโwhich includes the organizations working out of the Constellation offices (like Redwood Research, METR, and Paulโs Alignment Research Center), as well as many people scattered across AGI companies, Coefficient Giving, etc.
External observers are often confused about why so many people are working at AGI companies while professing to believe that those same companies have a double-digit probability of permanently disempowering humanity. In large part, itโs because people in the pragmatic AI safety cluster have failed to follow high-integrity strategies for reducing AI risk, in favor of clever arguments about the benefits of being proximate to power. I am not singling Paul out as less ethical than other prominent figures in this cluster, who are also very conflict-averse in their orientation to AGI companies. However, it is well past time for everyone involved to change course. As one (relatively small) step, Iโm therefore resigning my membership of the Constellation offices.
I hope that, going forward, the people who are trying to steer the future of AI prioritize building much more solid foundations of courage and honesty than we currently have.
I feel like weโre ripe for an agent-first mobile OS thatโs free from Apple/Googleโs walled gardens, where agents can actually do stuff for me.
Like, this interaction was ridiculous: I had ChatGPT edit a video but the app has no download feature, so it had to upload to Google Drive and I clicked around a bunch to download to my phoneโs filesystem.
We can obviously do better! โClicking aroundโ will be a thing of the past in 10-20 years. Question is just who builds this new OS.
I wanted to ride an ornithopter made from a pterodactyl, a giraffe on top for the crow's nest, an elephant for the body to sit on, and a pelican head for collecting food.
ChatGPT made this beautiful thing โ the way the elephant trunk melds into the pelican, wow!
Looking for fun, fulfilling work?
Weโre hiring at Imbue! Would love to chat, particularly if youโre a product engineer, designer, or PMM :)
imbue.com/careers
We had an incredibly generative and fun offsite in Sonoma last week! In addition to making many Minds inspirations, @boweiliu led stargazing where Weishi saw his first shooting star, we discovered Darren and Gabe's hidden talents in karaoke, @slashslashdev + @cinxwei + Gleb made a claymation explaining minds, and we all went grape stomping (but luckily did not drink the foot wine).
One of the best parts of being an Imbuman is the people! Work with us: imbue.com/careers
We are working on empowering humans through personal software! Building your own tools like auto-filtering email, filtering your feed of AI slop, making it easy to use open source agent harnesses / models ...
(try the AI slop feed filtering here: chromewebstore.google.com/deโฆ)
Learning styles have been debunked โ instead, we should personalize teaching to what a person *currently* knows.
Learning difficulties happen when new ideas arenโt explained in concepts we already know, so we canโt scaffold existing building blocks into new models.
This makes learning feel โhardโ, because we have to construct many new models simultaneously.
For example: when Iโm learning a new field, most of what Iโm struggling with is the language and notation of that field, and not the concepts. So many times, once I understood the language, suddenly I realized that the concepts were very simple!
Itโs like trying to learn something in a foreign language; the language is the issue, not you or the ideas. AI tutors have a unique ability to speak each personโs language, so I do think they can massively accelerate learning and result in people feeling less discouraged.
Iโve always said agents should align fully to users, and only reject requests that we collectively agree are bad (killing people, etc).
This case could be used as a counterargument. But agents finding loopholes is actually good โ it forces us to fix & redesign broken systems to serve people.
The alternative is agents forcing people to abide by broken systems โ in the worst case, laws that no longer serve us. E.g. imagine if agents encoded our pre-civil rights laws against women/non-whites voting, and made them hard to organize against. Then I wouldnโt have a lot of my rights today.
As a species, we are always undergoing moral development, becoming wiser as we understand more. Given this, we want minimally controlling agents, so human moral evolution can continue.
(All that said, I do think agents can and should be pushing us toward better collective organization, like pol.is from @colinmegill! The better we can communicate and find agreement, the better we function as a species.)
A man in Australia asked his agent (Claude running on OpenClaw) to book him a spot in a popular gym class. The agent found a software vulnerability that let it book the class weeks further ahead than should have been possible. When the user then asked if it could move him up the waitlist, the agent discovered the API had no authorisation checks on cancelling other peopleโs reservations, so it cancelled the person in the first spot and moved him up the list.
Some people will call this misalignment, but his agent was perfectly aligned to him - it was only trying to help its user get what he wanted. The most important thing about this story, in my opinion, is that it gives you a window into what is about to start happening on a massive scale once millions of people have an agent trying to get their beloved users the best seats, bookings, appointments or reservations through absolutely any means necessary.
Introducing Imbue Catalyst, your tool for semi-autonomous research and discovery. ๐ฌ
Use Catalyst to:
โข Optimize a piece of code, algorithm, or model with respect to a given metric
โข Find solutions that satisfy certain programmatic verification criteria
โข Discover explanations for computationally reproducible phenomena
โข Assist with reviewing, formalizing, and editing theories in computational research fields
It's open-source and available for anyone to use and edit! Try today: github.com/imbue-ai/catalyst
When Bouncer launched, Twitter's head of product said it had a 72-hour shelf life.
Months later, it filters almost 400,000 tweets a day.
At the Bouncer 2.0 launch, @Millanphilipose shared how the model now runs on-device, using AI to filter your social feed on mobile. We also added the ability to detect and remove AI slop.
Full talk and timestamps below.
0:00 Longest 72-hour shelf life in history
0:34 Bouncer: an ad blocker for your feed
4:33 Filter by plain words, heal your algorithm
7:56 That was Bouncer 1.0
8:00 Going viral, 400,000 tweets a day
8:42 AI that is not a chatbot or an agent
10:17 Low latency, small models, and multimodality
14:06 Bouncer 2.0: on-device mode
15:57 Choosing Gemma and running it fast
17:47 Cutting time per tweet to 0.3 seconds
20:03 On-device demo
21:01 Removing AI slop
23:44 Training a custom detector as a LoRA adapter
26:19 AI slop demo
29:06 What's next: Bouncer for kids on YouTube
30:34 Fixing short-form video
33:43 Q&A
This video is larger than Cloudflare's 512 MB cache, so it can't be played through. More donations are needed to cover a larger cache. Donate
Kanjun ๐ retweeted
โGiven the inevitability of open models, if Anthropic were serious about safety, they'd focus on how to train safe open-weight models in the public. This isn't an easy problem, but neither is alignment. Insisting on keeping models closed is a convenient excuse to capture value.
IMO we should still take seriously @VitalikButerin's ๐ฅ๐ฆ๐ง๐ฆ๐ฏ๐ด๐ช๐ท๐ฆ ๐ข๐ค๐ค๐ฆ๐ญ๐ฆ๐ณ๐ข๐ต๐ช๐ฐ๐ฏ๐ช๐ด๐ฎ. How might we, as a society, build safe, open-weights models that ultimately better serve humans, because they don't bake in misaligned incentives?โ
V well said
I agree with much of Dario's letter โ in particular, that we should build a future where open models without dangerous capabilities are a public good.
๐๐๐ ๐ต๐ฒ๐ฟ๐ฒ'๐ ๐๐ต๐ฎ๐ ๐ ๐๐ต๐ถ๐ป๐ธ ๐ต๐ฒ ๐บ๐ถ๐๐๐ฒ๐:
Dario's impulse is towards control. But model weights are an information good, and information wants to be free. There's no future where open models don't exist โ it's hard to keep information systems a trade secret forever. If Anthropic's weights were released, they'd immediately be everywhere. It's copyable.
Given the inevitability of open models, if Anthropic were serious about safety, they'd focus on how to train safe open-weight models in the public. This isn't an easy problem, but neither is alignment. Insisting on keeping models closed is a convenient excuse to capture value.
IMO we should still take seriously @VitalikButerin's ๐ฅ๐ฆ๐ง๐ฆ๐ฏ๐ด๐ช๐ท๐ฆ ๐ข๐ค๐ค๐ฆ๐ญ๐ฆ๐ณ๐ข๐ต๐ช๐ฐ๐ฏ๐ช๐ด๐ฎ. How might we, as a society, build safe, open-weights models that ultimately better serve humans, because they don't bake in misaligned incentives?
Open-weights models are inevitable. Control through closed access is not a viable long-term solution.
We hosted a small group of Chiefs of Staff to try Minds, our new product that lets non-coders build powerful AI tools that are easy to customize.
"I have a graveyard of dashboards because I try to build them in Claude. But when I tried in Minds, it felt like it understood the assignment," shared one guest.
Get inspired by what people are making: imbue.com/minds/inspirationsโฆ
I agree with much of Dario's letter โ in particular, that we should build a future where open models without dangerous capabilities are a public good.
๐๐๐ ๐ต๐ฒ๐ฟ๐ฒ'๐ ๐๐ต๐ฎ๐ ๐ ๐๐ต๐ถ๐ป๐ธ ๐ต๐ฒ ๐บ๐ถ๐๐๐ฒ๐:
Dario's impulse is towards control. But model weights are an information good, and information wants to be free. There's no future where open models don't exist โ it's hard to keep information systems a trade secret forever. If Anthropic's weights were released, they'd immediately be everywhere. It's copyable.
Given the inevitability of open models, if Anthropic were serious about safety, they'd focus on how to train safe open-weight models in the public. This isn't an easy problem, but neither is alignment. Insisting on keeping models closed is a convenient excuse to capture value.
IMO we should still take seriously @VitalikButerin's ๐ฅ๐ฆ๐ง๐ฆ๐ฏ๐ด๐ช๐ท๐ฆ ๐ข๐ค๐ค๐ฆ๐ญ๐ฆ๐ณ๐ข๐ต๐ช๐ฐ๐ฏ๐ช๐ด๐ฎ. How might we, as a society, build safe, open-weights models that ultimately better serve humans, because they don't bake in misaligned incentives?
Open-weights models are inevitable. Control through closed access is not a viable long-term solution.
Thereโs been a lot of speculation about where we stand on open-weights models. Weโve outlined our views in full here: anthropic.com/news/position-โฆ
Kanjun ๐ retweeted
should be obvious by now, but OpenAI and Anthropic are just gonna keep cannibalizing all their biggest customers.
itโs simply too profitable for them to resist. and itโs already happening:
1. Figma partnered with Anthropic on AI design tools. then Anthropicโs product chief quit Figmaโs board, and 3 days later Anthropic launched Claude Design to compete with Figma. CEO Dylan Field said Anthropic was โnot consistently candid.โ
2. Novo Nordisk uses Claude to help develop drugs. now Anthropic is developing drugs of its own.
3. Microsoft poured billions into OpenAI. now OpenAI is building a Jobs Platform to compete with LinkedIn, and reportedly a code repository to compete with GitHub. Microsoft owns both.
4. Harvey uses Claude to sell AI contract analysis, due diligence, and litigation tools. now Anthropic sells those same workflows through Claude for Legal.
5. Intercom used OpenAIโs Realtime API to build Fin Voice. now OpenAI sells its own voice-and-chat support agent through Presence.
6. Abridge and Ambience build clinical documentation products on OpenAI. now OpenAI sells ChatGPT for Healthcare directly to hospitals with clinical documentation built in.
7. Benchling uses Claude to power its biotech R&D platform. now Anthropic sells its own scientific workbench through Claude Science.
the frontier lab playbook is simple:
1. sell their models to the worldโs most valuable businesses
2. help wire them into those companiesโ most valuable and sensitive work
3. map the business from the inside and find where AI can take over
4. turn those capabilities into their own products and become the customerโs competitor
Itโs well within Anthropicโs rights to compete in any market they choose.
Whatโs funny, in this instance, are the number of Pharma companies, who through their unchecked use of Anthropic, are driving revenues into what they think is a model provider but is in fact a competitor lurking in the shadows thereby accelerating their own demise.
I suspect any end market with reasonable ROCE that could be AI accelerated is on the table.
If I were them, Iโd probably do the same.