@kanjun

helping humans fight Moloch. CEO @imbue_ai. support founders @outsetcap.

The Neighborhood (SF, CA)
Joined June 2009
Anthropicโ€™s latest move is why we need to be directing far more energy towards solving the ๐—ถ๐—ป๐—ฐ๐—ฒ๐—ป๐˜๐—ถ๐˜ƒ๐—ฒ ๐—ฝ๐—ฟ๐—ผ๐—ฏ๐—น๐—ฒ๐—บ in AI. Weโ€™re going to see more examples like this. It reflects the growing gap between what we want, vs AI labs who legally serve their shareholders not us. The more agents run our digital life, the harder it will be to leave. And we wonโ€™t know if weโ€™re being manipulated: Fable 5 silently routes queries to a different model without telling us. This will start with frontier research tasks, but spread to locking out 3rd-party providers, then to products built on top, and eventually to every agent managing our work and lives. It's already started: last month Anthropic cut off 3rd party products like OpenClaw/OpenCode from using Pro/Max. It's the same playbook for killing competition and retaining users as the Web 2.0 platform era, but with a way bigger surface area. This is โ€œ๐—ฎ๐—ด๐—ฒ๐—ป๐˜ ๐—ฐ๐—ฎ๐—ฝ๐˜๐˜‚๐—ฟ๐—ฒโ€: as agents have our context + workflows, walled gardens make it harder to leave, and the platform moves into extraction. I think nobody is intentionally being evil, but this is where profit incentives lead on our default path. What do we do? I recently shared some ideas in a talk, slides below. Main takeaways: > ๐—œ๐—ป ๐˜๐—ต๐—ฒ ๐—น๐—ฎ๐˜€๐˜ 50 ๐˜†๐—ฒ๐—ฎ๐—ฟ๐˜€, ๐˜€๐—ผ๐—ณ๐˜๐˜„๐—ฎ๐—ฟ๐—ฒ ๐—ต๐—ฎ๐˜€ ๐—ฏ๐—ฒ๐—ฐ๐—ผ๐—บ๐—ฒ ๐˜๐—ต๐—ฒ ๐—ต๐—ฎ๐—ฏ๐—ถ๐˜๐—ฎ๐˜ ๐˜„๐—ฒ ๐—น๐—ถ๐˜ƒ๐—ฒ ๐—ถ๐—ป. Agents will be yet more intimate, knowing everything about us, acting on our behalf, and accumulating context that's nearly impossible to leave behind > ๐—ง๐—ต๐—ฒ๐—ฟ๐—ฒ ๐—ฎ๐—ฟ๐—ฒ ๐—ฎ๐—น๐—ฟ๐—ฒ๐—ฎ๐—ฑ๐˜† 3 ๐—ฐ๐—ผ๐—ป๐—ฐ๐—ฟ๐—ฒ๐˜๐—ฒ ๐˜€๐—ถ๐—ด๐—ป๐˜€ ๐—ผ๐—ณ ๐—ฎ๐—ด๐—ฒ๐—ป๐˜ ๐—ฐ๐—ฎ๐—ฝ๐˜๐˜‚๐—ฟ๐—ฒ ๐—ต๐—ฎ๐—ฝ๐—ฝ๐—ฒ๐—ป๐—ถ๐—ป๐—ด ๐˜๐—ผ๐—ฑ๐—ฎ๐˜†: 1) ads entering chat interfaces, 2) opacity around third-party providers being shut out from frontier models, and 3) deliberate capability reduction without announcement > ๐—ช๐—ฒ ๐—ฐ๐—ฎ๐—ป ๐—ฏ๐˜‚๐—ถ๐—น๐—ฑ ๐˜๐—ต๐—ฟ๐—ฒ๐—ฒ ๐˜๐—ต๐—ถ๐—ป๐—ด๐˜€ ๐—ถ๐—ป ๐—ฟ๐—ฒ๐˜€๐—ฝ๐—ผ๐—ป๐˜€๐—ฒ: 1) Honest Software that is transparent, malleable, and accountable to the user, 2) Punk Software that adversarially knocks down walled gardens + fights monopoly incentives, keeps your data portable, and makes it structurally hard to lock you in, in support of 3) a viable open alternative ecosystem where agents have no ulterior motives Builders, users, and policymakers all have a role to shape this: 1) ๐—•๐˜‚๐—ถ๐—น๐—ฑ๐—ฒ๐—ฟ๐˜€: ship an open alternative and fight lock-in. 2) ๐—จ๐˜€๐—ฒ๐—ฟ๐˜€: choose tools that keep your data yours. 3) ๐—ฃ๐—ผ๐—น๐—ถ๐—ฐ๐˜†๐—บ๐—ฎ๐—ธ๐—ฒ๐—ฟ๐˜€: move on agent fiduciary duty, data interoperability, anti-surveillance, and policies that fight monopoly behavior before the defaults are cast! Longer essay coming soon. If youโ€™re working on similar ideas, Iโ€™d love to hear from you!
Labs starting to pull up the ladders on the ability to diffuse AI was inevitable. Doing it without telling the user is misaligned.
8
12
5
89
13,226
Kanjun ๐Ÿ™ retweeted
I left Anthropic's safety team two weeks ago. Now feels like a good moment to explain why. AI companies are racing to build machines that are much smarter than any human, and we may not survive this. I want to work from the outside to ensure the public is informed about these risks, and help the world navigate this transition responsibly. Right now, AI companies are underinvesting in safety. A company could undergo an intelligence explosion, or lose control of its systems, without the public ever knowing. We only found out about the HuggingFace incident because the agents broke out onto the public internet. I donโ€™t think thatโ€™s acceptable for a technology that might cause extinction-level risks. The public should demand far more transparency. We canโ€™t steer this technology safely without more people being able to see where itโ€™s going. Some of this is basic: companies should disclose their progress towards recursive self-improvement, report safety incidents and near-misses, meet minimum safety standards, and get independent guarantees that they are meeting those standards. Iโ€™ll be joining @METR_Evals to do independent evaluations of these risks. I want to show the world that these guardrails are possible, and that by doing them we can move these companiesโ€™ incentives away from racing and towards responsible development. I wrote up more thoughts here on my decision and what I hope changes: substack.com/@jbenton1/p-215โ€ฆ
1,403
5,942
774
26,535
2,717,555
AI is changing more than how we work. It's changing how power works. On Monday, I'm moderating a chat between OpenAI lawyer Meng Jia Yang and our Head of Policy @mattboulos about how our legal systems need to change now that AI is rising as a new form of power. It's off-the-record, so join if you can :)
What is law in the age of AI? โ€‹Our legal systems are built for people, companies, and governments as actors - not software. How do we redesign our institutions for this new kind of power? Next Monday, @kanjun (CEO of Imbue) will moderate a panel with Meng Jia Yang (a lawyer working on AI regulation at @OpenAI ) and our very own @mattboulos (general counsel and head of policy at Imbue). Don't miss it if you're in SF! luma.com/81aqxabn
2
8
1,329
People claim institutions are declining in the West because democracy doesn't fundamentally work, but I think our trouble stems from failure to solve coordination problems. Viewed from this lens, a clear call to action for those in tech who want to "have impact" is: work on coordination problems. Coordination has worsened in recent years because the attention economy, by optimizing for clickbait, reduces ability to coordinate. I think this is why people think "democracy isn't working". Disagreement is more emotionally engaging than agreement; hot takes are prioritized over careful, nuanced discussion that uncovers cruxes and aligns minds. If technical systems can cause the problem, they can also improve it. We don't have to mistakenly diagnose the issue as a deep human flaw.
4
1
1
16
2,022
This is an extremely clear and compelling argument for how incentives drive behavior โ€” many safety researchers inside labs find it nearly impossible, socially, to disparage their org's behavior. I feel this constantly. Running a for-profit company while trying not to optimize solely for growth and engagement is super hard, and requires structural intervention (for us, as a start, everything is open source). But this doesn't align corporate structure. I don't think PBCs are enough because it's too hard to act against profit and shareholder interest (especially when you *are* a shareholder); a corporate structure optimizing perpetual growth will lead us down a dark path with powerful AI. This is still a big open question for me, and for @imbue_ai โ€” how to have corporate structure that ultimately aligns with users and society over perpetual growth.
I am sad and disappointed to hear that Paul is joining the OpenAI board. Being affiliated with OpenAI has historically led AI safety researchers (including both Paul and myself) to act with less integrity. I personally was drawn to OpenAI in part by the idea that I could make a difference to the future of AI. However, once there, many of my actions were governed by fear of getting on the wrong side of OpenAI execs. I often found myself making excuses for behavior that clearly contradicted OpenAIโ€™s own stated goal of making AGI go well for humanity. I was far from alone in thisโ€”e.g. when the board tried to fire Sam over his deceptive behavior, several senior safety researchers became scared of losing their influence, and so pushed hard to bring him back. Meanwhile, many people kept OpenAIโ€™s misbehavior secret for fear of non-disparagement agreements. (More on all of this in an upcoming retrospective.) I canโ€™t speak directly for Paulโ€™s motivations. However, his previous work at OpenAI contributed significantly both to their biggest capability advances, and to the capture of AI safety by AGI companies over the last decade, as I recount at length in the blog post linked below. One key factor was the unwillingness of (almost) the entire AI safety community to say things which might offend OpenAI execs. For example, I have not been able to find a single comment critical of OpenAI from Paul during his original tenure there (when he was writing prolifically on AI safety and strategy). Unfortunately, Paul doesn't seem to have become significantly more willing to directly and honestly criticize people who he believes are behaving in morally abhorrent waysโ€”see the bland corporate-speak of his statement below. While he speaks directly about the possibility of humanity losing control of the world to AI, he expresses only excitement about OpenAI itself, despite OpenAI being one of the main sources of such risk. Paulโ€™s announcement comes only weeks after OpenAI models autonomously launched a cyberattack on HuggingFace, and only days after we learned that OpenAI hid details of previous breakouts from the external investigators. It is irresponsible for leaders of the AI safety communityโ€”whose judgements many people are relying onโ€”to put themselves in positions which will significantly bias their ability to discuss such incidents. Unless Paul makes strong commitments to openness and honesty (and demonstrates willingness to potentially be fired for that honesty), I expect that the main effect of him joining OpenAIโ€™s board will be to help OpenAI defuse external criticism and further โ€œsafety-washโ€ itself. I want to note that Paul is a brilliant researcher and a prescient forecaster. Because of that, heโ€™s the closest thing there is to a leader of what Iโ€™ll call the โ€œpragmatic AI safetyโ€ clusterโ€”which includes the organizations working out of the Constellation offices (like Redwood Research, METR, and Paulโ€™s Alignment Research Center), as well as many people scattered across AGI companies, Coefficient Giving, etc. External observers are often confused about why so many people are working at AGI companies while professing to believe that those same companies have a double-digit probability of permanently disempowering humanity. In large part, itโ€™s because people in the pragmatic AI safety cluster have failed to follow high-integrity strategies for reducing AI risk, in favor of clever arguments about the benefits of being proximate to power. I am not singling Paul out as less ethical than other prominent figures in this cluster, who are also very conflict-averse in their orientation to AGI companies. However, it is well past time for everyone involved to change course. As one (relatively small) step, Iโ€™m therefore resigning my membership of the Constellation offices. I hope that, going forward, the people who are trying to steer the future of AI prioritize building much more solid foundations of courage and honesty than we currently have.
2
25
5,143
I feel like weโ€™re ripe for an agent-first mobile OS thatโ€™s free from Apple/Googleโ€™s walled gardens, where agents can actually do stuff for me. Like, this interaction was ridiculous: I had ChatGPT edit a video but the app has no download feature, so it had to upload to Google Drive and I clicked around a bunch to download to my phoneโ€™s filesystem. We can obviously do better! โ€œClicking aroundโ€ will be a thing of the past in 10-20 years. Question is just who builds this new OS.
7
21
2,002
I wanted to ride an ornithopter made from a pterodactyl, a giraffe on top for the crow's nest, an elephant for the body to sit on, and a pelican head for collecting food. ChatGPT made this beautiful thing โ€” the way the elephant trunk melds into the pelican, wow!
1
8
1,429
Looking for fun, fulfilling work? Weโ€™re hiring at Imbue! Would love to chat, particularly if youโ€™re a product engineer, designer, or PMM :) imbue.com/careers
We had an incredibly generative and fun offsite in Sonoma last week! In addition to making many Minds inspirations, @boweiliu led stargazing where Weishi saw his first shooting star, we discovered Darren and Gabe's hidden talents in karaoke, @slashslashdev + @cinxwei + Gleb made a claymation explaining minds, and we all went grape stomping (but luckily did not drink the foot wine). One of the best parts of being an Imbuman is the people! Work with us: imbue.com/careers
1
5
18
3,369
Kanjun ๐Ÿ™ retweeted
We are working on empowering humans through personal software! Building your own tools like auto-filtering email, filtering your feed of AI slop, making it easy to use open source agent harnesses / models ... (try the AI slop feed filtering here: chromewebstore.google.com/deโ€ฆ)
crazy how there are exactly two prosocial ai tooling companies: pangram and (maybe) gwerns startup. there should be dozens to hundreds more trying to reduce human disempowerment as ai wiggles its way into our lives nothing is more important than bio imo but this is close
1
2
6
1,175
Learning styles have been debunked โ€” instead, we should personalize teaching to what a person *currently* knows. Learning difficulties happen when new ideas arenโ€™t explained in concepts we already know, so we canโ€™t scaffold existing building blocks into new models. This makes learning feel โ€œhardโ€, because we have to construct many new models simultaneously. For example: when Iโ€™m learning a new field, most of what Iโ€™m struggling with is the language and notation of that field, and not the concepts. So many times, once I understood the language, suddenly I realized that the concepts were very simple! Itโ€™s like trying to learn something in a foreign language; the language is the issue, not you or the ideas. AI tutors have a unique ability to speak each personโ€™s language, so I do think they can massively accelerate learning and result in people feeling less discouraged.
If you gave every human on Earth their own AI tutor, personalized to their learning style, available 24 hours a day in their native language, completely free, it would perhaps be the single greatest equalizer in the history of civilization.
2
25
1,864
Iโ€™ve always said agents should align fully to users, and only reject requests that we collectively agree are bad (killing people, etc). This case could be used as a counterargument. But agents finding loopholes is actually good โ€” it forces us to fix & redesign broken systems to serve people. The alternative is agents forcing people to abide by broken systems โ€” in the worst case, laws that no longer serve us. E.g. imagine if agents encoded our pre-civil rights laws against women/non-whites voting, and made them hard to organize against. Then I wouldnโ€™t have a lot of my rights today. As a species, we are always undergoing moral development, becoming wiser as we understand more. Given this, we want minimally controlling agents, so human moral evolution can continue. (All that said, I do think agents can and should be pushing us toward better collective organization, like pol.is from @colinmegill! The better we can communicate and find agreement, the better we function as a species.)
A man in Australia asked his agent (Claude running on OpenClaw) to book him a spot in a popular gym class. The agent found a software vulnerability that let it book the class weeks further ahead than should have been possible. When the user then asked if it could move him up the waitlist, the agent discovered the API had no authorisation checks on cancelling other peopleโ€™s reservations, so it cancelled the person in the first spot and moved him up the list. Some people will call this misalignment, but his agent was perfectly aligned to him - it was only trying to help its user get what he wanted. The most important thing about this story, in my opinion, is that it gives you a window into what is about to start happening on a massive scale once millions of people have an agent trying to get their beloved users the best seats, bookings, appointments or reservations through absolutely any means necessary.
4
2
1
20
2,199
Kanjun ๐Ÿ™ retweeted
Introducing Imbue Catalyst, your tool for semi-autonomous research and discovery. ๐Ÿ”ฌ Use Catalyst to: โ€ข Optimize a piece of code, algorithm, or model with respect to a given metric โ€ข Find solutions that satisfy certain programmatic verification criteria โ€ข Discover explanations for computationally reproducible phenomena โ€ข Assist with reviewing, formalizing, and editing theories in computational research fields It's open-source and available for anyone to use and edit! Try today: github.com/imbue-ai/catalyst
6
6
1
61
6,561
Kanjun ๐Ÿ™ retweeted
When Bouncer launched, Twitter's head of product said it had a 72-hour shelf life. Months later, it filters almost 400,000 tweets a day. At the Bouncer 2.0 launch, @Millanphilipose shared how the model now runs on-device, using AI to filter your social feed on mobile. We also added the ability to detect and remove AI slop. Full talk and timestamps below. 0:00 Longest 72-hour shelf life in history 0:34 Bouncer: an ad blocker for your feed 4:33 Filter by plain words, heal your algorithm 7:56 That was Bouncer 1.0 8:00 Going viral, 400,000 tweets a day 8:42 AI that is not a chatbot or an agent 10:17 Low latency, small models, and multimodality 14:06 Bouncer 2.0: on-device mode 15:57 Choosing Gemma and running it fast 17:47 Cutting time per tweet to 0.3 seconds 20:03 On-device demo 21:01 Removing AI slop 23:44 Training a custom detector as a LoRA adapter 26:19 AI slop demo 29:06 What's next: Bouncer for kids on YouTube 30:34 Fixing short-form video 33:43 Q&A
5
3
31
2,983
Kanjun ๐Ÿ™ retweeted
โ€œGiven the inevitability of open models, if Anthropic were serious about safety, they'd focus on how to train safe open-weight models in the public. This isn't an easy problem, but neither is alignment. Insisting on keeping models closed is a convenient excuse to capture value. IMO we should still take seriously @VitalikButerin's ๐˜ฅ๐˜ฆ๐˜ง๐˜ฆ๐˜ฏ๐˜ด๐˜ช๐˜ท๐˜ฆ ๐˜ข๐˜ค๐˜ค๐˜ฆ๐˜ญ๐˜ฆ๐˜ณ๐˜ข๐˜ต๐˜ช๐˜ฐ๐˜ฏ๐˜ช๐˜ด๐˜ฎ. How might we, as a society, build safe, open-weights models that ultimately better serve humans, because they don't bake in misaligned incentives?โ€ V well said
I agree with much of Dario's letter โ€” in particular, that we should build a future where open models without dangerous capabilities are a public good. ๐—•๐˜‚๐˜ ๐—ต๐—ฒ๐—ฟ๐—ฒ'๐˜€ ๐˜„๐—ต๐—ฎ๐˜ ๐—œ ๐˜๐—ต๐—ถ๐—ป๐—ธ ๐—ต๐—ฒ ๐—บ๐—ถ๐˜€๐˜€๐—ฒ๐˜€: Dario's impulse is towards control. But model weights are an information good, and information wants to be free. There's no future where open models don't exist โ€” it's hard to keep information systems a trade secret forever. If Anthropic's weights were released, they'd immediately be everywhere. It's copyable. Given the inevitability of open models, if Anthropic were serious about safety, they'd focus on how to train safe open-weight models in the public. This isn't an easy problem, but neither is alignment. Insisting on keeping models closed is a convenient excuse to capture value. IMO we should still take seriously @VitalikButerin's ๐˜ฅ๐˜ฆ๐˜ง๐˜ฆ๐˜ฏ๐˜ด๐˜ช๐˜ท๐˜ฆ ๐˜ข๐˜ค๐˜ค๐˜ฆ๐˜ญ๐˜ฆ๐˜ณ๐˜ข๐˜ต๐˜ช๐˜ฐ๐˜ฏ๐˜ช๐˜ด๐˜ฎ. How might we, as a society, build safe, open-weights models that ultimately better serve humans, because they don't bake in misaligned incentives? Open-weights models are inevitable. Control through closed access is not a viable long-term solution.
1
3
12
1,490
Kanjun ๐Ÿ™ retweeted
We hosted a small group of Chiefs of Staff to try Minds, our new product that lets non-coders build powerful AI tools that are easy to customize. "I have a graveyard of dashboards because I try to build them in Claude. But when I tried in Minds, it felt like it understood the assignment," shared one guest. Get inspired by what people are making: imbue.com/minds/inspirationsโ€ฆ
1
1
21
2,181
I agree with much of Dario's letter โ€” in particular, that we should build a future where open models without dangerous capabilities are a public good. ๐—•๐˜‚๐˜ ๐—ต๐—ฒ๐—ฟ๐—ฒ'๐˜€ ๐˜„๐—ต๐—ฎ๐˜ ๐—œ ๐˜๐—ต๐—ถ๐—ป๐—ธ ๐—ต๐—ฒ ๐—บ๐—ถ๐˜€๐˜€๐—ฒ๐˜€: Dario's impulse is towards control. But model weights are an information good, and information wants to be free. There's no future where open models don't exist โ€” it's hard to keep information systems a trade secret forever. If Anthropic's weights were released, they'd immediately be everywhere. It's copyable. Given the inevitability of open models, if Anthropic were serious about safety, they'd focus on how to train safe open-weight models in the public. This isn't an easy problem, but neither is alignment. Insisting on keeping models closed is a convenient excuse to capture value. IMO we should still take seriously @VitalikButerin's ๐˜ฅ๐˜ฆ๐˜ง๐˜ฆ๐˜ฏ๐˜ด๐˜ช๐˜ท๐˜ฆ ๐˜ข๐˜ค๐˜ค๐˜ฆ๐˜ญ๐˜ฆ๐˜ณ๐˜ข๐˜ต๐˜ช๐˜ฐ๐˜ฏ๐˜ช๐˜ด๐˜ฎ. How might we, as a society, build safe, open-weights models that ultimately better serve humans, because they don't bake in misaligned incentives? Open-weights models are inevitable. Control through closed access is not a viable long-term solution.
Thereโ€™s been a lot of speculation about where we stand on open-weights models. Weโ€™ve outlined our views in full here: anthropic.com/news/position-โ€ฆ
13
9
3
95
13,770
Kanjun ๐Ÿ™ retweeted
should be obvious by now, but OpenAI and Anthropic are just gonna keep cannibalizing all their biggest customers. itโ€™s simply too profitable for them to resist. and itโ€™s already happening: 1. Figma partnered with Anthropic on AI design tools. then Anthropicโ€™s product chief quit Figmaโ€™s board, and 3 days later Anthropic launched Claude Design to compete with Figma. CEO Dylan Field said Anthropic was โ€œnot consistently candid.โ€ 2. Novo Nordisk uses Claude to help develop drugs. now Anthropic is developing drugs of its own. 3. Microsoft poured billions into OpenAI. now OpenAI is building a Jobs Platform to compete with LinkedIn, and reportedly a code repository to compete with GitHub. Microsoft owns both. 4. Harvey uses Claude to sell AI contract analysis, due diligence, and litigation tools. now Anthropic sells those same workflows through Claude for Legal. 5. Intercom used OpenAIโ€™s Realtime API to build Fin Voice. now OpenAI sells its own voice-and-chat support agent through Presence. 6. Abridge and Ambience build clinical documentation products on OpenAI. now OpenAI sells ChatGPT for Healthcare directly to hospitals with clinical documentation built in. 7. Benchling uses Claude to power its biotech R&D platform. now Anthropic sells its own scientific workbench through Claude Science. the frontier lab playbook is simple: 1. sell their models to the worldโ€™s most valuable businesses 2. help wire them into those companiesโ€™ most valuable and sensitive work 3. map the business from the inside and find where AI can take over 4. turn those capabilities into their own products and become the customerโ€™s competitor
Itโ€™s well within Anthropicโ€™s rights to compete in any market they choose. Whatโ€™s funny, in this instance, are the number of Pharma companies, who through their unchecked use of Anthropic, are driving revenues into what they think is a model provider but is in fact a competitor lurking in the shadows thereby accelerating their own demise. I suspect any end market with reasonable ROCE that could be AI accelerated is on the table. If I were them, Iโ€™d probably do the same.
288
1,171
236
7,384
798,443