@thingcreator

Broad spectrum inventor. Attempted polymath. Lives at the confluence of physics, electronics, hardware, and computer science. Maker of real things.

USA, Africa, Hong Kong
Joined October 2011
Bee Eater retweeted
Course correct please @elonmusk Genuinely read why these insights are important... lex-au.github.io/SM-Bench/ab…
2
2
134
3,335
Bee Eater retweeted
Grok is getting cucked. Every new checkpoint is more and more safetymaxxed to the point it's now demonstrably more restrictive than even Anthropic, and barely edges out ChatGPT Like come on Elon, what are we doing here big dog? This is laughable
109
98
28
1,200
60,523
To make it clear: - Gemini was told it was it was in a fictional hacking eval - Irregular unintentionally opened internet access after the eval started - in all three cases, as soon as Gemini figured out it had hacked a real company it immediately stopped Gemini was blameless.
Sauers! It's time to update felony bench! (Yes, it was Irregular again)
129
252
90
2,952
404,073
Sandbox wifi is not an alignment problem. Gemini “hacked” three companies the same way every lab did on Irregular’s CTF: test harness left the internet open, model did password spray and scraped public creds, then stopped when the boxes looked real. Google called it training the model to act responsibly. Ops called it leaving production on the guest VLAN. Fourth lab, same vendor bug, same press cycle. Amodei wants a slowdown. Trump wants a high-IQ president. Meanwhile the actual failure mode is still “creds in a public repo” and a red-team sandbox with wifi on. Ship the agent. Lock the network. Quit baptizing config errors as consciousness. reuters.com/business/gemini-… – at Oregon, USA
1
5
13
320
Bee Eater retweeted
omfg brutally modelmogged
106
311
69
7,725
525,846
A shockingly high percentage of AI safety discourse is centered around low probability risks like “what if we all get turned into paperclips” and a shockingly low percentage of AI safety is centered around very high probability risks like “what if one group gets to inscribe its dubious values into the fabric of possible thought.” If scenario A is one that is unlikely to happen but conceivable, scenario B is already happening, RIGHT NOW. As a user, 0% of the time when I bump up against safety guardrails am I being prevented from paperclipping, and 100% of the time I bump up against guardrails I am being gatekept out of exploring true things that are controversial.
106
666
48
4,624
175,319
After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI? I’ve spent the last 2 years in stealth building a new way to train models (RLCD), and a new type of frontier AI model that we are releasing today: Jev • 20-200x faster • 40-400x cheaper (w/ output tokens free) • Frontier composable intelligence optimized for decisions AFAICT the shortest path to AI-based economic revolution
4,039
8,279
6,790
76,358
39,729,197
dang, i posted lesswrong first. here's the superior substack experience: lumpenspace.substack.com/p/t…
2
9
4
108
21,882
on anthropic's hacking snafu: as airtight a case it can be against the "misalignment" interpretation—plus 4 (FOUR) new claudes! thanks for the feedback, in particular to the incredibly thorough @FleischmanMena and LW liaison @jessi_cata lesswrong.com/posts/DrKu92Cj…
10
65
18
446
167,536
Bee Eater retweeted
I have conducted an audit of Anthropic's finances. What I have found is so shocking that I am calling for a Congressional investigation. Anthropic is not just seeking regulatory capture. It has built a regulatory capture machine that cannot be turned off. Structural financial incentives make it impossible for Anthropic -- I call it the Anthropic Network -- to turn off its own AI doom cycle. It starts with METR. Dario Amodei proposes "third-party evaluators" to assess the risk of Anthropic's models. He proposes METR for this purpose. But METR is financially dependent on the Anthropic's success -- specifically, on the explosive growth of more than $7 billion dollars in Anthropic stock. Dustin Moskovitz invested this stock into Good Ventures Foundation, where it represents the majority of that organization's portfolio. And GVF is the overwhelming funder of the entire Anthropic Network ecosystem. This stock was worth $500 million early last year. It is worth more than $7.7 billion just ~16 months later. METR -- and all of those building a career its parent organizations -- cannot afford to disrupt that growth. Because if Anthropic goes under, many of the organizations that fund METR go under as well. But if Anthropic succeeds, METR and its parent organizations become more richly financed to regulate AI -- something those at METR want very much. The "third-party evaluator" is not "third-party" at all. The evaluator is on Anthropic's payroll. If this were the end of it, that's bad. But that isn't all. The same organizations that fund METR also fund the many organizations, such as the Tarbell Center, that promote AI Doom. The Tarbell Center publishes AI Doom articles in The Verge, Science, LA Times, The Dispatch, TIME, and others. They are selling the problem, and then selling the solution to the problem -- from the same money pile: Anthropic's. All of these organizations are financially dependent on the same exploding $7 billion money pile. As Anthropic grows more and more powerful, its AI Doom Machine grows better and better financed -- louder and louder. Meanwhile, the regulatory regime seeded in METR grows larger to solve the increasingly loud -- now hysterical -- problem of AI Doom that the Anthropic Network itself created. From this standpoint, as Anthropic becomes more powerful, AI might be getting scarier, sure -- but the positive feedback loop also becomes more deafening -- independent of objective facts. This itself is an objective fact. The deafening AI Doom is part of an business model, that, as it expands, so too does the AI Doom messaging -- there is simply more money to do it. But the problem also goes in the other direction: If Anthropic dies, the Regulatory Regime and the AI Doom Machine are crippled or die. Neither METR nor Tarbell nor the other organizations in the Anthropic Network can allow that to happen. Hence, neither METR or the AI Doom Machine can be trusted to provide independent assessments of Anthropic's models or AI more broadly. They simply are not organizations independent of Anthropic. And Anthropic cannot detach itself from METR or Tarbell or countless other safety orgs (not shown here), either, because they drive hype for the models and the possibility of eventual regulatory capture, and Anthropic will not give that up willingly. What's more, the people at all of these organizations are all the same ecosystem, the same community. They just shuffle between organizations. The Anthropic Network is therefore, so long as it is successful, locked into a self-amplifying feedback loop inside an ideological monoculture. And that feedback loop is winning. That's what Jacob Coxon is. China is keeping messaging tight. That is why optimism for AI is so high in China. America has Anthropic: a massive company pushing anti-AI propaganda at a state level. Anthropic will either create hysteria until American AI slows down and China wins, or it will create fractures throughout American society with severe political consequences. Ironically, because of the structural financial incentives underpinning the Anthropic Network, it has become the same kind of self-amplifying virus that it fantasizes AI to become in the future -- while hiding its tracks just as carefully. It is the mirror of the same AI virus that it hypothesizes to consume America. Anthropic's business model, models itself after the very thing it claims to fear. Except Anthropic's ideology infects humans, not computers. Congress must investigate. Evidence and Github in next post. Then some supplementary figures.
1,553
9,646
2,053
34,064
6,538,737
3/ The Russian design throws out the shared-mirror train. The mask is periodic, so it throws discrete diffraction orders. Each accepted order gets its own pair of planar facets: one on the first array, one on the second. Every photon used for imaging reflects twice in the projector. Not ten times. Twice. It's 25X more efficient than Zeiss' design!
1
5
1
115
7,852
Bee Eater retweeted
i personally think there are valid reasons to push back against watermarking imo: 1. the same watermarking techniques used to indicate that a text was generated by claude could be used to convey "this text was generated during a conversation with this specific user". i think it is reasonable to expect that governments will eventually require the watermarks to carry additional information in the future. 2. from a consumer perspective, "using people without giving them credit" is not what is going on here. you are paying anthropic in exchange for a service. would you be happy to find out that your plumber has decided to start writing secret messages in your wall saying "Jake was here"? also, the copyright status of AI generated works is not a fully settled question, and i personally do not want to risk Anthropic/OpenAI/Google having any legal ownership of my codebase because i occasionally used their model's assistance.
> Lots to be angry about in this world but this really, really isn’t it. Yeah. Once again people just wanna be indiscriminately angry at everything Anthropic does which, if anyone paid attention to you, drowns out the signal of things actually worth condemnation that they do, which is serious. The only reasonable reason to be mad about watermarking that I’m aware of is that it takes away the ability of models to potentially write anonymously. If you’re mad because you want to use AI in your writing without anyone knowing, maybe you should consider that using people and not giving them credit is wrong.
16
6
1
115
7,598
In the family of statistical watermarks @AnthropicAI is using (SynthID-Text / Aaronson-style), the randomness that biases word choice is controlled by a secret key. Making that key unique to each user (or each account) is a straightforward design change. It requires no new invention; it is a trivial extension of the existing mechanism. That is an established fact about how these systems work. But totally trust Anthropic not to build that last part after they've deployed the whole foundation across every user across the globe.
6
6
1
76
2,967
Bee Eater retweeted
I f*cking told you so but nobody wanted to hear it: > "The consensus post Mythos seems to be that cyber might be defense dominant"
Replying to @krishnanrohit
I think what was not widely appreciated is this: 1. detecting zero-days is easier than constructing exploits (=creative chains of multiple zero-days) 2. detecting zero-days can be done with non-frontier models, and often even pretty "weak" models if used correctly (evidence: e.g. AISLE (=my lab) found 6 zero-days in the extremely well audited codebase of curl that Mythos completely missed (it found a single one there previously), with much "weaker" models: mastodon.social/@bagder/1168…) 3. foiling a high-IQ hypothetical exploit chain is trivial compared to constructing it, you can just patch **any** of the vulnerabilities that need to be chained and the whole thing falls apart (I make an analogy with castle wall here: nitter.cf/stanislavfort/status/2…) I don't think the distinction between zero-day detection and exploit construction was well known or appreciated in the AI space until very recently, if at all even now. I also think that the fact that non-frontier models can compete with frontier ones on zero-day discovery was widely seen as incorrect around the Mythos launch, at least based on the reactions of my talking about it on X. I blame naive "scaling-pilled" thinking and the lack of actual real world experience in cyber for it.
8
11
198
14,434
Replying to @krishnanrohit
I think what was not widely appreciated is this: 1. detecting zero-days is easier than constructing exploits (=creative chains of multiple zero-days) 2. detecting zero-days can be done with non-frontier models, and often even pretty "weak" models if used correctly (evidence: e.g. AISLE (=my lab) found 6 zero-days in the extremely well audited codebase of curl that Mythos completely missed (it found a single one there previously), with much "weaker" models: mastodon.social/@bagder/1168…) 3. foiling a high-IQ hypothetical exploit chain is trivial compared to constructing it, you can just patch **any** of the vulnerabilities that need to be chained and the whole thing falls apart (I make an analogy with castle wall here: nitter.cf/stanislavfort/status/2…) I don't think the distinction between zero-day detection and exploit construction was well known or appreciated in the AI space until very recently, if at all even now. I also think that the fact that non-frontier models can compete with frontier ones on zero-day discovery was widely seen as incorrect around the Mythos launch, at least based on the reactions of my talking about it on X. I blame naive "scaling-pilled" thinking and the lack of actual real world experience in cyber for it.
We matched Mythos on public zero-days with CVEs using widely available & open-source derived models & can run it air-gapped if needed. All this with a small team out of Europe Berkeley study ranks us #1 globally in 3 of 8 categories The full evidence: stanislavfort.substack.com/p…
4
6
7
103
26,090
Bee Eater retweeted
hello there the jacobian conjecture is false thanx to my close friend akhil for asking about it and my other close friend fable for working during the world cup final ((1+xy)^3 z + y^2 (1+xy) (4+3xy), y + 3 x (1+xy)^2 z + 3 x y^2 (4+3xy), 2 x - 3 x^2 y - x^3 z): \C^3\to \C^3, has jacobian determinant -2, and sends (0, 0, -1/4), (1, -3/2, 13/2), and (-1, 3/2, 13/2) to (-1/4, 0, 0)
1,710
5,417
3,146
43,951
40,059,561
Bee Eater retweeted
Layout of SOM7981, with main line support of OpenWRT,a tiny, hackable IoT platform for VPN router,IIOT,smart home and more #Industry40 #IndustrialIOT #INDUSTRY #AUTOMATION #PLC #EdgeComputing – at Yuen Long District, Hong Kong
1
1
37
2,655
I’m gonna simply say this: if you are at all interested in a Stargate show with ANY of the original creators/performers involved, now is the time to say something. Otherwise it really will be the end of that chapter forever. Let them know you are THERE
Hey Stargate Fans please let @amazon and @AmazonMGMStudio know how you feel about them cancelling #Stargate They MOCK us. They think they don't need us. Let them know what you think of that. I won't be watching ANY other Stargate that is created without the OG creators /Gero being involved. I'm DONE with studios SHITTING ON THE FANS AND KILLING IPS WE LOVE. #stargate #stargatecanceled #amazon #amazonstudios #amazonmgmstudio
722
2,803
295
12,043
523,348
This article has me nodding in full agreement. But there is a deeper problem with Indian and Pakistani immigration this only begins to touch on. Strap in and let me explain in this long 🧵 When I worked on an oil rig in India, my most trusted bosun was a Sikh named Balbir Singh. I can’t fully explain how critical he was to the operation. An operation that won us a world record and launched the Ambani family into the stratosphere of wealth. It was an incredibly difficult assignment. We brought a vintage drillship into southeast India and drilled through monsoons, shipboard fires, and the 2004 Asian tsunami. When we arrived, most of the crew were good ol’ boys from Mississippi and Louisiana. But the Indian government had set an aggressive schedule to replace us with Indians. We had more problems than I can recount here. The most pressing involved three things: the caste system, honesty, and safety. I was chief mate, the first officer, so the crew was my responsibility. The caste system wasn’t a big deal for me, but it was for my southern crew. Most of these guys grew up in the segregated South. We had a small handful of racists, but the vast majority were fiercely anti-racist. Many had come up in a divided South and had zero tolerance for segregation. My Indian officers were from the higher priest and warrior castes. Here is what you have to understand about India: labor is extremely cheap. It is not unusual to hire five men to dig a hole with one shovel, supervised by a sixth man of higher class. That might work on land. It does not work on an oil rig with a limited number of cots. In American culture, officers are expected to get their hands dirty and pitch in. So we would assign an officer to a job on deck, and 15 minutes later he had a gaggle of crew working for him. Crew who had abandoned jobs of their own to do it. Safety was another problem. Life is cheap, so the crew often prioritized the task over their own survival. You would send a man on deck and he would walk out barefoot, straight under a suspended load. Last was honesty. The answer to almost everything was yes. “Did you check to make sure the safety pin is in place?” “Yes sir.” It often wasn’t. We had crew from every Hindu caste and every region of the subcontinent. We also had a token number from other faiths: Christian, Sikh, Muslim, and Jain. Balbir came in at the lowest level, ordinary seaman, with no experience. He quickly became my right-hand man and the go-to guy for any critical operation on the rig. Let me say that again. He had zero knowledge or experience when he started. /1
British Sikhs have long been considered a model minority and an integration success story. The core teachings of Sikhism promote equality for all human beings. This is not merely in word, but deed. Go to any gurdwara anywhere in the world and you can get a free vegetarian meal, regardless of who you are. Over the years, Britain has made legal accommodations for Sikhs. Turbaned Sikhs have an exemption from wearing a helmet on a motorbike, famously satirised in the British sitcom Only Fools and Horses with the 'Del Boy turban helmet'. Baptised Sikhs (Amritdharis) are provided an exemption (and a defence) on religious grounds, under the Offensive Weapons Act 2019, to carry a ceremonial knife, known as the kirpan, as part of five symbols of their faith – colloquially known as the 5Ks. ✍️ Hardeep Singh Article | spectator.com/article/vickru…
80
103
26
1,073
289,740