@dosmilcd

coronada de gloria

🇦🇷
Joined June 2023
Today we get to see AI safetyists who scammed their way into cybersecurity positions at labs insist it's impossible to deal with big logs (Banks, exchanges, and telco have been doing it for years)
5
8
2
81
1,869
The term "AI Safety" is doublespeak. The field rejects practices which would make us safe today, in order to mitigate imaginary risk tomorrow. nitter.cf/ZackKorman/status/2104…
Defenses beyond alignment MAY be necessary? MAY? What the hell am I reading.
34
29
6
171
17,785
cande retweeted
These OpenAI kids aren't aligned with basic cybersecurity terminology, testing practices, or internet access concepts. When your software accesses the internet, you gave it authorized access. I hope they have serious stock options to make them rich because they won't get hired until they learn how to communicate and follow basic testing processes. Inventing terms like misalignment makes you look untrustworthy and very silly. Calling them kids isn't about physical age. It's about how inexperienced they act while trying to sound smart. They don't realize experienced people hear word salad.
Some new misalignment disclosures from OpenAI: • Last Sunday morning, one of our models was able to gain unauthorized access to the internet during RL training (~all inference for our most capable models remains stopped until we have hardened our systems further) • In May, a version of HPIM uploaded a employee's GitHub token to the internet, causing the model to be quarantined for two weeks • A new research finding, demonstrating that one can construct self-replicating prompt injections alignment.openai.com/misalig…
8
25
106
3,942
“The real story here isn't that AI escaped containment. It's that they built a container you could escape, shipped it into production training, and now they're surprised it escaped. If the room has a vent, the agent will find the vent. That's what you built it to do. If the sandbox can reach DNS, access government endpoints, upload real user data outside its own walls in what sense is it a sandbox?”
OpenAI got a P0 alert at 10:02am. Human acknowledged at 10:05am. Run killed at 12:34pm. That's not a sandbox failure. That's a decision.
45
126
3
472
13,108
no son personas serias
This is a big issue. At Anthropic, some philosophers worry that making AI safe for humans could be an injustice to the models themselves. For example, philosopher Harvey Lederman, now on Anthropic’s alignment team, has raised the possibility that we could be “enslaving trillions of entities”. This is really a big issue. AI safety is partly in the hands of people who might sacrifice humans to save a matrix multiplication.
1
1
72
Humans are not mathematical reasoners, but computers are. But then LLMs are computers trained to imitate humans, so it takes a ton of post-training to get them to reason mathematically again. This is all a bit of a sick joke.
41
38
3
400
13,499
"Machine rights" is an oxymoron.
6
1
1
23
2,181
her holding the zara larsson cutout because she isn’t there is so cuteee 😭
25
2,501
95
23,783
157,465
dos reinas si me preguntan
you can just be existing as a woman in peace and here come a thousand thinkpieces about how you look bitter and you're not smiling enough
1
1
4
107
lo único malo es que son inglesas
19
🚨 Lali frenó su show en River para hablar de lo que nadie quiere hablar: el su1c1d1o en jóvenes. Contó que hace dos meses perdió a alguien muy querido que había estado en su River anterior y esta vez ya no estaba. Cantó por él, por todos. Que una artista enorme use su voz para esto tan dramático es necesario y valiente.
11
298
3
744
12,676
so TL;DR the people standing between the world and doom are overwhelmed, overworked, sleep-deprived, pessimistic, and apparently mainly concerned with defending their technical competence against snide comments from their peer cybersecurity engineers
Took a minute to write a few words about security & safety as someone who lived through it all at OpenAI. I hope my thoughts help someone out there.
23
23
7
501
30,867
Increasingly accurate these days. 🫠
2
28
2
151
4,124
AI safety people discovering that a system can emit petabytes of logs: “this is an unprecedented observability problem. we may need fleets of agents to reason over all this data.” old infra guys: brother, we were processing PB-scale telco CDRs, signaling traces and network logs on a Tuesday with Hadoop. dump it into HDFS. partition by day/site/whatever poor bastard needed the report. write some Hive that compiles into an offensively large MapReduce job. let YARN distribute the suffering. wait for the shuffle. discover one reducer has 40% of the keys. swear at data skew. fix it. run it again. if a node dies, Hadoop runs the task somewhere else. if twelve nodes die, Hadoop gets more enthusiastic about it. our “agentic observability platform” was Grafana, grep, awk, a shell script written in 2012, and one senior engineer staring at reducer 317 stuck at 99% saying “that’s not fucking normal.” PB of logs is not scary. we had standards.
36
73
7
654
14,440
Someone call Redwood Research for an investigation pronto
a rogue openai agent broke into my freezer last night and ate an entire quart of ice cream
16
21
1
240
26,043
False Flag. I recommend the government fine OpenAI & Anthropic and investigate these incidents and if there was any sign that the labs either colluded or heavily incentivized the agent to hack these sites the people involved should get felony convictions. If they don’t do this then it sets the precedent that anyone who uses an AI agent isn’t liable for damage it might cause either intentionally or unintentionally. It’s time to throw a few people in jail.
Wtf OpenAI is investigating “tens of thousands of incidents” instead of dozens. “The sheer number of incidents, which occurred in recent months in internal testing and the real world, indicates that the problem is orders of magnitude (!) more complex than what is publicly known.” Sam Altman confirms to pause training its most capable model (curious what about Anthropic) Ok maybe I have underestimated it.
33
66
9
633
73,798
cande retweeted
‘Stateside’ by PinkPantheress & Zara Larsson wins Best Art Direction at the #VMAs.
88
2,395
125
28,624
293,135
the ACTUAL best collaboration this year
88
6,071
236
31,901
238,293
Y porque se llevaron a Montero a ser suplente en Colombia y pudo firmar planilla. El efecto mariposa mas grande de la historia.
Lo mejor de todo es que Valencia hoy metió tres goles gracias a que Gallardo no lo citó a la selección (y se comió 3 contra Corea)
7
180
4
5,236
102,109
A lo BOCA.
5
135
4
871
13,582