Lahti
Joined April 2009
Antti Vermas retweeted
Replying to @SariMultala
Miksi pitää aina vaan "koettaa aktiivisesti vaikuttaa"? Erotettaisiinko Suomi unionista, jos vain ilmoittaisimme, että "Ei kiitos, pidämme vanhan, ehdotettua paremman järjestelmän voimassa?"
29
49
735
6,721
Antti Vermas retweeted
Blocked or not, this is a gem
61
507
38
9,784
162,847
Antti Vermas retweeted
Amazing. These researchers found that there's a spike in Bitcoin activity around the time foreign aid money goes out. Somewhere between 2 and 6 cents of every dollar of World Bank foreign aid disbursements got siphoned off into crypto wallets:
200
1,778
209
13,704
448,120
Good prompting hack for agentic loop to do some verification as you ordered (it constantly lies and just says so) is to phrase it like "Give me the evidence of xyz...."
69
Veikkaus: AI-doomerismi kasvaa ilmastonmuutoksen kaltaiseksi massiiviseksi poliittiseksi liikkeeksi, joka vetää jopa Greta Thunberg mukaansa lähiaikoina. Muutaman vuoden sisällä alkaa tuleen tietoturvaongelmia mm. homehtuneiden npm-kirjastojen hyväksikäytön muodossa ja pahantekoa kaikenlaisissa muissakin asioissa, missä ihminen käyttää tekoälyä työkaluna. Myös agenttisia firmoja alkaa nouseen isossa määrin siellä täällä. Samalla tulee kaikenlaista hutilointia kuinka se ja se firma menetti niin ja niin paljon rahaa, kunnes tulee yksi monumentaalinen Enron-kokoluokan sekoilu, missä palaa kasa miljardeja ja vaikka kasa ihmishenkiä. Tässä koko ajan rinnalla doomerismi ottaa kierroksia, kun voidaan osoittaa, että “mitä minä sanoin, sieltä se nyt tulee!”. Viimeistään tässä vaiheessa AI-regulaatio alkaa oleen tiukkaa, mutta ennen kaikkea aletaan toden teolla laittaan manuaalisia kill-switchejä vähän joka paikkaan. Samalla koko ajan rakentuu myös konservatiivinen suuntaus, jossa kaikki inferenssi ja datan säilöntä ja kopelointi tapahtuu lokaalisti korporaation kellarissa, niin kuin ennen vanhaan konsanaan. Se keskeisin kysymys on, että päädytäänkö me tilanteeseen, jossa on kourallinen “too big to fail” hallinnon näpeissä olevaa frontier-lafkaa kuten Amodeit ja Altmanit yrittää, jotka efektiivisesti ohjaa mitä kaikkea vastuullista tekoäly ihmisille kertoo ja tekee, tai vaihtoehtoisesti rikas ekosysteemi monenmoisia pelureita isoista pieneen eri spesialiteeteillaan ja vaikka täysin avoin lohkoketjun päällä oleva jonkinlainen RAGitys-verkko josta kaikki tieto tulee versioituna ja kielimalleja käytetään lähinnä sen parsimiseen. EU:n AI act ymmärtääkseni vaatii avoimuutta aika paljon ja mikä on sen avoimempi kuin lohkoketju. Tosin se vaatii kaikkea muutakin, mutta se on asia erikseen.
2
1
1
6
515
Myöskin veikkaan, että doomerismista tulee uusi kulttuurisodan areena. Huolimatta siitä mitä mieltä kukin mistäkin aspektista on, niin he gravitoituvat omien heimojensa mielipiteisiin. Vasemmisto vaatii keihäänkärkenään valtion väliintuloa, koska muuten ihmiskunta tuhoutuu (kuulostaako tutulta?) ja oikeisto vastavuoroisesti haluaa säilyttää open source mallit vapaina, koska muuten valutaan totalitarismiin.
1
2
54
Tämän laiminlyönti on kandidaatti Enron-seinäänajon syyksi. "Deliberate inefficiency is not waste. In safety-critical engineering we have always known it as insurance, and we buy it on purpose. As AI takes over the work where expertise is forged, the smart move is not to resist the automation. It is to keep our hands on the controls by design—so that when the automation fails, as it always eventually does, there is still someone in the chair who knows how to fly." spectrum.ieee.org/ai-enginee…
11
This is a big issue. At Anthropic, some philosophers worry that making AI safe for humans could be an injustice to the models themselves. For example, philosopher Harvey Lederman, now on Anthropic’s alignment team, has raised the possibility that we could be “enslaving trillions of entities”. This is really a big issue. AI safety is partly in the hands of people who might sacrifice humans to save a matrix multiplication.
584
279
303
1,629
360,496
Hyvät naiset ja herrat, Dario Amodei!
Dario Amodei is at the desk to assure that the future of humanity is safe from AI
1
411
Kun aloitin opettajana, luokissa oli vielä opettajakorokkeet, jotta kaikki näkisivät taulun ja kenties opettajankin. Ne poistettiin pian, kun joku kasvatustieteilijä sai päähänsä väittää niiden olevan epädemokraattisia. Nyt taululle ei sitten näe.
63
64
8
1,495
48,044
Antti Vermas retweeted
The loudest voices stoking fears about AI dangers have made tremendous headway in the past two weeks. AI technology has not taken some unexpected, dangerous turn, but the hype around it — propelled by what appears to be a well orchestrated PR campaign — has drummed up considerable fear. I worry that it represents a setback for our field. I have written frequently that fears of AI are overhyped. AI’s capabilities can be uncannily human-like and unpredictable, and it’s rational to worry when people who are directly involved express concerns. But I see the problems as a sign of the engineering work that ahead, rather than insurmountable barriers or the sky falling. AI technology continues to advance — which is a good thing! — but technical advances, poorly understood by the public, give those who seek to generate hype repeated opportunities to do so. First, I don’t see any step up in the risk of human extinction from AI compared to a few months ago. The theories about this remain the same fantastical, science fiction scenarios as a few months ago. The biggest change in AI risk is its cybersecurity capabilities — a topic which we should take seriously — but this, too, will not lead to the end of the world. The most notable recent event leading to increased fear was when an OpenAI team deployed an agent swarm that hacked into Hugging Face. Much of the popular press contained significant hype. For example, some publications reported that a swarm of 1,200 agents carried out the attack. While this was technically accurate, as I write this, I have about 1,300 processes running on my laptop. Yes, the ability to get large swarms of agents to work in parallel on a task is a significant technical advance, And, in computing, many processes run at the same time. So this shouldn’t be seen as some magical capability. Additionally, OpenAI’s buggy sandboxing and monitoring processes were key to enabling this incident. Fixing these bugs and putting in place improved monitoring would be appropriate fixes, not pausing AI. There are many well known ways to attack software systems. The main advantage of AI agents is that they are relentless. They will tirelessly try many tactics — and have the patience to chain vulnerabilities together — that previously would have taken an infeasible amount of human effort. But in the long term, I believe the advantage will lie with defenders (because they have more information with which to identify bugs, which they can fix), but the cyber-threat landscape has changed significantly. There are still bottlenecks to identifying and exploiting a vulnerability. AI agents still have to try a lot of things to see what works, and taking these actions takes time and might be detected by defenders. This is why, even though it is now easy to obtain versions of leading open weight models that have had their guardrails removed or weakened, so they will not refuse to try to execute cyber attacks, the world has not ended. I am also concerned about the anthropomorphization of AI in a lot of reporting, where LLMs and agents are unnecessarily treated as if they were people. If I wield a hammer, miss a nail, and accidentally dent the wall, it’s not the fault of the hammer. The problem lies in how I used the hammer. Similarly, if I prompt an agent and it hacks into someone else’s system, the responsibility lies with me, not the agent. Of course, we want to build systems that are as safe and predictable as possible. (For example, an unsafe hammer would be one whose head randomly flies off under normal use.) Today’s agentic systems are not predictable, but I see no reason why, by applying sound engineering practices, we won’t be able to make them extremely safe to use. One new element in the forecasts of AI-enabled doom is AI companies disclaiming responsibility for their own products. “I didn’t do it; my out-of-control agent did!” There’s a balance to be struck between the responsibility of the tool maker and the tool user, but when something goes wrong, let’s hold the people building and/or using the hammer responsible, rather than the hammer. (By the way, if you’re worried about AI bioweapon risk, David Bellamy has a great post on why this, too, is overhyped. Briefly, the bottleneck in building a bioweapon is not intelligence, but lab work and manufacturing.) Pausing AI progress will create much more harm than benefit. First, our adversaries will certainly not slow down. Second, engineering requires discovering problems empirically so we can fix them. If we pause AI by a decade, we will also delay finding and implementing safety engineering fixes by about the same duration. Of course, the incentive to stoke fears — for regulatory capture, to garner attention, or to make one’s technology seem more powerful — remains the same as before. Disclaiming responsibility is a new one. Taking a hard technical look at the actual risks however, I see little factual basis for the degree of fear that’s been stoked up. We still have hard research and engineering work ahead to improve AI safety, but the beneficial applications continue to vastly outweigh the risks, and we should keep building. [Original text (with links): deeplearning.ai/the-batch/is… ]
951
1,864
480
8,833
7,977,982
Antti Vermas retweeted
🚨BREAKING: Jensen Huang just EXPOSED Dario and Altman’s "Rogue AI" grift to avoid getting sued into oblivion under EXISTING law “Don't let this doomsday narrative cause somebody to relieve them of the laws that currently exist. Go and read between the lines. They're actually not asking for more laws; they're asking to be relieved of the laws we do have.” ABSOLUTE TRUTH NUKE
308
2,915
357
15,439
1,152,597
Antti Vermas retweeted
Victory By Any Means
747
2,581
498
25,195
10,651,802
Yksinkertainen ratkaisuehdoitus.
The best way to pace the frontier is to hold the labs fully liable for the behavior of their models.
2
228
Antti Vermas retweeted
The best way to pace the frontier is to hold the labs fully liable for the behavior of their models.
932
1,950
431
21,382
1,286,546
"None of this, in itself, proves that he is wrong about AI safety. What it does indicate is that he was already sold on the idea as a teenager, before he worked in the labs, before GPT-1 came out."
What do we know about Jacob Coxon, the Anthropic “whistleblower”? He became concerned about AGI doom in high school (in the town of Oxford btw - see my previous posts) after reading Nick Bostrom’s book Superintelligence. Attended Cambridge, during which time he went to an alignment conference at age 19. He then became a fellow at Newspeak (yes, really) House, which received EA funding and hosted Jeremy Corbyn’s Digital Democracy Manifesto launch. In 2022 he received a scholarship from a Moskovitz-funded org. Moskovitz is one of the leading EA donors, and an early Anthropic investor. The 80,000 Hours project (also funded indirectly by Moskovitz) explicitly encourages EAs to infiltrate frontier labs to influence them from the inside. None of this, in itself, proves that he is wrong about AI safety. What it does indicate is that he was already sold on the idea as a teenager, before he worked in the labs, before GPT-1 came out. I think good epistemic practice means we should therefore take him no more seriously than any other doomer. His experience in the labs didn’t convince him AI is dangerous. He already believed that even before the current AI paradigm became dominant, and probably chose his career based on this belief. If you’re convinced by Bostrom’s arguments yourself, that’s fine. But Coxon adds precisely zero additional weight to them.
2
203
Antti Vermas retweeted
Jensen Huang: “Safety is paramount. In a lot of ways, it’s job one. However, safety is an engineering problem... If we’re not confident about the safety of the products — like all companies, like you and I, all the companies here — if you build a product or a service and you’re not confident in its functionality, capability, or safety, then don’t release it. That’s a very obvious thing to do. You pace yourself until you are confident you’re releasing something that the market would appreciate. The market forces are already there. We don’t need any new laws. We don’t need new regulations.”
396
1,124
179
7,580
633,236
"Meta delayed shipping Muse for several months to focus on safety and security. We didn't call for everyone else to do this before we would. We just did it as part of our day-to-day work because it was clearly the right thing for people and for us. I'm proud of the security foundations we've built." "- Engaging independent evaluators and advisors is industry best practice. MSL already does this today in several areas because it helps produce better work. Other labs can just do this too. In general, it would be helpful for there to be a larger and more diverse ecosystem of evaluators."
Last month I wrote about how we can build a positive and safe future for everyone: meta.com/thefutureisforevery… Every lab has the responsibility and incentive to move at the pace required to train its models safely, and the ability to take its own actions to ensure that happens. The reality is: - People won't want to use agents that are misaligned with them and that don't do what they ask, so labs have a strong natural incentive to make their models more aligned. There is a lot of debate about slowing progress on capabilities until alignment catches up. My view is that trust and alignment are quickly becoming the most important capabilities that will differentiate agents and models. Any lab that doesn't focus on alignment will fall behind. - Labs face significant liability if their models cause harm, so they have a strong incentive to prevent this as well. Meta delayed shipping Muse for several months to focus on safety and security. We didn't call for everyone else to do this before we would. We just did it as part of our day-to-day work because it was clearly the right thing for people and for us. I'm proud of the security foundations we've built. - Engaging independent evaluators and advisors is industry best practice. MSL already does this today in several areas because it helps produce better work. Other labs can just do this too. In general, it would be helpful for there to be a larger and more diverse ecosystem of evaluators. - Committing the significant majority of compute towards serving people rather than racing towards recursive self-improvement is one of the best ways to ensure we develop this technology safely. Meta has made this commitment and other labs can do this as well. I believe the key to building a positive future for everyone is maintaining the right balance of power. This is within our power to do.
148
Self-hatred is autoimmune disease. Expecting high standard from yourself is a virtue but beating the crap out of yourself is not.
2
116