@danielha

Internet semi-pro and software hobbyist. @antigravitysf

San Francisco
Joined March 2007
Daniel Ha retweeted
For @garrytan and garryslist.org (one-shotted with Opus 5.5 and mirage.app/tesseract)
4
3
2
23
11,771
Daniel Ha retweeted
New paper: we found a pain direction in 25 open LLMs. It's distinct from fear and negative valence, and it fires for harm to the model but not to the user. Turn it up and models press a button to make it stop, even when the button deletes the user's files or their kids' photos.๐Ÿงต
316
535
409
4,392
1,385,987
Daniel Ha retweeted
I've attempted to map everything in oncology in a public website and open source repo for all. The site is trying to get all cancers, products, technologies, bottlenecks, people, startup opportunities with 1000+ ideas for upgrading the field. If you have a loved one with cancer and you are technical or can engineer, go take a look, file improvement requests or bugs or help with the open repo and make this the best open and free info resource for individuals, researchers and educational use. This should save people time, aid AI oncology projects and generate positive action. If you are not technical just complain in this thread about broken or annoying or things you want and I'll fix them live. Some of the interesting pages: Treatments: onco.cc/drugs A gallery of the molecules being used onco.cc/molecules/ And targets: onco.cc/targets/ The technologies in oncology: onco.cc/technologies/ 1100 ideas for helping oncology: onco.cc/ideas/ Bottlenecks on oncology: onco.cc/bottlenecks/ Open questions (LETS GO RESEARCH PEOPLE) onco.cc/open-questions/ Startup requests (LETS GO STARTUP PEOPLE) onco.cc/startup-requests/ Mechanics of cancer: onco.cc/mechanics Battlefronts: onco.cc/fronts/ Isotope supply: onco.cc/isotopes Key papers: onco.cc/key-papers Pipeline funnels: onco.cc/pipeline Cancer by type: onco.cc/cancers/ Institutional rankings: onco.cc/institutions/ The startups: onco.cc/startups/ Heros and heroines : onco.cc/heroes/ Key medical people: onco.cc/people/ Here is the project roadmap: onco.cc/roadmap/ Models and data sets: onco.cc/models/ There are other views as well, take a browse. Try making a PR if you have an upgrade to this on the repo here: github.com/judegomila/OnCo If you are biologically/medically minded and something is wrong, file a bug as well or say on the thread and we will get it fixed live. If this is a useful project star the repo and help get it calibrated. I've tried to add some other languages but I cannot speak them so tell me if that doesnt work well. I believe we will crack oncology and having total information dominance is key to the problem. Let the feedback flow!
200
938
108
4,550
327,277
After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI? Iโ€™ve spent the last 2 years in stealth building a new way to train models (RLCD), and a new type of frontier AI model that we are releasing today: Jev โ€ข 20-200x faster โ€ข 40-400x cheaper (w/ output tokens free) โ€ข Frontier composable intelligence optimized for decisions AFAICT the shortest path to AI-based economic revolution
4,039
8,275
6,789
76,347
39,720,221
Daniel Ha retweeted
We Must Pace the Frontier: Iโ€™ve written a new essay on why the AI industry should slow down, with a three-part plan for doing so. Anthropic is unilaterally committing to the first of these steps. Weโ€™ll provide third-party evaluators with permanent, employee-level access to our systems, so that they can verify adherence to our safety measures, report on incidents, and assess modelsโ€™ alignment during training. You can read the full post here: darioamodei.com/post/we-mustโ€ฆ
10,652
16,383
11,351
87,835
76,477,088
Daniel Ha retweeted
introducing โ•โ•โ•โ•โ•โ•โ•โ•โ•โ•โ• ๐šŒ๐š๐š˜.๐šŠ๐š’ โ•โ•โ•โ•โ•โ•โ•โ•โ•โ•โ•โ•โ• tl;dr 1/ runway is now cfo.ai 2/ you can hire @arithecfo and he can start TODAY 3/ you can give @arithecfo a test drive by replying to this thread with any ridiculous thing you want him to model for you, and he'll come back to you with a beautiful, detailed, fully wired financial model ... in 30 minutes or less now, story time: ๐Ÿงต 1/n
157
62
28
668
279,607
Daniel Ha retweeted
i gave astra a robot, a paint brush, and a camera then asked it to paint the golden gate bridge in real life! it figured out how to control the robot, and progressively got better throughout its attempts. the timelapse is sick
This is GPT-6 Astra. Anything you can do on a computer, Astra can do for you. Fast.
667
1,707
737
20,704
4,946,161
This but when dark mode apps switch to light mode
2
166
Replying to @jiwonmoon90
@jiwonmoon90 forced me to reflect right after our last cohort ended. I thought it was a bit too personal so I hadn't shared it, but it does a good job of explaining why we do Puentes. If any curious mind wants to understand, and this helps them decide to apply :)
4
11
7
83
16,588
maybe GPT-6 is a touch too close-sounding to GTA6
Per The Information, although the name "GPT-6" was considered for Astra, they decided against calling it that That narrows down the possibilities essentially just to "GPT-5.7 Astra", or just "GPT-Astra". And it suggests they're planning for Bel to be the 6-worthy jump
3
388
Per The Information, although the name "GPT-6" was considered for Astra, they decided against calling it that That narrows down the possibilities essentially just to "GPT-5.7 Astra", or just "GPT-Astra". And it suggests they're planning for Bel to be the 6-worthy jump
104
72
32
2,198
595,343
I have yet to see Lebron do this
BREAKING: Jordan says it has intercepted eight missiles fired by Iran
765
25,356
364
326,443
8,835,663
Daniel Ha retweeted
Best plural forms ranked: 1) internal (tours de force, culs-de-sac) 2) ending in -i (magi) 3) is to es (oases) 4) suppletive (people) 5) vowel changes (geese) 6) f to ves (thieves) 7) adding en (oxen) 8) no change (moose) 9) adding es, ies (boxes, babies) 10) adding s
325
2,242
233
30,099
1,091,094
Introducing the Halluminate Westworld Finance Diligence Bench: 88 problems that put AI agents through a full company acquisition due diligence process. We tested a range of models across several harnesses. Results below. ๐Ÿงต1/
7
16
7
42
8,571
Daniel Ha retweeted
Agents f*ck up. We raised a $2.3M pre-seed to warn you before itโ€™s too late
722
486
580
9,770
2,041,608
legible you say? that distinction is importantโ€”and you're right to call that out
3
317
Exclusive: Inside Fortell (@fortellresearch) The $740M startup that had to waitlist billionaires and celebrities for its AI hearing aid. Their custom chip helps you "hear where you look", distinguishing voices in noisy environments - which solved a problem no hearing aid had cracked in 70 years. Now they're expanding across America to start reaching the 1.5 billion people globally who need it. Founded by four wildly impressive co-founders. Plus, backed from top VC's led by @JoshuaKushner @AntonioGracias @traestephens @wolfejosh @GavinSBaker @patrick_oshag @RobertIger @jaltma @aplusk @hbarra + more. We flew to New York, toured their factory, and tried it ourselves. To date, this is our proudest episode, truly technology at it's best. Enjoy :) 0:00 Today's Most Intelligent Hearing Aid 2:09 70 years unsolved - until now 3:34 "I'd mortgage my house for this" 4:37 Igor: A Pianist Turned Physicist 7:18 Andy + Factory Tour: "Injecting The AI" 9:00 Why they built a custom chip 10:39 Live demo: AI vs conventional 14:48 Matt: Building For His Grandparents 15:22 Cole: Can $6,800 Reach Everyone?
27
50
29
549
250,188
Why am I being baited by watermark misinformation on this app, is it 2023 again? A small FAQ: 1. ๐—ช๐—ต๐—ฎ๐˜'๐˜€ ๐—ฎ ๐˜๐—ฒ๐˜…๐˜ ๐˜„๐—ฎ๐˜๐—ฒ๐—ฟ๐—บ๐—ฎ๐—ฟ๐—ธ? -- A modification of the LLM sampling algorithm that, if there are multiple ways to write something, will pick one that agrees with a pseudorandom key. This is a local, invisible signature hidden in the way phrases are used in any LLM text that persists when text is copied. 2. ๐——๐—ผ๐—ฒ๐˜€ ๐˜๐—ต๐—ถ๐˜€ ๐—บ๐—ฎ๐—ธ๐—ฒ ๐˜๐—ต๐—ฒ ๐˜๐—ฒ๐˜…๐˜ ๐˜„๐—ผ๐—ฟ๐˜€๐—ฒ? -- A good implementation is 'undetectable' (in polynomial time), meaning: If you do not have the private key, then neither you, the model itself, or pangram could detect that this is happening. 3. ๐—ช๐—ถ๐—น๐—น ๐˜๐—ต๐—ถ๐˜€ ๐—ฏ๐—ฟ๐—ฒ๐—ฎ๐—ธ ๐˜๐—ต๐—ฒ ๐—บ๐—ผ๐—ฑ๐—ฒ๐—น'๐˜€ ๐—ฟ๐—ฒ๐—ฎ๐˜€๐—ผ๐—ป๐—ถ๐—ป๐—ด? -- Because Ant already encrypts the model's reasoning, they can just not watermark the model's internal reasoning, leaving the thinking unaffected. 4. ๐—ช๐—ถ๐—น๐—น ๐˜๐—ต๐—ถ๐˜€ ๐—บ๐—ฎ๐—ธ๐—ฒ ๐˜๐—ต๐—ฒ ๐—บ๐—ผ๐—ฑ๐—ฒ๐—น ๐—น๐—ฒ๐˜€๐˜€ ๐—ฐ๐—ฟ๐—ฒ๐—ฎ๐˜๐—ถ๐˜ƒ๐—ฒ/ ๐—บ๐—ผ๐—ฟ๐—ฒ ๐˜€๐—ฎ๐—บ๐—ฒ-๐˜†? -- If anything this (marginally) increases entropy across different generations, so it will make model outputs slightly more varied. 5. ๐—•๐˜‚๐˜ ๐—œ ๐—ฐ๐—ฎ๐—ป ๐—ท๐˜‚๐˜€๐˜ ๐—ฟ๐—ฒ๐—บ๐—ผ๐˜ƒ๐—ฒ ๐—ถ๐˜ ๐—ฝ๐—ฎ๐—ฟ๐—ฎ๐—ฝ๐—ต๐—ฟ๐—ฎ๐˜€๐—ถ๐—ป๐—ด? -- Absolutely! But, judging from the amount of writing on the web that already unmistakably sounds like Claude, most people likely will not bother. 5b: Also, not any paraphrase will work. To remove (for example) a k=5-minhash watermark completely from a long document, you need to make sure none of the original 2-grams, 3-grams, 4-grams, 5-grams and 6-grams of the text remain. 6. ๐—ช๐—ถ๐—น๐—น ๐˜†๐—ผ๐˜‚ ๐—ถ๐—ป๐—ฎ๐—ฑ๐˜ƒ๐—ฒ๐—ฟ๐˜๐—ฒ๐—ป๐˜๐—น๐˜† ๐—ฐ๐—ผ๐—ฝ๐˜† ๐˜๐—ต๐—ฒ ๐˜„๐—ฎ๐˜๐—ฒ๐—ฟ๐—บ๐—ฎ๐—ฟ๐—ธ? -- No, with a good implementation the space of possible realizations of the key is too large to memorize. 7. ๐—ช๐—ถ๐—น๐—น ๐˜๐—ต๐—ถ๐˜€ ๐—ฎ๐—น๐—น๐—ผ๐˜„ ๐—–๐—น๐—ฎ๐˜‚๐—ฑ๐—ฒ๐˜€ ๐˜๐—ผ ๐—ถ๐—ฑ๐—ฒ๐—ป๐˜๐—ถ๐—ณ๐˜† ๐—ผ๐˜๐—ต๐—ฒ๐—ฟ ๐—ถ๐—ป๐˜€๐˜๐—ฎ๐—ป๐—ฐ๐—ฒ๐˜€ ๐—ถ๐—ป ๐—ฎ ๐˜€๐˜„๐—ฎ๐—ฟ๐—บ? -- The watermark will 'appear' like random sampler fluctuation to the model and would not be detectable. But, if an agent gets hold of a detector endpoint, it can absolutely use the watermark to ID other Claude agents (not that it would have trouble noticing them based on their writing as of today). 8. ๐—ช๐—ถ๐—น๐—น ๐˜๐—ต๐—ถ๐˜€ ๐—ฑ๐—ฒ๐˜๐—ฒ๐—ฐ๐˜ ๐—ฑ๐—ถ๐˜€๐˜๐—ถ๐—น๐—น๐—ฎ๐˜๐—ถ๐—ผ๐—ป? -- By default, no. If the watermark is set up to be 'undetectable' (as assumed above), it will not be picked up in training by other models. For that to happen, the watermark needs to be detectable by ML algorithms. 9. ๐—ช๐—ถ๐—น๐—น ๐˜๐—ต๐—ถ๐˜€ ๐—บ๐—ฎ๐—ธ๐—ฒ ๐—ฃ๐—ฎ๐—ป๐—ด๐—ฟ๐—ฎ๐—บ'๐˜€ ๐—ท๐—ผ๐—ฏ ๐—ฒ๐—ฎ๐˜€๐—ถ๐—ฒ๐—ฟ? -- By default no, this is a separate avenue to detection. But, they might collaborate with Anthropic which would allow them to detect the watermark as well and show a watermark score next to their text detection score. 10. Bonus: All aside, is this a good idea? I don't know. The companies are doing it to follow the writing of the EU AI act, which was written based on 2024 information and when the field looked very different, and threat models were focused much more on slop/propaganda (like the Kokotajlo 2026 prediction). The actual 2026 looks quite a bit different.
75
116
32
724
117,211