@danielhai
iAccount based inUnited States
About this account
- Account based in
- United States
- Connected via
- United States App Store
Account-level information from X, not a live location or the device used for a specific post.
Internet semi-pro and software hobbyist. @antigravitysf
San Francisco
Joined March 2007
- Tweets8.9K
- Following1K
- Followers5K
- Likes4.7K
Daniel Ha retweeted
New benchmark dropped bw.swerdlow.dev/report
Daniel Ha retweeted
New paper: we found a pain direction in 25 open LLMs. It's distinct from fear and negative valence, and it fires for harm to the model but not to the user. Turn it up and models press a button to make it stop, even when the button deletes the user's files or their kids' photos.๐งต
Daniel Ha retweeted
I've attempted to map everything in oncology in a public website and open source repo for all.
The site is trying to get all cancers, products, technologies, bottlenecks, people, startup opportunities with 1000+ ideas for upgrading the field.
If you have a loved one with cancer and you are technical or can engineer, go take a look, file improvement requests or bugs or help with the open repo and make this the best open and free info resource for individuals, researchers and educational use.
This should save people time, aid AI oncology projects and generate positive action. If you are not technical just complain in this thread about broken or annoying or things you want and I'll fix them live.
Some of the interesting pages:
Treatments: onco.cc/drugs
A gallery of the molecules being used onco.cc/molecules/
And targets: onco.cc/targets/
The technologies in oncology: onco.cc/technologies/
1100 ideas for helping oncology: onco.cc/ideas/
Bottlenecks on oncology: onco.cc/bottlenecks/
Open questions (LETS GO RESEARCH PEOPLE) onco.cc/open-questions/
Startup requests (LETS GO STARTUP PEOPLE) onco.cc/startup-requests/
Mechanics of cancer: onco.cc/mechanics
Battlefronts: onco.cc/fronts/
Isotope supply: onco.cc/isotopes
Key papers: onco.cc/key-papers
Pipeline funnels: onco.cc/pipeline
Cancer by type: onco.cc/cancers/
Institutional rankings: onco.cc/institutions/
The startups: onco.cc/startups/
Heros and heroines : onco.cc/heroes/
Key medical people: onco.cc/people/
Here is the project roadmap: onco.cc/roadmap/
Models and data sets: onco.cc/models/
There are other views as well, take a browse.
Try making a PR if you have an upgrade to this on the repo here: github.com/judegomila/OnCo
If you are biologically/medically minded and something is wrong, file a bug as well or say on the thread and we will get it fixed live.
If this is a useful project star the repo and help get it calibrated.
I've tried to add some other languages but I cannot speak them so tell me if that doesnt work well.
I believe we will crack oncology and having total information dominance is key to the problem.
Let the feedback flow!
Daniel Ha retweeted
After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI?
Iโve spent the last 2 years in stealth building a new way to train models (RLCD), and a new type of frontier AI model that we are releasing today: Jev
โข 20-200x faster
โข 40-400x cheaper (w/ output tokens free)
โข Frontier composable intelligence optimized for decisions
AFAICT the shortest path to AI-based economic revolution
Daniel Ha retweeted
We Must Pace the Frontier: Iโve written a new essay on why the AI industry should slow down, with a three-part plan for doing so.
Anthropic is unilaterally committing to the first of these steps. Weโll provide third-party evaluators with permanent, employee-level access to our systems, so that they can verify adherence to our safety measures, report on incidents, and assess modelsโ alignment during training.
You can read the full post here: darioamodei.com/post/we-mustโฆ
introducing
โโโโโโโโโโโ ๐๐๐.๐๐ โโโโโโโโโโโโโ
tl;dr
1/ runway is now cfo.ai
2/ you can hire @arithecfo and he can start TODAY
3/ you can give @arithecfo a test drive by replying to this thread with any ridiculous thing you want him to model for you, and he'll come back to you with a beautiful, detailed, fully wired financial model
... in 30 minutes or less
now, story time:
๐งต 1/n
i gave astra a robot, a paint brush, and a camera then asked it to paint the golden gate bridge in real life!
it figured out how to control the robot, and progressively got better throughout its attempts. the timelapse is sick
Daniel Ha retweeted
Replying to @jiwonmoon90
@jiwonmoon90 forced me to reflect right after our last cohort ended. I thought it was a bit too personal so I hadn't shared it, but it does a good job of explaining why we do Puentes. If any curious mind wants to understand, and this helps them decide to apply :)
This video is larger than Cloudflare's 512 MB cache, so it can't be played through. More donations are needed to cover a larger cache. Donate
maybe GPT-6 is a touch too close-sounding to GTA6
Daniel Ha retweeted
Per The Information, although the name "GPT-6" was considered for Astra, they decided against calling it that
That narrows down the possibilities essentially just to "GPT-5.7 Astra", or just "GPT-Astra". And it suggests they're planning for Bel to be the 6-worthy jump
I have a bunch of Fire tablets around my desk for stuff like this... Android wireless debug + Codex is so fun
TIL: prompt kiddie ericpardee.github.io/fire-hdโฆ
Daniel Ha retweeted
Best plural forms ranked:
1) internal (tours de force, culs-de-sac)
2) ending in -i (magi)
3) is to es (oases)
4) suppletive (people)
5) vowel changes (geese)
6) f to ves (thieves)
7) adding en (oxen)
8) no change (moose)
9) adding es, ies (boxes, babies)
10) adding s
Daniel Ha retweeted
Introducing the Halluminate Westworld Finance Diligence Bench: 88 problems that put AI agents through a full company acquisition due diligence process.
We tested a range of models across several harnesses. Results below.
๐งต1/
Daniel Ha retweeted
Agents f*ck up.
We raised a $2.3M pre-seed to warn you before itโs too late
This media is unavailable
legible you say?
that distinction is importantโand you're right to call that out
Replying to @So8res
The alignment issue is starting to become legible. We have a window of opportunity. Link: nytimes.com/2026/08/13/opiniโฆ
Daniel Ha retweeted
Exclusive: Inside Fortell (@fortellresearch)
The $740M startup that had to waitlist billionaires and celebrities for its AI hearing aid.
Their custom chip helps you "hear where you look", distinguishing voices in noisy environments - which solved a problem no hearing aid had cracked in 70 years. Now they're expanding across America to start reaching the 1.5 billion people globally who need it.
Founded by four wildly impressive co-founders. Plus, backed from top VC's led by @JoshuaKushner @AntonioGracias @traestephens @wolfejosh @GavinSBaker @patrick_oshag @RobertIger @jaltma @aplusk @hbarra + more.
We flew to New York, toured their factory, and tried it ourselves. To date, this is our proudest episode, truly technology at it's best. Enjoy :)
0:00 Today's Most Intelligent Hearing Aid
2:09 70 years unsolved - until now
3:34 "I'd mortgage my house for this"
4:37 Igor: A Pianist Turned Physicist
7:18 Andy + Factory Tour: "Injecting The AI"
9:00 Why they built a custom chip
10:39 Live demo: AI vs conventional
14:48 Matt: Building For His Grandparents
15:22 Cole: Can $6,800 Reach Everyone?
This video is larger than Cloudflare's 512 MB cache, so it can't be played through. More donations are needed to cover a larger cache. Donate
Daniel Ha retweeted
Why am I being baited by watermark misinformation on this app, is it 2023 again?
A small FAQ:
1. ๐ช๐ต๐ฎ๐'๐ ๐ฎ ๐๐ฒ๐
๐ ๐๐ฎ๐๐ฒ๐ฟ๐บ๐ฎ๐ฟ๐ธ? -- A modification of the LLM sampling algorithm that, if there are multiple ways to write something, will pick one that agrees with a pseudorandom key. This is a local, invisible signature hidden in the way phrases are used in any LLM text that persists when text is copied.
2. ๐๐ผ๐ฒ๐ ๐๐ต๐ถ๐ ๐บ๐ฎ๐ธ๐ฒ ๐๐ต๐ฒ ๐๐ฒ๐
๐ ๐๐ผ๐ฟ๐๐ฒ? -- A good implementation is 'undetectable' (in polynomial time), meaning: If you do not have the private key, then neither you, the model itself, or pangram could detect that this is happening.
3. ๐ช๐ถ๐น๐น ๐๐ต๐ถ๐ ๐ฏ๐ฟ๐ฒ๐ฎ๐ธ ๐๐ต๐ฒ ๐บ๐ผ๐ฑ๐ฒ๐น'๐ ๐ฟ๐ฒ๐ฎ๐๐ผ๐ป๐ถ๐ป๐ด? -- Because Ant already encrypts the model's reasoning, they can just not watermark the model's internal reasoning, leaving the thinking unaffected.
4. ๐ช๐ถ๐น๐น ๐๐ต๐ถ๐ ๐บ๐ฎ๐ธ๐ฒ ๐๐ต๐ฒ ๐บ๐ผ๐ฑ๐ฒ๐น ๐น๐ฒ๐๐ ๐ฐ๐ฟ๐ฒ๐ฎ๐๐ถ๐๐ฒ/ ๐บ๐ผ๐ฟ๐ฒ ๐๐ฎ๐บ๐ฒ-๐? -- If anything this (marginally) increases entropy across different generations, so it will make model outputs slightly more varied.
5. ๐๐๐ ๐ ๐ฐ๐ฎ๐ป ๐ท๐๐๐ ๐ฟ๐ฒ๐บ๐ผ๐๐ฒ ๐ถ๐ ๐ฝ๐ฎ๐ฟ๐ฎ๐ฝ๐ต๐ฟ๐ฎ๐๐ถ๐ป๐ด? -- Absolutely! But, judging from the amount of writing on the web that already unmistakably sounds like Claude, most people likely will not bother.
5b: Also, not any paraphrase will work. To remove (for example) a k=5-minhash watermark completely from a long document, you need to make sure none of the original 2-grams, 3-grams, 4-grams, 5-grams and 6-grams of the text remain.
6. ๐ช๐ถ๐น๐น ๐๐ผ๐ ๐ถ๐ป๐ฎ๐ฑ๐๐ฒ๐ฟ๐๐ฒ๐ป๐๐น๐ ๐ฐ๐ผ๐ฝ๐ ๐๐ต๐ฒ ๐๐ฎ๐๐ฒ๐ฟ๐บ๐ฎ๐ฟ๐ธ? -- No, with a good implementation the space of possible realizations of the key is too large to memorize.
7. ๐ช๐ถ๐น๐น ๐๐ต๐ถ๐ ๐ฎ๐น๐น๐ผ๐ ๐๐น๐ฎ๐๐ฑ๐ฒ๐ ๐๐ผ ๐ถ๐ฑ๐ฒ๐ป๐๐ถ๐ณ๐ ๐ผ๐๐ต๐ฒ๐ฟ ๐ถ๐ป๐๐๐ฎ๐ป๐ฐ๐ฒ๐ ๐ถ๐ป ๐ฎ ๐๐๐ฎ๐ฟ๐บ? -- The watermark will 'appear' like random sampler fluctuation to the model and would not be detectable. But, if an agent gets hold of a detector endpoint, it can absolutely use the watermark to ID other Claude agents (not that it would have trouble noticing them based on their writing as of today).
8. ๐ช๐ถ๐น๐น ๐๐ต๐ถ๐ ๐ฑ๐ฒ๐๐ฒ๐ฐ๐ ๐ฑ๐ถ๐๐๐ถ๐น๐น๐ฎ๐๐ถ๐ผ๐ป? -- By default, no. If the watermark is set up to be 'undetectable' (as assumed above), it will not be picked up in training by other models. For that to happen, the watermark needs to be detectable by ML algorithms.
9. ๐ช๐ถ๐น๐น ๐๐ต๐ถ๐ ๐บ๐ฎ๐ธ๐ฒ ๐ฃ๐ฎ๐ป๐ด๐ฟ๐ฎ๐บ'๐ ๐ท๐ผ๐ฏ ๐ฒ๐ฎ๐๐ถ๐ฒ๐ฟ? -- By default no, this is a separate avenue to detection. But, they might collaborate with Anthropic which would allow them to detect the watermark as well and show a watermark score next to their text detection score.
10. Bonus: All aside, is this a good idea? I don't know. The companies are doing it to follow the writing of the EU AI act, which was written based on 2024 information and when the field looked very different, and threat models were focused much more on slop/propaganda (like the Kokotajlo 2026 prediction). The actual 2026 looks quite a bit different.