@calebmer

Wayfinding towards something new. Previously frontend infra @meta, early @airtable, and OSS contributor

NYC
Joined November 2013
I’ve replaced emdashes in my writing…with ellipsis Used to be a big emdash user but whereas before emdashes were a sign of careful narrative construction…now they’re a sign you offloaded your thinking …but an ellispsis gives a similar pause (and you can use it in new ways) Like at the end of a sentence to invite the reader to continue on… …or at the beginning of a sentence to pick back up where you left off And you can still…albeit in a slightly more awkward way…use the ellipsis for asides within a sentence Can’t wait for LLMs to adopt this quirk and I lose another favorite punctuation mark
1
1
380
Whoa. This seems well poised to help with two big problems: 1. Fraud detection 2. Model alignment confirmation Instead of METR using ChatGPT to understand if ChatGPT was misaligned during the HuggingFace hack now you have a cheap *independent* (???) model with confidence scores you can run against every agent swarm trace You could run this thing as a cheaper classifier for dangerous behaviors like bioweapon development too For fraud, imagine sending millions of logs and it can tell you which ones are fraudulent very fast Disclaimer: I haven’t used the model and don’t understand how it works, just evaluating based on the company’s own claims. If the claims hold up…seems cool
After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI? I’ve spent the last 2 years in stealth building a new way to train models (RLCD), and a new type of frontier AI model that we are releasing today: Jev • 20-200x faster • 40-400x cheaper (w/ output tokens free) • Frontier composable intelligence optimized for decisions AFAICT the shortest path to AI-based economic revolution
2
1
461
Great point
This is cool, but if your output domain is known in advance, why not just train a model to produce logprobs over enums?
1
129
More of this! What a rich, information dense way to learn about other places
Okay but there are three kinds of NC rich
1
3
410
Another one that taught me a lot
Bay Area starter pack - mom edition
2
160
The wayfinder skill by @mattpocockuk is great At first I was skeptical, I’ve managed big projects just fine by myself But what’s nice about this skill is I never have the blank canvas problem. There’s always something to sharpen my ideas against There’s more room for taste expression because the “default” decision is already there in front of you. You can either accept and move on or refine Collaborative wayfinding sessions are also a great way to build alignment 10/10, great skill
1
5
292
uh oh, i forgot how to write code
2
128
Who’s right? When exploring an idea with an LLM my cofounder will avoid saying what he thinks to avoid a “you’re absolutely right,” he fears the agent will stop exploring other options I’ve found that agents will push back when I’m wrong so I’m not afraid to say what I think
1
131
With coding agents instead of saying: "Change X to Y" …I unconsciously say: "What do you think about changing X to Y?" I'm used to phrasing suggestions as a question in code review with humans to be polite. Now changing code is cheap, I should tell the agent to do the thing
4
2
290
Most of the time it does the thing I ask, sometimes it says "good idea, I'd make this change" and then I have to tell it to go actually make the change
53
Some amusing light mode bugs this morning in Codex. Am I the only light mode Codex user? @thsottiaux
188
I still experience flow but it’s different! It’s more hectic and chunky, less smooth. But I’m still absolutely engrossed and it’s hard to tear my attention away. I only reach flow state at 3+ productive agent threads running. Below that there’s too much waiting.
It’s a weird complaint to have but with AI agents, that “flow state” during coding is gone (because the coding itself is gone.) Every time I talk with a dev about work now vs before, it seems to come up, like today with a former colleague. Are we over-glorifying it (I mean it wasn’t all roses, it was also frustration) or is it something important / relevant they we had but now don’t really have?
3
4
814
I've been thinking a lot about what "AI-native productivity suite" actually means I think everyone chasing this has been approaching it from the wrong angle No one has asked the question "what does the AI want?" and designed for that first and foremost
2
372
Most repeatable processes in Google Docs are still: copy the old doc → rename it → manually update everything Alpine templates let you use lightweight {{variables}} throughout a document. Enter values once and every reference updates automatically
1
1
284
You're losing hours every week reformatting docs into slides. In Alpine the doc *is* the slide deck! Write it once and present it immediately.
265
Long Slack threads fall apart fast. Someone replies to point 2. Someone else replies to point 7. Now everyone’s copy-pasting quotes trying to preserve context. In our app you highlight text and reply directly. Every response stays anchored to the exact phrase it references.
1
3
721
image gallery + inline video player = ❤️
199
In Google Docs when you share a doc it goes into the black hole of email In our integrated chat + docs + tasks work app when you share anything it goes into your chat! Right in your chat inbox along with the rest of your conversation history. I love this feature.
1
219