Founder, VideoFire (SPC F25), CTO & AI Architect, Scaled & Sold Datastreamer ($2M+ ARR, Acquired), 2 exits.

San Francisco, CA
Joined April 2018
Include

Only show posts containing:

Exclude

Hide posts containing:

Time range
-
Minimum likes
Kevin Burton
@inputneuron
23h
Replying to @trq212
Maybe you just invented a new feature for Claude. The ability to easily share the prompt + context. That would be really cool actually
1
15
Kevin Burton
@inputneuron
Sep 27
Literally said yesterday someone needs to ship a nerfbench . Did you steal my idea? If so I say run with it !
8
Kevin Burton
@inputneuron
Sep 27
Replying to @trq212
I've been working with agents 16 hours a day 7 days a week since Feb... speak for yourself :)
62
Kevin Burton
@inputneuron
Sep 25
It's arguable because I don't think they've ever guaranteed that you're going to see a full 16-bit quant, but I feel what you're saying. It should be.
29
Kevin Burton
@inputneuron
Sep 25
Replying to @jaredctate
We need a cheap 'nerfbench' that people can run so that you can spend say $20-50 to see if a specific model+provider has regressed.
1
588
Kevin Burton
@inputneuron
Sep 24
Replying to @trq212
Not your plan mode but mine which is which is derived from @mattpocockuk 's version and has some extra features ... it resolves ambiguity and forces you to think through complicated specs..
224
Kevin Burton
@inputneuron
Sep 23
Replying to @MiaAI_lab
I wish there was a way to cryptographically verify that the model you're running on is what you expect and isn't being yanked from you. 3rd party inference providers also lie about the models they sell and use a quantized/nerfed version but sell it at premium pricing.
107
Kevin Burton
@inputneuron
Sep 23
Replying to @iannuttall
And let us use the harness in any model
130
Kevin Burton
@inputneuron
Sep 22
Replying to @hany99dev
"I feel like intelligence is becoming too cheap" ... bro shut up! You're going to ruin it!
175
Kevin Burton
@inputneuron
Sep 22
Replying to @AlexFinn
I'm really hoping this is true. And you're right.. it's like "reading fog" ... you see something but then when you reach for it nothing's there. It's really weird.
165
Kevin Burton
@inputneuron
Sep 22
Switch over now to increase your token usage!
Replying to @claudeai
Opus 5.5 requires less compute to serve than Opus 5, and its pricing reflects that. Our tests show that at default settings it will cost 40% less than Opus 5 on typical workloads.
1
20
Kevin Burton
@inputneuron
Sep 22
Replying to @Shpigford
... when you want to waste more tokens and give more money to Anthropic.
1
7
Kevin Burton
@inputneuron
Sep 22
Replying to @claudeai
As a masochist, this is NOT the Claude I've come to love. I expect abuse and for you to ignore customer requests!!! How dare you!
6
1,150
Kevin Burton
@inputneuron
Sep 22
Replying to @scheemunai
WAIT .. you don't think 15 tokens per second is fast? /s This is on z.ai btw... Apparently, Fireworks.ai is faster and like 250 t/ps but I haven't used it yet. I think Claude is still a better value - but might be wrong.
1
1,191
Kevin Burton
@inputneuron
Sep 22
Replying to @synthwavedd
It's totally going to happen. Anthropic always does the right thing - but only after exhausting all possible options.
417
Kevin Burton
@inputneuron
Sep 20
Replying to @oldstackjournal
Compiling my own Linux kernel. I used to do this constantly in the 2000s but at some point I stopped doing it and it was my last time. Kind of sad to think about
1
3
133
Kevin Burton
@inputneuron
Sep 18
Replying to @polynoamial
"One lesson from the HF incident is that we put too much trust in sandbox isolation" ... could I go so far as to say it was negligent? It should have been obvious to anyone with a security background that this could have / would have happened.
1
9
386
Kevin Burton
@inputneuron
Sep 18
Replying to @thesupermannx
This is great and I'm glad more people are talking about this. I thought about this a few years ago as a form of 'neural compression' where you could put an LLM on mars and one on earth and just send embeddings for INSANELY fast compression.
2
438
Kevin Burton
@inputneuron
Sep 12
100% ... it's only accelerate. Don't listen to anyone trying to decelerate.
2
37
Kevin Burton
@inputneuron
Sep 9
Replying to @elonmusk
Understatement of the century... Completely agree.
7