AI Architect & Builder - Inference, LLMs, SLMs GSoC Org Admin - Jenkins | Docker Captain | CNCF Ambassador

New Delhi
Joined December 2018
What an amazing day! Presented a poster at @PyTorch Conference: how we optimised the RLM paper with pre-fix caching & batched Sub calls with @vllm_project And a full house for our Pytorch Conference talk on running LLMs using Executorch on Android!
4
5
2
84
12,840
Shivay Lamba retweeted
Theo - t3.gg
@theo
3h
Dev Day announcement overview, as short as reasonably possible. - Dots: personal assistant (similar to GrokBot, Muse etc) - $500/month plan including Ultrafast - GPT-6.1 Sol: really smart and cheap, way better than 6 was - Decisions API: Jev killer with vision capabilities - ultrafast is real (300tps on Astra). It’s way expensive though. - sign in with ChatGPT: use your sub and inference in 3rd party apps - extensions in ChatGPT can now be full apps - 1 banked reset for all I think that’s most of it!
223
207
52
5,393
299,146
I think Muse will stay
8
Shivay Lamba retweeted
.@Qualcomm 🫶 @Linux The commitment to our ecosystem never stops. Discover how our support for Linux deepens each year:
3
15
1
80
2,534
Shivay Lamba retweeted
Vaibhav (VB) Srivastav
@reach_vb
3h
GPT-6.1 Sol matches Astra on DeepSWE at ~1/5 the cost!! $2 input, $10 output and just $0.10 cached input per million tokens. Enjoy!!
24
15
4
313
11,002
Shivay Lamba retweeted
OpenAI
@OpenAI
4h
Introducing dots, powered by GPT-6 Astra. Remarkably capable, always-on agents built to handle everything.
1,316
1,952
1,837
24,591
5,813,217
Shivay Lamba retweeted
Introducing 6.1 Sol, near Astra intelligence at one fifth of the price of Astra and 95% cache read discount. It is an absolute workhorse. Combined with ultrafast for 8X speeded available today for Astra and coming soon for 6.1 Sol.
897
691
367
14,936
1,125,919
Shivay Lamba retweeted
I’ll explain the new Pro 200 plan differently, before I start live tweeting from DevDay on things that are going out! Today we are going to ship a number of things that increase what you can do across the Plus and Pro plans. A lot of compute is online for this increase. As we increase the floor, we are changing the relative difference between plans to be Plus = 1X Pro 100 = 5X Pro 200 = 10X and we are reopening subscriptions for Pro 200 (we had paused it). If you have an existing plan you will keep the 20X multiplier for a bit and also receive a lot of additional credits because we know changes are hard even if it means that everyone will get more in the end.
3,096
447
1,077
9,258
1,708,379
Best person in AI one can hire
Professional Update: I have transitioned from my full-time role at iii, though I remain committed to contributing to the company's ongoing success, as I deeply value the innovative work we have accomplished together. If you are developing next ambitious thing and require expertise in Engineering, Product, or Developer Relations, hit me up!
1
3
513
Shivay Lamba retweeted
OpenAI
@OpenAI
Sep 28
Get ready.
2,950
3,664
3,664
62,833
16,223,109
Shivay Lamba retweeted
Training AI agents takes more than fast token generation. Tools run. Tests stall. Training batches wait. We cut fixed-weight batch collection time by 66.9% on a Terminal-Bench-based workload by improving routing, sandbox execution and scheduling. Here’s how: nebius.com/blog/posts/in-age…
4
4
3
27
9,729
Shivay Lamba retweeted
Theo - t3.gg
@theo
Sep 28
A $200 Claude Code sub gets you ~$9,000 of Opus usage per month. I have managed to run 3 Claude accounts down to 0% since Opus 5.5 dropped. When analyzing their usage, costs were $2,086, $2,442, and $2,182. Average is $2,200/week, so ~$9,000/month
272
90
73
5,733
1,497,015
looks so nice, can you share the prompt?
1
27
Replying to @Astrodevil_
looks good!
1
21
Replying to @Sauain
Legend
1
68
Replying to @whyrohitwhy
No no I meant countries I have visited
1
205
Replying to @arsh_goyal
And Antartica
2
191