@pingToven

leading inference provider ops (i'm hiring!) @openrouter. opinions my own. aka tomas. 🇦🇷🇦🇷🇦🇷

vibe/acc
Joined October 2021
the provider operations team at openrouter is absolutely cooking on automations. 727 new endpoints added in august alone, 23 endpoints a day! our systems allow 75+ providers to self serve onboard their endpoints onto our platform (gated on quality and functionality tests)
9
8
3
63
12,748
man i kinda miss the days where i could read every single openrouter mention on here
4
1
28
1,309
Toven retweeted
Xiaomi MiMo-V2.6 from @XiaomiMiMo is live on OpenRouter. Three new models. A 1T+ parameter flagship, an open-source MoE, and a ~10x faster variant of the flagship. All three take text, image, video, and audio with 1M context. More in the thread 🧵
34
36
13
766
46,642
The Jev moment happening now is similar in energy to the OpenClaw moment that happened in January, and also to the Opus 4.5 moment before that. Developers are scrambling to find use cases for a new hot thing, and it's a sudden blooming of creativity. Models have typically NOT optimized for specific use cases and have gone the other direction, generalizing over all of them. This could be a very important moment for the whole AI ecosystem if this turns out to be the beginning of other model skews that make the market much more diverse, such as compaction, summarization, extraction and more, all of which can now be used to optimize harnesses and decouple them from provider lock-in. We'll see what happens over the course of the next few weeks when the model labs optimize their small models more or try new ways of branding and packaging them.
11
12
3
180
11,034
.@openrouter tip: you can get useful notifications like these directly in your team's Slack
Stay on top of model changes, low credits, workspace budgets, and API key limits with Notifications in OpenRouter! Pick which events notify you and how they're delivered 🧵 openrouter.ai/settings/notif…
3
2
1
34
7,668
Toven retweeted
Stay on top of model changes, low credits, workspace budgets, and API key limits with Notifications in OpenRouter! Pick which events notify you and how they're delivered 🧵 openrouter.ai/settings/notif…
11
3
2
80
16,642
Quite an interesting turn in LLM diversity, if the market starts looking for specialized models that are optimized and branded for different use cases. Should be very good for the market and open more room for non-frontier neolabs!
29
13
6
197
12,270
What is a decision model? Jev by @typesafeai answers yes/no and multiple-choice questions, with a confidence score. Much of software development are a sequence of decisions, and Jev is 10x cheaper and faster than an LLM. Let’s understand this through practical examples:
67
124
31
1,681
165,262
are @typesafeai twitter affiliates jevrels
6
1
65
1,994
y'all are going ham on jev damn
7
18
980
Jev by @typesafeai is now on OpenRouter, in beta. Jev is a System One model. Instead of generating text, it takes your app's state plus a typed question and returns a typed decision with a probability attached. There is no JSON prompting, parsing layer, and nothing to validate against.
143
283
124
3,860
544,771
Proudest moment of my career: being featured on @OpenRouter 's new Meme Benchmarks openrouter.ai/benchmarks/med… with @pingToven
2
2
11
623
Replying to @OpenRouter
Jev by @typesafeai is now on OpenRouter, in beta. Jev is a System One model. Instead of generating text, it takes your app's state plus a typed question and returns a typed decision with a probability attached. There is no JSON prompting, parsing layer, and nothing to validate against.
43
56
20
1,799
174,337
Unsurprising that the 2 flagship labs are making far more revenue than their mostly open-source competition. But rate of change also matters 🙏
SITUATION DETECTED: OpenAI and Anthropic are making 10 times more revenue than all Chinese AI models combined, per CNBC
1
4
1
40
11,221
"The most important thing about the intelligence layer is the model.” But what happens when developers can choose from hundreds of them? @vincentweisser of @PrimeIntellect joined us for New Defaults to talk open models, model choice, routing, economics, and what happens as intelligence becomes increasingly open and competitive.
12
7
4
119
16,757
ZCode now supports more model providers, with improved stability and performance. We’ll keep expanding integrations based on your feedback. Which models do you like most beyond the GLM series?
203
41
10
1,028
97,869
twitter when we launch a stealth model
4
1
21
3,183