@pingToveni
iAccount based inUnited States
About this account
- Account based in
- United States
- Connected via
- United States App Store
Account-level information from X, not a live location or the device used for a specific post.
leading inference provider ops (i'm hiring!) @openrouter. opinions my own. aka tomas. 🇦🇷🇦🇷🇦🇷
vibe/acc
Joined October 2021
- Tweets8.4K
- Following3.8K
- Followers7.5K
- Likes64.4K
Pinned Tweet
the provider operations team at openrouter is absolutely cooking on automations. 727 new endpoints added in august alone, 23 endpoints a day!
our systems allow 75+ providers to self serve onboard their endpoints onto our platform (gated on quality and functionality tests)
Toven retweeted
Xiaomi MiMo-V2.6 from @XiaomiMiMo is live on OpenRouter.
Three new models. A 1T+ parameter flagship, an open-source MoE, and a ~10x faster variant of the flagship. All three take text, image, video, and audio with 1M context.
More in the thread 🧵
Toven retweeted
The Jev moment happening now is similar in energy to the OpenClaw moment that happened in January, and also to the Opus 4.5 moment before that. Developers are scrambling to find use cases for a new hot thing, and it's a sudden blooming of creativity.
Models have typically NOT optimized for specific use cases and have gone the other direction, generalizing over all of them. This could be a very important moment for the whole AI ecosystem if this turns out to be the beginning of other model skews that make the market much more diverse, such as compaction, summarization, extraction and more, all of which can now be used to optimize harnesses and decouple them from provider lock-in.
We'll see what happens over the course of the next few weeks when the model labs optimize their small models more or try new ways of branding and packaging them.
Toven retweeted
.@openrouter tip: you can get useful notifications like these directly in your team's Slack
Stay on top of model changes, low credits, workspace budgets, and API key limits with Notifications in OpenRouter!
Pick which events notify you and how they're delivered 🧵 openrouter.ai/settings/notif…
Toven retweeted
Stay on top of model changes, low credits, workspace budgets, and API key limits with Notifications in OpenRouter!
Pick which events notify you and how they're delivered 🧵 openrouter.ai/settings/notif…
Toven retweeted
Quite an interesting turn in LLM diversity, if the market starts looking for specialized models that are optimized and branded for different use cases.
Should be very good for the market and open more room for non-frontier neolabs!
Toven retweeted
new york jeveloper night just got bigger: @OpenRouter is providing credits for attendees!
we’re capping the room at 150, and @jjacky will share how he’s been working directly with the @typesafeai team.
6pm this thursday at @betaworks.
demos + food + drinks.
register: luma.com/xogxfokf
new york jeveloper night
6pm this thursday at @betaworks
demos + food + drinks
register: luma.com/xogxfokf
jev by @typesafeai
Toven retweeted
What is a decision model?
Jev by @typesafeai answers yes/no and multiple-choice questions, with a confidence score. Much of software development are a sequence of decisions, and Jev is 10x cheaper and faster than an LLM.
Let’s understand this through practical examples:
Toven retweeted
Jev by @typesafeai is now on OpenRouter, in beta.
Jev is a System One model. Instead of generating text, it takes your app's state plus a typed question and returns a typed decision with a probability attached. There is no JSON prompting, parsing layer, and nothing to validate against.
Toven retweeted
Proudest moment of my career: being featured on @OpenRouter 's new Meme Benchmarks openrouter.ai/benchmarks/med… with @pingToven
Toven retweeted
Replying to @OpenRouter
Jev by @typesafeai is now on OpenRouter, in beta.
Jev is a System One model. Instead of generating text, it takes your app's state plus a typed question and returns a typed decision with a probability attached. There is no JSON prompting, parsing layer, and nothing to validate against.
(jev)ons paradox? openrouter.ai/typesafe/jev-1…
Toven retweeted
"The most important thing about the intelligence layer is the model.”
But what happens when developers can choose from hundreds of them?
@vincentweisser of @PrimeIntellect joined us for New Defaults to talk open models, model choice, routing, economics, and what happens as intelligence becomes increasingly open and competitive.
Toven retweeted
ZCode now supports more model providers, with improved stability and performance.
We’ll keep expanding integrations based on your feedback. Which models do you like most beyond the GLM series?