🚀 Introducing XOR
An open source, multimodal Jev-like decision model. Being developed with data sovereignty & enterprise grade in mind.
❶ Multimodal: great for text + visual classification
❷ Much bigger 260k context window
❸ Qwen3.6 based
huggingface.co/juspay/xor
↓↓↓ See it in action
joseph retweeted
Just watched a frontier LLM debug a website by literally turning my wifi off and on - but ofc it couldn’t actually turn the wifi back on because its connection was severed
joseph retweeted
Alexandr Wang@alexandr_wang
Sep 22United States
United StatesConnected via United States App StoreAccount-level information, not a live location or per-post device.
state of the art motherfuckers
joseph retweeted
636606729769440499166579950236036751749912014371509557713570027508971809534551913252252094954941974952859310861988904737359709200557919
is a factor of RSA-896
saweis.net/posts/rsa-896.htm…
joseph retweeted
Introducing Step 5 Preview: Advancing the Pareto Frontier.
Step 5 Preview is our new flagship model for agentic work, delivering frontier-level performance across software engineering and professional knowledge work, with particular strength in finance.
- 600B total / 27B active MoE, with 1M context + Vision
- Substantially lower task cost at comparable intelligence
- Broad software engineering capabilities with sustained execution over long horizons
Try Step 5 Preview: platform.stepfun.ai
Model page: stepfun.com/step-5-preview
Open weights on Oct 15.
I got early access to @typesafeai’s Jev—the new “System One” model that doesn’t generate text.
I tested the live API.
Headline: 50 semantic judgments in 226 ms.
One state, one request, all answers together.
The parallelism looks real. 🧵
After co-inventing ChatGPT, I kept asking myself: why have superhuman chat models not led to AGI?
I’ve spent the last 2 years in stealth building a new way to train models (RLCD), and a new type of frontier AI model that we are releasing today: Jev
• 20-200x faster
• 40-400x cheaper (w/ output tokens free)
• Frontier composable intelligence optimized for decisions
AFAICT the shortest path to AI-based economic revolution
Where this is useful:
• route tickets/events/docs
• score risk, urgency, relevance
• audit agent traces and claims
• gate cheap model → expensive model → human
• monitor huge streams and wake an agent only when a semantic condition hits
glm5.3 flash at 28-30 tok/s on a single spark 🥸
need some more tweaks and verifications but recipe is dropping soon
got it up to 60 tok/s try it out github.com/gitcommit90/glm-5…