We can’t wait to see you all for the SF Voice AI meetup today!
Casual AI chats and you take a voice app home!
3 hours to go, sign up below
AssemblyAI retweeted
AssemblyAI Universal 3.5 Pro topped our speech-to-text benchmark on both accuracy and latency.
We ran 15 models through the same audio and pipeline.
On 1,000 read-aloud clips from Pipecat FLEURS, Universal 3.5 Pro had the lowest word error rate at 1.93%.
It also had the fastest median time to first text at 489ms, so it started returning a transcript sooner than any other model we tested.
Congrats to the @AssemblyAI team.
Full results: benchmarks.cekura.ai/stt
8 more days left: we can’t wait to see what you all build!
Close your laptop. The agent keeps taking calls.
New tutorial: a voice agent on AssemblyAI's Voice Agent API with HTTP tools. No dispatcher, no WebSocket in your backend. AssemblyAI calls your API itself, mid-call.
lablab.ai/ai-tutorials/assem….
AssemblyAI retweeted
Can’t wait to see what amazing voice apps you build in 15 mins at the workshop!
See you all in San Francisco: luma.com/xwnkujzr
The fastest way to get text into a computer is to talk.If your product has a text field, it has somewhere for dictation to go.
Thursday in SF, @iHarnoorSingh shows you how to put it there in about 30 lines of Python.
Save your seat: luma.com/xwnkujzr
AssemblyAI retweeted
Universal-3.5 Pro from @AssemblyAI is live on OpenRouter, 50% off through September 29!
Speech-to-text in 19 languages in a single synchronous call. Ranks first for accuracy across independent speech-to-text benchmarks.
Send an audio clip of up to two minutes and quickly get the full transcript back with word-level timestamps and confidence scores. Steer transcription toward your domain by supplying keyterms and contextual prompting.
Try it: openrouter.ai/assemblyai/uni…
The fastest way to get text into a computer is to talk.If your product has a text field, it has somewhere for dictation to go.
Thursday in SF, @iHarnoorSingh shows you how to put it there in about 30 lines of Python.
Save your seat: luma.com/xwnkujzr
AssemblyAI retweeted
Honor to host NYC Voice AI Meetup this week for @AssemblyAI
100+ people showed up.
Great collaborating with @livekit and @boardyai
Building with voice is so much fun right now, comment down what you're building with voice, TTS, or voice agents :)
SF, check your mailboxes 👀
Invites to our SF Tech Week Voice Agent workshop went out this week.
Not sure "vintage" feels quite right, but we can't wait to see what you code into your voice-agent sidekick to clip onto your backpack and relive your childhood dreams.
Didn't get one? Request an invite 👇
luma.com/w9e4qgol
AssemblyAI retweeted
Most intake systems optimize for collecting fields, but guess what?
> HangON optimizes for preserving intent
A voice first front desk for operations teams, service desks, and community programs, built for the moments a misunderstood first handoff creates hours of rework
Fully built with @AssemblyAI's Voice Agent API
People rarely explain real requests in neat form fields
HangON lets callers explain naturally, then asks only the questions the workspace has configured
The result is a structured request, without forcing people through a rigid maze
Before anything is prepared, HangON reads the request back
The caller can correct it, clarify it, or confirm it.
That small interaction matters, because "I understood you" is not the same as "I am allowed to act."
The architecture runs from a browser microphone, through AudioWorklet capture, into a server minted AssemblyAI session, through workspace owned prompts and tools, to a signed confirmation, an idempotent request store, and an optional signed webhook
Confirmation is not just a boolean, because HangON signs approval against the exact workspace, summary, details, route, and idempotency key
Change one detail and the approval no longer matches
We built this with just one goal: make voice intake feel human without making systems less accountable
HangON is for teams that need better first handoffs, not more forms
For callers, it feels conversational
For operators, it produces a request they can inspect and trust
Try the public demo: tryhangon.vercel.app
Thank you
Blurt: push‑to‑talk dictation on the AssemblyAI Dictation API
The Dictation API converts a spoken clip into finished text. Filler words and false starts are removed, and the output can be formatted as notes, a commit message, or a reply to a customer.
Built on Universal‑3.5 Pro, it supports 19 languages, processes short clips in under a second, and costs $0.62/hr.
Link in comments below
AssemblyAI retweeted
I spent the morning building a small video transcription platform for myself.
Took a few back-and-forths with Astra to get this built and deployed on my Dokploy instance.
Part of my work to replace existing SaaS subscriptions with my own tools.
This saves ~120/year.
AssemblyAI retweeted
if you don't have anything to do this weekend,
We have something for you!
We’re kicking off Voice Hackathon Week: Hack into Dictation with @AssemblyAI
Join one of the most talented builder communities across the globe
luma.com/qwckwa01
AssemblyAI retweeted
Dylan from @AssemblyAI started building in the voice infra space 5 years before it became obvious.
Loved visiting the office and hearing his story - coming soon!
I found blurt today, open source, built by @AssemblyAI to showcase its dictation api, its fast and very accurate, you can just build your own with dictation api and you get $100 credits on assembly as a new user, crazy value
assemblyai.com/blurt
AssemblyAI retweeted
Voice is becoming the default way we communicate across all apps.
Let’s integrate dictation into your apps, such as lengthy forms (Google Form / Typeform), job applications, paperwork, and more.
Meet the fastest way to ship voice input @AssemblyAI Dictation API:
AssemblyAI retweeted
Built CareEcho for the @AssemblyAI Voice Hackathon
CareEcho is a multilingual, voice-first health memory that helps patients remember what happens before, during, and after doctor visits.
▪️Speak your symptoms --> build a health timeline
▪️Record a consultation → ask later:
“What did my doctor say about my medication?”
CareEcho answers from the actual visit evidence, not guesses.
Built for accessibility, multilingual users, older adults, and anyone who struggles to keep track of health information.
Your health memory, in your voice. 💙