Access powerful AI models to transcribe and understand speech via a simple API. Try our no-code playground for free 👉 https://nitter.cf/t.co/YPCK9mqDG6

Joined October 2017
We can’t wait to see you all for the SF Voice AI meetup today! Casual AI chats and you take a voice app home! 3 hours to go, sign up below
4
7
1,459
SF Voice AI Meetup today! Excited to have you all here! Sign up link below:
2
1
22
1,876
AssemblyAI Universal 3.5 Pro topped our speech-to-text benchmark on both accuracy and latency. We ran 15 models through the same audio and pipeline. On 1,000 read-aloud clips from Pipecat FLEURS, Universal 3.5 Pro had the lowest word error rate at 1.93%. It also had the fastest median time to first text at 489ms, so it started returning a transcript sooner than any other model we tested. Congrats to the @AssemblyAI team. Full results: benchmarks.cekura.ai/stt
2
4
21
5,762
8 more days left: we can’t wait to see what you all build!
Close your laptop. The agent keeps taking calls. New tutorial: a voice agent on AssemblyAI's Voice Agent API with HTTP tools. No dispatcher, no WebSocket in your backend. AssemblyAI calls your API itself, mid-call. lablab.ai/ai-tutorials/assem….
4
4
1,556
Can’t wait to see what amazing voice apps you build in 15 mins at the workshop! See you all in San Francisco: luma.com/xwnkujzr
The fastest way to get text into a computer is to talk.If your product has a text field, it has somewhere for dictation to go. Thursday in SF, @iHarnoorSingh shows you how to put it there in about 30 lines of Python. Save your seat: luma.com/xwnkujzr
1
3
1,258
AssemblyAI retweeted
Universal-3.5 Pro from @AssemblyAI is live on OpenRouter, 50% off through September 29! Speech-to-text in 19 languages in a single synchronous call. Ranks first for accuracy across independent speech-to-text benchmarks. Send an audio clip of up to two minutes and quickly get the full transcript back with word-level timestamps and confidence scores. Steer transcription toward your domain by supplying keyterms and contextual prompting. Try it: openrouter.ai/assemblyai/uni…
7
5
1
131
22,036
The fastest way to get text into a computer is to talk.If your product has a text field, it has somewhere for dictation to go. Thursday in SF, @iHarnoorSingh shows you how to put it there in about 30 lines of Python. Save your seat: luma.com/xwnkujzr
2
1
4
10,495
Honor to host NYC Voice AI Meetup this week for @AssemblyAI 100+ people showed up. Great collaborating with @livekit and @boardyai Building with voice is so much fun right now, comment down what you're building with voice, TTS, or voice agents :)
2
1
32
11,940
SF, check your mailboxes 👀 Invites to our SF Tech Week Voice Agent workshop went out this week. Not sure "vintage" feels quite right, but we can't wait to see what you code into your voice-agent sidekick to clip onto your backpack and relive your childhood dreams. Didn't get one? Request an invite 👇 luma.com/w9e4qgol
4
3
9
10,556
AssemblyAI retweeted
Most intake systems optimize for collecting fields, but guess what? > HangON optimizes for preserving intent A voice first front desk for operations teams, service desks, and community programs, built for the moments a misunderstood first handoff creates hours of rework Fully built with @AssemblyAI's Voice Agent API People rarely explain real requests in neat form fields HangON lets callers explain naturally, then asks only the questions the workspace has configured The result is a structured request, without forcing people through a rigid maze Before anything is prepared, HangON reads the request back The caller can correct it, clarify it, or confirm it. That small interaction matters, because "I understood you" is not the same as "I am allowed to act." The architecture runs from a browser microphone, through AudioWorklet capture, into a server minted AssemblyAI session, through workspace owned prompts and tools, to a signed confirmation, an idempotent request store, and an optional signed webhook Confirmation is not just a boolean, because HangON signs approval against the exact workspace, summary, details, route, and idempotency key Change one detail and the approval no longer matches We built this with just one goal: make voice intake feel human without making systems less accountable HangON is for teams that need better first handoffs, not more forms For callers, it feels conversational For operators, it produces a request they can inspect and trust Try the public demo: tryhangon.vercel.app Thank you
6
1
1
12
1,063
Blurt: push‑to‑talk dictation on the AssemblyAI Dictation API The Dictation API converts a spoken clip into finished text. Filler words and false starts are removed, and the output can be formatted as notes, a commit message, or a reply to a customer. Built on Universal‑3.5 Pro, it supports 19 languages, processes short clips in under a second, and costs $0.62/hr. Link in comments below
7
2
12
1,715
BTW, we intentionally spelled Minoxidil as “Minoxidril”; the model spells it exactly as provided in the keyterms.
2
234
I spent the morning building a small video transcription platform for myself. Took a few back-and-forths with Astra to get this built and deployed on my Dokploy instance. Part of my work to replace existing SaaS subscriptions with my own tools. This saves ~120/year.
7
2
1
30
6,403
if you don't have anything to do this weekend, We have something for you! We’re kicking off Voice Hackathon Week: Hack into Dictation with @AssemblyAI Join one of the most talented builder communities across the globe luma.com/qwckwa01
1
2
17
3,038
Dylan from @AssemblyAI started building in the voice infra space 5 years before it became obvious. Loved visiting the office and hearing his story - coming soon!
2
2
37
2,374
AssemblyAI retweeted
Replying to @maxktz @raycast
I found blurt today, open source, built by @AssemblyAI to showcase its dictation api, its fast and very accurate, you can just build your own with dictation api and you get $100 credits on assembly as a new user, crazy value assemblyai.com/blurt
1
239
Voice is becoming the default way we communicate across all apps. Let’s integrate dictation into your apps, such as lengthy forms (Google Form / Typeform), job applications, paperwork, and more. Meet the fastest way to ship voice input @AssemblyAI Dictation API:
2
1
21
22,969
Built CareEcho for the @AssemblyAI Voice Hackathon CareEcho is a multilingual, voice-first health memory that helps patients remember what happens before, during, and after doctor visits. ▪️Speak your symptoms --> build a health timeline ▪️Record a consultation → ask later: “What did my doctor say about my medication?” CareEcho answers from the actual visit evidence, not guesses. Built for accessibility, multilingual users, older adults, and anyone who struggles to keep track of health information. Your health memory, in your voice. 💙
8
8
22
617