Introducing our Voice Agent API: The fastest path to a working voice agent Learn more
Blog
All
Releases & Updates
Insights & Use Cases
News
Video
Newsletter
Thank you! Your submission has been received!
Oops! Something went wrong while submitting the form.
April 29, 2026
Create an ambient AI scribe that works during telehealth video calls
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 29, 2026
Voice agents in noisy environments (Drive-Thrus, Contact Centers, Field)
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 29, 2026
When to use Voice Agent API vs. Universal-3 Pro Streaming
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 29, 2026
Voice Agent Orchestrators Compared: Vapi vs Pipecat vs LiveKit with AssemblyAI
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 29, 2026
What streaming speech to text model is best for voice agents and why?
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 29, 2026
Build a voice agent with a chained STT-LLM-TTS architecture
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 29, 2026
Migrating from OpenAI Realtime API to AssemblyAI Voice Agent API
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 29, 2026
Introducing our Voice Agent API
Releases & Updates
By
Madison Bernstein
,
Product Marketing
April 29, 2026
How to choose the best speech-to-text API for voice agents
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 29, 2026
Top APIs and models for real-time speech recognition and transcription in 2026
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 28, 2026
Real-time vs batch transcription: What's the difference?
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 28, 2026
5 Google Cloud Speech-to-Text alternatives in 2026
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 28, 2026
Transformative use cases of AI in contact centers
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 27, 2026
Noise cancellation with speech-to-text: The pros and cons
Insights & Use Cases
By
David Lange
,
Applied AI Engineer
April 22, 2026
How to create an AI cold-calling agent
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 22, 2026
How to create a phone-based voice agent
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 21, 2026
What's the best medical transcription API?
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 21, 2026
What is the difference between speaker recognition and speaker verification?
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 21, 2026
5 Speechmatics alternatives in 2026
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 21, 2026
Top 8 open source STT options for voice applications in 2026
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 21, 2026
Conversational AI in healthcare: maturity model and 7 use cases
Insights & Use Cases
By
Jesse Sumrak
,
Featured writer
April 21, 2026
How to run OpenAI's Whisper speech recognition model
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
April 21, 2026
Conversation AI: What it is and top use cases
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 20, 2026
Voice AI Meetup recap: How Commure and Ona Health are building for healthcare
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 14, 2026
Word error rate is broken: How to actually evaluate speech-to-text in 2026
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 14, 2026
How to vibe code a voice agent (and why AI always recommends AssemblyAI)
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 14, 2026
Build a voice agent with function calling
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 14, 2026
Build a voice agent with LiveKit
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 13, 2026
Can transcripts be used to generate meeting agendas?
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 13, 2026
Beyond transcription: Combining speech-to-text with AI analysis
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 13, 2026
5 Deepgram alternatives in 2026
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 13, 2026
How accurate is speech-to-text in 2026?
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 13, 2026
Build a call center analytics pipeline in Python with AssemblyAI
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 13, 2026
Best AI playgrounds in 2026
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 13, 2026
Content moderation: What it is, how it works, and the best APIs
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 9, 2026
Why speech-to-text accuracy is the hidden bottleneck in your AI agent pipeline
By
,
April 8, 2026
Tutorial: How to easily build a voice agent with AssemblyAI
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 8, 2026
How to build a voice agent with Python in 5 minutes
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 8, 2026
When to stop self-hosting Whisper (and what you actually gain)
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 8, 2026
Edge cases in transcription: Offline mode, partial audio files and API limits
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 8, 2026
Building with transcripts: Search, indexing, display and downstream Integrations
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 8, 2026
Transcript output guide: SRT, VTT & TXT export formats + live caption sync
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 8, 2026
How to build an AI-Powered interview scoring system with speech-to-text
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 8, 2026
How to evaluate speech recognition models
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 6, 2026
Voice AI guardrails: Built-in protection for compliance, quality, and cost control
Releases & Updates
By
Kelsey Foster
,
Growth
April 6, 2026
Medical voice recognition: How AI solves terminology problems
Insights & Use Cases
By
,
April 6, 2026
AI voice agents: what they are and how they work in 2026
Insights & Use Cases
By
Jesse Sumrak
,
Featured writer
April 6, 2026
How to choose the best speech-to-text API
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 6, 2026
How to automatically redact PII from audio and video files with Python
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
April 6, 2026
The top free speech-to-text APIs, AI models, and open source engines
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 2, 2026
Agora voice agent with AssemblyAI Universal-3 Pro Streaming
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 2, 2026
Node.js voice agent with AssemblyAI Universal-3 Pro Streaming
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 2, 2026
Daily.co voice agent with AssemblyAI Universal-3 Pro Streaming
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 2, 2026
Raw WebSocket voice agent with AssemblyAI Universal-3 Pro Streaming
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 2, 2026
Retell AI + AssemblyAI: custom LLM and post-call analytics
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 2, 2026
Vapi voice agent with AssemblyAI Universal-3 Pro Streaming
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 2, 2026
Twilio phone agent with AssemblyAI Universal-3 Pro Streaming
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 2, 2026
Pipecat voice agent with AssemblyAI Universal-3 Pro Streaming
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 2, 2026
Speech-to-text for HR and recruiting: Interview transcription, screening and scoring
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 2, 2026
How to build a lecture capture system with speaker identification
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 2, 2026
Speech-to-Text for EdTech: Lectures, Captions, Accessibility & Assessments
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 2, 2026
Best Nuance Dragon medical alternatives for clinical documentation
Insights & Use Cases
By
,
April 2, 2026
AssemblyAI vs Rev AI: Accuracy, pricing and features compared
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 2, 2026
Handling transcript errors: Homophones, corrections and AI quality improvement
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 31, 2026
LiveKit voice agent with AssemblyAI Universal-3 Pro Streaming
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 31, 2026
AssemblyAI vs Deepgram for medical transcription
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 31, 2026
What is the best speech to text api to build ai medical ambient scribes?
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 31, 2026
AI medical transcription
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 31, 2026
The best audio file formats for speech-to-text: A guide
Insights & Use Cases
By
Patrick Loeber
,
Senior Developer Advocate
March 26, 2026
Medical transcription that actually works — Beyond generic STT
By
Kelsey Foster
,
Growth
March 26, 2026
What is LLM Gateway?
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 26, 2026
Best medical speech-to-text in 2026
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 26, 2026
Speech-to-text for healthcare developer guide
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 25, 2026
Introducing Medical Mode: Purpose-built accuracy for medical terminology
Releases & Updates
By
Madison Bernstein
,
Product Marketing
.png)
March 24, 2026
Turn detection vs forced endpoints in voice AI: Why getting this wrong tanks your UX
Insights & Use Cases
By
Kelsey Foster
,
Growth
.png)
March 24, 2026
Real-time transcription in Python with Universal-3 Pro Streaming
Insights & Use Cases
By
Kelsey Foster
,
Growth
%20benchmark%20might%20be%20lying%20to%20you%20(1).png)
March 24, 2026
Why your word error rate (WER) benchmark might be lying to you
Insights & Use Cases
By
Zackary Klebanoff
,
Applied AI Lead
.png)
March 23, 2026
How do I transcribe audio in languages like Spanish, French, or German?
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 23, 2026
Are there language-specific models for better accuracy?
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 23, 2026
What metrics can I get from transcribed call center data?
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 23, 2026
Can voice AI recognize the topic or themes of a conversation?
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 23, 2026
AssemblyAI Named a Leader in G2’s Spring 2026 Voice Recognition Report
News
By
Devon Malloy
,
Staff Growth Manager
March 23, 2026
Large-scale audio transcription: Handling hours of content efficiently
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 23, 2026
What is real-time speech to text?
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 23, 2026
Do I need a custom speech recognition model?
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 18, 2026
Speech-to-text API pricing guide: Per-minute, per-hour and feature costs explained
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 18, 2026
How to evaluate and choose the best speech to text API for enterprises
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 18, 2026
10 call center metrics you can extract from transcripts with AI
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 18, 2026
How real-time agent assist is changing conversation intelligence
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 18, 2026
Streaming speaker diarization: How to identify who's speaking in real time
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 18, 2026
Real-time entity extraction from speech: Capturing emails, phone numbers, and addresses in live audio
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 18, 2026
Building a production-ready voice agent: The developer's guide to real-time speech-to-text
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 18, 2026
10 best agent assist software in 2026
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 18, 2026
Real-time conversation intelligence: The shift from post-call analysis to live insights
Video
By
Kelsey Foster
,
Growth
.png)
March 17, 2026
Multilingual streaming with Universal-3 Pro: Native code switching across 6 languages
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 17, 2026
What is audio intelligence or speech understanding?
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 17, 2026
What is speaker diarization and how does it work? (Complete 2026 Guide)
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 6, 2026
Speech-to-text prompting with AssemblyAI Universal-3 Pro
Insights & Use Cases
By
Martin Schweiger
,
Senior Technical Product Marketing Manager
March 4, 2026
How accurate is AI transcription for pharmaceutical drug names?
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 4, 2026
Contact center AI trends for 2026
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 4, 2026
Best scalable voice AI solutions for customer service
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 3, 2026
Universal-3 Pro Streaming: The most accurate real-time transcription model for voice agents
Releases & Updates
By
Madison Bernstein
,
Product Marketing
March 3, 2026
Conversation Intelligence: The complete guide for 2026
Insights & Use Cases
By
,
March 3, 2026
Text Summarization for NLP: 5 Best APIs, AI Models, and AI Summarizers in 2026
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 3, 2026
8 best transcript summarizers powered by AI
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 3, 2026
What is speech to text? The complete guide
Insights & Use Cases
By
Jesse Sumrak
,
Featured writer
March 3, 2026
6 best named entity recognition APIs for entity detection
Insights & Use Cases
By
Kelsey Foster
,
Growth
February 26, 2026
How to get the most out of Universal-3 Pro with prompt engineering
Insights & Use Cases
By
Ryan Seams
,
VP, Customer Solutions
February 26, 2026
AssemblyAI Universal-3 Pro vs Deepgram Nova-3: An honest comparison for developers
Insights & Use Cases
By
Kelsey Foster
,
Growth
February 26, 2026
Multilingual speech recognition in 2026: How Universal-3 Pro handles accents, code-switching, and non-English audio
Insights & Use Cases
By
Kelsey Foster
,
Growth
February 26, 2026
What is speaker fingerprinting for Voice AI
Insights & Use Cases
By
Kelsey Foster
,
Growth
February 25, 2026
How do I build an AI medical scribe using speech-to-text?
Insights & Use Cases
By
Kelsey Foster
,
Growth
February 25, 2026
Multi-language voice agents: Building agents that speak to anyone
Insights & Use Cases
By
,
February 25, 2026
Voice agent feature prioritization: What customers actually use (and what they don’t)
Insights & Use Cases
By
Kelsey Foster
,
Growth
February 25, 2026
Best speech-to-text APIs for startups
Insights & Use Cases
By
Kelsey Foster
,
Growth
February 24, 2026
Top 7 meeting intelligence platforms in 2026
Insights & Use Cases
By
Jesse Sumrak
,
Featured writer
February 24, 2026
What is speech recognition? A comprehensive guide
Insights & Use Cases
By
Kelsey Foster
,
Growth
February 24, 2026
How to use Voice AI for healthcare market research
Insights & Use Cases
By
Jesse Sumrak
,
Featured writer
February 24, 2026
Speech-to-Text AI for product managers: How it works and key considerations
Insights & Use Cases
By
Julie Griffin
,
Featured writer
February 24, 2026
Top 3 benefits of Voice AI for revenue Intelligence
Insights & Use Cases
By
Kelsey Foster
,
Growth
February 20, 2026
Top 10 AI notetakers in 2026: Compare features, pricing, and accuracy
Insights & Use Cases
By
Kelsey Foster
,
Growth
February 17, 2026
Top text-to-speech APIs in 2026
Insights & Use Cases
By
Kelsey Foster
,
Growth
February 17, 2026
Best medical speech recognition software and APIs in 2026
Insights & Use Cases
By
Kelsey Foster
,
Growth
February 17, 2026
Speaker diarization: Speaker labels for mono channel files
Insights & Use Cases
By
Joe Zaghloul
,
February 17, 2026
How to Use Speech to Text AI for Ad Targeting and Brand Protection
Insights & Use Cases
By
Jesse Sumrak
,
Featured writer
February 12, 2026
AssemblyAI Universal-3 Pro vs Google Gemini: Speech-to-text API vs multimodal audio processing
Insights & Use Cases
By
Martin Schweiger
,
Senior Technical Product Marketing Manager
February 11, 2026
Healthcare voice agents: Complete implementation guide
Insights & Use Cases
By
Kelsey Foster
,
Growth
February 11, 2026
Building a medical scribe startup in 2026
Insights & Use Cases
By
Kelsey Foster
,
Growth
February 11, 2026
Latest trends and tools in medical transcription services
Insights & Use Cases
By
Kelsey Foster
,
Growth
February 11, 2026
The best 7 ambient AI scribes
Insights & Use Cases
By
Kelsey Foster
,
Growth
February 11, 2026
Top tools for live transcription
Insights & Use Cases
By
Kelsey Foster
,
Growth
February 11, 2026
Voice AI in 2026: Inside the companies and investments shaping the future of speech
News
By
Kelsey Foster
,
Growth
February 10, 2026
Top 8 speaker diarization libraries and APIs in 2026
Insights & Use Cases
By
Kelsey Foster
,
Growth
February 5, 2026
How to use AssemblyAI with Java
Insights & Use Cases
By
Martin Schweiger
,
Senior Technical Product Marketing Manager
February 5, 2026
How to use AssemblyAI with C#
Insights & Use Cases
By
Martin Schweiger
,
Senior Technical Product Marketing Manager
February 4, 2026
Inside AssemblyAI's NYC voice agents January 2026 meetup: Production insights from the front lines
News
By
Kelsey Foster
,
Growth
February 4, 2026
AssemblyAI Universal-3-Pro vs ElevenLabs Scribe v2 Compared
Insights & Use Cases
By
Martin Schweiger
,
Senior Technical Product Marketing Manager
February 3, 2026
Prompt engineering for Universal-3 Pro: A practical guide
Insights & Use Cases
By
Martin Schweiger
,
Senior Technical Product Marketing Manager
February 3, 2026
Introducing Universal-3 Pro: A new class of speech language model optimized for Voice AI
Releases & Updates
By
Madison Bernstein
,
Product Marketing
January 27, 2026
What is an Ambient AI Scribe and how do they work?
Insights & Use Cases
By
Kelsey Foster
,
Growth
January 27, 2026
The voice AI stack for building agents in 2026
Insights & Use Cases
By
Kelsey Foster
,
Growth
January 27, 2026
AI call centers: How AI voice agents are transforming contact centers
Insights & Use Cases
By
Jesse Sumrak
,
Featured writer
January 27, 2026
Biggest challenges in building AI voice agents (and how AssemblyAI & Vapi are solving them)
Insights & Use Cases
By
Smitha Kolan
,
Developer Educator
January 22, 2026
Make vs Zapier: Which platform for Voice AI workflows?
Insights & Use Cases
By
Griffin Sharp
,
Applied AI Engineer
January 22, 2026
New 2026 insights report: What actually makes a good voice agent
News
By
Kelsey Foster
,
Growth
January 22, 2026
n8n vs Postman: Which platform for Voice AI workflows?
Insights & Use Cases
By
Griffin Sharp
,
Applied AI Engineer
January 20, 2026
AI notetakers beyond transcription: How leading companies turn meetings into measurable business value
Insights & Use Cases
By
Kelsey Foster
,
Growth
January 20, 2026
8 best revenue intelligence platforms using AI in 2026
Insights & Use Cases
By
Jesse Sumrak
,
Featured writer
January 20, 2026
10 speech-to-text use cases to inspire your applications
Insights & Use Cases
By
Jesse Sumrak
,
Featured writer
January 14, 2026
Best real-time speech-to-text apps in 2026
Insights & Use Cases
By
Kelsey Foster
,
Growth
January 12, 2026
Top AI models for conversation intelligence
Insights & Use Cases
By
Kelsey Foster
,
Growth
January 7, 2026
How to build an AI medical scribe with AssemblyAI
Insights & Use Cases
By
Martin Schweiger
,
Senior Technical Product Marketing Manager
January 6, 2026
Real-time transcription that code-switches for multilingual speakers
Insights & Use Cases
By
Meredith Rauch
,
Growth
January 6, 2026
6 best orchestration tools to build AI voice agents in 2026
Insights & Use Cases
By
Jesse Sumrak
,
Featured writer
January 6, 2026
Best APIs for Sentiment Analysis in 2026
Insights & Use Cases
By
Kelsey Foster
,
Growth
December 16, 2025
Optimizing Voice AI costs: When to switch STT providers and what to expect
Insights & Use Cases
By
Kelsey Foster
,
Growth
December 16, 2025
Medical terminology accuracy: Techniques for domain-specific transcription
Insights & Use Cases
By
Kelsey Foster
,
Growth
December 16, 2025
What kinds of businesses use automatic transcription?
Insights & Use Cases
By
Kelsey Foster
,
Growth
December 16, 2025
The 300ms rule: Why latency makes or breaks voice AI applications
Insights & Use Cases
By
Martin Schweiger
,
Senior Technical Product Marketing Manager
December 16, 2025
Python speech recognition in 30 lines of code
Insights & Use Cases
By
Yujian Tang
,
Contributor
December 8, 2025
What is real-time agent assist? How AI transforms live customer support
By
Kelsey Foster
,
Growth
December 8, 2025
Automatically summarize audio and video files at scale with AI summarization
Insights & Use Cases
By
Kelsey Foster
,
Growth
December 8, 2025
Automatic speech-to-text punctuation, casing, and ITN to boost transcript readability
Insights & Use Cases
By
Kelsey Foster
,
Growth
December 8, 2025
7 best conversation intelligence software in 2026
Insights & Use Cases
By
Jesse Sumrak
,
Featured writer
December 5, 2025
New guide: Evaluating Voice AI for Ambient AI Scribes in healthcare
Releases & Updates
By
Kelsey Foster
,
Growth
December 2, 2025
How to remove or reduce background noise from audio for (stt) transcription
Insights & Use Cases
By
Kelsey Foster
,
Growth
December 2, 2025
How do I transcribe audio in languages like Spanish, French, or German?
By
Kelsey Foster
,
Growth
%20influence%20automatic%20speaker%20labeling_.png)
December 2, 2025
How does context (like names spoken) influence automatic speaker labeling?
Insights & Use Cases
By
Kelsey Foster
,
Growth
December 1, 2025
Is Word Error Rate Useful?
Insights & Use Cases
By
Dylan Fox
,
Founder, CEO
December 1, 2025
7 LLM use cases and applications in 2026
Insights & Use Cases
By
Jesse Sumrak
,
Featured writer
November 25, 2025
AssemblyAI ping pong tournament 2025: When NYC's tech community came to compete
By
,
November 25, 2025
Nov. Voice AI meetup recap: Real-world challenges of deploying voice AI agents
News
By
Maxinne Rillo
,
Senior Field & Campaign Marketing Manager
November 25, 2025
AI-powered call analytics: How to extract insights from customer conversations
Insights & Use Cases
By
Kelsey Foster
,
Growth
November 25, 2025
What Is media monitoring? (Definition, Benefits, and AI)
Insights & Use Cases
By
Jesse Sumrak
,
Featured writer
November 25, 2025
What is conversational intelligence AI?
Insights & Use Cases
By
Kelsey Foster
,
Growth
November 24, 2025
Speaker identification and diarization with AssemblyAI
Insights & Use Cases
By
Kelsey Foster
,
Growth
November 24, 2025
AI in customer service: Top use cases for 2026
Insights & Use Cases
By
Kelsey Foster
,
Growth
.png)
November 21, 2025
Why evals in voice AI are so hard (and how to fix them)
Insights & Use Cases
By
Ryan Seams
,
VP, Customer Solutions
November 20, 2025
Gemini 3 Pro vs GPT-5 vs Claude 4.5: Which model wins for audio workflows?
Insights & Use Cases
By
Meredith Rauch
,
Growth
November 18, 2025
Using multichannel and speaker diarization
Insights & Use Cases
By
Patrick Loeber
,
Senior Developer Advocate
November 18, 2025
How to use Google's Speech-to-Text API to transcribe audio in Python
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
November 18, 2025
Speech-to-text API accuracy for phone call transcription
Insights & Use Cases
By
Kelsey Foster
,
Growth
November 12, 2025
Voice agents in healthcare: Automating phone interactions for scheduling, billing, and more
Insights & Use Cases
By
Kelsey Foster
,
Growth
November 12, 2025
Introducing Multilingual Universal-Streaming: Go global with ultra-fast, ultra-accurate real-time speech-to-text
Releases & Updates
By
Madison Bernstein
,
Product Marketing
November 12, 2025
Speaker Diarization: Adding speaker labels for enterprise speech-to-text
Insights & Use Cases
By
Kelsey Foster
,
Growth
November 10, 2025
Python Speech-to-Text with Punctuation, Casing, and Formatting
Insights & Use Cases
By
Matt Makai
,
November 10, 2025
Transcribe a phone call in real-time using Python with AssemblyAI and Twilio
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
November 4, 2025
Top sales coaching software in 2025
By
Kelsey Foster
,
Growth
November 4, 2025
Troubleshooting the AssemblyAI API: The importance of retrying requests after server or upload errors
Insights & Use Cases
By
Michelle Asuamah
,
Senior API Support Engineer
November 4, 2025
Real-time transcription in Python with Universal-Streaming
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
.png)
November 3, 2025
AssemblyAI's October 2025 releases: Multilingual streaming, guardrails, and LLM gateway
Releases & Updates
By
Kelsey Foster
,
Growth
November 3, 2025
How Voice AI technology can improve transcription services
Insights & Use Cases
By
Jesse Sumrak
,
Featured writer
November 3, 2025
Speech AI use cases for Learning Management Systems
Insights & Use Cases
By
Amanda Smith
,
October 31, 2025
AI trends in 2025: Graph Neural Networks
Insights & Use Cases
By
Marco Ramponi
,
October 30, 2025
Summarize audio with LLMs in Node.js
Insights & Use Cases
By
Niels Swimberghe
,
October 29, 2025
Build a real-time medical transcription analysis app with AssemblyAI and LLM Gateway
Insights & Use Cases
By
Kelsey Foster
,
Growth
October 29, 2025
Speech Understanding tasks explained: Speaker ID, custom formatting, and translation
Releases & Updates
By
Kelsey Foster
,
Growth
October 29, 2025
What our customers shipped in October 2025
Releases & Updates
By
Kelsey Foster
,
Growth
October 28, 2025
How to summarize meetings with LLMs
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
October 28, 2025
Extract phone call insights with LLMs in Python
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
October 28, 2025
Convert Speech to Text in Python in 5 Minutes
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
October 28, 2025
Transcribe and generate subtitles for YouTube videos with Node.js
Insights & Use Cases
By
Niels Swimberghe
,
October 27, 2025
One platform, multiple models: Simplifying Voice AI with LLM Gateway
Releases & Updates
By
Kelsey Foster
,
Growth
October 27, 2025
Analyze Audio from Zoom Calls with AssemblyAI and Node.js
Insights & Use Cases
By
David Ekete
,
October 27, 2025
Build an AI-powered video conferencing app with Next.js and Stream
Insights & Use Cases
By
Stefan Blos
,
Developer Advocate at Stream
%20is%20Being%20Used%20Today.png)
October 27, 2025
10 ways streaming speech-to-text (live transcription) is being used today
Insights & Use Cases
By
Jesse Sumrak
,
Featured writer
October 27, 2025
Business use cases for Generative AI
Insights & Use Cases
By
Amanda Smith
,
October 27, 2025
Detect scam calls using Go with LLM Gateway and Twilio
Insights & Use Cases
By
Marcus Olsson
,
Senior Developer Educator
October 27, 2025
Automatic summarization with LLMs in Python
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
October 23, 2025
Build voice AI apps with LLM Gateway
Insights & Use Cases
By
Kelsey Foster
,
Growth
October 22, 2025
Introducing new products and model updates to help you build, deploy, and scale Voice AI applications
Releases & Updates
By
Madison Bernstein
,
Product Marketing
%20audio%20with%20timestamps%20for%20captions%20with%20AssemblyAI.png)
October 22, 2025
How to transcribe (stt) audio with timestamps for captions with AssemblyAI
Insights & Use Cases
By
Kelsey Foster
,
Growth
October 22, 2025
Video transcription made simple: From segments to timestamps
Insights & Use Cases
By
Kelsey Foster
,
Growth
October 22, 2025
JavaScript and Node.js Speech-to-Text
Insights & Use Cases
By
,
October 21, 2025
18 Ways Businesses are Launching New Products with Voice AI
By
,
October 21, 2025
How to use AI to automatically summarize meeting transcripts
Insights & Use Cases
By
Kelsey Foster
,
Growth
October 16, 2025
Speech recognition in the browser using Web Speech API
Insights & Use Cases
By
Patrick Loeber
,
Senior Developer Advocate
October 15, 2025
5 Amazon Transcribe alternatives in 2025
Insights & Use Cases
By
Kelsey Foster
,
Growth
October 15, 2025
What is Automatic Speech Recognition? A Comprehensive Overview of ASR Technology
Insights & Use Cases
By
Kelsey Foster
,
Growth
October 1, 2025
9 best AI subtitle generators for 2025
Insights & Use Cases
By
Kelsey Foster
,
Growth
September 30, 2025
How to convert an MP3 file to text with an API
Insights & Use Cases
By
Yujian Tang
,
Contributor
September 22, 2025
Voice agents take center stage: Highlights from the SF Voice Agent Hackathon
Insights & Use Cases
By
Devon Malloy
,
Staff Growth Manager
September 17, 2025
Speech-to-text AI: A complete guide to modern speech recognition technology
Insights & Use Cases
By
Kelsey Foster
,
Growth
September 17, 2025
Real-time speech recognition with Python
Insights & Use Cases
By
Yujian Tang
,
Contributor
September 16, 2025
AssemblyAI Named Leader in G2's Fall 2025 Voice Recognition Grid® Report
News
By
Devon Malloy
,
Staff Growth Manager
September 11, 2025
Introducing Keyterms Prompting to Streaming STT: Never miss the words that matter most
Releases & Updates
By
Madison Bernstein
,
Product Marketing
September 10, 2025
Speech AI for sales intelligence platforms: How to use AI in 2025
Insights & Use Cases
By
Jesse Sumrak
,
Featured writer
September 9, 2025
Introducing the In-App Playground: Test Speech-to-Text instantly—no code required
Releases & Updates
By
Devon Malloy
,
Staff Growth Manager
August 29, 2025
SF Hackathon + In-App Playground + 99 Language Support | August 29, 2025 Newsletter
Newsletter
By
Devon Malloy
,
Staff Growth Manager
August 28, 2025
How intelligent turn detection (endpointing) solves the biggest challenge in voice agent development
Insights & Use Cases
By
Martin Schweiger
,
Senior Technical Product Marketing Manager
August 27, 2025
The complete guide to speaker diarization APIs and tools
Insights & Use Cases
By
Kelsey Foster
,
Growth
August 26, 2025
How does real-time agent assist work? An implementation guide
Insights & Use Cases
By
Kelsey Foster
,
Growth
August 26, 2025
Now Available: 99 Languages, Advanced Features, One Price
Releases & Updates
By
Madison Bernstein
,
Product Marketing
August 20, 2025
The conversation intelligence value machine: How AI transforms every customer interaction
Insights & Use Cases
By
Kelsey Foster
,
Growth
August 14, 2025
How to perform speaker diarization in JavaScript
Insights & Use Cases
By
Kelsey Foster
,
Growth
August 14, 2025
The conversational AI evolution: How agentic systems are rewriting contact center operations
Insights & Use Cases
By
Kelsey Foster
,
Growth
August 14, 2025
Auto-tweet your words using speech recognition in Python
Insights & Use Cases
By
Mısra Turp
,
Developer Educator
.png)
August 11, 2025
Build and deploy real-time AI voice agents using LiveKit and AssemblyAI
Insights & Use Cases
By
Kelsey Foster
,
Growth
August 11, 2025
How to build and deploy a voice agent using Pipecat and AssemblyAI
Insights & Use Cases
By
Kelsey Foster
,
Growth
August 11, 2025
Transcribe phone calls in real-time in Go with Twilio and AssemblyAI
Insights & Use Cases
By
Marcus Olsson
,
Senior Developer Educator
August 7, 2025
These 7 voice AI projects just blew us away
Insights & Use Cases
By
Meredith Rauch
,
Growth
August 7, 2025
Offline speech recognition with Whisper: Browser + Node.js implementations
Insights & Use Cases
By
Tema Bolshakov
,
Contributer
August 7, 2025
How to use Whisper API to transcribe audio in JavaScript
Insights & Use Cases
By
Tema Bolshakov
,
Contributer
August 7, 2025
How to automatically transcribe Zoom calls in real-time with Recall.ai and AssemblyAI
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
August 7, 2025
Build a real-time AI voice bot using Python, AssemblyAI, and ElevenLabs
Insights & Use Cases
By
Smitha Kolan
,
Developer Educator
August 6, 2025
Python speech recognition in 2025
Insights & Use Cases
By
Yujian Tang
,
Contributor
August 1, 2025
Streaming STT Performance Update | August 1, 2025 Newsletter
Newsletter
By
,
July 31, 2025
How to do hotword detection with Universal-Streaming Speech-to-Text and Go
Insights & Use Cases
By
Yasoob Khalid
,
Featured writer
July 29, 2025
Easy C# Speech Recognition
Insights & Use Cases
By
Yujian Tang
,
Contributor
July 29, 2025
Transcribe audio and video files with Python and Universal
Insights & Use Cases
By
Patrick Loeber
,
Senior Developer Advocate
July 28, 2025
Universal model improvements: Introducing advanced contextual text formatting for Spanish and German
Releases & Updates
By
Madison Bernstein
,
Product Marketing
July 18, 2025
AI in Sales Calls: Ways Voice AI helps sales teams win more deals
Insights & Use Cases
By
Kelsey Foster
,
Growth
July 17, 2025
Enhanced diarization + Dovetail case study + G2 wins | July 18, 2025 Newsletter
Newsletter
By
,
July 16, 2025
G2's Summer 2025 Voice Recognition Reports: AssemblyAI receives top rankings across key categories
Releases & Updates
By
Devon Malloy
,
Staff Growth Manager
July 16, 2025
Introducing our most accurate Speaker Diarization yet—30% better in noisy, overlapping audio
Releases & Updates
By
Madison Bernstein
,
Product Marketing
July 15, 2025
How to Get YouTube Video Transcripts
Insights & Use Cases
By
Patrick Loeber
,
Senior Developer Advocate
July 15, 2025
Transcribe Twilio Phone Calls in Real-Time with AssemblyAI
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
July 14, 2025
How to build the lowest latency voice agent in Vapi: Achieving ~465ms end-to-end Latency
Insights & Use Cases
By
Daniel Ince
,
Product
July 11, 2025
Build an AI Voice Agent with DeepSeek R1, AssemblyAI, and ElevenLabs
Insights & Use Cases
By
Smitha Kolan
,
Developer Educator
July 10, 2025
Claude 4 models now available through our LeMUR API
Releases & Updates
By
Madison Bernstein
,
Product Marketing
July 9, 2025
OpenAI Whisper for developers: Choosing between API, local, or server-side transcription
Insights & Use Cases
By
Tema Bolshakov
,
Contributer
July 9, 2025
How to convert voice to text in real time using JavaScript
Insights & Use Cases
By
Patrick Loeber
,
Senior Developer Advocate
.png)
July 7, 2025
Real-time Speech Recognition with AssemblyAI
Insights & Use Cases
By
Mısra Turp
,
Developer Educator
July 4, 2025
29 questions to ask when building AI voice agents
Insights & Use Cases
By
Jesse Sumrak
,
Featured writer
.png)
July 2, 2025
The ongoing need for human-in-the-loop in conversation intelligence
Video
By
,
June 30, 2025
Build your first AI voice agent: 3 step-by-step examples
Insights & Use Cases
By
Kelsey Foster
,
Growth
June 24, 2025
Expanding Access: Slam-1 and LeMUR Now Available in the EU
Releases & Updates
By
Madison Bernstein
,
Product Marketing
June 16, 2025
New 2025 Insights Report: The State of Conversation Intelligence
Releases & Updates
By
Kelsey Foster
,
Growth
June 5, 2025
How to build a LiveKit AI Agent for real-time Speech-to-Text
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
June 2, 2025
Introducing Universal-Streaming: Ultra-Fast, Ultra-Accurate Speech-to-Text for Voice Agents
Releases & Updates
By
JD Prater
,
Head of Product Marketing
April 28, 2025
How to build an MCP voice agent with OpenAI and LiveKit Agents
Insights & Use Cases
By
Juan Luis Ruiz-Tagle
,
Contributor
%20Blog%20-%20Slam-1.png)
April 23, 2025
Slam-1 now in public beta: the most powerful prompt-based Speech Language Model to unlock real world outcomes
Releases & Updates
By
JD Prater
,
Head of Product Marketing
April 22, 2025
Model Context Protocol (MCP) - What it is, how it works, and why it matters
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
April 15, 2025
Introducing the new AssemblyAI Developer Hub: Everything in one place
Releases & Updates
By
Martin Schweiger
,
Senior Technical Product Marketing Manager
April 7, 2025
Predicting the future: What top AI founders have to say about innovation in 2025
Video
By
Kelsey Foster
,
Growth
April 7, 2025
Conversation intelligence in contact centers
Insights & Use Cases
By
Jesse Sumrak
,
Featured writer
April 7, 2025
Expanding our strategic partnership with AWS
Releases & Updates
By
Amy Deora
,
Head of Partnerships
March 31, 2025
Announcing dashboard revamp & multiple API keys
Releases & Updates
By
Ryan O'Connor
,
Senior Developer Educator
March 18, 2025
AssemblyAI named to Fast Company’s list of Most Innovative Companies for 2025
Releases & Updates
By
Dylan Fox
,
Founder, CEO
.webp)
March 5, 2025
Raising the bar for Speech AI: Announcing a first of its kind Speech Language Model and improved Streaming model
Releases & Updates
By
Dylan Fox
,
Founder, CEO
March 5, 2025
Monitor your SpeechAI app with OpenLIT
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
February 27, 2025
Building with AI in 2025: Top advice from leading founders
Video
By
Kelsey Foster
,
Growth
February 25, 2025
Google Cloud's Future of AI: Perspectives for Startups, featuring AssemblyAI
News
By
Dylan Fox
,
Founder, CEO
February 25, 2025
Modern Generative AI for images
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
February 25, 2025
New AI Models to summarize audio and video for any use case
Releases & Updates
By
Ryan O'Connor
,
Senior Developer Educator
February 21, 2025
Summarize meetings in 5 minutes with Python
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
February 20, 2025
Universal speech-to-text model leads in English, German, and Spanish
Releases & Updates
By
Ryan O'Connor
,
Senior Developer Educator
February 6, 2025
AI product strategy in 2025: Top advice from AI-first founders
Video
By
Kelsey Foster
,
Growth
February 4, 2025
Golden Gemini: A new approach in Voice AI
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
.webp)
February 3, 2025
Expanding Enterprise Security and Data Residency Capabilities
Releases & Updates
By
Madison Bernstein
,
Product Marketing
January 29, 2025
Building AI-first products, technology readiness, and a deep sense of curiosity
Video
By
,
January 23, 2025
Enterprise conversation intelligence: The power of superior Voice AI
Insights & Use Cases
By
Jesse Sumrak
,
Featured writer
January 23, 2025
Python Speech Recognition in 2025
Insights & Use Cases
By
Patrick Loeber
,
Senior Developer Advocate
January 9, 2025
Dev.to x AssemblyAI: Winter Speech-to-Text Challenge Winners
News
By
Kelsey Foster
,
Growth
January 9, 2025
Top 6 benefits of integrating LLMs for Conversation Intelligence platforms
By
Kelsey Foster
,
Growth
December 19, 2024
What is voice intelligence and how does it work?
By
,
December 18, 2024
Announcing the AssemblyAI integration for LiveKit
Releases & Updates
By
Ryan O'Connor
,
Senior Developer Educator
December 11, 2024
Top Voice AI projects and winners at 2024 AssemblyAI Hackathon
Insights & Use Cases
By
Whitney DeGraaf
,
Program Manager
December 4, 2024
Market timing, partnering with the right AI providers, and building your own competitive moat
Video
By
,
November 25, 2024
How to transcribe Zoom participant recordings (multichannel)
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
November 25, 2024
Voice content moderation with AI: Everything you need to know
Insights & Use Cases
By
Jesse Sumrak
,
Featured writer
November 21, 2024
Build or buy? What industry leaders are choosing
Insights & Use Cases
By
Chelsea Weber
,
November 15, 2024
Talk to ChatGPT on a Phone Call
Insights & Use Cases
By
Artem Oppermann
,
Featured writer
November 11, 2024
Universal in Action: Transforming Conversational Data Across Industries
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
November 7, 2024
The race to AI integration
Insights & Use Cases
By
Chelsea Weber
,
November 7, 2024
Universal-2 vs OpenAI's Whisper: Comparing Speech-to-Text models in real-world use cases
Insights & Use Cases
By
Patrick Loeber
,
Senior Developer Advocate
November 5, 2024
Auto-generate subtitles with Python and AssemblyAI
Insights & Use Cases
By
Marcus Olsson
,
Senior Developer Educator
October 31, 2024
Beyond Word Error Rate: Universal-2 Delivers Accuracy Where It Matters
Insights & Use Cases
By
JD Prater
,
Head of Product Marketing
October 23, 2024
New 2024 Insights Report: How AI is shaping product strategy
Releases & Updates
By
Chelsea Weber
,
October 22, 2024
How to build a free Whisper API with GPU backend
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
October 17, 2024
Building community and unlocking AI innovation
Video
By
Kelsey Foster
,
Growth
October 17, 2024
7 no-code and low-code ways to build AI-powered Speech-to-Text tools
Insights & Use Cases
By
Kelsey Foster
,
Growth
October 11, 2024
Introducing the AssemblyAI integration for Langflow
Releases & Updates
By
Patrick Loeber
,
Senior Developer Advocate
October 7, 2024
Put Voice AI on the roadmap
Insights & Use Cases
By
Jesse Sumrak
,
Featured writer
October 3, 2024
AI-powered meeting company Supernormal launches customizable Voice Agents
Insights & Use Cases
By
Kelsey Foster
,
Growth
October 2, 2024
Taking risks, disrupting categories, and building value for customers with AI
Video
By
,
September 27, 2024
Speech-to-Text with Django
Insights & Use Cases
By
Patrick Loeber
,
Senior Developer Advocate
September 26, 2024
Introducing the Postman collection for AssemblyAI
Releases & Updates
By
Niels Swimberghe
,
September 24, 2024
Speech recognition with Ruby using Universal-1
Insights & Use Cases
By
Niels Swimberghe
,
September 19, 2024
Building AI startups, crafting product strategy, and earning customer trust
Video
By
,
September 19, 2024
Introducing the AssemblyAI piece for Activepieces
Releases & Updates
By
Niels Swimberghe
,
September 13, 2024
Build Powerful Speech AI Apps with AssemblyAI & Speaker Diarization Tutorials
Newsletter
By
Smitha Kolan
,
Developer Educator
September 12, 2024
How to identify languages in audio data using Python
Insights & Use Cases
By
Patrick Loeber
,
Senior Developer Advocate
September 11, 2024
Voice AI apps: 8 new Voice AI tools, releases, updates, and more
Insights & Use Cases
By
Kelsey Foster
,
Growth
September 10, 2024
How to perform Speaker Diarization in Python
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
September 9, 2024
Speaker diarization vs speaker recognition - what's the difference?
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
September 6, 2024
AssemblyAI's C# .NET SDK + Latest Tutorials
Newsletter
By
Smitha Kolan
,
Developer Educator
September 5, 2024
Build a Discord Voice Bot to Add ChatGPT to Your Voice Channel
Insights & Use Cases
By
Michael Nyamande
,
September 3, 2024
Introducing the AssemblyAI C# .NET SDK
Releases & Updates
By
Niels Swimberghe
,
August 30, 2024
🚀 Upgraded Automatic Language Detection + Latest Tutorials
Newsletter
By
Smitha Kolan
,
Developer Educator
August 26, 2024
Automatic language detection improvements: increased accuracy & expanded language support
Releases & Updates
By
JD Prater
,
Head of Product Marketing
August 23, 2024
Build with AssemblyAI's Streaming Speech-to-Text + Latest Tutorials
Newsletter
By
Smitha Kolan
,
Developer Educator
August 22, 2024
Conversation intelligence: How to better understand the voice of the customer with Speech AI
Insights & Use Cases
By
Joseph Rendeiro
,
August 21, 2024
Decoding Strategies: How LLMs Choose The Next Word
Insights & Use Cases
By
Marco Ramponi
,
August 16, 2024
Build with AssemblyAI's Speaker Diarization Model + Latest Tutorials
Newsletter
By
Smitha Kolan
,
Developer Educator
August 12, 2024
What is Customer Success? The key role of technical customer success and support teams in winning and retaining customers
Insights & Use Cases
By
Jesse Sumrak
,
Featured writer
August 12, 2024
Introducing the AssemblyAI Ruby SDK
Releases & Updates
By
Niels Swimberghe
,
August 9, 2024
New LeMUR Claude 3 Endpoints & Latest Zapier Integration
Newsletter
By
Smitha Kolan
,
Developer Educator
August 6, 2024
Introducing the enhanced AssemblyAI app for Zapier
Releases & Updates
By
Niels Swimberghe
,
August 6, 2024
Generate subtitles with AssemblyAI and Zapier
Insights & Use Cases
By
Niels Swimberghe
,
August 5, 2024
How to evaluate AI models and systems: Why objective benchmarks are important
Insights & Use Cases
By
Kelly Moon
,
August 2, 2024
🎉 AssemblyAI's Python SDK Crosses 100K Monthly Downloads & Latest Tutorials
Newsletter
By
Smitha Kolan
,
Developer Educator
July 25, 2024
🔥 New PII Redaction and Entity Detection Features
Newsletter
By
Smitha Kolan
,
Developer Educator
July 19, 2024
Get started using Claude 3.5 Sonnet with audio data
Insights & Use Cases
By
Patrick Loeber
,
Senior Developer Advocate
July 18, 2024
Announcing New Language Support for PII Text Redaction and Expanding Entity Detection
Releases & Updates
By
JD Prater
,
Head of Product Marketing
July 18, 2024
Speech-to-Text security: Top foundational security questions to consider for your next project using speech
Insights & Use Cases
By
Miki Fukushima
,
July 15, 2024
Florence-2: How it works and how to use it
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
July 15, 2024
Use Claude 3.5 Sonnet With Audio Data & Latest Speech-to-Text Tutorials
Newsletter
By
Smitha Kolan
,
Developer Educator
July 12, 2024
How to Create a Real-Time Language Translation Service with AssemblyAI and DeepL in JavaScript
Insights & Use Cases
By
,
July 10, 2024
Claude 3 Models now available with LeMUR
Releases & Updates
By
JD Prater
,
Head of Product Marketing
July 8, 2024
Create Multi-Lingual Subtitles with AssemblyAI and DeepL
Insights & Use Cases
By
Aniket Bhattacharyea
,
July 8, 2024
Build Powerful Speech AI Apps with AssemblyAI and LLM Integrations
Newsletter
By
Smitha Kolan
,
Developer Educator
June 28, 2024
Get More from Audio Data with Conversational Intelligence
Newsletter
By
Mısra Turp
,
Developer Educator
June 21, 2024
🎙️ Speaker Diarization Now More Accurate & 🔔 Introducing Billing Alerts
Newsletter
By
Mısra Turp
,
Developer Educator
June 20, 2024
Speaker diarization improvements: new languages, increased accuracy
Releases & Updates
By
Ryan O'Connor
,
Senior Developer Educator
June 19, 2024
Announcing the AssemblyAI Starter App for Encore
Releases & Updates
By
Marcus Olsson
,
Senior Developer Educator
June 17, 2024
How to Create WebVTT Files for Videos in Node.js
Insights & Use Cases
By
Niels Swimberghe
,
June 17, 2024
How to Create SRT Files for Videos in Node.js
Insights & Use Cases
By
Niels Swimberghe
,
June 14, 2024
🇩🇪 New German STT & Improved PII Detection Models
Newsletter
By
Smitha Kolan
,
Developer Educator
June 12, 2024
Redact Personally Identifiable Information (PII) from audio with Node.js
Insights & Use Cases
By
Niels Swimberghe
,
June 12, 2024
Lower latency, reduced prices, and our Java SDK release
Newsletter
By
Smitha Kolan
,
Developer Educator
June 7, 2024
Newsletter #39: Build With AssemblyAI's Integrations
Newsletter
By
Smitha Kolan
,
Developer Educator
May 31, 2024
Newsletter #38: Apply LLMs To Voice Data
Newsletter
By
Smitha Kolan
,
Developer Educator
May 31, 2024
How to Transcribe Audio to Text Accurately at Scale
Insights & Use Cases
By
Jesse Sumrak
,
Featured writer
May 30, 2024
Node.js Speech-to-Text with Punctuation, Casing, and Formatting
Insights & Use Cases
By
Niels Swimberghe
,
May 28, 2024
Filter profanity from audio files using Node.js
Insights & Use Cases
By
Niels Swimberghe
,
May 27, 2024
Content moderation on audio files with Python
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
May 24, 2024
Newsletter #37: Speaker Diarization Now in 5 New Languages 🇨🇳🇮🇳🇯🇵🇰🇷🇻🇳 & Latest Speech AI tutorials
Newsletter
By
Smitha Kolan
,
Developer Educator
May 22, 2024
Filter profanity from audio files using Python
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
May 17, 2024
Newsletter #36: Latest Speech-to-text Model Benchmarks
Newsletter
By
Smitha Kolan
,
Developer Educator
May 10, 2024
Newsletter #35: Nano & Best: New Speech-to-text Pricing Options
Newsletter
By
Smitha Kolan
,
Developer Educator
May 8, 2024
Best and Nano Tiers: More Speech-to-Text and Pricing Options
Releases & Updates
By
Kelly Moon
,
May 3, 2024
Newsletter #34: AssemblyAI API Reference & Latest Tutorials
Newsletter
By
Smitha Kolan
,
Developer Educator
April 26, 2024
Newsletter #33: Make.com Voice AI Integration and Streaming STT Updates
Newsletter
By
Smitha Kolan
,
Developer Educator
April 26, 2024
Best Large Language Models (LLMs) & Frameworks in 2024
Insights & Use Cases
By
Jesse Sumrak
,
Featured writer
April 23, 2024
Redact PII in Audio with Make and AssemblyAI
Insights & Use Cases
By
Niels Swimberghe
,
April 23, 2024
Introducing the AssemblyAI app for Make (Integromat)
Releases & Updates
By
Niels Swimberghe
,
April 19, 2024
Newsletter #32:⚡️ Upgrades To Streaming Speech-to-Text
Newsletter
By
Smitha Kolan
,
Developer Educator
April 19, 2024
AssemblyAI + 🔗LangChain Go, Universal-1 Recap
Newsletter
By
Smitha Kolan
,
Developer Educator
April 15, 2024
Transcribe audio with Java using Universal-1
Insights & Use Cases
By
Niels Swimberghe
,
April 12, 2024
Newsletter #30: 🚀 Universal-1 Model Launch
Newsletter
By
Smitha Kolan
,
Developer Educator
April 10, 2024
Introducing the AssemblyAI integration for LangChain Go
Releases & Updates
By
Marcus Olsson
,
Senior Developer Educator
April 8, 2024
9 ways to transform contact center results with AI-powered speech analytics
Insights & Use Cases
By
Jesse Sumrak
,
Featured writer
April 5, 2024
Build Audio LLM Apps with AssemblyAI
Newsletter
By
Smitha Kolan
,
Developer Educator
April 3, 2024
Introducing Universal-1
By
Dylan Fox
,
Founder, CEO
March 21, 2024
Improved Streaming Speech-to-Text Pricing and Features
Newsletter
By
Mısra Turp
,
Developer Educator
March 19, 2024
Real-Time is now Streaming Speech-to-Text, with added customization and control for users
Releases & Updates
By
Kelsey Foster
,
Growth
March 14, 2024
🔥 New Free Video Course from Talk Python: Build an Audio AI App
Newsletter
By
Smitha Kolan
,
Developer Educator
March 13, 2024
A New Free Python Course to Build Real-World Audio AI Apps
Releases & Updates
By
Patrick Loeber
,
Senior Developer Advocate
March 13, 2024
AssemblyAI Go SDK v1.3.0: Utterance Detection and Word Search
Releases & Updates
By
Marcus Olsson
,
Senior Developer Educator
March 13, 2024
AssemblyAI Java SDK New Features & Improvements
Newsletter
By
Smitha Kolan
,
Developer Educator
March 8, 2024
Improved Audio LLM Docs & AssemblyAI Go SDK
Newsletter
By
Smitha Kolan
,
Developer Educator
March 7, 2024
AI tools for business: Top 6 considerations before building with AI models and LLMs
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 4, 2024
How to use AI to build powerful market research tools
Insights & Use Cases
By
Kelsey Foster
,
Growth
February 29, 2024
Top 3 ways to enhance AI video editing tools with Voice AI
Insights & Use Cases
By
Kelsey Foster
,
Growth
February 23, 2024
New Utterance Controls for Real-Time Transcription
Newsletter
By
Smitha Kolan
,
Developer Educator
February 23, 2024
Why product teams at top call tracking solutions are turning to AI
Insights & Use Cases
By
Kelsey Foster
,
Growth
February 17, 2024
PII Redaction and Entity Detection In 13 New Languages 🇫🇷🇩🇪🇮🇳
Newsletter
By
Smitha Kolan
,
Developer Educator
February 9, 2024
Improvements to Real-Time Transcription
Newsletter
By
Smitha Kolan
,
Developer Educator
February 1, 2024
Ask questions about your audio with LLMs
Newsletter
By
Smitha Kolan
,
Developer Educator
January 26, 2024
🚀 New AssemblyAI Go SDK & Speech-to-Text Tutorials
Newsletter
By
Smitha Kolan
,
Developer Educator
January 23, 2024
How to do Speech-To-Text with Go
Insights & Use Cases
By
Patrick Loeber
,
Senior Developer Advocate
January 19, 2024
Claude 2.1 Now Available with LeMUR + New Integrations
Newsletter
By
Smitha Kolan
,
Developer Educator
January 19, 2024
Announcing the AssemblyAI Go SDK
Releases & Updates
By
Marcus Olsson
,
Senior Developer Educator
January 16, 2024
Announcing the AssemblyAI Integration for Haystack
Releases & Updates
By
Mısra Turp
,
Developer Educator
January 10, 2024
Lower latency, lower cost, more possibilities
Releases & Updates
By
Ryan O'Connor
,
Senior Developer Educator
January 8, 2024
Announcing the AssemblyAI integration for Semantic Kernel .NET
Releases & Updates
By
Niels Swimberghe
,
January 8, 2024
Ask .NET Rocks! questions with Semantic Kernel, GPT, and Chroma DB
Insights & Use Cases
By
Niels Swimberghe
,
January 5, 2024
AssemblyAI's New Integrations & Latest Tutorials
Newsletter
By
Smitha Kolan
,
Developer Educator
January 2, 2024
Why Virtual Meeting Companies Should Use Voice AI
Insights & Use Cases
By
Julie Griffin
,
Featured writer
December 20, 2023
2023 at AssemblyAI - A Year in Review
By
Smitha Kolan
,
Developer Educator
December 15, 2023
How to Create VTT Files for Videos in Python
Insights & Use Cases
By
Patrick Loeber
,
Senior Developer Advocate
December 15, 2023
🚀 New Punctuation & Casing Model For Real-Time Transcription
Releases & Updates
By
Smitha Kolan
,
Developer Educator
December 14, 2023
How to Create SRT Files for Videos in Python
Insights & Use Cases
By
Patrick Loeber
,
Senior Developer Advocate
December 8, 2023
🎉 Announcing our $50M Series C to build superhuman Voice AI models
News
By
Smitha Kolan
,
Developer Educator
December 7, 2023
AI for Universal Audio Understanding: Qwen-Audio Explained
Insights & Use Cases
By
Marco Ramponi
,
December 6, 2023
Announcing the AssemblyAI integration for LlamaIndex.TS
Releases & Updates
By
Niels Swimberghe
,
December 6, 2023
How to integrate spoken audio into LlamaIndex.TS using AssemblyAI
Insights & Use Cases
By
,
December 3, 2023
Announcing our $50M Series C to build superhuman Voice AI models
Releases & Updates
By
Dylan Fox
,
Founder, CEO
December 1, 2023
Improved Hold Music Detection + Build LLM Audio Apps with LeMUR
Releases & Updates
By
Smitha Kolan
,
Developer Educator
December 1, 2023
5 Benefits of Voice AI for Video Editing Platforms
Insights & Use Cases
By
Amanda Smith
,
November 30, 2023
6 Ways Telehealth Platforms Can Leverage Speech-to-Text AI
Insights & Use Cases
By
Julie Griffin
,
Featured writer
November 27, 2023
Should I build or buy an AI speech recognition system?
Insights & Use Cases
By
Kelsey Foster
,
Growth
November 24, 2023
🚀 LeMUR's Custom Text Input + Revamped Playground
Newsletter
By
Smitha Kolan
,
Developer Educator
November 20, 2023
AssemblyAI is now on the Amazon Web Services (AWS) Marketplace
Releases & Updates
By
Kelsey Foster
,
Growth
November 16, 2023
Enhancing Our Speech-to-Text Models with Google v5e TPUs and 🎉100K on YouTube
Newsletter
By
Smitha Kolan
,
Developer Educator
November 15, 2023
7 best practices for product teams to consider when building with AI
Insights & Use Cases
By
Kelsey Foster
,
Growth
November 15, 2023
7 best practices for product teams to consider when building with AI
Insights & Use Cases
By
Kelsey Foster
,
Growth
November 8, 2023
Introducing Our New Punctuation Restoration and Truecasing Models
Releases & Updates
By
Marco Ramponi
,
November 7, 2023
Automatically determine video sections with AI using Python
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
November 7, 2023
Improved Punctuation Restoration & Truecasing Models
Newsletter
By
Smitha Kolan
,
Developer Educator
November 2, 2023
Key phrase detection in audio files using Python
Insights & Use Cases
By
Patrick Loeber
,
Senior Developer Advocate
November 1, 2023
Faster Audio File Handling and Improved Error Messages
Newsletter
By
Smitha Kolan
,
Developer Educator
October 30, 2023
Transcribe audio to text on Cloudflare Workers with AssemblyAI and TypeScript
Insights & Use Cases
By
Niels Swimberghe
,
October 27, 2023
How Bluedot built with AssemblyAI to increase user conversion rate
Insights & Use Cases
By
Kelsey Foster
,
Growth
October 27, 2023
Combining Speech Recognition and Diarization in one model
Insights & Use Cases
By
Marco Ramponi
,
October 26, 2023
New Code Cookbooks & AssemblyAI's Q4 Product Enhancements
Newsletter
By
Smitha Kolan
,
Developer Educator
October 19, 2023
New Multilingual Capabilities and TypeScript/JavaScript SDK
Newsletter
By
Smitha Kolan
,
Developer Educator
October 16, 2023
How to use audio data in LlamaIndex with Python
Insights & Use Cases
By
Patrick Loeber
,
Senior Developer Advocate
October 12, 2023
Announcing the AssemblyAI Node SDK 2.0
Releases & Updates
By
Niels Swimberghe
,
October 10, 2023
Building with Automatic Speech Recognition (ASR) models: Why accuracy matters
Insights & Use Cases
By
Kelsey Foster
,
Growth
October 9, 2023
🚀LlamaIndex Integration + Model-Specific Usage Dashboards
Newsletter
By
Smitha Kolan
,
Developer Educator
October 2, 2023
New Usage Dashboard + Mistral 7B First Look
Newsletter
By
Smitha Kolan
,
Developer Educator
September 29, 2023
How DALL-E 2 Actually Works
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
September 29, 2023
8 Ways Automatic Speech Recognition Can Increase Efficiency For Your Business
Insights & Use Cases
By
Julie Griffin
,
Featured writer
September 27, 2023
How to use Voice AI systems for podcast hosting, editing, and monetization
Insights & Use Cases
By
Kelsey Foster
,
Growth
September 26, 2023
Retrieval Augmented Generation on audio data with LangChain and Chroma
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
September 22, 2023
What AI Music Generators Can Do (And How They Do It)
Insights & Use Cases
By
Marco Ramponi
,
September 20, 2023
Announcing the AssemblyAI plugin for Rivet
Releases & Updates
By
Niels Swimberghe
,
September 14, 2023
How to get Zoom Transcripts with the Zoom API
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
September 4, 2023
What is Residual Vector Quantization?
Insights & Use Cases
By
Marco Ramponi
,
August 31, 2023
How to build an interactive lecture summarization app
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
August 31, 2023
How to use audio data in LangChain with Python
Insights & Use Cases
By
Patrick Loeber
,
Senior Developer Advocate
August 22, 2023
RLHF vs RLAIF for language model alignment
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
August 18, 2023
Why Language Models Became Large Language Models And The Hurdles In Developing LLM-based Applications
Insights & Use Cases
By
Marco Ramponi
,
August 15, 2023
How to integrate spoken audio into LangChain.js using AssemblyAI
Insights & Use Cases
By
Niels Swimberghe
,
August 15, 2023
Introducing the AssemblyAI integration for LangChain.js
Releases & Updates
By
Niels Swimberghe
,
August 14, 2023
Customer Stories: Conformer-2 in Action
Insights & Use Cases
By
Kelsey Foster
,
Growth
August 1, 2023
How Reinforcement Learning from AI Feedback works
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
July 27, 2023
LeMUR: Now Available for Early Access
Releases & Updates
By
Kelsey Foster
,
Growth
July 27, 2023
Recent developments in Generative AI for Audio
Insights & Use Cases
By
Marco Ramponi
,
July 19, 2023
Conformer-2: a state-of-the-art speech recognition model trained on 1.1M hours of data
Releases & Updates
By
Francis McCann
,
May 23, 2023
Large Language Models for Product Managers: 5 Things to Know
Insights & Use Cases
By
Marco Ramponi
,
May 23, 2023
How AI helps Marvin's users spend 60% less time analyzing research data
Insights & Use Cases
By
Kelsey Foster
,
Growth
May 17, 2023
Introduction to Large Language Models for Generative AI
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
May 16, 2023
3 ways to build and deploy AI tools and features faster
Insights & Use Cases
By
Kelsey Foster
,
Growth
May 3, 2023
The Full Story of Large Language Models and RLHF
Insights & Use Cases
By
Marco Ramponi
,
May 2, 2023
Everything you need to know about Generative AI
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
May 2, 2023
Introduction to Generative AI
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
April 19, 2023
How physics advanced Generative AI
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
April 3, 2023
How RLHF Preference Model Tuning Works (And How Things May Go Wrong)
Insights & Use Cases
By
Marco Ramponi
,
March 15, 2023
Conformer-1: A robust speech recognition model trained on 650K hours of data
Releases & Updates
By
Marco Ramponi
,
March 13, 2023
3 easy ways to add AI Summarization to Conversation Intelligence tools
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 7, 2023
Emergent Abilities of Large Language Models
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
February 21, 2023
Why every Fortune 500 business needs a chief AI officer
Insights & Use Cases
By
Dylan Fox
,
Founder, CEO
January 19, 2023
Build a free Stable Diffusion app with a GPU backend
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
January 18, 2023
AI research review – Locating and Editing Factual Associations in GPT
Insights & Use Cases
By
Gabriel Oexle
,
December 23, 2022
How ChatGPT actually works
Insights & Use Cases
By
Marco Ramponi
,
December 15, 2022
Winners and Honorable Mentions - AssemblyAI $50k Winter Hackathon
Releases & Updates
By
Ryan O'Connor
,
Senior Developer Educator
December 14, 2022
Releasing our new v9 transcription model - 11% better accuracy
Releases & Updates
By
Ryan O'Connor
,
Senior Developer Educator
December 12, 2022
Build standout call coaching features with AI Summarization
Insights & Use Cases
By
Kelsey Foster
,
Growth
December 6, 2022
Stable Diffusion 1 vs 2 - What you need to know
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
November 30, 2022
Stable Diffusion in Keras - A Simple Tutorial
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
November 22, 2022
DeepMind's AlphaTensor Explained
Insights & Use Cases
By
Marco Ramponi
,
November 16, 2022
AI research review - Merging Models Modulo Permutation Symmetries
Insights & Use Cases
By
Yash Khare
,
Deep Learning Researcher
November 7, 2022
AI for product managers: Today’s top terms to stay in the know
Insights & Use Cases
By
Kelsey Foster
,
Growth
October 26, 2022
An Introduction to Poisson Flow Generative Models
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
October 19, 2022
Transcribe audio or video files right from your terminal
Releases & Updates
By
Francisco Castillo
,
September 28, 2022
AssemblyAI Recognized as G2 High Performer, Momentum Leader for Fall 2022
Releases & Updates
By
Kelsey Foster
,
Growth
September 21, 2022
AI Research Review - Multistream CNN
Insights & Use Cases
By
Luka Chkhetiani
,
Deep Learning Research Lead
September 21, 2022
Getting Started with Hugging Face's Gradio
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
September 14, 2022
Introducing the AssemblyAI Creators Program
Releases & Updates
By
Patrick Loeber
,
Senior Developer Advocate
September 8, 2022
AI Research Review - Spelling and ASR
Insights & Use Cases
By
Taufiquzzaman Peyash
,
Deep Learning Engineer
September 6, 2022
New for Enterprise: Improved Accuracy, Always-on Support, and SOC 2 Type 2
Releases & Updates
By
Micky Teng
,
September 6, 2022
AssemblyAI Obtains SOC 2 Type 2 Compliance for 2022/2023
Releases & Updates
By
Mike Groves
,
September 2, 2022
2022 Benchmark Report
Insights & Use Cases
By
Lee Vaughn
,
API Support Engineer
September 1, 2022
Coming Soon in Fall 2022 at AssemblyAI
Releases & Updates
By
Kelsey Foster
,
Growth
August 24, 2022
Deep Learning Paper Recap - Diffusion and Transformer Models
Insights & Use Cases
By
Dillon Pulliam
,
August 23, 2022
How to Run Stable Diffusion Locally to Generate Images
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
August 17, 2022
Deep Learning Paper Recap - Redundancy Reduction and Sparse MoEs
Insights & Use Cases
By
Domenic Donato
,
August 17, 2022
MinImagen - Build Your Own Imagen Text-to-Image Model
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
August 10, 2022
Deep Learning Paper Recap - Transfer Learning
Insights & Use Cases
By
Michael Liang
,
August 3, 2022
Deep Learning Paper Recap - Automatic Speech Recognition
Insights & Use Cases
By
Taufiquzzaman Peyash
,
Deep Learning Engineer
August 3, 2022
Topic Detection in NLP: The Top APIs for 2023
By
,
July 27, 2022
Deep Learning Paper Recaps - Modality Matching and Masked Autoencoders
Insights & Use Cases
By
Luka Chkhetiani
,
Deep Learning Research Lead
July 20, 2022
AssemblyAI Named G2 Voice Recognition Software Leader in Winter 2022
Releases & Updates
By
Kelsey Foster
,
Growth
July 20, 2022
AssemblyAI Named a G2 High Performer and Momentum Leader for Summer 2022
Releases & Updates
By
Kelsey Foster
,
Growth
July 19, 2022
Creating Top Hiring Intelligence Platforms with AI Models
Insights & Use Cases
By
Kelsey Foster
,
Growth
July 14, 2022
Announcing our $30M Series B
Releases & Updates
By
Dylan Fox
,
Founder, CEO
July 7, 2022
Deep Learning Paper Recap - Language Models
Insights & Use Cases
By
Taufiquzzaman Peyash
,
Deep Learning Engineer
June 23, 2022
How Imagen Actually Works
By
Ryan O'Connor
,
Senior Developer Educator
June 22, 2022
Deep Learning Paper Recap - Streaming ASR and Summarization
Insights & Use Cases
By
Guru Rao
,
June 16, 2022
Review – TOXIGEN & Knowledge Distillation Meets Open-Set Semi-Supervised Learning
Insights & Use Cases
By
Domenic Donato
,
June 15, 2022
Hack with AssemblyAI: HawkHacks 2022
Releases & Updates
By
Kelsey Foster
,
Growth
June 8, 2022
Review - Decision Transformer & SPIRAL
Insights & Use Cases
By
Kevin Zhang
,
June 6, 2022
Getting Started with ESPnet
By
Ryan O'Connor
,
Senior Developer Educator
June 3, 2022
Building Standout Hybrid Event Solutions with AI Models
Insights & Use Cases
By
Kelsey Foster
,
Growth
May 12, 2022
Introduction to Diffusion Models for Machine Learning
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
May 11, 2022
Building an Intelligent Cloud-based Contact Center? How AI Models Can Help
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 18, 2022
Built with AssemblyAI - Real-time Speech-to-Image Generation
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 12, 2022
How to Build a JavaScript Audio Transcript Application
Insights & Use Cases
By
Stefan Rosanitsch
,
Contributor
April 7, 2022
MediaPipe for Dummies
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
April 4, 2022
JavaScript Text-to-Speech - The Easy Way
Insights & Use Cases
By
Stefan Rosanitsch
,
Contributor
March 29, 2022
AssemblyAI Recognized as G2 High Performer, Momentum Leader in Voice Recognition Software for Spring 2022
Releases & Updates
By
Kelsey Foster
,
Growth
March 28, 2022
A Beginner's Guide to TorchStudio, The PyTorch IDE
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
March 20, 2022
React Text to Speech - Simplified!
Insights & Use Cases
By
Stefan Rosanitsch
,
Contributor
March 17, 2022
Automate Meeting Notes with Python
Insights & Use Cases
By
Mısra Turp
,
Developer Educator
March 16, 2022
Review - ALBERT: A Lite BERT for Self-supervised Learning of Language Representations
Insights & Use Cases
By
Sergio Ramirez Martin
,
March 15, 2022
Built with AssemblyAI - YouTube Transcripts
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 15, 2022
Transcribe Audio Files in an S3 Bucket with AssemblyAI
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
March 11, 2022
Kaldi Install for Dummies
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
March 11, 2022
AI Models for Smart Media Monitoring
Insights & Use Cases
By
Kelsey Foster
,
Growth
March 4, 2022
Announcing Our $28M Series A Led by Accel
Releases & Updates
By
Dylan Fox
,
Founder, CEO
March 2, 2022
Differentiable Programming - A Simple Introduction
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
February 25, 2022
BitFit: Simple Parameter-efficient Fine-tuning for Transformer-based Masked Language-models
Insights & Use Cases
By
Taufiquzzaman Peyash
,
Deep Learning Engineer
February 24, 2022
Learn How To Get Started with OpenAI API and GPT-3
Insights & Use Cases
By
Patrick Loeber
,
Senior Developer Advocate
February 21, 2022
What is Gradient Clipping for Neural Networks?
Insights & Use Cases
By
Mısra Turp
,
Developer Educator
February 15, 2022
Why You Should (or Shouldn't) be Using Google's JAX in 2023
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
February 14, 2022
Hyperparameters of Neural Networks
Insights & Use Cases
By
Mısra Turp
,
Developer Educator
February 14, 2022
How to Build a Python Project that Summarizes Your Lectures
Insights & Use Cases
By
Patrick Loeber
,
Senior Developer Advocate
February 8, 2022
What is Layer Normalization?
Insights & Use Cases
By
Mısra Turp
,
Developer Educator
February 4, 2022
Review - Perceiver: General Perception with Iterative Attention
Insights & Use Cases
By
Dillon Pulliam
,
February 2, 2022
Reinforcement Learning With (Deep) Q-Learning Explained
By
Patrick Loeber
,
Senior Developer Advocate
February 2, 2022
Built with AssemblyAI - Rhetoric
Insights & Use Cases
By
Kelsey Foster
,
Growth
February 1, 2022
Machine Learning Concepts for Beginners
Insights & Use Cases
By
Kelsey Foster
,
Growth
January 31, 2022
What is Weight Initialization for Neural Networks?
Insights & Use Cases
By
Mısra Turp
,
Developer Educator
January 26, 2022
Review - data2vec: A General Framework for Self-supervised Learning in Speech, Vision, and Language
Insights & Use Cases
By
Guru Rao
,
January 26, 2022
Unsupervised Machine Learning For Beginners
Insights & Use Cases
By
Patrick Loeber
,
Senior Developer Advocate
January 25, 2022
How to Evaluate Machine Learning Models
Insights & Use Cases
By
Mısra Turp
,
Developer Educator
January 24, 2022
DeltaHacks - AssemblyAI at McMaster University Hackathon
Releases & Updates
By
Britney Xiu
,
January 24, 2022
What is BERT and How Does It Work?
Insights & Use Cases
By
Mısra Turp
,
Developer Educator
January 24, 2022
Supervised Machine Learning For Beginners
Insights & Use Cases
By
Patrick Loeber
,
Senior Developer Advocate
January 20, 2022
Kaldi Speech Recognition for Beginners - A Simple Tutorial
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
January 18, 2022
Recurrent Neural Networks (RNNs) Explained
Insights & Use Cases
By
Patrick Loeber
,
Senior Developer Advocate
January 10, 2022
Bias and Variance for Machine Learning
Insights & Use Cases
By
Mısra Turp
,
Developer Educator
January 6, 2022
Best Speech-to-Text Software
Insights & Use Cases
By
Patrick Loeber
,
Senior Developer Advocate
January 5, 2022
Built with AssemblyAI - Wordcab
Insights & Use Cases
By
Kelsey Foster
,
Growth
January 4, 2022
Backpropagation For Neural Networks Explained
Insights & Use Cases
By
Patrick Loeber
,
Senior Developer Advocate
January 4, 2022
Jupyter Notebooks Tips and Tricks
Insights & Use Cases
By
Mısra Turp
,
Developer Educator
January 3, 2022
Introduction to Variational Autoencoders Using Keras
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
December 29, 2021
Built with AssemblyAI - IronScribe
Insights & Use Cases
By
Kelsey Foster
,
Growth
December 28, 2021
What is GPT-3 and How Does It Work?
Insights & Use Cases
By
Mısra Turp
,
Developer Educator
December 27, 2021
Getting Started With Torchaudio
Insights & Use Cases
By
Patrick Loeber
,
Senior Developer Advocate
December 20, 2021
2021 at AssemblyAI - A Year in Review
Releases & Updates
By
Kelsey Foster
,
Growth
December 20, 2021
2022 at AssemblyAI - A Year in Review
Releases & Updates
By
Kelsey Foster
,
Growth
December 16, 2021
Introducing Sentiment Analysis - Detect Sentiments in Spoken Audio
Releases & Updates
By
Kelsey Foster
,
Growth
December 15, 2021
Review - JUST: Joint Unsupervised and Supervised Training For Multilingual ASR
By
,
December 15, 2021
Auto Chapters in Action - Build a Web App that Automatically Summarizes Podcasts
Insights & Use Cases
By
Patrick Loeber
,
Senior Developer Advocate
December 14, 2021
PyTorch vs TensorFlow in 2023
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
December 13, 2021
Sentiment Analysis in Action - Earnings Calls
Insights & Use Cases
By
Mısra Turp
,
Developer Educator
December 9, 2021
Activation Functions In Neural Networks Explained
Insights & Use Cases
By
Patrick Loeber
,
Senior Developer Advocate
December 7, 2021
Review - VICReg: Variance-Invariance-Covariance Regularization for Self-Supervised Learning
Insights & Use Cases
By
Kevin Zhang
,
December 6, 2021
PyTorch Lightning for Dummies - A Tutorial and Overview
Insights & Use Cases
By
Ryan O'Connor
,
Senior Developer Educator
December 1, 2021
Introducing Entity Detection - Detect Named Entities in Audio/Video
Releases & Updates
By
Kelsey Foster
,
Growth
November 30, 2021
Add Speech Recognition to Applications in 5 Minutes
Insights & Use Cases
By
Mısra Turp
,
Developer Educator
November 30, 2021
Quick Automatic Chapter Detection
Insights & Use Cases
By
Patrick Loeber
,
Senior Developer Advocate
November 30, 2021
Transformers for Beginners - An Introduction
Insights & Use Cases
By
Mısra Turp
,
Developer Educator
November 23, 2021
Review - SimCLS and RefSum - Summarization Techniques
Insights & Use Cases
By
Dillon Pulliam
,
November 22, 2021
What is Regularization? Overfitting and Neural Networks
Insights & Use Cases
By
Mısra Turp
,
Developer Educator
November 18, 2021
What is Sentiment Analysis?
Insights & Use Cases
By
Kelsey Foster
,
Growth
November 17, 2021
Podcasts About AI - Our Deep Learning Team’s Top Picks
Releases & Updates
By
Kelsey Foster
,
Growth
November 17, 2021
Review - Speech processing Universal Performance Benchmark Review
Insights & Use Cases
By
Guru Rao
,
November 16, 2021
Text Segmentation - Approaches, Datasets, and Evaluation Metrics
Insights & Use Cases
By
Taufiquzzaman Peyash
,
Deep Learning Engineer
November 11, 2021
Introducing Auto Chapters - Summarize Audio and Video Files
Releases & Updates
By
Dylan Fox
,
Founder, CEO
November 8, 2021
Top 7 Data Science Blogs for Data Scientists and Enthusiasts
Insights & Use Cases
By
Kelsey Foster
,
Growth
November 5, 2021
AssemblyAI at HackDuke and HackUMBC
By
,
November 5, 2021
An Overview of Transducer Models for ASR
Insights & Use Cases
By
Michael Nguyen
,
November 5, 2021
Batch Normalization for Neural Networks - How it Works
Insights & Use Cases
By
Mısra Turp
,
Developer Educator
November 2, 2021
Machine Learning Podcasts - The Ultimate Listening Guide
Insights & Use Cases
By
Kelsey Foster
,
Growth
November 1, 2021
How to Make a Web App that Transcribes YouTube Videos with Streamlit
Insights & Use Cases
By
Mısra Turp
,
Developer Educator
October 29, 2021
Data Science Podcasts to Listen to Now
Insights & Use Cases
By
Kelsey Foster
,
Growth
October 29, 2021
Deep Shallow Fusion for RNN-T Personalization
Insights & Use Cases
By
Michael Nguyen
,
October 26, 2021
Top 5 Machine Learning Blogs to Follow
Insights & Use Cases
By
Kelsey Foster
,
Growth
October 25, 2021
Deep Learning in 5 Minutes
Insights & Use Cases
By
Mısra Turp
,
Developer Educator
October 25, 2021
Hack the Valley - AssemblyAI at University of Toronto Hackathon
Releases & Updates
By
Britney Xiu
,
October 20, 2021
Review - A Graph Convolutional Neural Network for Emotion Recognition in Conversation
Insights & Use Cases
By
Shuyun Tang
,
October 19, 2021
Releasing our v8 Transcription Model - 18.72% Better Accuracy
Releases & Updates
By
Dylan Fox
,
Founder, CEO
October 13, 2021
DeepSpeech for Dummies - A Tutorial and Overview
Insights & Use Cases
By
Yujian Tang
,
Contributor
October 13, 2021
Review - Pretraining Representations for Data-Efficient Reinforcement Learning
Insights & Use Cases
By
Kevin Zhang
,
October 11, 2021
New - 8.37% Better Accuracy for Topic Detection and IAB Classification with V4 Update
Releases & Updates
By
Kelsey Foster
,
Growth
October 8, 2021
Top 3 Hackathon Projects Built with AssemblyAI’s Speech-to-Text API
Releases & Updates
By
Britney Xiu
,
October 6, 2021
Improved Accuracy on AssemblyAI’s Real Time Speech-to-Text API
Releases & Updates
By
Yujian Tang
,
Contributor
September 30, 2021
New: Improved Topic Detection and IAB Classification
Releases & Updates
By
Dillon Pulliam
,
September 27, 2021
AssemblyAI Recognized as G2 High Performer, Momentum Leader for Fall 2021
Releases & Updates
By
Kelsey Foster
,
Growth
September 24, 2021
Review - Text-Free Prosody-Aware Generative Spoken Language Modeling
Insights & Use Cases
By
Steven Hillis
,
Featured writer
September 21, 2021
How Well Does AI Transcribe Song Lyrics?
Insights & Use Cases
By
Yujian Tang
,
Contributor
September 3, 2021
8/31/2021 AWS Outage Post-Mortem
Releases & Updates
By
Mitch Anderson
,
September 2, 2021
Comparing Zoom Transcription Accuracy Across Speech-to-Text APIs
Insights & Use Cases
By
Joe Zaghloul
,
September 2, 2021
Can Podcasts Predict the Stock Market?
Insights & Use Cases
By
Yujian Tang
,
Contributor
August 18, 2021
How to Add Subtitles to Mux Videos with Python
Insights & Use Cases
By
Yujian Tang
,
Contributor
August 5, 2021
How to Set Up Twilio Voicemail
Insights & Use Cases
By
Yujian Tang
,
Contributor
August 2, 2021
How to Build a Burner Phone with Voicemail in Python
Insights & Use Cases
By
Yujian Tang
,
Contributor
July 28, 2021
The Definitive Guide to Python Click
Insights & Use Cases
By
Yujian Tang
,
Contributor
July 20, 2021
Improved Real-Time Transcription Speed and Accuracy
Releases & Updates
By
Andrew Galyan-Mann
,
Senior API Support Engineer
July 15, 2021
How to build a YouTube downloader in Python
Insights & Use Cases
By
,
June 15, 2021
Fine-Tuning Transformers for NLP
Insights & Use Cases
By
Dillon Pulliam
,
June 15, 2021
Speech-to-Text Accuracy on Podcasts, News Broadcasts, and Social Media
Insights & Use Cases
By
Joe Zaghloul
,
June 1, 2021
Voice Search Partnership with Algolia
Releases & Updates
By
Joe Zaghloul
,
June 1, 2021
Open Sourcing Drone Deploy ECS
Releases & Updates
By
Mitch Anderson
,
May 13, 2021
A New API Endpoint to Paginate Through Historical Transcripts
Releases & Updates
By
Dylan Fox
,
Founder, CEO
May 1, 2021
Improved WER - Speech-to-Text Accuracy vs. Google, AWS (May 2020 Update)
Releases & Updates
By
Joe Zaghloul
,
April 28, 2021
New! Break Transcripts into Paragraphs and Sentences
Releases & Updates
By
Andrew Galyan-Mann
,
Senior API Support Engineer
April 27, 2021
Open Sourcing our Drone CI/CD CloudWatch Auto Scaler
Releases & Updates
By
Mitch Anderson
,
April 8, 2021
Getting started with HttpClientFactory in C# and .NET 5
Insights & Use Cases
By
Yujian Tang
,
Contributor
April 1, 2021
Content Safety Detection is now GA!
Releases & Updates
By
Dylan Fox
,
Founder, CEO
March 23, 2021
Speech-to-Text with Postman and AssemblyAI
Insights & Use Cases
By
Yujian Tang
,
Contributor
March 12, 2021
Transcribing Local Audio Files with Node.js
Insights & Use Cases
By
Yujian Tang
,
Contributor
March 2, 2021
New Punctuation and Casing Model Released
Releases & Updates
By
Andrew Galyan-Mann
,
Senior API Support Engineer
January 1, 2021
Comparing End-To-End Speech Recognition Architectures in 2021
Insights & Use Cases
By
Michael Nguyen
,
December 1, 2020
Building an End-to-End Speech Recognition Model in PyTorch
Insights & Use Cases
By
Michael Nguyen
,
November 20, 2020
PII Redaction and Accuracy Improvements
Releases & Updates
By
Joe Zaghloul
,
September 15, 2020
AssemblyAI Wins Best Public API 2020
Releases & Updates
By
Joe Zaghloul
,
June 1, 2020
[Webinar] Conversation Intelligence with CallRail
Releases & Updates
By
Joe Zaghloul
,
May 1, 2020
Teaming up with Algolia for Effortless Voice Search
Releases & Updates
By
Joe Zaghloul
,
April 1, 2020
Automated SRT and VTT Video Captions (April 2020 Update)
Releases & Updates
By
Joe Zaghloul
,
March 1, 2020
PII Redaction for Speech-to-Text Transcriptions (March 2020 Update)
Releases & Updates
By
Joe Zaghloul
,
Whisper alternatives
By
,
Why AssemblyAI beats self-hosting Whisper
Insights & Use Cases
By
Kelsey Foster
,
Growth
No results found!
Please try different search parameters.
Featured
April 29, 2026
Introducing our Voice Agent API
Releases & Updates
By
Madison Bernstein
,
Product Marketing
April 8, 2026
How to evaluate speech recognition models
Insights & Use Cases
By
Kelsey Foster
,
Growth
%20benchmark%20might%20be%20lying%20to%20you%20(1).png)
March 24, 2026
Why your word error rate (WER) benchmark might be lying to you
Insights & Use Cases
By
Zackary Klebanoff
,
Applied AI Lead
March 3, 2026
Conversation Intelligence: The complete guide for 2026
Insights & Use Cases
By
,
Releases & updates
[See all\
April 29, 2026
Introducing our Voice Agent API
Releases & Updates
By
Madison Bernstein
,
Product Marketing
April 6, 2026
Voice AI guardrails: Built-in protection for compliance, quality, and cost control
Releases & Updates
By
Kelsey Foster
,
Growth
March 25, 2026
Introducing Medical Mode: Purpose-built accuracy for medical terminology
Releases & Updates
By
Madison Bernstein
,
Product Marketing
March 3, 2026
Universal-3 Pro Streaming: The most accurate real-time transcription model for voice agents
Releases & Updates
By
Madison Bernstein
,
Product Marketing
Insights & use cases
[See all\
April 29, 2026
Create an ambient AI scribe that works during telehealth video calls
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 29, 2026
Voice agents in noisy environments (Drive-Thrus, Contact Centers, Field)
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 29, 2026
When to use Voice Agent API vs. Universal-3 Pro Streaming
Insights & Use Cases
By
Kelsey Foster
,
Growth
April 29, 2026
Voice Agent Orchestrators Compared: Vapi vs Pipecat vs LiveKit with AssemblyAI
Insights & Use Cases
By
Kelsey Foster
,
Growth
News
[See all\
March 23, 2026
AssemblyAI Named a Leader in G2’s Spring 2026 Voice Recognition Report
News
By
Devon Malloy
,
Staff Growth Manager
February 11, 2026
Voice AI in 2026: Inside the companies and investments shaping the future of speech
News
By
Kelsey Foster
,
Growth
February 4, 2026
Inside AssemblyAI's NYC voice agents January 2026 meetup: Production insights from the front lines
News
By
Kelsey Foster
,
Growth
January 22, 2026
New 2026 insights report: What actually makes a good voice agent
News
By
Kelsey Foster
,
Growth
Video
[See all\
March 18, 2026
Real-time conversation intelligence: The shift from post-call analysis to live insights
Video
By
Kelsey Foster
,
Growth
.png)
July 2, 2025
The ongoing need for human-in-the-loop in conversation intelligence
Video
By
,
April 7, 2025
Predicting the future: What top AI founders have to say about innovation in 2025
Video
By
Kelsey Foster
,
Growth
February 27, 2025
Building with AI in 2025: Top advice from leading founders
Video
By
Kelsey Foster
,
Growth
Newsletter
[See all\
August 29, 2025
SF Hackathon + In-App Playground + 99 Language Support | August 29, 2025 Newsletter
Newsletter
By
Devon Malloy
,
Staff Growth Manager
August 1, 2025
Streaming STT Performance Update | August 1, 2025 Newsletter
Newsletter
By
,
July 17, 2025
Enhanced diarization + Dovetail case study + G2 wins | July 18, 2025 Newsletter
Newsletter
By
,
September 13, 2024
Build Powerful Speech AI Apps with AssemblyAI & Speaker Diarization Tutorials
Newsletter
By
Smitha Kolan
,
Developer Educator
Subscribe to AssemblyAI’s newsletter
By clicking “Submit” you agree to our TOS and Privacy Policy
Thank you for subscribing!
Oops! Something went wrong while submitting the form.
Unlock the value of voice data
Build what’s next on the platform powering thousands of the industry’s leading of Voice AI apps.