Introducing our Voice Agent API: The fastest path to a working voice agent  Learn more

Blog

All

Releases & Updates

Insights & Use Cases

News

Video

Newsletter

Thank you! Your submission has been received!

Oops! Something went wrong while submitting the form.

April 29, 2026

Create an ambient AI scribe that works during telehealth video calls

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 29, 2026

Voice agents in noisy environments (Drive-Thrus, Contact Centers, Field)

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 29, 2026

When to use Voice Agent API vs. Universal-3 Pro Streaming

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 29, 2026

Voice Agent Orchestrators Compared: Vapi vs Pipecat vs LiveKit with AssemblyAI

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 29, 2026

What streaming speech to text model is best for voice agents and why?

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 29, 2026

Build a voice agent with a chained STT-LLM-TTS architecture

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 29, 2026

Migrating from OpenAI Realtime API to AssemblyAI Voice Agent API

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 29, 2026

Introducing our Voice Agent API

Releases & Updates

By

Madison Bernstein

,

Product Marketing

April 29, 2026

How to choose the best speech-to-text API for voice agents

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 29, 2026

Top APIs and models for real-time speech recognition and transcription in 2026

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 28, 2026

Real-time vs batch transcription: What's the difference?

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 28, 2026

5 Google Cloud Speech-to-Text alternatives in 2026

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 28, 2026

Transformative use cases of AI in contact centers

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 27, 2026

Noise cancellation with speech-to-text: The pros and cons

Insights & Use Cases

By

David Lange

,

Applied AI Engineer

April 22, 2026

How to create an AI cold-calling agent

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 22, 2026

How to create a phone-based voice agent

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 21, 2026

What's the best medical transcription API?

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 21, 2026

What is the difference between speaker recognition and speaker verification?

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 21, 2026

5 Speechmatics alternatives in 2026

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 21, 2026

Top 8 open source STT options for voice applications in 2026

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 21, 2026

Conversational AI in healthcare: maturity model and 7 use cases

Insights & Use Cases

By

Jesse Sumrak

,

Featured writer

April 21, 2026

How to run OpenAI's Whisper speech recognition model

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

April 21, 2026

Conversation AI: What it is and top use cases

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 20, 2026

Voice AI Meetup recap: How Commure and Ona Health are building for healthcare

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 14, 2026

Word error rate is broken: How to actually evaluate speech-to-text in 2026

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 14, 2026

How to vibe code a voice agent (and why AI always recommends AssemblyAI)

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 14, 2026

Build a voice agent with function calling

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 14, 2026

Build a voice agent with LiveKit

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 13, 2026

Can transcripts be used to generate meeting agendas?

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 13, 2026

Beyond transcription: Combining speech-to-text with AI analysis

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 13, 2026

5 Deepgram alternatives in 2026

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 13, 2026

How accurate is speech-to-text in 2026?

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 13, 2026

Build a call center analytics pipeline in Python with AssemblyAI

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 13, 2026

Best AI playgrounds in 2026

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 13, 2026

Content moderation: What it is, how it works, and the best APIs

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 9, 2026

Why speech-to-text accuracy is the hidden bottleneck in your AI agent pipeline

By

,

April 8, 2026

Tutorial: How to easily build a voice agent with AssemblyAI

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 8, 2026

How to build a voice agent with Python in 5 minutes

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 8, 2026

When to stop self-hosting Whisper (and what you actually gain)

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 8, 2026

Edge cases in transcription: Offline mode, partial audio files and API limits

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 8, 2026

Building with transcripts: Search, indexing, display and downstream Integrations

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 8, 2026

Transcript output guide: SRT, VTT & TXT export formats + live caption sync

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 8, 2026

How to build an AI-Powered interview scoring system with speech-to-text

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 8, 2026

How to evaluate speech recognition models

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 6, 2026

Voice AI guardrails: Built-in protection for compliance, quality, and cost control

Releases & Updates

By

Kelsey Foster

,

Growth

April 6, 2026

Medical voice recognition: How AI solves terminology problems

Insights & Use Cases

By

,

April 6, 2026

AI voice agents: what they are and how they work in 2026

Insights & Use Cases

By

Jesse Sumrak

,

Featured writer

April 6, 2026

How to choose the best speech-to-text API

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 6, 2026

How to automatically redact PII from audio and video files with Python

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

April 6, 2026

The top free speech-to-text APIs, AI models, and open source engines

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 2, 2026

Agora voice agent with AssemblyAI Universal-3 Pro Streaming

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 2, 2026

Node.js voice agent with AssemblyAI Universal-3 Pro Streaming

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 2, 2026

Daily.co voice agent with AssemblyAI Universal-3 Pro Streaming

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 2, 2026

Raw WebSocket voice agent with AssemblyAI Universal-3 Pro Streaming

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 2, 2026

Retell AI + AssemblyAI: custom LLM and post-call analytics

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 2, 2026

Vapi voice agent with AssemblyAI Universal-3 Pro Streaming

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 2, 2026

Twilio phone agent with AssemblyAI Universal-3 Pro Streaming

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 2, 2026

Pipecat voice agent with AssemblyAI Universal-3 Pro Streaming

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 2, 2026

Speech-to-text for HR and recruiting: Interview transcription, screening and scoring

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 2, 2026

How to build a lecture capture system with speaker identification

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 2, 2026

Speech-to-Text for EdTech: Lectures, Captions, Accessibility & Assessments

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 2, 2026

Best Nuance Dragon medical alternatives for clinical documentation

Insights & Use Cases

By

,

April 2, 2026

AssemblyAI vs Rev AI: Accuracy, pricing and features compared

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 2, 2026

Handling transcript errors: Homophones, corrections and AI quality improvement

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 31, 2026

LiveKit voice agent with AssemblyAI Universal-3 Pro Streaming

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 31, 2026

AssemblyAI vs Deepgram for medical transcription

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 31, 2026

What is the best speech to text api to build ai medical ambient scribes?

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 31, 2026

AI medical transcription

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 31, 2026

The best audio file formats for speech-to-text: A guide

Insights & Use Cases

By

Patrick Loeber

,

Senior Developer Advocate

March 26, 2026

Medical transcription that actually works — Beyond generic STT

By

Kelsey Foster

,

Growth

March 26, 2026

What is LLM Gateway?

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 26, 2026

Best medical speech-to-text in 2026

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 26, 2026

Speech-to-text for healthcare developer guide

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 25, 2026

Introducing Medical Mode: Purpose-built accuracy for medical terminology

Releases & Updates

By

Madison Bernstein

,

Product Marketing

.png)

March 24, 2026

Turn detection vs forced endpoints in voice AI: Why getting this wrong tanks your UX

Insights & Use Cases

By

Kelsey Foster

,

Growth

.png)

March 24, 2026

Real-time transcription in Python with Universal-3 Pro Streaming

Insights & Use Cases

By

Kelsey Foster

,

Growth

%20benchmark%20might%20be%20lying%20to%20you%20(1).png)

March 24, 2026

Why your word error rate (WER) benchmark might be lying to you

Insights & Use Cases

By

Zackary Klebanoff

,

Applied AI Lead

.png)

March 23, 2026

How do I transcribe audio in languages like Spanish, French, or German?

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 23, 2026

Are there language-specific models for better accuracy?

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 23, 2026

What metrics can I get from transcribed call center data?

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 23, 2026

Can voice AI recognize the topic or themes of a conversation?

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 23, 2026

AssemblyAI Named a Leader in G2’s Spring 2026 Voice Recognition Report

News

By

Devon Malloy

,

Staff Growth Manager

March 23, 2026

Large-scale audio transcription: Handling hours of content efficiently

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 23, 2026

What is real-time speech to text?

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 23, 2026

Do I need a custom speech recognition model?

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 18, 2026

Speech-to-text API pricing guide: Per-minute, per-hour and feature costs explained

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 18, 2026

How to evaluate and choose the best speech to text API for enterprises

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 18, 2026

10 call center metrics you can extract from transcripts with AI

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 18, 2026

How real-time agent assist is changing conversation intelligence

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 18, 2026

Streaming speaker diarization: How to identify who's speaking in real time

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 18, 2026

Real-time entity extraction from speech: Capturing emails, phone numbers, and addresses in live audio

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 18, 2026

Building a production-ready voice agent: The developer's guide to real-time speech-to-text

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 18, 2026

10 best agent assist software in 2026

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 18, 2026

Real-time conversation intelligence: The shift from post-call analysis to live insights

Video

By

Kelsey Foster

,

Growth

.png)

March 17, 2026

Multilingual streaming with Universal-3 Pro: Native code switching across 6 languages

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 17, 2026

What is audio intelligence or speech understanding?

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 17, 2026

What is speaker diarization and how does it work? (Complete 2026 Guide)

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 6, 2026

Speech-to-text prompting with AssemblyAI Universal-3 Pro

Insights & Use Cases

By

Martin Schweiger

,

Senior Technical Product Marketing Manager

March 4, 2026

How accurate is AI transcription for pharmaceutical drug names?

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 4, 2026

Contact center AI trends for 2026

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 4, 2026

Best scalable voice AI solutions for customer service

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 3, 2026

Universal-3 Pro Streaming: The most accurate real-time transcription model for voice agents

Releases & Updates

By

Madison Bernstein

,

Product Marketing

March 3, 2026

Conversation Intelligence: The complete guide for 2026

Insights & Use Cases

By

,

March 3, 2026

Text Summarization for NLP: 5 Best APIs, AI Models, and AI Summarizers in 2026

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 3, 2026

8 best transcript summarizers powered by AI

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 3, 2026

What is speech to text? The complete guide

Insights & Use Cases

By

Jesse Sumrak

,

Featured writer

March 3, 2026

6 best named entity recognition APIs for entity detection

Insights & Use Cases

By

Kelsey Foster

,

Growth

February 26, 2026

How to get the most out of Universal-3 Pro with prompt engineering

Insights & Use Cases

By

Ryan Seams

,

VP, Customer Solutions

February 26, 2026

AssemblyAI Universal-3 Pro vs Deepgram Nova-3: An honest comparison for developers

Insights & Use Cases

By

Kelsey Foster

,

Growth

February 26, 2026

Multilingual speech recognition in 2026: How Universal-3 Pro handles accents, code-switching, and non-English audio

Insights & Use Cases

By

Kelsey Foster

,

Growth

February 26, 2026

What is speaker fingerprinting for Voice AI

Insights & Use Cases

By

Kelsey Foster

,

Growth

February 25, 2026

How do I build an AI medical scribe using speech-to-text?

Insights & Use Cases

By

Kelsey Foster

,

Growth

February 25, 2026

Multi-language voice agents: Building agents that speak to anyone

Insights & Use Cases

By

,

February 25, 2026

Voice agent feature prioritization: What customers actually use (and what they don’t)

Insights & Use Cases

By

Kelsey Foster

,

Growth

February 25, 2026

Best speech-to-text APIs for startups

Insights & Use Cases

By

Kelsey Foster

,

Growth

February 24, 2026

Top 7 meeting intelligence platforms in 2026

Insights & Use Cases

By

Jesse Sumrak

,

Featured writer

February 24, 2026

What is speech recognition? A comprehensive guide

Insights & Use Cases

By

Kelsey Foster

,

Growth

February 24, 2026

How to use Voice AI for healthcare market research

Insights & Use Cases

By

Jesse Sumrak

,

Featured writer

February 24, 2026

Speech-to-Text AI for product managers: How it works and key considerations

Insights & Use Cases

By

Julie Griffin

,

Featured writer

February 24, 2026

Top 3 benefits of Voice AI for revenue Intelligence

Insights & Use Cases

By

Kelsey Foster

,

Growth

February 20, 2026

Top 10 AI notetakers in 2026: Compare features, pricing, and accuracy

Insights & Use Cases

By

Kelsey Foster

,

Growth

February 17, 2026

Top text-to-speech APIs in 2026

Insights & Use Cases

By

Kelsey Foster

,

Growth

February 17, 2026

Best medical speech recognition software and APIs in 2026

Insights & Use Cases

By

Kelsey Foster

,

Growth

February 17, 2026

Speaker diarization: Speaker labels for mono channel files

Insights & Use Cases

By

Joe Zaghloul

,

February 17, 2026

How to Use Speech to Text AI for Ad Targeting and Brand Protection

Insights & Use Cases

By

Jesse Sumrak

,

Featured writer

February 12, 2026

AssemblyAI Universal-3 Pro vs Google Gemini: Speech-to-text API vs multimodal audio processing

Insights & Use Cases

By

Martin Schweiger

,

Senior Technical Product Marketing Manager

February 11, 2026

Healthcare voice agents: Complete implementation guide

Insights & Use Cases

By

Kelsey Foster

,

Growth

February 11, 2026

Building a medical scribe startup in 2026

Insights & Use Cases

By

Kelsey Foster

,

Growth

February 11, 2026

Latest trends and tools in medical transcription services

Insights & Use Cases

By

Kelsey Foster

,

Growth

February 11, 2026

The best 7 ambient AI scribes

Insights & Use Cases

By

Kelsey Foster

,

Growth

February 11, 2026

Top tools for live transcription

Insights & Use Cases

By

Kelsey Foster

,

Growth

February 11, 2026

Voice AI in 2026: Inside the companies and investments shaping the future of speech

News

By

Kelsey Foster

,

Growth

February 10, 2026

Top 8 speaker diarization libraries and APIs in 2026

Insights & Use Cases

By

Kelsey Foster

,

Growth

February 5, 2026

How to use AssemblyAI with Java

Insights & Use Cases

By

Martin Schweiger

,

Senior Technical Product Marketing Manager

February 5, 2026

How to use AssemblyAI with C#

Insights & Use Cases

By

Martin Schweiger

,

Senior Technical Product Marketing Manager

February 4, 2026

Inside AssemblyAI's NYC voice agents January 2026 meetup: Production insights from the front lines

News

By

Kelsey Foster

,

Growth

February 4, 2026

‍AssemblyAI Universal-3-Pro vs ElevenLabs Scribe v2 Compared

Insights & Use Cases

By

Martin Schweiger

,

Senior Technical Product Marketing Manager

February 3, 2026

Prompt engineering for Universal-3 Pro: A practical guide

Insights & Use Cases

By

Martin Schweiger

,

Senior Technical Product Marketing Manager

February 3, 2026

Introducing Universal-3 Pro: A new class of speech language model optimized for Voice AI

Releases & Updates

By

Madison Bernstein

,

Product Marketing

January 27, 2026

What is an Ambient AI Scribe and how do they work?

Insights & Use Cases

By

Kelsey Foster

,

Growth

January 27, 2026

The voice AI stack for building agents in 2026

Insights & Use Cases

By

Kelsey Foster

,

Growth

January 27, 2026

AI call centers: How AI voice agents are transforming contact centers

Insights & Use Cases

By

Jesse Sumrak

,

Featured writer

January 27, 2026

Biggest challenges in building AI voice agents (and how AssemblyAI & Vapi are solving them)

Insights & Use Cases

By

Smitha Kolan

,

Developer Educator

January 22, 2026

Make vs Zapier: Which platform for Voice AI workflows?

Insights & Use Cases

By

Griffin Sharp

,

Applied AI Engineer

January 22, 2026

New 2026 insights report: What actually makes a good voice agent

News

By

Kelsey Foster

,

Growth

January 22, 2026

n8n vs Postman: Which platform for Voice AI workflows?

Insights & Use Cases

By

Griffin Sharp

,

Applied AI Engineer

January 20, 2026

AI notetakers beyond transcription: How leading companies turn meetings into measurable business value

Insights & Use Cases

By

Kelsey Foster

,

Growth

January 20, 2026

8 best revenue intelligence platforms using AI in 2026

Insights & Use Cases

By

Jesse Sumrak

,

Featured writer

January 20, 2026

10 speech-to-text use cases to inspire your applications

Insights & Use Cases

By

Jesse Sumrak

,

Featured writer

January 14, 2026

Best real-time speech-to-text apps in 2026

Insights & Use Cases

By

Kelsey Foster

,

Growth

January 12, 2026

Top AI models for conversation intelligence

Insights & Use Cases

By

Kelsey Foster

,

Growth

January 7, 2026

How to build an AI medical scribe with AssemblyAI

Insights & Use Cases

By

Martin Schweiger

,

Senior Technical Product Marketing Manager

January 6, 2026

Real-time transcription that code-switches for multilingual speakers

Insights & Use Cases

By

Meredith Rauch

,

Growth

January 6, 2026

6 best orchestration tools to build AI voice agents in 2026

Insights & Use Cases

By

Jesse Sumrak

,

Featured writer

January 6, 2026

Best APIs for Sentiment Analysis in 2026

Insights & Use Cases

By

Kelsey Foster

,

Growth

December 16, 2025

Optimizing Voice AI costs: When to switch STT providers and what to expect

Insights & Use Cases

By

Kelsey Foster

,

Growth

December 16, 2025

Medical terminology accuracy: Techniques for domain-specific transcription

Insights & Use Cases

By

Kelsey Foster

,

Growth

December 16, 2025

What kinds of businesses use automatic transcription?

Insights & Use Cases

By

Kelsey Foster

,

Growth

December 16, 2025

The 300ms rule: Why latency makes or breaks voice AI applications

Insights & Use Cases

By

Martin Schweiger

,

Senior Technical Product Marketing Manager

December 16, 2025

Python speech recognition in 30 lines of code

Insights & Use Cases

By

Yujian Tang

,

Contributor

December 8, 2025

What is real-time agent assist? How AI transforms live customer support

By

Kelsey Foster

,

Growth

December 8, 2025

Automatically summarize audio and video files at scale with AI summarization

Insights & Use Cases

By

Kelsey Foster

,

Growth

December 8, 2025

Automatic speech-to-text punctuation, casing, and ITN to boost transcript readability

Insights & Use Cases

By

Kelsey Foster

,

Growth

December 8, 2025

7 best conversation intelligence software in 2026

Insights & Use Cases

By

Jesse Sumrak

,

Featured writer

December 5, 2025

New guide: Evaluating Voice AI for Ambient AI Scribes in healthcare

Releases & Updates

By

Kelsey Foster

,

Growth

December 2, 2025

How to remove or reduce background noise from audio for (stt) transcription

Insights & Use Cases

By

Kelsey Foster

,

Growth

December 2, 2025

How do I transcribe audio in languages like Spanish, French, or German?

By

Kelsey Foster

,

Growth

%20influence%20automatic%20speaker%20labeling_.png)

December 2, 2025

How does context (like names spoken) influence automatic speaker labeling?

Insights & Use Cases

By

Kelsey Foster

,

Growth

December 1, 2025

Is Word Error Rate Useful?

Insights & Use Cases

By

Dylan Fox

,

Founder, CEO

December 1, 2025

7 LLM use cases and applications in 2026

Insights & Use Cases

By

Jesse Sumrak

,

Featured writer

November 25, 2025

AssemblyAI ping pong tournament 2025: When NYC's tech community came to compete

By

,

November 25, 2025

Nov. Voice AI meetup recap: Real-world challenges of deploying voice AI agents

News

By

Maxinne Rillo

,

Senior Field & Campaign Marketing Manager

November 25, 2025

AI-powered call analytics: How to extract insights from customer conversations

Insights & Use Cases

By

Kelsey Foster

,

Growth

November 25, 2025

What Is media monitoring? (Definition, Benefits, and AI)

Insights & Use Cases

By

Jesse Sumrak

,

Featured writer

November 25, 2025

What is conversational intelligence AI?

Insights & Use Cases

By

Kelsey Foster

,

Growth

November 24, 2025

Speaker identification and diarization with AssemblyAI

Insights & Use Cases

By

Kelsey Foster

,

Growth

November 24, 2025

AI in customer service: Top use cases for 2026

Insights & Use Cases

By

Kelsey Foster

,

Growth

.png)

November 21, 2025

Why evals in voice AI are so hard (and how to fix them)

Insights & Use Cases

By

Ryan Seams

,

VP, Customer Solutions

November 20, 2025

Gemini 3 Pro vs GPT-5 vs Claude 4.5: Which model wins for audio workflows?

Insights & Use Cases

By

Meredith Rauch

,

Growth

November 18, 2025

Using multichannel and speaker diarization

Insights & Use Cases

By

Patrick Loeber

,

Senior Developer Advocate

November 18, 2025

How to use Google's Speech-to-Text API to transcribe audio in Python

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

November 18, 2025

Speech-to-text API accuracy for phone call transcription

Insights & Use Cases

By

Kelsey Foster

,

Growth

November 12, 2025

Voice agents in healthcare: Automating phone interactions for scheduling, billing, and more

Insights & Use Cases

By

Kelsey Foster

,

Growth

November 12, 2025

Introducing Multilingual Universal-Streaming: Go global with ultra-fast, ultra-accurate real-time speech-to-text

Releases & Updates

By

Madison Bernstein

,

Product Marketing

November 12, 2025

Speaker Diarization: Adding speaker labels for enterprise speech-to-text

Insights & Use Cases

By

Kelsey Foster

,

Growth

November 10, 2025

Python Speech-to-Text with Punctuation, Casing, and Formatting

Insights & Use Cases

By

Matt Makai

,

November 10, 2025

Transcribe a phone call in real-time using Python with AssemblyAI and Twilio

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

November 4, 2025

Top sales coaching software in 2025

By

Kelsey Foster

,

Growth

November 4, 2025

Troubleshooting the AssemblyAI API: The importance of retrying requests after server or upload errors

Insights & Use Cases

By

Michelle Asuamah

,

Senior API Support Engineer

November 4, 2025

Real-time transcription in Python with Universal-Streaming

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

.png)

November 3, 2025

AssemblyAI's October 2025 releases: Multilingual streaming, guardrails, and LLM gateway

Releases & Updates

By

Kelsey Foster

,

Growth

November 3, 2025

How Voice AI technology can improve transcription services

Insights & Use Cases

By

Jesse Sumrak

,

Featured writer

November 3, 2025

Speech AI use cases for Learning Management Systems

Insights & Use Cases

By

Amanda Smith

,

October 31, 2025

AI trends in 2025: Graph Neural Networks

Insights & Use Cases

By

Marco Ramponi

,

October 30, 2025

Summarize audio with LLMs in Node.js

Insights & Use Cases

By

Niels Swimberghe

,

October 29, 2025

Build a real-time medical transcription analysis app with AssemblyAI and LLM Gateway

Insights & Use Cases

By

Kelsey Foster

,

Growth

October 29, 2025

Speech Understanding tasks explained: Speaker ID, custom formatting, and translation

Releases & Updates

By

Kelsey Foster

,

Growth

October 29, 2025

What our customers shipped in October 2025

Releases & Updates

By

Kelsey Foster

,

Growth

October 28, 2025

How to summarize meetings with LLMs

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

October 28, 2025

Extract phone call insights with LLMs in Python

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

October 28, 2025

Convert Speech to Text in Python in 5 Minutes

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

October 28, 2025

Transcribe and generate subtitles for YouTube videos with Node.js

Insights & Use Cases

By

Niels Swimberghe

,

October 27, 2025

One platform, multiple models: Simplifying Voice AI with LLM Gateway

Releases & Updates

By

Kelsey Foster

,

Growth

October 27, 2025

Analyze Audio from Zoom Calls with AssemblyAI and Node.js

Insights & Use Cases

By

David Ekete

,

October 27, 2025

Build an AI-powered video conferencing app with Next.js and Stream

Insights & Use Cases

By

Stefan Blos

,

Developer Advocate at Stream

%20is%20Being%20Used%20Today.png)

October 27, 2025

10 ways streaming speech-to-text (live transcription) is being used today

Insights & Use Cases

By

Jesse Sumrak

,

Featured writer

October 27, 2025

Business use cases for Generative AI

Insights & Use Cases

By

Amanda Smith

,

October 27, 2025

Detect scam calls using Go with LLM Gateway and Twilio

Insights & Use Cases

By

Marcus Olsson

,

Senior Developer Educator

October 27, 2025

Automatic summarization with LLMs in Python

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

October 23, 2025

Build voice AI apps with LLM Gateway

Insights & Use Cases

By

Kelsey Foster

,

Growth

October 22, 2025

Introducing new products and model updates to help you build, deploy, and scale Voice AI applications

Releases & Updates

By

Madison Bernstein

,

Product Marketing

%20audio%20with%20timestamps%20for%20captions%20with%20AssemblyAI.png)

October 22, 2025

How to transcribe (stt) audio with timestamps for captions with AssemblyAI

Insights & Use Cases

By

Kelsey Foster

,

Growth

October 22, 2025

Video transcription made simple: From segments to timestamps

Insights & Use Cases

By

Kelsey Foster

,

Growth

October 22, 2025

JavaScript and Node.js Speech-to-Text

Insights & Use Cases

By

,

October 21, 2025

18 Ways Businesses are Launching New Products with Voice AI

By

,

October 21, 2025

How to use AI to automatically summarize meeting transcripts

Insights & Use Cases

By

Kelsey Foster

,

Growth

October 16, 2025

Speech recognition in the browser using Web Speech API

Insights & Use Cases

By

Patrick Loeber

,

Senior Developer Advocate

October 15, 2025

5 Amazon Transcribe alternatives in 2025

Insights & Use Cases

By

Kelsey Foster

,

Growth

October 15, 2025

What is Automatic Speech Recognition? A Comprehensive Overview of ASR Technology

Insights & Use Cases

By

Kelsey Foster

,

Growth

October 1, 2025

9 best AI subtitle generators for 2025

Insights & Use Cases

By

Kelsey Foster

,

Growth

September 30, 2025

How to convert an MP3 file to text with an API

Insights & Use Cases

By

Yujian Tang

,

Contributor

September 22, 2025

Voice agents take center stage: Highlights from the SF Voice Agent Hackathon

Insights & Use Cases

By

Devon Malloy

,

Staff Growth Manager

September 17, 2025

Speech-to-text AI: A complete guide to modern speech recognition technology

Insights & Use Cases

By

Kelsey Foster

,

Growth

September 17, 2025

Real-time speech recognition with Python

Insights & Use Cases

By

Yujian Tang

,

Contributor

September 16, 2025

AssemblyAI Named Leader in G2's Fall 2025 Voice Recognition Grid® Report

News

By

Devon Malloy

,

Staff Growth Manager

September 11, 2025

Introducing Keyterms Prompting to Streaming STT: Never miss the words that matter most

Releases & Updates

By

Madison Bernstein

,

Product Marketing

September 10, 2025

Speech AI for sales intelligence platforms: How to use AI in 2025

Insights & Use Cases

By

Jesse Sumrak

,

Featured writer

September 9, 2025

Introducing the In-App Playground: Test Speech-to-Text instantly—no code required

Releases & Updates

By

Devon Malloy

,

Staff Growth Manager

August 29, 2025

SF Hackathon + In-App Playground + 99 Language Support | August 29, 2025 Newsletter

Newsletter

By

Devon Malloy

,

Staff Growth Manager

August 28, 2025

How intelligent turn detection (endpointing) solves the biggest challenge in voice agent development

Insights & Use Cases

By

Martin Schweiger

,

Senior Technical Product Marketing Manager

August 27, 2025

The complete guide to speaker diarization APIs and tools

Insights & Use Cases

By

Kelsey Foster

,

Growth

August 26, 2025

How does real-time agent assist work? An implementation guide

Insights & Use Cases

By

Kelsey Foster

,

Growth

August 26, 2025

Now Available: 99 Languages, Advanced Features, One Price

Releases & Updates

By

Madison Bernstein

,

Product Marketing

August 20, 2025

The conversation intelligence value machine: How AI transforms every customer interaction

Insights & Use Cases

By

Kelsey Foster

,

Growth

August 14, 2025

How to perform speaker diarization in JavaScript

Insights & Use Cases

By

Kelsey Foster

,

Growth

August 14, 2025

The conversational AI evolution: How agentic systems are rewriting contact center operations

Insights & Use Cases

By

Kelsey Foster

,

Growth

August 14, 2025

Auto-tweet your words using speech recognition in Python

Insights & Use Cases

By

Mısra Turp

,

Developer Educator

.png)

August 11, 2025

Build and deploy real-time AI voice agents using LiveKit and AssemblyAI

Insights & Use Cases

By

Kelsey Foster

,

Growth

August 11, 2025

How to build and deploy a voice agent using Pipecat and AssemblyAI

Insights & Use Cases

By

Kelsey Foster

,

Growth

August 11, 2025

Transcribe phone calls in real-time in Go with Twilio and AssemblyAI

Insights & Use Cases

By

Marcus Olsson

,

Senior Developer Educator

August 7, 2025

These 7 voice AI projects just blew us away

Insights & Use Cases

By

Meredith Rauch

,

Growth

August 7, 2025

Offline speech recognition with Whisper: Browser + Node.js implementations

Insights & Use Cases

By

Tema Bolshakov

,

Contributer

August 7, 2025

How to use Whisper API to transcribe audio in JavaScript

Insights & Use Cases

By

Tema Bolshakov

,

Contributer

August 7, 2025

How to automatically transcribe Zoom calls in real-time with Recall.ai and AssemblyAI

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

August 7, 2025

Build a real-time AI voice bot using Python, AssemblyAI, and ElevenLabs

Insights & Use Cases

By

Smitha Kolan

,

Developer Educator

August 6, 2025

Python speech recognition in 2025

Insights & Use Cases

By

Yujian Tang

,

Contributor

August 1, 2025

Streaming STT Performance Update | August 1, 2025 Newsletter

Newsletter

By

,

July 31, 2025

How to do hotword detection with Universal-Streaming Speech-to-Text and Go

Insights & Use Cases

By

Yasoob Khalid

,

Featured writer

July 29, 2025

Easy C# Speech Recognition

Insights & Use Cases

By

Yujian Tang

,

Contributor

July 29, 2025

Transcribe audio and video files with Python and Universal

Insights & Use Cases

By

Patrick Loeber

,

Senior Developer Advocate

July 28, 2025

Universal model improvements: Introducing advanced contextual text formatting for Spanish and German

Releases & Updates

By

Madison Bernstein

,

Product Marketing

July 18, 2025

AI in Sales Calls: Ways Voice AI helps sales teams win more deals

Insights & Use Cases

By

Kelsey Foster

,

Growth

July 17, 2025

Enhanced diarization + Dovetail case study + G2 wins | July 18, 2025 Newsletter

Newsletter

By

,

July 16, 2025

G2's Summer 2025 Voice Recognition Reports: AssemblyAI receives top rankings across key categories

Releases & Updates

By

Devon Malloy

,

Staff Growth Manager

July 16, 2025

Introducing our most accurate Speaker Diarization yet—30% better in noisy, overlapping audio

Releases & Updates

By

Madison Bernstein

,

Product Marketing

July 15, 2025

How to Get YouTube Video Transcripts

Insights & Use Cases

By

Patrick Loeber

,

Senior Developer Advocate

July 15, 2025

Transcribe Twilio Phone Calls in Real-Time with AssemblyAI

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

July 14, 2025

How to build the lowest latency voice agent in Vapi: Achieving ~465ms end-to-end Latency

Insights & Use Cases

By

Daniel Ince

,

Product

July 11, 2025

Build an AI Voice Agent with DeepSeek R1, AssemblyAI, and ElevenLabs

Insights & Use Cases

By

Smitha Kolan

,

Developer Educator

July 10, 2025

Claude 4 models now available through our LeMUR API

Releases & Updates

By

Madison Bernstein

,

Product Marketing

July 9, 2025

OpenAI Whisper for developers: Choosing between API, local, or server-side transcription

Insights & Use Cases

By

Tema Bolshakov

,

Contributer

July 9, 2025

How to convert voice to text in real time using JavaScript

Insights & Use Cases

By

Patrick Loeber

,

Senior Developer Advocate

.png)

July 7, 2025

Real-time Speech Recognition with AssemblyAI

Insights & Use Cases

By

Mısra Turp

,

Developer Educator

July 4, 2025

29 questions to ask when building AI voice agents

Insights & Use Cases

By

Jesse Sumrak

,

Featured writer

.png)

July 2, 2025

The ongoing need for human-in-the-loop in conversation intelligence

Video

By

,

June 30, 2025

Build your first AI voice agent: 3 step-by-step examples

Insights & Use Cases

By

Kelsey Foster

,

Growth

June 24, 2025

Expanding Access: Slam-1 and LeMUR Now Available in the EU

Releases & Updates

By

Madison Bernstein

,

Product Marketing

June 16, 2025

New 2025 Insights Report: The State of Conversation Intelligence

Releases & Updates

By

Kelsey Foster

,

Growth

June 5, 2025

How to build a LiveKit AI Agent for real-time Speech-to-Text

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

June 2, 2025

Introducing Universal-Streaming: Ultra-Fast, Ultra-Accurate Speech-to-Text for Voice Agents

Releases & Updates

By

JD Prater

,

Head of Product Marketing

April 28, 2025

How to build an MCP voice agent with OpenAI and LiveKit Agents

Insights & Use Cases

By

Juan Luis Ruiz-Tagle

,

Contributor

%20Blog%20-%20Slam-1.png)

April 23, 2025

Slam-1 now in public beta: the most powerful prompt-based Speech Language Model to unlock real world outcomes

Releases & Updates

By

JD Prater

,

Head of Product Marketing

April 22, 2025

Model Context Protocol (MCP) - What it is, how it works, and why it matters

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

April 15, 2025

Introducing the new AssemblyAI Developer Hub: Everything in one place

Releases & Updates

By

Martin Schweiger

,

Senior Technical Product Marketing Manager

April 7, 2025

Predicting the future: What top AI founders have to say about innovation in 2025

Video

By

Kelsey Foster

,

Growth

April 7, 2025

Conversation intelligence in contact centers

Insights & Use Cases

By

Jesse Sumrak

,

Featured writer

April 7, 2025

Expanding our strategic partnership with AWS

Releases & Updates

By

Amy Deora

,

Head of Partnerships

March 31, 2025

Announcing dashboard revamp & multiple API keys

Releases & Updates

By

Ryan O'Connor

,

Senior Developer Educator

March 18, 2025

AssemblyAI named to Fast Company’s list of Most Innovative Companies for 2025

Releases & Updates

By

Dylan Fox

,

Founder, CEO

.webp)

March 5, 2025

Raising the bar for Speech AI: Announcing a first of its kind Speech Language Model and improved Streaming model

Releases & Updates

By

Dylan Fox

,

Founder, CEO

March 5, 2025

Monitor your SpeechAI app with OpenLIT

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

February 27, 2025

Building with AI in 2025: Top advice from leading founders

Video

By

Kelsey Foster

,

Growth

February 25, 2025

Google Cloud's Future of AI: Perspectives for Startups, featuring AssemblyAI

News

By

Dylan Fox

,

Founder, CEO

February 25, 2025

Modern Generative AI for images

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

February 25, 2025

New AI Models to summarize audio and video for any use case

Releases & Updates

By

Ryan O'Connor

,

Senior Developer Educator

February 21, 2025

Summarize meetings in 5 minutes with Python

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

February 20, 2025

Universal speech-to-text model leads in English, German, and Spanish

Releases & Updates

By

Ryan O'Connor

,

Senior Developer Educator

February 6, 2025

AI product strategy in 2025: Top advice from AI-first founders

Video

By

Kelsey Foster

,

Growth

February 4, 2025

Golden Gemini: A new approach in Voice AI

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

.webp)

February 3, 2025

Expanding Enterprise Security and Data Residency Capabilities

Releases & Updates

By

Madison Bernstein

,

Product Marketing

January 29, 2025

Building AI-first products, technology readiness, and a deep sense of curiosity

Video

By

,

January 23, 2025

Enterprise conversation intelligence: The power of superior Voice AI

Insights & Use Cases

By

Jesse Sumrak

,

Featured writer

January 23, 2025

Python Speech Recognition in 2025

Insights & Use Cases

By

Patrick Loeber

,

Senior Developer Advocate

January 9, 2025

Dev.to x AssemblyAI: Winter Speech-to-Text Challenge Winners

News

By

Kelsey Foster

,

Growth

January 9, 2025

Top 6 benefits of integrating LLMs for Conversation Intelligence platforms

By

Kelsey Foster

,

Growth

December 19, 2024

What is voice intelligence and how does it work?

By

,

December 18, 2024

Announcing the AssemblyAI integration for LiveKit

Releases & Updates

By

Ryan O'Connor

,

Senior Developer Educator

December 11, 2024

Top Voice AI projects and winners at 2024 AssemblyAI Hackathon

Insights & Use Cases

By

Whitney DeGraaf

,

Program Manager

December 4, 2024

Market timing, partnering with the right AI providers, and building your own competitive moat

Video

By

,

November 25, 2024

How to transcribe Zoom participant recordings (multichannel)

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

November 25, 2024

Voice content moderation with AI: Everything you need to know

Insights & Use Cases

By

Jesse Sumrak

,

Featured writer

November 21, 2024

Build or buy? What industry leaders are choosing

Insights & Use Cases

By

Chelsea Weber

,

November 15, 2024

Talk to ChatGPT on a Phone Call

Insights & Use Cases

By

Artem Oppermann

,

Featured writer

November 11, 2024

Universal in Action: Transforming Conversational Data Across Industries

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

November 7, 2024

The race to AI integration

Insights & Use Cases

By

Chelsea Weber

,

November 7, 2024

Universal-2 vs OpenAI's Whisper: Comparing Speech-to-Text models in real-world use cases

Insights & Use Cases

By

Patrick Loeber

,

Senior Developer Advocate

November 5, 2024

Auto-generate subtitles with Python and AssemblyAI

Insights & Use Cases

By

Marcus Olsson

,

Senior Developer Educator

October 31, 2024

Beyond Word Error Rate: Universal-2 Delivers Accuracy Where It Matters

Insights & Use Cases

By

JD Prater

,

Head of Product Marketing

October 23, 2024

New 2024 Insights Report: How AI is shaping product strategy

Releases & Updates

By

Chelsea Weber

,

October 22, 2024

How to build a free Whisper API with GPU backend

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

October 17, 2024

Building community and unlocking AI innovation

Video

By

Kelsey Foster

,

Growth

October 17, 2024

7 no-code and low-code ways to build AI-powered Speech-to-Text tools

Insights & Use Cases

By

Kelsey Foster

,

Growth

October 11, 2024

Introducing the AssemblyAI integration for Langflow

Releases & Updates

By

Patrick Loeber

,

Senior Developer Advocate

October 7, 2024

Put Voice AI on the roadmap

Insights & Use Cases

By

Jesse Sumrak

,

Featured writer

October 3, 2024

AI-powered meeting company Supernormal launches customizable Voice Agents

Insights & Use Cases

By

Kelsey Foster

,

Growth

October 2, 2024

Taking risks, disrupting categories, and building value for customers with AI

Video

By

,

September 27, 2024

Speech-to-Text with Django

Insights & Use Cases

By

Patrick Loeber

,

Senior Developer Advocate

September 26, 2024

Introducing the Postman collection for AssemblyAI

Releases & Updates

By

Niels Swimberghe

,

September 24, 2024

Speech recognition with Ruby using Universal-1

Insights & Use Cases

By

Niels Swimberghe

,

September 19, 2024

Building AI startups, crafting product strategy, and earning customer trust

Video

By

,

September 19, 2024

Introducing the AssemblyAI piece for Activepieces

Releases & Updates

By

Niels Swimberghe

,

September 13, 2024

Build Powerful Speech AI Apps with AssemblyAI & Speaker Diarization Tutorials

Newsletter

By

Smitha Kolan

,

Developer Educator

September 12, 2024

How to identify languages in audio data using Python

Insights & Use Cases

By

Patrick Loeber

,

Senior Developer Advocate

September 11, 2024

Voice AI apps: 8 new Voice AI tools, releases, updates, and more

Insights & Use Cases

By

Kelsey Foster

,

Growth

September 10, 2024

How to perform Speaker Diarization in Python

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

September 9, 2024

Speaker diarization vs speaker recognition - what's the difference?

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

September 6, 2024

AssemblyAI's C# .NET SDK + Latest Tutorials

Newsletter

By

Smitha Kolan

,

Developer Educator

September 5, 2024

Build a Discord Voice Bot to Add ChatGPT to Your Voice Channel

Insights & Use Cases

By

Michael Nyamande

,

September 3, 2024

Introducing the AssemblyAI C# .NET SDK

Releases & Updates

By

Niels Swimberghe

,

August 30, 2024

🚀 Upgraded Automatic Language Detection + Latest Tutorials

Newsletter

By

Smitha Kolan

,

Developer Educator

August 26, 2024

Automatic language detection improvements: increased accuracy & expanded language support

Releases & Updates

By

JD Prater

,

Head of Product Marketing

August 23, 2024

Build with AssemblyAI's Streaming Speech-to-Text + Latest Tutorials

Newsletter

By

Smitha Kolan

,

Developer Educator

August 22, 2024

Conversation intelligence: How to better understand the voice of the customer with Speech AI

Insights & Use Cases

By

Joseph Rendeiro

,

August 21, 2024

Decoding Strategies: How LLMs Choose The Next Word

Insights & Use Cases

By

Marco Ramponi

,

August 16, 2024

Build with AssemblyAI's Speaker Diarization Model + Latest Tutorials

Newsletter

By

Smitha Kolan

,

Developer Educator

August 12, 2024

What is Customer Success? The key role of technical customer success and support teams in winning and retaining customers

Insights & Use Cases

By

Jesse Sumrak

,

Featured writer

August 12, 2024

Introducing the AssemblyAI Ruby SDK

Releases & Updates

By

Niels Swimberghe

,

August 9, 2024

New LeMUR Claude 3 Endpoints & Latest Zapier Integration

Newsletter

By

Smitha Kolan

,

Developer Educator

August 6, 2024

Introducing the enhanced AssemblyAI app for Zapier

Releases & Updates

By

Niels Swimberghe

,

August 6, 2024

Generate subtitles with AssemblyAI and Zapier

Insights & Use Cases

By

Niels Swimberghe

,

August 5, 2024

How to evaluate AI models and systems: Why objective benchmarks are important

Insights & Use Cases

By

Kelly Moon

,

August 2, 2024

🎉 AssemblyAI's Python SDK Crosses 100K Monthly Downloads & Latest Tutorials

Newsletter

By

Smitha Kolan

,

Developer Educator

July 25, 2024

🔥 New PII Redaction and Entity Detection Features

Newsletter

By

Smitha Kolan

,

Developer Educator

July 19, 2024

Get started using Claude 3.5 Sonnet with audio data

Insights & Use Cases

By

Patrick Loeber

,

Senior Developer Advocate

July 18, 2024

Announcing New Language Support for PII Text Redaction and Expanding Entity Detection

Releases & Updates

By

JD Prater

,

Head of Product Marketing

July 18, 2024

Speech-to-Text security: Top foundational security questions to consider for your next project using speech

Insights & Use Cases

By

Miki Fukushima

,

July 15, 2024

Florence-2: How it works and how to use it

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

July 15, 2024

Use Claude 3.5 Sonnet With Audio Data & Latest Speech-to-Text Tutorials

Newsletter

By

Smitha Kolan

,

Developer Educator

July 12, 2024

How to Create a Real-Time Language Translation Service with AssemblyAI and DeepL in JavaScript

Insights & Use Cases

By

,

July 10, 2024

Claude 3 Models now available with LeMUR

Releases & Updates

By

JD Prater

,

Head of Product Marketing

July 8, 2024

Create Multi-Lingual Subtitles with AssemblyAI and DeepL

Insights & Use Cases

By

Aniket Bhattacharyea

,

July 8, 2024

Build Powerful Speech AI Apps with AssemblyAI and LLM Integrations

Newsletter

By

Smitha Kolan

,

Developer Educator

June 28, 2024

Get More from Audio Data with Conversational Intelligence

Newsletter

By

Mısra Turp

,

Developer Educator

June 21, 2024

🎙️ Speaker Diarization Now More Accurate & 🔔 Introducing Billing Alerts

Newsletter

By

Mısra Turp

,

Developer Educator

June 20, 2024

Speaker diarization improvements: new languages, increased accuracy

Releases & Updates

By

Ryan O'Connor

,

Senior Developer Educator

June 19, 2024

Announcing the AssemblyAI Starter App for Encore

Releases & Updates

By

Marcus Olsson

,

Senior Developer Educator

June 17, 2024

How to Create WebVTT Files for Videos in Node.js

Insights & Use Cases

By

Niels Swimberghe

,

June 17, 2024

How to Create SRT Files for Videos in Node.js

Insights & Use Cases

By

Niels Swimberghe

,

June 14, 2024

🇩🇪 New German STT & Improved PII Detection Models

Newsletter

By

Smitha Kolan

,

Developer Educator

June 12, 2024

Redact Personally Identifiable Information (PII) from audio with Node.js

Insights & Use Cases

By

Niels Swimberghe

,

June 12, 2024

Lower latency, reduced prices, and our Java SDK release

Newsletter

By

Smitha Kolan

,

Developer Educator

June 7, 2024

Newsletter #39: Build With AssemblyAI's Integrations

Newsletter

By

Smitha Kolan

,

Developer Educator

May 31, 2024

Newsletter #38: Apply LLMs To Voice Data

Newsletter

By

Smitha Kolan

,

Developer Educator

May 31, 2024

How to Transcribe Audio to Text Accurately at Scale

Insights & Use Cases

By

Jesse Sumrak

,

Featured writer

May 30, 2024

Node.js Speech-to-Text with Punctuation, Casing, and Formatting

Insights & Use Cases

By

Niels Swimberghe

,

May 28, 2024

Filter profanity from audio files using Node.js

Insights & Use Cases

By

Niels Swimberghe

,

May 27, 2024

Content moderation on audio files with Python

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

May 24, 2024

Newsletter #37: Speaker Diarization Now in 5 New Languages 🇨🇳🇮🇳🇯🇵🇰🇷🇻🇳 & Latest Speech AI tutorials

Newsletter

By

Smitha Kolan

,

Developer Educator

May 22, 2024

Filter profanity from audio files using Python

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

May 17, 2024

Newsletter #36: Latest Speech-to-text Model Benchmarks

Newsletter

By

Smitha Kolan

,

Developer Educator

May 10, 2024

Newsletter #35: Nano & Best: New Speech-to-text Pricing Options

Newsletter

By

Smitha Kolan

,

Developer Educator

May 8, 2024

Best and Nano Tiers: More Speech-to-Text and Pricing Options

Releases & Updates

By

Kelly Moon

,

May 3, 2024

Newsletter #34: AssemblyAI API Reference & Latest Tutorials

Newsletter

By

Smitha Kolan

,

Developer Educator

April 26, 2024

Newsletter #33: Make.com Voice AI Integration and Streaming STT Updates

Newsletter

By

Smitha Kolan

,

Developer Educator

April 26, 2024

Best Large Language Models (LLMs) & Frameworks in 2024

Insights & Use Cases

By

Jesse Sumrak

,

Featured writer

April 23, 2024

Redact PII in Audio with Make and AssemblyAI

Insights & Use Cases

By

Niels Swimberghe

,

April 23, 2024

Introducing the AssemblyAI app for Make (Integromat)

Releases & Updates

By

Niels Swimberghe

,

April 19, 2024

Newsletter #32:⚡️ Upgrades To Streaming Speech-to-Text

Newsletter

By

Smitha Kolan

,

Developer Educator

April 19, 2024

AssemblyAI + 🔗LangChain Go, Universal-1 Recap

Newsletter

By

Smitha Kolan

,

Developer Educator

April 15, 2024

Transcribe audio with Java using Universal-1

Insights & Use Cases

By

Niels Swimberghe

,

April 12, 2024

Newsletter #30: 🚀 Universal-1 Model Launch

Newsletter

By

Smitha Kolan

,

Developer Educator

April 10, 2024

Introducing the AssemblyAI integration for LangChain Go

Releases & Updates

By

Marcus Olsson

,

Senior Developer Educator

April 8, 2024

9 ways to transform contact center results with AI-powered speech analytics

Insights & Use Cases

By

Jesse Sumrak

,

Featured writer

April 5, 2024

Build Audio LLM Apps with AssemblyAI

Newsletter

By

Smitha Kolan

,

Developer Educator

April 3, 2024

Introducing Universal-1

By

Dylan Fox

,

Founder, CEO

March 21, 2024

Improved Streaming Speech-to-Text Pricing and Features

Newsletter

By

Mısra Turp

,

Developer Educator

March 19, 2024

Real-Time is now Streaming Speech-to-Text, with added customization and control for users

Releases & Updates

By

Kelsey Foster

,

Growth

March 14, 2024

🔥 New Free Video Course from Talk Python: Build an Audio AI App

Newsletter

By

Smitha Kolan

,

Developer Educator

March 13, 2024

A New Free Python Course to Build Real-World Audio AI Apps

Releases & Updates

By

Patrick Loeber

,

Senior Developer Advocate

March 13, 2024

AssemblyAI Go SDK v1.3.0: Utterance Detection and Word Search

Releases & Updates

By

Marcus Olsson

,

Senior Developer Educator

March 13, 2024

AssemblyAI Java SDK New Features & Improvements

Newsletter

By

Smitha Kolan

,

Developer Educator

March 8, 2024

Improved Audio LLM Docs & AssemblyAI Go SDK

Newsletter

By

Smitha Kolan

,

Developer Educator

March 7, 2024

AI tools for business: Top 6 considerations before building with AI models and LLMs

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 4, 2024

How to use AI to build powerful market research tools

Insights & Use Cases

By

Kelsey Foster

,

Growth

February 29, 2024

Top 3 ways to enhance AI video editing tools with Voice AI

Insights & Use Cases

By

Kelsey Foster

,

Growth

February 23, 2024

New Utterance Controls for Real-Time Transcription

Newsletter

By

Smitha Kolan

,

Developer Educator

February 23, 2024

Why product teams at top call tracking solutions are turning to AI

Insights & Use Cases

By

Kelsey Foster

,

Growth

February 17, 2024

PII Redaction and Entity Detection In 13 New Languages 🇫🇷🇩🇪🇮🇳

Newsletter

By

Smitha Kolan

,

Developer Educator

February 9, 2024

Improvements to Real-Time Transcription

Newsletter

By

Smitha Kolan

,

Developer Educator

February 1, 2024

Ask questions about your audio with LLMs

Newsletter

By

Smitha Kolan

,

Developer Educator

January 26, 2024

🚀 New AssemblyAI Go SDK & Speech-to-Text Tutorials

Newsletter

By

Smitha Kolan

,

Developer Educator

January 23, 2024

How to do Speech-To-Text with Go

Insights & Use Cases

By

Patrick Loeber

,

Senior Developer Advocate

January 19, 2024

Claude 2.1 Now Available with LeMUR + New Integrations

Newsletter

By

Smitha Kolan

,

Developer Educator

January 19, 2024

Announcing the AssemblyAI Go SDK

Releases & Updates

By

Marcus Olsson

,

Senior Developer Educator

January 16, 2024

Announcing the AssemblyAI Integration for Haystack

Releases & Updates

By

Mısra Turp

,

Developer Educator

January 10, 2024

Lower latency, lower cost, more possibilities

Releases & Updates

By

Ryan O'Connor

,

Senior Developer Educator

January 8, 2024

Announcing the AssemblyAI integration for Semantic Kernel .NET

Releases & Updates

By

Niels Swimberghe

,

January 8, 2024

Ask .NET Rocks! questions with Semantic Kernel, GPT, and Chroma DB

Insights & Use Cases

By

Niels Swimberghe

,

January 5, 2024

AssemblyAI's New Integrations & Latest Tutorials

Newsletter

By

Smitha Kolan

,

Developer Educator

January 2, 2024

Why Virtual Meeting Companies Should Use Voice AI

Insights & Use Cases

By

Julie Griffin

,

Featured writer

December 20, 2023

2023 at AssemblyAI - A Year in Review

By

Smitha Kolan

,

Developer Educator

December 15, 2023

How to Create VTT Files for Videos in Python

Insights & Use Cases

By

Patrick Loeber

,

Senior Developer Advocate

December 15, 2023

🚀 New Punctuation & Casing Model For Real-Time Transcription

Releases & Updates

By

Smitha Kolan

,

Developer Educator

December 14, 2023

How to Create SRT Files for Videos in Python

Insights & Use Cases

By

Patrick Loeber

,

Senior Developer Advocate

December 8, 2023

🎉 Announcing our $50M Series C to build superhuman Voice AI models

News

By

Smitha Kolan

,

Developer Educator

December 7, 2023

AI for Universal Audio Understanding: Qwen-Audio Explained

Insights & Use Cases

By

Marco Ramponi

,

December 6, 2023

Announcing the AssemblyAI integration for LlamaIndex.TS

Releases & Updates

By

Niels Swimberghe

,

December 6, 2023

How to integrate spoken audio into LlamaIndex.TS using AssemblyAI

Insights & Use Cases

By

,

December 3, 2023

Announcing our $50M Series C to build superhuman Voice AI models

Releases & Updates

By

Dylan Fox

,

Founder, CEO

December 1, 2023

Improved Hold Music Detection + Build LLM Audio Apps with LeMUR

Releases & Updates

By

Smitha Kolan

,

Developer Educator

December 1, 2023

5 Benefits of Voice AI for Video Editing Platforms

Insights & Use Cases

By

Amanda Smith

,

November 30, 2023

6 Ways Telehealth Platforms Can Leverage Speech-to-Text AI

Insights & Use Cases

By

Julie Griffin

,

Featured writer

November 27, 2023

Should I build or buy an AI speech recognition system?

Insights & Use Cases

By

Kelsey Foster

,

Growth

November 24, 2023

🚀 LeMUR's Custom Text Input + Revamped Playground

Newsletter

By

Smitha Kolan

,

Developer Educator

November 20, 2023

AssemblyAI is now on the Amazon Web Services (AWS) Marketplace

Releases & Updates

By

Kelsey Foster

,

Growth

November 16, 2023

Enhancing Our Speech-to-Text Models with Google v5e TPUs and 🎉100K on YouTube

Newsletter

By

Smitha Kolan

,

Developer Educator

November 15, 2023

7 best practices for product teams to consider when building with AI

Insights & Use Cases

By

Kelsey Foster

,

Growth

November 15, 2023

7 best practices for product teams to consider when building with AI

Insights & Use Cases

By

Kelsey Foster

,

Growth

November 8, 2023

Introducing Our New Punctuation Restoration and Truecasing Models

Releases & Updates

By

Marco Ramponi

,

November 7, 2023

Automatically determine video sections with AI using Python

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

November 7, 2023

Improved Punctuation Restoration & Truecasing Models

Newsletter

By

Smitha Kolan

,

Developer Educator

November 2, 2023

Key phrase detection in audio files using Python

Insights & Use Cases

By

Patrick Loeber

,

Senior Developer Advocate

November 1, 2023

Faster Audio File Handling and Improved Error Messages

Newsletter

By

Smitha Kolan

,

Developer Educator

October 30, 2023

Transcribe audio to text on Cloudflare Workers with AssemblyAI and TypeScript

Insights & Use Cases

By

Niels Swimberghe

,

October 27, 2023

How Bluedot built with AssemblyAI to increase user conversion rate

Insights & Use Cases

By

Kelsey Foster

,

Growth

October 27, 2023

Combining Speech Recognition and Diarization in one model

Insights & Use Cases

By

Marco Ramponi

,

October 26, 2023

New Code Cookbooks & AssemblyAI's Q4 Product Enhancements

Newsletter

By

Smitha Kolan

,

Developer Educator

October 19, 2023

New Multilingual Capabilities and TypeScript/JavaScript SDK

Newsletter

By

Smitha Kolan

,

Developer Educator

October 16, 2023

How to use audio data in LlamaIndex with Python

Insights & Use Cases

By

Patrick Loeber

,

Senior Developer Advocate

October 12, 2023

Announcing the AssemblyAI Node SDK 2.0

Releases & Updates

By

Niels Swimberghe

,

October 10, 2023

Building with Automatic Speech Recognition (ASR) models: Why accuracy matters

Insights & Use Cases

By

Kelsey Foster

,

Growth

October 9, 2023

🚀LlamaIndex Integration + Model-Specific Usage Dashboards

Newsletter

By

Smitha Kolan

,

Developer Educator

October 2, 2023

New Usage Dashboard + Mistral 7B First Look

Newsletter

By

Smitha Kolan

,

Developer Educator

September 29, 2023

How DALL-E 2 Actually Works

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

September 29, 2023

8 Ways Automatic Speech Recognition Can Increase Efficiency For Your Business

Insights & Use Cases

By

Julie Griffin

,

Featured writer

September 27, 2023

How to use Voice AI systems for podcast hosting, editing, and monetization

Insights & Use Cases

By

Kelsey Foster

,

Growth

September 26, 2023

Retrieval Augmented Generation on audio data with LangChain and Chroma

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

September 22, 2023

What AI Music Generators Can Do (And How They Do It)

Insights & Use Cases

By

Marco Ramponi

,

September 20, 2023

Announcing the AssemblyAI plugin for Rivet

Releases & Updates

By

Niels Swimberghe

,

September 14, 2023

How to get Zoom Transcripts with the Zoom API

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

September 4, 2023

What is Residual Vector Quantization?

Insights & Use Cases

By

Marco Ramponi

,

August 31, 2023

How to build an interactive lecture summarization app

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

August 31, 2023

How to use audio data in LangChain with Python

Insights & Use Cases

By

Patrick Loeber

,

Senior Developer Advocate

August 22, 2023

RLHF vs RLAIF for language model alignment

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

August 18, 2023

Why Language Models Became Large Language Models And The Hurdles In Developing LLM-based Applications

Insights & Use Cases

By

Marco Ramponi

,

August 15, 2023

How to integrate spoken audio into LangChain.js using AssemblyAI

Insights & Use Cases

By

Niels Swimberghe

,

August 15, 2023

Introducing the AssemblyAI integration for LangChain.js

Releases & Updates

By

Niels Swimberghe

,

August 14, 2023

Customer Stories: Conformer-2 in Action

Insights & Use Cases

By

Kelsey Foster

,

Growth

August 1, 2023

How Reinforcement Learning from AI Feedback works

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

July 27, 2023

LeMUR: Now Available for Early Access

Releases & Updates

By

Kelsey Foster

,

Growth

July 27, 2023

Recent developments in Generative AI for Audio

Insights & Use Cases

By

Marco Ramponi

,

July 19, 2023

Conformer-2: a state-of-the-art speech recognition model trained on 1.1M hours of data

Releases & Updates

By

Francis McCann

,

May 23, 2023

Large Language Models for Product Managers: 5 Things to Know

Insights & Use Cases

By

Marco Ramponi

,

May 23, 2023

How AI helps Marvin's users spend 60% less time analyzing research data

Insights & Use Cases

By

Kelsey Foster

,

Growth

May 17, 2023

Introduction to Large Language Models for Generative AI

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

May 16, 2023

3 ways to build and deploy AI tools and features faster

Insights & Use Cases

By

Kelsey Foster

,

Growth

May 3, 2023

The Full Story of Large Language Models and RLHF

Insights & Use Cases

By

Marco Ramponi

,

May 2, 2023

Everything you need to know about Generative AI

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

May 2, 2023

Introduction to Generative AI

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

April 19, 2023

How physics advanced Generative AI

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

April 3, 2023

How RLHF Preference Model Tuning Works (And How Things May Go Wrong)

Insights & Use Cases

By

Marco Ramponi

,

March 15, 2023

Conformer-1: A robust speech recognition model trained on 650K hours of data

Releases & Updates

By

Marco Ramponi

,

March 13, 2023

3 easy ways to add AI Summarization to Conversation Intelligence tools

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 7, 2023

Emergent Abilities of Large Language Models

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

February 21, 2023

Why every Fortune 500 business needs a chief AI officer

Insights & Use Cases

By

Dylan Fox

,

Founder, CEO

January 19, 2023

Build a free Stable Diffusion app with a GPU backend

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

January 18, 2023

AI research review – Locating and Editing Factual Associations in GPT

Insights & Use Cases

By

Gabriel Oexle

,

December 23, 2022

How ChatGPT actually works

Insights & Use Cases

By

Marco Ramponi

,

December 15, 2022

Winners and Honorable Mentions - AssemblyAI $50k Winter Hackathon

Releases & Updates

By

Ryan O'Connor

,

Senior Developer Educator

December 14, 2022

Releasing our new v9 transcription model - 11% better accuracy

Releases & Updates

By

Ryan O'Connor

,

Senior Developer Educator

December 12, 2022

Build standout call coaching features with AI Summarization

Insights & Use Cases

By

Kelsey Foster

,

Growth

December 6, 2022

Stable Diffusion 1 vs 2 - What you need to know

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

November 30, 2022

Stable Diffusion in Keras - A Simple Tutorial

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

November 22, 2022

DeepMind's AlphaTensor Explained

Insights & Use Cases

By

Marco Ramponi

,

November 16, 2022

AI research review - Merging Models Modulo Permutation Symmetries

Insights & Use Cases

By

Yash Khare

,

Deep Learning Researcher

November 7, 2022

AI for product managers: Today’s top terms to stay in the know

Insights & Use Cases

By

Kelsey Foster

,

Growth

October 26, 2022

An Introduction to Poisson Flow Generative Models

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

October 19, 2022

Transcribe audio or video files right from your terminal

Releases & Updates

By

Francisco Castillo

,

September 28, 2022

AssemblyAI Recognized as G2 High Performer, Momentum Leader for Fall 2022

Releases & Updates

By

Kelsey Foster

,

Growth

September 21, 2022

AI Research Review - Multistream CNN

Insights & Use Cases

By

Luka Chkhetiani

,

Deep Learning Research Lead

September 21, 2022

Getting Started with Hugging Face's Gradio

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

September 14, 2022

Introducing the AssemblyAI Creators Program

Releases & Updates

By

Patrick Loeber

,

Senior Developer Advocate

September 8, 2022

AI Research Review - Spelling and ASR

Insights & Use Cases

By

Taufiquzzaman Peyash

,

Deep Learning Engineer

September 6, 2022

New for Enterprise: Improved Accuracy, Always-on Support, and SOC 2 Type 2

Releases & Updates

By

Micky Teng

,

September 6, 2022

AssemblyAI Obtains SOC 2 Type 2 Compliance for 2022/2023

Releases & Updates

By

Mike Groves

,

September 2, 2022

2022 Benchmark Report

Insights & Use Cases

By

Lee Vaughn

,

API Support Engineer

September 1, 2022

Coming Soon in Fall 2022 at AssemblyAI

Releases & Updates

By

Kelsey Foster

,

Growth

August 24, 2022

Deep Learning Paper Recap - Diffusion and Transformer Models

Insights & Use Cases

By

Dillon Pulliam

,

August 23, 2022

How to Run Stable Diffusion Locally to Generate Images

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

August 17, 2022

Deep Learning Paper Recap - Redundancy Reduction and Sparse MoEs

Insights & Use Cases

By

Domenic Donato

,

August 17, 2022

MinImagen - Build Your Own Imagen Text-to-Image Model

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

August 10, 2022

Deep Learning Paper Recap - Transfer Learning

Insights & Use Cases

By

Michael Liang

,

August 3, 2022

Deep Learning Paper Recap - Automatic Speech Recognition

Insights & Use Cases

By

Taufiquzzaman Peyash

,

Deep Learning Engineer

August 3, 2022

Topic Detection in NLP: The Top APIs for 2023

By

,

July 27, 2022

Deep Learning Paper Recaps - Modality Matching and Masked Autoencoders

Insights & Use Cases

By

Luka Chkhetiani

,

Deep Learning Research Lead

July 20, 2022

AssemblyAI Named G2 Voice Recognition Software Leader in Winter 2022

Releases & Updates

By

Kelsey Foster

,

Growth

July 20, 2022

AssemblyAI Named a G2 High Performer and Momentum Leader for Summer 2022

Releases & Updates

By

Kelsey Foster

,

Growth

July 19, 2022

Creating Top Hiring Intelligence Platforms with AI Models

Insights & Use Cases

By

Kelsey Foster

,

Growth

July 14, 2022

Announcing our $30M Series B

Releases & Updates

By

Dylan Fox

,

Founder, CEO

July 7, 2022

Deep Learning Paper Recap - Language Models

Insights & Use Cases

By

Taufiquzzaman Peyash

,

Deep Learning Engineer

June 23, 2022

How Imagen Actually Works

By

Ryan O'Connor

,

Senior Developer Educator

June 22, 2022

Deep Learning Paper Recap - Streaming ASR and Summarization

Insights & Use Cases

By

Guru Rao

,

June 16, 2022

Review – TOXIGEN & Knowledge Distillation Meets Open-Set Semi-Supervised Learning

Insights & Use Cases

By

Domenic Donato

,

June 15, 2022

Hack with AssemblyAI: HawkHacks 2022

Releases & Updates

By

Kelsey Foster

,

Growth

June 8, 2022

Review - Decision Transformer & SPIRAL

Insights & Use Cases

By

Kevin Zhang

,

June 6, 2022

Getting Started with ESPnet

By

Ryan O'Connor

,

Senior Developer Educator

June 3, 2022

Building Standout Hybrid Event Solutions with AI Models

Insights & Use Cases

By

Kelsey Foster

,

Growth

May 12, 2022

Introduction to Diffusion Models for Machine Learning

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

May 11, 2022

Building an Intelligent Cloud-based Contact Center? How AI Models Can Help

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 18, 2022

Built with AssemblyAI - Real-time Speech-to-Image Generation

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 12, 2022

How to Build a JavaScript Audio Transcript Application

Insights & Use Cases

By

Stefan Rosanitsch

,

Contributor

April 7, 2022

MediaPipe for Dummies

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

April 4, 2022

JavaScript Text-to-Speech - The Easy Way

Insights & Use Cases

By

Stefan Rosanitsch

,

Contributor

March 29, 2022

AssemblyAI Recognized as G2 High Performer, Momentum Leader in Voice Recognition Software for Spring 2022

Releases & Updates

By

Kelsey Foster

,

Growth

March 28, 2022

A Beginner's Guide to TorchStudio, The PyTorch IDE

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

March 20, 2022

React Text to Speech - Simplified!

Insights & Use Cases

By

Stefan Rosanitsch

,

Contributor

March 17, 2022

Automate Meeting Notes with Python

Insights & Use Cases

By

Mısra Turp

,

Developer Educator

March 16, 2022

Review - ALBERT: A Lite BERT for Self-supervised Learning of Language Representations

Insights & Use Cases

By

Sergio Ramirez Martin

,

March 15, 2022

Built with AssemblyAI - YouTube Transcripts

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 15, 2022

Transcribe Audio Files in an S3 Bucket with AssemblyAI

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

March 11, 2022

Kaldi Install for Dummies

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

March 11, 2022

AI Models for Smart Media Monitoring

Insights & Use Cases

By

Kelsey Foster

,

Growth

March 4, 2022

Announcing Our $28M Series A Led by Accel

Releases & Updates

By

Dylan Fox

,

Founder, CEO

March 2, 2022

Differentiable Programming - A Simple Introduction

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

February 25, 2022

BitFit: Simple Parameter-efficient Fine-tuning for Transformer-based Masked Language-models

Insights & Use Cases

By

Taufiquzzaman Peyash

,

Deep Learning Engineer

February 24, 2022

Learn How To Get Started with OpenAI API and GPT-3

Insights & Use Cases

By

Patrick Loeber

,

Senior Developer Advocate

February 21, 2022

What is Gradient Clipping for Neural Networks?

Insights & Use Cases

By

Mısra Turp

,

Developer Educator

February 15, 2022

Why You Should (or Shouldn't) be Using Google's JAX in 2023

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

February 14, 2022

Hyperparameters of Neural Networks

Insights & Use Cases

By

Mısra Turp

,

Developer Educator

February 14, 2022

How to Build a Python Project that Summarizes Your Lectures

Insights & Use Cases

By

Patrick Loeber

,

Senior Developer Advocate

February 8, 2022

What is Layer Normalization?

Insights & Use Cases

By

Mısra Turp

,

Developer Educator

February 4, 2022

Review - Perceiver: General Perception with Iterative Attention

Insights & Use Cases

By

Dillon Pulliam

,

February 2, 2022

Reinforcement Learning With (Deep) Q-Learning Explained

By

Patrick Loeber

,

Senior Developer Advocate

February 2, 2022

Built with AssemblyAI - Rhetoric

Insights & Use Cases

By

Kelsey Foster

,

Growth

February 1, 2022

Machine Learning Concepts for Beginners

Insights & Use Cases

By

Kelsey Foster

,

Growth

January 31, 2022

What is Weight Initialization for Neural Networks?

Insights & Use Cases

By

Mısra Turp

,

Developer Educator

January 26, 2022

Review - data2vec: A General Framework for Self-supervised Learning in Speech, Vision, and Language

Insights & Use Cases

By

Guru Rao

,

January 26, 2022

Unsupervised Machine Learning For Beginners

Insights & Use Cases

By

Patrick Loeber

,

Senior Developer Advocate

January 25, 2022

How to Evaluate Machine Learning Models

Insights & Use Cases

By

Mısra Turp

,

Developer Educator

January 24, 2022

DeltaHacks - AssemblyAI at McMaster University Hackathon

Releases & Updates

By

Britney Xiu

,

January 24, 2022

What is BERT and How Does It Work?

Insights & Use Cases

By

Mısra Turp

,

Developer Educator

January 24, 2022

Supervised Machine Learning For Beginners

Insights & Use Cases

By

Patrick Loeber

,

Senior Developer Advocate

January 20, 2022

Kaldi Speech Recognition for Beginners - A Simple Tutorial

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

January 18, 2022

Recurrent Neural Networks (RNNs) Explained

Insights & Use Cases

By

Patrick Loeber

,

Senior Developer Advocate

January 10, 2022

Bias and Variance for Machine Learning

Insights & Use Cases

By

Mısra Turp

,

Developer Educator

January 6, 2022

Best Speech-to-Text Software

Insights & Use Cases

By

Patrick Loeber

,

Senior Developer Advocate

January 5, 2022

Built with AssemblyAI - Wordcab

Insights & Use Cases

By

Kelsey Foster

,

Growth

January 4, 2022

Backpropagation For Neural Networks Explained

Insights & Use Cases

By

Patrick Loeber

,

Senior Developer Advocate

January 4, 2022

Jupyter Notebooks Tips and Tricks

Insights & Use Cases

By

Mısra Turp

,

Developer Educator

January 3, 2022

Introduction to Variational Autoencoders Using Keras

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

December 29, 2021

Built with AssemblyAI - IronScribe

Insights & Use Cases

By

Kelsey Foster

,

Growth

December 28, 2021

What is GPT-3 and How Does It Work?

Insights & Use Cases

By

Mısra Turp

,

Developer Educator

December 27, 2021

Getting Started With Torchaudio

Insights & Use Cases

By

Patrick Loeber

,

Senior Developer Advocate

December 20, 2021

2021 at AssemblyAI - A Year in Review

Releases & Updates

By

Kelsey Foster

,

Growth

December 20, 2021

2022 at AssemblyAI - A Year in Review

Releases & Updates

By

Kelsey Foster

,

Growth

December 16, 2021

Introducing Sentiment Analysis - Detect Sentiments in Spoken Audio

Releases & Updates

By

Kelsey Foster

,

Growth

December 15, 2021

Review - JUST: Joint Unsupervised and Supervised Training For Multilingual ASR

By

,

December 15, 2021

Auto Chapters in Action - Build a Web App that Automatically Summarizes Podcasts

Insights & Use Cases

By

Patrick Loeber

,

Senior Developer Advocate

December 14, 2021

PyTorch vs TensorFlow in 2023

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

December 13, 2021

Sentiment Analysis in Action - Earnings Calls

Insights & Use Cases

By

Mısra Turp

,

Developer Educator

December 9, 2021

Activation Functions In Neural Networks Explained

Insights & Use Cases

By

Patrick Loeber

,

Senior Developer Advocate

December 7, 2021

Review - VICReg: Variance-Invariance-Covariance Regularization for Self-Supervised Learning

Insights & Use Cases

By

Kevin Zhang

,

December 6, 2021

PyTorch Lightning for Dummies - A Tutorial and Overview

Insights & Use Cases

By

Ryan O'Connor

,

Senior Developer Educator

December 1, 2021

Introducing Entity Detection - Detect Named Entities in Audio/Video

Releases & Updates

By

Kelsey Foster

,

Growth

November 30, 2021

Add Speech Recognition to Applications in 5 Minutes

Insights & Use Cases

By

Mısra Turp

,

Developer Educator

November 30, 2021

Quick Automatic Chapter Detection

Insights & Use Cases

By

Patrick Loeber

,

Senior Developer Advocate

November 30, 2021

Transformers for Beginners - An Introduction

Insights & Use Cases

By

Mısra Turp

,

Developer Educator

November 23, 2021

Review - SimCLS and RefSum - Summarization Techniques

Insights & Use Cases

By

Dillon Pulliam

,

November 22, 2021

What is Regularization? Overfitting and Neural Networks

Insights & Use Cases

By

Mısra Turp

,

Developer Educator

November 18, 2021

What is Sentiment Analysis?

Insights & Use Cases

By

Kelsey Foster

,

Growth

November 17, 2021

Podcasts About AI - Our Deep Learning Team’s Top Picks

Releases & Updates

By

Kelsey Foster

,

Growth

November 17, 2021

Review - Speech processing Universal Performance Benchmark Review

Insights & Use Cases

By

Guru Rao

,

November 16, 2021

Text Segmentation - Approaches, Datasets, and Evaluation Metrics

Insights & Use Cases

By

Taufiquzzaman Peyash

,

Deep Learning Engineer

November 11, 2021

Introducing Auto Chapters - Summarize Audio and Video Files

Releases & Updates

By

Dylan Fox

,

Founder, CEO

November 8, 2021

Top 7 Data Science Blogs for Data Scientists and Enthusiasts

Insights & Use Cases

By

Kelsey Foster

,

Growth

November 5, 2021

AssemblyAI at HackDuke and HackUMBC

By

,

November 5, 2021

An Overview of Transducer Models for ASR

Insights & Use Cases

By

Michael Nguyen

,

November 5, 2021

Batch Normalization for Neural Networks - How it Works

Insights & Use Cases

By

Mısra Turp

,

Developer Educator

November 2, 2021

Machine Learning Podcasts - The Ultimate Listening Guide

Insights & Use Cases

By

Kelsey Foster

,

Growth

November 1, 2021

How to Make a Web App that Transcribes YouTube Videos with Streamlit

Insights & Use Cases

By

Mısra Turp

,

Developer Educator

October 29, 2021

Data Science Podcasts to Listen to Now

Insights & Use Cases

By

Kelsey Foster

,

Growth

October 29, 2021

Deep Shallow Fusion for RNN-T Personalization

Insights & Use Cases

By

Michael Nguyen

,

October 26, 2021

Top 5 Machine Learning Blogs to Follow

Insights & Use Cases

By

Kelsey Foster

,

Growth

October 25, 2021

Deep Learning in 5 Minutes

Insights & Use Cases

By

Mısra Turp

,

Developer Educator

October 25, 2021

Hack the Valley - AssemblyAI at University of Toronto Hackathon

Releases & Updates

By

Britney Xiu

,

October 20, 2021

Review - A Graph Convolutional Neural Network for Emotion Recognition in Conversation

Insights & Use Cases

By

Shuyun Tang

,

October 19, 2021

Releasing our v8 Transcription Model - 18.72% Better Accuracy

Releases & Updates

By

Dylan Fox

,

Founder, CEO

October 13, 2021

DeepSpeech for Dummies - A Tutorial and Overview

Insights & Use Cases

By

Yujian Tang

,

Contributor

October 13, 2021

Review - Pretraining Representations for Data-Efficient Reinforcement Learning

Insights & Use Cases

By

Kevin Zhang

,

October 11, 2021

New - 8.37% Better Accuracy for Topic Detection and IAB Classification with V4 Update

Releases & Updates

By

Kelsey Foster

,

Growth

October 8, 2021

Top 3 Hackathon Projects Built with AssemblyAI’s Speech-to-Text API

Releases & Updates

By

Britney Xiu

,

October 6, 2021

Improved Accuracy on AssemblyAI’s Real Time Speech-to-Text API

Releases & Updates

By

Yujian Tang

,

Contributor

September 30, 2021

New: Improved Topic Detection and IAB Classification

Releases & Updates

By

Dillon Pulliam

,

September 27, 2021

AssemblyAI Recognized as G2 High Performer, Momentum Leader for Fall 2021

Releases & Updates

By

Kelsey Foster

,

Growth

September 24, 2021

Review - Text-Free Prosody-Aware Generative Spoken Language Modeling

Insights & Use Cases

By

Steven Hillis

,

Featured writer

September 21, 2021

How Well Does AI Transcribe Song Lyrics?

Insights & Use Cases

By

Yujian Tang

,

Contributor

September 3, 2021

8/31/2021 AWS Outage Post-Mortem

Releases & Updates

By

Mitch Anderson

,

September 2, 2021

Comparing Zoom Transcription Accuracy Across Speech-to-Text APIs

Insights & Use Cases

By

Joe Zaghloul

,

September 2, 2021

Can Podcasts Predict the Stock Market?

Insights & Use Cases

By

Yujian Tang

,

Contributor

August 18, 2021

How to Add Subtitles to Mux Videos with Python

Insights & Use Cases

By

Yujian Tang

,

Contributor

August 5, 2021

How to Set Up Twilio Voicemail

Insights & Use Cases

By

Yujian Tang

,

Contributor

August 2, 2021

How to Build a Burner Phone with Voicemail in Python

Insights & Use Cases

By

Yujian Tang

,

Contributor

July 28, 2021

The Definitive Guide to Python Click

Insights & Use Cases

By

Yujian Tang

,

Contributor

July 20, 2021

Improved Real-Time Transcription Speed and Accuracy

Releases & Updates

By

Andrew Galyan-Mann

,

Senior API Support Engineer

July 15, 2021

How to build a YouTube downloader in Python

Insights & Use Cases

By

,

June 15, 2021

Fine-Tuning Transformers for NLP

Insights & Use Cases

By

Dillon Pulliam

,

June 15, 2021

Speech-to-Text Accuracy on Podcasts, News Broadcasts, and Social Media

Insights & Use Cases

By

Joe Zaghloul

,

June 1, 2021

Voice Search Partnership with Algolia

Releases & Updates

By

Joe Zaghloul

,

June 1, 2021

Open Sourcing Drone Deploy ECS

Releases & Updates

By

Mitch Anderson

,

May 13, 2021

A New API Endpoint to Paginate Through Historical Transcripts

Releases & Updates

By

Dylan Fox

,

Founder, CEO

May 1, 2021

Improved WER - Speech-to-Text Accuracy vs. Google, AWS (May 2020 Update)

Releases & Updates

By

Joe Zaghloul

,

April 28, 2021

New! Break Transcripts into Paragraphs and Sentences

Releases & Updates

By

Andrew Galyan-Mann

,

Senior API Support Engineer

April 27, 2021

Open Sourcing our Drone CI/CD CloudWatch Auto Scaler

Releases & Updates

By

Mitch Anderson

,

April 8, 2021

Getting started with HttpClientFactory in C# and .NET 5

Insights & Use Cases

By

Yujian Tang

,

Contributor

April 1, 2021

Content Safety Detection is now GA!

Releases & Updates

By

Dylan Fox

,

Founder, CEO

March 23, 2021

Speech-to-Text with Postman and AssemblyAI

Insights & Use Cases

By

Yujian Tang

,

Contributor

March 12, 2021

Transcribing Local Audio Files with Node.js

Insights & Use Cases

By

Yujian Tang

,

Contributor

March 2, 2021

New Punctuation and Casing Model Released

Releases & Updates

By

Andrew Galyan-Mann

,

Senior API Support Engineer

January 1, 2021

Comparing End-To-End Speech Recognition Architectures in 2021

Insights & Use Cases

By

Michael Nguyen

,

December 1, 2020

Building an End-to-End Speech Recognition Model in PyTorch

Insights & Use Cases

By

Michael Nguyen

,

November 20, 2020

PII Redaction and Accuracy Improvements

Releases & Updates

By

Joe Zaghloul

,

September 15, 2020

AssemblyAI Wins Best Public API 2020

Releases & Updates

By

Joe Zaghloul

,

June 1, 2020

[Webinar] Conversation Intelligence with CallRail

Releases & Updates

By

Joe Zaghloul

,

May 1, 2020

Teaming up with Algolia for Effortless Voice Search

Releases & Updates

By

Joe Zaghloul

,

April 1, 2020

Automated SRT and VTT Video Captions (April 2020 Update)

Releases & Updates

By

Joe Zaghloul

,

March 1, 2020

PII Redaction for Speech-to-Text Transcriptions (March 2020 Update)

Releases & Updates

By

Joe Zaghloul

,

Whisper alternatives

By

,

Why AssemblyAI beats self-hosting Whisper

Insights & Use Cases

By

Kelsey Foster

,

Growth

No results found!

Please try different search parameters.

Previous Next

Featured

April 29, 2026

Introducing our Voice Agent API

Releases & Updates

By

Madison Bernstein

,

Product Marketing

April 8, 2026

How to evaluate speech recognition models

Insights & Use Cases

By

Kelsey Foster

,

Growth

%20benchmark%20might%20be%20lying%20to%20you%20(1).png)

March 24, 2026

Why your word error rate (WER) benchmark might be lying to you

Insights & Use Cases

By

Zackary Klebanoff

,

Applied AI Lead

March 3, 2026

Conversation Intelligence: The complete guide for 2026

Insights & Use Cases

By

,

Releases & updates

[See all\

April 29, 2026

Introducing our Voice Agent API

Releases & Updates

By

Madison Bernstein

,

Product Marketing

April 6, 2026

Voice AI guardrails: Built-in protection for compliance, quality, and cost control

Releases & Updates

By

Kelsey Foster

,

Growth

March 25, 2026

Introducing Medical Mode: Purpose-built accuracy for medical terminology

Releases & Updates

By

Madison Bernstein

,

Product Marketing

March 3, 2026

Universal-3 Pro Streaming: The most accurate real-time transcription model for voice agents

Releases & Updates

By

Madison Bernstein

,

Product Marketing

Insights & use cases

[See all\

April 29, 2026

Create an ambient AI scribe that works during telehealth video calls

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 29, 2026

Voice agents in noisy environments (Drive-Thrus, Contact Centers, Field)

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 29, 2026

When to use Voice Agent API vs. Universal-3 Pro Streaming

Insights & Use Cases

By

Kelsey Foster

,

Growth

April 29, 2026

Voice Agent Orchestrators Compared: Vapi vs Pipecat vs LiveKit with AssemblyAI

Insights & Use Cases

By

Kelsey Foster

,

Growth

News

[See all\

March 23, 2026

AssemblyAI Named a Leader in G2’s Spring 2026 Voice Recognition Report

News

By

Devon Malloy

,

Staff Growth Manager

February 11, 2026

Voice AI in 2026: Inside the companies and investments shaping the future of speech

News

By

Kelsey Foster

,

Growth

February 4, 2026

Inside AssemblyAI's NYC voice agents January 2026 meetup: Production insights from the front lines

News

By

Kelsey Foster

,

Growth

January 22, 2026

New 2026 insights report: What actually makes a good voice agent

News

By

Kelsey Foster

,

Growth

Video

[See all\

March 18, 2026

Real-time conversation intelligence: The shift from post-call analysis to live insights

Video

By

Kelsey Foster

,

Growth

.png)

July 2, 2025

The ongoing need for human-in-the-loop in conversation intelligence

Video

By

,

April 7, 2025

Predicting the future: What top AI founders have to say about innovation in 2025

Video

By

Kelsey Foster

,

Growth

February 27, 2025

Building with AI in 2025: Top advice from leading founders

Video

By

Kelsey Foster

,

Growth

Newsletter

[See all\

August 29, 2025

SF Hackathon + In-App Playground + 99 Language Support | August 29, 2025 Newsletter

Newsletter

By

Devon Malloy

,

Staff Growth Manager

August 1, 2025

Streaming STT Performance Update | August 1, 2025 Newsletter

Newsletter

By

,

July 17, 2025

Enhanced diarization + Dovetail case study + G2 wins | July 18, 2025 Newsletter

Newsletter

By

,

September 13, 2024

Build Powerful Speech AI Apps with AssemblyAI & Speaker Diarization Tutorials

Newsletter

By

Smitha Kolan

,

Developer Educator

Subscribe to AssemblyAI’s newsletter

By clicking “Submit” you agree to our TOS and Privacy Policy

Thank you for subscribing!

Oops! Something went wrong while submitting the form.

Unlock the value of voice data

Build what’s next on the platform powering thousands of the industry’s leading of Voice AI apps.

Try our API for free Contact sales