DEV Community

#speechtotext

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
The Hard Cases in Speaker Diarization: Overlap & Noise

The Hard Cases in Speaker Diarization: Overlap & Noise

Comments
9 min read
Per-Tenant Cost Ledger for One API Key, Speech-to-Text, and Model Gateway Summaries

Per-Tenant Cost Ledger for One API Key, Speech-to-Text, and Model Gateway Summaries

Comments
7 min read
Sales Call CRM Actions: Node.js MP3/WAV Speech-to-Text Uploads Across US/EU

Sales Call CRM Actions: Node.js MP3/WAV Speech-to-Text Uploads Across US/EU

Comments
7 min read
EU Speech-to-Text API Procurement: A Startup Test Beyond Per-Minute Pricing

EU Speech-to-Text API Procurement: A Startup Test Beyond Per-Minute Pricing

Comments
7 min read
One API Key for Speech-to-Text Plus Model-Routed Transcript Briefs: An Engineering Test

One API Key for Speech-to-Text Plus Model-Routed Transcript Briefs: An Engineering Test

Comments
5 min read
Audio Transcription API 404 or 501: A Node.js Speech-to-Text Triage Guide

Audio Transcription API 404 or 501: A Node.js Speech-to-Text Triage Guide

Comments
5 min read
EU Speech-to-Text API Procurement: Compare Startup Pricing, SLOs, and Exit Tests

EU Speech-to-Text API Procurement: Compare Startup Pricing, SLOs, and Exit Tests

Comments
7 min read
From meeting audio to structured minutes in health settings

From meeting audio to structured minutes in health settings

Comments
6 min read
Why Realtime Is the Future of Speech-to-Text

Why Realtime Is the Future of Speech-to-Text

Comments
11 min read
Fast ASR for Voice Agents: Bring Your Own Turn Detection

Fast ASR for Voice Agents: Bring Your Own Turn Detection

Comments
6 min read
Sync vs. Async Transcription: Which to Use (2026)

Sync vs. Async Transcription: Which to Use (2026)

Comments
7 min read
Grok audio on Vercel, Go 1.26 green GC, voice APIs

Grok audio on Vercel, Go 1.26 green GC, voice APIs

Comments
4 min read
Build a Reliable AI Transcription Pipeline: A Developer’s Field Guide

Build a Reliable AI Transcription Pipeline: A Developer’s Field Guide

Comments 1
5 min read
Introducing Solaria-3: The most accurate speech-to-text model for European languages

Introducing Solaria-3: The most accurate speech-to-text model for European languages

Comments
9 min read
Turn Your Phone Into Voice Input for Any React Text Field

Turn Your Phone Into Voice Input for Any React Text Field

1
Comments
3 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.