DEV Community

#audio

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Voice Cloning: Sample Quality Beats Sample Length, and Prosody Beats Both

Voice Cloning: Sample Quality Beats Sample Length, and Prosody Beats Both

1
Comments
4 min read
Stitching Hundreds of TTS Segments Into an Audiobook Without Seams

Stitching Hundreds of TTS Segments Into an Audiobook Without Seams

Comments
4 min read
Barge-In, VAD, and the Latency Budget: Engineering Realtime Voice

Barge-In, VAD, and the Latency Budget: Engineering Realtime Voice

Comments
4 min read
Multi-Voice Audiobooks: The Character Attribution Problem

Multi-Voice Audiobooks: The Character Attribution Problem

Comments
2 min read
Making Two TTS Voices Sound Like an Actual Conversation

Making Two TTS Voices Sound Like an Actual Conversation

Comments
2 min read
Test Voicebox Across Permission Loss, Bluetooth Handoff, and Backgrounding

Test Voicebox Across Permission Loss, Bluetooth Handoff, and Backgrounding

Comments
2 min read
Can YAMNet Detect Unseen Sudden Sounds in Real Time? A 48-Stream Evaluation

Can YAMNet Detect Unseen Sudden Sounds in Real Time? A 48-Stream Evaluation

Comments
7 min read
Transcribe Audio to Text Like a Developer: From File to Final Text

Transcribe Audio to Text Like a Developer: From File to Final Text

Comments
5 min read
How to Transcribe Audio to Text: A Practical Workflow for Developers

How to Transcribe Audio to Text: A Practical Workflow for Developers

Comments
5 min read
Half-Time vs Double-Time BPM Detection: How We Fixed Spotify's Known Accuracy Gap

Half-Time vs Double-Time BPM Detection: How We Fixed Spotify's Known Accuracy Gap

Comments
6 min read
Going Below Android's Volume Floor: DynamicsProcessing, LoudnessEnhancer, and Reproducible F-Droid Builds

Going Below Android's Volume Floor: DynamicsProcessing, LoudnessEnhancer, and Reproducible F-Droid Builds

Comments
6 min read
Recognizing every song played on internet radio

Recognizing every song played on internet radio

Comments
4 min read
Build a Voice Assistant with Python and Whisper

Build a Voice Assistant with Python and Whisper

Comments
5 min read
FFmpeg dynaudnorm Filter: Dynamic Audio Normalization Guide

FFmpeg dynaudnorm Filter: Dynamic Audio Normalization Guide

Comments
2 min read
FFmpeg atempo Filter: Change Audio Speed Without Pitch Shift

FFmpeg atempo Filter: Change Audio Speed Without Pitch Shift

Comments
5 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.