FA-Bench Leaderboard for ASR with Word Timestamps
We ranked 28 speech recognition pipelines, 19 built from open models and 9 that call a commercial API, by how closely their word timestamps land on the boundar…
FA-Bench: A Benchmark for Evaluating Phone- and Word-Level Timestamp Accuracy in Forced Aligners
Most forced aligner results are hard to reproduce. FA-Bench is an open, standardized baseline. We scored 30 systems on phone- and word-level timestamp accuracy…
What Makes Olewave's Speech Data Cleaning Pipeline Effective and Unique
How Olewave's Olign pipeline turns raw audio into production-ready ASR/TTS training data with validated transcripts, word-level timestamps and confidence score…
Olewave at ICASSP 2024
We were a patron and exhibitor at IEEE ICASSP 2024 in Seoul, and our Founder & CEO Wei Chu delivered a spotlight talk on training speech models from wild, web-…
Harnessing Wild Data for High-ROI Speech R&D — Our ICASSP 2024 Spotlight
Our Founder & CEO Wei Chu's ICASSP 2024 spotlight talk on why traditional speech-service builds are low-ROI — and how filtering wild data + auto-labeling priva…