매일 아침, 어제의 AI를 한 통으로 정리해 보내드립니다메일로 받아보기

METAL LAB

Easper: An Accessible ASR Pipeline for Language Documentation

arXiv:2608.116292026-08-13

arXiv:2608.11629v1 Announce Type: new Abstract: Audio transcription is a critical bottleneck in language documentation. While multilingual Automatic Speech Recognition (ASR) models like Whisper offer solutions, field linguists often lack the expertise to utilise them. We present Easper, an open-source, no-code workflow enabling linguists to iteratively fine-tune ASR models via cloud resources directly from ELAN annotations. Deploying ASR also raises a cold start problem: deciding which recording

저자 · Aso Mahmudi, Ting Dang, Ekaterina Vylomova, Nick Thieberger

arXiv에서 원문 보기