LLMs1 min read
TamilEOT Dataset and Model Released for Tamil Telephone Speech
A new dataset, TamilEOT, and two audio-only detectors are released for semantic end-of-turn detection in Tamil telephone speech. The models achieve accuracy improvements of 13.41% on a held-out test set, with inference times under 150ms.
From arXiv cs.CL