Building a Production Greek-English Speech Recognizer

  • 类型:arxiv
  • 标识:2609.13498
  • 链接:https://arxiv.org/abs/2609.13498
  • 主分类:multimodal
  • 形态:position
  • TLDR:We report a multi-month engineering program to build Sophea, a production bilingual Greek-English automatic speech recognition system. We evaluate the system against nine production gates covering Greek and English word error rate, language identification, and hallucinations on non-speech audio. Across twenty-three training iterations and two model architectures, no training-data composition passed all nine gates simultaneously. Meeting the Greek noisy-environment target required about 1,500 steps of dense domain exposure, while preserving English language identification tolerated only about 2
  • 副分类:engineering
  • 待LLM分类:否
  • 标题中文:构建生产级希腊语-英语语音识别器
  • TLDR中文:我们报告了一项历时数月的工程项目,用于构建 Sophea,一个生产级希腊语-英语双语自动语音识别系统。我们依据九个生产门控评估该系统,涵盖希腊语和英语词错误率、语言识别及非语音音频的幻觉问题。经过二十三次训练迭代和两种模型架构,没有任何训练数据组合能同时通过全部九个门控。满足希腊语噪声环境目标需要约1,500步密集领域暴露,而保持英语语言识别仅能容忍约2%
  • 来源文件
  • /inbox/tom/_candidates/2026-09-15-agent-rag-longcontext-candidates.json