跳到正文
arXiv:cs.CL· Zirui Li, Jens Edlund, Yicheng Gu, Nhan Phan, Lauri Juvela, Mikko Kurimo·· 6 小时前AI 评分40

Nord-Parl-TTS:基于北欧议会演讲的芬兰语和瑞典语 TTS 数据集

Nord-Parl-TTS: Finnish and Swedish TTS Dataset from Parliament Speech

AI 导读

Nord-Parl-TTS 是一个面向芬兰语和瑞典语的开放 TTS 数据集,从北欧议会演讲录音中提取了 900 小时芬兰语和 5090 小时瑞典语语音。该数据集基于改编的 Emilia 数据处理流程构建,并包含统一评测集,用于支持模型开发与基准测试,以缩小高资源与低资源语言之间的 TTS 数据差距。该工作已被 ICASSP 2026 接收。

正文

View PDF HTML (experimental)

Abstract:Text-to-speech (TTS) development is limited by scarcity of high-quality, publicly available speech data for most languages outside a few high-resource languages. We present Nord-Parl-TTS, an open TTS dataset for Finnish and Swedish based on speech found in the wild. Using recordings of Nordic parliamentary proceedings, we extract 900 hours of Finnish and 5090 hours of Swedish speech suitable for TTS training. The dataset is built using an adapted version of the Emilia data processing pipeline and includes unified evaluation sets to support model development and benchmarking. By offering open, large-scale data for Finnish and Swedish, Nord-Parl-TTS narrows the resource gap in TTS between high- and lower-resourced languages.
Comments: Accepted by ICASSP 2026. 5 pages, 2 figures
Subjects: Audio and Speech Processing (eess.AS); Computation and Language (cs.CL)
Cite as: arXiv:2509.17988 [eess.AS]
  (or arXiv:2509.17988v2 [eess.AS] for this version)
  https://doi.org/10.48550/arXiv.2509.17988

arXiv-issued DOI via DataCite

Submission history

From: Zirui Li [view email]
[v1] Mon, 22 Sep 2025 16:30:26 UTC (2,495 KB)
[v2] Sat, 7 Feb 2026 19:21:08 UTC (2,493 KB)

来源:arXiv:cs.CL · arxiv.org