Research

Paper

TESTING March 25, 2026

How Open is Open TTS? A Practical Evaluation of Open Source TTS Tools for Romanian

Authors

Teodora Răgman, Adrian Bogdan Stânea, Horia Cucu, Adriana Stan

Abstract

Open-source text-to-speech (TTS) frameworks have emerged as highly adaptable platforms for developing speech synthesis systems across a wide range of languages. However, their applicability is not uniform -- particularly when the target language is under-resourced or when computational resources are constrained. In this study, we systematically assess the feasibility of building novel TTS models using four widely adopted open-source architectures: FastPitch, VITS, Grad-TTS, and Matcha-TTS. Our evaluation spans multiple dimensions, including qualitative aspects such as ease of installation, dataset preparation, and hardware requirements, as well as quantitative assessments of synthesis quality for Romanian. We employ both objective metrics and subjective listening tests to evaluate intelligibility, speaker similarity, and naturalness of the generated speech. The results reveal significant challenges in tool chain setup, data preprocessing, and computational efficiency, which can hinder adoption in low-resource contexts. By grounding the analysis in reproducible protocols and accessible evaluation criteria, this work aims to inform best practices and promote more inclusive, language-diverse TTS development. All information needed to reproduce this study (i.e. code and data) are available in our git repository: https://gitlab.com/opentts_ragman/OpenTTS

Metadata

arXiv ID: 2603.24116
Provider: ARXIV
Primary Category: eess.AS
Published: 2026-03-25
Fetched: 2026-03-26 06:02

Related papers

Raw Data (Debug)
{
  "raw_xml": "<entry>\n    <id>http://arxiv.org/abs/2603.24116v1</id>\n    <title>How Open is Open TTS? A Practical Evaluation of Open Source TTS Tools for Romanian</title>\n    <updated>2026-03-25T09:27:52Z</updated>\n    <link href='https://arxiv.org/abs/2603.24116v1' rel='alternate' type='text/html'/>\n    <link href='https://arxiv.org/pdf/2603.24116v1' rel='related' title='pdf' type='application/pdf'/>\n    <summary>Open-source text-to-speech (TTS) frameworks have emerged as highly adaptable platforms for developing speech synthesis systems across a wide range of languages. However, their applicability is not uniform -- particularly when the target language is under-resourced or when computational resources are constrained. In this study, we systematically assess the feasibility of building novel TTS models using four widely adopted open-source architectures: FastPitch, VITS, Grad-TTS, and Matcha-TTS. Our evaluation spans multiple dimensions, including qualitative aspects such as ease of installation, dataset preparation, and hardware requirements, as well as quantitative assessments of synthesis quality for Romanian. We employ both objective metrics and subjective listening tests to evaluate intelligibility, speaker similarity, and naturalness of the generated speech. The results reveal significant challenges in tool chain setup, data preprocessing, and computational efficiency, which can hinder adoption in low-resource contexts. By grounding the analysis in reproducible protocols and accessible evaluation criteria, this work aims to inform best practices and promote more inclusive, language-diverse TTS development.\n  All information needed to reproduce this study (i.e. code and data) are available in our git repository: https://gitlab.com/opentts_ragman/OpenTTS</summary>\n    <category scheme='http://arxiv.org/schemas/atom' term='eess.AS'/>\n    <published>2026-03-25T09:27:52Z</published>\n    <arxiv:comment>Published in IEEE Access</arxiv:comment>\n    <arxiv:primary_category term='eess.AS'/>\n    <author>\n      <name>Teodora Răgman</name>\n    </author>\n    <author>\n      <name>Adrian Bogdan Stânea</name>\n    </author>\n    <author>\n      <name>Horia Cucu</name>\n    </author>\n    <author>\n      <name>Adriana Stan</name>\n    </author>\n    <arxiv:doi>10.1109/ACCESS.2025.3637322</arxiv:doi>\n    <link href='https://doi.org/10.1109/ACCESS.2025.3637322' rel='related' title='doi'/>\n  </entry>"
}