ResearchTrend.AI
  • Papers
  • Communities
  • Events
  • Blog
  • Pricing
Papers
Communities
Social Events
Terms and Conditions
Pricing
Parameter LabParameter LabTwitterGitHubLinkedInBlueskyYoutube

© 2025 ResearchTrend.AI, All rights reserved.

  1. Home
  2. Papers
  3. 2504.07776
22
0

SlimSpeech: Lightweight and Efficient Text-to-Speech with Slim Rectified Flow

10 April 2025
K. Wang
Wenhao Guan
Shenghui Lu
Jianglong Yao
Lin Li
Q. Hong
ArXivPDFHTML
Abstract

Recently, flow matching based speech synthesis has significantly enhanced the quality of synthesized speech while reducing the number of inference steps. In this paper, we introduce SlimSpeech, a lightweight and efficient speech synthesis system based on rectified flow. We have built upon the existing speech synthesis method utilizing the rectified flow model, modifying its structure to reduce parameters and serve as a teacher model. By refining the reflow operation, we directly derive a smaller model with a more straight sampling trajectory from the larger model, while utilizing distillation techniques to further enhance the model performance. Experimental results demonstrate that our proposed method, with significantly reduced model parameters, achieves comparable performance to larger models through one-step sampling.

View on arXiv
@article{wang2025_2504.07776,
  title={ SlimSpeech: Lightweight and Efficient Text-to-Speech with Slim Rectified Flow },
  author={ Kaidi Wang and Wenhao Guan and Shenghui Lu and Jianglong Yao and Lin Li and Qingyang Hong },
  journal={arXiv preprint arXiv:2504.07776},
  year={ 2025 }
}
Comments on this paper