ResearchTrend.AI
  • Papers
  • Communities
  • Events
  • Blog
  • Pricing
Papers
Communities
Social Events
Terms and Conditions
Pricing
Parameter LabParameter LabTwitterGitHubLinkedInBlueskyYoutube

© 2025 ResearchTrend.AI, All rights reserved.

  1. Home
  2. Papers
  3. 2303.09048
8
2

Improving Perceptual Quality, Intelligibility, and Acoustics on VoIP Platforms

16 March 2023
Joseph Konan
Ojas Bhargave
Shikhar Agnihotri
Hojeong Lee
Ankit Parag Shah
Shuo Han
YUNYANG ZENG
Amanda Shu
Haohui Liu
Xuankai Chang
Hamza Khalid
Minseon Gwak
Kawon Lee
Minjeong Kim
Bhiksha Raj
ArXivPDFHTML
Abstract

In this paper, we present a method for fine-tuning models trained on the Deep Noise Suppression (DNS) 2020 Challenge to improve their performance on Voice over Internet Protocol (VoIP) applications. Our approach involves adapting the DNS 2020 models to the specific acoustic characteristics of VoIP communications, which includes distortion and artifacts caused by compression, transmission, and platform-specific processing. To this end, we propose a multi-task learning framework for VoIP-DNS that jointly optimizes noise suppression and VoIP-specific acoustics for speech enhancement. We evaluate our approach on a diverse VoIP scenarios and show that it outperforms both industry performance and state-of-the-art methods for speech enhancement on VoIP applications. Our results demonstrate the potential of models trained on DNS-2020 to be improved and tailored to different VoIP platforms using VoIP-DNS, whose findings have important applications in areas such as speech recognition, voice assistants, and telecommunication.

View on arXiv
Comments on this paper