Papers
Communities
Events
Blog
Pricing
Search
Open menu
Home
Papers
2505.15670
Cited By
v1
v2 (latest)
Efficient and Direct Duplex Modeling for Speech-to-Speech Language Model
21 May 2025
Ke Hu
Ehsan Hosseini-Asl
Chen Chen
Edresson Casanova
Subhankar Ghosh
Piotr .Zelasko
Zhiwen Chen
Jia-Nan Li
Jagadeesh Balam
Boris Ginsburg
AuLLM
Re-assign community
ArXiv (abs)
PDF
HTML
Papers citing
"Efficient and Direct Duplex Modeling for Speech-to-Speech Language Model"
4 / 4 papers shown
Title
Koel-TTS: Enhancing LLM based Speech Generation with Preference Alignment and Classifier Free Guidance
Shehzeen Samarah Hussain
Paarth Neekhara
Xuesong Yang
Edresson Casanova
Subhankar Ghosh
Mikyas T. Desta
Roy Fejgin
Rafael Valle
Jason Chun Lok Li
151
5
0
07 Feb 2025
MinMo: A Multimodal Large Language Model for Seamless Voice Interaction
Qian Chen
Yafeng Chen
Yanni Chen
Mengzhe Chen
Yuxiao Chen
...
Shiliang Zhang
Nan Zhao
Pei Zhang
Chuxu Zhang
Jinren Zhou
AuLLM
MLLM
112
24
0
10 Jan 2025
OmniFlatten: An End-to-end GPT Model for Seamless Voice Conversation
Qinglin Zhang
Luyao Cheng
Chong Deng
Qian Chen
Wen Wang
...
Jiaqing Liu
Hai Yu
Chaohong Tan
Zhihao Du
Shiliang Zhang
SyDa
BDL
AuLLM
VLM
143
20
0
23 Oct 2024
Chain-of-Thought Prompting for Speech Translation
Ke Hu
Zhehuai Chen
Chao-Han Huck Yang
Piotr Żelasko
Oleksii Hrinchuk
Vitaly Lavrukhin
Jagadeesh Balam
Boris Ginsburg
LRM
160
9
0
17 Sep 2024
1