Papers

1

Total Citations

5

H-Index

1

About

Yakun Gao is a rising researcher in the field of speech and audio processing, with a primary focus on expressive and emotionally nuanced text-to-speech (TTS) synthesis. Their work addresses a critical gap in current TTS technology: the ability to generate audio that is not only high in quality but also capable of conveying complex emotions and controlled, detailed content. As a key organizer and contributor to the **Inspirational and Convincing Audio Generation Challenge 2024 (ICAGC 2024)**, part of the ISCSLP 2024 Competitions and Challenges track, Gao has helped define new benchmarks for evaluating and advancing emotionally intelligent speech synthesis. This challenge, already garnering 5 citations, underscores their commitment to pushing the boundaries of how machines produce human-like, persuasive, and inspirational audio. By spearheading efforts to move beyond flat, neutral TTS outputs, Yakun Gao is shaping a future where synthetic voices can truly resonate with listeners, making their contributions vital for applications in virtual assistants, audiobooks, and interactive media. Their work represents a significant step toward more authentic and engaging human-computer interaction.

Research Focus

Key Achievements

1
H-Index
1
Papers
5
Total Citations
5
Avg Citations/Paper
🏆 Most Cited Paper
ICAGC 2024: Inspirational and Convincing Audio Generation Challenge 2024
5 citations · 2024
📈 Most Prolific Year: 2024 (1 Papers)
🤝 Key Collaborators: 13
🏛 Institutions: Beijing University of Posts and Telecommunications

Top Papers

  1. 1

Key Collaborators

Contact & Links

Available for collaboration
Content generated · 14 days ago