Publications

2026

  1. MoLGE: Mixture of Language Group Experts for Efficient Scaling of Massively Multilingual Speech Recognition
    Sangmin Lee, Woojin Chung, Woongjib Choi, and Hong-Goo Kang
    In Conference on Language Modeling (COLM), 2026
  2. UR-BERT: Scaling Text Encoders for Massively Multilingual TTS Through Universal Romanization and Speech Token Prediction
    Sangmin Lee, Eekgyun Ahn, Woongjib Choi, and Hong-Goo Kang
    In INTERSPEECH, 2026
  3. UniverSR: Unified and Versatile Audio Super-Resolution via Vocoder-Free Flow Matching
    Woongjib Choi, Sangmin Lee, Hyungseob Lim, and Hong-Goo Kang
    In IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), 2026

2025

  1. Preprint
    SAGE-LD: Towards Scalable and Generalizable End-to-End Language Diarization via Simulated Data Augmentation
    Sangmin Lee, Woongjib Choi, Jihyun Kim, and Hong-Goo Kang
    arXiv preprint, 2025
  2. Neural Spectral Band Generation for Audio Coding
    Woongjib Choi, Byeong Hyeon Kim, Hyungseob Lim, Inseon Jang, and Hong-Goo Kang
    In INTERSPEECH, 2025

2024

  1. DNN-based Speech Codec Enhancement Utilizing Conditional Flow Matching
    Jaehoon Shin, Woongjib Choi, Byeong Hyeon Kim, Inseon Jang, and Hong-Goo Kang
    In Proceedings of the Korean Institute of Broadcast and Media Engineers (KIBME) Conference, 2024
  2. Non-Blind Speech Bandwidth Extension Using Pre-Trained Self-Supervised Features
    Woongjib Choi, Byeong Hyeon Kim, and Hong-Goo Kang
    In Proceedings of the Institute of Electronics and Information Engineers (IEIE) Conference, 2024