Tony Woo

prof_pic.jpg

Hello everyone! My name is Tony Woo (or if you prefer my Korean name, Sang Hoon Woo) and I’m a first-year PhD student at Georgia Tech, where I’m advised by Professor Larry Heck. Previously, I was a research intern at the Vision & Learning Lab at Seoul National University.

My works cover a range of topics, but my central research area is conversational AI. Broadly, my research goal is to create natural, context-aware conversational agents that improve human capability. Some of my current subtopics of interest include:

  • Spoken Dialogue Interface: Humans perform most of their communication through speech, not text. Conversational systems should therefore support natural, voice-based interaction, not just by converting text to speech, but by understanding the characteristics of speech as a medium and adapting or exploiting them for more effective communication.
  • Multimodal Context: Human perception is inherently multimodal, shaped by simultaneous cues from vision, audio, and more. I am interested in developing conversational agents that can interpret these diverse signals, integrate them meaningfully, and, when appropriate, generate multimodal expressions themselves.
  • Dynamic Conversational Context: Conversations are dynamic. Users reveal new information about themselves, revise opinions, and shift preferences over time. Even external factors, e.g. temporal, social, or situational, change as interactions unfold. A robust conversational agent must track these evolving contexts and adapt its responses accordingly.

At Seoul National University, I was advised by Professor Gunhee Kim. Before that, I worked at a couple of Korean startups for my alternative military service, focusing on speech and language technologies. I received my B.S. in Computer Science from University of California Davis, where I had the opportunity to work with Professor Prem Devanbu and Professor Chen-Nee Chuah, that introduced me to the world of research.

news

Aug 20, 2026 Two papers accepted to EMNLP 2026: DuplexGen: Adaptive Synthesis of Human-AI Turn-Taking Dialogues (main) and StreamAlign: Streaming Text-Aligned Speech Tokenization (Findings)!
Apr 15, 2026 Excited to share that I’ve committed to Georgia Tech for my PhD, starting Fall 2026!
Apr 07, 2026 Our paper WoW-Bench: Evaluating Fine-Grained Acoustic Perception in Audio-Language Models via Marine Mammal Vocalizations got accepted to ACL 2026 Findings!
Oct 13, 2025 Just started my research internship at SK Telecom!
Aug 21, 2025 Our paper Think, Verbalize, then Speak: Bridging Complex Thoughts and Comprehensible Speech got accepted to EMNLP 2025!

selected publications

  1. Jaeyeon Kim, Heeseung Yun, Sang Hoon Woo, Chao-Han Huck Yang, and Gunhee Kim
    In ACL 2026 Findings
  2. Takyoung Kim, Kang-wook Kim, Sang Hoon Woo, Julia Hirschberg, Gunhee Kim, and Dilek Hakkani-Tür
    In EMNLP 2026
  3. Tony Woo*, Sehun Lee*, Kang-wook Kim, and Gunhee Kim
    In EMNLP 2025
  4. Jaeyeon Kim, Jaeyoon Jung, Jinjoo Lee, and Sang Hoon Woo
    In ICASSP 2024
  5. Hyoung-Kyu Song*, Sang Hoon Woo*, Junhyeok Lee, Seungmin Yang, Hyunjae Cho, Youseong Lee, Dongho Choi, and Kang-wook Kim
    In CVPR 2022