Members

The publication list under each member is fetched automatically from Google Scholar and shows only that member's most recent papers, so it may be incomplete or out of date, and titles and venues are not always correct. For an accurate and complete list, please see each member's own website or Google Scholar profile, linked on their card. The lab's curated list of papers is on the Publications page.

Faculty

Publications (842)
  1. Full-Duplex-Bench v1.5: Evaluating Overlap Handling for Full-Duplex Speech Models ICASSP, 2026
  2. AHa-bench: Benchmarking audio hallucinations in large audio-language models Advances in Neural Information Processing Systems 38, 2026
  3. Full-Duplex-Bench-v2: A Multi-Turn Evaluation Framework for Duplex Dialogue Systems with an Automated Examiner ACL, 2026
  4. Stream RAG: Instant and Accurate Spoken Dialogue Systems with Streaming Tool Usage ICML, 2026
  5. UALM: Unified Audio Language Model for Understanding, Generation and Reasoning ICLR, 2026
  6. POWSM: A Phonetic Open Whisper-Style Speech Foundation Model ACL, 2026
  7. Low-resource audio codec (lrac): 2025 challenge description ICASSP 2026-2026 IEEE International Conference on Acoustics, Speech and …, 2026
  8. Tmt: Tri-modal translation between speech, image, and text by processing different modalities as different languages IEEE Transactions on Multimedia, 2026
  9. ICASSP 2026 URGENT Speech Enhancement Challenge ICASSP, 2026
  10. Bagpiper: Solving open-ended audio tasks via rich captions arXiv preprint arXiv:2602.05220, 2026
  11. An end-to-end integration of speech separation and recognition with self-supervised learning representation Computer Speech & Language 95, 101813, 2026
  12. Music Arena: Live evaluation for text-to-music Advances in Neural Information Processing Systems 38, 2026
  13. Urgentmos: Unified multi-metric and preference learning for robust speech quality assessment arXiv preprint arXiv:2601.18438, 2026
  14. Optimizing Conversational Quality in Spoken Dialogue Systems with Reinforcement Learning from AI Feedback ACLFindings, 2026
  15. PRiSM: Benchmarking Phone Realization in Speech Models ACL, 2026
  16. Speech-Hands: A Self-Reflection Voice Agentic Approach to Speech Recognition and Audio Reasoning with Omni Perception ACL, 2026
  17. Reasoning Beyond Majority Vote: An Explainable SpeechLM Framework for Speech Emotion Recognition ICASSP, 2026
  18. AudioChat: Unified Audio Storytelling, Editing, and Understanding with Transfusion Forcing ICML, 2026
  19. Mind the gap: Impact of synthetic conversational data on multi-talker ASR and speaker diarization arXiv preprint arXiv:2605.15442, 2026
  20. BSCodec: A Band-Split Neural Codec for High-Quality Universal Audio Reconstruction EACLFindings, 2026

Showing 20 of 842. View all on Google Scholar ↗


Post-Doc

Publications (102)
  1. ICASSP 2026 URGENT Speech Enhancement Challenge ICASSP, 2026
  2. Se-dicow: Self-enrolled diarization-conditioned whisper ICASSP 2026-2026 IEEE International Conference on Acoustics, Speech and …, 2026
  3. An end-to-end integration of speech separation and recognition with self-supervised learning representation Computer Speech & Language 95, 101813, 2026
  4. Urgentmos: Unified multi-metric and preference learning for robust speech quality assessment arXiv preprint arXiv:2601.18438, 2026
  5. Mind the gap: Impact of synthetic conversational data on multi-talker ASR and speaker diarization arXiv preprint arXiv:2605.15442, 2026
  6. Dissecting sensitivity to training language in self-supervised speech learning using neural audio codec tokens arXiv preprint arXiv:2607.26350, 2026
  7. Cross-Talk Speech Reduction, by Separation, for Separation arXiv preprint arXiv:2605.19695, 2026
  8. 2025 URGENT Speech Enhancement Challenge Multilingual Ṗ808 Listening Tests: Approach and Results ICASSP, 2026
  9. Ring Mixing with Auxiliary Signal-to-Consistency-Error Ratio Loss for Unsupervised Denoising in Speech Separation arXiv preprint arXiv:2604.08415, 2026
  10. The CMU-AIST submission for the ICME 2025 Audio Encoder Challenge arXiv preprint arXiv:2601.16273, 2026
  11. PlanRAG-Audio: Planning and Retrieval Augmented Generation for Long-form Audio Understanding ACLFindings, 2026
  12. ESPnet3: Infrastructure for Scalable Speech and Audio Research in the Foundation Model Era arXiv preprint arXiv:2606.21854, 2026
  13. Grounding Spoken LLMs in Multi-Speaker Audio via Diarization Conditioning arXiv preprint arXiv:2606.18134, 2026
  14. Exploiting Noise Inseparability for Weakly-Supervised Discriminative Speech Denoising Using Noisy Targets arXiv preprint arXiv:2606.02327, 2026
  15. MAPSS: Manifold-based Assessment of Perceptual Source Separation ICLR, 2026
  16. Who Spoke What When? Evaluating Spoken Language Models for Conversational ASR with Semantic and Overlap-Aware Metrics arXiv preprint arXiv:2603.22709, 2026
  17. Modeling Overlapped Speech with Shuffles arXiv preprint arXiv:2603.17769, 2026
  18. Interspeech 2025 URGENT Speech Enhancement Challenge Interspeech, 2025
  19. Lessons Learned from the URGENT 2024 Speech Enhancement Challenge Interspeech, 2025
  20. ESPnet-SpeechLM: An Open Speech Language Model Toolkit NAACL, 2025

Showing 20 of 102. View all on Google Scholar ↗


PhD Students

Publications (63)
  1. Omnivinci: Enhancing architecture and data for omni-modal understanding llm International Conference on Learning Representations 2026, 56101-56138, 2026
  2. UALM: Unified Audio Language Model for Understanding, Generation and Reasoning ICLR, 2026
  3. Bagpiper: Solving open-ended audio tasks via rich captions arXiv preprint arXiv:2602.05220, 2026
  4. Optimizing Conversational Quality in Spoken Dialogue Systems with Reinforcement Learning from AI Feedback ACLFindings, 2026
  5. Speech-Hands: A Self-Reflection Voice Agentic Approach to Speech Recognition and Audio Reasoning with Omni Perception ACL, 2026
  6. Reasoning Beyond Majority Vote: An Explainable SpeechLM Framework for Speech Emotion Recognition ICASSP, 2026
  7. BSCodec: A Band-Split Neural Codec for High-Quality Universal Audio Reconstruction EACLFindings, 2026
  8. Online Register for Dual-Mode Self-Supervised Speech Models: Mitigating the Lack of Future Context ICASSP, 2026
  9. Do Neural Codecs Generalize? A Controlled Study Across Unseen Languages and Non-Speech Tasks arXiv preprint arXiv:2601.12205, 2026
  10. An Efficient vLLM-Based Inference Pipeline for Unified Audio Understanding and Generation arXiv preprint arXiv:2607.02119, 2026
  11. Bagpiper-TTS: Natural Language Guided Universal Speech Synthesis arXiv preprint arXiv:2606.22811, 2026
  12. ESPnet3: Infrastructure for Scalable Speech and Audio Research in the Foundation Model Era arXiv preprint arXiv:2606.21854, 2026
  13. Online Predictive Coding for Dual-Mode Self-Supervised Speech Model arXiv preprint arXiv:2606.21268, 2026
  14. Bagpiper-Edit: Zero-Shot Open-Ended Audio Editing via Rich-Caption arXiv preprint arXiv:2606.21227, 2026
  15. VERSA: A Versatile Evaluation Toolkit for Speech, Audio, and Music NAACL, 2025
  16. Spoofceleb: Speech deepfake detection and sasv in the wild IEEE Open Journal of Signal Processing 6, 68-77, 2025
  17. OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning Interspeech, 2025
  18. Preference Alignment Improves Language Model-Based TTS ICASSP, 2025
  19. OWLS: Scaling Laws for Multilingual Speech Recognition and Translation Models ICML, 2025
  20. OpusLM: A Family of Open Unified Speech Language Models Interspeech, 2025

Showing 20 of 63. View all on Google Scholar ↗

Publications (78)
  1. Steerable vision-language-action policies for embodied reasoning and hierarchical control arXiv preprint arXiv:2602.13193, 2026
  2. Bagpiper: Solving open-ended audio tasks via rich captions arXiv preprint arXiv:2602.05220, 2026
  3. AudioChat: Unified Audio Storytelling, Editing, and Understanding with Transfusion Forcing ICML, 2026
  4. Why Does Action Chunking Improve Behavioral Cloning Performance in Robotic Control? arXiv preprint arXiv:2608.02547, 2026
  5. Dissecting sensitivity to training language in self-supervised speech learning using neural audio codec tokens arXiv preprint arXiv:2607.26350, 2026
  6. Improving Robotic Generalist Policies via Flow Reversal Steering arXiv preprint arXiv:2606.13675, 2026
  7. An Empirical Recipe for Universal Phone Recognition arXiv preprint arXiv:2603.29042, 2026
  8. Adapting Generalist Robot Policies with Semantic Reinforcement Learning arXiv preprint arXiv:2606.31958, 2026
  9. ESPnet3: Infrastructure for Scalable Speech and Audio Research in the Foundation Model Era arXiv preprint arXiv:2606.21854, 2026
  10. Bagpiper-Edit: Zero-Shot Open-Ended Audio Editing via Rich-Caption arXiv preprint arXiv:2606.21227, 2026
  11. Evaluating Large Language Models Abilities for Addressee, Turn-change, and Next Speaker Prediction in Meetings arXiv preprint arXiv:2606.17542, 2026
  12. Dynamic-SUPERB Phase-2: A Collaboratively Expanding Benchmark for Measuring the Capabilities of Spoken Language Models with 180 Tasks ICLR, 2025
  13. Quantum-enabled microwave-to-optical transduction via silicon nanomechanics Nature nanotechnology 20 (5), 602-608, 2025
  14. OWSM v4: Improving Open Whisper-Style Speech Models via Data Scaling and Cleaning Interspeech, 2025
  15. OWLS: Scaling Laws for Multilingual Speech Recognition and Translation Models ICML, 2025
  16. Polaris: Scalable real-to-sim evaluations for generalist robot policies arXiv preprint arXiv:2512.16881, 2025
  17. Vulbinllm: Llm-powered vulnerability detection for stripped binaries arXiv preprint arXiv:2505.22010, 2025
  18. OpusLM: A Family of Open Unified Speech Language Models Interspeech, 2025
  19. Measuring general intelligence with generated games arXiv preprint arXiv:2505.07215, 2025
  20. Findings of the IWSLT 2025 evaluation campaign PROCEEDINGS OF THE 22ND INTERNATIONAL CONFERENCE ON SPOKEN LANGUAGE …, 2025

Showing 20 of 78. View all on Google Scholar ↗

Shikhar Bharadwaj
Shikhar Bharadwaj (co-supervising)
Publications (26)
  1. POWSM: A Phonetic Open Whisper-Style Speech Foundation Model ACL, 2026
  2. PRiSM: Benchmarking Phone Realization in Speech Models ACL, 2026
  3. An Empirical Recipe for Universal Phone Recognition arXiv preprint arXiv:2603.29042, 2026
  4. The CMU-AIST submission for the ICME 2025 Audio Encoder Challenge arXiv preprint arXiv:2601.16273, 2026
  5. Phone Segmentation and Recognition through Phonological Activation Mapping arXiv preprint arXiv:2607.09020, 2026
  6. Gemini 2.5: Pushing the frontier with advanced reasoning, multimodality, long context, and next generation agentic capabilities arXiv preprint arXiv:2507.06261, 2025
  7. OpusLM: A Family of Open Unified Speech Language Models Interspeech, 2025
  8. ESPnet-SpeechLM: An Open Speech Language Model Toolkit NAACL, 2025
  9. ESPnet-SDS: Unified Toolkit and Demo for Spoken Dialogue Systems NAACL, 2025
  10. OpenBEATs: A Fully Open-Source General-Purpose Audio Encoder WASPAA, 2025
  11. Context-driven dynamic pruning for large speech foundation models arXiv preprint arXiv:2505.18860, 2025
  12. Identifying and Mitigating Mismatched Language Code in Multilingual ASR ICASSP 2025-2025 IEEE International Conference on Acoustics, Speech and …, 2025
  13. Identifying and Mitigating Mismatched Language Signal in Multilingual Automated Speech Recognition US Patent App. 19/173,179, 2025
  14. EmoNews: A Spoken Dialogue System for Expressive News Conversations Proceedings of the 26th Annual Meeting of the Special Interest Group on …, 2025
  15. VERSA-v2: A Modular and Scalable Toolkit for Speech and Audio Evaluation with Expanded Metrics, Visualization, and LLM Integration ASRU, 2025
  16. Indicgenbench: A multilingual benchmark to evaluate generation capabilities of llms on indic languages Proceedings of the 62nd Annual Meeting of the Association for Computational …, 2024
  17. Codequeries: A dataset of semantic queries over code Proceedings of the 17th Innovations in Software Engineering Conference, 1-11, 2024
  18. STAB: speech tokenizer assessment benchmark arXiv preprint arXiv:2409.02384, 2024
  19. Multimodal modeling for spoken language identification ICASSP 2024-2024 IEEE International Conference on Acoustics, Speech and …, 2024
  20. Label aware speech representation learning for language identification arXiv preprint arXiv:2306.04374, 2023

Showing 20 of 26. View all on Google Scholar ↗

Publications (16)
  1. Desta2. 5-audio: Toward general-purpose large audio language model with self-generated cross-modal alignment IEEE Transactions on Audio, Speech and Language Processing, 2026
  2. Bagpiper: Solving open-ended audio tasks via rich captions arXiv preprint arXiv:2602.05220, 2026
  3. Causal tracing of audio-text fusion in large audio language models arXiv preprint arXiv:2603.13768, 2026
  4. PlanRAG-Audio: Planning and Retrieval Augmented Generation for Long-form Audio Understanding ACLFindings, 2026
  5. Dynamic-SUPERB Phase-2: A Collaboratively Expanding Benchmark for Measuring the Capabilities of Spoken Language Models with 180 Tasks ICLR, 2025
  6. A preliminary exploration with gpt-4o voice mode arXiv preprint arXiv:2502.09940, 2025
  7. SpeechCaps: Advancing Instruction-Based Universal Speech Models with Multi-Talker Speaking Style Captioning ICASSP 2025-2025 IEEE International Conference on Acoustics, Speech and …, 2025
  8. Dynamic-Superb: Towards a Dynamic, Collaborative, and Comprehensive Instruction-Tuning Benchmark for Speech ICASSP, 2024
  9. Fusion of Discrete Representations and Self-Augmented Representations for Multilingual Automatic Speech Recognition SLT, 2024
  10. Prompting and adapter tuning for self-supervised encoder-decoder speech model 2023 IEEE Automatic Speech Recognition and Understanding Workshop (ASRU), 1-8, 2023
  11. Toward degradation-robust voice conversion ICASSP 2022-2022 IEEE International Conference on Acoustics, Speech and …, 2022
  12. Investigating on incorporating pretrained and learnable speaker representations for multi-speaker multi-style text-to-speech ICASSP 2021-2021 IEEE International Conference on Acoustics, Speech and …, 2021
  13. Defending your voice: Adversarial attack on voice conversion 2021 IEEE Spoken Language Technology Workshop (SLT), 552-559, 2021
  14. Utilizing self-supervised representations for MOS prediction Interspeech 2021, 2021
  15. How far are we from robust voice conversion: A survey 2021 IEEE Spoken Language Technology Workshop (SLT), 514-521, 2021
  16. Improving cross-lingual reading comprehension with self-training arXiv preprint arXiv:2105.03627, 2021
Jaeyeon Kim
Jaeyeon Kim (co-supervising)
Publications (13)
  1. The interspeech 2026 audio reasoning challenge: Evaluating reasoning process quality for audio reasoning models and agents arXiv preprint arXiv:2602.14224, 2026
  2. Wow-bench: Evaluating fine-grained acoustic perception in audio-language models via marine mammal vocalizations Findings of the Association for Computational Linguistics: ACL 2026, 31208-31229, 2026
  3. K-BrowseComp: A Web Browsing Agent Benchmark Grounded in Korean Contexts arXiv preprint arXiv:2606.02404, 2026
  4. Comparative Reasoning: Making an Audio Language Model Better at Comparing Emotions arXiv preprint arXiv:2606.24082, 2026
  5. Towards Scene-Aware Video-to-Spatial Audio Generation International Journal of Computer Vision 134 (4), 184, 2026
  6. Visage: Video-to-spatial audio generation ICLR 2025, 2025
  7. Multi-Domain Audio Question Answering Benchmark Toward Acoustic Content Reasoning ICASSP 2026, 2025
  8. Gaze Beyond the Frame: Forecasting Egocentric 3D Visual Span NeurIPS 2025, 2025
  9. Enclap: Combining neural audio codec and audio-text joint embedding for automated audio captioning ICASSP 2024, 2024
  10. Enclap++: Analyzing the enclap framework for optimizing automated audio captioning performance DCASE2024 Workshop, 2024
  11. Expanding on EnCLAP with auxiliary retrieval model for automated audio captioning DCASE2024 Challenge Technical Report, 2024
  12. Learning Semantic Information from Raw Audio Signal Using Both Contextual and Phonetic Representations ICASSP 2024, 2024
  13. PITS: Variational pitch inference without fundamental frequency for end-to-end pitch-controllable TTS ICML 2023 Workshop on SPIGM, 2023
Haeri Kim
Haeri Kim (co-supervising)
Publications (3)
  1. BBPE16: UTF-16-based byte-level byte-pair encoding for improved multilingual speech recognition ICASSP 2026-2026 IEEE International Conference on Acoustics, Speech and …, 2026
  2. A More Accurate Internal Language Model Score Estimation for the Hybrid Autoregressive Transducer Proc. Interspeech 2023, 869-873, 2023
  3. Self-diagnosing gan: Diagnosing underrepresented samples in generative adversarial networks Advances in Neural Information Processing Systems 34, 1925-1938, 2021

Master Students


Visitors

Publications (31)
  1. Test-time alignment for large language models via textual model predictive control International Conference on Learning Representations 2026, 317-346, 2026
  2. Webgen-v bench: Structured representation for enhancing visual design in llm-based web generation and evaluation Proceedings of the 32nd ACM SIGKDD Conference on Knowledge Discovery and …, 2026
  3. Benchmarking Agentic Newswriting via Journalistic Workflows Findings of the Association for Computational Linguistics: ACL 2026, 36450-36463, 2026
  4. Visual Prompt Discovery via Semantic Exploration arXiv preprint arXiv:2603.16250, 2026
  5. Adapting to Evolving Data: Test-Time Expert Aggregation for Imbalanced Tabular Regression Proceedings of the Nineteenth ACM International Conference on Web Search and …, 2026
  6. Multi-Agent Reinforcement Learning Correctable Strategy: A Framework with Correctable Strategies for Portfolio Management Engineering Proceedings 120 (1), 11, 2026
  7. Template-based financial report generation in agentic and decomposed information retrieval Proceedings of the 48th International ACM SIGIR Conference on Research and …, 2025
  8. Extending automatic machine translation evaluation to book-length documents Proceedings of the 2025 Conference on Empirical Methods in Natural Language …, 2025
  9. Mixture experts with test-time self-supervised aggregation for tabular imbalanced regression arXiv preprint arXiv:2506.07033, 2025
  10. Tree-of-report: Table-to-text generation for sports game reports with tree-structured prompting ACL 2025 Student Research Workshop, 2025
  11. NEWSAGENT: Benchmarking Multimodal Agents as Journalists with Real-World Newswriting Tasks arXiv preprint arXiv:2509.00446, 2025
  12. NVIDIA-NeMo’s WMT 2025 Metrics Shared Task Submission Proceedings of the Tenth Conference on Machine Translation, 920-925, 2025
  13. Ddot: A derivative-directed dual-decoder ordinary differential equation transformer for dynamic system modeling Pacific-Asia Conference on Knowledge Discovery and Data Mining, 434-445, 2025
  14. APAR: Modeling Irregular Target Functions in Tabular Regression via Arithmetic-Aware Pre-Training and Adaptive-Regularized Fine-Tuning Proceedings of the AAAI Conference on Artificial Intelligence 39 (20), 21536 …, 2025
  15. RallyDiffuser: A Representation-Guided Diffusion Model Framework for Strategic Planning in Badminton Proceedings of the 24th International Conference on Autonomous Agents and …, 2025
  16. Imitation Learning of Correlated Policies in Stackelberg Games arXiv preprint arXiv:2503.08883, 2025
  17. Plan2Align: Predictive Planning Based Test-Time Preference Alignment in Paragraph-Level Machine Translation. CoRR, 2025
  18. Root Cause Analysis In Microservice Using Neural Granger Causal Discovery Proceedings of the AAAI Conference on Artificial Intelligence 38 (1), 206-213, 2024
  19. BADGE: BADminton report Generation and Evaluation with LLM arXiv preprint arXiv:2406.18116, 2024
  20. The coachai badminton environment: Bridging the gap between a reinforcement learning environment and real-world badminton games Proceedings of the AAAI Conference on Artificial Intelligence 38 (21), 23844 …, 2024

Showing 20 of 31. View all on Google Scholar ↗


Industrial Collaborators

Publications (97)
  1. ASVspoof 5: Design, collection and validation of resources for spoofing, deepfake, and adversarial attack detection using crowdsourced speech Computer Speech & Language 95, 101825, 2026
  2. WildSpoof: advancing in-the-wild data in Text-to-Speech generation and Spoofing-aware automatic speaker verification ICASSP 2026-2026 IEEE International Conference on Acoustics, Speech and …, 2026
  3. Which Data Matter? Embedding-Based Data Selection for Speech Recognition arXiv preprint arXiv:2603.05819, 2026
  4. Spoofceleb: Speech deepfake detection and sasv in the wild IEEE Open Journal of Signal Processing 6, 68-77, 2025
  5. X3A: Efficient Multimodal Deepfake Detection with Score-Level Fusion ACM/SIGAPP Symposium on Applied Computing, 767-774, 2025
  6. Chain-of-Thought Training for Open E2E Spoken Dialogue Systems ISCA Interspeech 2025, 2025
  7. WildSpoof Challenge Evaluation Plan arXiv preprint arXiv:2508.16858, 2025
  8. Speaker-IPL: Unsupervised Learning of Speaker Characteristics with i-Vector based Pseudo-Labels ICASSP, 2025
  9. SEED: Speaker Embedding Enhancement Diffusion Model ISCA Interspeech 2025, 2025
  10. Context-Driven Dynamic Pruning for Large Speech Foundation Models ISCA Interspeech 2025, 2025
  11. Token-based Attractors and Cross-attention in Spoof Diarization IEEE ASRU, 2025
  12. ASVspoof 5: Crowdsourced Speech Data, Deepfakes, and Adversarial Attacks at Scale ASVspoof 2024 Workshop (ISCA Interspeech 2024 Satellite), 2024
  13. OWSM v3.1: Better and Faster Open Whisper-Style Speech Models based on E-Branchformer Interspeech, 2024
  14. Voxtlm: Unified Decoder-Only Models for Consolidating Speech Recognition, Synthesis and Speech, Text Continuation Tasks ICASSP, 2024
  15. Exploring Speech Recognition, Translation, and Understanding with Discrete Speech Units: A Comparative Study ICASSP, 2024
  16. ESPnet-SPK: full pipeline speaker embedding toolkit with reproducible recipes, self-supervised front-ends, and off-the-shelf models Interspeech, 2024
  17. ESPnet-Codec: Comprehensive Training and Evaluation of Neural Codecs for Audio, Music, and Speech SLT, 2024
  18. One Model to Rule Them All? Towards End-to-End Joint Speaker Diarization and Speech Recognition ICASSP, 2024
  19. Improving Audio Captioning Models with Fine-Grained Audio Features, Text Embedding Supervision, and LLM Mix-Up Augmentation ICASSP, 2024
  20. The VoxCeleb Speaker Recognition Challenge: A Retrospective IEEE/ACM Transactions on Audio, Speech, and Language Processing, 2024

Showing 20 of 97. View all on Google Scholar ↗


Alumni

Visting Faculty
  • 2023. 09 -- 2024. 06: Karen Livescu (TTIC)
Post-Docs
  • 2024. 02 - 2025. 05: Hye-jin Shim (CMU)
  • 2023. 03 - 2024. 09: Jeeweon Jung (CMU)
  • 2022. 03 - 2024. 08: Soumi Maiti (CMU)
  • 2021. 09 - 2024. 07: Zhong-Qiu Wang (CMU)
PhD
  • 2021. 09 - 2026. 04: Brian Yan (CMU)
  • 2022. 05 - 2026. 03: Li-Wei Chen (CMU, co-supervising)
  • 2020. 08 - 2026. 03: Siddhant Arora (CMU)
  • 2019. 09 - 2025. 12: Jiatong Shi (CMU)
  • 2020. 09 - 2025. 04: Yifan Peng (CMU)
  • 2019. 09 - 2024. 05: Xuankai Chang (CMU)
  • 2021. 09 - 2024. 05: Muqiao Yang (CMU, co-supervisor)
  • 2021. 01 - 2023. 06: Xinjian Li (CMU, co-supervisor)
  • 2020. 09 - 2023. 06: Jessica Huynh (CMU, co-supervisor)
  • 2020. 09 - 2022. 08: Siddharth Dalmia (CMU, co-supervisor)
  • 2017. 09 - 2021. 08: Aswin Shanmugam Subramanian (JHU)
  • 2017. 10 - 2021. 08: Matthew Maciejewski (JHU, co-supervisor)
  • 2017. 12 - 2020. 12: Matthew Wiesner (JHU, co-supervising)
MS & Undergraduate
  • 2024. 08 - 2026. 05: Chyi-Jiunn Lin (CMU)
  • 2023. 05 - 2025. 05: Kwanghee Choi (CMU)
  • 2022. 09 - 2024. 05: Shih-Lun Wu (CMU)
  • 2021. 09 - 2023. 05: Dan Berrebbi (CMU)
  • 2021. 09 - 2022. 12: Dorsa Zeinali (CMU)
  • 2021. 09 - 2022. 12: Karthik Ganesan (MIIS directed study, CMU)
  • 2021. 01 - 2022. 08: Chaitanya Narisetty (CMU)
  • 2021. 01 - 2022. 08: Peter Wu (CMU, co-supervisor)
  • 2021. 09 - 2022. 08: Sujay Suresh Kumar (MIIS directed study, CMU)
  • 2021. 09 - 2022. 08: Debayan Ghosh (CMU)
  • 2020. 08 - 2021. 08: Tianzi Wang (JHU)
  • 2018. 07 - 2019. 06: Zhiqi Wang (JHU)
  • 2017. 09 - 2018. 12: Szu-Jui Chen (JHU)
Visitors & Collaborators
  • 2026. 02 - 2026. 07: Dahee Yang (Hanyang University)
  • 2026. 01 - 2026. 05: Alexander Polok (Brno University of Technology)
  • 2025. 12 - 2026. 03: Xun Gong (Shanghai Jiao Tong University)
  • 2025. 08 - 2025. 12: Haoran Wang (Shanghai Jiao Tong University)
  • 2025. 04 - 2025. 12: Bo-Hao Su (National Tsing Hua University)
  • 2025. 05 - 2025. 11: Ji-Hoon Kim (Korea Advanced Institute of Science and Technology)
  • 2025. 02 - 2025. 07: Jialu Li (University of Illinois Urbana-Champaign)
  • 2025. 01 - 2025. 04: Pu Wang (KU Leuven)
  • 2024. 08 - 2025. 02: Kalvin Chang (UC Berkeley)
  • 2024. 09 - 2025. 02: Holger Severin Bovbjerg (Aalborg University)
  • 2024. 11 - 2024. 12: Junyi Peng (Brno University of Technology)
  • 2024. 08 - 2024. 12: Carlos Carvalho (Instituto Superior Técnico )
  • 2024. 04 - 2024. 12: Shuichiro Shimizu (Kyoto University)
  • 2024. 08 - 2024. 11: Shih-Heng Wang (National Taiwan University)
  • 2024. 08 - 2024. 11: Yoshiaki Bando (National Institute of Advanced Industrial Science and Technology)
  • 2023. 09 - 2024. 08: Yihan Wu (Renmin University)
  • 2023. 11 - 2024. 04: Chenda Li (Shanghai Jiaotong University)
  • 2023. 08 - 2024. 03: Roshan Sharma (Carnegie Mellon University)
  • 2023. 03 - 2024. 03: Wangyou Zhang (Shanghai Jiaotong University)
  • 2023. 08 - 2023. 10: Minsu Kim (Korea Advanced Institute of Science and Technology)
  • 2023. 05 - 2023. 08: Kohei Saijo (Waseda University)
  • 2022. 10 - 2023. 01: Takaaki Saeki (University of Tokyo)
  • 2021. 12 - 2022. 12: Yosuke Kashiwagi (Sony)
  • 2022. 04 - 2022. 09: Samuele Cornell (Universita Politecnica delle Marche)
  • 2022. 03 - 2022. 06: Yoshiki Masuyama (Tokyo Metropolitan University)
  • 2021. 07 - 2022. 07: Yushi Ueda (Japan Patent Office)
  • 2021. 08 - 2021. 12: Yen-Ju Lu (Research Center for Information Technology Innovation, Academia Sinica)
  • 2020. 01 - 2021. 01: Pengcheng Guo (Northwestern Polytechnical University)
  • 2019. 12 - 2020. 03 & 2022. 03 - 2022. 06: Yosuke Higuchi (Waseda University)
  • 2019. 12 - 2020. 12: Jing Shi (Chinese Academy of Science)
  • 2019. 07 - 2019. 10: Katsuki Inoue (Okayama University)
  • 2018. 11 - 2019. 05: Murali Karthick Baskar (Brno University of Technology)
  • 2018. 09 - 2018. 12: Xuankai Chang (Shanghai Jiao Tong University)
  • 2018. 08 - 2018. 09: Hirofumi Inaguma (Kyoto University)
  • 2018. 07 - 2020. 03: Yusuke Fujita (Hitachi Ltd.)
  • 2018. 04 - 2018. 09: Nelson Enrique Yalta Soplin (Waseda University)