Minje Kim
Computer Science · Indiana University
Publications
171
Citations
2,700
Est. group size
—
Recurring co-author estimate
Active years
25
Publishing since 2002
Minje Kim's research focuses on processing speech and audio signals using machine learning, including separating individual speakers from mixed audio, removing background noise, and compressing audio efficiently for transmission or storage. Much of the recent work explores generative AI methods (such as diffusion models) and techniques to make these systems smaller and faster to run on limited hardware. The work is largely published through computer science and signal processing venues, with some outputs shared as preprints before formal publication.
Publication output has remained fairly steady over the past decade, with a peak around 2021 and consistent activity of roughly 11-18 papers per year in recent years.
Generated by claude-sonnet-5 from public bibliographic data · Jul 20, 2026
- Adaptive Deterministic Flow Matching for Target Speaker Extraction
2026
- Development and Application of an AI Diagnostic Assessment Platform for Primary English Literacy (AIDAPEL)
The Korea English Language Testing Association · 2026
- Generative Data Augmentation Challenge: Synthesis of Room Acoustics for Speaker Distance Estimation
2025
- DTA: Dual Temporal-channel-wise Attention for Spiking Neural Networks
2025
- AI Patent Network Analysis for the Advancement of Investigative Technologies
Criminal Investigation Studies · 2025
- Adaptive Slimming for Scalable and Efficient Speech Enhancement
2025
- TGIF: Talker Group-Informed Familiarization of Target Speaker Extraction
2025
- Perceptual Audio Coding: A 40-Year Historical Perspective
2025
- Generative Data Augmentation Challenge: Synthesis of Room Acoustics for Speaker Distance Estimation
arXiv (Cornell University) · 2025
- Adaptive Slimming for Scalable and Efficient Speech Enhancement
arXiv (Cornell University) · 2025
- DTA: Dual Temporal-channel-wise Attention for Spiking Neural Networks
arXiv (Cornell University) · 2025
- Combolutional Neural Networks
2025
- TGIF: Talker Group-Informed Familiarization of Target Speaker Extraction
arXiv (Cornell University) · 2025
- About Whom Are We Talking When We Use Intellectual and Developmental Disabilities?
JAMA Pediatrics · 2024
- Neural Speech and Audio Coding: Modern AI technology meets traditional codecs
IEEE Signal Processing Magazine · 2024
- arXiv (Cornell University)×43
- ICASSP 2022 - 2022 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)×5
- Cancer Research×5
- Zenodo (CERN European Organization for Nuclear Research)×4
- IEEE Journal of Selected Topics in Signal Processing×2
- Donald S. Williamson
Computer Science · Indiana University
- Heming Wang
Computer Science · The Ohio State University
- Hassan Taherian
Computer Science · The Ohio State University
- DeLiang Wang
Computer Science · The Ohio State University
- Anurag Kumar
Computer Science · The Ohio State University
This profile was generated automatically from public scholarly data (OpenAlex). Group size and activity levels are estimates derived from co-authorship patterns.
Last updated Jul 19, 2026.
Claim or correct this profile