Publications
Recent Publications (Refer to this Google Scholar Page for the Full List of Publications)
Authors with underlined bold text represent the first authors of the publication.
The superscript asterisk (*) indicates the corresponding author.
End-to-End Sleep Sound Event Detection from Ambient Audio via Adaptive Line Enhancer Confidence Fusion
IEEE Trans. On Audio, Speech and Language Processing (TASLP), Sept. 2026. (JIF Top 8.5 % in the JCR category of "Acoustics")
HiGraph: Explicit Discrepancy Reasoning with Hierarchical Graph Propagation for Joint Pronunciation Assessment
Proc. Empirical Methods in Natural Language Processing (EMNLP) (Findings), Oct. 2026. (한국정보과학회 우수학술대회)
TOOLDF: Tool-Integrated Reasoning for Mixed-Authenticity Audio Deepfake Detection
Proc. Empirical Methods in Natural Language Processing (EMNLP) (Findings), Oct. 2026. (한국정보과학회 우수학술대회)
Delta2Gamma: Band-Wise Adaptive Contrastive Learning of EEG Rhythms for Alzheimer's Disease Detection
Proc. IEEE Biomedical Circuits And Systems (BioCAS), Oct. 2026.
Adaptive Line Enhancer-Derived ECG Separation and Weighted Reconstruction for Sleep Apena Detection
Proc. IEEE Biomedical Circuits And Systems (BioCAS), Oct. 2026.
LMMSE: A Clinically Inspired Framework for Cognitive Evaluation of Large Language Models
Information Sciences, July 2026 (JIF Top 18.2% in the JCR category of "Computer Science, Information Systems")
Sleep Sound Event Detection Powered by Learnable Multi-Resolution Adaptive Line Enhancer
Proc. INTERSPEECH, (The Long Paper Track), Sept. 2026 (한국정보과학회 우수학술대회)
SISER : Speaker Invariant Speech Emotion Recognition with Entropy-Based Adversarial Training
Proc. INTERSPEECH, Sept. 2026. (한국정보과학회 우수학술대회)
Beyond Short Segments : Expanding Speaker Embeddings with Vector Archives
Proc. INTERSPEECH, Sept. 2026. (한국정보과학회 우수학술대회)
From Masking to Merging: Rethinking SpecAugment for Efficient Audio Spectrogram Transformer
Proc. INTERSPEECH, Sept. 2026. (한국정보과학회 우수학술대회)
Efficient Punctuation Restoration via Weighted Lookahead Scoring Method for Streaming ASR Systems
Proc. International Joint Conference on Neural Networks (IJCNN), June, 2026.
Enhancing Document-Level Machine Translation via filtered synthetic corpora and two-stage LLM adaptation
Proc. IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP), May, 2026. (한국정보과학회 우수학술대회)
Controllable Singing Voice Synthesis using Phoneme-Level Energy Sequence
Proc. ASRU, Dec. 2025.
A Novel Chain-of-Thought Reasoning Approach for Alzheimer’s Disease Detection Using Large Language and Vision-Language Models
IEEE Trans. Neural Systems and Rehabilitation Engineering Nov. 2025 (TNSRE) (JIF Top 1.4 % in the JCR category of "Rehabilitation")
Reasoning-Based Approach with Chain-of-Thought for Alzheimer’s Detection Using Speech and Large Language Models
Proc. INTERSPEECH, 2025. (한국정보과학회 우수학술대회)
Wave-U-Mamba: An End-To-End Framework For High-Quality And Efficient Speech Super Resolution
Proc. IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP), 2025. (한국정보과학회 우수학술대회)
Mels-Tts: Multi-Emotion Multi-Lingual Multi-Speaker Text-To-Speech System Via Disentangled Style Tokens
Proc. IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP), 2024. (한국정보과학회 우수학술대회)
Latent Filling: Latent Space Data Augmentation for Zero-Shot Speech Synthesis
Proc. IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP), 2024. (한국정보과학회 우수학술대회)
Hierarchical Timbre-Cadence Speaker Encoder for Zero-shot Speech Synthesis
Proc. INTERSPEECH, 2023. (한국정보과학회 우수학술대회)
Self-Supervised Accent Learning for Under-Resourced Accents Using Native Language Data
Proc. IEEE International Conference on Acoustics, Speech, and Signal Processing (ICASSP), 2023. (한국정보과학회 우수학술대회)
Conformer-Based on-Device Streaming Speech Recognition with KD Compression and Two-Pass Architecture
Proc. SLT, 2022.
Macro-Block Dropout for Improved Regularization in Training End-to-End Speech Recognition Models
Proc. SLT, 2022.