Jee-weon Jung
Senior Research Scientist, Apple
Speech deepfake detection · speaker recognition · speech LLMs

For over a decade I've worked on speech technology: keeping it trustworthy by detecting speech deepfakes and verifying who is really speaking, and making sense of conversations by tracking who spoke when. My teams have taken first place in international challenges, and my research spans industry and academia — Naver Clova, a postdoc at Carnegie Mellon, and now Senior Research Scientist at Apple. Today I'm increasingly focused on building speech LLMs to be personalized and to understand paralinguistics.

At a glance

Now — Senior Research Scientist, Apple.

Focus — speech LLMs, speech deepfake detection & speaker recognition.

Publications — 90+ at ICASSP & INTERSPEECH.

Community — co-organizer of the SASV, VoxSRC, ASVspoof 5 & WildSpoof challenges.

[email protected]

Trustworthy speech

Detecting speech deepfakes and verifying who is really speaking, so voice interfaces can be trusted.

Understanding conversations

Knowing who spoke when, and making sense of multi-speaker audio.

Speech LLMs

Personalized spoken-language models that understand paralinguistics.

97Publications
5,034Citations
31h-index
66i10-index
§ 01

Experience & Education

Sep 2024 — present

Senior Research Scientist

Apple
May 2023 — Sep 2024

Postdoctoral Research Associate

Carnegie Mellon University
Dec 2020 — Apr 2023

Research Scientist

Naver Clova
2017 — 2021

PhD, Computer Science & Engineering

University of Seoul
2013 — 2017

BS, Computer Science & Engineering · BBA

University of Seoul
§ 02

Selected Honors

1ST PLACEDCASE 2024 Challenge, Task 6 — automated audio captioning2024
1ST PLACEDCASE 2023 Challenge, Task 6-a — automated audio captioning2023
3RD PLACEThird DIHARD Challenge, Track 1 core — speaker diarization2021
§ 03

Selected Publications

  1. AASIST: Audio anti-spoofing using integrated spectro-temporal graph attention networks
    J. Jung, H. Heo, H. Tak, H. Shim, J. S. Chung, B. Lee, H. Yu, N. Evans
    Proc. ICASSP · 2022
  2. SpoofCeleb: Speech Deepfake Detection and SASV in the Wild
    J. Jung, Y. Wu, X. Wang, J.-H. Kim, S. Maiti, Y. Matsunaga, H. Shim, J. Tian, N. Evans, J. S. Chung, et al.
    IEEE Open Journal of Signal Processing · 2025
  3. ASVspoof 5: Crowdsourced speech data, deepfakes, and adversarial attacks at scale
    X. Wang, H. Delgado, H. Tak, J. Jung, H. Shim, M. Todisco, N. Evans, T. Kinnunen, et al.
    Computer Speech & Language · 2024
  4. SASV 2022: The First Spoofing-Aware Speaker Verification Challenge
    J. Jung, H. Tak, H. Shim, H. Heo, B. Lee, S. Chung, H. Yu, N. Evans, T. Kinnunen
    Proc. INTERSPEECH · 2022
  5. RawNet: End-to-end deep neural network using raw waveforms for text-independent speaker verification
    J. Jung, H. Heo, J. Kim, H. Shim, H. Yu
    Proc. INTERSPEECH · 2019
  6. ESPnet-SPK: Full pipeline speaker embedding toolkit with reproducible recipes and self-supervised front-ends
    J. Jung, W. Zhang, J. Shi, Z. Aldeneh, et al.
    Proc. INTERSPEECH · 2024
  7. OWSM v3.1: Better and faster open Whisper-style speech models based on E-Branchformer
    Y. Peng, J. Tian, W. Chen, J. Jung, et al.
    Proc. INTERSPEECH · 2024

Full list of 90+ publications available in the CV.

§ 04

Mentoring

I've had the privilege of working with extremely talented students, fostering their growth as researchers. These collaborations have produced impactful publications across diverse speech-processing tasks. I'm always excited to mentor new people — reach out if you'd like to explore working together.

Masao Someki2025profile ↗

Dynamic pruning of LLMs

  • Masao Someki, Shikhar Bharadwaj, Atharva Anand Joshi, Chyi-Jiunn Lin, Jinchuan Tian, Jee-weon Jung, Markus Müller, Nathan Susanj, Jing Liu, Shinji Watanabe, "Context-Driven Dynamic Pruning for Large Multi-Modal Foundation Model," Interspeech 2025.
Siddhant Arora2024 — 2025profile ↗

Spoken language understanding · spoken dialogue systems

  • Siddhant Arora, Jinchuan Tian, Hayato Futami, Jee-weon Jung, Jiatong Shi, Yosuke Kashiwagi, Emiru Tsunoo, Shinji Watanabe, "A Chain-of-Thought Reasoning Approach to E2E Spoken Dialogue Systems with an Open-Source Toolkit," Interspeech 2025.
  • Siddhant Arora, Ankita Pasad, Chung-Ming Chien, Jionghao Han, Roshan Sharma, Jee-weon Jung, Hira Dhamyal, William Chen, Suwon Shon, Hung-yi Lee, Karen Livescu, Shinji Watanabe, "On the Evaluation of Speech Foundation Models for Spoken Language Understanding," ACL Findings 2024.
  • Siddhant Arora, Hayato Futami, Jee-weon Jung, Yifan Peng, Roshan Sharma, Yosuke Kashiwagi, Emiru Tsunoo, Shinji Watanabe, "UniverSLU: Universal Spoken Language Understanding for Diverse Classification and Sequence Generation Tasks with a Single Network," NAACL 2024.
KiHyun Nam2023 — 2025profile ↗

Robust automatic speaker verification

  • KiHyun Nam, Jungwoo Heo, Jee-weon Jung, Gangin Park, Chaeyoung Jung, Ha-Jin Yu, Joon Son Chung, "SEED: Speaker Embedding Enhancement Diffusion Model," Interspeech 2025.
  • KiHyun Nam, Hee-Soo Heo, Jee-weon Jung, Joon Son Chung, "Disentangled Representation Learning for Environment-agnostic Speaker Recognition," Interspeech 2024.
  • KiHyun Nam, Youkyum Kim, Jaesung Huh, Hee-Soo Heo, Jee-weon Jung, Joon Son Chung, "Disentangled Representation Learning for Multilingual Speaker Recognition," Interspeech 2023.
Sung Hwan Mun2022 — 2023profile ↗

Speaker verification · spoofing-robust ASV

  • Sung Hwan Mun, Hye-jin Shim, Hemlata Tak, Xin Wang, Xuechen Liu, Md Sahidullah, Myeonghun Jeong, Min Hyun Han, Massimiliano Todisco, Kong Aik Lee, Junichi Yamagishi, Nicholas Evans, Tomi Kinnunen, Nam Soo Kim, Jee-weon Jung, "Towards Single Integrated Spoofing-aware Speaker Verification Embeddings," Interspeech 2023.
  • Sung Hwan Mun, Jee-weon Jung, Min Hyun Han, Nam Soo Kim, "Frequency and Multi-Scale Selective Kernel Attention for Speaker Verification," SLT 2022.
§ 05

Resources

Challenge & workshop organization

Datasets

§ 06

Contact

For collaboration, questions, or correspondence on speech, audio security, and machine learning:

[email protected]