Zhifeng Kong

Zhifeng Kong

Senior Research Scientist @ NVIDIA

About

I am a research scientist at NVIDIA. I work on audio understanding in the context of multi modalities including language, audio, and vision. I also work on deep generative models for audio, with a focus on audio diffusion models. Before I joined NVIDIA, I obtained a Ph.D. in Computer Science at UC San Diego advised by Professor Kamalika Chaudhuri. During my Ph.D., I have conducted research on understanding deep generative models, including their expressivity, reliability, controllability, and trustworthiness. I hold a B.S. in mathematics and applied mathematics from Xi'an Jiaotong University, China, in which I was luckily enrolled in the Honors Youth Program (少年班) and the National Honors Science Program (珠峰计划). My Erdős number is 3.

Country

United States

City

Santa Clara

Industry

Computer Software

Skill

Deep Learning Research, 人工智能, 机器学习, 数据分析, Python, MATLAB, Mathematica, LaTeX, 深度学习, 数学建模

Experience

NVIDIA

Senior Research Scientist

NVIDIA

LinkedIn
2025-3 - Present · 1 yr 7 mos

美国 加利福尼亚 圣塔克拉拉

NVIDIA

Research Scientist

NVIDIA

LinkedIn
2023-8 - 2025-3 · 1 yr 8 mos

美国 加利福尼亚州 圣克拉拉

Research scientist in the Applied Deep Learning Research

UC San Diego

PHD Student

UC San Diego

LinkedIn
2018-9 - 2023-6 · 4 yrs 10 mos

La Jolla

I work on deep generative models, including Diffusion Model, GAN, Normalizing Flow, and VAE. Specifically, I build state-of-the-art generative (diffusion) models for different applications. Besides diffusion models, I conduct research on interpretability, privacy, post-editing, and expressivity of deep generative models. I am also widely interested in conditional and multi-modal generative models.

NVIDIA

Research Intern

NVIDIA

LinkedIn
2022-6 - 2022-9 · 4 mos

美国 加利福尼亚州 圣克拉拉

Conduct research on speech denoising and music generation.

NVIDIA

Research Intern

NVIDIA

LinkedIn
2021-6 - 2021-9 · 4 mos

美国 加利福尼亚州 圣克拉拉

Conduct research on speech denoising: Zhifeng Kong, Wei Ping, Ambrish Dantrey, Bryan Catanzaro. Speech Denoising in the Waveform Domain with Self-Attention. In ICASSP 2022.

Baidu USA

Research Intern

Baidu USA

LinkedIn
2020-6 - 2020-9 · 4 mos

美国 加利福尼亚州 桑尼维尔

Conduct research on speech synthesis. Zhifeng Kong, Wei Ping, Jiaji Huang, Kexin Zhao, Bryan Catanzaro. DiffWave: A Versatile Diffusion Model for Audio Synthesis. In ICLR 2021 (oral).

Education

UC San Diego

UC San Diego

LinkedIn

Computer Science and Engineering

2018 - 2023 · 5 yrs

Ph.D. Student Research topic: deep generative models theory and application

Xi'an Jiaotong University

Xi'an Jiaotong University

LinkedIn

数学与应用数学(试验班)

2014 - 2018 · 4 yrs
Georgia Institute of Technology

Georgia Institute of Technology

LinkedIn

Mathematics

2017 - 2017

Jan. 2017 - May. 2017

University of Alberta

University of Alberta

LinkedIn

Summer Visiting

2016 - 2016

July 2016 - Aug. 2016

Xi'an Jiaotong University

Xi'an Jiaotong University

LinkedIn

少年班

2012 - 2014 · 2 yrs

Zhifeng Kong's Contact Information

Email

******@***.com

Phone

(**) *** ****

Find the Right Leads
Find Verified Contact Data

Try with: Jensen Huang @ nvidia.com Click to autofill
LeadContact awards, five-star ratings, and GDPR compliance badges

What LeadContact does well

Find verified emails, phone numbers, and decision-makers with 98% accuracy.

Find Leads

Find Leads

Find the right people by company, role, industry, location, and more.

925M+ professional profiles

Find Leads
Find Emails

Find Emails

Access verified email addresses for your target contacts.

657M+ emails

Find Emails
Find Phone Numbers

Find Phone Numbers

Get cross-validated phone data from multiple top sources.

239M+ phone numbers

Find Phone Numbers

More Accurate. Lower Cost.

Find contact data in 1 tool with 98% accuracy

LeadContact integrates leading enrichment tools to deliver more accurate contact data—without paying for each one.

LeadContact Logo
Competitor Tools

All these = $289 per month

Great conversations start with the right contact.

It’s time to find yours.