Zhenyu Wang
AI/ML Researcher @ Apple
United States
Boston
Internet
deep learning, Speech Signal Processing, pattern recognition, Python, C++, Perl, Latex, Pytorch, Keras, kaldi, Shell, ASR, speaker recognition, Algorithm Analysis, Machine Learning
Experience

Research Assistant
United States
Speaker Recognition, Speech Recognition, Model Post-training, Audio Anti-spoofing, Speech Signal Processing, and Deep Learning @ CRSS The Ph.D. thesis was successfully defended on July 17th, 2024. The title is "Robust Text-independent Forensic Speaker Recognition Against Multiple Acoustic Domain Mismatches and Spoofing Attacks." Major publications: * Z. Wang, and John H. L. Hansen. Multi-source Domain Adaptation for Text-independent Forensic Speaker Recognition [J]. IEEE/ACM Transactions on Audio, Speech, and Language Processing, vol. 30, pp. 60-75, 2021. * Z. Wang and J. H. L. Hansen, Toward Improving Synthetic Audio Spoofing Detection Robustness via Meta-Learning and Disentangled Training With Adversarial Examples [J]. IEEE ACCESS, vol. 12, pp. 99894-99911, 2024.

TE/TRH/Fundamental Research Intern
Bellevue, Washington, United States
Audio Generation R&D with diffusion model and LLM. Responsibilities include data engineering/analysis, research, model design, implementation, data parallel distributed training, optimization, evaluation, results & demo presentations. The internship work titled as "Textually-Guided Audio Generation Using A Dual-Conditioned Latent Diffusion Model" was accepted by NeurIPS 2024 Workshop.

Research Scientist Intern (AI)
Seattle, Washington, United States
Customized keyword spotting system R&D. Responsibilities cover implementing the entire pipeline (data I/O, model design, implementation, parallel distributed training, evaluation), committing code to the FB codebase, and presentations. The internship work titled "Hardware-efficient Customized Keyword Spotting with Spectral-Temporal Graph Attentive Pooling and a hybrid Loss" accepted by INTERSPEECH 2024.

Research Scientist Intern (AI)
Seattle, Washington, United States
Robust small-footprint KWS system R&D @ AI Speech Team based on disentangled training with adversarial examples. Responsibilities include model architecture design, improvement and implementation, research, model training, and testing. This internship work, titled "Disentangled Training with Adversarial Examples For Robust Small-footprint Keyword Spotting", was accepted by ICASSP 2023.

AI Research Intern
Beijing, China
@ Speech Interactive Technology Team Online decoding platform development based on the Kaldi Nnet3 recipe for the real-time acoustic model assessment. Conducting research on ASR, including DNN pruning based speech recognition acoustic model iterative training and testing, revising the acoustic dictionary and language model for addressing issues introduced by English liaison and reduction in the mixed Chinese-English scenario, feature engineering, model architecture improvement, and training acceleration.
Zhenyu Wang's Contact Information
Phone
Find the Right Leads
Find Verified Contact Data
What LeadContact does well
Find verified emails, phone numbers, and decision-makers with 98% accuracy.
Find Leads
Find the right people by company, role, industry, location, and more.
925M+ professional profiles

Find Emails
Access verified email addresses for your target contacts.
657M+ emails

Find Phone Numbers
Get cross-validated phone data from multiple top sources.
239M+ phone numbers

More Accurate. Lower Cost.
Find contact data in 1 tool with 98% accuracy
LeadContact integrates leading enrichment tools to deliver more accurate contact data—without paying for each one.
Great conversations start with the right contact.
It’s time to find yours.






