About Me
Hello! My name is Xiaoqin Feng. I am an M.S. student in Artificial Intelligence at the University of Southern California (USC), graduating in May 2027, and an AI Engineer Intern at Wyze. I build applied AI systems spanning real-time conversational audio agents, LLM-based agentic systems, generative speech, large-scale data pipelines, model evaluation, and reliable backend services. Previously, I was a Research Assistant at USC and spent six years at Mobvoi, most recently as a Tech Lead. My interests center on LLM agents, speech and NLP, multimodal learning, and human–AI collaboration.
AI Product: Audio Story Visualizer 👨💻 Indie Dev
Building an AI-powered audio visualization and dubbing platform that turns music, podcasts, and audiobooks into editable videos with high-fidelity speech processing. Try demo ▶
Education
- [Sep. 2025 – May 2027 (expected)] M.S. in Artificial Intelligence at University of Southern California (USC)
- [Sep. 2016 – May 2019] M.S. in Software Engineering at Beijing University of Technology (BJUT)
- [Sep. 2012 – Jul. 2016] B.E. in Computer Science at Southwest Minzu University (SMU); ranked Top 5 among 154 students
Professional Experience
- [Jul. 2026 – Present] AI Engineer Intern, Wyze. Developing and integrating core components for real-time conversational audio agents on edge devices, including model integration, automated testing, language identification, and speech-to-intent optimization.
- [Oct. 2025 – May 2026] Research Assistant, University of Southern California. Researched computer-use and research agents, with a focus on tool-use planning, execution robustness, and failure modes.
- [Aug. 2025 – Mar. 2026] Part-time Intern, Mobvoi. Optimized production speech models and LLM-based multi-agent workflows, and supported data pipelines and system integration.
- [May 2023 – Jul. 2025] Tech Lead, Mobvoi. Led production LLM agent frameworks, speech-LLM models and services, multimodal data engineering, and model evaluation systems; mentored junior engineers and interns.
- [Jul. 2019 – May 2023] Speech Algorithm Engineer, Mobvoi. Built multilingual speech and NLP models, scalable backend services, and standardized data pipelines for production systems serving tens of millions of users.
- [Aug. 2018 – Dec. 2018] Research Intern, TAL Education Group. Developed graph-enhanced deep knowledge tracing models and co-authored work published at AIED 2019.
Publications [Google Scholar]
-
SparkTTS
arXiv preprint 2025, 22 pages -
APIN
Applied Intelligence, 2024, 20 pages -
APSIPA
APSIPA ASC, 2023, 5 pages -
PRML
-
AIED
In proceedings of AIED 2019, Springer, Cham, 5 pages -
Journal of Physics
Journal of Physics: Conference Series. JPCS 2019, 6 pages -
CCIOT
-
CCIS
In proceedings of CCIS 2018, 5 pages