Our Digital Life Podcast: A series by IEEE-SPS
Our Digital Life Podcast: A series by IEEE-SPS
Podcast Description
As the world's largest professional organization, IEEE plays a significant role in enhancing the quality of our lives. Specifically, the IEEE signal processing society or SPS focuses on research and development of audio and speech processing, biomedical analysis, and wireless communication technologies, all of which are key enablers to today's modern society. In this series, we explore more about the works of signal processing and engage with various global speakers.
Podcast Insights
Content Themes
The podcast centers around themes of audio and speech processing, biomedical analysis, and advancements in wireless communication technologies. Specific episodes delve into topics like AI-driven methods in medical imaging, with featured discussions on how innovations improve diagnostic accuracy and therapeutic approaches, as well as challenges in data constraints and solutions like FDA-approved imaging algorithms.

As the world’s largest professional organization, IEEE plays a significant role in enhancing the quality of our lives. Specifically, the IEEE signal processing society or SPS focuses on research and development of audio and speech processing, biomedical analysis, and wireless communication technologies, all of which are key enablers to today’s modern society. In this series, we explore more about the works of signal processing and engage with various global speakers.
In this episode, Prof. Hung-yi Lee, Professor at National Taiwan University, interviews Dr. Jinyu Li, Partner Applied Science Manager at Microsoft. Their conversation traces the evolution of speech translation from traditional cascaded pipelines to today's end-to-end, large language model (LLM)-powered systems.
Jinyu Li
Dr. Jinyu Li is Partner Applied Science Manager at Microsoft, where he leads a science team advancing speech, translation, and language technologies. He is an IEEE Fellow and an ISCA Fellow for contributions to deep-learning-based speech technology innovation, adoption, and commercialization. He has served on the IEEE Speech and Language Processing Technical Committee and as Vice Chair starting in 2026. He previously served as Associate Editor of IEEE/ACM Transactions on Audio, Speech, and Language Processing and is a Distinguished Industry Speaker for the IEEE Signal Processing Society. He has also been awarded the IEEE SPS Best Paper Award (2025), the APSIPA Industrial Distinguished Leader Award (2021), and the APSIPA Sadaoki Furui Prize Paper Award (2023).
In this episode, Dr. Li discusses the evolution of AI-powered speech translation, from conventional cascaded systems to modern end-to-end architectures, while highlighting the role of large language models in enabling real-time multilingual communication. He also explores the latest applications of speech translation and examines the technical challenges that remain.

Disclaimer
This podcast’s information is provided for general reference and was obtained from publicly accessible sources. The Podcast Collaborative neither produces nor verifies the content, accuracy, or suitability of this podcast. Views and opinions belong solely to the podcast creators and guests.
For a complete disclaimer, please see our Full Disclaimer on the archive page. The Podcast Collaborative bears no responsibility for the podcast’s themes, language, or overall content. Listener discretion is advised. Read our Terms of Use and Privacy Policy for more details.