I am currently a Staff Researcher/Tech Lead at Alibaba ATH, where I lead the Speech and Omni LLM Applied Research team. Previously, I was a researcher at the Mohamed bin Zayed University of Artificial Intelligence (MBZUAI), focusing on multilingual and multimodal large language models. I earned my Ph.D. in Machine Learning from Dublin City University's ML-Labs in 2023, following a Bachelor of Engineering from Northeastern University in China in 2018.

My research lies primarily in natural language processing, with a focus on LLMs across multilingual, multimodal, and speech settings, as well as LLM agents. My research work includes CoQuIR, selected for an ACL 2026 SAC Highlights Award, Marco-MoE, Trust No Tool, and Crayotter, alongside papers accepted at COLM 2026, CVPR 2026, and ICASSP 2026.

GPT4Video was nominated for the Best Paper Award at ACM-MM 2024. My open-source projects and contributions have collectively earned 4k+ GitHub stars, and I won two championships and two runner-up prizes in the IWSLT 2025 speech translation competition. Prior to my current role, I gained extensive research experience through positions as a research assistant and visiting scholar at Tencent AI Lab, the National Institute of Informatics (NII), and IBM Research-China.

I have received several personal honors, including the German DAAD AInet Fellowship, the 2023 Young AI Role Model of the Year award at the Irish AI Awards, and an SFI PhD Scholarship. My research has also been covered by RTÉ, Slator, and Irish Tech News, including a podcast interview on LLMs.

Portrait of Chenyang Lyu
Photo: X / Twitter profile

Research

Speech & Omni Models

Speech, vision, video and audio intelligence across understanding and generation.

LLM Agents and Reliability

Traceable multi-agent workflows, tool-feedback defense and robust evaluation.

Multilingual Models

Efficient model adaptation and culturally grounded intelligence across languages.


Education


Industry & Research Experience

Alibaba logo

Staff Researcher / Tech Lead

Alibaba ATH · Speech and Omni LLM Applied Research

Visiting & Research Roles

Tencent logo
Tencent AI LabResearch Assistant / Visiting Scholar
National Institute of Informatics logo
National Institute of InformaticsVisiting Scholar

Research Internships

Huawei logo
Huawei Noah’s Ark LabResearch Intern · 2020–2021
IBM logo
IBM Research-ChinaResearch Intern · 2018

News

🏆 CoQuIR was selected for an ACL 2026 SAC Highlights Award.

🎉 Spurious Rewards Paradox was accepted to ICML 2026.

🎉 ElasticFormer was accepted to CVPR 2026.

🎉 Marco-Voice, LongSpeech and MECap-R1 were accepted to ICASSP 2026.

🎤 Invited industry expert talk on Marco Models at ACM Multimedia Asia 2025.

🏆 Our team secured two championships and two runner-up prizes at IWSLT 2025.

🎉 Four papers on multilingual LLMs and hallucination detection were accepted to ACL 2025.

🎙️ CVQA was featured in a Microsoft Research podcast on culturally aware and linguistically diverse multimodal evaluation.


Selected Publications

* denotes corresponding or equal contribution. See Google Scholar and DBLP for the complete record.

ACL
2026
CoQuIR: A Comprehensive Benchmark for Code Quality-Aware Information Retrieval

Jiahui Geng, Fengyu Cai, Shaobo Cui, Qing Li, Liangwei Chen, Chenyang Lyu, et al.

Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics.

COLM
2026
Marco-MoE: Open Multilingual Mixture-of-Expert Language Models with Efficient Upcycling

Fan Jiang, Yu Zhao, Chenyang Lyu, Tianqi Shi, Yichao Du, et al.

Conference on Language Modeling.

CVPR
2026
ElasticFormer: Detecting Objects in HRW Shots via Elastic Computing Vision Transformer

Xiang Li, Wenxi Li, Yuetong Wang, Chenyang Lyu, et al.

IEEE/CVF Conference on Computer Vision and Pattern Recognition.

ICASSP
2026
LongSpeech: A Scalable Benchmark for Transcription, Translation and Understanding in Long Speech

Fei Yang, Xuanfan Ni, Renyi Yang, Jiahui Geng, Qing Li, Chenyang Lyu*, et al.

IEEE International Conference on Acoustics, Speech, and Signal Processing.

ICASSP
2026
Marco-Voice Technical Report

Fengping Tian, Chenyang Lyu, Xuanfan Ni, Haoqin Sun, Qingjuan Li, et al.

IEEE International Conference on Acoustics, Speech, and Signal Processing.

ACM MM
2024
GPT4Video: A Unified Multimodal Large Language Model for Instruction-Followed Understanding and Safety-Aware Generation

Zhanyu Wang, Longyue Wang, Zhen Zhao, Minghao Wu, Chenyang Lyu, et al.

Proceedings of the 32nd ACM International Conference on Multimedia.

NeurIPS
2024
CVQA: Culturally-diverse Multilingual Visual Question Answering Benchmark

David Romero*, Chenyang Lyu*, Haryo Akbarianto Wibowo, Teresa Lynn, et al.

NeurIPS Datasets and Benchmarks Track.

View the complete publication list →