Homepage

I am a CS PhD student at the School of Computing, National University of Singapore. I am supervised by Prof. Mong-Li Lee (Director of CTIC) and Prof. Wynne Hsu (Director of IDS) at Center for Trusted Internet and Community, and I also work with Dr. Hao Fei, Dr. Shengqiong Wu, Dr. Bobo Li, Dr. Hongzhan Lin and Dr. Tianjie Ju.

Prior to this, I received my master’s degree from NUS and my bachelor’s degree from Wuhan University, where I also completed a minor in Business Administration as part of the Ziqiang Entrepreneurship Program.

My research interest includes Bridging Physical and Mental Worlds toward Human-Like Intelligence through Multimodal (Video) Understanding, Reasoning, and Generation.

I am always exploring new collaboration opportunities. I am always happy to discuss potential collaborations — feel free to drop me an email at mluo@u.nus.edu.

News

  • Accepted at NeurIPS 2026

    From Finding to Linking: Benchmarking and Advancing Cross-Long-Video Reasoning for Multimodal LLMs

  • Will Release a Survey on Video World Model

    Awesome-Video-World-Model

  • Release a Survey on AI for Games

    AI for Games in the Foundation Model Era

  • Accepted at EMNLP (Findings) 2026

    RIDGE: Region-Informed Derivative-Guided Evidence Selection for Long Video Understanding

  • Accepted at EMNLP (Findings) 2026

    OmniHallu: Unified Hallucination Detection for Cross-Modal Comprehension and Generation in Multimodal Large Language Models

  • Accepted at ECCV 2026

    From Evaluation to Enhancement: Benchmarking and Improving Think-with-Video Reasoning for Video Generative Models

  • Accepted at ECCV 2026

    No Place to Hide: Benchmarking Video Hallucination with Background-Controlled Pairs

  • Accepted at ICML 2026 DL4C Workshop

    Nexus: Execution-Grounded Multi-Agent Test Oracle Synthesi

  • Accepted at IJCV

    Dr.V: A Hierarchical Perception-Temporal-Cognition Framework to Diagnose Video Hallucination by Fine-grained Spatial-Temporal Grounding

  • Accepted at ICLR 2026

    Unveiling the Cognitive Compass: Theory-of-Mind-Guided Multimodal Emotion Reasoning

2025 and earlier (17 updates)Hide earlier updates
  • Accepted at TIFS

    Poisoning Attacks to Knowledge Distillation-based Federated Learning under Robust Aggregation Rules

  • Accepted at ACL 2025 FEVER Workshop

    EMULATE: A Multi-Agent Framework for Determining the Veracity of Atomic Claims by Emulating Human Actions

  • Accepted at ACL 2025 (Oral)

    Aristotle: Mastering Logical Reasoning with A Logic-Complete Decompose-Search-Resolve Framework

  • Accepted at ICML 2025 (Oral, Spotlight)

    On Path to Multimodal Generalist: Levels and Benchmarks

  • Accepted at ICML 2025

    VistaDPO: Video Hierarchical Spatial-Temporal Direct Preference Optimization for Large Video Models

  • Accepted at ICML 2025

    SWIFTCODE: Enhancing Code Generation in Large Language Models through Efficiency-Aware Fine-tuning

  • Co-organizing a Grand Challenge at ACM MM 2025

    Multimodal Conversational Aspect-based Sentiment Analysis (MCABSA 2025)

  • Co-organizing a Workshop at ACM MM 2025

    The 1st Cognition-oriented Multimodal Affective and Empathetic Computing (CogMAEC 2025) Workshop

  • Accepted at ICLR 2025

    PAD: Personalized Alignment at Decoding-Time

  • Accepted at WWW 2025

    Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark

  • New Paper Published on arxiv

    A Survey on Benchmarks of Multimodal Large Language Models

  • Accepted at ACM MM Workshop (MIS24) (Best Paper Award)

    Fine-grained Structural Hallucination Detection for Unified Visual Comprehension and Generation in Multimodal LLM

  • Accepted at ACM MM 2024 (Oral)

    PanoSent: A Panoptic Sextuple Extraction Benchmark for Multimodal Conversational Aspect-based Sentiment Analysis

  • 2nd Place at SemEval-2024

    NUS-Emo at SemEval-2024 Task 3: Instruction-Tuning LLM for Multimodal Emotion-Cause Analysis in Conversations

  • Accepted at TDSC

    Towards Class-Balanced Privacy Preserving Heterogeneous Model Aggregation

Selected Works

Google Scholar

From Finding to Linking: Benchmarking and Advancing Cross-Long-Video Reasoning for Multimodal LLMs

Meng Luo, Zikang Zhou, Shanqing Xu, Shize Zhang, Bobo Li, Hao Fei, Mong-Li Lee, Wynne Hsu

NeurIPS 2026

AI for Games in the Foundation Model Era

Meng Luo, Yanlin Li, Hao Li, Hongzhan Lin, Pengfei Zhou, Tianjie Ju, Ran Zhang, Yeying Jin, Mong-Li Lee, Wynne Hsu

arXiv 2026 HF Daily Paper #2

From Evaluation to Enhancement: Benchmarking and Improving Think-with-Video Reasoning for Video Generative Models

Meng Luo, Yicheng Liu, Jiahao Wang, Yuanxing Zhang, Xin Tao, Pengfei Wan, Kun Gai, and Hao Fei

ECCV 2026

Unveiling the Cognitive Compass: Theory-of-Mind-Guided Multimodal Emotion Reasoning

Meng Luo, Bobo Li, Shanqing Xu, …, Hao Fei, Mong-Li Lee, Wynne Hsu

ICLR 2026

Dr.V: A Hierarchical Perception-Temporal-Cognition Framework to Diagnose Video Hallucination by Fine-grained Spatial-Temporal Grounding

Meng Luo, Shengqiong Wu, Liqiang Jing, …, Jiebo Luo, William Yang Wang, Hao Fei, Mong-Li Lee, Wynne Hsu

IJCV 2026

RIDGE: Region-Informed Derivative-Guided Evidence Selection for Long Video Understanding

Shanqing Xu, Meng Luo, Mengchen Qian, …, Xiaojin Zhang, Zhongyu Wei, Wei Chen, Xiang Bai

EMNLP (Findings) 2026

OmniHallu: Unified Hallucination Detection for Cross-Modal Comprehension and Generation in Multimodal Large Language Models

Jianjiang Yang, Peihang Li, Shanqing Xu, Mengchen Qian, Lu Zhang, Meng Luo*

EMNLP (Findings) 2026

No Place to Hide: Benchmarking Video Hallucination with Background-Controlled Pairs

Haojian Huang, Harold Haodong Chen, Meng Luo, Junjia Du, Shanqing Xu, Ziheng Chen, Yanxiang Huang, Yinchuan Li, Ying-Cong Chen

ECCV 2026

On Path to Multimodal Generalist: General-Level and General-Bench

Hao Fei, Yuan Zhou, Juncheng Li, Xiangtai Li, …, Meng Luo, Jiebo Luo, Tat‑Seng Chua, Hanwang Zhang, Shuicheng Yan

ICML 2025 Oral · Spotlight

VistaDPO: Video Hierarchical Spatial-Temporal Direct Preference Optimization for Large Video Models

Haojian Huang, Haodong Chen, Shengqiong Wu, Meng Luo, Jinlan Fu, Xinya Du, Hanwang Zhang, Hao Fei

ICML 2025

Aristotle: Mastering Logical Reasoning with A Logic-Complete Decompose-Search-Resolve Framework

Jundong Xu, Hao Fei, Meng Luo, Qian Liu, Liangming Pan, William Yang Wang, Preslav Nakov, Mong-Li Lee, Wynne Hsu

ACL 2025 Oral

PanoSent: A Panoptic Sextuple Extraction Benchmark for Multimodal Conversational Aspect-based Sentiment Analysis

Meng Luo, Hao Fei, Bobo Li, Shengqiong Wu, Qian Liu, Soujanya Poria, Erik Cambria, Mong-Li Lee, Wynne Hsu

ACM MM 2024 Oral

SWIFTCODE: Enhancing Code Generation in Large Language Models through Efficiency-Aware Fine-tuning

Dong Huang, Guangtao Zeng, Jianbo Dai, Meng Luo, Han Weng, Yuhao Qing, Heming Cui, Zhijiang Guo, Jie M. Zhang

ICML 2025

PAD: Personalized Alignment at Decoding-Time

Ruizhe Chen, Xiaotian Zhang, Meng Luo, Wenhao Chai, Zuozhu Liu

ICLR 2025

Towards Multimodal Empathetic Response Generation: A Rich Text-Speech-Vision Avatar-based Benchmark

Han Zhang, Zixiang Meng, Meng Luo, Hong Han, Lizi Liao, Erik Cambria, Hao Fei

WWW 2025

NUS-Emo at SemEval-2024 Task 3: Instruction-Tuning LLM for Multimodal Emotion-Cause Analysis in Conversations

Meng Luo, Han Zhang, Shengqiong Wu, Bobo Li, Hong Han, Hao Fei

SemEval @ ACL 2024 2nd Place

Fine-grained Structural Hallucination Detection for Unified Visual Comprehension and Generation in Multimodal LLM

Hao Fei, Meng Luo, Jundong Xu, Shengqiong Wu, Wei Ji, Mong-Li Lee, Wynne Hsu

ACM MM Workshop 2024 Best Paper Award

A Survey on Benchmarks of Multimodal Large Language Models

Jian Li, Weiheng Lu, Hao Fei, Meng Luo, Ming Dai, Min Xia, Yizhang Jin, Zhenye Gan, Ding Qi, Chaoyou Fu, Ying Tai, Wankou Yang, Yabiao Wang, Chengjie Wang

arXiv 2024

Professional Activity

Multi-Modality Research Intern (青云计划)

Tencent IEG · Singapore

Video Generation Algorithm Research Intern

Kling Team

Mentored by Xintao Wang and Jiahao Wang.

Academic Service and Honors

  • Reviewer for TPAMI, NeurIPS, ICLR, ICML, CVPR, ICCV, ACL, ACM MM, AAAI, WWW, ECCV, EMNLP, Neurocomputing, ACM TOMM, KBS, ACM TALLIP, and various workshops.

Honors and Awards

During my undergraduate studies

Academic Achievements

  • Huawei Scholarship, Wuhan University (Top 5%)
  • First-class Excellence Scholarship, Wuhan University (Ranked 2nd)
  • Merit Student, Wuhan University (Top 10%)
  • Outstanding Student, Wuhan University

Competitions and Recognitions

  • Silver Award, Hubei Challenge Cup, Wuhan University
  • Gold Award, Ziqiang Cup College, Wuhan University
  • National First Prize, Citi Cup Financial Innovation Application Contest
  • Bole Award, ByteTop Summit Project, ByteDance
  • Top 10 Book Ambassador, Wuhan University Library
  • The First Prize, HP Dream Factory Innovation Hackathon Wuhan Station, HP

Leadership and Social Activities

  • Chairman, Wuhan University Campus Ambassador, ByteDance
  • Excellent Campus Ambassador, WePie Team
  • Online Course on Interdisciplinary Communication, University of Cambridge
Research figure

Open image