Welcome to my personal website. I’m Lu Hongliang (卢红亮), and I received an M.E. in Mechanical Engineering from Peking University (2023–2026), where I was advised by Prof. Zaiwen Wen. I also hold a B.E. in Robotics Engineering from PKU (2019–2023).

My research lies at the intersection of Reinforcement Learning and Large Language Models, with a focus on RL for LLMs, Agentic RL, self-evolving and multi-agent systems, and LLMs for optimization modeling and decision-making. My recent work has been accepted at NeurIPS, ICLR, and ICML. I am currently at Tencent, where I focus on enhancing foundation models’ coding capabilities and their ability to complete long-horizon tasks.

news

Sep 25, 2026 Our paper “PEARL: Solver-in-the-Loop Interactive Optimization Modeling from Natural Language” has been accepted to NeurIPS 2026 Main Track as a poster! 🎉
Jul 30, 2026 We have released our technical report “Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents” on arXiv! Learn more on the Qwen-UI-Agent project page. 🤖
May 14, 2026 Our paper “PEARL: Solver-in-the-Loop Interactive Optimization Modeling from Natural Language” is now available on arXiv! We have also open-sourced the code and released the PEARL-Qwen3-4B-Instruct-2507 model on Hugging Face. 🔓
May 01, 2026 Our paper “Constructing Industrial-Scale Optimization Modeling Benchmark” has been accepted to ICML 2026! 🎉
Jan 28, 2026 Our paper “Search Self-Play: Pushing the Frontier of Agent Capability without Supervision” has been accepted to ICLR 2026! 🎉

education

Peking University · College of Engineering

2023 — 2026
M.E. in Mechanical Engineering

Peking University · College of Engineering

2019 — 2023
B.E. in Robotics Engineering

experience

Alibaba Group · Tongyi Lab

Mar 2026 — Jun 2026
Research Intern (校招提前实习)

Alibaba Group · QuarkLLM

May 2025 — Sep 2025
Research Intern

Moonshot AI · RL Team

Jan 2025 — May 2025
Research Intern

selected publications

  1. Technical Report
    qwen_ui_agent_overview.png
    Qwen-UI-Agent Technical Report: Toward Next-Generation Real-World Centric Foundation GUI Agents
    Hanzhang Zhou*, Panrong Tong*, Xu Zhang* , and 24 more authors
    arXiv preprint arXiv:2607.28227, 2026
  2. NeurIPS 2026
    pearl_pipeline.png
    PEARL: Solver-in-the-Loop Interactive Optimization Modeling from Natural Language
    Hongliang Lu*, Zhong Li*, Yuxuan Chen , and 3 more authors
    arXiv preprint arXiv:2607.18256, 2026
  3. ICLR 2026
    ssp_pipeline.png
    Search Self-Play: Pushing the Frontier of Agent Capability without Supervision
    Hongliang Lu*, Yuhang Wen*, Pengyu Cheng , and 7 more authors
    The Fourteenth International Conference on Learning Representations, 2026
  4. ICML 2025
    optmath_pipeline.png
    OptMATH: A Scalable Bidirectional Data Synthesis Framework for Optimization Modeling
    Hongliang Lu*, Zhonglin Xie*, Yaoyu Wu , and 3 more authors
    Forty-Second International Conference on Machine Learning, 2025

latest posts