Skip to content
View Mubuky's full-sized avatar

Block or report Mubuky

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Mubuky/README.md

Hi there 👋, I'm Mingzhe Li!

  • 🔭 I am a Ph.D. student jointly at the College of Computer Science and Artificial Intelligence, Fudan University, and the Shanghai Innovation Institute, advised by Prof. Xipeng Qiu. Before that, I received my B.Eng. degree in Artificial Intelligence from Harbin Institute of Technology, where I worked with Prof. Yanyan Zhao.
  • 🌱 I am broadly interested in Natural Language Processing and Machine Learning. My current research focuses on Reinforcement Learning, Self-Evolving Agent and Synthetic Data Generation. I am particularly interested in leveraging reinforcement learning and its derivative techniques to stimulate the self-evolving capabilities of LLM-based agents in real-world environments.
  • 💬 If you are interested in any form of academic collaboration, please feel free to email me at mzli@ir.hit.edu.cn.

Education

Mingzhe Li's GitHub Streak

Pinned Loading

  1. Self-Foveate Self-Foveate Public

    [Findings of ACL 2025] Self-Foveate: Enhancing Diversity and Difficulty of Synthesized Instructions from Unsupervised Text via Multi-Level Foveation

    Python 24

  2. tongjingqi/AI-Can-Learn-Scientific-Taste tongjingqi/AI-Can-Learn-Scientific-Taste Public

    We propose Reinforcement Learning from Community Feedback (RLCF), a training paradigm that uses large-scale community signals as supervision, and formulate scientific taste learning as a preference…

    423 11

  3. tongjingqi/Thinking-with-Video tongjingqi/Thinking-with-Video Public

    We introduce 'Thinking with Video', a new paradigm leveraging video generation for multimodal reasoning. Our VideoThinkBench shows that Sora-2 surpasses GPT5 by 10% on eyeballing puzzles and reache…

    Python 317 5

  4. realZillionX/InspireSkill realZillionX/InspireSkill Public

    启智平台(qz.sii.edu.cn)的 Agent 驾驶舱:Skill + CLI,一条命令直达。Agent cockpit for the Inspire ML platform — one command, every operation, straight from chat.

    Python 187 19