张乐涵 Lehan Zhang

多模态大模型 · 具身智能

Multimodal LLMs · Embodied Intelligence

嗨,我是张乐涵。2026 年毕业于北京服装学院信息管理与信息系统专业,获北京市优秀毕业生称号。 目前正在备考中国科学院大学(UCAS)计算技术研究所人工智能专硕。

Hi, I am Lehan Zhang. I graduated in 2026 from Beijing Institute of Fashion Technology with a B.S. in Information Management & Information Systems, honored as a Beijing Outstanding Graduate. Currently preparing for the M.S. in AI at Institute of Computing Technology, UCAS.

研究方向涵盖多模态大模型与具身智能。我认为具身智能的大脑能力高度依赖多模态大模型, 二者相辅相成。同时通过 ysyx 项目学习处理器设计,探索从底层硬件到上层算法的完整技术栈。

My research spans multimodal LLMs and embodied intelligence. I believe the brain of embodied agents fundamentally relies on multimodal LLMs — the two go hand in hand. I also study processor design through the ysyx project, exploring the full technology stack from hardware to algorithms.

张乐涵 张乐涵 · Lehan Zhang

📢 动态

📢 News

2026.08开始参与 ysyx 项目,系统学习处理器设计Started the ysyx project on processor design
2026.08MRAFnd 论文上线 arXiv:2608.01430MRAFnd on arXiv: 2608.01430
2026.06本科毕业,获北京市优秀毕业生称号 🎓Graduated with Beijing Outstanding Graduate honors 🎓
2026.01聚焦多模态大模型与具身智能方向Focusing on multimodal LLMs & embodied intelligence
2025.10论文被 MMM 2026 接收 🎉Paper accepted at MMM 2026 🎉

📖 教育

📖 Education

2026 – 至今 2026 – Present
备考 中国科学院大学(UCAS)计算技术研究所 人工智能专硕 Preparing for M.S. in AI, Institute of Computing Technology, UCAS
2022 – 2026
北京服装学院 · 信息管理与信息系统 本科 BIFT · B.S. in Information Management & Information Systems 北京市优秀毕业生 Beijing Outstanding Graduate

🔬 研究方向

🔬 Research Interests

点击卡片展开细分方向

Click a card to expand

🧠

多模态大模型

Multimodal LLMs

多模态感知理解、跨模态生成与检索增强。研究如何融合视觉、语言等多种模态信息,构成具身智能的「大脑」基础。

Multimodal perception & understanding, cross-modal generation, and retrieval augmentation. Building the foundational "brain" for embodied intelligence by fusing vision, language, and other modalities.

+
🤖

具身智能

Embodied Intelligence

从芯片设计到系统架构再到算法策略的完整技术栈。以多模态大模型为大脑,驱动机器人感知、决策与行动。

Full technology stack from chip design to system architecture to algorithm strategies. Multimodal LLMs as the brain driving robot perception, planning, and action.

+

📝 论文

📝 Publications

多模态大模型 Multimodal LLMs

MRAFnd: Multimodal Retrieval-Augmented Framework for Zero-Shot Fake News Detection

感知Perception

Lehan Zhang, Yinlei Cheng, Shiqi Hu, Yiheng Zhou, Shangxi Li, Naidong Zhao

提出基于多模态检索增强的零样本假新闻检测框架,无需训练即可跨域识别虚假信息。

A multimodal retrieval-augmented framework for zero-shot fake news detection, enabling cross-domain identification without training.

✅ MMM 2026 [arXiv]
具身智能 Embodied Intelligence
Paimon

「前面的区域,我正在探索!」

"The area ahead — I'm exploring it!"

具身智能相关工作推进中,敬请期待。

Embodied intelligence work in progress — stay tuned.

🎨 兴趣爱好

🎨 Hobbies

编程 💻  ·  摄影 📷  ·  旅行 ✈️  ·  原神 🎮

Coding 💻  ·  Photography 📷  ·  Travel ✈️  ·  Genshin Impact 🎮

"Stay hungry, stay foolish."

"Stay hungry, stay foolish."