Hi! I am Hao Lin (Chinese: ๆž—ๆตฉ), a third-year undergraduate student majoring in Software Engineering at the School of Software Engineering, Huazhong University of Science and Technology(HUST) (Rank: 2/116 , GPA: 93.06/100) .

I am currently exploring how multimodal and video foundation models can reason more reliably, generate more efficiently, and maintain consistent dynamic world states over long contexts.

๐Ÿ” Research

My academic exploration focuses on multimodal foundation models, video reasoning, reinforcement learning and efficient generative systems. Currently, my research is focused on:

  • Multimodal Large Language Models and Vision-Language Reasoning ๐Ÿ”ญ
  • Video Reasoning and Reinforcement Learning ๐ŸŽฌ
  • Efficient Video Generation and Inference โšก
  • World Models ๐ŸŒ

๐Ÿ”ฅ News

๐Ÿ“ Publications

CLIP negation paper figure

Not Just What's There: Enabling CLIP to Comprehend Negated Visual Descriptions Without Fine-tuning

Junhao Xiao, Zhiyu Wu, Hao Lin, Yi Chen, Yahui Liu, Xiaoran Zhao, Zixu Wang, Zejiang He

A plug-and-play framework for improving CLIP's understanding of negated visual descriptions while keeping the CLIP backbone frozen.

VISD paper figure

VISD: Enhancing Video Reasoning via Structured Self-Distillation

Hao Lin, Kunyang Lv, Xu Jiang, Jingqi Tian, Zhongjing Du, Jiayu Ding, Qiaoman Zhang, Hongbo Jin

A structured on-policy self-distillation framework for video reasoning RLVR, using teacher feedback on student rollouts for fine-grained credit assignment.

Focused Forcing paper figure

Focused Forcing: Content-Aware Per-Frame KV Selection for Efficient Autoregressive Video Diffusion

Peiliang Cai, Evelyn Zhang, Jiacheng Liu, Hao Lin, Ruiqi Zhang, Weile Mo, Yue Ma, Shikang Zheng, Jiehang Huang, Dongrui Liu, Linfeng Zhang

A training-free KV selection method that focuses cached history along generated-frame and attention-head dimensions for faster long video generation.

Domino paper figure

Domino: Decoupling Causal Modeling from Autoregressive Drafting in Speculative Decoding

Jianuo Huang, Yaojie Zhang, Qituan Zhang, Hao Lin, Hanlin Xu, Linfeng Zhang

A speculative decoding framework that decouples causal dependency modeling from autoregressive drafting, improving draft quality while keeping parallel drafting efficient.

MemeSleuth-Bench paper figure

MemeSleuth-Bench: Can Models Detect Chinese Internet Meme Origins Through Web Retrieval?

Shengjie Xu, Tianyi Wang, ..., Hao Lin, Mengran Zhu, Zhenghao Gao, Chengrui Hu, Zehua Lyu

A benchmark for evaluating whether multimodal models can trace Chinese internet meme origins through web retrieval and culturally grounded evidence seeking.

๐ŸŽ– Honors and Awards

  • 2026.03: ๐ŸŽ–๏ธ FiberHome Telecommunication Scholarship(็ƒฝ็ซ้€šไฟกไผไธšๅฅ–ๅญฆ้‡‘),FiberHome
  • 2025.11: ๐Ÿฅˆ National Second Prize, Challenge Cup โ€œAI+โ€ Special Competition
  • 2025.09: ๐ŸŽ–๏ธ National Scholarship(ๅ›ฝๅฎถๅฅ–ๅญฆ้‡‘),Ministry of Education of China
  • 2025.09: ๐ŸŽ–๏ธAcademic Excellence Scholarship,HUST
  • 2025.08: ๐Ÿฅ‡ National First Prize, China Robotics and Artificial Intelligence Competition, AI Innovation Track
  • 2025.08: ๐Ÿฅ‰ National Third Prize, China Robotics and Artificial Intelligence Competition, AI Innovation Track
  • 2025.06: ๐Ÿฅ‰ National Third Prize, Lanqiao Cup AI Practical Competition
  • 2024.12: ๐ŸŽ–๏ธ National Scholarship(ๅ›ฝๅฎถๅฅ–ๅญฆ้‡‘),Ministry of Education of China
  • 2024.11: ๐Ÿฅ‡ First Prize in Hubei Province, Chinese Mathematics Competition(CMC)
  • 2024.09: ๐ŸŽ–๏ธMerit Student Scholarship,HUST
  • 2024.03: ๐ŸŽ–๏ธFreshman Self-Reliance Scholarship,HUST

๐Ÿ’ป Experience

Baidu PaddleOCR

Internship, 2025.07 - 2025.08

Topic: Multimodal Document Understanding and OCR Benchmarking

๐Ÿ“– Educations

2023.09 - Now

Undergraduate, School of Software Engineering, Huazhong University of Science and Technology

Major: Software Engineering