I am a third-year undergraduate student in the AI Elite Class at the College of Computer Science and Artificial Intelligence, Fudan University.
My research interests lie in Large Language Models (LLMs), AI Agents, and Interpretability. I am currently working as a research intern at the Fudan NLP Group, led by Professor Qi Zhang and Associate Professor Tao Gui.
“The only thing we have to fear is fear itself—nameless, unreasoning, unjustified terror which paralyzes needed efforts to convert retreat into advance.”

Submitted to EMNLP 2026 Under Review
NovGauge is a human-anchored benchmark for diagnosing LLM-based scientific novelty assessment across task, problem, and method dimensions. Evaluation of 18 LLMs on 619 paper pairs and 50 multi-paper sets shows that faithfulness verification reduces most models' raw F1 by more than half, exposing substantial hallucination and unsupported-evidence failures.
Submitted to EMNLP 2026 Under Review
NovGauge is a human-anchored benchmark for diagnosing LLM-based scientific novelty assessment across task, problem, and method dimensions. Evaluation of 18 LLMs on 619 paper pairs and 50 multi-paper sets shows that faithfulness verification reduces most models' raw F1 by more than half, exposing substantial hallucination and unsupported-evidence failures.