职称 副教授
研究方向 大语言模型/智能体安全
导师类别 硕士生导师
主讲课程 无
联系邮箱 liuyr26@tju.edu.cn
教育与工作背景
(按时间倒序填写,标注时间段、单位、职务/学历)
2026.07-至今 天津大学人工智能学院 副教授
2021.09-2026.06 清华大学 计算机专业 博士研究生
2017.09-2021.06 清华大学 数学与应用数学专业 本科生
个人简介
(100-200字,突出核心成果及行业影响力)
长期从事大语言模型安全与可信人工智能研究,致力于应用因果推断理论系统性解决模型偏见、幻觉及对抗脆弱性等核心风险。参与国家自然科学基金专项项目、科技创新2030"新一代人工智能"重大项目等国家级课题多项,在NeurIPS、ICML、AAAI等CCF A类国际顶级会议及Games and Economic Behavior等权威期刊发表论文多篇,授权国家发明专利1项。曾获中国信息经济学会创新论文奖、清华大学毕业生启航奖金奖,兼任ICLR、NeurIPS、ICML、AAAI等国际会议审稿人。
代表性成果
(一)代表性论文
Liu Y, Yang K, Qi Z, Liu X, Yu Y, Zhai C. Bias and Volatility: A Statistical Framework for Evaluating Large Language Model's Stereotypes and the Associated Generation Inconsistency. Advances in Neural Information Processing Systems (NeurIPS), 2024.
Liu Y, Xu X, Hou Z, Yu Y. Causality Based Front-door Defence Against Backdoor Attack on Language Model. Proceedings of the 41st International Conference on Machine Learning (ICML), 2024.
Liu Y, Hou Z, Xu X, Wang S, Wu H, Yu K, Zhai C, Yu Y. HEV Generative Sandbox: A Framework for Assessing Domain-Specific Social Risks through Human-LLM Simulation. Proceedings of the AAAI Conference on Artificial Intelligence (AAAI), 2026.
Jiao Q, Kuang Z, Liu Y (共同一作), Yu Y. Optimal Tree Contest Design and Winner-Take-All. Games and Economic Behavior, 2025.
(二)主要获奖
中国信息经济学会创新论文奖,中国信息经济学会,2025年