Can you trust the model you deployed?
你部署的模型,还值得信任吗?
I am an undergraduate student in computer science at the University of Adelaide,
Australia, and currently a research assistant at Monash University working with
Dr. Jingwen Ye on trustworthy AI. I lead research on
backdoor and bit-flip attacks against vision-language-action models,
with earlier stints at HKUST (model IP protection) and the IoT Laboratory of China
University of Petroleum (East China). My agenda spans both sides of the fight:
mapping the attack surfaces of embodied AI and large language models, and building
provenance, watermarking, and defense mechanisms that make
deployed models verifiable and trustworthy.
我是澳大利亚阿德莱德大学计算机科学专业本科生,目前在莫纳什大学担任研究助理,
与 Jingwen Ye 博士合作开展可信 AI 研究。我主导
针对视觉-语言-动作(VLA)模型的后门与位翻转攻击研究,
此前曾在香港科技大学(模型知识产权保护)和中国石油大学(华东)物联网实验室从事研究。
我的研究议程横跨攻防两端:一边刻画具身智能与大语言模型的攻击面,
一边构建溯源、水印与防御机制,让部署中的模型可验证、可信任。
A robot that follows language is also a robot that can be told the wrong thing —
not through its microphone, but through its weights. That failure mode is what I study.
一台听得懂语言的机器人,也是一台可能被"说错话"的机器人——
指令不经过麦克风,而是直接写进权重。这种失效模式,正是我研究的对象。
researcher.yaml
- focus
- VLA security · model IP
- methods
- backdoors · bit-flips · watermarks · red-teaming
- stack
- PyTorch · MuJoCo · LIBERO
- affil
- B.CompSc @ U. Adelaide · RA @ Monash
- advisor
- Dr. Jingwen Ye (Monash)
- langs
- Mandarin · English
- status
- ● open to collaboration