Yuzheng Xu

Large language model evaluation, human–AI complementarity, computer vision, and human–computer interaction.

Research Interests

I work on how we tell whether a language model is actually good at something, and on what happens when a person and a model have to reach an answer together rather than one replacing the other — evaluation reliability, and human–AI complementarity. Before that, computer vision: video understanding and annotation, and interfaces that give people real-time feedback on their own performance.

Education

Selected Publications

Toolkit

Python · PyTorch · OpenRLHF · FastAPI · Vue.js · n8n · Docker · Git

English · 日本語 · 中文 · Français