Yuzheng Xu

Research Interests

I work on how we tell whether a language model is actually good at something, and on what happens when a person and a model have to reach an answer together rather than one replacing the other — evaluation reliability, and human–AI complementarity. Before that, computer vision: video understanding and annotation, and interfaces that give people real-time feedback on their own performance.

Selected Publications

Keywords

LLM evaluation · Human–AI complementarity · Self-awareness · AI for science