Yucheng Du
yucheng-du
AI & ML interests
LLM reliability, mechanistic interpretability, AI auditing, computational linguistics
Recent Activity
submitted a paper about 1 month ago
Recognition-Refusal Misalignment in LLMs: Why Models Answer Structurally Unanswerable Questions authored a paper about 1 month ago
Geometric Deviation as an Unsupervised Pre-Generation Reliability Signal: Probing LLM Representations for Answerability authored a paper about 1 month ago
Recognition-Refusal Misalignment in LLMs: Why Models Answer Structurally Unanswerable Questions