Yucheng Du
yucheng-du
AI & ML interests
LLM reliability; mechanistic interpretability; AI auditing; computational linguistics
Recent Activity
submitted a paper about 2 hours ago
Recognition-Refusal Misalignment in LLMs: Why Models Answer Structurally Unanswerable Questions authored a paper about 8 hours ago
Geometric Deviation as an Unsupervised Pre-Generation Reliability Signal: Probing LLM Representations for Answerability authored a paper about 8 hours ago
Recognition-Refusal Misalignment in LLMs: Why Models Answer Structurally Unanswerable Questions