Yucheng Du
yucheng-du
AI & ML interests
LLM reliability, mechanistic interpretability, AI auditing, computational linguistics
Recent Activity
submitted a paper 10 days ago
Recognition-Refusal Misalignment in LLMs: Why Models Answer Structurally Unanswerable Questions authored a paper 10 days ago
Recognition-Refusal Misalignment in LLMs: Why Models Answer Structurally Unanswerable Questions