AI & ML interests

Mechanistic interpretability, LLM security, indirect prompt injection detection

AICordon 's datasets

None public yet