MechAudit-40: White-Box Auditing across 40 LLM Attack Mechanisms
A systematic study of internal attack signatures across 40 mechanisms, with a runtime auditor evaluated on previously unseen attacks.
SAINT LOUIS UNIVERSITY
I study AI security and trustworthy AI, with a focus on reasoning backdoors, model auditing, and privacy. I am supervised by Dr. Reza Tourani.
PAPERS & PREPRINTS
Research in AI security, privacy, and communication systems.
A systematic study of internal attack signatures across 40 mechanisms, with a runtime auditor evaluated on previously unseen attacks.
A compact reasoning firewall that audits intermediate steps and identifies where a reasoning trace first becomes unsupported.
Investigating how backdoors survive continual model updates through Blind Task Backdoor and Latent Task Backdoor attacks.
Exposing latent backdoors that activate within a model’s chain of thought to alter its output without changing the user’s query.
Vib-Sound combines vibration and acoustic signals with collaborative cross-channel demodulation for communication between commodity devices.
CONTACT
For questions about my work or research collaborations, please get in touch.