Hi, I'm Mansi Phute

CS PhD at Georgia Tech
My research interests are Responsible AI and ML safety. I work on developing explanations for ML systems, analyzing them to identify vulnerabilities, and finding solutions to mitigate these issues. My UNDREAM system system offers a way to bridge differentiable rendering and photorealistic simulation for end-to-end adversarial attacks, thus enabling beter transferability of attacks to the physical world. My work includes VISOR and VISOR++ which create a universal, transferable steering image that can steer architecturally diverse models without requiring access to their internal model parameters at runtime.
I am a PhD student at Georgia Tech advised by Polo Chau as a part of the Polo Club of Data Science.
I have collaborated with designers, developers, and scientists at IBM, HiddenLayer, Intel Labs, and Nanyang Technological University

Featured Publications

Transferrable Visual Input based Steering for Output Redirection in Large Vision Language Models
ECCV 2026, 2026
Bridging Differentiable Rendering and Photorealistic Simulation for End-to-end Adversarial Attacks
Pacific Asia Conference on Knowledge and Data Discovery, 2026
Visual Input based Steering for Output Redirection in Large Vision Language Models
AAAI AIR-FM Workshop, 2026
A Large-Scale Dataset for Testing Robustness of Image Classifiers
NeurIPS, 2024
By Self Examination, LLMs Know They Are Being Tricked!
ICLR Tiny Paper, 2024