Hi, I'm Mansi Phute

CS PhD at Georgia Tech
My research is in the field of Trustworthy, Responsible, and Secure AI. My PhD thesis aims to identify vulnerabilities in AI systems and develop practical, scalable defenses against them, with the goal of securing large-scale AI models and increasing trust in the deployed systems. My work has produced two patents and two industry-deployed defenses along with papers at multiple top tier AI conferences. My work includes VISOR and VISOR++ which create a universal, transferable steering image that can steer architecturally diverse models without requiring access to their internal model parameters at runtime.
I am a PhD student at Georgia Tech advised by Polo Chau as a part of the Polo Club of Data Science.
I have collaborated with scientists at IBM, HiddenLayer, Intel Labs, and Nanyang Technological University

News


Featured Publications

Transferrable Visual Input based Steering for Output Redirection in Large Vision Language Models
ECCV 2026, 2026
A Unified Evaluation System for Configurable Attacks in Differentiable Environments
arXiv, 2026
Bridging Differentiable Rendering and Photorealistic Simulation for End-to-end Adversarial Attacks
Pacific Asia Conference on Knowledge and Data Discovery, 2026
Visual Input based Steering for Output Redirection in Large Vision Language Models
AAAI AIR-FM Workshop, 2026
A Large-Scale Dataset for Testing Robustness of Image Classifiers
NeurIPS, 2024
By Self Examination, LLMs Know They Are Being Tricked!
ICLR Tiny Paper, 2024