Shreyash Dhoot

AI Safety Alignment Adversarial Robustness
I'm fascinated by the world of AI. My focus is on understanding how to make these models robust against adversarial jailbreaks, aligned with human intent, and ultimately safe for the world to use. When I'm not buried behind my laptop, you can usually find me trail running, sketching, or cooking.
Shreyash on a night forest trail