research intern
7 months, mar 2026 – sep 2026AI4Bharat, IIT Madras
- investigating mechanistic interpretability and circuit tracing in SoTA open models (Gemma, Qwen) on reasoning tasks across modalities, languages and model scales, focused on vision-language models over Indic inputs.
- ran 40,000+ instrumented forward passes for activation capture on multi-billion-parameter models under a hard 24 GB VRAM ceiling, with component-attribution scripts to isolate critical circuits before committing to full runs.
- engineered a memory-efficient activation-patching framework: disabled fused and attention-kernel optimisations to expose per-layer activations, then managed the non-linear VRAM growth that followed.
stackPythonPyTorchHugging FaceCUDAmixed precisionLinux