Machine Learning Engineer with 5+ years across ML systems, distributed model training, and deep learning research. Builds and troubleshoots multi-GPU / multi-node training pipelines (PyTorch, SLURM) with reproducible tracking, checkpointing, and failure recovery. Delivered production ML at enterprise scale at ServiceNow, including an agentic ticket-resolution system with a neurosymbolic verification layer. Published at AAAI 2026, TMLR 2025, and KDD/ASONAM 2025.