Tech lead for agentic evaluation on Gemini in Chrome — building the infrastructure that measures what AI systems can actually do, reliably, at scale. Eight years at Google, from Chromium internals relied on by billions to the eval platforms that decide what ships.
I lead model evaluation and agentic quality for Gemini in Chrome, where my job is to turn fuzzy questions about capability and reliability into signals a team can trust: deterministic offline replay so results are reproducible, autoraters calibrated against human judgment, distributed eval pipelines with health dashboards and on-call alerting, and the observability tooling used to root-cause anomalous results.
Before that, three years deep in Chromium internals on architecture relied on by an estimated two billion people, and research in multi-agent systems and multi-objective optimization — published work and a granted patent.
I grew up in Hyderabad in an entrepreneurial family that shaped how I lead: growth mindset, resilience, service to community. Tennis and cricket since age ten taught me teamwork; Vipassana clarified the purpose underneath it — build technology that makes the world more equitable. As an angel investor, I back founders who share that conviction.
If it matters, it should be measurable and defensible. I'd rather build the metric than argue about the anecdote.
Absolute truth in the architecture and in the team. Good systems and good cultures both depend on it.
Empathy and constant learning, bringing out the best in the people around me. Service to community is the point.