Highlights
Joining the DevNetwork Advisory Board
Selected to DevNetwork's 2026 AI/ML Advisory Board, helping shape content across its global developer and AI conferences.
Talks, publications, presentations, and recognition.
Highlights
Selected to DevNetwork's 2026 AI/ML Advisory Board, helping shape content across its global developer and AI conferences.
Highlights
My take in Tech Magazine's expert roundup on balancing technical debt against new features.
Highlights
I spoke at a PMI San Francisco Bay Area session on how PMs can move beyond execution to become strategic technology leaders in the AI era.
Highlights
A morning of AI research, from rocket design to AI governance, and what stood out from the chair's seat.
Highlights
CTO Sync featured my approach to winning stakeholder support for technical debt paydown: weighing blast radius, exposing consequences to leadership, and reserving a fixed slice of every cycle for paydown.
Highlights
Invited speaker at CloudX 2026 (Santa Clara, Sept 1-3), co-located with API World + AI TechWorld. Speaking on how to forecast hardware capacity for AI infrastructure at billion-user scale.
Highlights
Mashable profiled my work leading AI infrastructure at billion-user scale: lifting model deployment success from 60% to 99%, cutting scale-up time from three days to two hours, and a GPU efficiency redesign worth an estimated $20M in savings.
Highlights
Invited speaker at the Kong API + AI Summit 2026 in Los Angeles (Sep 30 – Oct 1). A practitioner session on the operational realities of running large-scale AI inference: deployment coordination, GPU and infrastructure constraints, performance tuning, and reliability under rapid model iteration.
Highlights
Upcoming: I'm speaking at AI Infra Summit 2026 (Santa Clara, Sept 15-17) on what it takes to run AI infrastructure at billion-user scale.
Highlights
I'm serving as Session Chair at the Sixth IEEE International Conference on Intelligent Technologies (CONIT 2026) in Hubballi, India, on June 20, 2026.
Highlights
My paper on a replica-centric capacity-planning framework for GPU and Multi-Instance GPU (MIG) based AI inference platforms was accepted at IEEE MetroInd 4.0 & IoT 2026 in Rome, Italy.
Highlights
I spoke at the Budapest Data + AI Forum 2026 on why ML deployments fail in production and how to engineer reliability at scale. Session abstract, key takeaways, and links.