Architecting Intelligence LogoAI Labs

Media Kit

For conference organizers, podcast hosts, journalists, and media professionals. Everything you need to book or feature Pawan K Jha.

PJ

Headshot available on request

Pawan K Jha

Sr. Principal AI/ML Scientist & Systems Architect

Founder, Architecting Intelligence Labs

LLM InferenceML InfrastructureProduction AIAgentic AIML Systems Architecture

15+

Years in Production ML

Growing

Substack Subscribers

@PawanMLEng

X / Twitter Followers

Architecting Intelligence

YouTube Channel

Short Bio (50 words)

Copy-ready

Pawan K Jha is a Sr. Principal AI/ML Scientist and Systems Architect with 15+ years building large-scale ML platforms, LLM inference systems, and production AI architecture. He is the founder of Architecting Intelligence Labs, where he publishes deep technical research on LLM inference, ML infrastructure, and production AI systems. He writes at pawankjha.substack.com and can be found on X at @PawanMLEng.

Full Bio (200 words)

Copy-ready
Pawan K Jha is a Sr. Principal AI/ML Scientist and Systems Architect with over 15 years of experience building large-scale machine learning platforms, LLM inference systems, search and ranking, forecasting systems, and production AI architecture. Throughout his career, Pawan has designed and shipped ML systems that operate at massive scale — from feature platforms serving billions of predictions per day, to LLM serving infrastructure handling complex multi-model inference pipelines. He has deep expertise in the full ML stack: GPU scheduling, KV cache management, tensor parallelism, model quantization, continuous batching, and the operational layer that makes it all reliable in production. He is the founder of Architecting Intelligence Labs, an independent research and consulting practice focused on the hardest problems in production AI. Through deep technical writing, courses, tools, and consulting, he translates years of production experience into content and products that make serious ML engineers better at their craft. Pawan publishes a widely-read technical newsletter on Substack (pawankjha.substack.com) and is active on X as @PawanMLEng, where he shares daily insights on LLM inference, ML infrastructure, and production AI systems.

Key Talking Points

Topics Pawan speaks and writes about with authority:

  • Why LLM inference is fundamentally different from traditional ML serving
  • The real bottlenecks in large-scale LLM deployment (hint: it's not what most people think)
  • How KV cache management changed the economics of LLM serving
  • What 'production-ready' actually means for agentic AI systems
  • The gap between research and production in AI/ML — and how to bridge it
  • Career paths for senior ML engineers in the age of LLMs

Available For

Conference Keynote

30–60 min

Corporate Workshop

Half-day to full-day

Podcast / Interview

30–90 min

Get in Touch

For speaking inquiries, podcast bookings, or media requests, reach out directly.