Work With Me
15+ years building production ML systems — now available to help your team architect, scale, and operate LLM infrastructure and AI platforms.
Services
Architecture Review
Deep-dive analysis of your AI/ML system architecture with a written report and actionable recommendations. Ideal for teams hitting scaling issues or preparing for a major infrastructure change.
Includes:
- Written architecture assessment (15–25 pages)
- Prioritized recommendations by impact
- 2-hour review call with your team
- 30-day follow-up Q&A via email
Consulting Engagement
Hands-on consulting for specific AI/ML projects. From LLM serving infrastructure design to ML platform architecture to production deployment strategy. Scope defined upfront, delivered on time.
Includes:
- Custom scope and timeline agreed upfront
- Architecture and design documents
- Implementation guidance and code review
- Weekly sync calls during engagement
- Written summary and handoff documentation
Advisory Retainer
Ongoing strategic and technical guidance for your AI/ML team. Regular sessions to help make better architecture decisions, unblock hard problems, and stay ahead of the curve on LLM infrastructure.
Includes:
- 2× 60-min strategy sessions per month
- Async Q&A via Slack/email (48-hr response)
- Architecture review of key decisions
- Access to all course and research materials
- Introductions to relevant experts in my network
1:1 Technical Advisory
Individual sessions for senior ML engineers, engineering managers, or founders who need expert input on specific technical decisions. Book a single session or a package of hours.
Includes:
- 60-min deep-dive session
- Notes and action items sent after call
- Option to book follow-up sessions
- Async follow-up Q&A for 2 weeks
Corporate Workshops
Custom training workshops for engineering teams. From LLM inference deep dives to ML platform design to production AI systems. Delivered on-site or virtually.
Includes:
- Custom curriculum for your team's needs
- Hands-on exercises and architecture labs
- All slide decks and reference materials
- Post-workshop Q&A session
- Team assessment and learning summary
Speaking Engagements
Conference keynotes, corporate events, podcasts, and panel discussions. Technical talks on LLM inference, production AI, and ML infrastructure grounded in real engineering experience.
Includes:
- Custom talk tailored to your audience
- 30–60 min keynote or breakout session
- Q&A and audience interaction
- Slide deck provided to organizers
How It Works
Fill out the form below
Tell me about your project, team, and what you're trying to achieve. The more context the better.
Discovery call (30 min)
We'll have a quick call to assess fit, understand your situation, and determine the right service and scope.
Proposal & agreement
I'll send a clear proposal with scope, deliverables, timeline, and pricing. No surprises.
Engagement begins
We kick off on an agreed date. I work systematically, communicate proactively, and deliver on time.
FAQ
How do I know if my project is a good fit?
I focus on LLM inference infrastructure, ML platform design, production AI systems, and scaling ML at the team/org level. If you're building in one of these areas and need senior-level architectural guidance, we're probably a good fit. Fill out the form and describe your situation — I'll let you know honestly.
What's the typical engagement timeline?
Architecture Reviews take 2–3 weeks from kickoff to delivery. Consulting Engagements are scoped upfront — typically 4–12 weeks. Advisory Retainers are rolling 3-month commitments. Corporate Workshops can be scheduled 3–6 weeks out.
Do you work with early-stage startups?
Yes, though the fit depends on what you're building. If you're a seed/Series A company building LLM-powered products or ML infrastructure, I can often add a lot of value. If you're pre-product or pre-technical, it may be too early for most of my services.
Can you sign an NDA?
Yes. I routinely sign standard mutual NDAs before engagements begin.