ML / AI Engineering · Kiran Kumar

Where should the signal go?

I spent a PhD asking that about human attention. Now I ask it about LLM routing, marketing attribution, and the production ML systems that answer it at scale, most recently at Chewy, EquipmentShare, and Convoy.

Currently building

Two open source projects, one underlying question.

A common interface evaluation harness for comparing agentic LLM routers (LiteLLM, RouteLLM, vLLM Semantic Router) across live benchmarks. Scores on a deployment focused metric suite (cost, latency, tool call accuracy, Pareto frontier) instead of a single leaderboard number.

  • RouterBench · BFCL v4 · tau²-bench · WebArena · SWE-bench · Terminal-Bench
  • Simulated (seeded, offline) and live backends
View on GitHub →

calmmm

Python · PyMC

A Bayesian media mix modeling package for measuring marketing effectiveness across geographies and KPIs at once: hierarchical geo×KPI pooling, adstock/saturation curves, and lift calibration from real incrementality experiments.

  • Full MCMC, variational, and MAP inference
  • Integrated attribution, ROI, and diagnostic reporting
View on GitHub →

Experience

10+ years shipping production ML, from underwriting to LLM agents.

2024 to Present

ML Engineer Tech Lead · Chewy

Marketing measurement and causal MMM/incrementality tooling behind $500M+ in annual media spend; architected an LLM/agent prototype that helped set the org's AI roadmap.

2023 to 2024

ML Engineer Tech Lead · EquipmentShare

Architected an LLM powered RAG search & recommendation engine and a production LLM gateway on Kubernetes handling 10K+ daily requests at 99.9% uptime.

2021 to 2023

Research Scientist · Convoy

Built a cost forecasting platform on Tab Transformer foundation models and adapters; prototyped a human in the loop SHAP based explainability loop that lifted revenue 3%.

2021

Senior ML Engineer · Metromile

Defined and shipped AI driven renewal underwriting, cutting loss ratio 5%.

2021

Data Scientist · Viaduct

Improved vehicle failure prediction recall 12% through telematics feature engineering across multiple automotive OEM product lines.

2019 to 2021

Data Scientist · SAP

Led Industry 4.0 and explainability initiatives across operations, finance, and procurement, including an NLP driven order status chatbot.

2013 to 2019

Doctoral Researcher · Indiana University

PhD research on visual attention, the throughline for the routing and attribution work above. See Research.

Full role history, metrics, and skills in the CV (PDF) →

Research · 2013 to 2019

Where visual attention goes, and why.

Plus 3 conference presentations at the Cognitive Science Society and CEMS · Google Scholar →

Contact

Say hello, or talk shop about routing, attribution, or attention.

[email protected]