Work

Production AI systems, upstream work in the open-source infrastructure they depend on, and the research that connects them.

Selected work

  1. AI platformsProductionA consumer-facing AI assistant on Amazon Bedrock AgentCoreAgentic assistant with grounded retrieval, governed tool access, and conversation memory, plus the shared AI platform internal tools now build on.
  2. Platform engineeringProductionCloud-native microservices for mortgage loan automationDesigned and built many of the platform's scalable microservices on AWS, owning each from design and Terraform to code and deployment.
  3. AI infrastructureOpen-source projectkvfleet: cache-aware routing for LLM fleetsA routing control plane that keeps KV-cache locality, policy, and health in one explainable decision.
  4. Data systemsUpstreamCheckpoint safety for change-data-capture in Apache SeaTunnelPinned down what a CDC checkpoint is allowed to contain when a fetcher fails mid-handoff.
  5. Query enginesUpstreamFaster, bounded query planning in Apache DataFusionRemoved two planner hot spots: repeated name resolution on wide queries and runaway regex compilation.
  6. Build platformsUpstreamNew public extension APIs for Apache Maven 4Gave Maven extensions stable, Maven-owned contracts for repository events and settings parsing.
  7. ResearchEventFlowSentry: reproducible fault testing for streamingTurns event-time failures in streaming pipelines into reproducible, explainable experiments.

Open-source tools I maintain

Small, focused libraries for AI systems, testing, and developer workflows.

Also contributed to

Apache Spark, Apache Beam, Apache Pinot, Apache Arrow, CPython, Go, Rust Cargo, PyTorch ExecuTorch, ClickHouse, Spring AI, Spring Cloud Gateway, SQLAlchemy, Meta Pyrefly, OWASP, OpenTofu, Kubernetes inference-perf.

Full history on GitHub