experience.
Where I have worked and what I shipped.
June, 2025 - Present
- Building and improving distributed backend systems where reliability, consistency, and failure recovery matter.
- Designing infrastructure for long-running workflows and asynchronous processing, with a focus on making failures recoverable rather than disruptive.
- Working on reusable systems for metadata lineage, real-time state propagation, and engineering automation.
- Building practical AI tooling that can interact with real engineering environments, including code execution, Kubernetes operations, debugging, and automated workflows.
- Taking ownership of problems across service boundaries, from understanding the root cause to designing, implementing, and shipping the fix.
Python/
Kafka/
Numaflow/
Kubernetes/
AI Agents/
Distributed Systems
December, 2024 - June, 2025
- Worked on workflow orchestration and reliability, including the transition toward more durable execution and better failure handling.
- Debugged large-scale production workflows where seemingly small issues could cascade into repeated failures across millions of assets.
- Improved the security and reliability of platform integrations by addressing how credentials and operational data were handled.
- Strengthened observability and debugging workflows so engineers could get from an incident to its root cause with less guesswork.
- Built a strong foundation in backend engineering by working on problems that crossed application code, infrastructure, and distributed systems.
Python/
Argo Workflows/
Temporal/
HashiCorp Vault/
Data Governance