Senior Data Engineer · Open-source maintainer · Apache Airflow PMC member
I’m a Senior Data Engineer at Datadog, working on large-scale billing data pipelines and the infrastructure behind their scheduling, computation, storage, and reliability.
I’m also an Apache Airflow PMC member, committer, and security team member. I contribute across Airflow’s core, providers, execution infrastructure, APIs, and developer experience. A large part of my open-source work is reviewing contributions, helping contributors, and improving the project’s reliability and security.
My interests sit at the intersection of:
- Workflow orchestration and distributed systems
- Data platforms, lakehouse architecture, and stream processing
- Apache Spark and Kubernetes
- Open-source sustainability and contributor experience
| Project | What it does |
|---|---|
| Apache Airflow | A platform to programmatically author, schedule, and monitor workflows. |
| spark-on-k8s | A Python package for submitting and managing Apache Spark applications on Kubernetes. |
| airflow-duckdb | An Apache Airflow integration for running DuckDB queries. |
| async-batcher | Asynchronous batching utilities for reducing request overhead and improving throughput. |
Open-source maintenance is more than writing code. It includes reviewing pull requests, supporting contributors, investigating regressions and security issues, maintaining releases, and improving documentation.
If my work has helped you or your organization, sponsoring me on GitHub helps me dedicate more time to work that benefits the wider data engineering community.






