Senior Site Reliability Engineer
Airbyte
Airbyte is the data and action layer for AI agents. We give agents fast, accurate, authenticated access to business data across hundreds of sources, so they can discover the entities that matter, reason over real-time context, and take action in the systems they read from, not just observe them.
We started as the open-source standard for data movement and proved the economics of data integration at scale: hundreds of connectors, thousands of companies, and, since 2020, have raised $181M from leading investors including Benchmark, Accel, Altimeter, Coatue, and Y Combinator. As our CEO Michel Tricot puts it, "the last ten years were all about structured data. The future is all about context." We're now building that context infrastructure for production-grade agents on the same open foundation, as agents become the primary consumers of enterprise data.
Our mission is unchanged: make data available and actionable to everyone, everywhere. That everyone now includes AI agents.
THE ROLE:
You'll be the infrastructure and reliability engineer on the Data Replication team - a full-stack product team running over 3 million sync jobs a week powering thousands of data use cases across multiple regions and clouds. You’ll build and maintain the infrastructure, set reliability standards, drive down incidents, and make it easier and safer for engineers to ship through tooling. You're equally comfortable in a Terraform file, a Kubernetes cluster, and a postmortem doc.
We expect engineers here to actively use AI as a force multiplier - agentic tools to automate toil, augment incident response, and build smarter internal tooling. If you're not already doing this, you should be excited to start. We care as much about how you work as what you build. Trust, directness, and craftsmanship matter here.
WHAT YOU’LL DO:
- Own the infrastructure underpinning the Data Replication platform - Kubernetes clusters, CI/CD pipelines, secrets management, networking, and cloud resource configuration across AWS and GCP.
- Partner with product engineers to reliably integrate product features with infrastructure.
- Maintain and enhance observability, alerting, and anomaly detection with an eye towards LLM automation.
Don't want to miss the next one?
Subscribe to daily email alerts for roles matching your interests.



