About the role
As a Platform Engineer on the Compute team, you’ll improve how engineers build and run services, helping ensure our compute platform is reliable at scale for our learners and fellow Duo engineers.
The Compute team builds and operates the core compute platform that powers our applications. We build paved roads (self-service platform primitives, templates, and tooling) and we operate them so product engineers can focus on shipping features with confidence.
Responsibilities
- Build and operate core compute primitives that power production workloads based on Kubernetes, across different regions and cloud providers.
- Improve the delivery pipeline from commit to production by evolving our GitOps and deployment patterns, making rollouts safer and faster.
- Raise the operational bar: define SLOs, build dashboards/alerts, write runbooks, and participate in on-call/incident response to keep the platform dependable.
- Enable effective self-service by turning platform capabilities into reusable abstractions.
Requirements
- Strong problem-solving skills and experience delivering and shipping pragmatic solutions in production environments.
- Experience with distributed systems fundamentals (networking, service-to-service communication, failure modes, caching/storage concepts).
- Experience working on production systems (through industry roles, internships, or substantial open-source work).
- Experience with one or more CI/CD tools (Jenkins, Argo CD, GitHub Actions) and infrastructure management tools (Terraform, CloudFormation).
- Clear written and verbal communication skills—comfortable writing docs/runbooks and collaborating across teams.
- Functional knowledge of Linux system administration and automation.
Conditions
- Base salary is supplemented by equity compensation.
- Benefits include holistic well-being support.