← Все вакансии/Senior/Okta
SeniorOfficeBellevue, Washington; San Francisco, California

Site Reliability Engineer

O
Okta
Уровень
Senior
Формат
Office
О роли

Описание вакансии

About the role

The Senior Site Reliability Engineer will help build, improve, and maintain our cloud platform services by designing and implementing complex cloud-based engineering enablement systems. With a strong focus on automation, testing, and operational excellence, you will deliver foundational infrastructure capabilities that enable corporate engineering teams to operate securely, reliably, and at scale.

Responsibilities
  • Secure Cloud Infrastructure & Pipelines: Design, build, and modernize scalable cloud environments and development tools while strictly enforcing security policies and standards for regulated environments.
  • Cross-Functional Collaboration & Advocacy: Partner with software engineering teams to champion DevOps and SRE best practices, deliver excellent internal customer service, and actively contribute to Agile workflows (e.g., demos, architecture sessions).
  • Technical Documentation & Operations: Create and maintain comprehensive technical documentation, including network diagrams, runbooks, and disaster recovery procedures to ensure system reliability and knowledge sharing.
Requirements
  • Professional Experience & Scale: 5+ years of experience in SRE, DevOps, or Systems Engineering roles with a proven track record of delivering complex, large-scale infrastructure projects.
  • AWS Expertise & Centralized Governance: Expert in building and managing AWS multi-account environments (spanning hundreds of accounts), with deep proficiency in authentication, governance, and organization management (AWS Orgs, IAM, Identity Center, StackSets).
  • Automation & CI/CD Pipelines: Highly skilled in infrastructure as code (Terraform), writing secure automation tools in Python, and building Git-based CI/CD workflows (GitLab, GitHub Actions).
  • Containerization & Observability: Strong hands-on experience managing container orchestration environments (Kubernetes) and utilizing monitoring and logging tools (Splunk, CloudWatch, Grafana stack).
Conditions
  • This position requires the ability to access federal environments and/or have access to protected federal data. As a condition of employment for this position, the successful candidate must be able to submit documentation establishing U.S. Person status (e.g. a U.S. Citizen, National, Lawful Permanent Resident, Refugee, or Asylee. 22 CFR 120.15) upon hire.
Стек и навыки

С чем работаем