About the company
At Klaviyo, we value the unique backgrounds, experiences and perspectives each Klaviyo brings to our workplace each and every day. We believe everyone deserves a fair shot at success and appreciate the experiences each person brings beyond the traditional job requirements.
Responsibilities
- Design, build, and maintain cloud-native security services used across Klaviyo.
- Own meaningful components and services end to end, from implementation through production operation.
- Improve the availability, scalability, latency, and efficiency of infrastructure security systems.
- Collaborate with senior engineers on system design and architecture.
- Identify performance, reliability, and security issues in distributed systems and drive improvements.
- Work extensively with technologies such as Python, Golang, AWS, Kubernetes, Terraform, and modern data stores.
- Participate actively in design reviews, code reviews, and whiteboarding sessions.
- Contribute to operational excellence through monitoring, alerting, and incident response.
- Participate in an on-call rotation.
- Build tooling, automation, and documentation that improve developer experience and security posture.
- Partner with product-facing and platform engineers to ship impactful, secure solutions.
Requirements
- 6+ years of solid experience building and operating cloud-native, distributed systems in production.
- Comfortable writing production-quality code in a language such as Python or Go.
- Hands-on experience with AWS (or a similar cloud provider) and understand managed services, networking, and IAM concepts.
- Experience working with containers and orchestration platforms such as Kubernetes.
- Experience writing infrastructure as code using languages such as Terraform.
- Comfortable owning services and features independently.
- Understand the fundamentals of scalable, multi-tenant architectures and secure system design.
- Care about reliability, performance, observability, and security.
- Comfortable participating in on-call and responding to production issues.
- Excellent communicator and collaborator.
- Ability to handle complex systems in outage situations and drive failures to root cause analysis.
- You've already experimented with AI in work or personal projects.
- You question convention and proactively look for ways to improve.
- You mentor and support other engineers.
Conditions
- This role is based in Boston, Massachusetts.
- Klaviyo supports work authorization and relocation for this position.