About the company
Sezzle is revolutionizing the shopping experience beyond payments, blending cutting-edge tech with seamless, interest-free installment plans. We're transforming payments and redefining how people discover, interact with, and purchase the things they love.
About the role
We are seeking an exceptional Principal Infrastructure Engineer to design, build, operate, and scale the platform that powers Sezzle. Your focus will be the hardest infrastructure problems: increasing throughput, reducing latency, removing capacity bottlenecks, strengthening resilience, and making production operations more automated and predictable.
Responsibilities
- Own the technical architecture and evolution of core infrastructure: identify system limits, prioritize technical improvements, and implement changes that support increasing traffic, data volume, and workload complexity
- Connect business understanding to improvements across the system
- Engineer for scale and performance: build capacity models, run load and stress tests, diagnose bottlenecks across compute, networking, Kubernetes, and databases
- Design and build AWS infrastructure: implement resilient account, IAM, network, and service architectures
- Build and operate the Kubernetes platform: improve cluster architecture, lifecycle automation, workload isolation, resource allocation, autoscaling, safe upgrades, and deployment reliability
- Scale and optimize Aurora RDS for MySQL and Postgres: tune queries and indexes, address connection and replication bottlenecks, plan capacity, improve failover behavior
- Improve reliability through engineering: define and instrument service-level objectives and error budgets
- Participate in on-call and drive technical recovery during serious incidents
- Implement and test disaster recovery
- Build infrastructure-as-code and operational automation
Requirements
- 12+ years of experience
- Deep knowledge of AWS, Kubernetes, Aurora RDS (MySQL and Postgres)
- Comfortable moving between cloud architecture, networking, cluster internals, database performance, and application behavior
- Experience writing code and infrastructure-as-code
- Experience debugging production systems
- Experience with on-call rotation and incident recovery
- Experience building AI-assisted infrastructure and SRE tooling
Conditions
- Remote, based in Brazil
- Compensation: $12,500 - $20,800 USD per month (gross)
- On-call rotation required