← Все вакансии/Lead/jobgether
LeadRemoteBrazil

Staff Engineer, Platform & Infrastructure

J
jobgether
Уровень
Lead
Формат
Remote
О роли

Описание вакансии

About the company

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Staff Engineer, Platform & Infrastructure based in Brazil.

Responsibilities
  • Own and scale the cloud infrastructure supporting the platform, including compute, networking, storage, and core data services.
  • Lead the development and productization of BYOC and private cloud deployments, including provisioning, upgrades, observability, and operational scalability.
  • Build and evolve infrastructure-as-code and GitOps foundations that enable a small engineering team to deploy safely and efficiently.
  • Establish reliability as a core product capability by defining meaningful service-level objectives, improving observability, and developing trusted incident response processes.
  • Lead the infrastructure and operational strategy for Postgres, Redis, Elasticsearch, ClickHouse, and other critical data systems as usage grows.
  • Own infrastructure-related security and compliance initiatives, including SOC 2, GDPR, HIPAA, penetration testing, and enterprise security reviews.
  • Design scalable Terraform modules, multi-account cloud architectures, and safe infrastructure state-management practices.
  • Operate and improve production Kubernetes environments while making pragmatic decisions about when alternative technologies are more appropriate.
  • Support software deployments in customer-controlled environments, including single-tenant, on-premises, air-gapped, and other constrained environments.
  • Collaborate directly with customers and engineering teams to understand infrastructure challenges and turn customer needs into practical technical solutions.
  • Contribute to product specifications, technical roadmaps, and broader company strategy through strong technical judgment and ownership of key decisions.
  • Provide technical leadership during incidents, unblock teammates, and establish engineering practices that improve reliability and delivery speed.
Requirements
  • 10+ years of professional experience in platform engineering, infrastructure, DevOps, SRE, or closely related disciplines, with demonstrated experience operating high-traffic production systems.
  • Deep hands-on experience with Kubernetes in production, combined with strong judgment around its appropriate use and alternatives.
  • Proven experience operating software in environments outside your direct control, such as BYOC, private cloud, single-tenant, on-premises, or air-gapped environments.
  • Advanced AWS expertise; familiarity with GCP and Azure is a plus.
  • Strong Terraform experience, including reusable modules, secure state management, and multi-account cloud architectures.
  • Significant database expertise, particularly with PostgreSQL, including performance optimization, migrations under load, replication, and production operations.
  • Practical experience with security and compliance programs such as SOC 2 or ISO audits, penetration tests, and enterprise security reviews.
  • Product-oriented mindset with strong developer empathy and the ability to turn customer problems into technical solutions.
  • Familiarity with TypeScript and Node.js is beneficial for working effectively within the existing technology environment.
  • Excellent communication and collaboration skills, with the ability to work effectively in a remote-first, distributed environment.
  • Proactive, pragmatic, and low-ego approach, with a willingness to take ownership during incidents and help unblock colleagues.
  • Strong autonomy, ownership mentality, and comfort navigating ambiguity in a fast-moving startup environment.
  • Ability to influence technical direction and communicate complex infrastructure decisions clearly to both technical and non-technical stakeholders.
  • Availability to work remotely from North America, LATAM, or Europe.
Conditions
  • Fully remote, remote-first working environment.
  • Opportunity to work from North America, LATAM, or Europe.
  • High level of autonomy and ownership over critical platform and infrastructure decisions.
  • Staff-level opportunity to shape technical strategy, infrastructure standards, and product direction.
  • Work on challenging developer infrastructure and cloud technologies at a growing technology company.
  • Exposure to Kubernetes, AWS, Terraform, GitOps, BYOC, private cloud, distributed systems, and large-scale data infrastructure.
  • Opportunity to build scalable infrastructure products rather than maintaining bespoke solutions.
  • Close collaboration with an experienced, highly autonomous engineering team.
  • Significant influence on reliability, security, compliance, and the future scalability of the platform.
  • Startup environment offering broad technical scope, rapid decision-making, and the opportunity to make a direct impact.
Стек и навыки

С чем работаем