About the company
We are seeking a highly skilled and experienced DevOps Engineer to join a critical enterprise digital platform project within the banking sector. This platform is characterized by its multiple integrated systems, demanding a robust focus on automation, deployment reliability, comprehensive observability, seamless scalability, and unwavering operational stability. In this role, you will be an integral part of our engineering teams, responsible for the design, development, and maintenance of our infrastructure, sophisticated CI/CD workflows, and essential operational processes. Your expertise will span across both cloud-based (AWS) and on-premises environments, ensuring the smooth and efficient operation of our cutting-edge banking platform.
Responsibilities
- Design, build, and maintain robust automation tools, sophisticated deployment pipelines, and streamlined operational processes across both AWS and on-premises systems.
- Provide comprehensive support for application deployments and infrastructure management across all environments, including development, staging, UAT, and production.
- Develop and continuously improve CI/CD pipelines tailored for a diverse range of application technologies, ensuring efficient and reliable software delivery.
- Implement and manage GitOps-based deployment workflows and advanced release strategies to enhance control and visibility.
- Support and manage containerized workloads effectively on platforms such as OpenShift, Kubernetes, or equivalent technologies.
- Establish and optimize monitoring, logging, alerting, and distributed tracing systems to significantly enhance system reliability and expedite troubleshooting efforts.
- Actively support high availability, disaster recovery planning, rigorous load testing, and overall platform stability initiatives.
- Collaborate closely with backend, frontend, data, and ML engineering teams to provide seamless support for both batch and real-time workloads.
- Prepare detailed technical documentation, comprehensive operational runbooks, and clear troubleshooting guidelines to empower the team and ensure knowledge transfer.
Requirements
- A minimum of 4+ years of hands-on experience in DevOps, Cloud Engineering, Platform Engineering, or Site Reliability Engineering roles.
- Demonstrated strong experience with AWS cloud infrastructure and managing production environments.
- Proven hands-on experience in designing, implementing, and managing CI/CD pipelines, deployment automation, and effective release strategies.
- Solid experience with container orchestration platforms such as OpenShift, Kubernetes, or similar technologies.
- Experience with GitOps tools and deployment practices, including but not limited to ArgoCD, Helm Charts, or Kustomize.
- Proficiency with automation or configuration-management tools like Terraform/CDK and Ansible.
- Experience utilizing monitoring and logging tools such as Prometheus, Grafana, ELK/EFK stack, or Splunk.
- A good understanding of Linux operating systems, networking principles, load balancing, cloud security best practices, and general system troubleshooting.
- The ability to work independently, take initiative, and collaborate effectively with multiple cross-functional engineering teams.
- Proficient English communication skills, sufficient for engaging in technical discussions and producing clear documentation.
Conditions
- Salary: Up to 50 million VND Gross - open to discuss if you're a strong fit
- Opportunity to work on a large-scale enterprise platform facing complex infrastructure and integration challenges.
- Collaborate closely with highly experienced engineering and architecture teams, fostering professional growth.
- Take ownership of significant DevOps initiatives, directly impacting the quality and efficiency of our delivery processes.
- We offer competitive compensation packages and flexible working arrangements for suitable candidates, promoting a healthy work-life balance.