Based in our Manila office (or remote in PHT), you will be working in a global automation and operations team, focusing on building, deploying, and managing mission-critical infrastructure.
Design, implement, and maintain infrastructure services supporting 30+ development teams.
Develop and manage infrastructure as code using GitOps principles.
Manage and optimize AWS cloud infrastructure for scalability, reliability, and security.
Participate in on-call rotations and lead incident response within the team.
Monitor, troubleshoot, and resolve issues across infrastructure, networks, and applications.
Collaborate closely with development teams to deliver robust, production-ready solutions.
Set a strong technical example through code reviews, architecture discussions, and hands-on delivery.
3+ years of demonstrable experience in Platform Engineering, DevOps, or Site Reliability Engineering (SRE) roles.
Solid experience in maintaining large-scale, business-critical cloud production environments (AWS preferred).
Experience with infrastructure automation (Chef, Terraform, etc.).
Deep Hands-on experience on managing and scaling containers like Docker and Kubernetes.
Proven track record of designing secure and efficient CI/CD (like github actions)
Proficiency as a developer (python, ruby and go lang)
Strong passion for automation.
Experience with ELK (Elasticsearch, Logstash, Kibana), RabbitMQ, NoSQL or similar applications.
Comfortable working in distributed and global teams
Comfortable participating in a periodic on-call rotation to support critical production systems.
Ubuntu operating system expertise.
Fluent in English (written and verbal).
Copyright © 2026 Grabjobs Pte.Ltd. All Rights Reserved.