# Operations Support Engineer - (Public Sector)

**Company:** [Xtremax Pte. Ltd.](http://jobs.workable.com/companies/iLMpt7c7UbsmNQeNARY8E4.md)
**Location:** Singapore, Singapore
**Workplace:** on site
**Employment type:** Contract
**Department:** Infra

[Apply for this job](http://jobs.workable.com/view/de613014-fbaa-44e2-b718-9918bb14ce4f)

## Description

At Xtremax, we're looking for an **Operations Support Engineer** to help build, operate, and continuously improve the cloud platforms that power large-scale, mission-critical digital services.

In this role, you'll work across modern infrastructure technologies including AWS, virtualisation, container platforms, Infrastructure as Code, and observability tooling. You'll partner closely with software engineers and platform teams to ensure our environments remain secure, reliable, and highly available while driving automation and operational excellence.

This is an opportunity for infrastructure engineers who enjoy solving complex operational challenges, improving system reliability through automation, and working with modern DevOps and Site Reliability Engineering (SRE) practices. You'll contribute to cloud infrastructure that supports enterprise and public sector digital transformation initiatives while expanding your expertise across cloud-native technologies, platform engineering, and infrastructure automation.

### **Responsibilities**

-   Design, deploy, and maintain cloud and hybrid infrastructure across development, staging, and production environments.
-   Administer AWS services including EC2, ECS, S3, RDS, Lambda, IAM, VPC, CloudFormation, and CloudWatch.
-   Manage enterprise virtualisation platforms including VMware vSphere and Hyper-V, ensuring platform performance, capacity, and lifecycle management.
-   Implement and maintain monitoring and observability solutions using CloudWatch, Prometheus, Grafana, ELK Stack, StackOps, and related technologies.
-   Automate infrastructure provisioning and configuration using Terraform, Ansible, and Infrastructure as Code (IaC) practices.
-   Support containerised workloads using Docker, Kubernetes, and Amazon ECS.
-   Maintain platform security through system hardening, access management, compliance monitoring, and security tooling such as CyberArk.
-   Improve platform reliability by implementing SRE practices, defining SLOs and SLIs, reducing operational toil, and enhancing automation.
-   Support backup, disaster recovery, high availability, and resilience testing, including multi-AZ deployments and AWS Fault Injection Simulator (FIS).
-   Collaborate with development teams to troubleshoot platform issues, optimise system performance, and maintain operational documentation and runbooks.

## Requirements

-   Bachelor Degree in Computer Science, Information Technology, Software Engineering, or a related discipline.
-   4–7 years of hands-on experience in infrastructure engineering, cloud operations, platform engineering, or systems administration.
-   Strong experience administering Linux and Windows Server environments.
-   Hands-on experience with AWS services such as EC2, ECS, S3, RDS, IAM, VPC, Lambda, CloudFormation, and CloudWatch.
-   Experience managing VMware vSphere and/or Hyper-V environments.
-   Experience implementing Infrastructure as Code using Terraform, AWS CloudFormation, and/or Ansible.
-   Experience with container platforms such as Docker, Kubernetes, or Amazon ECS.
-   Experience implementing monitoring and observability solutions using CloudWatch, Prometheus, Grafana, ELK Stack, or similar platforms.
-   Good understanding of networking concepts including TCP/IP, DNS, DHCP, VPN, and routing.
-   Proficiency in scripting using Python, PowerShell, Bash, or similar languages.
-   Experience with GitHub and CI/CD workflows.
-   Strong analytical, troubleshooting, communication, and documentation skills.

### Preferred

-   Experience implementing Site Reliability Engineering (SRE) practices, including SLOs, SLIs, and error budgets.
-   Experience with AWS Fault Injection Simulator (FIS) or resilience testing.
-   Experience working in government, regulated, or security-sensitive environments.
-   Experience with configuration management tools such as Puppet or Chef.
-   Professional certifications such as:

-   AWS Certified Solutions Architect
-   AWS Certified SysOps Administrator
-   VMware Certified Professional (VCP)
-   Microsoft Certified: Windows Server
-   Red Hat Certified Engineer (RHCE)
-   Networking or security certifications

## Benefits

By submitting your resume/CV, you consent and agree to allow the information provided to be used and processed by or on behalf of Xtremax Pte Ltd for purposes related to your registration of interest in current or future employment with us and for the processing of your application for employment.

You also represent to us that you have obtained the consent of your referees when you disclose to us their personal data for the purpose of conducting reference checks.

The personal data held by us relating to your application will be kept strictly confidential and in accordance with the PDPA. You may also refer to our Privacy Policy for more details here: 

We regret to inform you that should you not consent to providing the necessary data required for us to process your application, your application will be considered void.
