Search Jobs

Search by job, company or skills

5-7 Years
SGD 6,500 - 13,000 per month
Early Applicant
  • Posted 10 days ago
  • Be among the first 10 applicants

Job Description

About the Role

We're hiring a Site Reliability Engineer (SRE) to join a global engineering team supporting mission-critical financial markets platforms undergoing a large-scale cloud, cyber security, and platform modernisation programme.

This is a hybrid Production Engineering and Cloud Operations role where you'll help maintain the stability of business-critical systems while driving automation, operational improvements, and cloud transformation initiatives. It's an excellent opportunity for engineers who enjoy troubleshooting complex production environments, improving platform reliability, and reducing manual operational effort through automation within a highly regulated, enterprise-scale environment.

Key Responsibilities

  • Provide L2/L3 support for business-critical production applications
  • Troubleshoot and resolve application, infrastructure, and platform incidents
  • Perform root cause analysis and implement preventative improvements
  • Build and maintain CI/CD pipelines and deployment automation
  • Develop automation solutions using Python and Shell scripting
  • Support AWS-based cloud environments and platform modernisation initiatives
  • Improve monitoring, observability, and operational processes
  • Partner with Engineering, Security, and Operations teams to deliver platform improvements
  • Support vulnerability remediation, patching, and security compliance activities
  • Contribute to platform reliability, resilience, and continuous improvement initiatives

Requirements

  • 5+ years of experience in Site Reliability Engineering, DevOps, Production Engineering, Cloud Operations, or Infrastructure Engineering
  • Strong experience supporting production-critical environments
  • Hands-on AWS experience
  • Strong Linux administration experience
  • Python and Shell scripting experience
  • Experience with CI/CD pipelines and deployment automation
  • Strong incident management and root cause analysis skills
  • Experience with monitoring and observability tools such as Datadog, BigPanda, or Splunk
  • Excellent communication and stakeholder management skills

Preferred Skills

  • Red Hat Enterprise Linux (RHEL)
  • Ansible
  • Terraform
  • HashiCorp Vault
  • Artifactory
  • Security remediation or cyber security programme experience
  • Financial Services experience
  • Experience supporting global platforms across multiple regions

We regret to inform that only shortlisted candidates will be notified.

EA registration number : ANDREW JONAS MATTHEW, R21103843

Allegis Group Singapore Pte Ltd, Company Reg No. 200909448N, EA Licence No. 10C4544

More Info

Job Type:
Industry:
Employment Type:

Key Skills

Security remediation

HashiCorp Vault

Cyber security programme experience

Red Hat Enterprise Linux (RHEL)

Deployment automation

CI/CD pipelines

Similar Jobs

5-7 yrs
Singapore
Skills:
Oauth, Agile Methodology, Prometheus, Jwt, Grafana, Linux Administration, Jenkins, Git, Docker, Ansible, Teamcity, Python, Kubernetes, Test-driven development, Splunk SIEM, CI/CD tools, Public Cloud platforms, Open-source frameworks and libraries relevant to backend and platform development, Proxy networking, SPIFFE, GitHub Actions, regular expressions, Identity authentication concepts
5-7 yrs
Singapore
Skills:
S3, Elk, Prometheus, Kafka, Vpc, Grafana, Shell, Terraform, MySQL, Aws Ec2, Python, RDS, Redis, Jenkins, Cloudwatch, Linux, Iam, Kubernetes, Go, CI CD, ElastiCache, RocketMQ, EKS, GitLab CI, ALB
5-7 yrs
Singapore
Skills:
Java, Quality assurance, C++, High Availability, Python, Operation deployment, Maintainability, Disaster Recovery, Systems observability, Go, Large-scale distributed systems
5-7 yrs
Singapore
Skills:
Load Balancing, Routing, Dns, VLAN, DHCP, Vpn, Docker, Network Troubleshooting, Tcp/ip, Kubernetes, AWS cloud infrastructure, Linux administration and troubleshooting
3-6 yrs
SGD 9,000 - 14,000 per month
Singapore, Alexandra Road
Skills:
Bash, Gcp, Docker, Linux, Distributed Systems, Azure, Kubernetes, Python, AWS, High-availability architecture, Go, oci, Monitoring and observability systems