Search by job, company or skills

IT Infrastructure Engineer

3-5 Years
SGD 4,000 - 6,000 per month
  • Posted an hour ago
  • Be among the first 10 applicants

Job Description

Responsibilities

  • Manage and operate centralized monitoring and observability platforms across applications, databases, infrastructure, networks, and cloud environments to ensure 24/7 service availability.

  • Monitor system health using metrics, logs, and alerts proactively identify anomalies, performance issues, and service degradation.

  • Perform alert triage, impact assessment, and incident coordination, escalating issues to the appropriate technical teams to meet SLA requirements.

  • Design and enhance monitoring strategies, dashboards, alerting frameworks, and observability standards to improve service visibility and reduce alert noise.

  • Support major incident management by providing diagnostics, cross-team coordination, and driving service reliability improvements through trend analysis and root cause identification.

  • Monitor and optimize cloud and infrastructure costs, implementing tagging, budgeting, cost allocation, and identifying opportunities for cost savings.

  • Develop operational and cost reports, dashboards, and forecasts to support service management, leadership, and operational decision-making.

  • Drive continuous improvement by expanding monitoring coverage, automating observability processes, maintaining documentation, and supporting after-hours operational activities.

Requirements

  • 3-5 years of experience in IT operations, NOC, service assurance, system monitoring, or cloud/infrastructure operations.

  • Hands-on experience with monitoring and observability tools such as CloudWatch, Grafana, Prometheus, Splunk, ELK Stack, or equivalent platforms.

  • Strong understanding of hybrid infrastructure (on-premises and AWS), system/network monitoring, application performance, and log/metric analysis.

  • Experience with AWS cost management, including Cost Explorer, budgeting, tagging strategies, and cloud cost optimization practices.

  • Familiarity with ITIL processes (Incident, Problem, and Change Management) AWS Associate-level certification or AWS FinOps Certified Practitioner is preferred.

    • Willingness to support after-hours operations, including deployments, maintenance, and incident response.

      GMP Recruitment Services (S) Pte Ltd | EA Licence: 09C3051 | VO UYEN AI LINH | Registration No: R22109232

More Info

Job Type:
Industry:
Function:
Employment Type:

Job ID: 153750759

Similar Jobs

Singapore

Skills:

Server TroubleshootingComputer EngineeringBackup SolutionsAdministrationInfrastructure MonitoringAnalytical and Problem-Solving SkillsConfiguration SolutionsInfrastructure Automationout-of-hours supportSystem Health ChecksVirtualization PlatformInfrastructure DeploymentAble To Work IndependentlyCollaborate With Internal Teaminfrastructure Incident Managementserver knowledge

Singapore

Skills:

cohesity SanAvamarVmware VsphereVeeamVlansGcpIsilonData DomainAzureAWSPowerStoreociRubrikVCFDell NetWorker

Middle Road, Singapore

Skills:

cohesity SanVmware VsphereAvamarVeeamVlansGcpIsilonData DomainAzureAWSPowerStoreociRubrikVCFDell NetWorker

Singapore

Skills:

VMwareBackupWindows ServerLoggingDnsFirewallsRed Hat LinuxroutingNutanixLoad BalancingAzureAWSHyper-VMonitoringalertingenterprise storageSegmentationobservability

Anson, Singapore

Skills:

VMwareWindows ServerCisco NetworkingPrometheusGrafanaBackup RecoveryOpenshiftRhelKubernetesFortiGate firewallsinfrastructure monitoringDell servers

Beware of Scammers

We don’t charge money for job offers