Brief Summary
Join a dynamic team as a Cloud Operations Specialist, where you will leverage your expertise to ensure optimal performance and reliability of multi-cloud infrastructure. Play a key role in driving automation and best practices in cloud operations.
Responsibilities
- Operate and maintain cloud-native services across multiple platforms, including AWS, Microsoft Azure, and Google Cloud. Monitor and troubleshoot infrastructure performance, availability, and scalability. Participate in round-the-clock operational support and assist senior engineers with hands-on technical troubleshooting. Maintain infrastructure deployment pipelines with tools such as Terraform or Ansible and troubleshoot environment issues. Conduct OS patching operations on various platforms and ensure compliance with organizational standards. Collaborate with application teams to optimize OS-level performance and troubleshoot deployment issues. Implement security hardening measures and manage vulnerability remediation across cloud platforms. Document systems, create standard operating procedures, and ensure audit readiness. Provide mentorship and technical guidance to junior engineers. Integrate monitoring and observability tools, ensuring effective metrics and alert management across cloud environments.
Requirements
- Bachelor's degree in Computer Science, Information Systems, or a related field. At least 3 years of experience in cloud engineering roles focusing on AWS, Azure, or GCP. Minimum: 2 years of experience in regulated cloud environments or public sector roles. Proven experience in 24/7 operational support with a shift rotation. Familiarity with ITIL processes and tools for incident, problem, and change management. Strong capability in mentoring and leading junior engineers to enhance team performance. Experience in security practices and vulnerability management across cloud services.