As an OpenStack Platform Engineer, you will be responsible for designing, implementing, automating, and maintaining reliable and scalable OpenStack infrastructure. You will work closely with cross-functional teams to ensure platform availability, security, performance, and operational efficiency. The ideal candidate will have strong expertise in OpenStack, Kubernetes, and Red Hat Linux, with solid troubleshooting, problem-solving, and communication skills.
Key Responsibilities include:
- Design, implement, and maintain Red Hat OpenStack-based cloud infrastructure solutions
- Develop and optimize automation scripts for deployment and management of OpenStack environments
- Collaborate with cross-functional teams to integrate OpenStack with other systems, including Kubernetes and OpenShift
- Troubleshoot and resolve complex issues related to OpenStack components and underlying infrastructure
- Implement security best practices and ensure compliance with organizational policies
- Contribute to capacity planning, performance optimization, and scalability of OpenStack environments
- Assist in change management of OpenStack platform for new versions, hotfixes, SysAdmin tasks, etc
- Explore and implement Agentic AI and AIOps solutions to enhance OpenStack administration and operational efficiency
- Leverage AI-assisted automation for monitoring, troubleshooting, incident analysis, and remediation of OpenStack environments
- Develop, Integrate AI agents with OpenStack APIs and automation tools to streamline infrastructure operations
- Work with experienced team members to conduct root cause analysis of issues, review new and existing code and/or perform unit testing
- Author/edit system documentation/playbook(s)
- May serve as a mentor on procedural matters to less experienced internal and third-party team members
Requirements
- At least 5+ years of relevant working experience
- Expertise with Red Hat or any OpenStack platform
- Expertise with Kubernetes
- Expertise in Red Hat Linux administration
- Experience in infrastructure provisioning with VMware ESX environment and monitoring with Prometheus and Ops
- Experience integrating with Elastic Stack
- Familiarity with Generative AI, LLMs, Agentic AI, and AIOps concepts for infrastructure operations
- Experience or exposure to AI agents and AI-assisted automation for cloud/OpenStack administration
- Understanding of integrating AI/LLM capabilities with APIs, Python, Ansible, or Terraform is preferred
- Ability to succeed in a fast-paced, high-demand environment
- Red Hat/OpenStack certifications in Linux and OpenStack administration preferred
- Excellent oral and written communication skills
- Experience working with infrastructure-as-code technologies such as Terraform is preferred
- Demonstrates strong customer service awareness and orientation
- Ability to establish new standards for quality, performance, or productivity
- Must have excellent writing and communication skills, with the ability to maintain open communication with internal employees, contractors, managers, third parties, and customers as needed
- Understanding of application development lifecycles, as well as practical experience working with continuous integration and continuous deployment (CI/CD) tools as part of the OpenStack platform lifecycle, will be useful
- Strong analytical and problem-solving capabilities
- Excellent project management and organizational skills
- Outstanding written and verbal communication abilities
- Proven leadership experience in technical environments