The Infrastructure & Platforms Engineer will play a critical role in administering, automating, and supporting enterprise Linux and infrastructure environments, with a strong emphasis on RHEL, Red Hat Satellite, and Ansible Automation. This position ensures the availability, security, performance, patching, and lifecycle management of Linux platforms across production, non-production, and disaster recovery environments. Additionally, the role supports virtualization, storage, backup, networking, and other key enterprise infrastructure technologies to maintain a robust and efficient IT ecosystem.
Key Responsibilities
- Administer and support enterprise RHEL environments across production, non-production, and disaster recovery platforms.
- Perform Linux installation, configuration, patching, upgrades, hardening, performance tuning, and troubleshooting.
- Manage Linux users, groups, filesystems, LVM, services, processes, permissions, SSH, SELinux, and overall system security.
- Oversee the RHEL lifecycle using Red Hat Satellite, including host registration, repositories, Content Views, Lifecycle Environments, Activation Keys, errata, and subscription management.
- Plan and execute enterprise Linux patching, remediation, and compliance activities through Satellite.
- Develop and maintain Ansible playbooks, roles, inventories, variables, templates, and handlers to automate infrastructure tasks.
- Automate server provisioning, configuration management, patching, compliance, application prerequisites, and routine operations using Ansible.
- Support and troubleshoot Ansible Automation Platform, Ansible Tower, or AWX environments, including failed automation jobs.
- Troubleshoot complex Linux OS, networking, filesystem, performance, and infrastructure-related issues.
- Support Linux high-availability and cluster environments where applicable.
- Administer and troubleshoot VMware vSphere/ESXi environments, including VM provisioning, migration, snapshots, resource management, and high availability.
- Support enterprise SAN/NAS storage systems, including Fibre Channel, LUNs, multipathing, NFS, and storage provisioning.
- Support enterprise backup and recovery environments using Commvault, Veeam, or equivalent technologies.
- Participate in infrastructure migration, technology refresh, disaster recovery, capacity planning, and platform upgrade projects.
- Monitor infrastructure health, availability, capacity, and performance metrics.
- Perform root-cause analysis and implement permanent solutions for recurring infrastructure issues.
- Coordinate with OEM and vendor support teams for Linux, hardware, virtualization, storage, backup, and platform-related issues.
- Maintain infrastructure documentation, standard operating procedures, technical guides, and operational runbooks.
- Participate in incident, problem, change, and release management processes.
- Provide technical support during critical production activities and planned maintenance windows.
Required Qualifications
- Bachelor’s degree in Computer Science, IT, Engineering, or a related discipline.
- Minimum of 5 years of relevant experience in enterprise infrastructure and Linux administration.
- Strong hands-on expertise with RHEL versions 7, 8, 9, or later.
- Proficient in Linux system administration and troubleshooting.
- Experience with Red Hat Satellite, including Content Views, Lifecycle Environments, Activation Keys, repositories, host registration, errata, and patch management.
- Solid practical experience developing Ansible playbooks, roles, inventories, variables, templates, and handlers.
- Good knowledge of Linux components such as LVM, filesystems, systemd, SELinux, SSH, RPM/YUM/DNF package management, NFS, DNS/NTP, networking, user/access management, and Bash scripting.
- Experience with Linux security hardening, vulnerability remediation, and OS performance monitoring.
- Hands-on experience with VMware vSphere/ESXi; familiarity with Hyper-V is a plus.
- Understanding of enterprise SAN/NAS storage, Fibre Channel, LUNs, multipathing, and NFS.
- Experience with backup technologies such as Commvault, Veeam, or equivalents, including backup, recovery, disaster recovery, RPO, and RTO concepts.
- Solid understanding of TCP/IP, DNS, NTP, VLANs, routing, firewalls, load balancing, and Linux network configuration.
- Strong troubleshooting, analytical, communication, documentation, and problem-solving skills.
- Ability to work independently and collaborate effectively with cross-functional infrastructure and application teams.
- Willingness to participate in on-call and after-hours support as needed.
Preferred Qualifications and Benefits
- Experience with Ansible Automation Platform, Ansible Tower, or AWX is preferred.
- Exposure to Red Hat OpenShift or Kubernetes is an advantage.
- Relevant certifications such as RHCSA, RHCE, Red Hat Ansible Automation, VMware VCP, or ITIL are highly desirable.
This role offers the opportunity to work in a dynamic enterprise environment, contributing to critical infrastructure operations and automation initiatives that drive efficiency and reliability.