Role Overview:We are seeking a highly skilled System Engineer to support and enhance enterprise infrastructure platforms in a client-facing environment. This role emphasizes expertise in OpenShift Container Platform and Linux Administration, requiring close collaboration with a distributed team and direct engagement with client stakeholders.
Key Responsibilities:- Design, build, configure, and maintain on-premises OpenShift Container Platform clusters.
- Design, build, configure, and maintain Red Hat Enterprise Linux Server infrastructure across on-premises, Amazon AWS, and Microsoft Azure environments.
- Automate server build, configuration, and deployment processes using shell scripting and Ansible.
- Develop, maintain, and enhance orchestration using Rundeck for automated operational activities.
- Act as an on-call escalation engineer for after-hours support and rapid resolution of high-severity incidents.
- Perform advanced troubleshooting, root-cause analysis, and performance tuning for OpenShift Container Platform and Red Hat Enterprise Linux.
- Engage in daily client interactions, providing technical updates and supporting service reviews.
- Collaborate with cross-functional engineering teams to deliver reliable and scalable solutions.
- Support change management activities, releases, and production deployments.
- Create and maintain technical documentation, build standards, automation artifacts, and operational runbooks.
- Identify opportunities for automation, optimization, and service improvement.
Required Skills:- Strong hands-on experience with Red Hat OpenShift Container Platform (installation, upgrades, cluster administration, troubleshooting, and lifecycle management).
- Strong experience with Red Hat Enterprise Linux (RHEL) administration (build, patching, performance tuning, and troubleshooting).
- Solid understanding of Kubernetes architecture and core concepts (pods, deployments, networking, storage, RBAC).
- Experience with automation and configuration management using shell scripting.
- Experience supporting production environments, including change management, incident response, root cause analysis, and on-call rotations.
- Strong communication and stakeholder engagement skills.
Qualifications:- 5+ years of experience in System Engineering, Infrastructure Engineering, or Platform Engineering roles.
- Bachelor's degree in Computer Science, Engineering, or equivalent practical experience.
Preferred Skills:- Experience developing and maintaining Ansible playbooks and automation frameworks.
- Experience with CI/CD tools such as GitHub.
- Familiarity with monitoring and observability tools (e.g., Prometheus, Grafana, Nagios).
- Understanding of security best practices.
- Experience with Rundeck or similar orchestration tools.