What You'll Do:
RHEL Platform Management
- Own and maintain the RHEL OS fleet underpinning OCP worker/control-plane nodes and non-containerised application servers
- Plan and execute RHEL OS patching cycles (EUS channel management, kernel upgrades, reboot sequencing) with minimal service disruption
- Implement and validate OS hardening benchmarks — SELinux policies, firewalld rules, auditd configuration, and cryptographic policy settings
- Troubleshoot system-level issues: kernel parameters, resource limits (ulimits, cgroups), storage mounts (NFS, iSCSI, SMB/CIFS PVs on OCP)
- Manage RHEL subscriptions, entitlements, and satellite/image registries
OCP Cluster Administration
- Deploy, upgrade, and maintain OCP 4.x clusters — including control plane, etcd health, and worker node lifecycle
- Configure and manage OCP Machine Config Operator (MCO) for node-level OS customisation: chrony NTP, SSH hardening, kernel arguments, and custom MachineConfigs
- Manage cluster certificates, ingress/route configurations, HAProxy timeout tuning for long-running enterprise workloads (e.g., Pega BIX, batch jobs)
- Administer Quay mirror registry — image mirroring, vulnerability scanning results, VAPT remediation for bundled components (e.g., jQuery CVEs)
- Implement OCP RBAC, SCCs, network policies, and resource quotas across multiple namespaces and project environments
- Support JVM and container resource tuning for Java workloads on OCP (heap sizing, Metaspace limits, G1GC flags, container memory budgeting)
Application Platform & Integration Support
- Take ownership of one or more application platforms (e.g., Pega Infinity, WebSphere, IBM MQ) — covering setup in HA mode, application patching, and ongoing operational support
- Participate in architecture sessions with build teams, advising on integration points (web services, MQ, SFTP, API gateway)
- Support container provisioning, image lifecycle, and deployment of containerised services across DEV / SIT / UAT / PROD environments
Operational Governance & Security
- Define and execute operational processes: OS patching, application-level patching, performance monitoring, housekeeping, backup and recovery
- Support VAPT engagements — assess Nessus/vulnerability findings, coordinate remediation, and produce risk acceptance or remediation reports
What We're Looking For
- Degree in Computer Science, Information Technology, Information Systems, or a related discipline
- Hands-on experience administering RHEL 8/9 in enterprise production environments
- With minimum 3 years in TA / infra / cloud working experience
- Proficiency in OS-level patching, package management (dnf/yum), SELinux enforcement, and system hardening per CIS/DISA benchmarks
- Experience with RHEL Extended Update Support (EUS) lifecycle management and version upgrades
- Familiarity with RHEL subscription management, satellite server, and image-based deployments (RHEL Image Mode / Bootc)
- Hands-on experience deploying, configuring, and administering OCP 4.x clusters (bare metal or IaaS)
- Proficiency with OCP Day-2 operations: node management, Machine Config Operator (MCO), cluster upgrades, certificate rotation, and resource quota management
- Designing and operating multi-tier containerised architectures on OCP/Kubernetes
- Working knowledge of IBM MQ messaging — queue manager setup, channel administration, troubleshooting, and HA configuration