Exp: 6 to 8 Years
We are seeking a highly skilled Linux Principal Engineer II to provide senior technical leadership in designing, implementing, and supporting enterprise-grade Linux platforms. This role is critical to maintaining the reliability, scalability, and security of infrastructure that underpins business-critical services. The successful candidate will bring deep expertise in Linux systems, IBM Spectrum Scale (GPFS), and Red Hat OpenShift/Kubernetes, and will play a key role in driving innovation, operational excellence, and platform modernization.
Key Responsibilities
Serve as the technical lead for business-critical projects, owning architecture, design decisions, and successful delivery.
Design, implement, and maintain highly available, scalable Linux environments across on-premises and hybrid cloud platforms.
Administer and optimise IBM Spectrum Scale (GPFS) for high-performance, distributed storage and data management.
Deploy, manage, and enhance container platforms using OpenShift/Kubernetes, ensuring secure and efficient application delivery.
Ensure system stability, performance, availability, and data integrity across all Linux-based infrastructure.
Develop and implement security hardening standards, ensuring compliance with organizational and industry best practices.
Lead troubleshooting and root cause analysis for complex system and performance issues in distributed environments.
Drive automation initiatives using scripting and infrastructure-as-code to improve efficiency and reduce operational risk.
Collaborate with application, DevOps, security, and infrastructure teams to align platform capabilities with business needs.
Establish and enforce engineering standards, best practices, and documentation.
Mentor and provide technical guidance to engineers, fostering skill development and knowledge sharing.
Required Qualifications:
Extensive experience in Linux system engineering and administration (e.g., RHEL, CentOS, Oracle Linux, or equivalent).
Proven expertise with IBM Spectrum Scale (GPFS), including design, deployment, and performance tuning.
Strong experience with OpenShift and Kubernetes, including cluster management and container orchestration.
Deep understanding of system performance tuning, monitoring, and capacity planning.
Experience implementing security hardening, patching strategies, and compliance frameworks.
Strong scripting and automation skills (e.g., Bash, Python, Ansible, Terraform).
Solid networking knowledge (TCP/IP, DNS, load balancing, firewalls).
Experience with high-availability and disaster recovery architectures.
Demonstrated ability to lead complex technical projects and influence architectural decisions.
Key Competencies:
Strong problem-solving and analytical skills.
Excellent communication and stakeholder management abilities.
Leadership mindset with the ability to drive initiatives and mentor teams.
Proactive approach to continuous improvement and innovation.
Ability to perform effectively in fast-paced, high-impact environments.
Experience with hybrid or multi-cloud environments (e.g., Azure, AWS, GCP).
Familiarity with CI/CD pipelines and DevOps practices.
Knowledge of enterprise monitoring and logging tools (e.g., Prometheus, Grafana, ELK stack).
Relevant certifications (e.g., RHCE, RHCA, CKA, or equivalent).