T
- Experience
- Any
- Salary
- —
- Openings
- 1
- Posted
- 32 minutes ago
- Work mode
- In office
- Resume
- Required to apply
Where you'll work
Sign in to tell us what does and doesn't work for you here — it sharpens every match we show you.
Job description
Job Overview
We are looking for an experienced DevOps Engineer to provide advanced Linux system support and infrastructure management at our Thuwal, Makkah facility. This position involves highly technical troubleshooting, performance optimization, automation development, and support for specialized computing environments.
Key Responsibilities
- Deliver second and third-level support for Linux systems by diagnosing and resolving complex operational incidents.
- Troubleshoot various Linux operating system issues including failures during boot, kernel complications, service disruptions, package management challenges, system upgrades, and migrations.
- Execute routine operating system patching, enforce vulnerability management, and implement system hardening to maintain security compliance.
- Enhance Linux system performance across CPU, memory, storage subsystems, and NVIDIA GPU workloads to ensure efficient resource utilization.
- Create and maintain automation scripts and frameworks using Ansible, Puppet, Bash, and Python to streamline operations and deployments.
- Manage infrastructure life cycle with GitLab CI/CD pipelines and employ Infrastructure as Code practices to maintain consistent environments.
- Support virtual desktop infrastructure technologies including SLURM workload manager, AWS DCV, Docker containers, and Kasm Workspaces.
- Configure and maintain scientific runtime environments through Lmod and Environment Modules to facilitate user workflows.
- Administer Linux agents like Automox for patch management, Nessus for vulnerability scanning, and Puppet for configuration management.
- Maintain and troubleshoot NFS and SMB storage solutions as well as containerized workloads ensuring high availability and performance.
- Conduct root cause analysis on incidents and develop preventative automation solutions to reduce recurrence.
- Collaborate with infrastructure teams to address complex issues that extend beyond the scope of VDI layers.
Skills and Expertise
- Strong proficiency in Linux system administration focusing on Ubuntu LTS and Rocky Linux distributions.
- Experience managing Linux environments that utilize NVIDIA GPUs for computing workloads.
- Hands-on expertise with automation tools such as Ansible and Puppet.
- Familiarity with patch management tools like Automox and vulnerability scanners like Nessus.
- Proficient with GitLab CI/CD workflows and Infrastructure as Code methodologies.
- Advanced scripting skills in Bash and Python for automation and operational improvements.
- Practical knowledge of container technologies including Docker and workspace management with Kasm Workspaces.
- Working experience with SLURM workload manager and AWS DCV for virtual desktop infrastructure support.
- Knowledge of networked file system protocols including NFS and SMB.
- Experience configuring scientific runtime environments using Lmod and Environment Modules.
- Skills in Linux OS patching, security hardening, and overall system performance tuning.
Tools & software
Python
required