Linux Specialist
Thuwal, Makkah Province, Saudi Arabia · Full Time
Be the first to apply
- Experience
- Any
- Salary
- —
- Openings
- 1
- Posted
- 4時間前
- Work mode
- In office
- Resume
- Required to apply
Where you'll work
Sign in to tell us what does and doesn't work for you here — it sharpens every match we show you.
Job description
Overview
This role focuses on providing advanced Linux system administration and support, specifically at Level 2 and Level 3, to troubleshoot and resolve complex issues. The position involves maintaining and optimizing Linux infrastructure with a focus on scientific computing and GPU workloads.
Key Duties
- Deliver expert support for Linux systems, addressing complex incidents and system errors.
- Troubleshoot core Linux operating system challenges, including boot failures, kernel problems, and service disruptions.
- Manage OS patching processes, vulnerability assessments, and enhance system security through hardening practices.
- Optimize performance metrics related to CPU, memory, storage, and particularly NVIDIA GPU workloads.
- Create and maintain automation scripts and configurations using Ansible, Puppet, Bash, and Python.
- Oversee infrastructure lifecycle management via GitLab CI/CD pipelines and infrastructure-as-code principles.
- Support virtual desktop infrastructure tools such as SLURM, AWS DCV, Docker, and Kasm Workspaces.
- Configure scientific runtime environments utilizing Lmod or Environment Modules systems.
- Administer and manage Linux agents including Automox, Nessus, and Puppet for security and maintenance.
- Support network storage solutions such as NFS and SMB; also troubleshoot issues with containerized workloads.
- Conduct root cause analysis to identify underlying problems and develop automation to prevent recurrence.
- Liaise with broader infrastructure teams to resolve issues beyond the VDI environment.
Required Expertise
- Proficient in Linux administration of Ubuntu LTS and Rocky Linux distributions.
- Experience working with NVIDIA GPU-enabled Linux systems.
- Strong familiarity with automation tools such as Ansible, Puppet, Automox, and GitLab CI/CD.
- Skilled in scripting languages including Bash and Python for system automation.
- Practical knowledge of Docker, Kasm Workspaces, SLURM, and AWS DCV virtual desktop technologies.
- Expertise in Linux OS patching, system hardening, and performance tuning.
- Experience managing scientific environment modules using Lmod or similar.
- Competence administering NFS and SMB network storage.
- Capability to provide onsite support in an infrastructure environment.
Additional Notes
The position requires robust troubleshooting skills and the ability to collaborate with multi-disciplinary infrastructure teams. Experience with scientific runtime environments and GPU compute workloads is essential due to the specialized nature of the infrastructure support.