Summary
A senior infrastructure engineering role focused on designing and operating high-performance computing environments. The position requires expertise in servers, storage, GPU workloads, container platforms, backup systems, and large-scale infrastructure optimization.
Highlights
Work on advanced high-performance computing infrastructure with enterprise-scale systems, GPU acceleration, and opportunities to mentor engineers.
Description
๐จ We're Hiring | Senior Server and Storage
โ Abu Dhabi
๐ FOR PEOPLE IN JORDAN ONLY
Position: Senior Server and Storage / HPC System Engineer
Location: Abu Dhabi, UAE
We are currently looking for an experienced Senior Server and Storage
Engineer for an opportunity based in Abu Dhabi, UAE.
This opportunity is open to
candidates currently based in Jordan.
Jordanian professionals who meet the
Job Overview
below requirements are encouraged to apply.
The Senior Server and Storage Engineer is responsible for
designing, implementing, managing, and optimizing advanced HPC infrastructure
solutions.
The role focuses on Dell servers, AMD GPUs, Pure Storage systems,
Commvault backup solutions, and Kubernetes environments to deliver scalable, high-performance,
and reliable infrastructure solutions for enterprise and HPC workloads.
Key Responsibilities
Dell Servers:
Architect and deploy HPC systems using DellPowerEdge servers, ensuring high availability and optimized performance for compute-intensive applications.
Manage server hardware lifecycle, including deployment, upgrades, and diagnostics.
Configure HPC cluster nodes for seamless integration with Kubernetes and GPU workloads.
AMD GPUs (MI210)
Deploy and optimize AMD GPU-based servers to accelerate AI/ML, HPC, and data-intensive applications.
Monitor GPU utilization, troubleshoot performance bottlenecks, and optimize workloads for GPU acceleration.
Integrate GPUs into Kubernetes environments for containerized GPU-based applications.
Pure Storage
Design and manage Pure Storage solutions, including FlashBlade, to support HPC and data-intensive workloads.
Implement multitenancy configurations for isolated, secure, and efficient resource utilization.
Monitor storage health and ensure performance optimization for high-speed data access.
Commvault Backup
Architect and manage enterprise-wide Commvault backup solutions, ensuring data integrity and readiness for disaster recovery.
Implement backup and retention policies for HPC environments, including containerized and GPU-accelerated workloads.
Kubernetes Container Management
Deploy and manage Kubernetes clusters for HPC applications, ensuring scalability and fault tolerance.
Configure persistent storage for containerized workloads and integrate storage with GPUs for high-performance data processing.
Monitor cluster performance and troubleshoot HPC-specific Kubernetes challenges.
System Optimization And Monitoring
Implement advanced monitoring solutions for servers, GPUs, storage, and Kubernetes clusters to ensure peak performance.
Develop and enforce policies for system security,resource allocation, and compliance with industry standards.
Lead capacity planning and scaling initiatives for HPC infrastructure.
Team Leadership And Collaboration
Mentor and guide junior engineers on HPC best practices, system design, and troubleshooting techniques.
Collaborate with cross-functional teams, including data scientists and DevOps teams, to align infrastructure capabilitieswith organizational goals.
Required Qualifications
Technical Skills:
Extensive experience with Dell PowerEdge servers in HPC or enterprise environments.
Proven expertise in AMD GPUs (MI210), including integration and optimization for AI/ML and HPC workloads.
Advanced knowledge of Pure Storage systems, including multitenancy and high-performance configurations.
Expertise in Commvault backup systems, including design, deployment, and disaster recovery.
Strong proficiency in Kubernetes container orchestration, particularly for GPU-accelerated applications.
Knowledge of high-performance interconnects (RDMA, InfiniBand) and networking for HPC environments.
Soft Skills
Strong problem-solving and analytical skills for addressing HPC-specific challenges.
Effective communication and collaboration skills with technical and non-technical stakeholders.
Leadership skills for mentoring and guiding junior team members.
Preferred Qualifications
Certifications in Dell EMC Proven Professional, AMD GPUs, Pure Storage, and Commvault.
Work Environment
Onsite role requiring hands-on management of HPC infrastructure, including Dell servers, AMD GPUs, Pure Storage, and Kubernetes clusters.
Opportunity to work on state-of-the-art HPC systems and GPU-accelerated solutions.
๐ฉ How To Apply
If you meet the above requirements and are currently based in Jordan,
Please Send Your Updated CV To
๐ง info@thequickwins.com
Email Subject: Senior Server and Storage
Only shortlisted candidates will be contacted.
#Hiring #ServerEngineer #StorageEngineer #HPC #Kubernetes #DellServers
#PureStorage #AMD #Commvault #ITJobs #JordanJobs #Jordan #AbuDhabi #UAEJobs
#InfrastructureEngineer #TheQuickWins