Senior Server and Storage / HPC System Engineer

Thequickwins โ€” United Arab Emirates ยท Posted ~2 weeks ago

Senior Full-time Onsite

Skills

Dell PowerEdge servers AMD GPUs Pure Storage Commvault Kubernetes HPC infrastructure RDMA InfiniBand Dell PowerEdge AMD MI210 GPU

๐Ÿ”“ Log in to save this job, tailor your resume & track your apply process โ€” 7 days free, no card needed.

Log in to add to target list

Summary

A senior infrastructure engineering role focused on designing and operating high-performance computing environments. The position requires expertise in servers, storage, GPU workloads, container platforms, backup systems, and large-scale infrastructure optimization.

Highlights

Work on advanced high-performance computing infrastructure with enterprise-scale systems, GPU acceleration, and opportunities to mentor engineers.

Description

๐Ÿšจ We're Hiring | Senior Server and Storage โ€“ Abu Dhabi ๐Ÿ“ FOR PEOPLE IN JORDAN ONLY Position: Senior Server and Storage / HPC System Engineer Location: Abu Dhabi, UAE We are currently looking for an experienced Senior Server and Storage Engineer for an opportunity based in Abu Dhabi, UAE. This opportunity is open to candidates currently based in Jordan. Jordanian professionals who meet the Job Overview below requirements are encouraged to apply. The Senior Server and Storage Engineer is responsible for designing, implementing, managing, and optimizing advanced HPC infrastructure solutions. The role focuses on Dell servers, AMD GPUs, Pure Storage systems, Commvault backup solutions, and Kubernetes environments to deliver scalable, high-performance, and reliable infrastructure solutions for enterprise and HPC workloads. Key Responsibilities Dell Servers: Architect and deploy HPC systems using DellPowerEdge servers, ensuring high availability and optimized performance for compute-intensive applications. Manage server hardware lifecycle, including deployment, upgrades, and diagnostics. Configure HPC cluster nodes for seamless integration with Kubernetes and GPU workloads. AMD GPUs (MI210) Deploy and optimize AMD GPU-based servers to accelerate AI/ML, HPC, and data-intensive applications. Monitor GPU utilization, troubleshoot performance bottlenecks, and optimize workloads for GPU acceleration. Integrate GPUs into Kubernetes environments for containerized GPU-based applications. Pure Storage Design and manage Pure Storage solutions, including FlashBlade, to support HPC and data-intensive workloads. Implement multitenancy configurations for isolated, secure, and efficient resource utilization. Monitor storage health and ensure performance optimization for high-speed data access. Commvault Backup Architect and manage enterprise-wide Commvault backup solutions, ensuring data integrity and readiness for disaster recovery. Implement backup and retention policies for HPC environments, including containerized and GPU-accelerated workloads. Kubernetes Container Management Deploy and manage Kubernetes clusters for HPC applications, ensuring scalability and fault tolerance. Configure persistent storage for containerized workloads and integrate storage with GPUs for high-performance data processing. Monitor cluster performance and troubleshoot HPC-specific Kubernetes challenges. System Optimization And Monitoring Implement advanced monitoring solutions for servers, GPUs, storage, and Kubernetes clusters to ensure peak performance. Develop and enforce policies for system security,resource allocation, and compliance with industry standards. Lead capacity planning and scaling initiatives for HPC infrastructure. Team Leadership And Collaboration Mentor and guide junior engineers on HPC best practices, system design, and troubleshooting techniques. Collaborate with cross-functional teams, including data scientists and DevOps teams, to align infrastructure capabilitieswith organizational goals. Required Qualifications Technical Skills: Extensive experience with Dell PowerEdge servers in HPC or enterprise environments. Proven expertise in AMD GPUs (MI210), including integration and optimization for AI/ML and HPC workloads. Advanced knowledge of Pure Storage systems, including multitenancy and high-performance configurations. Expertise in Commvault backup systems, including design, deployment, and disaster recovery. Strong proficiency in Kubernetes container orchestration, particularly for GPU-accelerated applications. Knowledge of high-performance interconnects (RDMA, InfiniBand) and networking for HPC environments. Soft Skills Strong problem-solving and analytical skills for addressing HPC-specific challenges. Effective communication and collaboration skills with technical and non-technical stakeholders. Leadership skills for mentoring and guiding junior team members. Preferred Qualifications Certifications in Dell EMC Proven Professional, AMD GPUs, Pure Storage, and Commvault. Work Environment Onsite role requiring hands-on management of HPC infrastructure, including Dell servers, AMD GPUs, Pure Storage, and Kubernetes clusters. Opportunity to work on state-of-the-art HPC systems and GPU-accelerated solutions. ๐Ÿ“ฉ How To Apply If you meet the above requirements and are currently based in Jordan, Please Send Your Updated CV To ๐Ÿ“ง info@thequickwins.com Email Subject: Senior Server and Storage Only shortlisted candidates will be contacted. #Hiring #ServerEngineer #StorageEngineer #HPC #Kubernetes #DellServers #PureStorage #AMD #Commvault #ITJobs #JordanJobs #Jordan #AbuDhabi #UAEJobs #InfrastructureEngineer #TheQuickWins