Principal Engineer, Service Delivery

Senior GenAI & HPC Engineer (Service Delivery Principal Engineer)
Dell Technologies customers rely on our products and services to drive progress. So, we take the service we provide extremely seriously. Service Delivery is all about making sure our technical solutions help clients fulfil their priorities, challenges and initiatives. As trusted advisors, we build in-depth knowledge of what each client wants to achieve. Then we make sure the services delivered by Dell Technologies deliver on all our promises. We also work closely with Sales and Global Services colleagues to develop strategic account growth plans, and to identify and pursue sales opportunities.
What you’ll achieve:
We’re seeking a Senior GenAI & HPC Engineer with deep experience in GPU accelerated systems, Linux performance tuning, and benchmarking. This role is highly hands on and customer facing, supporting onsite deployments across the South East Asia/APJ for advanced HPC and GenAI solutions. You will work as a part of a team to help build, integrate, and test some of the world’s largest multi GPU systems, benchmark them using industry standard tools, make suggestions on how to optimize performance, and deliver the next generations of AI/HPC infrastructure.
Join us to do the best work of your career and make a profound social impact as a Senior GenAI & HPC Engineer on our Service Delivery Team in Malaysia.
Responsibilities
You will:
- Design and deliver advanced service solutions.
- Develop automation and monitoring; reduce toil.
- Conduct root-cause analyses and corrective actions.
- Prepare handover and operational documentation.
Qualifications
Take the first step towards your dream career
Every Dell Technologies team member brings something unique to the table. Here’s what we are looking for with this role:
Essential Requirements
- 8+ years of related experience
- Experience in deploying GPU accelerated compute clusters for AI with NVIDIA Base Command Manager especially in a NVL72 environment
- Experience in implementing large scale GPU networking (more than 100,000 connections)
- Experience in configuring L2/L3 Leaf/Spine networking using Nvidia Spectrum switches/Cumulus OS
- Experience in InfiniBand switches
- Experience in performing L2/L3 networking using Sonic OS (Dell/Nvidia Switches)
- Strong troubleshooting, problem-solving, and stakeholder management skills
- Network cabling design will be an added experience
- Experience in deployment of Air Cooled and/or Liquid cooled racks will be an added advantage
- Knowledge of Kubernetes, Ubuntu, Openshift will be advantageous
- High amount of travel across SEA
- Flexibility to support project activities outside office hours when required
Desirable Requirements
- Bachelor’s degree in Engineering, Computer Science, or related field
- Knowledge of Kubernetes, Ubuntu, Openshift will be advantageous
Who We Are
We believe that each of us has the power to make an impact. That’s why we put our team members at the center of everything we do. If you’re looking for an opportunity to grow your career with some of the best minds and most advanced tech in the industry, we’re looking for you.
Dell Technologies is a unique family of businesses that helps individuals and organizations transform how they work, live and play. Join us to build a future that works for everyone because Progress Takes All of Us.
Dell Technologies is committed to the principle of equal employment opportunity for all employees and to providing employees with a work environment free of discrimination and harassment. Read the full Equal Employment Opportunity Policy.
Visit our Culture Code page to learn more about how we work and lead.
You'll be redirected to
the company's application page