Senior HPC Support Engineer - Compute and GPU Platform at NVIDIA | California, United States | Rezi

Senior HPC Support Engineer - Compute and GPU Platform at NVIDIA

Senior HPC Support Engineer - Compute and GPU Platform

NVIDIA · California, United States

1 months ago

Senior HPC Support Engineer - Compute and GPU Platform

NVIDIA · California, United States

a month ago
Resume preview

Impress employers and recruiters.
Choose from hundreds of resume examples.

Target Resume Now

About the Role

Provide comprehensive solutions for sophisticated installations, maintenance, or operations for a broad scope of AI hardware and software products. Act as a primary point of contact for customers, assisting with technical questions, debugging, and issue resolution. Interact regularly with Engineering, Marketing, and Support teams on technical issues.

Responsibilities

  • Resolve sophisticated customer concerns and technical issues through research, reproduction, and problem-solving for customers installing and supporting systems using Linux Operating Systems.
  • Debug and respond to user-reported issues via telephone, email, or conference calls on the DGX Platform (hardware and software).
  • Resolve customer issues during installation, operation, maintenance, product application, or interoperability.
  • Apply industry-standard AI tools to share debugging results, create knowledge base articles, and analyze customer issues.
  • Develop, redefine, and document standard methodologies for internal teams (Support/R&D) for support process improvements.
  • Participate in multi-functional team meetings and provide feedback to engineering and marketing regarding product requirements, customer experience, and support tools.

Requirements

  • 5+ years in providing in-depth customer support and debugging experience for hardware and software products.
  • Strong organizational skills and ability to prioritize/multi-task easily with limited supervision.
  • Proven use of established AI technologies in day-to-day job responsibilities.
  • Established knowledge of Enterprise platform and systems engineering.
  • Understanding of Linux triage and servers.
  • Ability to resolve hardware and/or OS internal issues.
  • Excellent verbal and written English skills.
  • Intellectual curiosity, positive attitude, flexibility, analytical ability, self-motivation, and team-oriented.
  • Professional-level communication skills, interpersonal skills with a passion to solve problems.

Skills

  • Linux System Administration on engineering and networking level (Red Hat Enterprise Linux and Ubuntu distributions preferred).
  • Deep understanding of at least two of the following: data centers, servers, distributed systems, virtualization, deep learning frameworks, containers/containerization (Docker, Kubernetes).
  • Proficient working in various Linux environments.
  • Experience with other operating systems like Microsoft Windows and VMWARE.
  • Knowledge and working experience with InfiniBand, RDMA/RoCEv2 and GPU Technology.
  • Clustering or HPC Data-Center technologies including Upper Layer Protocols (MPI, NCCL).
  • Shell scripting (Bash/Python).
  • Ethernet and Distributed File System Storage technologies.
  • Configuration and operational expertise with traditional network switch/router and Open platforms.

Location

  • Remote

Work Type

  • Full-time

Experience Level

  • Senior

Education Level

  • Academic degree from an accredited university or college in Networking, Computer Science/Engineering, or Electrical/IT (or equivalent experience).

Salary/Compensations

  • 108,000 USD - 172,500 USD for Level 3
  • 120,000 USD - 207,000 USD for Level 4

Benefits

  • Equity
  • Comprehensive benefits package

About the Company

  • NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years.
  • Today, we’re tapping into the unlimited potential of AI to define the next era of computing.
  • NVIDIA uses AI tools in its recruiting processes.
  • NVIDIA pioneered accelerated computing. Today, our AI infrastructure powers global intelligence, transforming every industry.

Equal Opportunity

  • NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer.
  • As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.