MNC InsiderMNC Insider
NVIDIA logo

Senior Performance Engineer

NVIDIA

Senior Performance Engineer

full-timePosted: Jul 28, 2026Updated: Aug 27, 2026Yokneam, Israel

Job Description

NVIDIA is seeking a highly skilled Senior Performance Engineer to join our Performance and R&D organizations. In this role, you will help build and evolve systems that support performance analysis, telemetry, and optimization for large-scale GPU- and CPU-based clusters used in AI and high-performance computing environments. You will work closely with hardware, networking, firmware, and software teams to collect, analyze, and interpret performance data from live systems. This is a fast-paced R&D environment where system behavior and requirements evolve rapidly, requiring adaptable engineering solutions and strong analytical thinking.What you’ll be doing:Profile, benchmark, and analyze AI and HPC workloads on GPU and CPU clustersExplore performance characteristics of high-performance networking and collective communications (e.g., NCCL, RDMA, MPI, RoCE)Identify performance bottlenecks across networking, compute, memory, and system architectureDevelop and enhance performance analysis, benchmarking, and diagnostic toolsDefine performance test plans and establish expectations for new technologies and platformsCollaborate across hardware, firmware, networking, systems, and software teams to provide actionable performance insightsSupport telemetry collection and data refinement efforts to enable accurate performance analysisMaintain high standards for data quality, reproducibility, and traceability of performance resultsWhat we need to see:B.Sc. or M.Sc. in Computer Science, Computer Engineering, Software Engineering, or equivalent experience5+ years of experience in performance analysis, systems engineering, or HPC/AI infrastructureDemonstrated expertise in performance analysis skills and methodologiesHands-on experience with high-performance networking (RDMA, MPI, NCCL, congestion control)Strong understanding of system performance metrics (latency, throughput, resource utilization)Exposure to hardware, firmware, or embedded telemetry environmentsStrong analytical, problem-solving, and communication skillsAbility to work effectively in cross-functional, fast-paced R&D teamsWays to stand out from the crowd:Knowledge of CUDA, NCCL internals, and congestion control algorithmsDeep system-level understanding of CPU architectures, GPUs, HCAs, memory, and PCIeExperience with NVIDIA GPUs, CUDA, and deep learning frameworks such as PyTorch or TensorFlowExperience with cloud platforms Proficiency in Python; experience with Bash and C/C++ is a plus as well as a strong experience working in Linux environments

Locations

  • Yokneam, Israel

Responsibilities

  • Profile, benchmark, and analyze AI and HPC workloads on GPU and CPU clusters
  • Explore performance characteristics of high-performance networking and collective communications (e.g., NCCL, RDMA, MPI, RoCE)
  • Identify performance bottlenecks across networking, compute, memory, and system architecture
  • Develop and enhance performance analysis, benchmarking, and diagnostic tools
  • Define performance test plans and establish expectations for new technologies and platforms
  • Collaborate across hardware, firmware, networking, systems, and software teams to provide actionable performance insights
  • Support telemetry collection and data refinement efforts to enable accurate performance analysis
  • Maintain high standards for data quality, reproducibility, and traceability of performance results
  • Profile, benchmark, and analyze AI and HPC workloads on GPU and CPU clusters
  • Explore performance characteristics of high-performance networking and collective communications (e.g., NCCL, RDMA, MPI, RoCE)
  • Identify performance bottlenecks across networking, compute, memory, and system architecture
  • Develop and enhance performance analysis, benchmarking, and diagnostic tools
  • Define performance test plans and establish expectations for new technologies and platforms
  • Collaborate across hardware, firmware, networking, systems, and software teams to provide actionable performance insights
  • Support telemetry collection and data refinement efforts to enable accurate performance analysis
  • Maintain high standards for data quality, reproducibility, and traceability of performance results

Target Your Resume for "Senior Performance Engineer" , NVIDIA

Get personalized recommendations to optimize your resume specifically for Senior Performance Engineer. Takes only 15 seconds!

AI-powered keyword optimization
Skills matching & gap analysis
Experience alignment suggestions

Check Your ATS Score for "Senior Performance Engineer" , NVIDIA

Find out how well your resume matches this job's requirements. Get comprehensive analysis including ATS compatibility, keyword matching, skill gaps, and personalized recommendations.

ATS compatibility check
Keyword optimization analysis
Skill matching & gap identification
Format & readability score

Tags & Categories

GeneralGeneral

Answer 10 quick questions to check your fit for Senior Performance Engineer @ NVIDIA.

Quiz Challenge
10 Questions
~2 Minutes
Instant Score

Related Books and Jobs

No related jobs found at the moment.