Senior Performance Analysis Engineer - #2214969
NVIDIA
Date: vor 4 Stunden
Stadt: München
Vertragstyp: Ganztags
Arbeitsplan: Volle Tag

Intelligent machines powered by Artificial Intelligence computers that can learn, reason and interact with people are no longer science fiction. GPU Deep Learning has provided the foundation for machines to learn, perceive, reason and solve problems. Today, visual computing is a crucial tool in helping people get along with technology, and NVIDIA has extended its technology into datacenters, mobile devices and cars. There has never been a more exciting time to join our team - if this role sounds like a fit for you, we'd love to hear from you!
NVIDIA is seeking a Senior High Performance Computing (HPC) and AI Networking Performance Research and Analysis Engineer to join our Performance group. In this exciting role, you will profile and analyze AI workloads on large GPUs and CPUs scale clusters for distributed Deep Learning LLM training focused on collectives communication and networking. You will interact with many types of hardware and platforms, such as HCAs, Switches, CPUs, GPUs, and Systems. You will develop performance analysis tools and methodologies to dive deeply into the details and understand performance expectations, limitations, and bottlenecks.
What You'll Be Doing
JR1998185
NVIDIA is seeking a Senior High Performance Computing (HPC) and AI Networking Performance Research and Analysis Engineer to join our Performance group. In this exciting role, you will profile and analyze AI workloads on large GPUs and CPUs scale clusters for distributed Deep Learning LLM training focused on collectives communication and networking. You will interact with many types of hardware and platforms, such as HCAs, Switches, CPUs, GPUs, and Systems. You will develop performance analysis tools and methodologies to dive deeply into the details and understand performance expectations, limitations, and bottlenecks.
What You'll Be Doing
- Exploring and researching AI workloads and DL models specifically tailored for large-scale deep learning LLM training on NVIDIA supercomputers and distributed systems focusing on high-performance networking and Nvidia Collective Communications Library (NCCL).
- Benchmarking, Profiling, and Analyzing the performance to find bottlenecks and identify areas of improvement and optimizations, with a strong emphasis on networking aspects.
- Implementing performance analysis tools.
- Collaborating with many teams from hardware to software to provide performance analysis insights.
- Defining performance test planning , setting performance expectations for new technologies and solutions, and working to reach the performance targets limits.
- B.Sc in Computer Science or Software Engineering or equivalent experience
- 5+ years of experience with high-performance Networking (RDMA, MPI, NCCL, Congestion Control Algorithms)
- Demonstrated Performance Analysis skills and methodologies.
- Experience with NVIDIA GPUs, CUDA library, deep learning frameworks like TensorFlow or PyTorch, combined with expertise in networking collective communication libraries (such as NCCL) and protocols (such as RoCE and RDMA).
- Fast and self-learning capabilities with strong analytical and problem-solving skills.
- Programming Languages: Python, Bash and C languages
- Experience with Linux OS distros.
- Great teammate with good communication and interpersonal skills
- In-depth knowledge and experience with AI workloads and benchmarking for distributed LLM training.
- Knowledge in CUDA, and NCCL libraries.
- Knowledge in Congestion Control algorithms.
- In-depth System knowledge and understanding (Intel / AMD / ARM CPUs, NVIDIA GPUs, HCA, Memory, PCI).
- Strong Performance Analysis skills and methodologies using modern tools.
JR1998185
Wie bewerbe ich mich?
Um sich für diesen Job zu bewerben, müssen Sie auf unserer Website autorisieren. Wenn Sie noch kein Konto haben, registrieren Sie sich bitte.
Veröffentlichen Sie einen LebenslaufÄhnliche Jobs
Senior Controller (mensch)
Omnicom Media Group,
vor 1 Stunde
Über die Hearts & Science… Wir sind eine Full-Funnel-Agentur, die die wichtigsten Marketing-Disziplinen in den Bereichen Performance Media, Technik und Kreativität miteinander verbindet, um die besten Marken- und Geschäftsergebnisse zu erzielen. Wir schätzen es, das große Ganze zu sehen und...

IT Governance & Compliance Specialist (d/w/m)
JobRad,
vor 2 Stunden
Richtlinienentwicklung: Konzeption, Implementierung und Pflege von IT-Richtlinien sowie Sicherheitsstandards zur Sicherstellung der Einhaltung gesetzlicher Vorgaben (z. B. DSGVO, ISO 27001, BSI IT-Grundschutz) Governance-Überwachung: Anwendung und Weiterentwicklung etablierter IT-Governance-Rahmenwerke (z. B. COBIT, ITIL), um einen strukturierten und regelkonformen IT-Betrieb zu gewährleisten...

Mobiler Fahrzeugbewerter – Region München & bundesweit (f/m/d)
CarOnSale,
vor 2 Stunden
CarOnSale ist ein Start-Up im Herzen von Berlin. Als Tech-Unternehmen haben wir es uns zur Aufgabe gemacht, reibungslose digitale Prozesse im Automobilhandel europaweit zu ermöglichen. Sei dabei und entwickle mit uns neue, innovative Lösungen für den B2B-Bereich. Wir suchen ab...
