Principal AI Developer Technology Engineer
Munich, Bavaria, Germany · Full Time
Be the first to apply
- Experience
- 15+ yrs
- Salary
- —
- Openings
- 1
- Posted
- 19 hours ago
- Work mode
- In office
- Education
- Advanced degree in Computer Science, Computer Engineering or related computational science
- Resume
- Required to apply
Where you'll work
Sign in to tell us what does and doesn't work for you here — it sharpens every match we show you.
Job description
About the Role
We are searching for a Principal Developer Technology Engineer specializing in Artificial Intelligence to join our team. This position involves researching parallel algorithms to accelerate AI workloads on advanced computing architectures and optimizing systems to remove performance bottlenecks on leading-edge hardware. The role offers an opportunity to engage with the developer community and contribute to cutting-edge technology advancements at NVIDIA.
Key Responsibilities
- Conduct research and develop innovative techniques to accelerate workloads related to deep learning, machine learning, or other AI fields using GPUs.
- Collaborate closely with industry and academic experts to analyze and optimize complex AI and high-performance computing algorithms, ensuring peak performance on modern CPU and GPU platforms.
- Share optimized methods by publishing blogs or presenting at relevant conferences to educate and engage the developer community.
- Provide influential input on the design of future hardware architectures, software, and programming models by working alongside research, hardware, system software, libraries, and tools teams within NVIDIA.
Qualifications and Experience
- Advanced degree in Computer Science, Computer Engineering, or a related computational science discipline, or equivalent practical experience.
- Over 15 years of experience in software development or research.
- Proficient in programming with C/C++ and possess a comprehensive understanding of algorithms and software construction.
- Experience with parallel programming paradigms such as CUDA, OpenACC, OpenMP, MPI, or pthreads.
- Hands-on expertise in low-level performance optimization techniques.
- Strong foundational knowledge of CPU and GPU architecture.
- Excellent communication, organizational, logical problem-solving, time management, and prioritization capabilities.
Preferred Qualifications
- Specialized knowledge in parallelizing and optimizing Deep Learning models, especially for Natural Language Processing, Computer Vision, and Recommender Systems.
- Robust understanding of linear algebra concepts.
Equal Opportunity Statement
We are committed to creating a diverse workplace and are proud to be an equal opportunity employer. Our hiring and promotion policies do not discriminate based on race, religion, color, national origin, gender or gender expression, sexual orientation, age, marital status, veteran status, disability, or any other legally protected category.
Level
Lead
Minimum education
Master's Degree