We are now looking for a Senior Deep Learning Architect for LLM Inference!
NVIDIA is at the forefront of the generative AI revolution. The Inference Benchmarking (IB) team specifically focuses on advanced inference server performance for Large Language Models (LLMs). If you’re passionate about pushing the boundaries of GPU hardware and software performance and understand terms like pre-fill phase, generation phase, paged attention, MoE, Tensor Parallel, Llama, Mixtral, and HuggingFace, then this could be a great role for you!
What you’ll be doing:
What we need to see:
Ways to stand out from the crowd:
NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have a team of highly skilled and motivated individuals who excel in their work. If you have a proactive and independent approach, we want to hear from you!
The base salary range is 148,000 USD – 276,000 USD. Your base salary will be determined based on your location, experience, and the pay of employees in similar positions.
You will also be eligible for equity and benefits. NVIDIA accepts applications on an ongoing basis.
NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
Joining Razer will place you on a global mission to revolutionize the way the world games. Razer is a place...
How to applyWith over 18,000 employees worldwide, the Microsoft Customer Experience & Success (CE&S) organization is responsible for the strategy, design, and...
How to applyThe OpenShift Sales Architect SSA for AI (AI SSA) assumes a crucial role in providing expert deep technical sales support...
How to applyCLICK HERE TO APPLY Job Description Overview Help build the world’s most advanced reinforcement learning systems at Microsoft AI. We’re...
How to applyAre you passionate about Artificial Intelligence, Machine Learning and Deep Learning? Are you passionate about helping customers build solutions leveraging...
How to applyData and AI Solutions Architect Location: Dallas, TX Employment Type: Full Time About KiZAN We make technology personal! KiZAN is...
How to apply