Principal Engineer, Efficient GenAI

San Jose, CA, US • Posted 9 days ago • Updated 10 hours ago
Full Time
On-site
USD 210,000.00 per year
Fitment

Dice Job Match Score™

🔢 Crunching numbers...

Job Details

Skills

  • Health Care
  • Large Language Models (LLMs)
  • 3D Computer Graphics
  • Video
  • Art
  • Algorithms
  • Transformer
  • Optimization
  • Caching
  • Open Source
  • Computer Hardware
  • Workflow
  • Collaboration
  • PyTorch
  • JAX
  • Innovation
  • Publications
  • Generative Artificial Intelligence (AI)
  • Training
  • Presentations
  • Deep Learning
  • Software Development
  • Computer Science
  • Electrical Engineering
  • Mathematics
  • Machine Vision
  • Military
  • Law
  • Recruiting
  • Artificial Intelligence

Summary

ADVANCE YOUR CAREER. ADVANCE THE WORLD.

At AMD, we believe technology has the power to solve the world's most important challenges. From advancing healthcare and scientific discovery to powering AI and the technologies people rely on every day, innovation at AMD is shaping the future.

Whether you're designing next-gen processors, enabling AI breakthroughs, or bringing leading edge products to market, every role at AMD contributes to something bigger - technology that moves the world forward. Join us and, together, we'll advance your career.

THE ROLE:

The AI Models and Applications team at AMD is looking for a specialized Principal Engineer who is passionate about enabling innovative and efficient Generative AI training and inference at scale. You will be part of a core team of incredibly talented specialists and work on scaling training and inference for the latest Generative AI models.

THE PERSON:

You have a deep technical understanding and hands-on experience with the latest Generative AI applications in at least one of the following areas: large language models (LLMs), 3D World and Action Models, or image/video generation models. You have experience training models at scale and are passionate about developing efficient approaches to enable distributed training and inference on AMD devices.

Why Join Us?
  • Exciting Opportunities: As a senior member of the team, you will be at the forefront of innovation, working with the latest Generative AI models and algorithms. You will have the opportunity to shape the future of AI model training and inference optimization across a variety of applications.
  • Talented Team: Join a team of highly skilled industry specialists who are passionate about pushing the boundaries of AI. Collaborate with like-minded professionals and learn from the best in the field.
  • Cutting-Edge Technology: Work with state-of-the-art Generative AI algorithms and software, enabling you to stay ahead of the curve and drive advancements in AI model training at scale and deployment.
  • Impactful Work: Your contributions will directly influence how cutting-edge Generative AI models across the industry are efficiently trained at scale, as well as how inference solutions are deployed to serve millions of customers, making a significant difference across various industries and applications.

KEY RESPONSIBILITIES:
  • Propose and apply innovative techniques to support both training and inference, including innovative transformer architectures, parallelism strategies for training on large clusters, inference optimization techniques such as speculative decoding, and optimal KV-caching strategies.
  • Implement novel, efficient architectures for Generative AI models for training and inference and showcase the benefits on AMD platforms.
  • Work with open-source frameworks and communities (e.g., PyTorch, JAX, vLLM, SGLang) to integrate AMD-optimized models and libraries and publish training recipes.
  • Collaborate with software and hardware teams to co-optimize end-to-end performance on current and future AMD solutions.
  • Increase adoption of agentic workflows for optimizing and deploying Generative AI applications at scale on AMD platforms.
  • Publish and promote your work at external venues, including major conferences.
  • Collaborate with researchers within AMD and across industry and academia to promote innovation on AMD platforms.

PREFERRED EXPERIENCE:
  • Strong technical expertise in Generative AI model training and inference, with familiarity working with deep learning frameworks such as PyTorch, JAX, vLLM, SGLang, and MuJoCo.
  • Strong technical expertise in algorithmic innovation for efficient Generative AI applications across both training and inference.
  • Expertise and publications in one or more of the following preferred areas: efficient model architectures, optimized training, innovative parallelism strategies, or low-precision training.
  • Additional plus if publications have been presented at conferences such as NeurIPS, CVPR, ECCV, ICCV, ICML, or ICLR.
  • Experience productizing Generative AI models and training foundation models at scale.
  • Excellent written, verbal, and presentation skills, with the ability to coordinate effectively both internally and externally.
  • Several years of experience in AI, deep learning, and related software development.

ACADEMIC CREDENTIALS:

PhD or master's degree in computer science, Electrical Engineering, Mathematics, or a related field.

LOCATION:

San Jose, CA (Hybrid)

Alternative locations in Seattle, WA, or Austin, TX may be considered.

#LI-MV1

#LI-HYBRID

Benefits offered are described: AMD benefits at a glance.

AMD does not accept unsolicited resumes from headhunters, recruitment agencies, or fee-based recruitment services. AMD and its subsidiaries are equal opportunity, inclusive employers and will consider all applicants without regard to age, ancestry, color, marital status, medical condition, mental or physical disability, national origin, race, religion, political and/or third-party affiliation, sex, pregnancy, sexual orientation, gender identity, military or veteran status, or any other characteristic protected by law. We encourage applications from all qualified candidates and will accommodate applicants' needs under the respective laws throughout all stages of the recruitment and selection process.

AMD may use Artificial Intelligence to help screen, assess or select applicants for this position. AMD's "Responsible AI Policy" is available here.

This posting is for an existing vacancy.
Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 10127278
  • Position Id: d0364ace3c7c80c287ed11b96aff419c
  • Posted 9 days ago
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

San Jose, California

•

Today

Full-time

USD 151,800.00 - 265,350.00 per year

Mountain View, California

•

Today

Full-time

USD 102,100.00 - 202,200.00 per year

Foster City, California

•

13d ago

Full-time

USD 66.00 - 73.00 per hour

Foster City, California

•

8d ago

Full-time

USD 169,100.00 - 270,800.00 per year

Search all similar jobs