AI Accelerator, Software Engineer- Graph Optimization/Compilers

Santa Clara, CA, US • Posted 2 hours ago • Updated 2 hours ago
Full Time
On-site
USD $159,000.00 - 239,000.00 per year
Fitment

Dice Job Match Score™

🔢 Crunching numbers...

Job Details

Skills

  • Semiconductors
  • CPU
  • Energy
  • High Performance Computing
  • Cloud Computing
  • Scalability
  • PyTorch
  • Fusion
  • Regression Analysis
  • Collaboration
  • Design Review
  • Documentation
  • Computer Science
  • Computer Engineering
  • Mathematics
  • Data Structure
  • Algorithms
  • Python
  • C
  • C++
  • Open Source
  • Benchmarking
  • Computer Hardware
  • Deep Learning
  • Neural Network
  • Layout
  • CUDA
  • OpenCL
  • GPU
  • NPU
  • Transformer
  • Caching
  • Optimization
  • Problem Solving
  • Conflict Resolution
  • Platinum DB2 for z/OS
  • Research
  • Production Engineering
  • Analytical Skill
  • Quick Learner
  • Artificial Intelligence
  • Testing
  • Debugging
  • Code Review
  • Health Insurance
  • Insurance
  • Finance
  • Military
  • SAP BASIS
  • Law

Summary

Description

Invent the future with us.

Ampere is a semiconductor design company for a new era, leading the future of computing with an innovative approach to CPU design focused on high-performance, energy efficient AI compute.

As a pioneer in the new frontier of energy efficient high-performance computing, Ampere is part of the Softbank Group of companies driving sustainable computing for AI, Cloud, and edge applications.

Join us at Ampere and work alongside a passionate and growing team - we'd love to have you apply!

About the Role:

As a Software Engineer on Ampere's AI Accelerator team, you will optimize deep learning computational graphs to maximize the performance, efficiency, and scalability of Ampere's AI accelerator hardware. You will work across the software stack, from model frameworks and inference-serving systems to graph optimization, compiler infrastructure, runtimes, and compute kernels.

What You'll Achieve:
  • Optimize computational graphs for performance, throughput, latency, memory efficiency, and power efficiency on Ampere AI accelerators
  • Enable and optimize models, frameworks, and inference platforms, including PyTorch, Llama.cpp, vLLM, and SGLang
  • Develop graph-level optimizations such as operator fusion, pattern matching, redundancy elimination, constant folding, layout optimization, memory planning, quantization, and accelerator offload
  • Optimize transformer and LLM workloads, including dynamic shapes, attention mechanisms, KV-cache management, and mixed-precision execution
  • Analyze end-to-end performance across frameworks, compilers, runtimes, kernels, and hardware
  • Build profiling, benchmarking, validation, and performance-regression infrastructure
  • Identify bottlenecks using traces, compiler diagnostics, microbenchmarks, and hardware performance data
  • Collaborate with compiler, runtime, kernel, architecture, hardware, and applications teams on hardware/software co-design
  • Contribute to architecture, design reviews, code reviews, documentation, and engineering best practices

About You:
  • Bachelor's degree in Computer Science, Computer Engineering, Mathematics, or a related technical field & 5 years of relevant experience; or a Master's degree with & 3 years of relevant experience
  • Strong foundations in algorithms, data structures, graph algorithms, computational complexity, and systems programming
  • Proficiency in Python and C/C++, demonstrated through internships, research, coursework, open-source contributions, or personal projects
  • Strong ability to reason about execution dependencies, memory movement, numerical correctness, and hardware execution behavior
  • Experience diagnosing performance issues through profiling, benchmarking, tracing, or hardware-level analysis is a plus
  • Familiarity with deep learning concepts, neural-network architectures, tensor operations, numerical precision, quantization, and memory layout
  • Experience with CUDA, ROCm, OpenCL, SYCL, Triton, GPU programming, NPU programming, or other accelerator architectures is a plus
  • Familiarity with transformer models, LLM inference, attention mechanisms, KV-cache optimization, speculative decoding, mixed-precision execution, or sparsity is a plus
  • Demonstrated exceptional problem-solving ability-IOI medal, ACM ICPC medal, Codeforces Grandmaster, USACO Platinum, or equivalent achievement in research or production engineering is a strong plus
  • Strong analytical and debugging skills, with the ability to investigate ambiguous technical problems and deliver robust solutions
  • Fast learner who can quickly understand new architectures, frameworks, compilers, and workloads
  • Experience using AI-assisted development tools to accelerate implementation, testing, debugging, and code review while maintaining technical ownership and code quality

What We'll Offer:

At Ampere we believe in taking care of our employees and providing a competitive total rewards package that includes base pay, cash long-term incentive, and comprehensive benefits. The full base pay range for this role is between $159,000 and $239,000. Our benefits include health, wellness, and financial programs that support employees through every stage of life.

Benefit highlights include:
  • Premium medical insurance, dental insurance, vision insurance, as well as income protection and a 401K retirement plan, so that you can feel secure in your health and financial future.
  • Unlimited Flextime and 10+ paid holidays so that you can embrace a healthy work-life balance.
  • A variety of healthy snacks, energizing espresso, and refreshing drinks to keep you fueled and focused throughout the day.

And there is much more than compensation and benefits. At Ampere, we foster an inclusive culture that empowers our employees to do more and grow more. We are excited to share more about our career opportunities with you through the interview process. Our benefits include health, wellness, and financial programs that support employees through every stage of life.

#LI-Hybrid #LI-DR

#LI-Hybrid

Ampere is an inclusive and equal opportunity employer and welcomes applicants from all backgrounds. All qualified applicants will receive consideration for employment without regard to race, color, national origin, citizenship, religion, age, veteran and/or military status, sex, sexual orientation, gender, gender identity, gender expression, physical or mental disability, or any other basis protected by federal, state or local law.
Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 80184296
  • Position Id: 5277ba2d494ae0583eb359dfa0a3faa
  • Posted 2 hours ago
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Santa Clara, California

Today

Full-time

USD 195,000.00 - 292,000.00 per year

San Jose, California

Today

Full-time

USD 210,000.00 per year

Mountain View, California

Today

Full-time

USD 142,800.00 - 274,800.00 per year

Cupertino, California

Today

Full-time

Search all similar jobs