Research Engineer (AI Inference)

San Francisco, CA, US • Posted 9 hours ago • Updated 9 hours ago
Full Time
On-site
USD $175,000.00 - 225,000.00 per year
Fitment

Dice Job Match Score™

👤 Reviewing your profile...

Job Details

Skills

  • Health Care
  • Performance Tuning
  • Computer Hardware
  • High Performance Computing
  • Optimization
  • CUDA
  • Deep Learning
  • Benchmarking
  • Python
  • C++
  • GPU
  • Machine Learning (ML)
  • Performance Engineering
  • Research
  • Startups
  • Military
  • SAP BASIS
  • Authorization
  • Law
  • LOS
  • Recruiting
  • Legal
  • Artificial Intelligence
  • Privacy

Summary

This Jobot Job is hosted by: Grant Greenhalgh
Are you a fit? Easy Apply now by clicking the "Apply Now" button and sending us your resume.
Salary: $175,000 - $225,000 per year

A bit about us:

We're a well-funded AI infrastructure startup developing modern software at the intersection of artificial intelligence, high-performance computing, and specialized hardware. The team is tackling complex performance challenges associated with running modern AI workloads across emerging compute architectures.

Why join us?
  • Well-funded by leading tech investors
  • Cutting edge technical problems with complex solutions
  • Lucrative Equity in a seed stage startup
  • Competitive compensation
  • Excellent benefits (healthcare, vision, dental)


Job Details

We're looking for a Research Engineer focused on AI inference and performance optimization. This person will work close to the hardware, developing and optimizing the low-level software responsible for running modern AI models efficiently across high-performance compute platforms. This is an ideal opportunity for an engineer who enjoys working at the intersection of AI systems, GPU programming, performance engineering, and low-level optimization.

What We're Looking For

  • Strong experience with GPU programming, high-performance computing, or AI inference optimization
  • Hands-on experience with CUDA, ROCm, Triton, or similar low-level compute frameworks
  • Experience developing or optimizing GPU kernels
  • Strong understanding of modern deep learning and LLM inference workloads
  • with performance-critical operations including matrix multiplication and attention
  • Experience benchmarking, profiling, and optimizing compute-intensive workloads
  • Strong Python and/or C++ programming experience
  • Understanding of GPU architecture, memory hierarchy, parallelism, and performance bottlenecks
  • Experience with inference frameworks or serving infrastructure is highly valuable
  • Background in ML systems, distributed systems, compilers, or performance engineering is a plus
  • Ability to operate effectively in a highly technical, research-oriented startup environment


Interested in hearing more? Easy Apply now by clicking the "Apply Now" button.

Jobot is an Equal Opportunity Employer. We provide an inclusive work environment that celebrates diversity and all qualified candidates receive consideration for employment without regard to race, color, sex, sexual orientation, gender identity, religion, national origin, age (40 and over), disability, military status, genetic information or any other basis protected by applicable federal, state, or local laws. Jobot also prohibits harassment of applicants or employees based on any of these protected categories. It is Jobot's policy to comply with all applicable federal, state and local laws respecting consideration of unemployment status in making hiring decisions.

Sometimes Jobot is required to perform background checks with your authorization. Jobot will consider qualified candidates with criminal histories in a manner consistent with any applicable federal, state, or local law regarding criminal backgrounds, including but not limited to the Los Angeles Fair Chance Initiative for Hiring and the San Francisco Fair Chance Ordinance.

Information collected and processed as part of your Jobot candidate profile, and any job applications, resumes, or other information you choose to submit is subject to Jobot's Privacy Policy, as well as the Jobot California Worker Privacy Notice and Jobot Notice Regarding Automated Employment Decision Tools which are available at jobot.com/legal.

By applying for this job, you agree to receive calls, AI-generated calls, text messages, or emails from Jobot, and/or its agents and contracted partners. Frequency varies for text messages. Message and data rates may apply. Carriers are not liable for delayed or undelivered messages. You can reply STOP to cancel and HELP for help. You can access our privacy policy here: jobot.com/privacy-policy
Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 91113390
  • Position Id: 612269340
  • Posted 9 hours ago
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

San Francisco, California

Today

Full-time

USD 175,000.00 - 250,000.00 per year

San Francisco, California

Today

Full-time

USD 160,000.00 - 230,000.00 per year

San Francisco, California

Today

Full-time

USD 160,000.00 - 230,000.00 per year

San Francisco, California

Today

Full-time

USD 250,000.00 - 300,000.00 per year

Search all similar jobs