TixelJobs
I
Inferactvia Ashby

Member of Technical Staff, TPU Performance Engineering

San Francisco$200K - $400K/yrPosted 3mo ago
OtherStaff+Full-time

About the Role

Inferact's mission is to grow vLLM as the world's AI inference engine and accelerate AI progress by making inference cheaper and faster. Founded by the creators and core maintainers of vLLM, we sit at the intersection of models and hardware, a position that took years to build. About the Role We're looking for a TPU performance engineer to make vLLM a first-class inference engine on Google TPUs. You'll build and optimize TPU backends, compiler integrations, runtime paths, and benchmarking infrastructure using JAX,…

Limited-time offerMembers only

Unlock the apply link on every AI job

Browsing is free. Membership is $9 a year or $4.90 every three months and unlocks the apply link and the full description on every job. Cancel any time from your billing page.

  • The apply link on every job, straight to the official posting
  • Full job descriptions instead of the preview
  • Every AI job, every day: thousands of listings from 1,000+ company career pages
  • Save jobs and get daily alerts for your categories

Secure checkout by Dodo Payments·Cancel anytime from your billing page

Share