← all roles

Who's hiring GPU, inference & ML-systems engineers right now

29 AI labs and infrastructure companies have 130 open GPU, CUDA, ML-systems, inference and performance engineering roles between them. Ranked by openings, refreshed daily from each company's own public job feed.

open roles
130
companies
29
list salary
73 · $139K–$850K
visa mention
24
remote
13

Observed across current open postings, refreshed daily — not a survey. Salary band is drawn only from roles that publish a range. Salary breakdown →

The useful question for someone in this field usually isn't “what jobs exist” — a search box answers that — but which companies are actually staffing up, and for what. That is what this page tracks. The shape of the demand is fairly consistent: dedicated AI-silicon companies (the Graphcores and Tenstorrents of the list) skew heavily toward performance, compiler and kernel work, because their whole product is making specific hardware fast. GPU-cloud and infrastructure providers (CoreWeave, Nebius and similar) hire for the systems and reliability side — operating accelerator fleets, not designing them. Frontier labs concentrate in inference and ML-systems, where the bottleneck is serving large models efficiently and training them across thousands of GPUs.

Two numbers worth reading off the table: how many of a company's roles disclose a salary band (a rough proxy for hiring maturity and US-comp transparency norms) and how many mention visa sponsorship or relocation — the single most decision-relevant fact for an engineer hiring across borders. Counts move as companies open and close postings; this is a snapshot of the current open set, not a survey.

CompanyOpenGPUML-sysInfPerfRemVisa$
Anthropic17228611714
Nebius1161337·6
Together AI111·91··6
Graphcore10···8·1·
Tenstorrent10···7··7
CoreWeave91·63··8
d-Matrix8··24···
SambaNova7·1251·7
Crusoe6··14··6
Databricks6··51··6
xAI61222··6
FluidStack32······
Lambda3·······
Prime Intellect31·1·12·
Baseten2··2····
Lightning AI211··112
Liquid AI21·1····
OpenAI2···1·1·
Scale AI2·11···1
Cognition1·1·····
Cohere11·····1
EnCharge AI1···11·1
Etched1···1·11
Lightmatter1·1····1
Modal1··1····
Perplexity1··1····
Poolside1··1·1··
SF Compute11····1·
World Labs11·11···

Browse by cut: Salaries · Remote · Europe · Senior · Staff · GPU & CUDA Engineering · ML Systems & Infrastructure · Inference & Model Serving · Performance & Kernel Engineering

Refreshed 2026-09-03 10:46 UTC