Who's hiring GPU, inference & ML-systems engineers right now
29 AI labs and infrastructure companies have 130 open GPU, CUDA, ML-systems, inference and performance engineering roles between them. Ranked by openings, refreshed daily from each company's own public job feed.
- open roles
- 130
- companies
- 29
- list salary
- 73 · $139K–$850K
- visa mention
- 24
- remote
- 13
Observed across current open postings, refreshed daily — not a survey. Salary band is drawn only from roles that publish a range. Salary breakdown →
The useful question for someone in this field usually isn't “what jobs exist” — a search box answers that — but which companies are actually staffing up, and for what. That is what this page tracks. The shape of the demand is fairly consistent: dedicated AI-silicon companies (the Graphcores and Tenstorrents of the list) skew heavily toward performance, compiler and kernel work, because their whole product is making specific hardware fast. GPU-cloud and infrastructure providers (CoreWeave, Nebius and similar) hire for the systems and reliability side — operating accelerator fleets, not designing them. Frontier labs concentrate in inference and ML-systems, where the bottleneck is serving large models efficiently and training them across thousands of GPUs.
Two numbers worth reading off the table: how many of a company's roles disclose a salary band (a rough proxy for hiring maturity and US-comp transparency norms) and how many mention visa sponsorship or relocation — the single most decision-relevant fact for an engineer hiring across borders. Counts move as companies open and close postings; this is a snapshot of the current open set, not a survey.
| Company | Open | GPU | ML-sys | Inf | Perf | Rem | Visa | $ |
|---|---|---|---|---|---|---|---|---|
| Anthropic | 17 | 2 | 2 | 8 | 6 | 1 | 17 | 14 |
| Nebius | 11 | 6 | 1 | 3 | 3 | 7 | · | 6 |
| Together AI | 11 | 1 | · | 9 | 1 | · | · | 6 |
| Graphcore | 10 | · | · | · | 8 | · | 1 | · |
| Tenstorrent | 10 | · | · | · | 7 | · | · | 7 |
| CoreWeave | 9 | 1 | · | 6 | 3 | · | · | 8 |
| d-Matrix | 8 | · | · | 2 | 4 | · | · | · |
| SambaNova | 7 | · | 1 | 2 | 5 | 1 | · | 7 |
| Crusoe | 6 | · | · | 1 | 4 | · | · | 6 |
| Databricks | 6 | · | · | 5 | 1 | · | · | 6 |
| xAI | 6 | 1 | 2 | 2 | 2 | · | · | 6 |
| FluidStack | 3 | 2 | · | · | · | · | · | · |
| Lambda | 3 | · | · | · | · | · | · | · |
| Prime Intellect | 3 | 1 | · | 1 | · | 1 | 2 | · |
| Baseten | 2 | · | · | 2 | · | · | · | · |
| Lightning AI | 2 | 1 | 1 | · | · | 1 | 1 | 2 |
| Liquid AI | 2 | 1 | · | 1 | · | · | · | · |
| OpenAI | 2 | · | · | · | 1 | · | 1 | · |
| Scale AI | 2 | · | 1 | 1 | · | · | · | 1 |
| Cognition | 1 | · | 1 | · | · | · | · | · |
| Cohere | 1 | 1 | · | · | · | · | · | 1 |
| EnCharge AI | 1 | · | · | · | 1 | 1 | · | 1 |
| Etched | 1 | · | · | · | 1 | · | 1 | 1 |
| Lightmatter | 1 | · | 1 | · | · | · | · | 1 |
| Modal | 1 | · | · | 1 | · | · | · | · |
| Perplexity | 1 | · | · | 1 | · | · | · | · |
| Poolside | 1 | · | · | 1 | · | 1 | · | · |
| SF Compute | 1 | 1 | · | · | · | · | 1 | · |
| World Labs | 1 | 1 | · | 1 | 1 | · | · | · |
Browse by cut: Salaries · Remote · Europe · Senior · Staff · GPU & CUDA Engineering · ML Systems & Infrastructure · Inference & Model Serving · Performance & Kernel Engineering