Careers

Build the GPU-native AI runtime with us.

Small team. Generational ambition. We hire operators who think in systems and ship GPU infrastructure every week.

How we work

Compounding teams. Calm intensity.

Weekly ship cadence

Every Friday is a release. Velocity is a feature.

Direct, written, async

Decisions live in docs. Meetings are rare and ruthless.

Hybrid — SF, London, Singapore

Three offices, two days in-person, one global summit a year.

Top-of-band cash + equity

We index salaries to the 90th percentile of Levels.fyi for your role.

Forward-deployed mindset

Everyone — engineering, design, GTM — spends time with customers.

4 weeks PTO + sabbatical

Plus closed company weeks twice a year. Rest is a system.

Open roles

Where you fit.

Founding Engineer · GPU Inference

Engineering · San Francisco

Apply →

Staff Engineer · Triton/TensorRT-LLM

Engineering · Remote (US)

Apply →

Senior Designer · Product

Design · London

Apply →

Forward-Deployed Engineer

Engineering · San Francisco

Apply →

Applied AI Researcher · NeMo/NVIDIA Stack

Research · San Francisco / Remote

Apply →

Security Engineer

Security · Remote (EU)

Apply →

Enterprise Account Executive

GTM · New York

Apply →

Developer Advocate

DevRel · Remote

Apply →

Head of Marketing

Marketing · San Francisco

Apply →