Careers
Build the GPU-native AI runtime with us.
Small team. Generational ambition. We hire operators who think in systems and ship GPU infrastructure every week.
How we work
Compounding teams. Calm intensity.
Weekly ship cadence
Every Friday is a release. Velocity is a feature.
Direct, written, async
Decisions live in docs. Meetings are rare and ruthless.
Hybrid — SF, London, Singapore
Three offices, two days in-person, one global summit a year.
Top-of-band cash + equity
We index salaries to the 90th percentile of Levels.fyi for your role.
Forward-deployed mindset
Everyone — engineering, design, GTM — spends time with customers.
4 weeks PTO + sabbatical
Plus closed company weeks twice a year. Rest is a system.
Open roles
Where you fit.
Founding Engineer · GPU Inference
Engineering · San Francisco
Staff Engineer · Triton/TensorRT-LLM
Engineering · Remote (US)
Senior Designer · Product
Design · London
Forward-Deployed Engineer
Engineering · San Francisco
Applied AI Researcher · NeMo/NVIDIA Stack
Research · San Francisco / Remote
Security Engineer
Security · Remote (EU)
Enterprise Account Executive
GTM · New York
Developer Advocate
DevRel · Remote
Head of Marketing
Marketing · San Francisco