
Engineering Jobs at Inference
Inference has 2 open engineering roles, 1 of them remote, alongside 3 other open roles.
About Inference
Inference.net provides AI infrastructure for AI-native teams to run, monitor, and optimize production AI workloads. The company also trains and hosts specialized language models for companies, offering an end-to-end platform that covers distillation, training, evaluation, and hosting. Its models are smaller and faster than GPT-5, with up to 90% lower cost.
The open roles
- Senior Software Engineer - Model Performance: San Francisco, USD 220k to 320k, Senior, posted 21 January 2026.
- Fullstack Engineer - Frontend Focus: remote, USD 120k to 180k, posted 23 July 2025.
Also hiring at Inference
Inference is also hiring in data and AI (2) and sales and business development (1).
Its roles are in San Francisco, CA, US (5).
What engineering work involves at an early-stage startup
You own whole parts of the product rather than one service inside a large system. Expect to write the code, deploy it, and get paged when it breaks. Reviews are quick and the codebase is small, so a decision made on Tuesday can ship on Thursday.
Experience level
Open roles by the experience the posting asks for.
Questions people ask
Can these roles be done remotely?
1 of the 2 open engineering roles at Inference are remote.
What else is Inference hiring for?
Data and AI and sales and business development: 3 more roles.