Software Engineer, Inference
AI · 11 roles open
- Location: San Francisco, CA
- Pay: $230k–$390k a year
- Posted: Posted
About the role
The Inference team builds and maintains the systems that enable AI agents to operate quickly and reliably. The engineer on this team helps design the inference architecture, working on systems for serving, routing, and optimizing model performance. This role involves managing both self-hosted and third-party inference solutions.
Our summary of the posting; the original is on the Sierra careers page.
What the posting requires
- Required
- Distributed systems
- Production systems
- Infrastructure
- Large-scale systems
- Production operation
- Systems thinking
- Latency
- AI infrastructure
- Learning
- Reliability
- Preferred
- vLLM
- SGLang
- Inference frameworks
- MLOps
- LLMs
- GPU infrastructure
These tags are our reading of the posting. They can miss something; the original posting is the reference.
Do you qualify?
Each requirement checked against your resume, in about a minute. No account. Your resume is deleted when the verdict appears.
