Skip to content

Software Engineer, Inference

Sierra

AI · 11 roles open

  • Location: San Francisco, CA
  • Pay: $230k–$390k a year
  • Posted: Posted
Check your fit

About the role

The Inference team builds and maintains the systems that enable AI agents to operate quickly and reliably. The engineer on this team helps design the inference architecture, working on systems for serving, routing, and optimizing model performance. This role involves managing both self-hosted and third-party inference solutions.

Our summary of the posting; the original is on the Sierra careers page.

What the posting requires

Required
  • Distributed systems
  • Production systems
  • Infrastructure
  • Large-scale systems
  • Production operation
  • Systems thinking
  • Latency
  • AI infrastructure
  • Learning
  • Reliability
Preferred
  • vLLM
  • SGLang
  • Inference frameworks
  • MLOps
  • LLMs
  • GPU infrastructure

These tags are our reading of the posting. They can miss something; the original posting is the reference.

Do you qualify?

Each requirement checked against your resume, in about a minute. No account. Your resume is deleted when the verdict appears.

Check your fit