Senior Network Production Engineer, Operations
AI · 25 roles open in software and data
- Location: San Francisco, CA - US
- Pay: Pay not stated
- Experience: 5+ yrs
- Posted: Posted
About the role
The team supports Crusoe's global network infrastructure, including data centers and GPU clusters. The engineer responds to network incidents, performs root cause analysis, and develops automation to maintain the reliability of AI workloads. This role focuses on keeping large-scale AI infrastructure operational.
Our summary of the posting; the original is on the Crusoe careers page.
What the posting requires
- Required
- Python
- Observability
- Scripting
- PFC
- BGP
- Arista
- Leaf-spine
- Monitoring
- API
- Preferred
- NVIDIA
- Kentik
- Mellanox
- Arbor
- GPU cluster
- Traffic analysis
These tags are our reading of the posting. They can miss something; the original posting is the reference.
Do you qualify?
Each requirement checked against your resume, in about a minute. No account. Your resume is deleted when the verdict appears.
