Machine Learning Performance Engineer, Training
Quant trading · 14 postings open
- Location: New York, NY
- Pay: Pay not stated
- Experience: 3+ yrs
- Posted: Posted
About the role
The team optimizes systems for training machine learning models at scale. The engineer analyzes and improves the entire training process, from data handling to hardware utilization, to speed up model development.
Our summary of the posting; the original is on the Tower Research Capital careers page.
What the posting requires
- Required
- Machine learning frameworks
- Python
- C++
- GPU kernel development
- GPU architecture
- Distributed training
- Performance analysis tools
- High-performance networking
- Benchmarking
- Preferred
- Kubernetes
- Large-scale checkpointing
- Experiment reproducibility
- Gpu cluster observability
- Compiler technologies
- Slurm
These tags are our reading of the posting. They can miss something; the original posting is the reference.
Do you qualify?
Each requirement checked against your resume, in about a minute. No account. Your resume is deleted when the verdict appears.
