Sr Staff Engineer
Uber is seeking a Sr Staff Engineer to focus on developing and deploying large-scale ML models, optimizing inference performance, and collaborating with research teams to bring cutting-edge AI advancements to production.
What you'd actually do
- Design, build, and maintain scalable and reliable ML infrastructure to support the training and serving of large-scale models.
- Optimize ML model inference performance, focusing on latency, throughput, and cost-efficiency.
- Collaborate with research scientists and other engineering teams to productionize novel ML algorithms and techniques.
- Develop and implement robust monitoring, evaluation, and alerting systems for ML models in production.
- Mentor junior engineers and contribute to the overall technical strategy of the ML platform team.
Skills
Required
- Experience with large-scale ML model deployment and serving
- Strong understanding of ML inference optimization techniques (quantization, pruning, etc.)
- Proficiency in Python and ML frameworks (e.g., TensorFlow, PyTorch)
- Experience with cloud platforms (AWS, GCP, or Azure)
- Excellent problem-solving and debugging skills
Nice to have
- Experience with distributed systems and big data technologies
- Familiarity with MLOps best practices
- Experience with GPU optimization
Other signals
- Develop and deploy large-scale ML models
- Optimize inference performance
- Collaborate with research teams