Skills
About the Role
Join Stitch Fix as an ML Platform Engineer and help build the infrastructure that powers machine learning and AI across the organization. You’ll create reliable, scalable platform capabilities that enable model training, deployment, feature workflows, and AI-driven experiences.
Responsibilities
- Design, develop, and maintain services and frameworks for ML training and production deployment
- Build and support feature engineering and feature serving pipelines
- Enable candidate generation and AI agent deployment capabilities
- Implement observability to monitor performance, reliability, and model/agent behavior
- Collaborate with data science and engineering teams to improve platform usability and scalability
Requirements
- Experience building or operating machine learning infrastructure at scale
- Strong software engineering skills with an emphasis on reliability and performance
- Knowledge of ML workflows such as training, deployment, and serving
- Familiarity with monitoring/observability practices for production systems
Benefits
- Opportunity to work on personalized retail and Generative AI at scale
- Collaborative team environment with investment in employee growth
- Remote work flexibility