Skills
About the Role
Join our team as an AI Researcher focused on model distillation. You’ll help us transform large, compute-heavy models into smaller, faster, and more deployable systems—without sacrificing quality.
This role is a great fit if you enjoy publishing research, working at the intersection of papers and production, and iterating quickly from ideas to working implementations.
Responsibilities
- Design and evaluate model distillation methods such as teacher–student training, self-distillation, layer-wise distillation, and representation matching
- Study and document tradeoffs across model size, latency, memory usage, and accuracy
- Create and test new distillation approaches for large language models and long-context or specialized architectures
- Assess distillation outcomes with rigorous experiments and clear reporting
Requirements
- Strong research experience in machine learning with a focus on distillation or compression techniques
- Proficiency with modern deep learning frameworks and experimentation workflows
- Experience evaluating models using practical metrics relevant to deployment (e.g., speed, memory, quality)
- Ability to communicate results through publications, technical write-ups, or conference-style documentation
Benefits
- Opportunity to work on frontier research with real-world deployment goals
- Collaborative environment where research ideas can move from paper to code to production
- Impactful work improving efficiency and performance of AI systems