Boost your chances at reddit
Tailor your resume to this exact job and generate a matching cover letter in about 60 seconds with JobEase — our AI application assistant.
- ATS-optimized resume
- Personalized cover letter
- Match score & keyword gaps
Free to start · no card required
Get more other jobs in your inbox
Verified daily — no ghost listings.
About This RoleAI processing…
Reddit has a flexible workforce! If you happen to live close to one of our physical office locations our doors are open for you to come into the office as often as you'd like. Don't live near one of our offices? No worries: You can apply to work remotely in any country in which we have a physical presence.
Key Responsibilities
- 1Lead & Grow: Hire, mentor, and retain a high-performing team of ML engineers / systems-oriented engineers working on model optimization and ML efficiency.
- 2Set Technical Direction: Define the roadmap for training optimization, inference optimization, launch-readiness tooling, and reusable efficiency primitives across Ads ML.
- 3Deliver Measurable Wins: Drive reductions in model training time, online latency, serving cost, and infra-driven launch risk.
- 4Build Systems and Tooling: Guide the development of profiling, benchmarking, load testing, observability, cost analysis, debugging, and efficiency certification systems.
- 5Operate in the Critical Path: Partner with model owners and platform teams to accelerate high-priority launches and remove bottlenecks from the path to production.
- 6Shape the Team’s Evolution: Balance near-term white-glove optimization work with medium-term platformization and automation.
- 7Build XFN Alignment: Work closely with MLP, AMP, Ranking, and serving teams to clarify boundaries, upstream generic wins, and keep Ads needs on track.
- 8Raise the Bar: Establish engineering rigor around measurement, performance debugging, launch safety, and technical decision-making for efficiency work.
Requirements
- Deep ML Engineering Experience: The candidate should have been close to the models themselves and understand training, serving, debugging, and optimization in depth.
- Hands-on Optimization Background: Direct experience improving training loops, serving systems, profiling workflows, model/inference efficiency, or GPU utilization.
- Strong Managerial Ability: Experience building and leading teams, coaching engineers, managing delivery, and making prioritization tradeoffs under ambiguity.
- Distributed Systems Fluency: Proven ability to reason about production-scale ML systems and the tradeoffs that govern reliability, speed, cost, and scale.
- Customer and Platform Instincts: Able to work as a service provider to modeling teams while still building reusable systems rather than only heroic one-offs.
- Strong Communication: Can explain technical tradeoffs clearly to engineers, PMs, and senior stakeholders.
- Ads experience: Experience in ads ranking, recommender systems, marketplace ML, or adjacent production ML domains is strongly preferred.
- Experience with GPU training and serving migrations.
- Experience with PyTorch, distributed training frameworks, or kernel/performance optimization.
- Experience building efficiency benchmarking or launch certification frameworks.
- Experience working in organizations where ML platform and applied modeling responsibilities are split across multiple teams.
Perks & Benefits
Apply to This Job in Minutes
Generate ATS-optimized resume + cover letter + interview prep with Jobease.ca AI. Complete your application faster.
75% of AI Resumes Get Rejected
Beat the ATS with Jobease.ca's AI Resume Builder. Optimized for real hiring systems.
Build My ResumeProfile Match
Loading…Checking your profile against this job…
Job Overview
Share This Job
Track All Your Applications
Never lose track again. Jobease.ca organizes every application, interview, and follow-up.
Organize My Search