Engineering Manager, ML
Engineering Manager · Manager · Full Time
Opens Cursor's application page
Role
What you'll do.
Lead a high-impact ML infrastructure team at Cursor, where you'll own the training, testing, and evaluation systems that power the world's leading AI-assisted coding platform. This role sits at the intersection of infrastructure and model behavior, requiring both strong technical leadership and hands-on systems expertise to build scalable ML training pipelines, robust evaluation frameworks, and optimized environments for model iteration at scale.
Responsibilities
- ML Infrastructure Leadership & Technical Direction: Set and execute technical strategy for model training, testing, and evaluation infrastructure at scale. Make critical architectural decisions that balance latency, quality, and cost tradeoffs. Own the codebase alongside your team, conducting technical reviews and debugging complex distributed systems issues that span both infrastructure and model behavior boundaries.
- Reinforcement Learning & Evaluation Infrastructure: Design and implement rollout infrastructure enabling researchers to conduct large-scale RL experiments efficiently. Build evaluation pipelines that catch regressions before production deployment and provide rapid, trustworthy signal on experimental changes. Create reproducible, sandboxed training and testing environments optimized for iteration speed.
- Team Leadership & Engineering Development: Source, interview, and hire exceptional infrastructure engineers aligned with Cursor's mission and values. Develop engineers through mentorship, code review, coaching, and strategic project assignments. Build a small, talent-dense team capable of operating with high autonomy in ambiguous environments while maintaining rigorous engineering standards.
- Cross-Functional Collaboration & Communication: Partner closely with research teams to translate model-level requirements into concrete infrastructure solutions. Communicate fluently across both research and engineering domains, identifying technical ownership boundaries and ensuring clarity on whether issues stem from systems design or model behavior. Collaborate with product and research to prioritize infrastructure investments.
- Measurement & Rigor: Establish rigorous measurement frameworks for infrastructure quality and team progress, especially in areas where shipping code doesn't guarantee impact. Instrument systems for observability, implement monitoring for reliability and performance under production load. Drive data-driven decision-making about infrastructure investments and prioritization.
Qualifications
What we look for.
Technical
ML Infrastructure & Training Frameworks
Production experience building or leading infrastructure for model training, evaluation, or serving systems. Deep knowledge of distributed training frameworks, experiment tracking, model checkpoint management, and optimization for large-scale ML workloads.
Distributed Systems & Reliability Engineering
Strong fundamentals in distributed systems, including consistency, fault tolerance, and performance optimization. Proven ability to design systems that perform reliably under real production load, not just in theory. Experience with containerization, orchestration platforms, or microservices architecture.
Software Engineering Excellence
Demonstrated ability to write clean, maintainable production code and conduct thorough technical code reviews. Comfortable with infrastructure-as-code principles, CI/CD pipelines, and modern development workflows. Proficiency in at least one systems-level programming language (e.g., Python, Go, Rust, C++).
Debugging Complex Systems
Strong debugging skills for production systems spanning multiple layers of abstraction. Ability to trace issues from application code through infrastructure to identify root causes. Experience with observability tools, profiling, and performance analysis.
Education
Computer Science or Related Field
Bachelor's degree in Computer Science, Computer Engineering, Mathematics, or equivalent professional experience demonstrating advanced technical depth in systems design and software engineering.
Experience
ML Infrastructure Leadership
Led engineering teams building infrastructure for training, evaluating, or serving machine learning models in production environments. Track record of shipping infrastructure that unblocked research or production teams and improved iteration velocity.
Team Building & Development
Proven track record hiring and developing infrastructure engineers who have grown into senior roles. Experience mentoring engineers through both technical skill development and career growth. Demonstrated ability to build high-performing, autonomous teams.
Production Systems at Scale
Hands-on experience owning or building production systems handling significant scale or complexity. Deep understanding of reliability, performance characteristics, and operational challenges in production environments.
Cross-Functional Technical Communication
Experience translating between research and engineering domains, explaining technical tradeoffs and architectural decisions to non-infrastructure specialists. Ability to communicate with researchers fluently about model behavior and systems behavior.
Skills
Required
ML Model Training Infrastructure
Production experience with distributed training systems, experiment management, model checkpointing, and training pipeline orchestration for large-scale machine learning workloads.
Distributed Systems Design
Core expertise in designing and operating distributed systems including data consistency, fault tolerance, load balancing, and performance optimization under real-world constraints.
Infrastructure-as-Code & DevOps
Proficiency with containerization (Docker), orchestration platforms, CI/CD systems, and infrastructure automation. Experience managing deployment pipelines and production infrastructure.
Technical Leadership
Ability to set technical direction, make architectural decisions, conduct thorough code reviews, and stay hands-on with implementation alongside the team.
Team Leadership & Hiring
Experience sourcing, interviewing, and hiring high-quality engineers. Track record of developing engineers through mentorship and strategic project assignments.
Preferred
Reinforcement Learning Infrastructure
Nice to haveHands-on experience building or maintaining infrastructure specifically for reinforcement learning training, including rollout systems, reward computation, and experiment orchestration at scale.
Evaluation & Testing Frameworks
Nice to haveExperience designing evaluation pipelines for machine learning models, building regression detection systems, or creating comprehensive testing infrastructure for model behavior validation.
Simulated Environments & Sandboxing
Nice to haveExperience building and maintaining sandboxed or simulated environments for model training or testing, including reproducibility, isolation, and performance optimization.
AI/Coding Tools Ecosystem
Nice to haveFamiliarity with AI-assisted coding tools, language models, or code completion systems. Experience integrating with or building on top of modern AI development tools.
Observability & Monitoring
Nice to haveDeep expertise in instrumentation, logging, tracing, and monitoring systems. Experience building observability platforms that provide actionable insights into system behavior and performance.
Cost Optimization for ML Infrastructure
Nice to haveExperience optimizing infrastructure costs for machine learning systems, including compute resource allocation, efficient data movement, and cost-aware architectural decisions.
Tech stack
Languages
Frameworks
Databases
Tools
Other
Compensation
Pay and benefits.
Base·USD 185,000 – 280,000
Equity·Stock options
Benefits
Equity & Ownership
Meaningful equity stake in Cursor as an early-stage company with significant growth potential in the AI developer tools space. Direct impact on the company's success.
Cutting-Edge Technology & Impact
Lead infrastructure powering the world's leading AI-assisted coding platform. Work at the intersection of distributed systems and machine learning on problems with significant real-world impact.
Small, Talent-Dense Team
Work in a flat organizational structure with exceptionally talented engineers and researchers. High autonomy, rapid decision-making, and direct impact on company direction.
Technical Leadership Opportunity
Rare opportunity to remain deeply technical while leading infrastructure strategy. Set architectural direction and stay hands-on with code and debugging alongside your team.
Mentorship & Growth
Develop and mentor exceptional infrastructure engineers. Build a team culture focused on learning, shipping, and creative problem-solving.
Process
Interview steps.
- 01
Initial Screening Conversation
Preliminary call with Cursor's recruiting team to discuss your background in ML infrastructure and team leadership experience. Focus on your technical depth and experience scaling machine learning systems.
- 02
Technical Deep Dive
Conversation with current infrastructure team members or engineering leadership. Expect discussion of past projects, architectural decisions, debugging experiences, and how you've handled complex distributed systems challenges.
- 03
Leadership & Vision Interview
Discussion with Cursor's leadership about your approach to team building, technical direction-setting, and cross-functional collaboration. Emphasis on your vision for ML infrastructure at scale and how you'd approach ambiguity.
- 04
Research & Engineering Collaboration
Conversation with members of Cursor's research team to assess your ability to translate between research and engineering domains, understand model training dynamics, and make infrastructure tradeoff decisions.
- 05
Executive Alignment
Final conversation with senior leadership to discuss long-term vision, team scaling strategy, and alignment on Cursor's mission to automate coding through cutting-edge AI and infrastructure.
Full posting
Original listing.
Our mission is to automate coding. The first step in our journey is to build the best tool for professional programmers, using a combination of inventive research, design, and engineering. Our organization is very flat, and our team is small and talent dense. We particularly like people who are truth-seeking, passionate, and creative. We enjoy spirited debate, crazy ideas, and shipping code.
About the Role
You will lead a team of engineers building the infrastructure used to train, test, and evaluate our models. This is one of the few places at Cursor where infrastructure and model behavior meet directly: when something breaks, it's rarely obvious whether it's a systems bug or the model doing exactly what it was trained to do, and your team has to be good at telling the difference before they can fix it.
You'll set technical direction for how we train and evaluate models at scale, stay close enough to the code to debug alongside your team, and work daily with researchers to turn tradeoffs in latency, quality, and cost into infrastructure that actually gets built. We're hiring across a range of scope for this role, depending on experience and the size of problem you're ready to own.
Example projects include..
Building the rollout infrastructure that lets researchers run RL experiments at scale without fighting the plumbing.
Designing eval pipelines that catch regressions before they ship, and give researchers fast, trustworthy signal on whether a change actually helped.
Owning the environments in which models are trained and tested: sandboxed, reproducible, and fast enough that iteration speed isn't the bottleneck.
Bringing rigor to how the team measures quality and progress, in places where "did it ship" isn't the same as "did it work?"
Partnering with research to translate model-level tradeoffs (latency, quality, cost) into concrete infrastructure decisions.
Hiring and growing the team: sourcing, interviewing, and closing exceptional infrastructure engineers, while developing your engineers through coaching, mentorship, and high-leverage project assignments.
You may be a fit if
You've led engineering teams building infrastructure that trains, evaluates, or serves ML models in production.
You have strong infrastructure and distributed systems fundamentals: you know what reliability and performance look like under real load, not just in a design doc.
You genuinely want to stay technical: you're comfortable writing code, reviewing PRs with depth, and using tools like Cursor itself to move fast.
You’re comfortable operating in ambiguity: you ask the right questions, make sound decisions with incomplete information, and help the team find a path forward.
You have a track record of hiring and developing engineers who are better than you were at their stage.
You can talk fluently with researchers about model behavior and with engineers about systems design, and you know when a problem is actually the other team's.
Bonus: hands-on experience with RL training infrastructure, eval frameworks, or building and maintaining simulated environments for model training or testing.
Redirects to Cursor's application page.
Other roles
More at Cursor.
Software Engineer, ML Platform
Senior
Software Engineer, RL Data
Mid
Software Engineer, Pretraining
Senior
Field Engineer, Life Sciences
Mid
Field Engineer - India
Mid