About this role.
Nebius is hiring a Senior Technical Project Manager for its Token Factory team to lead complex cross-functional projects in AI cloud infrastructure. The role involves coordinating new region launches, engineering quality initiatives, and the delivery pipeline for AI models. The ideal candidate has strong project management skills, technical background, and ability to manage dependencies and risks in a fast-paced environment. Experience with GPU infrastructure and AI/ML platforms is a plus. The position offers competitive compensation and opportunities to work on impactful AI projects.
Role DNA
A quick view of the complexity, pace, ownership and collaboration implied by the job description.
Job Complexity
4/5Pace & Pressure
4/5Autonomy Level
5/5Communication Load
5/5Salary analysis
Estimated compensation compared with the broader US market for similar roles.
Core skills
Skills and capabilities most closely associated with this opportunity.
Cover letter sample
Dear Hiring Team,
I am writing to express my strong interest in the Senior Technical Project Manager - Token Factory position at Nebius. With extensive experience in managing complex cross-functional projects in cloud infrastructure and AI platforms, I am confident in my ability to coordinate the delivery of new region launches and AI model pipelines. My background in risk management, dependency tracking, and technical communication aligns perfectly with your requirements. I am excited about the opportunity to contribute to Nebius's mission of leading AI cloud infrastructure. Thank you for your consideration.
Sample interview questions
I led a data center expansion project involving engineering, operations, and procurement teams. I created a detailed Gantt chart, identified critical path items, held weekly sync meetings, and used a risk register to proactively address blockers. By maintaining clear communication and ownership, we delivered on schedule.
I prioritize by impact and urgency, using a structured framework. I communicate changes to stakeholders and adjust plans accordingly. For example, during a model launch, when a dependency shifted, I reallocated resources and updated the timeline transparently, ensuring no critical path was missed.
In my previous role, I coordinated the deployment of GPU clusters for training large language models. I worked with infrastructure teams to ensure hardware readiness, managed capacity planning, and facilitated collaboration between ML engineers and SRE to optimize inference performance.
I organize decision forums with clear agendas, invite relevant stakeholders, and document trade-offs. For instance, when choosing between two networking architectures, I led a discussion comparing cost, latency, and scalability, and we reached consensus on the best approach aligned with project goals.
I noticed our release cycles were unpredictable due to manual testing. I introduced CI/CD pipelines and automated regression tests, reducing release time by 40%. I also implemented retrospectives to continuously refine the process, leading to more reliable and faster deliveries.
About Nebius:
Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure.
Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI.
Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D.
Role
As a Technical Project Manager in Token Factory, your primary focus will be coordinating complex cross-functional projects, including new region launches, engineering quality and platform-wide improvement projects, and the end-to-end delivery pipeline for new AI models. You will bring together multiple engineering teams, understand the critical path, proactively manage dependencies and risks, and ensure that ambitious technical goals are delivered predictably, on time, and without unnecessary operational overhead.
Your responsibilities will include:
Leading infrastructure and capacity delivery projects, including new region deployments, capacity expansion, or maintenance framework initiatives.
Coordinating model onboarding and production delivery, from infrastructure readiness to successful customer availability.
Building and maintaining execution plans, identifying risks early, and ensuring blockers are resolved before they impact delivery.
Working closely with engineering managers, technical leads, product managers, SRE, infrastructure, networking, security, and other platform teams.
Facilitating technical decision-making and ensuring ownership and accountability across complex cross-team projects.
Continuously improving engineering delivery processes to make execution more predictable while maintaining a sustainable pace for engineering teams.
Requirements
Excellent project management and delivery skills with the ability to break down ambiguous initiatives into executable plans
Ability to identify critical paths, manage complex dependency graphs, and coordinate multiple parallel workstreams
Strong risk management and prioritisation skills, with the ability to make progress in fast-changing environments
Excellent written and verbal communication skills in English
Comfortable leading incident coordination, facilitating discussions, documenting decisions, and driving follow-up actions
Strong technical background that allows you to understand engineering discussions, infrastructure dependencies, and architectural trade-offs without being the primary implementer
It will be an added bonus if you have
Experience working with GPU infrastructure, AI/ML platforms, or large-scale inference systems
Experience delivering cloud infrastructure or data center deployment projects
Familiarity with capacity planning, production operations, and reliability engineering
Previous experience in high-growth infrastructure or platform engineering organisations where priorities change quickly and execution speed matters
Benefits & Perks:
- Competitive compensation
- Career growth and learning opportunities
- Flexibility and ownership
- Collaborative and innovative culture
- Opportunity to work on impactful AI projects
- International environment and talented teams
What’s it like to work at Nebius:
Fast moving – Bold thinking – Constant growth – Meaningful impact – Trust and real ownership – Opportunity to shape the future of AI
Equal Opportunity Statement:
Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law.
Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire.
If you need accommodations during the application process, please let us know.
Annual salary information is not provided for this position. Explore salary ranges for similar roles in our Salary Directory ›
This job listing has been manually reviewed by the Jobicy Trust & Safety Team for compliance with our posting guidelines, including verification of the company's legitimacy, accuracy of job details, clarity of remote work policy, and absence of misleading or fraudulent content.






