As a Technical Project Manager in Token Factory, your primary focus will be coordinating complex cross-functional projects, including new region launches, engineering quality and platform-wide improvement projects, and the end-to-end delivery pipeline for new AI models. You will bring together multiple engineering teams, understand the critical path, proactively manage dependencies and risks, and ensure that ambitious technical goals are delivered predictably, on time, and without unnecessary operational overhead.
Your responsibilities will include:
Leading infrastructure and capacity delivery projects, including new region deployments, capacity expansion, or maintenance framework initiatives.
Coordinating model onboarding and production delivery, from infrastructure readiness to successful customer availability.
Building and maintaining execution plans, identifying risks early, and ensuring blockers are resolved before they impact delivery.
Working closely with engineering managers, technical leads, product managers, SRE, infrastructure, networking, security, and other platform teams.
Facilitating technical decision-making and ensuring ownership and accountability across complex cross-team projects.
Continuously improving engineering delivery processes to make execution more predictable while maintaining a sustainable pace for engineering teams.
Excellent project management and delivery skills with the ability to break down ambiguous initiatives into executable plans
Ability to identify critical paths, manage complex dependency graphs, and coordinate multiple parallel workstreams
Strong risk management and prioritisation skills, with the ability to make progress in fast-changing environments
Excellent written and verbal communication skills in English
Comfortable leading incident coordination, facilitating discussions, documenting decisions, and driving follow-up actions
Strong technical background that allows you to understand engineering discussions, infrastructure dependencies, and architectural trade-offs without being the primary implementent.
Experience working with GPU infrastructure, AI/ML platforms, or large-scale inference systems
Experience delivering cloud infrastructure or data center deployment projects
Familiarity with capacity planning, production operations, and reliability engineering
Previous experience in high-growth infrastructure or platform engineering organisations where priorities change quickly and execution speed matters
Benefits & Perks:
Competitive compensation
Career growth and learning opportunities
Flexibility and ownership
Collaborative and innovative culture
Opportunity to work on impactful AI projects
International environment and talented teams
What's it like to work at Nebius:
Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI
Published on: 7/30/2026

Nebius
The Nebius AI Cloud brings powerful full-stack infrastructure for AI developers and practitioners across startups, enterprises and science institutes to build and deploy generative AI applications and rapidly deliver scientific breakthroughs by training and running ML models within a secure, high-performance, and cost-optimized cloud environment.
Please let Nebius know you found this job on Wantapply.com. It helps us to get more jobs on our site. Thanks!