← All jobs · Crusoe

Senior Technical Program Manager

Crusoe ·
66
AI-Agency
B62 U72
📍 Bellevue, US 🛠 AI tools welcome at work Senior 5–10+ yrs
PyTorchRayCUDAROCmGPU firmwareBMCBIOS
TL;DR

Senior Technical Program Manager at Crusoe managing AI infrastructure programs. Owns GPU cluster commissioning, IaaS feature delivery, and cross-functional execution for an energy-first AI cloud platform. Requires hands-on technical depth in compute infrastructure and 5–10 years at hyperscalers or neoclouds.

Apply at Crusoe →
share:
you'll be redirected to the company's career page

Job description

Crusoe is on a mission to accelerate the abundance of energy and intelligence. As the only vertically integrated AI infrastructure company built from the ground up, we own and operate each layer of the stack — from electrons to tokens — to power the world's most ambitious AI workloads. When you join Crusoe, you join a team that is building the future, faster.

We're in the midst of the greatest industrial revolution of our time. The demand for AI compute is boundless, and power is a bottleneck. We're solving that — with an energy-first approach that makes AI infrastructure better for the world and faster for the people innovating with AI.

We're looking for problem-solving, opportunity-finding teammates with a sense of urgency, who believe in the scale of our ambition and thrive on a path not fully paved — people who want to grow their careers alongside a team of experts across energy, manufacturing, data center construction, and cloud services.

If you want to do the most meaningful work of your career, help our customers and partners advance their AI strategies, and be part of a high-performing team that believes in each other, come build with us at Crusoe.

About the Role:

Crusoe's Cloud Product team is hiring a Senior Technical Program Manager to own and drive technical programs across our AI IaaS platform, aligning Product and Engineering on delivery. This role requires a hands-on individual contributor with genuine technical depth in AI infrastructure, comfortable operating in ambiguous, fast-moving environments and capable of building execution structure where little exists.

Our vision is the easiest-to-use AI purpose-built cloud. We offer IaaS products, letting AI/ML engineers focus on AI model frameworks (PyTorch, Ray) and compute stacks (CUDA, ROCm) while Crusoe manages the underlying complexity. Our platform standardizes firmware/OS bundles and automates component orchestration for consistent, scalable infrastructure. We also offer AI Managed Services, like SLA-bounded Managed Inference. The TPM connects engineering, product, procurement, and data center operations to deliver a reliable platform where customers run AI workloads without managing low-level system details.

At this level, TPM engagement focuses on defined technical programs and workstreams within larger cross-functional initiatives: GPU cluster commissioning, feature delivery within IaaS products, NPI workstream ownership, and coordination across hardware and software dependencies.

The ideal candidate is engineer-rooted with hands-on technical depth in compute infrastructure, firmware, or hardware platform delivery. You will own defined programs end-to-end, build lightweight execution frameworks, govern dependencies within your scope, and grow into broader program ownership over time.

What You'll Be Working On:

What You'll Bring to the Team:

Bonus Points

Benefits:

Compensation Range

Compensation will be paid in the range of up to $161,700 - $196,000 + Bonus. Restricted Stock Units are included in all offers. Compensation to be determined by the applicants knowledge, education, and abilities, as well as internal equity and alignment with market data.

Crusoe is an Equal Opportunity Employer. Employment decisions are made without regard to race, color, religion, disability, genetic information, pregnancy, citizenship, marital status, sex/gender, sexual preference/ orientation, gender identity, age, veteran status, national origin, or any other status protected by law or regulation.

Apply at Crusoe →

More open roles at Crusoe

Crusoe · 🔄 synced 3h ago
Staff Enterprise AI Automation Engineer
📍 San Francisco, US 💰 $190K–$230K 🛠 AI tools welcome at work · Staff
Staff Enterprise AI Automation Engineer at Crusoe designing and building agentic AI systems that orchestrate workflows across enterprise platforms. Focus on LLM integration, agent architecture, and scalable automation infrastructure.
PythonRESTGraphQLWorkatoAnthropic ClaudeGoogle Gemini
84
AI-core
Crusoe · 🔄 synced 3h ago
Principal Software Engineer, AI Model LifeCycle
📍 San Francisco, US 💰 $260K–$326K · Principal
Principal Software Engineer at Crusoe building managed platforms for LLM fine-tuning, training pipelines, and model lifecycle management. Focus on multi-node orchestration, reinforcement learning, and dataset/experiment versioning at scale.
PyTorchGolangPythonvLLMGPU systems
82
AI-core
Crusoe · 🔄 synced 3h ago
Senior Director of Engineering, Developer Experience
📍 San Francisco, US 💰 $301K–$355K 🛠 AI tools welcome at work · Director
Senior Director of Engineering for Developer Experience at Crusoe, an AI infrastructure company. Lead strategy and execution of internal developer platforms, CI/CD infrastructure, and AI-powered tooling to accelerate engineering velocity across the organization.
CICDDevOpsinternal APIsrepositoriesinfrastructure
76
AI-core
Crusoe · 🔄 synced 3h ago
Staff Product Manager, Managed Intelligence (SF/Sunnyvale)
📍 San Francisco, US 💰 $204K–$247K 🛠 AI tools welcome at work · Staff
Staff Product Manager at Crusoe leading product strategy for Managed Intelligence services. Focus on defining AI and agentic capabilities, model lifecycle, and scaling cloud products for AI-native companies.
PyTorchJAXTensorFlowKubernetesAWS
76
AI-core
Crusoe · 🔄 synced 3h ago
Senior Staff Software Engineer, AI Model LifeCycle
📍 San Francisco, US 💰 $237K–$318K · Staff
Senior Staff Software Engineer at Crusoe building managed platforms for AI model lifecycle, including fine-tuning systems, training pipelines, and reinforcement learning infrastructure for large language models.
PyTorchGolangPythonvLLMGPU systems
73
AI-fluent