Senior Staff Technical Program Manager
Job description
About the role
You will own the programs that directly enhance the KPI infrastructure and provide leadership with actionable insight into operational health across all Crusoe deployments. You will act as the connective tissue and the single source of truth for cross-functional alignment between DCOps, Cloud Engineering, and the TPM organization. In this capacity, you ensure quality hand-offs between teams with the authority to pull in necessary resources and prevent delays in compute capacity delivery. You do not merely react to breakdowns; you proactively identify friction points and fix the process as the fleet scales to meet demand. This role requires a technical background that allows you to understand physical data center constraints and translate them into operational rigor for software-defined teams. You will serve as a strategic advisor to leadership, determining whether teams are truly executing against the defined operational plan. The work is not limited to a single site or dashboard; it spans the entire infrastructure stack from electrons to tokens.
Key facts
What you'll do
- Run the programs that improve the KPI infrastructure, defining the metrics that tell the true story of deployment health before anyone has to ask.
- Support the Chief of Staff on strategic initiatives and cross-cutting projects that lack a natural single owner, ensuring continuity and execution.
- Partner with engineering, SRE, customer success, and data engineering to keep operational data accurate, consistent, and reliable across all reporting streams.
- Work toward a single pane of glass that makes the health of the organization and the status of deployments easy to understand at a glance for every stakeholder.
- Own the checks that guarantee a quality hand-off between teams, with the authority to pull in resources needed to deliver compute capacity to customers on schedule.
- Act as the single point of authority when issues surface during a deployment, replacing fragmented, ticket-driven communication with decisive coordination.
- Own the deployment process end to end, driving improvements as headcount and the number of sites grow over time.
- Translate messy, multi-source data into a clear picture of operational health, using facts to drive decisions rather than assumptions.
- Understand, or ramp rapidly on, the full data center build lifecycle from initial project planning through the physical-to-digital handoff.
- Build and evolve program tooling and processes yourself when they do not exist, prioritizing outcomes over ownership of credit.
Requirements
- 12+ years in software engineering, technical program management, or a technical product role, close enough to production systems to know what an SLO breach actually means.
- You have operated in high-growth infrastructure environments where processes are still being built, and ambiguity does not paralyze your decision-making.
- You are a natural coordinator who works across teams without formal authority, earning trust through consistent follow-through with DCOps, Network, and engineering leads.
- You can turn messy, multi-source data into a clear picture of organizational health, and you understand or can quickly learn the full data center build lifecycle.
- You are scrappy, low-ego, and high-drive, building the program and tooling yourself when necessary.
- You care more about the outcome than the credit, focusing on delivering reliable compute capacity to customers.
- You have experience working closely with physical deployments alongside software-defined infrastructure.
- You communicate with clarity and precision, aligning cross-functional teams around a shared definition of done.
Nice to have
- Time inside AWS, GCP, Azure, CoreWeave, Lambda Labs, or a similar cloud provider.
- Experience running or improving a data center deployment or build-to-turnup process.
- Familiarity with incident.io, Opsgenie, PagerDuty, or similar incident management platforms at scale.
- Background in AI/ML infrastructure.
Practical notes
This role is based in San Francisco, California, and is a full-time engagement. The position requires proximity to Crusoe operations in the United States and may involve travel between data center sites as needed. Candidates must be authorized to work in the United States without sponsorship at this time. There are no specific visa sponsorship options available for this role. The compensation package includes competitive base salary and equity awards, along with comprehensive benefits, paid time off, paid holidays, and leave of absence programs.