Senior Engineering Manager, Cloud Storage
Job description
About the role
You will lead the architecture, delivery, and operations of Crusoe's high performance cloud storage environment. You will manage a team with the dual mandate of operating our current petabyte scale infrastructure and building the next generation, bespoke cloud storage platform optimized for AI and HPC workloads. This role requires you to translate product requirements into actionable engineering plans and coordinate execution across partner teams. You will partner cross functionally with Compute, SDN, SRE, and CCX to ensure storage integrates cleanly into the broader Crusoe Cloud platform. You will serve as the primary technical representative for Storage in cross-functional initiatives and engage directly with hardware vendors and enterprise customers. Your leadership, strategy, and decisions directly impact revenue and the success of our customers' most critical AI pipelines and performance requirements.
Key facts
What you'll do
- Lead and grow a team of software engineers spanning all levels with a focus on career development, performance management, and building a high-performing engineering culture.
- Set clear team goals, run structured performance reviews, and create individualized growth paths that help engineers advance their careers over time.
- Foster a team environment where ownership is distributed, calculated risk-taking is encouraged, and engineers are empowered to solve hard problems without unnecessary constraints.
- Drive recruiting and hiring to grow the team; eventually owning the full hiring process for storage and related roles.
- Clearly translate product requirements into actionable engineering plans, from architecture and design through task breakdown, estimation, and delivery within defined timelines.
- Maintain a clear view of dependencies, resourcing, and sequencing across parallel workstreams to keep projects on track and aligned with business priorities.
- Coordinate across partner teams and external vendors to ensure storage integrates cleanly into the broader Crusoe Cloud platform and meets enterprise expectations.
- Lead architecture and design reviews, providing substantive technical input and holding the bar for quality across your team's work and deliverables.
- Maintain enough technical depth to conduct meaningful code reviews and engage credibly with your most senior engineers on complex design decisions for storage systems.
- Drive incident response for storage-related issues, ensuring fast resolution and durable fixes for business-critical systems that support AI and HPC workloads.
Requirements
- 4+ years of engineering management experience, with a demonstrated track record of growing teams and developing engineers, managing performance, and delivering cross-functional projects in a demanding environment.
- Experience leading teams of 6+ engineers at multiple levels, including senior individual contributors who look to you for technical as well as career guidance.
- A background as a strong individual contributor - software engineer, tech lead, or architect who moved into management and still engages deeply with technical work and code quality.
- 4+ years of hands-on experience with distributed storage systems such as Ceph, GlusterFS, or MinIO in production or large-scale environments.
- Hands on experience with production block, file, and object storage systems and the operational challenges of running them at scale.
- Experience with performant, scalable storage architectures that meet the needs of AI and HPC workloads and submillisecond latency requirements.
- A strong understanding of the full storage stack from NAND to GPU, focusing on performance characteristics, reliability, and data integrity in high-throughput scenarios.
- Comfort working in a fast-paced, rapidly scaling environment where priorities can shift based on energy, hardware, and infrastructure constraints.
Nice to have
- Experience building and operating cloud native platforms, including control planes and observability for storage services.
- Familiarity with AI and HPC access patterns and data workflows that stress storage throughput, latency, and consistency.
- Background working directly with hardware vendors and enterprise customers to address performance issues and influence next-generation hardware-software co-design.
- Experience with cost optimization and efficiency considerations for storage infrastructure in energy-constrained environments.
- Knowledge of open source storage projects and their integration into broader cloud infrastructure.
Practical notes
This is a full time role based in San Francisco, CA.
Output completed.