Principal Engineer, Cloud Infrastructure
Job description
About the role
We are building the cloud infrastructure and delivery platform that one of healthcare's most ambitious technology transformations runs on, and we need a Principal Engineer, Cloud Infrastructure to own the infrastructure behind it all. You will both set the technical direction and do the work hands-on, while leading a small team of infrastructure engineers. This role is about infrastructure and developer enablement. As our AI strategy evolves, we see a world where many outside of traditional engineering and data roles will contribute to our technology ecosystem. We believe reliability, security, and cost discipline come from good platform design and automation, and that ticket queues and runbooks are signs we haven't solved the underlying problem. You will bring a clear point of view on what a modern, AI-native infrastructure org should look like, the judgment and technical skill to make it real, and the leadership instincts to build and develop the contributors around you.
What you'll do
- Own cloud infrastructure end-to-end across our cloud environments, staying hands-on with IaC, CI/CD, and observability alongside the team.
- Own the SDLC surface area (dev environments, lower envs, release mechanics, monitoring and response) so engineers can ship in minutes rather than days.
- Build the substrate that lets AI agents observe, reason, and act on business systems with safe defaults, and the rails that let AI-assisted apps from people outside of tech be deployed and supported without becoming operational liabilities.
- Participate in security operations alongside our SecOps team on CSPM, threat management, and patch management as the infrastructure-side owner of remediations.
- Frame infrastructure decisions in terms of business outcomes (operational efficiency, financial impact, clinical results), and make progress and tradeoffs visible to leadership as a natural byproduct of how you work.
- Lead and develop a small, high-impact team, and operate as a peer to our SecOps, Data, IT Systems, and App Engineering leaders as a unifying technical force across the org.
- Translate complex infrastructure constraints into clear guardrails and self-service tools that allow fast, safe experimentation by non-infrastructure teams.
- Partner deeply with product and data leaders to align long-term platform vision with short-term delivery needs while maintaining rigorous cost and risk discipline.
Requirements
- You have an engineering background with deep, hands-on experience operating production infrastructure across at least two major clouds (GCP, AWS, Azure) and with modern infrastructure-as-code (Terraform or similar).
- You have built and operated modern CI/CD systems, dev environments, and observability stacks. What you've shipped matters more than the size of the fleet you've managed.
- You have built and operated infrastructure across different company sizes and business contexts. Cross-domain breadth is a real asset; experience in regulated environments is a plus.
- You are comfortable in a text editor and terminal, and you use scripting and programming to solve infrastructure problems and automate away the repetitive.
- You understand how networks, identity, and data storage systems work in the cloud, and you care deeply about how they are secured and monitored.
- You have a track record of making sound technical decisions under uncertainty, balancing delivery speed with reliability, security, and cost.
- You are comfortable working asynchronously and communicating clearly in writing, with an appreciation for structured thinking and concise documentation.
- You are a team player who elevates the work of others, provides context for decisions, and seeks to leave a healthier system behind for the next person who inherits your work.
Practical notes
LOCATION: Remote (USA)
USA
This role operates as a fully remote position within the United States. There are no specified working hours, travel requirements, visa sponsorship details, or application deadlines provided in the source material. You are expected to align with the operational cadence of the team and be available for synchronous collaboration during standard overlapping business hours as determined by internal agreements. The engagement type is full-time, and compensation details are not specified in the source. All practical arrangements regarding location, communication, and execution are expected to conform to internal policies and guidelines already in place at Clover Health.