Senior Software Engineer, Cloud Infrastructure
Job description
About the role
Oscar Health is a technology-driven health insurance company that builds digital platforms to improve the healthcare experience for its members. The Senior Software Engineer, Cloud Infrastructure role focuses on designing, building, and maintaining the cloud systems that power Oscar Health's digital products and services. This position is based in Tempe, Arizona and involves working closely with engineering teams across the organization to ensure reliable, scalable, and secure infrastructure. The role requires a senior-level engineer who can take ownership of complex cloud architecture decisions, drive technical strategy, mentor other team members through knowledge sharing, and play a key part in shaping the foundation that enables Oscar Health to deliver innovative healthcare solutions to its members.
Key facts
What you'll do
- Design and implement cloud infrastructure solutions that support Oscar Health's digital health platform and member-facing applications
- Collaborate with cross-functional engineering teams across the organization to define infrastructure requirements for new product features
- Monitor system performance and reliability metrics across cloud environments to identify optimization opportunities and potential bottlenecks
- Write infrastructure-as-code configurations using modern declarative approaches to ensure reproducible and version-controlled deployment environments
- Troubleshoot production issues related to networking, compute resources, storage systems, and service mesh components under pressure
- Conduct thorough code reviews and provide meaningful technical guidance to junior and mid-level engineers on the team
- Evaluate and recommend cloud services and tooling options to improve overall system resilience, scalability, and cost efficiency
- Participate in on-call rotations to ensure timely and effective response to infrastructure alerts and production incidents
- Contribute to architectural decision records that clearly document the rationale and trade-offs behind major infrastructure choices
- Automate deployment pipelines and infrastructure provisioning workflows to reduce manual intervention, errors, and time to production
- Partner with security and compliance teams to ensure infrastructure configurations meet regulatory and data protection standards
- Plan and execute capacity forecasting and resource allocation strategies to support growing user demand and traffic patterns
Requirements
- Demonstrated experience designing and operating cloud infrastructure in a production environment with high availability requirements
- Strong understanding of distributed systems concepts including networking protocols, storage architectures, and compute orchestration frameworks
- Proficiency with at least one major cloud provider and its core compute, networking, and storage services
- Hands-on experience writing infrastructure-as-code using declarative configuration management languages for automated provisioning and deployment
- Ability to diagnose and resolve complex system issues in high-availability and low-latency production environments
- Bachelor's degree in computer science, software engineering, or equivalent practical experience gained through professional work
- A minimum of five years of professional software engineering experience focused on infrastructure or platform engineering
- Comfort with on-call responsibilities and following structured incident response procedures during production system disruptions
- Experience working in fast-paced settings where prioritization and clear communication are essential for success
- Familiarity with continuous integration and continuous delivery practices and their application to infrastructure changes
Nice to have
- Familiarity with container orchestration platforms and service mesh technologies for managing distributed microservices at scale
- Experience with observability tooling including metrics collection, distributed tracing, and centralized logging for production systems
- Background in healthcare technology or working within regulated industries with strict compliance and security requirements
- Contributions to open-source infrastructure projects or active participation in community tooling and platform development efforts
- Knowledge of security best practices for cloud-native applications including network segmentation and access control policies
Skills & tools
- Cloud platform services including virtual machines, networking, load balancing, and managed storage offerings
- Infrastructure-as-code frameworks and configuration management systems for declarative provisioning and state management
- Containerization and orchestration technologies for deploying, scaling, and managing application workloads efficiently
- Monitoring and observability platforms for distributed systems including dashboards, alerting, and health checks
- Version control systems and collaborative development workflows using pull request-based code review processes
- Scripting and automation languages for operational tasks, infrastructure management, and repetitive process elimination
- Networking concepts including DNS, load balancing, service discovery, and traffic routing configurations
Practical notes
- This role is based in Tempe, Arizona and requires on-site presence during standard business hours each week
- The engineering team at Oscar Health follows an agile development process with regular sprint planning and retrospective meetings
- Candidates should expect to participate in design reviews, architecture discussions, and technical brainstorming sessions with peer engineers
- Oscar Health offers competitive benefits and a collaborative work environment that values diversity, inclusion, and continuous learning
- The infrastructure team works closely with product and engineering leadership to align technical roadmaps with business objectives