Senior DevOps Engineer
Job description
About the role
You will own the reliability and performance of the core systems that power our distributed EV charging infrastructure across the United States. You will design and operate the software and infrastructure that turns a portfolio of real-world charging sites into a single intelligent and resilient network. You will work closely with product, hardware, and operations teams to ensure that the systems we build meet strict uptime and performance guarantees for our customers. You will implement the data pipelines and monitoring that provide insight into millions of telemetry events generated by chargers in the field. You will ensure that our infrastructure can scale reliably as we rapidly expand our footprint and deploy new capabilities. You will play a key role in bridging the gap between cloud-native software practices and the physical constraints of power systems and site operations. You will contribute to a culture of ownership, rigor, and continuous improvement across the engineering organization. You will help define the architectural direction of the platform while maintaining the highest standards of security, observability, and reliability.
Key facts
What you'll do
Design and maintain the core infrastructure that orchestrates power delivery and telemetry across thousands of EV charging points nationwide.
Implement robust data ingestion pipelines capable of processing millions of telemetry events per day from distributed charging hardware.
Build and operate monitoring, alerting, and observability systems that provide real-time insight into the health and performance of the charging network.
Collaborate closely with hardware and site operations teams to ensure software solutions align with real-world constraints and operational workflows.
Lead incident response and on-call responsibilities for critical production systems, driving rapid resolution and post-incident improvements.
Define and enforce infrastructure standards, deployment pipelines, and security best practices across cloud and on-premises environments.
Partner with product teams to translate business requirements into scalable technical solutions that support new features and fleet deployments.
Optimize the performance and reliability of our power management systems to meet strict uptime guarantees for autonomous and electric vehicle fleets.
Contribute to the development of tools that enable engineers to deploy, test, and operate services with high levels of automation and confidence.
Champion continuous improvement by identifying bottlenecks, evaluating new technologies, and driving adoption of best-in-class practices.
Support the expansion of our charging infrastructure by ensuring that our systems scale efficiently as we add new sites and increase capacity.
Work with cross-functional stakeholders to coordinate releases, manage dependencies, and ensure smooth rollouts of critical updates.
Provide technical leadership in the design of data models, APIs, and integration layers that connect software systems with physical infrastructure.
Mentor and guide other engineers by sharing knowledge, conducting code reviews, and promoting a culture of quality and reliability.
Requirements
You must have a Bachelor's degree in Computer Science, Engineering, or a related technical field or equivalent practical experience.
You must have extensive experience with infrastructure as code tools such as Terraform, CloudFormation, or equivalent platforms.
You must have deep expertise in container orchestration platforms, particularly Kubernetes, including production-grade cluster management.
You must have strong proficiency in at least one modern programming language such as Python, Go, or JavaScript for building scalable services.
You must have a proven track record of designing and operating distributed systems that handle high volumes of data and traffic.
You must have experience with cloud platforms such as AWS, Azure, or GCP, including networking, compute, and storage services.
You must have a strong understanding of databases, both relational and NoSQL, and experience with designing data models for scale.
You must have excellent problem-solving skills and the ability to debug complex issues in distributed environments.