Senior Engineer, Platform Infrastructure
Job description
About the role
This role is responsible for building and operating the platform that engineering teams and customers rely on every day. The primary focus is the continuous improvement of deployment pipelines, automation frameworks, and overall system reliability. The hire will engage in critical thinking, collaborate closely with cross-functional partners, and drive ongoing refinement of both platform capabilities and engineering practices. Success in this position requires a mindset that treats infrastructure as a product with long term maintainability and clear user expectations. The work involves deep collaboration with software engineers to ensure platform changes enable them to deliver value safely and quickly. You will be expected to communicate technical tradeoffs in plain language to align stakeholders and simplify complex system concepts. The role is centered on ensuring the foundational systems are observable, secure, recoverable, and easy to evolve over time.
Key facts
What you'll do
Enhance Kubernetes platforms and surrounding infrastructure to improve performance, reliability, and security at scale.
Design and implement deployment automation that reduces operational risk and eliminates manual steps for routine operations.
Partner with other engineers to design, implement, and troubleshoot platform capabilities in a collaborative environment.
Write and maintain automated tests for infrastructure and deployment workflows to increase confidence and prevent regressions.
Continuously improve observability, reliability, security, and recoverability of the platform through iterative enhancements.
Document system designs and reasoning so that future engineers can understand intent without needing to know every command syntax.
Apply infrastructure as code practices using tools such as Ansible, Terraform, Helm, and Zarf to manage complex platform configurations.
Implement and manage containerized workloads, PKI and certificate management, and secrets management to support secure operations.
Contribute to trunk-based development processes supported by GitLab CI to enable frequent and low risk integrations.
Improve platform usability through clear interfaces, robust testing strategies, and thoughtful automation of repetitive tasks.
Support incident response and investigations to identify root causes and implement preventative measures for long term stability.
Champion simplicity and evidence based decision making to reduce operational complexity and maintenance burden.
Collaborate with product and engineering teams to ensure platform roadmaps align with customer needs and business objectives.
Mentor and enable other engineers by sharing knowledge, conducting code reviews, and promoting best practices across the organization.
Requirements
The posting states a bachelor's degree requirement. A degree is required as stated in the listed qualifications.
You build for change, prioritizing code and infrastructure that remain easy to modify in the future.
You prefer automation, treating repeated manual work as a signal for system improvement and process evolution.
You think in small batches, decomposing large problems into safe, incremental steps that can be validated quickly.
You work effectively as part of a team, clearly communicating reasoning, tradeoffs, and shared understanding with colleagues.
You improve the system by investigating root causes and implementing changes that prevent future issues from recurring.
You are comfortable with Kubernetes, Linux, networking fundamentals, trunk based development, GitLab CI, infrastructure as code, Ansible, Terraform, Helm, Zarf, containers, PKI and certificate management, secrets management, observability, infrastructure testing, and troubleshooting distributed systems.
You write clean, maintainable code and infrastructure definitions that are easy to test, review, and extend over time.
You understand how version control, code reviews, and continuous integration practices support high quality software delivery.
Practical notes
Employment in the United States is listed.
Typical interview steps
Hiring for engineering roles usually starts with a recruiter screen, followed by one or two technical rounds. Candidates often solve a coding problem, discuss past projects, and answer system design questions. Some loops include a take-home task. Final rounds typically cover team fit and give candidates a chance to ask questions. Interviewers look for how you break down an unfamiliar problem, not just whether you reach the answer. Practicing a few problems aloud and reviewing your own past projects are the best preparation.
Good to know
Platform infrastructure work centers on automation, reliability, and continuous delivery.
Extreme programming practices such as pair programming and small batch integration are the default workflow.
Infrastructure as code tools like Ansible, Terraform, and Helm are central to daily work.
Success is measured by frequent, safe deployments, faster recovery, and reduced operational complexity.
The culture favors simplicity, evidence, learning, and team outcomes over individual heroics.
Career growth
Engineering careers usually progress from individual contributor to senior, staff, and principal levels. Some engineers move into management and lead teams of five to twenty people. Others stay on the technical track. Growth follows demonstrated impact, not tenure alone. A typical engineering ladder has clear levels with defined expectations for scope, quality, and mentorship. Moving up usually requires owning outcomes end to end rather than completing assigned tickets.