Associate Infrastructure Engineer
Job description
About the role
You will empower engineers on other teams to take control of their services by maintaining monitoring tooling and collaborating on internal best practices for observability. You will enhance reliability of applications running in Kubernetes by optimizing resource allocation, streamlining upgrade processes, and ensuring scalability and fault tolerance. Occasionally you will dive into the main Webflow application in Node, Python, or Go to better discern and sometimes fix behavior in production. You will work with peers on Webflow's Customer Support, Partnerships, and Sales teams to enable customers using Webflow's services in production. You will participate in and continuously improve on-call and incident response processes. You will contribute to infrastructure improvements that underpin projects launched each month across a global user base.
Key facts
What you'll do
Empower engineers on other teams to take control of their services by maintaining monitoring tooling and collaborating on internal best practices for observability. Enhance reliability of applications running in Kubernetes by optimizing resource allocation, streamlining upgrade processes, and ensuring scalability and fault tolerance. Occasionally dive into the main Webflow application in Node, Python, or Go to better discern and sometimes fix behavior in production. Work with peers on Webflow's Customer Support, Partnerships, and Sales teams to enable customers using Webflow's services in production. Participate in and continuously improve on-call and incident response processes. Analyze infrastructure metrics to identify trends and drive proactive improvements. Collaborate with cross-functional teams to define and implement reliability standards for customer-facing services. Evaluate and integrate new infrastructure tools that support scaling needs and operational efficiency. Translate complex system behavior into clear documentation and runbooks for broader team consumption. Support deployment pipelines and configuration management to reduce lead time for changes. Conduct postmortem analysis and translate findings into concrete remediation tasks. Assist in capacity planning and forecasting for upcoming product initiatives. Mentor junior engineers through code reviews and pair debugging sessions. Champion security and compliance practices across infrastructure components.
Requirements
You must hold a BA/BS degree or equivalent experience. You have a background as an ops engineer with growing interest in code, or a software engineer background with growing interest in systems and infrastructure. You have 2+ years of experience operating or debugging distributed systems in a production environment. You are comfortable working in at least one major cloud provider (AWS or GCP) and want to go deeper. You have hands-on experience with Kubernetes workloads, and ideally exposure to GitOps tools like ArgoCD or Argo Workflows. You are comfortable with scripting and automation using languages common in infrastructure contexts. You understand networking fundamentals and observability concepts including metrics, logs, and traces. You have experience with containerization technologies and container orchestration platforms. You are able to work asynchronously and communicate clearly across distributed teams. You are comfortable making decisions with incomplete information and prioritizing work in a fast-paced environment.
Nice to have
Experience with Webflow-like SaaS products or headless content platforms. Familiarity with modern frontend build pipelines and hosting concepts. Contribution to open source projects related to infrastructure and developer tools. Background in marketing technology or content management systems. Experience with cost optimization and performance tuning at scale.
Practical notes
This role is full-time and permanent. Remote-first with the option to work from California, British Columbia, or Ontario; U.S. Remote locations are also eligible. Employment is exempt. Applications are accepted on an ongoing basis until the position is closed and filled.