Software Engineer, Workers Deploy & Config
Job description
About the role
Join the team responsible for the control plane of Cloudflare's serverless edge platform. You will build and maintain the systems that allow developers to deploy and manage global applications, directly impacting services like R2 and Pages. In this capacity, you will own the design and execution of features that power the developer experience for one of the internet's largest infrastructure platforms. You will be responsible for ensuring that the deployment workflows are reliable, scalable, and secure for a global user base. The role requires deep collaboration with product and infrastructure teams to translate complex requirements into robust technical solutions. You will be expected to act as a systems thinker, considering the long-term implications of architectural decisions on performance and maintainability. This position is centered on creating the tools that enable thousands of engineers to ship code with confidence at the edge.
Key facts
What you'll do
- Architect and implement features for the Workers control plane, ensuring alignment with product roadmaps and technical standards.
- Enhance the scalability, availability, and performance of the platform through careful analysis of system bottlenecks and user behavior.
- Create storage schemas and I/O strategies to support long-term growth and data integrity across distributed environments.
- Participate in design reviews for foundational infrastructure, providing critical feedback on system resilience and operational excellence.
- Manage production reliability and join the on-call rotation to respond to incidents and ensure high uptime for critical services.
- Collaborate with product teams to turn ambiguous requirements into concrete technical specifications and implementation plans.
- Manage projects from initial concept through to final release, coordinating tasks and stakeholders to meet aggressive deadlines.
- Mentor junior team members and interns, fostering a culture of learning and knowledge sharing within the engineering organization.
- Evaluate and integrate new technologies to address technical challenges and keep the platform competitive in the edge computing space.
- Build and maintain RESTful APIs that serve as the primary interface for internal and external consumers of the platform.
- Implement observability solutions using Prometheus and Grafana to provide deep insights into system health and performance metrics.
- Work with SQL and relational databases like PostgreSQL to design data models that support complex edge computing workflows.
- Utilize Kubernetes or similar orchestration tools to manage the deployment and scaling of backend services efficiently.
- Ensure that all code changes are verified and reviewed, including code generated by AI assistants, to maintain the highest quality standards.
- Contribute to the evolution of the control plane by exploring GraphQL and RPC patterns to improve API flexibility and developer ergonomics.
Requirements
- Production experience with Go is mandatory for this role, as the core platform services are built using this language.
- Proficiency in JavaScript and TypeScript is required to build and maintain both backend services and frontend tooling.
- Experience with observability and metrics using Prometheus and Grafana is necessary to monitor system performance and reliability.
- A background in SQL and relational databases like PostgreSQL is essential for managing persistent data and complex queries.
- Experience with Kubernetes or similar orchestration tools is required to manage containerized applications at scale.
- A solid understanding of distributed systems principles is critical for designing resilient and performant architectures.
- Ability to manage projects from start to finish, including scoping, planning, and execution, is a hard requirement for success.
- Experience building and using RESTful APIs is necessary to interact with various internal and external services.
- Ability to verify and review code generated by AI assistants is required to ensure security, correctness, and maintainability.
- Eligibility to work in the specified locations (Austin, TX; New York City, NY; or Washington, DC) is required for this position.
- Full-time engagement is mandatory, and the role is not eligible for part-time or contract arrangements.
- Willingness to participate in on-call rotations is required to support production incidents as they arise.
- Final interview stages may require in-person attendance at a Cloudflare office, and candidates must be prepared to accommodate this requirement.
- Candidates must be authorized to access software or technology under U.S. export control laws to be considered for the role.
Nice to have
- Experience scaling systems for high performance in large-scale distributed environments.
- Background working on data planes or control planes in networking or cloud infrastructure.
- Experience as a user of Cloudflare Workers or Pages, providing firsthand insight into the developer experience.
- A product-focused mindset with experience in customer-facing communication and stakeholder management.
- Familiarity with GraphQL and RPC to help design flexible and efficient API layers.
Practical notes
- This role is eligible for the Cloudflare equity plan.
- Benefits include medical, dental, and vision insurance, 401(k) retirement savings, flexible paid time off, parental leave, and mental health support.
- Offers may be conditional based on authorization to access software or technology under U.S. export control laws.
- Final interview stages may require in-person attendance at a Cloudflare office.