Senior Manager Site Reliability Engineer
Job description
Senior Manager Site Reliability Engineer at Careers.
About the role
You will lead Careers's SRE Center of Excellence in India as a senior member of the overall engineering leadership team. You will support Careers's growth of its engineering footprint in India and develop a high-performing, cross-time-zone, globally unified SRE team. The objective of the SRE Center of Excellence is to establish a durable, scalable, secure, and cost-efficient cloud operations capability. This role supports Careers's ongoing progression from monolithic architectures to modular, multi-service platforms, enabling both product agility and operational performance at scale. You will champion a platform-as-a-product outlook, promoting engineering team independence while maintaining reliability, resilience, cost efficiency, and solid operational standards. Leading a unit of Infrastructure, Automation, Cloud Operations, and SRE engineers, your leadership will directly influence the daily experience of Careers engineering teams and the digital journeys of millions of Careers customers, entrepreneurs, small business owners, and merchants worldwide. This is a rare opportunity to build and expand a global SRE capability from India, collaborating with dedicated engineers across regions to address complex distributed systems challenges.
Key facts
What you'll do
Establish and evolve a Center of Excellence for Site Reliability Engineering in India, defining standards, practices, and measurable outcomes for reliability and operational excellence. Partner with product and infrastructure teams to translate business requirements into scalable, resilient, and cost-optimized platform solutions. Design and implement automation frameworks that streamline operations, reduce manual toil, and improve system observability, monitoring, and incident response capabilities. Own the architecture and governance of shared platform services, ensuring alignment with security, compliance, and performance objectives across global teams. Lead capacity planning, cost management, and infrastructure optimization initiatives to support efficient and sustainable growth of engineering workloads. Facilitate cross-time-zone collaboration, creating clear operating models, ownership structures, and escalation paths that enable distributed teams to maintain high service quality. Drive adoption of platform thinking by building product-like interfaces, self-service tooling, and clear ownership models that empower development teams. Mentor and develop SRE, automation, and cloud engineering talent, fostering a culture of continuous learning, experimentation, and data-driven decision-making. Partner with finance, security, and business stakeholders to align technology operations with strategic objectives, budget constraints, and regulatory requirements. Represent Careers's SRE function in cross-company forums, contributing to industry best practices and shaping the future of cloud operations.
Requirements
You must possess a Bachelor's degree in Computer Science, Engineering, or a related technical field, or equivalent practical experience. Bring 8+ years of experience in software engineering, systems engineering, or platform roles, with a strong track record in reliability and operations. Demonstrate 8+ years of hands-on experience with infrastructure, automation, and cloud operations in large-scale, distributed environments. Show deep expertise in Site Reliability Engineering practices, including monitoring, alerting, incident management, capacity planning, and performance optimization. Have proven ability to design and manage cloud-native architectures on major public cloud platforms, ensuring scalability, resilience, and security. Exhibit strong leadership experience managing and developing technical teams, with a history of enabling engineers to deliver high-quality outcomes. Display excellent judgment in balancing trade-offs between speed, reliability, cost, and operational risk in fast-paced environments. Communicate effectively in English, both written and verbal, to influence and collaborate with stakeholders across geographies and levels of the organization.
Nice to have
Experience building and operating platform engineering organizations and SRE Centers of Excellence. Familiarity with DevOps, automation frameworks, and infrastructure-as-code practices at scale. Knowledge of observability platforms, logging, metrics, and tracing standards. Understanding of security and compliance frameworks relevant to global cloud operations. Experience mentoring senior engineers and leading cross-functional, globally distributed teams.
Practical notes
This is a remote position, so you will be working remotely from your home. You may occasionally visit a Careers office to meet with your team for events or meetings.