Senior DevOps Engineer
Job description
About the role
Join a team focused on developing and maintaining a scalable, reliable, and secure cloud contact center platform. This role involves enhancing our infrastructure and deployment processes to support global operations. You will contribute to the evolution of our CI/CD pipelines and automation strategies. The position requires deep ownership of infrastructure lifecycle management and a commitment to operational excellence. You will work closely with cross-functional engineering groups to align technology solutions with business objectives. This role is integral in driving reliability, performance, and security for customer-facing platforms. Expect to solve complex infrastructure challenges in a dynamic and fast-paced environment. Your work will directly impact the availability and scalability of critical communication services.
Key facts
What you'll do
- Design, implement, and manage infrastructure solutions within a cloud environment to support high-availability architectures.
- Automate deployment, scaling, and operational tasks for various services using modern infrastructure-as-code methodologies.
- Monitor system performance and troubleshoot complex issues across the platform to ensure seamless user experiences.
- Collaborate with development teams to integrate new features and services efficiently while maintaining operational stability.
- Ensure the security and compliance of our infrastructure and applications through proactive controls and best practices.
- Participate in on-call rotations to support critical systems and respond to incidents with urgency and precision.
- Evaluate, test, and recommend new tools and technologies that enhance infrastructure resilience and operational efficiency.
- Partner with site reliability engineers to develop observability strategies and improve monitoring frameworks.
- Contribute to the standardization of deployment patterns and architectural principles across engineering teams.
- Drive automation initiatives that reduce manual effort and improve consistency in environment management.
- Support the evolution of CI/CD pipelines to accelerate release cycles without compromising quality or security.
- Act as a technical leader in defining runbooks, operational playbooks, and disaster recovery procedures.
Requirements
- Bring at least 5 years of experience in DevOps or SRE roles with a strong track record of delivering reliable systems.
- Demonstrate proven experience with public cloud platforms, with a preference for deep knowledge in AWS services and ecosystems.
- Show a strong background in scripting and automation using languages such as Python or Bash to solve operational challenges.
- Have familiarity with containerization technologies including Docker and orchestration platforms like Kubernetes in production settings.
- Possess hands-on experience with CI/CD tools and practices to enable continuous software delivery and deployment.
- Show a solid understanding of networking concepts and security best practices to design robust infrastructure solutions.
- Exhibit the ability to work effectively in a hybrid model, balancing independent work with collaborative team efforts.
- Communicate clearly and professionally in English, both in writing and during technical discussions with stakeholders.
- Adhere to established processes while actively contributing to improvements in operational workflows.
- Maintain a strong sense of ownership for production systems and take responsibility for end-to-end outcomes.
Nice to have
- Bring experience with infrastructure as code tools such as Terraform or CloudFormation to accelerate environment provisioning.
- Demonstrate knowledge of monitoring and logging systems including Prometheus, Grafana, and the ELK stack for observability.
- Have prior work with microservices architectures and a clear understanding of distributed system challenges.
- Show familiarity with cloud security frameworks and compliance standards relevant to global operations.
- Demonstrate experience in optimizing cloud costs while maintaining performance and reliability targets.
- Show background in managing large-scale distributed deployments and handling complex operational scenarios.
Practical notes
This role operates under a full-time engagement model with a hybrid work arrangement based in Bengaluru, India. The position may require participation in on-call rotations that fall within standard working hours as defined by business needs. Some travel may be required for team meetings, training sessions, or project-related activities as determined by operational requirements. Candidates must be eligible to work in India without visa sponsorship for this specific engagement. Please ensure your application reflects direct experience with the technologies and practices outlined in the requirements.