Staff Engineer, DevOps
Job description
About the role
As a Staff Engineer in DevOps at Tekion, you will play a pivotal role in shaping and maintaining the scalable infrastructure that supports our cloud-native automotive platform. You will lead the design, development, and implementation of advanced automation solutions, including CI/CD pipelines and infrastructure-as-code, to ensure seamless deployment and high system availability. This position requires collaborating closely with cross-functional teams such as engineering, security, and product management to ensure the platform's scalability, security, and compliance. You will also be responsible for mentoring engineering teams, driving DevOps best practices, and evaluating emerging technologies to enhance platform reliability and velocity. The role is on-site at our Bangalore HQ and involves working with cutting-edge cloud technologies to support our innovative automotive retail ecosystem.
Key facts
What you'll do
- Architect and lead the development of sophisticated CI/CD pipelines, ensuring automated, reliable, and repeatable deployment processes across multiple environments.
- Design and implement cloud-native infrastructure solutions that support high-availability, fault-tolerance, and scalability for large-scale applications.
- Lead incident response efforts, perform performance tuning, and conduct capacity planning to optimize system performance and reliability.
- Collaborate with engineering, product, and security teams to ensure the platform adheres to security standards, regulatory compliance, and industry best practices.
- Develop and implement observability frameworks, including monitoring, logging, and alerting systems, to enable proactive incident detection and resolution.
- Mentor and coach engineers across teams, fostering a culture of DevOps excellence, and establishing standards, reusable frameworks, and internal tooling to streamline operations.
- Evaluate emerging technologies and trends in cloud computing, automation, and container orchestration, and drive their adoption to improve development velocity and system robustness.
- Support the deployment and management of container orchestration platforms such as Kubernetes, ensuring efficient resource utilization and system stability.
- Partner with security teams to implement security controls, manage vulnerabilities, and ensure compliance with industry standards, including ITAR where applicable.
- Lead efforts to automate infrastructure provisioning, configuration management, and deployment workflows, reducing manual intervention and human error.
- Collaborate with support teams to develop disaster recovery plans, backup strategies, and incident management procedures.
- Work closely with legal and compliance teams to ensure platform security and data privacy standards are maintained.
- Contribute to the continuous improvement of DevOps processes, tools, and documentation to enhance team productivity and system reliability.
- Stay updated on the latest industry trends, tools, and best practices, and incorporate them into the organization's DevOps strategy.
Requirements
- 8+ years of experience in DevOps, infrastructure automation, or platform engineering roles.
- Bachelor's or Master's degree in Computer Science, Engineering, or a related technical field.
- Extensive experience working with cloud platforms such as AWS, GCP, and Azure.
- Proven expertise in infrastructure-as-code tools like Terraform, Pulumi, or similar frameworks.
- Strong background in designing and implementing CI/CD pipelines using tools such as Jenkins, GitLab CI, or similar.
- Hands-on experience with container orchestration platforms, especially Kubernetes, including deployment, scaling, and management.
- Deep understanding of system observability, including logging, monitoring, and alerting tools such as Prometheus, Grafana, ELK stack, or equivalent.
- Knowledge of security best practices, compliance standards, and vulnerability management.
- Excellent problem-solving skills, with the ability to troubleshoot complex systems and incidents efficiently.
- Strong mentorship and leadership skills, with experience guiding engineering teams on DevOps practices.
- Effective communication skills, capable of working collaboratively across teams and conveying technical concepts clearly.
- Ability to work on-site at the Bangalore HQ and participate actively in team activities and meetings.
Nice to have
- Experience working with service mesh architectures like Istio or Linkerd.
- Familiarity with AI and machine learning deployment pipelines within cloud environments.
- Knowledge of legal and support considerations related to cloud deployments, including compliance with ITAR.
- Experience with large-scale enterprise platforms and complex microservices architectures.
- Understanding of security frameworks and standards applicable to automotive or enterprise cloud systems.
- Exposure to automation and orchestration tools beyond Kubernetes, such as OpenShift or Rancher.
Skills & tools
- Cloud platforms: AWS, Azure, GCP
- Container orchestration: Kubernetes
- Infrastructure-as-code: Terraform, Pulumi
- CI/CD tools: Jenkins, GitLab CI, or similar
- Monitoring and observability: Prometheus, Grafana, ELK stack, or equivalents
- Cloud-native infrastructure design and deployment
- Security and compliance frameworks
- Automation scripting and configuration management
- Incident response and troubleshooting tools
Practical notes
- This position is based at our Bangalore headquarters and requires on-site presence.
- Candidates should have a proven track record of designing and managing large-scale cloud infrastructure and automation solutions.
- The role involves working with advanced cloud-native technologies and supporting mission-critical automotive retail platforms.
- Tekion is an equal opportunity employer committed to diversity and inclusion. We encourage candidates from all backgrounds to apply.
- The successful candidate will be expected to collaborate effectively with cross-functional teams, mentor junior engineers, and stay current with industry best practices.
- This role offers an exciting opportunity to influence the future of automotive retail technology through innovative DevOps practices and infrastructure excellence.