Platform DR & Capacity Engineer
Job description
About the role
Anaplan is seeking a dedicated Platform Disaster Recovery and Capacity Engineer to enhance our crisis management solutions and ensure business continuity. This role involves collaborating with various teams to implement effective disaster recovery strategies and optimize resource allocation to meet the demands of our platform. The successful candidate will own the design and execution of critical initiatives that safeguard our services against disruptive events. You will be responsible for ensuring that our infrastructure remains resilient and performant under varying load conditions. This position requires a proactive mindset to identify potential weaknesses before they escalate into significant issues. You will partner closely with cross-functional stakeholders to align technical solutions with business continuity objectives. Your work will directly contribute to the reliability and stability of the Anaplan platform, which serves clients globally.
Key facts
What you'll do
- Collaborate with the Engineering and Service Management teams to develop and implement disaster recovery strategies for both Cloud and Data Center environments.
- Create tools that track the progress and effectiveness of disaster recovery plans against established metrics.
- Ensure that disaster recovery solutions are properly maintained and tested as part of ongoing operations.
- Provide insights and recommendations for risk management and mitigation strategies.
- Regularly update stakeholders on disaster recovery initiatives and activities.
- Design and implement frameworks and tools for capacity planning, policies, and strategies.
- Assess capacity requirements and impacts for new services or modifications.
- Address capacity discrepancies and initiate improvement projects.
- Work with other platform managers to achieve objectives outlined in our platform evolution roadmap.
- Monitor system performance to identify trends that may impact future capacity needs.
- Coordinate with vendors and internal teams to resolve issues related to infrastructure constraints.
- Implement best practices to optimize the efficiency of our disaster recovery and capacity management processes.
- Conduct regular reviews of operational procedures to ensure alignment with industry standards.
- Support the development of runbooks and documentation for disaster recovery scenarios.
Requirements
- A minimum of 5 years of experience in IT operations or Production Engineering.
- Strong background in disaster recovery planning and execution.
- Proficiency in capacity planning and resource management.
- Experience in monitoring performance metrics and managing risks associated with capacity.
- Familiarity with tools and technologies relevant to disaster recovery and capacity management.
- Excellent communication and collaboration skills.
- Ability to work effectively in a fast-paced environment with competing priorities.
- Strong analytical and problem-solving skills to address complex infrastructure challenges.
- Willingness to adhere to strict organizational policies and procedures.
- Capability to manage multiple tasks and projects simultaneously without compromising quality.
Nice to have
- Experience with cloud platforms and services.
- Knowledge of automation tools and scripting languages.
- Familiarity with performance monitoring tools and methodologies.
Practical notes
- This position is open to candidates requiring work visas.
- Anaplan offers a competitive salary and benefits package, including health insurance and retirement plans.
- Interested candidates should submit their applications by the specified deadline.
- The role is based in Gurugram, India, and is available for full-time employment.
- Candidates must be eligible to work in India without requiring sponsorship for a long-term visa.
- The work environment is professional, requiring adherence to standard corporate policies and operational hours.
- Travel may be required occasionally for business purposes, although this is not a primary expectation for this role.
- No specific visa sponsorship details are provided beyond the initial eligibility to work in the country of operation.
- Applications must be submitted through the official channels outlined by the hiring team.
- Deadlines for application submission are strict and must be met to ensure consideration.
- The selection process may include technical assessments and interviews to evaluate suitability for the role.
- Successful candidates will be required to complete any necessary onboarding procedures as dictated by company policy.
- This position is part of the Engineering & Service Management team, emphasizing cross-functional collaboration.
- The role involves a significant level of responsibility and ownership over critical business continuity processes.
- The candidate must be comfortable working with distributed teams and remote communication tools.
- Regular participation in meetings and status updates is expected to ensure alignment with organizational goals.
- The position requires a commitment to continuous improvement and learning in the areas of disaster recovery and capacity engineering.
- The successful applicant will need to demonstrate a history of reliability and attention to detail.
- The role is essential to maintaining the integrity and performance of Anaplan's platform infrastructure.
- All activities related to this position must comply with company policies and industry regulations.