Senior Platform Engineer
Job description
About the role
Omaze operates a platform that connects users with luxury prize draws while generating significant funding for major charities. As a Senior Platform Engineer, you will manage the cloud infrastructure and reliability standards that support our global operations. You will own the design and execution of resilient systems that ensure seamless user experiences during high traffic campaign launches. This role requires a proactive approach to infrastructure evolution and long-term platform scalability. You will partner closely with product and engineering teams to translate business requirements into robust technical solutions. Your work will directly influence the reliability and performance of critical user journeys across our digital touchpoints. You will champion best practices in automation and infrastructure governance across the technology organization.
Key facts
What you'll do
- Architect and evolve cloud infrastructure on AWS using Terraform or CDK, focusing on serverless patterns that meet current and future demand.
- Develop and maintain efficient CI/CD pipelines via GitHub Actions, ensuring rapid and reliable delivery of infrastructure and application changes.
- Implement observability frameworks, including alerting and incident response tooling, to provide clear insight into system behavior and failures.
- Manage security, compliance, IAM policies, and network boundaries to uphold a strong security posture and meet regulatory obligations.
- Monitor and optimize cloud resource utilization and infrastructure costs, balancing performance needs with financial efficiency.
- Improve developer workflows by reducing manual toil and creating reusable patterns that accelerate delivery and reduce complexity.
- Perform code reviews for infrastructure changes to ensure long-term maintainability, security, and alignment with architectural standards.
- Utilize AI-assisted tools to generate and validate infrastructure code, integrating these capabilities into existing development practices.
- Collaborate with cross-functional teams to define and drive improvements in deployment strategies and release management processes.
- Troubleshoot complex issues in production environments, coordinating with relevant teams to restore service and prevent recurrence.
- Contribute to the development of internal tooling that enhances operational efficiency and supports platform teams.
- Support the implementation of disaster recovery and business continuity strategies to safeguard critical services.
- Document infrastructure designs, operational procedures, and architectural decisions to ensure clarity and knowledge sharing.
- Mentor junior engineers on infrastructure best practices, cloud technologies, and effective problem-solving techniques.
Requirements
- Significant professional background in Platform, DevOps, or Infrastructure engineering, with a proven track record of delivering reliable systems.
- Deep expertise in AWS serverless architectures, specifically Lambda, DynamoDB, and API Gateway, including their configuration and optimization.
- Proven experience managing infrastructure as code using Terraform or AWS CDK, with a strong understanding of modular design patterns.
- Strong knowledge of CI/CD pipeline management and deployment reliability, including strategies for testing and progressive delivery.
- Experience with observability, monitoring, and participating in on-call rotations, including the analysis of metrics, logs, and traces.
- Understanding of cloud security, secrets management, and regulatory compliance, with an awareness of relevant frameworks and standards.
- Ability to communicate technical infrastructure goals to diverse stakeholders, translating complex concepts into clear and actionable information.
- Experience working in agile environments, collaborating effectively with product managers, developers, and operations teams.
- Commitment to maintaining high standards of code quality, documentation, and operational excellence in all infrastructure-related activities.
- Willingness to engage in continuous learning and adapt to new tools, technologies, and methodologies as the platform evolves.
Skills & tools
- AWS (Lambda, DynamoDB, API Gateway)
- Terraform
- AWS CDK
- GitHub Actions
- Infrastructure as Code (IaC)
- Observability and Incident Response tooling
Practical notes
Benefits include a stock options scheme, private medical and dental insurance, and a 9% employer pension contribution when you contribute 2%. You will also receive a personal learning and development budget, a home office equipment budget, and life assurance set at 4x your salary. Enhanced family leave policies are provided. This is a hybrid role based in Holborn, London, requiring 3 days per week in the office. The position is full-time and permanent.