Senior DevOps Engineer
Job description
About the role
You will design, develop, and evolve our AWS-based infrastructure using best-in-class tooling, leading the operation and optimisation of our container orchestration platforms with a focus on multi-cluster networking and service-to-service communication. You will take ownership of our cross-cluster Istio service mesh, implementing policies, improving traffic management, and enhancing mesh-level networking performance and security. You will also take ownership of the performance and scaling of our Confluent Kafka messaging deployments, proactively identifying bottlenecks, improving configurations, and collaborating with engineering teams to tune applications. You will monitor, optimise, and automate systems to ensure high availability, strong security posture, and superior performance using Datadog, Opsgenie, and AI-powered tooling. You will build scalable, robust solutions for complex technical challenges, ensuring infrastructure design aligns with best practices and long-term platform goals. You will mentor junior and mid-level engineers, promote engineering excellence, and contribute to team knowledge sharing and process improvements.
Key facts
What you'll do
- Design, develop, and evolve our AWS-based infrastructure using best-in-class tooling, including Terraform and ArgoCD, leading the operation and optimisation of our container orchestration platforms, with a focus on multi-cluster networking and service-to-service communication.
- Take ownership of our cross-cluster Istio service mesh - implementing policies, improving traffic management, and enhancing mesh-level networking performance and security.
- Take ownership of the performance and scaling of either our Confluent Kafka messaging deployments proactively identifying bottlenecks, improving configurations, and collaborating with engineering teams to tune applications.
- Monitor, optimise, and automate systems to ensure high availability, strong security posture, and superior performance using Datadog, Opsgenie, and AI-powered tooling.
- Build scalable, robust solutions for complex technical challenges, ensuring infrastructure design aligns with best practices and long-term platform goals.
- Mentor junior and mid-level engineers, promote engineering excellence, and contribute to team knowledge sharing and process improvements.
- Collaborate closely with product and platform teams to translate business requirements into resilient, secure, and scalable infrastructure solutions.
- Drive infrastructure roadmap initiatives, including cost optimisation, reliability improvements, and disaster recovery strategies across global environments.
Requirements
- 5+ years of hands-on experience building and managing AWS infrastructure (EC2, VPC, IAM, RDS, KMS, and related services).
- Strong expertise in Docker, Kubernetes, and Infrastructure as Code (Terraform/CloudFormation).
- 3+ years of hands-on experience building, managing, and optimizing either cluster-level Kafka deployments (Confluent preferred) or cloud relational/NoSQL database infrastructure in AWS, with a strong understanding of underlying engine mechanics and performance tuning.
- Advanced understanding of cloud networking, including routing, load balancing, VPC design, hybrid connectivity, and service-mesh-based networking - familiarity with multi-cluster service mesh (Istio) and related capabilities is preferred.
- Proficiency in CI/CD systems, with deep experience in GitLab CI/CD and working knowledge of ArgoCD.
- Proven experience deploying, maintaining, and scaling critical cloud infrastructure components in production environments.
- Strong experience with monitoring, logging, and alerting tools such as Datadog and Opsgenie, including creating meaningful dashboards and alert rules.
- Solid understanding of security best practices for cloud infrastructure, including identity and access management, encryption, and network segmentation.
Nice to have
- Experience with AI-powered tooling for operations, observability, or anomaly detection.
- Deep knowledge of financial services or regulated environments and their specific infrastructure requirements.
- Contributions to open source projects or active involvement in technical communities.
- Experience with multi-region and multi-account AWS architectures, including Control Tower or Landing Zone patterns.
- Familiarity with FinOps practices and cost governance in cloud environments.
- Background in building and supporting high-availability, low-latency trading or fintech platforms.
Practical notes
-
Engagement: Full-time.
-
Location: Vienna, Vienna, Austria.
- We offer a comprehensive relocation package to support candidates who are moving to join our team. Our goal is to make your transition as smooth as possible, and we provide assistance throughout the whole process.