Senior Platfrom Engineer
EverstarRemoteFull Time3w ago
Engineeringremotecurated-jd
Job description
Senior Platform Engineer at Everstar
About the role
Join a dedicated R&D team focused on enhancing proven products that have supported countless missions. This role offers the chance to contribute to advanced solutions, experiment with new methodologies, and witness your work scale across numerous applications.
Key facts
What you'll do
- Design and build a Kubernetes platform on GCP from its initial stages, including GKE clusters, multi-cluster setups, and foundational deployment templates.
- Develop self-service tools for development teams, creating templates and automated workflows that enable independent service deployment and operation.
- Manage infrastructure as code, ensuring reproducible environments and automating all infrastructure modifications.
- Establish network security standards, implementing default-deny policies, access controls, secrets management, and software supply chain security.
- Implement observability tools for metrics, logs, traces, SLO monitoring, and alerting.
- Maintain and optimize the infrastructure, including GKE updates, add-on management, and cost control.
- Lead incident response, identifying root causes and automating recovery processes.
- Collaborate with product teams to translate their technical needs into user-friendly platform solutions that improve developer experience.
Requirements
- Minimum of 5 years in infrastructure, DevOps, SRE, or Platform Engineering, with at least 3 years specifically working with Kubernetes in a production setting.
- Extensive experience managing GKE clusters throughout their lifecycle using Infrastructure as Code, covering deployment, configuration, and updates.
- Proven ability to build and manage multi-cluster or fleet infrastructure for diverse environments and teams.
- Practical experience creating self-service platform tools that have been adopted by internal development teams.
- Expertise in GKE networking, including network policies, ingress/gateway, services, DNS, load balancing, and Dataplane V2.
- Strong Linux system administration skills with a focus on troubleshooting compute, network, and storage issues.
- Professional proficiency with Terraform or OpenTofu for module design, remote state management, and multi-environment configurations.
- Experience building and maintaining CI/CD pipelines (GitHub Actions, GitLab CI) with automated deployments and rollbacks, using Helm and Kustomize.
- Hands-on experience with Google Cloud services beyond GKE, such as IAM, VPC, Cloud DNS, and GCS.
- Understanding of observability concepts and experience with Prometheus, Grafana, centralized logging, SLO monitoring, and alerting.
- Knowledge of Kubernetes security best practices, including secrets management, RBAC, network policies, image scanning, and supply chain security.
- Scripting proficiency in Bash, with programming experience in Python or Go.
- Experience planning and testing disaster recovery scenarios, including full cluster or environment rebuilds from code.
Skills & tools
- Kubernetes
- GKE
- GCP
- Terraform / OpenTofu
- GitHub Actions / GitLab CI
- Helm / Kustomize
- Prometheus / Grafana
- Bash
- Python / Go
- Linux
Practical notes
- Employment via Diia.City
- Competitive compensation with bonuses and premiums
- Support for military personnel (if Reserve+)
- Relocation assistance for candidates from other cities or countries
- Medical insurance after the probation period
- Psychological consultations
- 24 vacation days plus a birthday holiday
- Flexible working hours
- Up to 10 remote days per year
- Office located in Kyiv (right bank)
- Buddy and mentorship program
- Training and course compensation
- Access to a corporate library