Semi-senior DevOps Engineer
Job description
About the role
This position is responsible for designing, implementing, and maintaining cloud infrastructure on AWS for a leading SaaS company. The role focuses heavily on Infrastructure as Code, CI/CD automation, and platform reliability to ensure the modern SaaS platform remains scalable and efficient. You will collaborate with multiple engineering teams to identify bottlenecks and enhance both scalability and developer productivity on an ongoing basis. The position requires a proactive mindset to troubleshoot complex infrastructure issues before they impact end users. You will play a key role in automating operational tasks to reduce manual effort and improve system resilience. This job demands strong ownership of the entire cloud lifecycle from initial architecture to decommissioning. Effective communication is essential to align technical solutions with business goals and to document processes clearly for the team.
Key facts
What you'll do
Design and build robust AWS infrastructure components including compute, storage, and networking resources using a code-first approach.
Implement and manage Infrastructure as Code templates using Terraform and CloudFormation to ensure environments are reproducible and version controlled.
Create and optimize CI/CD pipelines with GitHub Actions, Jenkins, GitLab CI, and Bitbucket Pipelines to enable rapid and reliable software deployments.
Automate the provisioning and scaling of serverless services such as AWS Lambda and API Gateway to meet fluctuating demand without overprovisioning.
Establish monitoring and observability frameworks using Prometheus, Grafana, and Datadog to track system health and performance metrics in real time.
Manage and secure identity and access through IAM policies and roles, ensuring least privilege principles are applied across all services.
Maintain high availability architectures by configuring Route 53 for DNS routing and VPC networking components to isolate and protect critical workloads.
Optimize storage solutions on S3 and relational databases on RDS by implementing lifecycle policies, backups, and performance tuning strategies.
Collaborate with development teams to integrate security best practices such as secrets management and data encryption into infrastructure workflows.
Support incident response efforts during on-call rotations by investigating alerts and implementing temporary and permanent fixes to restore service stability.
Requirements
Candidates must possess between 2 and 4 years of professional experience in DevOps, Cloud Engineering, or a closely related field.
Hands on experience with AWS services is mandatory, including ECS, Lambda, S3, RDS, IAM, Route 53, and VPC configurations.
You must have demonstrated ability working with Infrastructure as Code tools, specifically Terraform and CloudFormation, to manage cloud resources.
A strong understanding of cloud infrastructure automation and established DevOps best practices is required to ensure consistency across environments.
Experience building and maintaining CI/CD pipelines using GitHub Actions, Jenkins, GitLab CI, or Bitbucket Pipelines is essential for success in this role.
Previous work with serverless architectures and event driven systems is necessary to design efficient and cost effective solutions.
Hands on experience with monitoring and observability tools such as Prometheus, Grafana, or Datadog is required to maintain visibility into production systems.
Programming or scripting experience in Python, Bash, or Ruby is required to automate tasks and interact with cloud APIs effectively.
Knowledge of cloud security best practices, including IAM, secrets management, and encryption, is required to protect data and infrastructure.
Advanced English communication skills are required to collaborate with global teams and document technical decisions clearly.
Nice to have
Experience with CloudFront and API Gateway is preferred to optimize content delivery and API management strategies.
Familiarity with AWS cost optimization techniques and tools is preferred to help manage and reduce overall cloud spend.
Experience working within Agile methodologies and participating in sprint planning or retrospectives is preferred to improve team workflows.
Practical notes
This is a 100% remote work position, allowing you to perform duties from any location with a reliable internet connection.
Paid time off is provided to ensure you have adequate rest and recovery outside of work hours.
You will be on-call for 2 weeks out of every 10-week rotation, responding to alerts and addressing critical issues as they arise.