SRE - CDI Paris - Theodo Cloud
Job description
About the role
Theodo Cloud represents a dedicated practice within Theodo Group focused on DevSecOps, specializing in the design, construction, and management of cloud infrastructures. The mission is to assist both scale-ups and large enterprises in resolving business challenges through technological innovation. As a pioneer in containerization, Theodo Cloud positions innovation at the core of its approach. The practice develops offers targeting current industry challenges, including accelerating cloud migrations and modernization using AI, building and managing regulated industry environments (such as HDS for healthcare and SecNumCloud), and unlocking the full potential of Large Language Models (LLMs) while optimizing performance, costs, and industrial security.
The SRE role is embedded within this structure. Theodo Cloud is a young entity where it is possible to build things and invest in personal motivations. Beyond the project, dedicated time is allocated for initiatives related to team training, diversity, or technical recruitment. The evolution roadmap leads first to a Lead position, then to specialized tracks in Management or Expertise. The methodology is based on Lean Tech®, which emphasizes continuous learning and places human development before product development.
This SRE position is intended for curious individuals who wish to rapidly increase their skills, working with a wide range of infrastructures and tools (AWS, GCP, Azure, Scaleway, Outscale, OVH, etc.) and complex bugs in production.
About the role
You are a Software Engineer who will design, implement, and operate cloud infrastructure as part of Theodo Cloud practice. You will work hand in hand with a Lead DevOps and an Agile Project Manager to ensure impactful delivery for clients in both build and run contexts. You will react to production incidents, run technical investigations, and own post-mortems with clear and structured communication. You will translate business needs into reliable infrastructure, leveraging cloud platforms and automation to remove bottlenecks and accelerate delivery. You will actively contribute to an environment where infrastructure is a growth lever rather than a constraint.
Key facts
-
Location: France
-
Engagement: CDI
What you'll do
- React to incidents in production and lead investigations to identify root causes and remediation plans.
- Design and decompose technical features to align with project constraints on quality, cost, and deadlines.
- Realize post-mortems and communicate findings clearly and transparently to both technical teams and clients.
- Build and maintain pipelines for CI/CD to streamline deployments and improve reliability.
- Set up monitoring, logging, and alerting solutions to ensure observability and rapid response.
- Advise clients on infrastructure strategies that improve quality, reliability, security, and performance.
- Migrate and modernize workloads onto Cloud and Kubernetes platforms with a focus on scalability and resilience.
- Participate actively in Agile rituals such as Scrum to ensure alignment with project goals.
- Optimize cloud costs and performance while maintaining strict security and compliance standards.
- Explore and prototype innovations using AI to accelerate infrastructure transformations and create business value.
- Maintain and evolve runbooks and technical documentation to support sustainable operations.
- Collaborate across teams to share knowledge and drive best practices in DevOps and SRE.
- Contribute to the continuous improvement of platforms as part of the run team's build activities.
- Conduct technical investigations and manage communications following established processes.
Requirements
- You hold a solid and formal training in Computer Science, Engineering, or a related field.
- You possess a strong foundation in computer science fundamentals such as algorithms, data structures, and distributed systems.
- You have hands-on experience with cloud platforms including AWS, GCP, Azure, Scaleway, Outscale, or OVH.
- You are comfortable working with container technologies such as Docker and orchestration tools like Kubernetes.
- You have a proven track record of operating production systems and responding to critical incidents.
- You understand infrastructure as code principles and have used tools such as Terraform, Ansible, or similar.
- You are fluent in at least one programming or scripting language for automation and tooling.
- You are able to communicate clearly in French and English, both verbally and in writing.
Nice to have
- Experience in regulated industries such as healthcare with knowledge of standards like HDS or SecNumCloud.
- Familiarity with Large Language Models and strategies to optimize their performance and cost in production.
- Background in performance optimization, cost management, and security best practices in cloud environments.
- Participation in open source projects or contributions to technical communities.
- Experience with advanced monitoring, logging, and alerting solutions such as Prometheus, Grafana, or ELK.
Practical notes
- This is a full-time position (CDI) based in Paris.
- The role may involve on-call duties on a voluntary basis.
- Travel is not required for this position.
- Candidates must be eligible to work in France without sponsorship.
- The application deadline is not specified; interested candidates are encouraged to apply promptly.