Senior Site Reliability Engineer
Job description
About the role
The plays a critical role in ensuring the stability and performance of our digital platforms under demanding conditions. You are responsible for designing robust systems that maintain high availability while supporting rapid delivery cycles. This position requires you to invent and implement strategies that keep services online and responsive as user demand scales unpredictably. You will take ownership of key features within both Travel and HR platforms, directly influencing how these services behave at scale and interact with millions of users. Decisions in this role prioritize clear business outcomes over adherence to specific technologies or rigid methodologies. The work you do will shape the reliability foundation across multiple business lines, impacting how products are delivered and experienced. You will partner closely with cross-functional teams to translate complex requirements into resilient technical designs that balance speed with stability. Your contributions will be measured by reduced friction for professionals and higher satisfaction across the company's suite of offerings.
Scope of impact
You will influence how Swile serves more than 5.5 million users across France and Brazil, impacting fintech, travel, HR, and employee benefits products. Your work ensures that critical services remain performant and trustworthy for 85,000 companies relying on our platforms. Success is defined by improved system resilience, faster incident response, and smoother user journeys across all product lines. You contribute to building a technical culture where reliability is treated as a first-class product requirement. Your designs will help reduce operational risk and enable product teams to move quickly without sacrificing stability. You play a key role in aligning infrastructure decisions with long-term business strategy and regulatory expectations. Through your efforts, you help create digital experiences that feel seamless and reliable for professionals in everyday situations. Your influence extends beyond infrastructure into how engineering teams think about reliability, monitoring, and operational excellence.
Daily responsibilities
You design intake pipelines that capture detailed requirements for Travel and HR services, ensuring clarity before implementation begins. These pipelines translate stakeholder needs into technical specifications that guide development teams toward correct implementation. You establish build standards that enforce code reliability and quality before features are promoted to production environments. Each standard must be clear, enforceable, and aligned with company quality goals and operational expectations. You create review checklists that validate performance and security for every deployment, helping the team catch risks early. These checklists prevent delays by streamlining reviews without compromising on safety or compliance standards. You orchestrate ship procedures that minimize downtime while enabling the release of new functionality and improvements. This includes coordinating releases, defining rollback strategies, and verifying outcomes at every stage of the deployment pipeline.
You form partner collaboration rhythms that align engineering teams with business objectives through consistent communication. Regular interaction with product, design, and operations groups keeps everyone informed and reduces misunderstandings. You implement monitoring views that highlight anomalies across fintech and employee benefits offerings, providing clarity into system behavior. These views give engineers and leaders timely insight into performance, helping them respond faster to issues. You automate repetitive tasks to free engineers from manual work, accelerating iteration cycles and reducing human error. Automation allows the team to focus on complex problems that require creative thinking and deep technical expertise. You document architectural decisions so future teams can understand and extend current solutions without unnecessary rework. Clear documentation supports continuity, reduces knowledge silos, and helps new members ramp quickly into demanding responsibilities.
Required background
You hold at least 3 years of experience maintaining services used by large user groups, handling real traffic and diverse user load scenarios. This background should include working with cloud infrastructure and container orchestration platforms such as Kubernetes. You have a solid understanding of networking concepts and distributed system communication patterns, including service discovery and load balancing. You write scripts and small programs to validate infrastructure behavior and automate operational tasks as part of your daily work. Proficiency in at least one scripting language, such as Python, Bash, or similar, supports these activities and increases your effectiveness. You communicate clearly in French and English during discussions with global colleagues and stakeholders located in different regions. Strong written and verbal skills in both languages help ensure alignment across teams and reduce ambiguity in critical situations.
You demonstrate ownership and accountability for production incidents, participating actively in on-call rotations when required. Experience with observability tools, monitoring solutions, CI systems, and container platforms is expected as part of your core skill set. You understand the importance of security and compliance in platform design and operate in a way that respects data protection requirements. Familiarity with the travel industry and employee benefit platforms is valued, as it helps you ask better questions and propose practical solutions. You are comfortable working in fast-paced environments where priorities can shift quickly based on business needs. You enjoy collaborating with cross-functional partners and translating ambiguous requirements into reliable technical approaches. You are proactive in identifying risks, suggesting improvements, and driving initiatives that enhance overall system reliability over time.
Additional signals
Experience with travel industry systems and employee benefit platforms is valued and helps you integrate more quickly into existing workflows. Familiarity with these domains enables you to understand constraints, anticipate edge cases, and design solutions that work in real-world conditions. Knowledge of monitoring solutions, CI systems, container platforms, and scripting languages is part of the expected skill set for this role. You should be comfortable using these tools to build, observe, and maintain reliable services across distributed environments. Exposure to financial services and regulated environments is considered a strong signal due to the sensitivity and consistency requirements involved. Your background should reflect an ability to work independently while collaborating closely with teams located in multiple countries and time zones.
Practical information
Location: France
Engagement: Permanent
Compensation: 60000 to 70000 euros
We Candidates are encouraged to review that page for exact application steps and required documents. The information here summarizes expectations and day to day responsibilities.