Senior Software Engineer, Site Reliability
Job description
About the role
This role centers on owning the reliability and operational excellence of Asana's systems from the ground up in Warsaw. You will be instrumental in shaping the practices, processes, and culture of a brand new SRE team from its earliest days. The position requires a strong software engineer who is passionate about building distributed systems that are both fast and highly resilient. You will collaborate daily with a small SRE team in San Francisco and infrastructure engineers in Reykjavik to define standards and execution. Your work will directly influence how the company manages reliability, incidents, and long-term infrastructure health across all products.
What you'll do
- Define and implement Asana's incident management process, establishing clear runbooks and communication norms for the entire company.
- Lead reliability-focused initiatives across the full stack, driving architectural improvements that reduce risk and increase system resilience.
- Build and maintain internal platforms and frameworks that empower other engineering teams to improve their service reliability.
- Own key reliability metrics and work closely with product and infrastructure teams to drive measurable improvements over time.
- Help shape and sustain a sustainable on-call rotation shared across teams in Warsaw, San Francisco, and Reykjavik, balancing coverage with quality of life.
- Act as a primary technical owner for critical infrastructure components, ensuring they meet stringent standards for performance and availability.
- Collaborate with infrastructure engineers in Warsaw to translate reliability requirements into robust technical designs and implementations.
- Work with our core technology stack, including AWS, Kubernetes (EKS), Datadog, MySQL (RDS), ElasticSearch (OpenSearch), Redis, DynamoDB, Terraform, TypeScript, Scala, Go, and Python.
- Partner with product and engineering teams to embed reliability practices early in the lifecycle of new features and services.
- Continuously explore and adopt new tools and techniques to improve observability, automation, and recovery workflows.
- Contribute to the broader engineering culture by documenting practices, mentoring peers, and promoting knowledge sharing across locations.
Requirements
- You are a strong and experienced software engineer who is comfortable writing and reading code, as this role is firmly rooted in engineering rather than pure operations.
- You care deeply about reliability, scalability, and long-term maintainability, prioritizing sustainable solutions over quick fixes.
- You have either worked as an SRE before or been a product engineer who regularly engaged with infrastructure challenges due to a strong sense of responsibility for system behavior.
- You have seen or actively want to work with systems at scale and enjoy solving complex infrastructure problems that affect many users.
- You are curious, take initiative, and are comfortable operating in ambiguous environments, which is critical for a founding team member in a new center.
- You collaborate effectively across distributed teams and are motivated by helping others build more reliable systems.
- You are eager to learn the full scope of our technology stack, even if you do not already know every tool, and you enjoy mastering new technologies quickly.
- You demonstrate curiosity about AI tools and emerging technologies, with a willingness to learn and leverage them to enhance productivity, collaboration, or decision-making.
Why this role
- Founding
team: You'll be one of the first SREs in Warsaw - and a key player in a growing team.
- Drive real change: This isn't a role where you'll just patch up legacy systems. We are expecting (and supporting) real architectural changes to improve reliability, scalability, and long-term operability.
- Big impact: Our product is scaling fast, and reliability is a top company priority.
- Room to grow: As this team grows, so will your influence - whether you want to lead projects, mentor others, or help shape how we scale.
- Global collaboration: Work closely with experienced engineers in San Francisco, Reykjavik, and Warsaw, while helping build the future of SRE at Asana.
What we offer
- Generous, tr
Requirements
- Strong and experienced software engineer, comfortable with reading and writing code.
- Deep care for reliability, scalability, and long-term maintainability.
- Experience as an SRE or product engineer who regularly handled infrastructure challenges.
- Proven ability to work with systems at scale and solve complex infrastructure problems.
- Curiosity, initiative, and comfort with ambiguity, especially in a founding team environment.
- Strong collaboration skills across distributed teams.
- Eagerness to learn the full technology stack and master new technologies quickly.
- Demonstrated curiosity about AI tools and emerging technologies for productivity and decision-making enhancement.
Practical notes
-
Location: Poland
-
Engagement: Contract of Employment (UoP) offered for employees in Poland
- Schedule: Office-centric hybrid with standard in-office days on Monday, Tuesday, and Thursday; option to work from home on Wednesdays; Friday work-from-home arrangements vary by work type; recruiter provides specific in-office requirements.