Senior Site Reliability Engineer
Job description
About the role
DualEntry is creating the foundational operating system for finance, automating accounting processes with AI-native ERP software. This role contributes to building and maintaining the core infrastructure that powers financial systems for businesses. You will enhance system reliability and security within a fast-paced engineering environment. The position demands a proactive mindset focused on long-term system integrity and performance. You will work closely with engineering teams to ensure that financial operations remain seamless and secure. This is an opportunity to shape the infrastructure of a rapidly scaling fintech company from the ground up. Your work will directly influence how businesses manage their most critical financial workflows.
Key facts
What you'll do
- Improve the stability and performance of our systems through rigorous analysis and optimization.
- Manage and expand our primary infrastructure using Terraform to define and provision cloud resources.
- Implement and refine our continuous integration and deployment pipelines to accelerate release cycles safely.
- Develop monitoring and insights pipelines following OpenTelemetry standards to capture system telemetry.
- Strengthen system security and enhance the developer experience by automating guardrails and access controls.
- Collaborate with cross-functional teams to translate business requirements into scalable technical solutions.
- Troubleshoot complex production incidents and lead post-mortem analyses to prevent future occurrences.
- Optimize cloud resource utilization to reduce costs while maintaining high availability and performance.
- Mentor junior engineers on best practices for infrastructure management and observability.
- Partner with product and infrastructure teams to plan capacity and roadmap initiatives for future growth.
Requirements
- Demonstrate a strong work ethic and proactive approach to problem-solving in a dynamic environment.
- Willingness to participate in an on-call rotation to support critical production systems 24 hours a day.
- Possess at least 5 years of experience focused on system scalability and improvements in large-scale environments.
- Bring at least 3 years of hands-on experience with Terraform for infrastructure as code and version control.
- Accumulate at least 5 years of experience working with Amazon Web Services (AWS) services and architecture patterns.
- Maintain a solid understanding of backend fundamentals in Python to write and review application code.
- Show experience with observability tools, specifically OpenTelemetry, for tracing, metrics, and logging.
- Exhibit a deep commitment to security and compliance standards relevant to financial services technology.
- Communicate effectively with both technical and non-technical stakeholders during high-pressure situations.
Nice to have
- Participation in a project that supported 1 million users or a similar scale of high-availability system.
- Practical experience in backend development and system design principles applied in production.
- An active GitHub profile showcasing contributions to open source or personal infrastructure projects.
- Familiarity with financial services technology and regulatory considerations in the fintech space.
- Experience with container orchestration platforms such as Kubernetes in cloud environments.
- Knowledge of database systems and query optimization for transactional workloads.
- Background in implementing CI/CD best practices for regulated industries.
Practical notes
- Visa and permanent residence sponsorship are available for candidates who qualify.
- Relocation support is provided for those moving to New York City to start the position.
- Benefits include medical, dental, vision, mental health support (Talkspace), fertility and family-building support (Kindbody), virtual healthcare (Teladoc Health), 401(k), commuter benefits, and a learning and development budget.
- Flexible PTO includes 27 days, encompassing public holidays to ensure adequate rest and personal time.
- The work environment is collaborative and in-office, with snacks, drinks, and regular team events to foster community.
- Candidates must be authorized to work in the United States without sponsorship requirements for the role specifics.
- The position requires consistent availability during standard business hours for collaboration with distributed teams.
- Relocation packages are subject to eligibility and specific terms outlined during the hiring process.
- This role is based in New York City and requires physical presence within the designated office location.
- The company maintains an equal opportunity employment policy and encourages diverse candidates to apply.
- Successful candidates will undergo standard background checks as required by finance industry regulations.
- The start date is contingent upon business needs and can be negotiated based on candidate circumstances.
- Employment terms are subject to the execution of standard offer letters and legal documentation.
- The equity component is part of the total compensation package and detailed during the offer stage.
- Continuous learning is encouraged through the allocated learning and development budget for professional growth.
- Team collaboration is a core value, requiring participation in daily standups, weekly planning, and monthly reviews.
- The role involves working with sensitive financial data, necessitating strict adherence to data protection protocols.
- Candidates should be prepared to demonstrate technical expertise through practical exercises or live coding sessions.
- The interview process may include discussions around past infrastructure projects and architecture decisions.
- This position reports to the senior leadership within the Engineering organization and impacts strategic decisions.
- The company reserves the right to adjust benefits and compensation structures as the business evolves.