Site Reliability Engineer
Job description
Site Reliability Engineer at Offchain Labs.
About the role
Offchain Labs seeks a Site Reliability Engineer to own the reliability and performance of critical infrastructure underpinning the Arbitrum ecosystem. The hire will own the design and stability of deployment pipelines that facilitate seamless chain operations across development and production environments. You will solve intricate infrastructure challenges that arise from supporting high-stakes financial applications and global scale. Your daily work will directly support ecosystem expansion and ensure alignment with key technical partners in a rapidly growing landscape. You will engage with pioneering decentralized systems while advancing a more equitable and accessible digital future for all participants. This role demands a commitment to the principles of decentralization, security, and transparency that define the Offchain Labs mission. Your contributions will help establish the foundational infrastructure that enables the next era of commerce and governance on Ethereum.
Key facts
What you'll do
- Architect and maintain deployment pipelines essential for the stability of chain operations within the Offchain Labs environment.
- Engineer foundational components that facilitate chain infrastructure and secure wallet interactions for diverse applications.
- Scrutinize deployment changes to ensure strict adherence to the technical specifications of Arbitrum One.
- Roll out updates that preserve the integrity and stability of Arbitrum stacks utilized by major institutional clients.
- Liaise with technical partners to synchronize objectives and drive alignment with broader ecosystem growth initiatives.
- Execute performance testing under heavy load conditions for high-profile projects such as Aave and Robinhood.
- Author comprehensive operational procedures for Prysm and ZeroDev environment management and maintenance.
- Investigate and resolve incidents that impact custom chains built upon Offchain Labs technology and associated tooling.
- Monitor and optimize data intake mechanisms to ensure efficient and reliable information flow into core systems.
- Develop and implement robust monitoring strategies to provide deep visibility into infrastructure health and performance metrics.
- Collaborate with cross-functional teams to identify bottlenecks and implement scalable solutions for infrastructure challenges.
- Conduct in-depth analysis of system logs and telemetry data to preemptively identify potential failures or security concerns.
- Evaluate emerging technologies and methodologies to continuously improve the resilience and efficiency of deployment processes.
- Serve as a technical expert to support other teams in understanding infrastructure constraints and capabilities.
Requirements
- Must possess three years of experience managing complex distributed systems within production environments at scale.
- Demonstrate a deep understanding of Ethereum scaling solutions specifically as they apply to large financial applications and workflows.
- Have a proven history of working with infrastructure that securely handles high transaction volumes across demanding use cases.
- Bring experience with deployment strategies for networks that support major financial entities such as Securitize and Ethena Labs.
- Exhibit strong proficiency with core infrastructure components including Arbitrum, Prysm, and ZeroDev.
- Show capability to work effectively within a fast-paced, agile environment while maintaining rigorous standards.
- Possess excellent written and verbal communication skills to articulate complex technical concepts to diverse stakeholders.
- Display a strong sense of ownership and accountability for system reliability and operational excellence.
- Adhere to best practices in security, ensuring all infrastructure components maintain robust postures against evolving threats.
- Commit to continuous learning and adaptation in the face of rapidly changing blockchain and infrastructure technologies.
Nice to have
- Experience with tools utilized by teams building on Arbitrum ecosystem infrastructure is valued and provides a distinct advantage.
Practical notes
This role is based in the United States and requires adherence to standard full-time working hours as defined by company policy. No specific travel requirements are outlined in the current listing. Visa sponsorship information and application deadlines should be verified on the official application page. Please About the company
Offchain provides technical infrastructure used by global companies. The focus is on protocol level breakthroughs. These breakthroughs support movement toward secure, programmable onchain infrastructure. Current work centers on foundational components that chains and applications can build upon. Services target organizations needing scalable base layers. Progress is measured by improvements in security and programmability. Teams rely on these core systems to operate with dependable structure and predictable behavior in production environments.