Senior Site Reliability Engineer - Telephony & Communications Platform
Job description
Senior Site Reliability Engineer for Telephony and Communications Platform at Filevine.
About the role
Filevine is a Legal AI company that builds a unified platform for modern legal work. The company uses LOIS, the Legal Operating Intelligence System, to connect context across every matter and transform operations from reactive to proactive. This role focuses specifically on the Telephony and Communications Platform within the Reliability team. You will help create autonomous systems that ensure reliability, scalability, and security for professionals. The ultimate goal is to take care of difficult daily details so professionals can focus entirely on what they love doing. The team applies the principle of continuous improvement to make each system iteration better than the previous one. Being a professional is hard, which is why Filevine dreams of a day when all of these details are taken care of. The company has earned AI awards and recognition from Deloitte and Inc. as one of the most innovative and fastest growing technology companies.
Key facts
What you'll do
Collaborate closely with a cross functional team to understand key responsibilities for specific system portions.
Apply software engineering principles to solve complex reliability and scalability challenges for the core platform.
Design and build autonomous systems that handle operational details without requiring any manual intervention.
Focus on continuous improvement to make each iteration of the autonomous systems better than the last.
Gain valuable context during the first year to move at speed effectively within the company.
Take ownership of mission critical objectives in successive years to grow and improve the systems.
Ensure the platform is cost effective and performs reliably under heavy and unpredictable user load.
Implement strong security measures to protect sensitive legal data across the entire communications platform.
Build comprehensive disaster recovery mechanisms to ensure the system can recover quickly from unexpected failures.
Maintain and improve foundational pieces that were recently completed or are currently under active development.
Requirements
Demonstrate strong software engineering skills to apply them toward solving complex reliability engineering problems daily.
Possess substantial experience working with cross functional teams to deliver key system responsibilities effectively and consistently.
Understand the core principles of building autonomous systems that significantly reduce manual operational toil and burden.
Have a mindset focused on continuous improvement for iterative system development and ongoing maintenance of services.
Be capable of gaining valuable context quickly to move at speed in new and challenging environments.
Show aptitude for taking on mission critical objectives over successive years of professional engineering work.
Know how to ensure distributed systems are reliable, scalable, and cost effective for all types of users.
Have experience maintaining systems that can grow to internet scale over time and increasing demand.
Nice to have
Familiarity with legal technology or legal operating intelligence platforms is a highly valued and welcome addition to the team.
Experience with telephony systems and communications infrastructure would be extremely beneficial for succeeding in this important role.
Background in building foundational pieces for nascent autonomous system architectures is always valued highly by the reliability team.
Prior recognition or awards for innovation in fast growing technology companies is a nice and helpful bonus to have.
Skills & tools
Proficiency in software engineering practices to build reliable and scalable distributed systems with confidence and ease over time.
Knowledge of continuous improvement methodologies to enhance autonomous system operations over long periods of sustained growth and stability.
Ability to reason across complex data to surface insights and automate complex operational processes efficiently and effectively.
Skill in designing disaster recovery mechanisms for mission critical production environments that must stay online always without failure.
Expertise in ensuring cost effectiveness while maintaining high performance and strict security standards for sensitive legal data always.
Comfort with operating in a fast growing environment where foundational pieces are under active and constant development and change.
Practical notes
The state of autonomous systems at Filevine is currently nascent and evolving very rapidly on a daily basis.
You will be embedded with a cross functional team during your first year of employment to learn the environment thoroughly.
Successive years involve taking on specific mission critical objectives to build out the platform further and improve it consistently.
The ultimate goal is to allow professionals to focus on their work by handling operational details automatically and reliably.