Filevine is a leading Legal AI company that provides a unified platform for legal operations, combining data, documents, and workflows into a single system of truth. As a Staff Site Reliability Engineer, you will serve as the senior technical authority, shaping engineering culture and defining the standards for how our systems run at scale.
Key responsibilities
- Define and execute the technical strategy for observability, alerting, and platform infrastructure.
- Lead the evolution of reliable, scalable, and secure cloud platforms and distributed systems.
- Champion SLIs, SLOs, error budgets, and capacity planning across the service lifecycle.
- Lead the organization through complex production incidents and drive permanent engineering improvements.
- Build self-service platform capabilities to reduce toil and improve engineering velocity.
Requirements
- Over 12 years of experience in software engineering, infrastructure, or SRE, with at least 6 years specifically in SRE.
- Expert-level depth in observability, platform infrastructure, and incident response.
- Advanced experience with container orchestration platforms like Kubernetes and observability tools such as Datadog or New Relic.
- Strong software engineering skills in languages such as Python, Go, or Bash for building automation and tooling.
- Proven ability to mentor engineers and communicate technical risks to diverse stakeholders.
What we offer
- Opportunity to work in a dynamic, rapidly growing company focused on innovation.
- Comprehensive medical, dental, and vision insurance plans.
- Supportive leave policies including maternity and paternity options.
- Access to a dedicated leadership team and professional growth opportunities.