Staff Site Reliability Engineer

Filevine · Worldwide

Apply ↗full time

Filevine is a Legal AI company delivering Legal Operating Intelligence for the future of legal work. Grounded in a singular system of truth, Filevine brings together data, documents, workflows, and teams into one unified platform—where modern legal work happens with clarity and consistency.

Powered by LOIS, the Legal Operating Intelligence System, Filevine connects context across every matter to transform legal operations from reactive to proactive. LOIS reads, understands, and reasons across your data to surface insight, automate complexity, and give professionals the clarity and confidence to see more, know more, and do more. Fueled by a team of exceptional collaborators and innovators, Filevine’s rapid growth has earned AI awards and recognition from Deloitte and Inc. as one of the most innovative and fastest-growing technology companies in the country.

Role Summary

As a Staff Site Reliability Engineer at Filevine, you are the senior technical authority on the SRE team

and a strategic partner to engineering leadership. You don’t just maintain systems — you shape

engineering culture, define the technical standard for how Filevine runs in production, and bridge the

gap between high-level business goals and robust, internet-scale technical execution. You bring a

forward-looking perspective — actively shaping how AI and machine learning drive the future of

reliability practice.

You own the roadmap across two critical SRE domains — Observability & Alerting and Platform

Infrastructure — and are accountable for ensuring the team solves reliability problems permanently

rather than absorbing them as toil. You operate as the senior IC counterpart to the Engineering

Manager: technical correctness lives with you. You partner with the Reliability Architect and engineering

leadership on significant technical decisions, mentor engineers across experience levels, and influence

reliability strategy across the broader organization. Reliability at Filevine protects revenue. You are the

senior technical voice responsible for ensuring that uptime, incident response, and every production

change meet the operational standard the business demands.

This role does not participate in on-call rotation, but you are deeply invested in the engineers who do —

shaping the on-call strategy, tooling, and culture that make production support sustainable and

effective.

Who You Are

The Technical Authority

• Master of the Craft: You bring deep expertise in distributed systems, cloud infrastructure,

observability, and reliability engineering. You raise the technical standard for every engineer

around you and thrive where the challenges are complex and the stakes are real.

• Technical Leader and Mentor: You are passionate about mentoring engineers and investing in

their growth. You influence technical direction and communicate production risk clearly across

engineering, product, and executive audiences.

• Forward-Thinking & AI/ML Fluent: You bring deep knowledge of AIOps and drive the use of

AI and machine learning in observability, anomaly detection, incident response, automated

remediation, and resource optimization.

• Production-Scale Problem Solver: You turn ambiguous, complex reliability challenges into

durable solutions for systems where availability, performance, and production changes carry

meaningful business impact.

• Software-Minded Builder: You use software, automation, Infrastructure as Code, and platform

capabilities to eliminate toil and make systems safer, more scalable, and easier to operate.

What you will do

Define and execute the technical strategy for Observability & Alerting, Platform Infrastructure,

and ope…

Don't just apply to this one.

Cold email the hiring manager and skip the pile. I'll write the email for you.

Book a call ↗

Related jobs