Responsibilities
• Design and improve the monitoring, logging, distributed tracing, dashboards, alerting, SLIs,
and SLOs that give teams meaningful visibility into production health and customer impact.
• Build and maintain automation, internal tools, and CI/CD systems that increase engineering
efficiency, reduce toil, and support reliable deployments at scale. Take responsibility for the
quality and reliability of tools and services you support.
• Drive the implementation and continuous improvement of reliable systems for building,
deploying, testing, and operating Filevine products, proactively identifying and resolving
reliability, performance, scalability, and security risks before they impact customers.
• Own complex production incidents through detection, triage, communication, resolution,
and follow-up. Turn incident learning into durable corrective actions, stronger runbooks and
operating practices, and improvements that reduce recurring incidents and operational
burden.
implementation and adoption. Coordinate work across engineers and teams, communicate
tradeoffs and risks, and help ensure the work delivers the intended results.
• Mentor other Site Reliability Engineers through design reviews, incident follow-ups, paired
problem-solving, and meaningful delegation. Help engineers develop stronger technical
judgment and become increasingly capable of handling complex production work
independently.
• Participate in the shared on-call rotation and help ensure production systems are prepared
to operate reliably at scale through capacity planning, operational readiness, and
continuous improvements to resilience and recovery.
• Apply AI and machine learning to analyze operational signals, identify patterns, forecast
reliability and capacity risks, and implement improvements that make systems more
reliable, efficient, and easier to operate.
Qualifications
engineering, DevOps, or related technical roles, including at least 5 years in a Site Reliability
Engineering or reliability-focused role.
• Strong knowledge of distributed systems and hands-on experience operating Kubernetes
workloads and cloud infrastructure in AWS or a comparable platform, with proficiency in
Infrastructure as Code, monitoring, logging, alerting, distributed tracing, SLIs, and SLOs.
• Strong proficiency with Python, Go, Bash, or a similar language, with demonstrated
experience building and maintaining production tooling, automation, CI/CD pipelines, and
deployment systems that reduce toil, improve reliability, and simplify ongoing operations.
• Demonstrated ability to lead troubleshooting, incident response, root cause analysis, and
long-term reliability improvements for complex production systems, including the
elimination of recurring incidents and operational work.
• Proven ability to mentor Site Reliability Engineers, help others build stronger technical
judgment, communicate clearly with technical and business stakeholders, and lead complex
initiatives from planning through delivery.
• Demonstrated experience applying AI and machine learning to operational data and
engineering workflows to identify patterns, forecast reliability or capacity risks, and
implement measurable improvements with appropriate safeguards.
Skills Required
- 8+ years hands-on experience in software engineering, cloud infrastructure, platform engineering, DevOps, or related roles, including at least 5 years in SRE or reliability-focused role
- Strong knowledge of distributed systems and hands-on experience operating Kubernetes workloads
- Hands-on experience with cloud infrastructure in AWS or a comparable platform
- Proficiency in Infrastructure as Code
- Experience with monitoring, logging, alerting, distributed tracing, SLIs, and SLOs
- Strong proficiency with Python, Go, Bash, or a similar language and experience building production tooling and automation
- Experience building and maintaining CI/CD pipelines and deployment systems
- Proven ability to lead troubleshooting, incident response, root cause analysis, and long-term reliability improvements
- Proven ability to mentor Site Reliability Engineers and lead cross-team technical initiatives
- Demonstrated experience applying AI and machine learning to operational data to forecast risks and improve reliability
Filevine Compensation & Benefits Highlights
The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about Filevine and has not been reviewed or approved by Filevine.
-
Healthcare Strength — Health coverage is described as covering the major bases (medical, dental, vision) and is often framed as decent quality. In some cases, premiums and copays are portrayed as relatively favorable, suggesting tangible value from the plans.
-
Parental & Family Support — Paid parental leave is positioned as a standard, clearly offered benefit. The presence of parental leave alongside disability coverage signals baseline family-support provisions typical of growth-stage tech employers.
-
Fair & Transparent Compensation — Compensation is sometimes framed as fair or reasonable relative to role expectations, with technical roles in particular appearing closer to market-aligned ranges. This creates pockets where pay is perceived as competitive even if not consistently top-of-market across the company.
Filevine Insights
What We Do
Filevine is case management software built for and inspired by real attorneys. As a fully-featured suite of tools, it comes ready to manage every part of a moving case. Assign tasks, upload files or images, monitor staff productivity, and communicate with your client directly from within their case file. Our software is built on the truth that every law firm functions differently. That’s why Filevine is so customizable. Build new case-type templates, design automatic workflows, and receive customized reports on a schedule that fits your needs. Accessing your information is never a problem, because Filevine is hosted on The Cloud. To ensure security, your law firm’s data is protected through state-of-the-art encryption on redundant servers. All you need to get started is an internet connection and your favorite web browser. Learn more at filevine.com.
Gallery









