Filevine

Filevine

Filevine is a top legal tech company revolutionizing legal work with AI-powered case management software, empowering law firms to streamline operations and enhance client services.

Specialized Consumer Services
251-1K
Founded 2015
$226M raised

Description

  • Own and evolve the observability strategy, including monitoring, alerting, dashboards, logging, and distributed tracing.
  • Define and manage SLIs, SLOs, and reliability metrics.
  • Lead incident response, postmortems, and continuous improvement initiatives.
  • Improve MTTD and MTTR through automation and operational excellence.
  • Integrate observability into CI/CD pipelines and software delivery workflows.
  • Build and maintain reliable cloud infrastructure on AWS and Kubernetes.
  • Mentor engineers and promote SRE best practices across the organization.
  • Support reliability engineering efforts across software delivery and operations.

Requirements

  • 8+ years of experience in software engineering, infrastructure, or operations.
  • 5+ years of Site Reliability Engineering experience.
  • Deep expertise with observability platforms such as New Relic, Datadog, Dynatrace, Grafana, or Prometheus.
  • Strong experience with monitoring, alerting, incident management, and reliability engineering practices.
  • Hands-on experience with AWS, Kubernetes, and cloud-native technologies.
  • Proficiency in Python, Bash, PowerShell, or similar scripting languages.
  • Excellent communication and collaboration skills.
  • Experience leading observability platform implementations or migrations at scale (preferred).
  • Experience building SLI/SLO frameworks and reliability programs (preferred).
  • Experience with OpenTelemetry, distributed tracing, and modern observability architectures (preferred).

Benefits

  • $175,000 - $195,000 annual salary.
  • Medical, Dental, and Vision insurance for full-time employees.
  • Competitive and fair pay.
  • Maternity and paternity leave for full-time employees.
  • Short- and long-term disability coverage.
  • Opportunity to learn from a dedicated leadership team.
  • Top-of-the-line company swag.
  • Remote full-time work arrangement.

Interested in this position?

Apply directly on the company website

Apply Now

Similar Roles

Vice President, Global Production Operations & Reliability

Everbridge 1K-5K Internet Software & Services

Everbridge is hiring a Vice President, Global Production Operations & Reliability to lead the company’s global production operations for its cloud-native SaaS platform and drive reliability, scalability, security, and operational excellence.

AWS CI/CD Kubernetes
20 hours, 25 minutes ago

DevOps Engineer - SRE Observability

Lingaro 5K-10K IT Services

An infrastructure-focused role at Lingaro responsible for monitoring, automating, and designing cloud systems within an Azure-based environment.

Azure Azure Pipelines CI/CD Docker GitHub GitHub Actions Grafana Kubernetes MySQL PostgreSQL Prometheus SQL Terraform
20 hours, 25 minutes ago

Staff Field Reliability Engineer

Honeycomb.io 51-250 Internet Software & Services

Honeycomb is hiring a Field Reliability Engineer to lead complex customer escalations, managed infrastructure operations, and observability strategy for its cloud-based platform.

Ansible AWS Chef EC2 Go Helm Honeycomb Java Kubernetes .NET Node.js OpenTelemetry Python Serverless Terraform TypeScript
20 hours, 40 minutes ago

Site Reliability Engineer

Yuno 51-200 Payment Processing Software

Yuno is seeking a Staff Site Reliability Engineer to define and lead reliability for its AWS-based platform that provisions and manages AI agents powering global payments at scale.

Apache Airflow AWS Databricks Datadog Docker EC2 Fly.io GCP Go Kafka Kubernetes MLflow MLOps MongoDB NATS OpsGenie PagerDuty PostgreSQL Prefect Pulumi Python RabbitMQ Railway Redis Snowflake SQL Terraform
21 hours, 10 minutes ago

You're on a roll! Sign up now to keep applying.

Sign Up

Already have an account? Log in

Used by 14,729+ remote workers