Remote Jobs RockRemote Jobs Rock

Site Reliability Engineer

📅 Jun 10
Site Reliability EngineeringAutomationCI/CDMonitoring

📜 Description

  • Design, build, and maintain monitoring, logging, and alerting systems for production health.
  • Develop automation and CI/CD improvements to increase engineering efficiency and support reliable deployments.
  • Implement and maintain systems for deploying and operating Filevine products, addressing reliability and performance risks.
  • Participate in a 24/7 on-call rotation, using insights to drive automation and reliability improvements.
  • Take ownership of technical initiatives, developing expertise in critical areas of the platform.

🛠️ Requirements

  • 4+ years of experience in software engineering, cloud infrastructure, or related roles, with at least 2 years in Site Reliability Engineering.
  • Working knowledge of distributed systems and ability to troubleshoot production issues.
  • Proficiency in Python, Bash, or similar scripting languages.
  • Experience with Kubernetes and cloud infrastructure in AWS or similar platforms.
  • Familiarity with Infrastructure as Code tools like Terraform or AWS CloudFormation.

Benefits

  • A dynamic, rapidly growing company, focused on helping organizations thrive - Medical
Full job description
Filevine is a Legal AI company delivering Legal Operating Intelligence for the future of legal work. Grounded in a singular system of truth, Filevine brings together data, documents, workflows, and teams into one unified platform—where modern legal work happens with clarity and consistency.
Powered by LOIS, the Legal Operating Intelligence System, Filevine connects context across every matter to transform legal operations from reactive to proactive. LOIS reads, understands, and reasons across your data to surface insight, automate complexity, and give professionals the clarity and confidence to see more, know more, and do more. Fueled by a team of exceptional collaborators and innovators, Filevine’s rapid growth has earned AI awards and recognition from Deloitte and Inc. as one of the most innovative and fastest-growing technology companies in the country.

Role Summary

As a Site Reliability Engineer at Filevine, you will improve the reliability, scalability, and
operational maturity of the Filevine platform. You’ll design automation that reduces toil,
strengthen observability, support reliable deployments at scale, and solve production challenges
that keep Filevine running for legal teams across the country.
This role is built for engineers who apply software engineering principles to infrastructure problems, thrive on complex
technical challenges, and are energized by taking ownership of the systems they build and
operate — growing into deeper expertise as their platform knowledge expands
Filevine is a Legal AI company delivering Legal Operating Intelligence for the future of legal work. Grounded in a singular system of truth, Filevine brings together data, documents, workflows, and teams into one unified platform—where modern legal work happens with clarity and consistency.   Powered by LOIS, the Legal Operating Intelligence System, Filevine connects context across every matter to transform legal operations from reactive to proactive. LOIS reads, understands, and reasons across your data to surface insight, automate complexity, and give professionals the clarity and confidence to see more, know more, and do more. Fueled by a team of exceptional collaborators and innovators, Filevine’s rapid growth has earned AI awards and recognition from Deloitte and Inc. as one of the most innovative and fastest-growing technology companies in the country. Compensation Information: $160,000 - $190,000   The base salary range represents the low and high end of the salary range for this position. The total compensation package for this position will be determined by each individual’s location, qualifications, education, work experience, skills and performance. We believe in the importance of pay equity - the range listed is just one component of Filevine’s total compensation package for employees. This position is also eligible for a paid time off policy, as well as a comprehensive benefits package.   Filevine is an Equal Opportunity Employer. Qualifications for employment, promotion and other terms and conditions of employment are based upon the ability to perform the job. Equal-employment opportunities are provided to all applicants and employees without regard to race, creed, religion, color, age, national origin, sex, disability, veteran status, or other legally protected class. Filevine is committed to providing reasonable accommodations for qualified individuals with disabilities. If you need assistance or accommodation due to disability, or if you have concerns related to Filevine’s equal employment opportunities, you may contact us at legal@filevine.com Work Location Expectation:This position is based out of our San Francisco, Chicago, Salt Lake City or New York offices. Candidates residing in or near these metro areas are expected to work on a hybrid (2-3 days in-office/week) or fully onsite basis.     Cool Company Benefits: - A dynamic, rapidly growing company, focused on helping organizations thrive  - Medical, Dental, & Vision Insurance (for full-time employees) - Competitive & Fair Pay - Maternity & paternity leave (for full-time employees) - Short & long-term disability - Opportunity to learn from a dedicated leadership team - Top-of-the-line company swag   Privacy Policy Notice Filevine will handle your personal information according to what’s outlined in our Privacy Policy.   Communication about this opportunity, or any open role at Filevine, will only come from representatives with email addresses using "filevine.com". Other addresses reaching out are not affiliated with Filevine and should not be responded to.

Responsibilities

  • Design, build, and maintain the monitoring, logging, distributed tracing, dashboards, and
    alerting that give teams meaningful visibility into production health.
  • Build automation, tooling, and CI/CD improvements that increase engineering efficiency,
    reduce toil, and support reliable deployments at scale.
  • Design, implement, and maintain reliable systems for building, deploying, testing, and
    operating Filevine products — proactively identifying and resolving reliability,
    performance, scalability, and security risks before they impact customers.
  • Participate in a shared 24/7 on-call rotation, using operational insights to drive automation
    and long-term reliability improvements; continuously improve runbooks, documentation,
    and engineering standards.
  • Take ownership of technical initiatives from design through implementation, develop deep
    expertise in critical areas of the Filevine platform, and communicate clearly with technical
    and business stakeholders.
  • What we are looking for

  • 4+ years of hands-on experience in software engineering, cloud infrastructure, platform
    engineering, DevOps, or related technical roles, including at least 2 years in a Site Reliability
    Engineering or reliability-focused role.
  • Working knowledge of distributed systems and how applications, infrastructure, and cloud
    services interact in production; demonstrated ability to troubleshoot production issues,
    perform root cause analysis, and drive long-term reliability improvements.
  • Proficiency with Python, Bash, or similar scripting languages; experience building
    production tooling, automation, or CI/CD pipelines and deployment automation.
  • Hands-on experience operating Kubernetes-based workloads and cloud infrastructure in
    AWS or a comparable platform, including compute, container orchestration, networking,
    IAM, object storage, and cloud-native monitoring.
  • Experience with Infrastructure as Code tools such as Terraform, Pulumi, or AWS
    CloudFormation, and familiarity with modern observability practices including monitoring,
    logging, alerting, distributed tracing, and incident response.
  • Experience using AI-assisted engineering tools to improve productivity, accelerate
    troubleshooting, or automate operational tasks; curiosity, ownership, and a passion for
    building reliable systems through continuous improvement.
  • Strong written and verbal communication skills; Bachelor’s degree in Computer Science,
    Information Systems, or a related field, equivalent industry certifications, or comparable
    professional experience.
  • Aristanetworks

    Site Reliability Engineer (SRE) - Engineering Productivity

    Aristanetworks👥 5001 - 10,000 employees🏢 Computer Networking
    🕒 6 days ago

    As a Site Reliability Engineer in the Engineering Productivity team, you will design, build, and operate scalable and reliable systems to enhance developer experience and support Arista's product development.

    GoPythonShell ScriptingLinux
    Anthropic

    Staff+ Site Reliability Engineer, Safeguards ML Infra

    Anthropic👥 10,000+ employees🏢 Research Services🤝 B2B
    🕒 4 days ago

    You'll lead the deployment and verification of safety systems for AI model launches, ensuring safeguards are effectively configured and operational across various platforms.

    Production Change ManagementDeploy PipelinesConfig Management SystemsCanary Analysis
    Cyberark1

    Principal Site Reliability Engineer (Sovereign Cloud)

    Cyberark1👥 10,000+ employees🏢 Computer & Network Security
    🕒 5 days ago

    Join Palo Alto Networks as a Principal Site Reliability Engineer to support and enhance our Sovereign Cloud infrastructure through automation, security, and reliability.

    Site Reliability EngineeringInfrastructure AutomationCloud Native ApplicationsKubernetes
    Servicetitan

    Senior Site Reliability Engineer

    Servicetitan👥 1001 - 5000 employees🏢 Software
    🕒 6 days ago

    Join our Site Reliability & Infrastructure Engineering team as a Senior Site Reliability Engineer, where you'll ensure the reliability and health of cloud applications while driving efficiency and innovation.

    KubernetesSRE PrinciplesAWSAzure

    Trusted by Remote Workers