Harvey

Senior Software Engineer, Site Reliability Engineer

$200K–$260KFull-time · San Francisco
✓ Verified live on the employer's own system · added 251 days ago
Save search
Mid-level · 5+ yrs exp

Requirements

Experience: 5+ years

Skills & tools

OperationsManagementRoot Cause AnalysisSecurityDevopsCloud PlatformsPython

Benefits — mentioned in this posting

Relocation
Apply on company site ↗ See your fit → free

Full job description

As a Software Engineer on the Site Reliability team at Harvey, you will ensure the reliability, scalability, and performance of our legal AI platform. You’ll join a high-leverage team that sits at the intersection of infrastructure and product, owning the systems that keep our platform fast, secure, and always on. From scaling across 50+ regions to automating mission-critical operations, your work will ensure that Harvey remains resilient as we grow.

If you’re passionate about building robust systems and reducing complexity through automation, we’d love to work with you.

This role is based in San Francisco, CA. We use an in-person work model and offer relocation assistance to new employees.

- Design, implement, and manage monitoring, alerting, and infrastructure resources (compute, storage, networking) across 50+ global regions

- Lead incident management processes, including postmortems, root cause analyses, and driving actionable improvements

- Automate operational tasks and workflows, building tools and processes for capacity planning, graceful rollouts, and safe data access to maintain high reliability and reduce manual intervention

- Collaborate across teams to drive reliability, security, and compliance throughout the software lifecycle

- Optimize infrastructure costs through strategic capacity planning and build-versus-buy decisions while maintaining system performance, reliability, and functionality.

- 5+ years of experience in Site Reliability Engineering or similar roles supporting production environments

- Expertise in infrastructure as code(IaC) tools (Pulumi, Terraform, CloudFormation, etc.).

- Deep familiarity with observability tools (Datadog, Sentry, etc.) and incident response practices (PagerDuty, IncidentIO, etc.)

- Proficiency with cloud infrastructure platforms (Azure, GCP, AWS, etc.)

- Strong programming skills (Python, Bash, Go, or similar languages)

- Proven track record of diagnosing complex system problems and implementing durable solutions

- Solid understanding of CI/CD, Kubernetes, containerization, networking, databases, and cloud security principles

- Excellent problem-solving skills, meticulous attention to detail, and a commitment to operational excellence

More jobs at Harvey

Similar jobs near San Francisco

Tell me when more Senior Site Reliability Engineer jobs post near San Francisco We re-check every listing against the employer’s own board — no résumé needed.

Search Senior Software Engineer, Site Reliability Engineer jobs near San Francisco → Browse all live jobs

This posting was published by Harvey on their own careers system and is shown here with a direct link to apply there. Employers: for corrections or removal, contact jobs@veritahire.com.