Role at a glance
- Salary
- $200K – $260K/yr
- Location
- San Francisco, United States
- Work arrangement
- On-site
- Employment
- Full-time
- Experience
- 5+ years of experience in Site Reliability Engineering or similar roles supporting production environments
Spotted an issue?
We’ll check it against the original posting.
Role Summary
This Software Engineer role on Harvey’s Site Reliability team focuses on the reliability, scalability, and performance of its legal AI platform. The team owns infrastructure and operational systems supporting the platform across more than 50 global regions, with an emphasis on automation, security, and resilience.
What You'll Do
- Design, implement, and manage monitoring, alerting, and infrastructure resources across 50+ global regions
- Lead incident management processes, including postmortems, root cause analyses, and actionable improvements
- Automate operational tasks and workflows for capacity planning, graceful rollouts, and safe data access
- Collaborate across teams to drive reliability, security, and compliance throughout the software lifecycle
- Optimize infrastructure costs through capacity planning and build-versus-buy decisions while maintaining system performance,...
Generated from the employer's posting. Verify important details before applying.
View full postingQualifications
5+ years of experience in Site Reliability Engineering or similar production-supporting roles; expertise with infrastructure-as-code tools; familiarity with observability and incident response tools; proficiency with cloud infrastructure platforms; programming skills in Python, Bash, Go, or similar languages; experience diagnosing complex system problems and implementing durable solutions; understanding of CI/CD, Kubernetes, containerization, networking, databases, and cloud security principles.
Required
- 5+ years of experience in Site Reliability Engineering or similar roles supporting production environments
- Infrastructure as code tools, including Pulumi, Terraform, or CloudFormation
- Observability tools, including Datadog or Sentry
- Incident response practices and tools, including PagerDuty or IncidentIO
- Cloud infrastructure platforms, including Azure, GCP, or AWS
- Programming skills in Python, Bash, Go, or similar languages
- Diagnosing complex system problems and implementing durable solutions
- CI/CD, Kubernetes, containerization, networking, databases, and cloud security principles
Original job description
Content provided by the employer
Original job description
Content provided by the employer
Why Harvey
At Harvey, we’re transforming how legal and professional services operate. By combining frontier agentic AI, an enterprise-grade platform, and deep domain expertise, we’re reshaping how critical knowledge work gets done for decades to come.
This is a rare chance to help build a generational company at a true inflection point. We have strong product-market fit and world-class investor support. We’re scaling fast and defining a new category in real time. The work is ambitious, the bar is high, and the opportunity for growth — personal, professional, and financial — is unmatched.
Our team moves fast, takes ownership, and is deeply committed to the mission — operating with intensity, staying close to our customers, and pushing each other for excellence. We live by three values: Decisiveness, Simplicity, and Job's Not Finished. We act quickly on clear judgment over perfect information, we believe simplicity is what scales, and we're never satisfied with where we are. If you want to do the best work of your career alongside people who share that drive, we'd love to build with you.
At Harvey, the future of professional services is being written today — and we’re just getting started.
Role Overview
As a Software Engineer on the Site Reliability team at Harvey, you will ensure the reliability, scalability, and performance of our legal AI platform. You’ll join a high-leverage team that sits at the intersection of infrastructure and product, owning the systems that keep our platform fast, secure, and always on. From scaling across 50+ regions to automating mission-critical operations, your work will ensure that Harvey remains resilient as we grow. If you’re passionate about building robust systems and reducing complexity through automation, we’d love to work with you.
This role is based in San Francisco, CA. We use an in-person work model and offer relocation assistance to new employees.
What You’ll Do
Design, implement, and manage monitoring, alerting, and infrastructure resources (compute, storage, networking) across 50+ global regions
Lead incident management processes, including postmortems, root cause analyses, and driving actionable improvements
Automate operational tasks and workflows, building tools and processes for capacity planning, graceful rollouts, and safe data access to maintain high reliability and reduce manual intervention
Collaborate across teams to drive reliability, security, and compliance throughout the software lifecycle
Optimize infrastructure costs through strategic capacity planning and build-versus-buy decisions while maintaining system performance, reliability, and functionality.
What You Have
5+ years of experience in Site Reliability Engineering or similar roles supporting production environments
Expertise in infrastructure as code(IaC) tools (Pulumi, Terraform, CloudFormation, etc.).
Deep familiarity with observability tools (Datadog, Sentry, etc.) and incident response practices (PagerDuty, IncidentIO, etc.)
Proficiency with cloud infrastructure platforms (Azure, GCP, AWS, etc.)
Strong programming skills (Python, Bash, Go, or similar languages)
Proven track record of diagnosing complex system problems and implementing durable solutions
Solid understanding of CI/CD, Kubernetes, containerization, networking, databases, and cloud security principles
Excellent problem-solving skills, meticulous attention to detail, and a commitment to operational excellence
Compensation Range
$200,000 - $260,000 USD
Depending on your location, an Applicant Privacy Notice may apply to you. You can find all of our Applicant Privacy Notices [here].
#LI-AN2
Harvey is an equal opportunity employer and does not discriminate on the basis of race, gender, sexual orientation, gender identity/expression, national origin, disability, age, genetic information, veteran status, marital status, pregnancy or related condition, or any other basis protected by law.
We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made by emailing [email protected]
About the company
Harvey
Startup
Harvey is an innovative technology company focused on streamlining the recruitment process through cutting-edge artificial intelligence solutions. By leveraging data-driven insights, Harvey enhances the connection between employers and job seekers, making it easier for companies to find the right talent efficiently. With a commitment to improving the hiring experience for both candidates and organizations, Harvey is at the forefront of transforming traditional recruitment methods into a more effective, user-friendly approach.
Harvey is an innovative technology company focused on streamlining the recruitment process through cutting-edge artificial intelligence solutions. By leveraging data-driven insights, Harvey enhances the connection between employers and job seekers, making it easier for companies to find the right talent efficiently. With a commitment to improving the hiring experience for both candidates and organizations, Harvey is at the forefront of transforming traditional recruitment methods into a more effective, user-friendly approach.