Staff Software Engineer, Production Engineering

Harvey


Job Location:

New York City, NY - USA

Yearly Salary: USD 231000 - 340000
Posted on: 3 hours ago
Vacancies: 1 Vacancy

Department:

Engineering

Job Summary

Why Harvey

At Harvey were transforming how legal and professional services operate. By combining frontier agentic AI an enterprise-grade platform and deep domain expertise were reshaping how critical knowledge work gets done for decades to come.

This is a rare chance to help build a generational company at a true inflection point. We have strong product-market fit and world-class investor support. Were scaling fast and defining a new category in real time. The work is ambitious the bar is high and the opportunity for growth personal professional and financial is unmatched.

Our team moves fast takes ownership and is deeply committed to the mission operating with intensity staying close to our customers and pushing each other for excellence. We live by three values: Decisiveness Simplicity and Jobs Not Finished. We act quickly on clear judgment over perfect information we believe simplicity is what scales and were never satisfied with where we are. If you want to do the best work of your career alongside people who share that drive wed love to build with you.

At Harvey the future of professional services is being written today and were just getting started.

Role Overview

Harvey is building the AI platform trusted by the worlds leading law firms and enterprises. Our infrastructure is the foundation that powers every customer interaction every model inference and every production workload.

Were looking for a Production Engineer to help build and operate Harveys core compute and networking infrastructure Kubernetes platform workflow orchestration platform and production infrastructure foundations. Youll work on the systems that enable engineering teams to move quickly and operate reliable services at scale.

In this role youll improve the reliability scalability security and efficiency of Harveys infrastructure platform. Youll solve complex production challenges across compute fleet management capacity planning infrastructure automation and production operations. Youll partner closely with Product Engineering Security AI Infrastructure and Platform teams to ensure our infrastructure scales with Harveys rapid growth.

At Harvey we value Decisiveness Simplicity and the belief that Jobs Not Finished. We move quickly prioritize clarity and continuously raise the bar for engineering excellence.

What Youll Do

Infrastructure Engineering & Technical Leadership

  • Design build and operate the production infrastructure that powers Harveys products and AI workloads.

  • Drive technical direction across compute infrastructure networking Kubernetes workflow orchestration and production operations.

  • Lead complex cross-functional technical initiatives that improve reliability scalability security operational efficiency and infrastructure cost.

  • Partner with Product Engineering Security AI Infrastructure and Platform teams to translate product and business requirements into resilient infrastructure solutions.

  • Establish reusable patterns tooling and paved paths that help engineering teams ship and operate production services safely.

  • Raise the engineering bar through thoughtful design reviews clear technical documentation operational rigor and mentorship.

Infrastructure Foundation & Production Operations

  • Build and operate Harveys global compute and network infrastructure ensuring high availability scalability reliability and performance.

  • Improve compute utilization performance and service availability while supporting rapidly growing AI workloads.

  • Develop capacity models demand forecasts and fleet lifecycle automation to help infrastructure scale efficiently with business growth.

  • Operate and continuously improve Harveys Kubernetes platform including cluster provisioning upgrades networking monitoring reliability performance and operational automation.

  • Drive infrastructure cost efficiency through capacity management resource rightsizing workload optimization and utilization monitoring.

  • Build secure infrastructure foundations including identity and access management network isolation secrets management auditing and compliance controls.

  • Develop scalable Infrastructure-as-Code and automation frameworks using technologies such as Terraform and Pulumi.

  • Improve observability monitoring alerting incident response and operational readiness across the infrastructure platform.

  • Participate in the on-call rotation lead incident response when needed and turn production learnings into durable engineering improvements.

What You Have

  • 10 years of software infrastructure site reliability or production engineering experience.

  • Deep experience building and operating large-scale cloud infrastructure on AWS Azure or Google Cloud Platform.

  • Strong hands-on experience operating Kubernetes in production including cluster lifecycle management networking and reliability.

  • Experience building and operating distributed systems with strong reliability scalability and performance characteristics.

  • Experience with infrastructure automation and Infrastructure-as-Code using tools such as Terraform or Pulumi.

  • Strong understanding of compute infrastructure networking capacity planning fleet management and production operations.

  • Experience designing and operating observability systems including monitoring logging alerting and incident response.

  • Strong understanding of infrastructure security including IAM network security secrets management and compliance best practices.

  • A track record of driving complex cross-functional technical initiatives and influencing engineering decisions without relying on formal authority.

  • Excellent communication skills and the ability to explain technical concepts clearly to engineering partners and other stakeholders.

  • A systems-thinking mindset and a passion for building simple reliable and scalable infrastructure platforms.

Nice to Have

  • Experience supporting AI/ML or LLM infrastructure at scale.

  • Experience operating GPU fleets high-performance compute infrastructure or large-scale capacity planning.

  • Experience with multi-cloud infrastructure or hybrid cloud environments.

  • Experience building internal platforms or developer tooling that improves engineering velocity and production safety.

Compensation

$231000 - $340000 USD

Depending on your location an Applicant Privacy Notice may apply to you. You can find all of our Applicant Privacy Notices here.

#LI-AN2

Harvey is an equal opportunity employer and does not discriminate on the basis of race gender sexual orientation gender identity/expression national origin disability age genetic information veteran status marital status pregnancy or related condition or any other basis protected by law.

We are committed to providing reasonable accommodations to applicants with disabilities and requests can be made by emailing


Required Experience:

Staff IC

Why HarveyAt Harvey were transforming how legal and professional services operate. By combining frontier agentic AI an enterprise-grade platform and deep domain expertise were reshaping how critical knowledge work gets done for decades to come.This is a rare chance to help build a generational compa...

About Company

Company Logo

Professional Class AI – Harvey is the platform built to meet the standards of the world’s leading professional service firms.

View Profile View Profile