Staff Software Engineer, Core Infrastructure
About this role
Staff Software Engineer, Core Infrastructure
At Harvey, we're transforming legal and professional services with frontier agentic AI and enterprise-grade infrastructure. As a Staff Software Engineer on the Core Infrastructure team, you'll design and scale the systems powering billions of prompt tokens and millions of daily requests across our global platform. Your work will directly impact the reliability, scalability, and security that enables Harvey to serve the world's leading law firms and professional service providers.
What you'll do
- Design and build scalable, fault-tolerant infrastructure systems that power Harvey's AI platform across multiple cloud regions
- Own and evolve our multi-cloud infrastructure (Azure, GCP), including Kubernetes orchestration, networking, and container management
- Lead technical initiatives around observability, incident response, and operational excellence to enable rapid issue detection and resolution
- Architect and optimize distributed systems for reliability, including load balancing, quota management, and failover mechanisms
- Partner with Product Engineering and Security teams to ensure infrastructure accelerates rather than constrains product delivery
- Drive infrastructure-as-code practices using Terraform and Pulumi to enable reproducible, auditable deployments
- Mentor engineers and raise the technical bar across the organization through code reviews, design reviews, and technical leadership
What Harvey is looking for
- 10+ years of experience in Infrastructure or Platform Engineering in production environments
- Proven track record building and scaling complex, large-scale distributed systems
- Deep proficiency with cloud infrastructure platforms (Azure preferred; GCP or AWS experience transfers well)
- Strong fluency in Infrastructure as Code tools (Terraform, Pulumi, or CloudFormation)
- Solid understanding of Kubernetes, container orchestration, networking, and cloud security at scale
- Experience with observability tools (Datadog, Sentry) and incident response practices
- Strong programming skills in Python, Go, or similar languages
- Excellent problem-solving skills and commitment to operational excellence
- Nice to have: Experience with AI/ML infrastructure, distributed rate limiting, multi-tenant platforms, or leading complex cross-functional projects
What happens next
Skip the application pile. I get you in front of the people who decide.
Confirm the fit
A few questions to make sure this role is the right shape for you. Two minutes.
I pitch you to the company
I write the intro, send it to the founder, and handle the back-and-forth.
A meeting lands on your calendar
When the company wants to meet, I get the call on your calendar. You just show up.
Know someone who'd be great for this?

