Live opening · Posted 27 days ago

Staff Software Engineer

Harvey · Bangalore
Instahyre 8-12 yrs
You are 27 days behind. JobBeeper subscribers saw this role while it was still new.

At a glance

The key details from the original listing.

Posted 27 days ago
CompanyHarvey
LocationBangalore
Experience8-12 yrs
SourceInstahyre
Listed27 days ago

Your early-applicant advantage

Live timing from JobBeeper.

Live data
0 min from Instahyre publishing this role to us finding it
8 min median time from a role going live to a subscriber being told
6 hours subscribers had this role before this page existed
15,865 roles found in the last 24 hours — the newest are not on this site yet
Start your free trial →

About the role

Description supplied by the original job listing.

As a Staff Software Engineer on the Core Infrastructure team at Harvey, you'll play a critical role in designing and building new infrastructure systems while equally scaling and strengthening our existing infrastructure. Our infrastructure is the foundation that powers every user interaction with Harvey, processing billions of prompt tokens and millions of daily requests across our global legal AI platform. You'll work in an environment balanced between innovation, building new systems, and operational excellence, ensuring that Harvey remains resilient and efficient as it scales products, regions, customers, and usage. Your contributions will directly impact the reliability, scalability, and security of our platform as we serve the world's leading law firms and professional service providers.
Responsibilities:
Design and build scalable, fault-tolerant infrastructure systems that power Harvey's AI platform across multiple cloud regions.
Own and evolve our multi-cloud infrastructure (Azure, GCP), including Kubernetes orchestration, networking, and container management.
Lead technical initiatives around observability, incident response, and operational excellence, building systems that enable rapid detection and resolution of issues.
Architect and optimize our distributed systems for reliability, including load balancing, quota management, and failover mechanisms.
Partner with Product Engineering and Security teams to ensure our infrastructure is an accelerant, not a constraint.
Drive infrastructure-as-code practices using tools like Terraform and Pulumi to enable reproducible, auditable deployments.
Mentor engineers and raise the technical bar across the organization through code reviews, design reviews, and technical leadership.
Representative Projects: Design and implement a next-generation model proxy architecture that routes millions of daily inference requests while maintaining model API compatibility and enabling seamless model integration.
Build distributed rate limiting and quota management systems using Redis-backed algorithms to handle bursty traffic patterns without degrading user experience.
Architect multi-region deployment strategies that meet strict data residency requirements for global enterprise customers.
Develop a comprehensive observability infrastructure with granular SLA monitoring, burn rate alerts, and detailed token attribution for cost tracking.
Lead the evolution of our CI/CD pipelines to improve developer velocity while maintaining production stability.
Requirements:
8+ years of experience in Infrastructure Engineering or Platform Engineering in a production environment.
Long track record building and scaling complex, large-scale distributed systems.
Deep proficiency with cloud infrastructure platforms (Azure preferred; GCP or AWS experience transfers well).
Strong fluency in Infrastructure as Code (IaC) tools, such as Terraform, Pulumi, or CloudFormation.
Solid understanding of Kubernetes, container orchestration, networking, and cloud security at scale.
Experience with observability tools (Datadog, Sentry) and incident response practices (PagerDuty, Incident.io ).
Strong programming skills in Python, Go, or similar languages.
Excellent problem-solving skills, a "spidey sense" of where things could go wrong, and a commitment to operational excellence.
Experience building infrastructure for AI/ML workloads or high-throughput inference systems.
Background with distributed rate limiting, load balancing, or quota management systems.
Experience operating multi-tenant platforms with strict security and compliance requirements.
Track record of leading complex cross-functional projects and delivering measurable impact.

Experience
8-12 yrs

Get JobBeeper Mobile App

Never miss a job opening! Get instant job alerts on your phone.

Subscribers see fresh openings within minutes. Download the JobBeeper App on Google Play to get real-time push notifications and apply before anyone else.

⚡ Instant Push Alerts 🎯 Tailored Filters 🚀 Direct Employer Links
GET IT ON Google Play

More openings worth a look

Recently tracked roles with full details and direct application links.

6 roles
Good roles move before most people even see them. Tell JobBeeper what you want and get fresh matches delivered in minutes.
Start your free trial →
⚡ Get fresh job alerts 📱 Get App