Live opening · Posted 6 days ago

ML Infrastructure Engineer

Clera · San Mateo
Ashby No FullTime
You are 6 days behind. JobBeeper subscribers saw this role while it was still new.

At a glance

The key details from the original listing.

Posted 6 days ago
CompanyClera
LocationSan Mateo
Job typeFullTime
Work modeNo
SourceAshby
Listed6 days ago

Your early-applicant advantage

Live timing from JobBeeper.

Live data
11 min from Ashby publishing this role to us finding it
10 min median time from a role going live to a subscriber being told
6 hours subscribers had this role before this page existed
73,932 roles found in the last 24 hours — the newest are not on this site yet
Start your free trial →

About the role

Description supplied by the original job listing.

ABOUT THE ROLE
This is a hands-on ML Infrastructure Engineer role at an early-stage enterprise AI company building a context and data governance layer that makes AI agents reliable in production. You will own the inference and model-serving infrastructure end to end, ensuring agents run fast and reliably at increasing concurrency. The work is squarely production-focused with real-world impact across regulated industries like insurance, banking, healthcare, and asset management.
WHAT YOU'LL DO
- Design, build, and scale inference and model-serving infrastructure from the ground up through production deployment.
- Optimize systems for latency, throughput, and reliability under high concurrency.
- Collaborate closely with ML and infrastructure teams to ensure seamless integration and surface performance bottlenecks.
- Drive solutions to infrastructure challenges across a fast-moving, cross-functional team.
WHAT WE'RE LOOKING FOR
- 5 or more years building and operating machine learning inference systems, model-serving platforms, or ML infrastructure in production environments.
- Hands-on experience designing and scaling inference-serving systems using tools such as TensorFlow Serving, TorchServe, Triton, KServe, or equivalent custom solutions.
- Strong distributed systems fundamentals, including containerization and orchestration with Docker and Kubernetes.
- Proficiency with monitoring and observability tooling for production systems, such as Prometheus, Grafana, or distributed tracing frameworks.
- Experience deploying and managing ML workloads on cloud platforms (AWS, GCP, or Azure).
- Proficiency in at least one systems or backend language: Python, Go, Rust, C++, or Java.
- Comfort collaborating across both ML and infrastructure disciplines in a fast-paced environment.
- Nice to have: experience with knowledge graphs, semantic search, or graph databases; real-time or low-latency inference systems; agentic or multi-step AI pipelines; enterprise data integration or pipeline infrastructure.
LOCATION
On-site in San Mateo, California, United States. Visa sponsorship is not available for this role.

Employment type
FullTime

Work arrangement
No

Get JobBeeper Mobile App

Never miss a job opening! Get instant job alerts on your phone.

Subscribers see fresh openings within minutes. Download the JobBeeper App on Google Play to get real-time push notifications and apply before anyone else.

⚡ Instant Push Alerts 🎯 Tailored Filters 🚀 Direct Employer Links
GET IT ON Google Play

More openings worth a look

Recently tracked roles with full details and direct application links.

6 roles
Good roles move before most people even see them. Tell JobBeeper what you want and get fresh matches delivered in minutes.
Start your free trial →
⚡ Get fresh job alerts 📱 Get App