---
title: "ML Infrastructure Engineer"
company: "Clera"
company_url: "https://www.remjobs.works/companies/clera"
url: "https://www.remjobs.works/job/clera-ml-infrastructure-engineer-966a5e31-ffc2-449c-b9db-8fb4fc2b3ff8"
apply_url: "https://jobs.ashbyhq.com/clera/88276d4c-0d4b-47d8-984c-0b315403d107"
workplace: onsite
location: "San Mateo"
employment_type: full-time
seniority: mid
role: ai-machine-learning
region: united-states
date_posted: 2026-09-22T17:23:34.142Z
first_seen_by_remjobs: 2026-09-22T17:30:12.209Z
---

# ML Infrastructure Engineer

**Clera** · San Mateo

Apply: https://jobs.ashbyhq.com/clera/88276d4c-0d4b-47d8-984c-0b315403d107

## About Clera

Stop applying to startups. Start getting introduced. Clera connects you directly with hiring managers at the companies you want to work for.

## About the role

##### About the Role

This is a hands-on ML Infrastructure Engineer role at an early-stage enterprise AI company building a context and data governance layer that makes AI agents reliable in production. You will own the inference and model-serving infrastructure end to end, ensuring agents run fast and reliably at increasing concurrency. The work is squarely production-focused with real-world impact across regulated industries like insurance, banking, healthcare, and asset management.

##### What You'll Do

- Design, build, and scale inference and model-serving infrastructure from the ground up through production deployment.

- Optimize systems for latency, throughput, and reliability under high concurrency.

- Collaborate closely with ML and infrastructure teams to ensure seamless integration and surface performance bottlenecks.

- Drive solutions to infrastructure challenges across a fast-moving, cross-functional team.

##### What We're Looking For

- 5 or more years building and operating machine learning inference systems, model-serving platforms, or ML infrastructure in production environments.

- Hands-on experience designing and scaling inference-serving systems using tools such as TensorFlow Serving, TorchServe, Triton, KServe, or equivalent custom solutions.

- Strong distributed systems fundamentals, including containerization and orchestration with Docker and Kubernetes.

- Proficiency with monitoring and observability tooling for production systems, such as Prometheus, Grafana, or distributed tracing frameworks.

- Experience deploying and managing ML workloads on cloud platforms (AWS, GCP, or Azure).

- Proficiency in at least one systems or backend language: Python, Go, Rust, C++, or Java.

- Comfort collaborating across both ML and infrastructure disciplines in a fast-paced environment.

- Nice to have: experience with knowledge graphs, semantic search, or graph databases; real-time or low-latency inference systems; agentic or multi-step AI pipelines; enterprise data integration or pipeline infrastructure.

##### Location

On-site in San Mateo, California, United States. Visa sponsorship is not available for this role.

---

Source: Clera's own career page, read by RemJobs. Canonical HTML version: https://www.remjobs.works/job/clera-ml-infrastructure-engineer-966a5e31-ffc2-449c-b9db-8fb4fc2b3ff8
