ML Infrastructure Engineer
Clera · San Mateo
- Location
- San Mateo
- Experience
- 5+ years
- Funding
- N/A
- Posted
- Sep 1, 2026
Clera is hiring a ML Infrastructure Engineer based in San Mateo. Every apply link on Engg.space goes straight to the company's own careers page - no recruiter middleman, no generic job-board form.
Apply directly at CleraRole details
ABOUT THE ROLE This is a hands-on infrastructure engineering role at an early-stage enterprise AI company building a context layer that makes AI agents reliable, accurate, and secure for mission-critical business operations. You'll own the systems that keep those agents running fast and reliably in production — from design through deployment — working closely with ML and infrastructure teams to scale inference at increasing concurrency. WHAT YOU'LL DO - Own inference and model-serving infrastructure end to end, from architecture design through production deployment. - Build and scale systems that enable AI agents to run reliably and efficiently under high concurrency in production environments. - Collaborate with ML and infrastructure teams to ensure seamless integration and drive performance optimization. - Identify infrastructure bottlenecks and lead the engineering effort to resolve them. WHAT WE'RE LOOKING FOR - 5+ years of experience building and operating machine learning inference systems, model-serving platforms, or ML infrastructure in production environments. - Hands-on experience designing and scaling inference-serving infrastructure using tools such as TensorFlow Serving, TorchServe, Triton, KServe, or equivalent custom systems. - Demonstrated ability to optimize production ML systems for latency, throughput, and reliability at scale. - Strong proficiency with containerization and orchestration technologies — Docker and Kubernetes — for deploying ML workloads. - Experience building or maintaining distributed systems that handle concurrent requests and manage resource allocation under load. - Solid command of monitoring, observability, and debugging tooling for production systems (e.g., Prometheus, Grafana, ELK, distributed tracing). - Experience deploying and managing ML systems on cloud platforms such as AWS, GCP, or Azure. - Proficiency in at least one systems or backend language: Python, Go, Rust, C++, or Java. - Experience with knowledge graphs, semantic search, or graph databases (e.g., Neo4j, Amazon Neptune) is a plus. - Familiarity with real-time or low-latency inference systems, agentic AI pipelines, or enterprise data infrastructure is a plus. LOCATION On-site in San Mateo, California, United States. Visa sponsorship is not available for this role.
More roles at Clera
| Link | ||||||
|---|---|---|---|---|---|---|
| Data Engineer | yesterday | Berlin | N/A | N/A | Apply | |
| Founding Engineer | yesterday | $170,000 to $220,000 USD | N/A | 4 openings | ||
| Founding AI Engineer | yesterday | Amsterdam | $225,000 to $255,000 USD | N/A | Apply | |
| Founding Engineer, Infrastructure | yesterday | San Francisco | $175,000 to $300,000 USD | N/A | Apply | |
| Full-Stack Software Engineer | yesterday | Berlin | N/A | N/A | Apply | |
| Platform Engineer | yesterday | Singapore | $150,000 to $250,000 USD | N/A | Apply | |
| Cloud / DevOps Engineer | yesterday | Berlin | N/A | N/A | Apply | |
| Founding Engineer - ML Demand Generation | 2 days ago | Mountain View | $220,000 – $300,000 USD | N/A | Apply |