ML Infrastructure Engineer
Clera- Location
- San Mateo
- Workplace
- —
- Employment
- Full Time
- Salary
- —
Posted 2d ago
About the Role
This is a hands-on infrastructure engineering role at an early-stage enterprise AI company building a context and data governance layer for AI agents deployed in highly regulated industries. You will own the inference and model-serving infrastructure end to end, making production AI agents fast, reliable, and scalable as concurrency grows.
What You'll Do
- Design, build, and own inference and model-serving infrastructure from initial architecture through production deployment.
- Scale systems that enable AI agents to run reliably and efficiently under increasing concurrent load.
- Identify and resolve infrastructure bottlenecks in collaboration with ML and platform engineering teams.
- Drive performance optimization across latency, throughput, and reliability for production workloads.
What We're Looking For
- 5+ years building and operating ML inference systems, model-serving platforms, or ML infrastructure in production environments.
- Hands-on experience designing and scaling inference-serving systems using frameworks such as TensorFlow Serving, TorchServe, Triton, KServe, or equivalent custom solutions.
- Strong distributed systems fundamentals, including experience managing concurrent requests and resource allocation under load.
- Proficiency with containerization and orchestration technologies, particularly Docker and Kubernetes, for ML workloads.
- Experience with cloud infrastructure platforms (AWS, GCP, or Azure) for deploying and managing ML systems.
- Solid monitoring and observability skills using tools such as Prometheus, Grafana, ELK, or distributed tracing solutions.
- Proficiency in at least one systems or backend language: Python, Go, Rust, C++, or Java.
- Familiarity with knowledge graphs, semantic search, or graph databases is a plus.
- Background in agentic or autonomous AI systems, real-time inference, or enterprise data infrastructure is a plus.
Location
On-site in San Mateo, California, United States. Visa sponsorship is not available for this role.
Skills
- Machine Learning
- TensorFlow
- Triton
- Docker
- Kubernetes
- AWS
- GCP
- Azure
- Prometheus
- Grafana
- ELK Stack
- Python
- Go
- Rust
- C++
- Java
More jobs at Clera
All 223Founding Engineer
Clera · San Francisco · yesterday
Forward Deployed Engineer
Clera · New York City · USD 150,000–230,000/yr · yesterday
Product Engineer (Software Engineer)
Clera · San Francisco · USD 120,000–200,000/yr · yesterday
Talent AI Operator
Clera · San Francisco · yesterday
Co-Founding CTO
Clera · Munich · yesterday
Similar roles
Staff AI Engineering - Enterprise Architecture
American Express · Phoenix, AZ, United States · USD 144,250–256,250/yr · today
Field Service Technician II
Toshiba America Business Solutions · East Syracuse, NY, United States · today
Field Service Technician I
Toshiba America Business Solutions · Rochester, NY, United States · today
Field Service Technician II
Toshiba America Business Solutions · Buffalo, NY, United States · today
Incentives Analyst
MicroStrategy · Tysons Corner, VIRGINIA, United States · USD 66,400–119,600/yr · today
Senior Azure Cloud Engineer
Vaxcyte · San Carlos, California, United States · today