Skip to main content
Zobhira
Home
Jobs
Certifications
Zobhira
JobsCertificationsTodayAbout
Log inSign up
Zobhira

New job and contest openings, updated every morning on one searchable board.

Find work

  • All jobs
  • Fresher roles
  • Remote roles
  • Certifications

Compete

  • Added today

Popular cities

  • India
  • Bangalore, Karnataka, India
  • Hyderabad, Telangana, India
  • Pune, Maharashtra, India
  • Chennai, Tamil Nadu, India
  • Mumbai, Maharashtra, India

Company

  • About
  • Contact
  • Privacy
  • Terms

Stay updated

One email a week with new roles.

Secure infrastructure
Free to use, no account needed to search

Board updated daily · © 2026 Zobhira. All rights reserved.

Privacy PolicyTerms of Service
Home / Jobs / Clera

ML Infrastructure Engineer

Clera

WorldwideremotePosted 17 days ago
Clera logo

Skill Required

ML-Infrastructure-EngineerML-AI-Infrastructure-EngineerAI-ML-Infrastructure-EngineerMachine-Learning-Infrastructure-EngineerAI-Infrastructure-EngineerDeep-Learning-Infrastructure-EngineerML-Infrastructure-EngineeringAI-ML-Infrastructure-EngineeringMachine Learning EngineerInfrastructure EngineerMachine LearningSystem DesignDockerObservabilityKubernetesEngineeringTensorFlowPrometheusdesigningsimilar)debuggingbuildingAzurePythondesignC++Neo4jCloudJavaRustAWSGCPandElasticsearchGoFulltime

Key highlights

  • 5+ years of hands-on experience building and operating machine learning inference systems required
  • Must have experience with containerization and orchestration (Docker, Kubernetes)
  • Full-time, on-site role based in San Mateo, CA
  • Visa sponsorship is not available

Role overview

We're a seed-stage enterprise AI infrastructure company building the context layer that makes AI agents reliable, accurate, and secure for critical business operations — including highly regulated industries like insurance, banking, asset management, and healthcare. Our platform automatically constructs a governed, real-time domain model across all enterprise data, enabling AI agents to make confident, auditable decisions in production. As an ML Infrastructure Engineer, you'll own the systems that keep our agents running reliably and fast at scale. This is a hands-on production engineering role — focused on real-world impact, not research. You'll design, build, and scale our inference and model-serving infrastructure as concurrency and customer demands grow.

Responsibilities

  • Own inference and model-serving infrastructure end to end — from initial design through production deployment and ongoing scaling.
  • Build and scale systems that enable AI agents to run reliably and efficiently under high and increasing concurrency.
  • Identify and resolve infrastructure bottlenecks in collaboration with ML and platform engineering teams.
  • Optimize systems for latency, throughput, and reliability across cloud-hosted production environments.
  • Drive observability, monitoring, and debugging practices across our production ML stack.

Requirements

  • 5+ years of hands-on experience building and operating machine learning inference systems, model-serving platforms, or ML infrastructure in production environments.
  • Demonstrated experience designing and scaling inference-serving infrastructure using tools such as TensorFlow Serving, TorchServe, Triton, KServe, or equivalent custom systems.
  • Proven ability to optimize production ML systems for latency, throughput, and reliability at scale.
  • Experience with containerization and orchestration (Docker, Kubernetes) for deploying and scaling ML workloads.
  • Background in distributed systems that handle high concurrency and dynamic resource allocation under load.
  • Proficiency with monitoring and observability tooling — e.g., Prometheus, Grafana, ELK stack, distributed tracing.
  • Experience deploying and managing ML systems on cloud platforms (AWS, GCP, or Azure).
  • Proficiency in at least one systems or backend language: Python, Go, Rust, C++, or Java.

Nice to have

  • Experience with knowledge graphs, semantic search, or graph databases (e.g., Neo4j, Amazon Neptune, or similar).
  • Background in real-time inference or low-latency serving requirements.
  • Familiarity with agentic AI systems, autonomous agents, or multi-step reasoning pipelines.
  • Experience with enterprise data infrastructure, data pipelines, or data integration platforms.

Additional details

  • This is a full-time, on-site role based in San Mateo, CA.
  • On-site collaboration is an important part of how this small, fast-moving team operates.
  • Visa sponsorship is not available for this position.

Similar jobs open now

ABB logo
Senior Project Engineer -SIS
ABB
Bangalore, Karnataka, India
View details
Cognyte logo
Product Manager
Cognyte
Pune, Maharashtra, India
View details
IBM logo
Data Engineer-Data Platforms-AWS
IBM
Pune, Maharashtra, India
View details
IBM logo
Technical Consultant-AI Integration
IBM
Pune, Maharashtra, India
View details
Apply now
LocationWorldwide
TypeFulltime
Posted8/21/2026
Apply by10/20/2026

Links are checked every day. If this one stops working, tell us and we'll pull it.

More like this

View all
ABB logo
Senior Project Engineer -SIS
ABB
Cognyte logo
Product Manager
Cognyte
IBM logo
Data Engineer-Data Platforms-AWS
IBM
Apply now