Model Serving jobs

5 open roles for model serving, updated continuously as employers post them.

Applied AI Engineer, Inference

CoreWeave · Bellevue, United States · San Francisco, United States · Sunnyvale, United States

Applied AI engineering role in Bellevue, San Francisco or Sunnyvale for an engineer with 4 or more years of systems, machine-learning or performance experience. The work improves inference latency, throughput, quality and cost for production workloads.

On-siteMid levelFull-time $188,000 – $275,000 a year
PythonLLM InferencevLLMSGLang +8
4 days ago

MLOps engineer

standard chartered india · Bengaluru, India · Bengaluru / Bangalore, Karnataka, India

An MLOps engineering role at Standard Chartered in Bengaluru, building and operating ML/GenAI platforms. Suits an engineer experienced with Python, Kubernetes, CI/CD and LangChain/LangGraph.

On-siteFull-time
PythonKubernetesCI/CDLangChain +6
5 days ago

Member of Technical Staff (Software Engineer, Infrastructure)

Perplexity · San Francisco, United States · London, United Kingdom · Austin, United States · Berlin, Germany · Seattle, United States · New York City, United States

Infrastructure software engineering role for an experienced engineer who can solve distributed systems problems across Perplexity’s platform. The position offers U.S. remote work and multiple office locations, with compensation of $220,000 to $405,000 plus equity for U.S.-based employees.

RemotePrincipalFull-time $220,000 – $405,000 a year
PythonGoRustC++ +11
5 days ago

MLOps engineer

Standard Chartered · Bengaluru, India · Bangalore, India

MLOps Engineer role in Bengaluru building and operating production ML and GenAI platforms. It suits engineers experienced in Python, cloud, Kubernetes, CI/CD, observability and governed model delivery.

On-siteFull-time
MLOpsPythonCloud ComputingKubernetes +11
8 days ago

Machine Learning Engineer

Nykaa · Bengaluru, India

Nykaa is hiring a mid-level machine learning engineer in Bengaluru to productionize ML systems used in e-commerce. The role requires 5 to 7 years of experience along with Python, MLOps, model serving, monitoring, and AWS or cloud skills.

HybridMid level
Machine LearningMLOpsPythonAWS +3
Date not stated

Browse by work mode