VLLM jobs
7 open roles for vllm, updated continuously as employers post them.
Staff AI Engineer
An onsite Staff AI Engineer role at Gnani.ai in Bengaluru, leading the agentic AI and RAG stack behind a voice-first enterprise platform.
Principal Engineer - Perf and Benchmarking
Principal engineer role in Sunnyvale or Bellevue leading performance engineering and benchmarking for CoreWeave's AI cloud infrastructure. The position is for a senior technical leader with ten-plus years of distributed systems or HPC experience and deep GPU, Kubernetes, and ML-platform expertise.
Applied AI Engineer, Inference
Applied AI engineering role in Bellevue, San Francisco or Sunnyvale for an engineer with 4 or more years of systems, machine-learning or performance experience. The work improves inference latency, throughput, quality and cost for production workloads.
Account Solutions Architect - Greenfield
Account Solutions Architect role for an AI-focused solutions engineer helping greenfield customers build and scale workloads on CoreWeave. The position combines technical pre-sales, hands-on demonstrations, production AI architecture, and customer relationship ownership.
Account Solutions Architect - Engaged
Account Solutions Architect role in Seattle supporting existing customers as they deploy and scale AI workloads on CoreWeave. The role suits a technically strong customer partner with Python, deep-learning, LLM, and cloud experience.
Account Solution Architect
San Francisco account solution architect role supporting existing customers running production AI workloads on CoreWeave. It suits AI-focused customer engineers who combine Python and deep-learning experience with enterprise advisory and presentation skills.
LLM Engineer
LLM Engineer role in Bangalore focused on enterprise AI applications, inference optimization, and document intelligence. It suits software engineers with Python, transformer, LLM API, and semantic-search experience.