LLM Engineer
Skills
About the role
This is a full-time LLM engineering position for a software engineer with roughly three to seven years of experience, based in Bangalore. You will build enterprise AI features, improve inference performance, and create systems for document understanding, extraction, and semantic search. APPIT Software Solutions presents the role as part of its AI, cloud computing, and digital transformation work.
What you’ll do
- Build LLM-powered capabilities for enterprise applications.
- Improve inference speed using quantization, caching, and batching.
- Create prompt templates and reasoning workflows.
- Validate and parse structured model responses.
- Compare LLM providers across cost, quality, and latency.
- Develop extraction and document-understanding pipelines.
What they’re looking for
- Three to five years of software engineering experience with solid machine learning fundamentals.
- Hands-on experience with LLM APIs or open-source models.
- Understanding of transformers and tokenization.
- Experience with vLLM, TGI, or Ollama.
- Python and asynchronous programming experience.
- Knowledge of embeddings and semantic search.
Nice to have
- Experience with Llama, Mistral, Qwen, or other open-source LLMs.
- Model distillation or compression knowledge.
- Familiarity with HELM or lm-eval.
What’s on offer
- Full-time employment in Bangalore.
- The company describes flexible work arrangements and career growth opportunities.
- The posting does not state salary, visa sponsorship, or relocation support.
Questions about this role
Where is this LLM Engineer role based?
The posting lists Bangalore.
What experience range is listed?
The job page lists three to seven years, while its requirements describe three to five years of software engineering experience.
Which LLM serving tools are mentioned?
The posting names vLLM, TGI, and Ollama.
Does the role involve semantic search?
Yes. Knowledge of embeddings and semantic search is required.
Related roles
Applied AI Engineer (LLM Systems)
Nference is hiring a mid-level Applied AI Engineer in Bengaluru to build and evaluate retrieval-augmented clinical reasoning systems. The role is suited to engineers working with Python, RAG, LLM evaluation, and PyTorch.
Solutions Engineer AI Engineer
Remote solutions engineering role in India for building conversational voice and video AI applications in life sciences. The role suits engineers with at least three years of Python, API, LLM, or AI solutions experience.
Engineer Engine Design & Application
Automotive engine design and application engineer role in Chennai covering powertrain calibration, validation, emissions, and release targets. It suits an engineer with two to nine years of relevant experience and hands-on MATLAB, Simulink, CAN, UDS, and engine-testing skills.
Engineer - Engine Reliability
Onsite engine-reliability engineering role in Chennai for an automotive product-development professional with two to nine years of relevant experience. The work covers powertrain calibration, testing, diagnostics, reliability, emissions, and vehicle integration.
Engineer - Engine Testing
Automotive engine-testing engineer role in Chennai for professionals who can plan and execute validation, analyse failures and support issue closure. The position suits candidates with hands-on experience in calibration, NVH, durability, instrumentation and test planning.
LLM Engineer/4-8 Years/ Hyderabad.
Full-time LLM Engineer role in Hyderabad for an experienced machine learning professional. The work covers model tuning, production deployment, intelligent agents, and generative AI research.