Data Scientist
Skills
About the role
This is a full-time Data Scientist role for someone building production AI systems in Chennai, Bengaluru, Pune, Noida, or Coimbatore, India. You'll work on LLM applications, retrieval systems, agentic workflows, and the services needed to deploy them at scale. The role supports healthcare, financial services, and retail projects and includes collaboration with product, UX, data science, and operations teams.
What you’ll do
- Build RAG pipelines using LLMs, vector stores, and knowledge graphs.
- Create multi-step agents and decision workflows with LangGraph or LangChain.
- Write tested Python for model serving, data pipelines, and microservices.
- Connect Java components through gRPC or REST.
- Deploy containerized services with Docker and Kubernetes.
- Maintain GitLab CI/CD pipelines covering tests, security checks, and blue-green releases.
- Improve indexing, graph traversal, and retrieval latency.
- Work with cross-functional teams to turn requirements into technical systems.
- Mentor junior engineers through reviews and technical presentations.
What they’re looking for
- At least five years of professional Python development, including asynchronous programming, type hints, and testing.
- Strong knowledge of data structures and algorithms.
- Practical expertise with LLMs, tokenization, fine-tuning, prompting, and evaluation.
- Experience delivering RAG systems with embeddings, similarity search, vector databases, and knowledge graphs.
- Hands-on experience with LangGraph or LangChain.
- Strong communication and collaboration skills in Agile teams.
Nice to have
- Java microservices experience using gRPC or REST.
- Docker, Kubernetes, and Helm experience.
- GitLab CI/CD, cloud platforms, and MLOps experience.
- Prometheus, Grafana, and container-security experience.
- Open-source contributions to AI tooling.
What’s on offer
- Full-time employment across Chennai, Bengaluru, Pune, Noida, or Coimbatore.
- Exposure to AI, data analytics, cloud, IoT, and telehealth work.
- The official requisition requests four to seven years of experience.
Questions about this role
Where is the Data Scientist role based?
The role is open in Chennai, Bengaluru, Pune, Noida, and Coimbatore, India.
How much experience is requested?
The requisition requests four to seven years of experience, while the requirements also call for at least five years of professional Python development.
What technologies are central to the role?
Python, LLMs, RAG, LangGraph, LangChain, vector databases, knowledge graphs, Docker, Kubernetes, and MLOps.
Related roles
Data Analyst - Data Quality & Databricks
A data analyst role centered on data quality, profiling and reconciliation using Databricks, SQL and PySpark, open across Bangalore, Chennai and Pune for candidates with 7-11 years of experience.
Data Scientist-Vice President-Finance Data & Analytics
A Vice President-level data science role at JPMorgan Chase in Columbus, Ohio, focused on financial analytics and AI tooling. Suits an experienced financial analytics professional skilled in SQL, cloud platforms and Databricks.
Lead Data Engineer – Databricks & AI | Remote Canada
Lead data engineer role for a financial services client, requiring 10+ years of Databricks expertise and experience collaborating with AI/ML teams to build scalable data infrastructure.
Senior Data Scientist / Lead Data Scientist
A senior/lead data scientist role at Celebal Technologies across Noida, Jaipur or Bengaluru, leading enterprise AI/ML and generative AI delivery.
Data Careers Hub - Lead Data Scientist
A senior technical leadership data science role at XPO based in Hyderabad or Pune, architecting enterprise GenAI and RAG solutions for supply chain operations. Suits an experienced data scientist ready to lead AI strategy and mentor a team.
Data Engineer, Data Platform & ML (Remote)
A Poland-remote Data Engineer role building a cloud-based corporate data warehouse for business analytics and machine learning. It suits a data engineer with at least three years of experience in Python, SQL, warehouse design, and modern data tooling.