Data Scientist III
On-siteGurugram, Haryana, India
Job Summary
Design, train, and optimize custom machine learning algorithms and deep neural networks to solve complex business predictive problems. Build and orchestrate production-level autonomous AI Agents and Multi-Agent structures using Python frameworks. Architect high-throughput RAG systems, managing the complete pipeline across semantic chunking, multi-stage retrieval, embedding generation, reranking, and vector store integration. Write clean, modular, and performance-optimized Python code while enforcing high-standard engineering practices including unit testing, Git version control, and CI/CD alignment. Assist in porting localized AI and ML workflows into cloud environments to ensure proper resource allocation for heavy training and inference workloads.
Required Qualifications
- 4–5 years of commercial experience blending traditional data science with cutting-edge artificial intelligence
- autonomous developer who writes production-grade Python
- deep foundations in Machine Learning (ML) and Deep Learning (DL)
- hands-on experience building generative workflows like Retrieval-Augmented Generation (RAG) and autonomous AI Agents
- 3 to 4 years of professional, hands-on experience working as a Data Scientist, ML Engineer, or AI Developer in a production software environment
- Proven track record building applications using Python and foundational libraries such as scikit-learn, PyTorch, or TensorFlow
- Direct experience manipulating LLMs
- implementing vector databases (e.g., Pinecone, Qdrant, Milvus, Chroma)
- tuning systemic prompt graphs
- Proficient in writing optimized SQL
- data cleaning, and structuring pipelines utilizing pandas or NumPy
Desired Qualifications
- Experience building, testing, or serving models using cloud-native platforms like Amazon SageMaker AI
- orchestrating managed endpoints via Amazon Bedrock
Hiring someone like this?
Get your role in front of qualified candidates on Sorce.