Senior Data Scientist at Unicornis AI · Navi Mumbai · 5 - 7 years · ₹9L - ₹15L / yr · Bootstrapped · Posted 26 Mar 2025

Note: We are looking for immediate joiners with 6+ years of experience.
Job Description
UnicornisAI is seeking a Senior Data Scientist with expertise in chatbot development using Retrieval-Augmented Generation (RAG) and OpenAI. This role is ideal for someone with a strong background in machine learning, natural language processing (NLP), and AI model deployment. If you are passionate about developing cutting-edge AI-driven solutions, we’d love to have you on our team.
Key Responsibilities
- Design and develop AI-powered chatbots using Retrieval-Augmented Generation (RAG), OpenAI models (GPT-4, etc.), and vector databases
- Build and fine-tune large language models (LLMs) to improve chatbot performance
- Implement document retrieval and knowledge management systems for chatbot responses
- Optimize NLP pipelines and model performance using state-of-the-art techniques
- Work with structured and unstructured data to enhance chatbot intelligence
- Deploy and maintain AI models in cloud environments such as AWS, Azure, or GCP
- Collaborate with engineering teams to integrate AI solutions into products
- Stay updated with the latest advancements in AI, NLP, and RAG-based architectures
Required Skills & Qualifications
- 6+ years of experience in data science, AI, or a related field
- Strong knowledge of RAG, OpenAI APIs (GPT-4, GPT-3.5, etc.), LLM fine-tuning, and embeddings
- Proficiency in Python, TensorFlow, PyTorch, and other ML frameworks
- Experience with vector databases such as FAISS, Pinecone, or Weaviate
- Expertise in NLP techniques such as Named Entity Recognition (NER), text summarization, and semantic search
- Hands-on experience in building and deploying AI models in production
- Knowledge of cloud platforms like AWS Sagemaker, Azure AI, or Google Vertex AI
- Strong problem-solving and analytical skills
Nice-to-Have Skills
- Experience with MLOps tools for model monitoring and retraining
- Understanding of prompt engineering and LLM chaining techniques
- Exposure to LangChain or similar frameworks for RAG-based chatbots
Location & Work Mode
- Open to remote or hybrid work, based on location
Interested candidates can email their resumes to Sachin at unicornisai.com

Similar jobs (10)
🚨 Hiring – Data Scientist | Python + Agentic AI
💼 Experience: 5+ Years
Must Have:
• Strong Data Science experience
• Python
• Agentic AI / AI Agents
• Generative AI / LLMs
• RAG / Vector Databases
• LangChain / LangGraph or similar Agent Frameworks
• Machine Learning & NLP
Roles & Responsibilities
- Design and develop intelligent AI-based applications using advanced NLP and LLM techniques to solve real-world business challenges in financial services.
- Build and optimize Retrieval-Augmented Generation (RAG) pipelines leveraging structured and unstructured financial data.
- Integrate and orchestrate LLMs/SLMs for question-answering, summarization, semantic search, and document understanding.
- Develop and maintain RESTful APIs (sync and async) to serve NLP models and chatbot interfaces using frameworks like FastAPI, Flask, etc.
- Should have knowledge of advanced prompting techniques.
- Implement semantic search, hybrid search, and text retrieval systems using Elasticsearch and vector databases (e.g., FAISS, Pinecone, Weaviate).
- Perform NLP tasks such as entity recognition, text classification, intent detection, embedding generation, and sentiment analysis where required.
- Monitor and fine-tune LLM/SLM performance with real-world user data to improve relevance, latency, and accuracy.
- Exposure to LLMOps tools for monitoring, evaluation, and versioning of AI models in production.
- Build, train, and evaluate deep learning models for NLP tasks including classification, NER, summarization, and embedding generation.
- Develop traditional machine learning models (e.g., regression, decision trees, clustering) for structured data analysis and prediction tasks.
- Interact with cross-functional teams to understand system issues and follow up with respective teams to get them fixed.
- Understand and identify areas of improvement across businesses and participate in solution identification and implementation.
- Should be able to work as an Individual Contributor on new and existing projects.
- Positive and problem-solving attitude, must work as an independent contributor.
Ideal Candidate
1.Strong Data Scientist / AI Engineer / Generative AI Engineer profile.
2.Mandatory (Experience 1) - Must have 3+ years of hands-on experience in Data Science, Artificial Intelligence, Machine Learning, Deep Learning, NLP, or Generative AI application development.
3.Mandatory (Experience 2) - Must have strong hands-on experience in Python programming, backend development, API development, and production-grade application support
4.Mandatory (Experience 3) - Must have experience working with Machine Learning and Deep Learning frameworks such as PyTorch, TensorFlow, Keras, or Scikit-learn.
5.Mandatory (Experience 4) - Must have hands-on experience in NLP use cases such as text classification, sentiment analysis, entity recognition (NER), semantic search, embeddings, or document understanding.
6.Mandatory (Experience 5) - Must have experience working with Large Language Models (LLMs) such as GPT, LLaMA, Mistral, Phi, Claude, Gemini, or similar models
.7.Mandatory (Experience 6) - Must have hands-on experience building or implementing Retrieval Augmented Generation (RAG) solutions, vector search, semantic search, or knowledge-based AI applications.
8.Mandatory (Experience 7) - Must have experience with Prompt Engineering and Generative AI frameworks such as LangChain, LangGraph, AI Agents, Azure OpenAI, or similar technologies.
9.Mandatory (Experience 8) - Must have experience developing, consuming, or integrating APIs using Python frameworks such as FastAPI, Flask, or similar technologies.
10.Mandatory (CTC) - The CTC breakup offered will be 75% fixed + 25% variable, as per company policy.
11.Preferred (Experience 1) - Experience with LLMOps/MLOps tools for monitoring, evaluation, experimentation, and versioning of AI models.
12.Preferred (Experience 2) - Exposure to Azure OpenAI, Azure Kubernetes Service (AKS), Kubernetes, cloud-native AI deployments, or distributed systems.
13.Preferred (Experience 3) - Experience working with PostgreSQL, MongoDB, Redis, Kafka, or large-scale data platforms.
14.Preferred (Experience 4) - Familiarity with Docker, Kubernetes, cloud platforms, and scalable deployment architecture.
15.Preferred (Company) - Candidates from AI-first startups, product companies, SaaS organizations, fintech, or data-driven technology companies.
16.Mandatory ( Age ) - Candidate Should be Below 28 Years.
17.Mandatory ( Pedigree) - B.TECH / M.TECH from Tier 1 Colleges (IIT's, NIT's, BITS) are Considered.
Strong Data Scientist / AI Engineer / Generative AI Engineer profile.
2
Mandatory (Experience 1) - Must have 3+ years of hands-on experience in Data Science, Artificial Intelligence, Machine Learning, Deep Learning, NLP, or Generative AI application development.
3
Mandatory (Experience 2) - Must have strong hands-on experience in Python programming, backend development, API development, and production-grade application support.
4
Mandatory (Experience 3) - Must have experience working with Machine Learning and Deep Learning frameworks such as PyTorch, TensorFlow, Keras, or Scikit-learn.
5
Mandatory (Experience 4) - Must have hands-on experience in NLP use cases such as text classification, sentiment analysis, entity recognition (NER), semantic search, embeddings, or document understanding.
6
Mandatory (Experience 5) - Must have experience working with Large Language Models (LLMs) such as GPT, LLaMA, Mistral, Phi, Claude, Gemini, or similar models.
7
Mandatory (Experience 6) - Must have hands-on experience building or implementing Retrieval Augmented Generation (RAG) solutions, vector search, semantic search, or knowledge-based AI applications.
8
Mandatory (Experience 7) - Must have experience with Prompt Engineering and Generative AI frameworks such as LangChain, LangGraph, AI Agents, Azure OpenAI, or similar technologies.
9
Mandatory (Experience 8) - Must have experience developing, consuming, or integrating APIs using Python frameworks such as FastAPI, Flask, or similar technologies.
10
Mandatory (CTC) - The CTC breakup offered will be 75% fixed + 25% variable, as per company policy.
11
Preferred (Experience 1) - Experience with LLMOps/MLOps tools for monitoring, evaluation, experimentation, and versioning of AI models.
12
Preferred (Experience 2) - Exposure to Azure OpenAI, Azure Kubernetes Service (AKS), Kubernetes, cloud-native AI deployments, or distributed systems.
The recruiter has not been active on this job recently. You may apply but please expect a delayed response.
Strong Data Scientist / AI Engineer / Generative AI Engineer profile.
2
Mandatory (Experience 1) - Must have 3+ years of hands-on experience in Data Science, Artificial Intelligence, Machine Learning, Deep Learning, NLP, or Generative AI application development.
3
Mandatory (Experience 2) - Must have strong hands-on experience in Python programming, backend development, API development, and production-grade application support.
4
Mandatory (Experience 3) - Must have experience working with Machine Learning and Deep Learning frameworks such as PyTorch, TensorFlow, Keras, or Scikit-learn.
5
Mandatory (Experience 4) - Must have hands-on experience in NLP use cases such as text classification, sentiment analysis, entity recognition (NER), semantic search, embeddings, or document understanding.
6
Mandatory (Experience 5) - Must have experience working with Large Language Models (LLMs) such as GPT, LLaMA, Mistral, Phi, Claude, Gemini, or similar models.
7
Mandatory (Experience 6) - Must have hands-on experience building or implementing Retrieval Augmented Generation (RAG) solutions, vector search, semantic search, or knowledge-based AI applications.
8
Mandatory (Experience 7) - Must have experience with Prompt Engineering and Generative AI frameworks such as LangChain, LangGraph, AI Agents, Azure OpenAI, or similar technologies.
9
Mandatory (Experience 8) - Must have experience developing, consuming, or integrating APIs using Python frameworks such as FastAPI, Flask, or similar technologies.
10
Mandatory (CTC) - The CTC breakup offered will be 75% fixed + 25% variable, as per company policy.
11
Preferred (Experience 1) - Experience with LLMOps/MLOps tools for monitoring, evaluation, experimentation, and versioning of AI models.
12
Preferred (Experience 2) - Exposure to Azure OpenAI, Azure Kubernetes Service (AKS), Kubernetes, cloud-native AI deployments, or distributed systems.
13
Preferred (Experience 3) - Experience working with PostgreSQL, MongoDB, Redis, Kafka, or large-scale data platforms.
14
Preferred (Experience 4) - Familiarity with Docker, Kubernetes, cloud platforms, and scalable deployment architecture.
15
Preferred (Company) - Candidates from AI-first startups, product companies, SaaS organizations, fintech, or data-driven technology companies.
16
Mandatory ( Age ) - Candidate Should be Below 28 Years.
17
Mandatory ( Pedigree) - B.TECH / M.TECH from Tier 1 Colleges (IIT's, NIT's, BITS) are Considered.
The recruiter has not been active on this job recently. You may apply but please expect a delayed response.
Strong Data Scientist / AI Engineer / Generative AI Engineer profile.
2
Mandatory (Experience 1) - Must have 3+ years of hands-on experience in Data Science, Artificial Intelligence, Machine Learning, Deep Learning, NLP, or Generative AI application development.
3
Mandatory (Experience 2) - Must have strong hands-on experience in Python programming, backend development, API development, and production-grade application support.
4
Mandatory (Experience 3) - Must have experience working with Machine Learning and Deep Learning frameworks such as PyTorch, TensorFlow, Keras, or Scikit-learn.
5
Mandatory (Experience 4) - Must have hands-on experience in NLP use cases such as text classification, sentiment analysis, entity recognition (NER), semantic search, embeddings, or document understanding.
6
Mandatory (Experience 5) - Must have experience working with Large Language Models (LLMs) such as GPT, LLaMA, Mistral, Phi, Claude, Gemini, or similar models.
7
Mandatory (Experience 6) - Must have hands-on experience building or implementing Retrieval Augmented Generation (RAG) solutions, vector search, semantic search, or knowledge-based AI applications.
8
Mandatory (Experience 7) - Must have experience with Prompt Engineering and Generative AI frameworks such as LangChain, LangGraph, AI Agents, Azure OpenAI, or similar technologies.
9
Mandatory (Experience 8) - Must have experience developing, consuming, or integrating APIs using Python frameworks such as FastAPI, Flask, or similar technologies.
10
Mandatory (CTC) - The CTC breakup offered will be 75% fixed + 25% variable, as per company policy.
11
Preferred (Experience 1) - Experience with LLMOps/MLOps tools for monitoring, evaluation, experimentation, and versioning of AI models.
12
Preferred (Experience 2) - Exposure to Azure OpenAI, Azure Kubernetes Service (AKS), Kubernetes, cloud-native AI deployments, or distributed systems.
13
Preferred (Experience 3) - Experience working with PostgreSQL, MongoDB, Redis, Kafka, or large-scale data platforms.
14
Preferred (Experience 4) - Familiarity with Docker, Kubernetes, cloud platforms, and scalable deployment architecture.
15
Preferred (Company) - Candidates from AI-first startups, product companies, SaaS organizations, fintech, or data-driven technology companies.
16
Mandatory ( Age ) - Candidate Should be Below 28 Years.
17
Mandatory ( Pedigree) - B.TECH / M.TECH from Tier 1 Colleges (IIT's, NIT's, BITS) are Considered.
Strong AI Engineer / Machine Learning Engineer profiles.
2
Mandatory (Experience 1) – Must have minimum 3+ years of hands-on experience in Data Science, Machine Learning, Applied AI, NLP, Deep Learning, or Generative AI solutions.
3
Mandatory (Experience 2) – Must have strong hands-on experience in Python programming, SQL, data analysis, feature engineering, model development, and production-grade ML applications.
4
Mandatory (Experience 3) – Must have experience working with Machine Learning and Deep Learning frameworks such as PyTorch, TensorFlow, Keras, Scikit-learn, or equivalent.
5
Mandatory (Experience 4) – Must have hands-on experience working on NLP, embeddings, semantic search, text classification, document understanding, recommendation systems, or similar AI/ML use cases.
6
Mandatory (Experience 5) – Must have experience working with Large Language Models (LLMs) such as GPT, Llama, Mistral, Claude, Gemini, Phi, or similar foundation models.
7
Mandatory (Experience 6) – Must have hands-on experience building or implementing RAG (Retrieval Augmented Generation) systems, vector search, knowledge retrieval, embeddings, chunking, indexing, or semantic retrieval solutions.
8
Mandatory (Experience 7) – Must have experience working with Git, CI/CD practices, production environments, and scalable AI/ML systems.
9
Mandatory (CTC) – The CTC breakup offered will be 75% fixed + 25% variable, as per company policy.
10
Mandatory (Age) - Candidate's Age should be below 28 Years
The recruiter has not been active on this job recently. You may apply but please expect a delayed response.
Strong AI Engineer / Machine Learning Engineer profiles.
2
Mandatory (Experience 1) – Must have minimum 5+ years of hands-on experience in Data Science, Machine Learning, Applied AI, NLP, Deep Learning, or Generative AI solutions.
3
Mandatory (Experience 2) – Must have strong hands-on experience in Python programming, SQL, data analysis, feature engineering, model development, and production-grade ML applications.
4
Mandatory (Experience 3) – Must have experience working with Machine Learning and Deep Learning frameworks such as PyTorch, TensorFlow, Keras, Scikit-learn, or equivalent.
5
Mandatory (Experience 4) – Must have hands-on experience working on NLP, embeddings, semantic search, text classification, document understanding, recommendation systems, or similar AI/ML use cases.
6
Mandatory (Experience 5) – Must have experience working with Large Language Models (LLMs) such as GPT, Llama, Mistral, Claude, Gemini, Phi, or similar foundation models.
7
Mandatory (Experience 6) – Must have hands-on experience building or implementing RAG (Retrieval Augmented Generation) systems, vector search, knowledge retrieval, embeddings, chunking, indexing, or semantic retrieval solutions.
8
Mandatory (Experience 7) – Must have experience working with Git, CI/CD practices, production environments, and scalable AI/ML systems.
9
Mandatory (CTC) – The CTC breakup offered will be 75% fixed + 25% variable, as per company policy.
10
Mandatory (Age) - Candidate's Age should be below 30 Years
11
Preferred (Experience 1) – Experience with MLFlow, Kubeflow, Airflow, Prefect, Feature Stores, Model Registry, or MLOps/LLMOps frameworks.
12
Preferred (Experience 2) – Experience working with Vector Databases, Spark, PySpark, distributed ML pipelines, large-scale data processing, or real-time ML systems..
13
Preferred (Experience 3) – Familiarity with Docker, Kubernetes, Azure, AWS, GCP, cloud-native AI deployments, and scalable ML architecture.
14
Preferred (Company) – Candidates from AI-first startups, Fintech, Banking, Lending, Fraud Analytics, Risk Analytics, Product Companies, SaaS organizations, or data-driven technology companies
15
Mandatory ( Pedigree) - B.TECH / M.TECH from Tier 1 Colleges (IIT's, NIT's, BITS) are Considered.
About the Role:
We are looking for an ideal candidate with 5+ years of experience in Data Science / Machine Learning, with strong hands-on experience in Generative AI, Large Language Models (LLMs), NLP, and AI-powered applications. The candidate should be comfortable working across the complete AI lifecycle—from understanding business requirements and experimenting with models to building, evaluating, deploying, and monitoring production-grade GenAI solutions.
The role requires a combination of strong technical expertise, business understanding, problem-solving ability, and stakeholder management skills.
Key Responsibilities:
Generative AI & LLM
· Design, develop, and deploy Generative AI and LLM-based solutions for enterprise use cases.
· Work with models such as OpenAI, Azure OpenAI, Llama, Mistral, Gemini, or equivalent LLM platforms.
· Develop applications using prompt engineering, structured outputs, function/tool calling, and LLM orchestration.
· Design and implement Retrieval-Augmented Generation (RAG) solutions.
· Work with vector databases and semantic search for enterprise knowledge retrieval.
· Develop and evaluate AI agents and multi-step AI workflows.
· Implement techniques such as prompt optimization, context management, grounding, and hallucination reduction.
· Develop AI solutions for text classification, summarization, information extraction, question answering, document intelligence, and other enterprise use cases.
Machine Learning & Data Science
· Develop and optimize traditional Machine Learning and statistical models where appropriate.
· Perform data exploration, feature engineering, model selection, training, validation, and evaluation.
· Apply appropriate ML and statistical techniques to solve business problems.
· Work with structured, unstructured, and semi-structured data.
· Develop scalable data pipelines to support AI/ML solutions.
· Collaborate with Data Engineers to prepare and manage data for AI applications.
AI Evaluation & Productionization
· Design evaluation frameworks to measure LLM accuracy, relevance, groundedness, toxicity, latency, and cost.
· Implement guardrails and responsible AI practices.
· Monitor model and application performance in production.
· Identify model/data drift and implement appropriate improvement strategies.
· Optimize AI solutions for performance, scalability, reliability, and cost.
· Support deployment and productionization of AI/ML solutions.
· Client & Delivery Responsibilities
· Work closely with the CEO, Delivery team, Solution Architects, Engineering teams, and clients to understand business problems and identify AI opportunities.
· Translate business requirements into practical AI/ML solutions.
· Participate in client discussions, solution presentations, technical workshops, and POCs.
· Develop rapid prototypes and demonstrate the feasibility of GenAI solutions.
· Convert successful POCs into scalable, production-ready applications.
· Provide technical guidance and contribute to AI solution architecture.
· Prepare technical documentation, solution approaches, and project estimates where required.
· Stay current with developments in Generative AI, LLMs, Agentic AI, and AI engineering.
Required Skills:
· 5+ years of hands-on experience in Data Science, Machine Learning, AI, or a related field.
· Strong practical experience in Generative AI and LLM-based applications.
· Strong proficiency in Python.
· Strong understanding of Machine Learning and statistical concepts.
· Hands-on experience with:
o LLMs
o Prompt Engineering
o RAG
o Vector Databases
o Embeddings
o Semantic Search
o LLM Evaluation
o AI Guardrails
· Experience with frameworks/tools such as LangChain, LangGraph, LlamaIndex, or equivalent.
· Experience with APIs and integrating LLMs into enterprise applications.
· Strong SQL and data handling skills.
· Experience working with large and complex datasets.
· Strong understanding of NLP concepts.XX
Technical Skills:
· Experience with OpenAI / Azure OpenAI / AWS Bedrock / Google Vertex AI.
· Experience with vector databases such as Pinecone, Weaviate, Milvus, FAISS, or equivalent.
· Experience with Databricks, Snowflake, or cloud data platforms.
· Experience with Docker and CI/CD.
· Exposure to AWS, Azure, or GCP.
· Experience with ML/AI deployment and MLOps.
· Knowledge of AI security, data privacy, governance, and responsible AI.
· Experience building AI Agents / Agentic AI workflows.
· Experience with multimodal AI is an added advantage
Key Competencies
· Strong analytical and problem-solving ability.
· Ability to translate business problems into practical AI solutions.
· Strong communication and presentation skills.
· Ability to interact confidently with senior stakeholders and clients.
· Strong ownership and delivery mindset.
· Ability to work independently in a fast-paced environment.
- Strong experimentation and innovation mindset.
- Ability to balance technical feasibility, business value, scalability, and cost.
Required Education & Experience:
· Bachelor's or Master's degree in Computer Science, Data Science, Artificial Intelligence, Statistics, Mathematics, Engineering, or a related discipline
The recruiter has not been active on this job recently. You may apply but please expect a delayed response.
Strong AI Engineer / Machine Learning Engineer profiles.
2
Mandatory (Experience 1) – Must have minimum 3+ years of hands-on experience in Data Science, Machine Learning, Applied AI, NLP, Deep Learning, or Generative AI solutions.
3
Mandatory (Experience 2) – Must have strong hands-on experience in Python programming, SQL, data analysis, feature engineering, model development, and production-grade ML applications.
4
Mandatory (Experience 3) – Must have experience working with Machine Learning and Deep Learning frameworks such as PyTorch, TensorFlow, Keras, Scikit-learn, or equivalent.
5
Mandatory (Experience 4) – Must have hands-on experience working on NLP, embeddings, semantic search, text classification, document understanding, recommendation systems, or similar AI/ML use cases.
6
Mandatory (Experience 5) – Must have experience working with Large Language Models (LLMs) such as GPT, Llama, Mistral, Claude, Gemini, Phi, or similar foundation models.
7
Mandatory (Experience 6) – Must have hands-on experience building or implementing RAG (Retrieval Augmented Generation) systems, vector search, knowledge retrieval, embeddings, chunking, indexing, or semantic retrieval solutions.
8
Mandatory (Experience 7) – Must have experience working with Git, CI/CD practices, production environments, and scalable AI/ML systems.
9
Mandatory (CTC) – The CTC breakup offered will be 75% fixed + 25% variable, as per company policy.
10
Mandatory (Age) - Candidate's Age should be below 30 Years
11
Preferred (Experience 1) – Experience with MLFlow, Kubeflow, Airflow, Prefect, Feature Stores, Model Registry, or MLOps/LLMOps frameworks.
12
Preferred (Experience 2) – Experience working with Vector Databases, Spark, PySpark, distributed ML pipelines, large-scale data processing, or real-time ML systems..
13
Preferred (Experience 3) – Familiarity with Docker, Kubernetes, Azure, AWS, GCP, cloud-native AI deployments, and scalable ML architecture.
14
Preferred (Company) – Candidates from AI-first startups, Fintech, Banking, Lending, Fraud Analytics, Risk Analytics, Product Companies, SaaS organizations, or data-driven technology companies
15
Mandatory ( Pedigree) - B.TECH / M.TECH from Tier 1 Colleges (IIT's, NIT's, BITS) are Considered.
Support with design and build to prove out agentic AI solution flow by working with other data
scientists and engineers to build, train Large Language Model (LLM) architectures, RAG
systems, and autonomous agentic workflows
Key qualifications:
>> AI solution design & Development: Design Agentic AI solutions using RAG (Retrieval-
Augmented Generation) and orchestration frameworks like LangGraph or LangChain.
>> Model Fine-Tuning: Solid understanding and experience with Pre-train, fine-tune, and
optimize open-source like BERT, LLama, and other proprietary foundation models for domain-
specific tasks
>> Solid Stats and ML foundations and (vibe) coding skills with Python, PySpark
>> Implement validation frameworks and tracing practices (using tools like Arize) to monitor
agent behavior, guard against model drift, and ensure compliance
>> Collaborate with Engineering to deploy models securely on cloud and on-prem ecosystems






