Data scientist at NeoGenCode Technologies Pvt Ltd · Gurugram · 4 - 6 years · ₹5L - ₹14L / yr · Raised funding · Posted 5 Aug 2025

We’re searching for an experienced Data Scientist with a strong background in NLP and large language models to join our innovative team! If you thrive on solving complex language problems and are hands-on with spaCy, NER, RNN, LSTM, Transformers, and LLMs (like GPT), we want to connect.
What You’ll Do:
- Build & deploy advanced NLP solutions: entity recognition, text classification, and more.
- Fine-tune and train state-of-the-art deep learning models (RNN, LSTM, Transformer, GPT).
- Apply libraries like spaCy for NER and text processing.
- Collaborate across teams to integrate AI-driven features.
- Preprocess, annotate, and manage data workflows.
- Analyze model performance and drive continuous improvement.
- Stay current with AI/NLP breakthroughs and advocate innovation.
What You Bring:
- 4-5+ years of industry experience in data science/NLP.
- Strong proficiency in Python, spaCy, NLTK, PyTorch or TensorFlow.
- Hands-on with NER, custom pipelines, and prompt engineering.
- Deep understanding and experience with RNN, LSTM, Transformer, and LLMs/GPT.
- Collaborative and independent problem solver.
Nice to Have:
- Experience deploying NLP models (Docker, cloud).
- MLOps, vector databases, RAG, semantic search.
- Annotation tools and team management.
Why Join Us?
- Work with cutting-edge technology and real-world impact.
- Flexible hours, remote options, and a supportive, inclusive culture.
- Competitive compensation and benefits.
Ready to push the boundaries of AI with us? Apply now or DM for more info!

About NeoGenCode Technologies Pvt Ltd
About
Welcome to Neogencode Technologies, an IT services and consulting firm that provides innovative solutions to help businesses achieve their goals. Our team of experienced professionals is committed to providing tailored services to meet the specific needs of each client. Our comprehensive range of services includes software development, web design and development, mobile app development, cloud computing, cybersecurity, digital marketing, and skilled resource acquisition. We specialize in helping our clients find the right skilled resources to meet their unique business needs. At Neogencode Technologies, we prioritize communication and collaboration with our clients, striving to understand their unique challenges and provide customized solutions that exceed their expectations. We value long-term partnerships with our clients and are committed to delivering exceptional service at every stage of the engagement. Whether you are a small business looking to improve your processes or a large enterprise seeking to stay ahead of the competition, Neogencode Technologies has the expertise and experience to help you succeed. Contact us today to learn more about how we can support your business growth and provide skilled resources to meet your business needs.
Candid answers by the company
IT & Engineering Talent Staffing
- Provides full-time and contract-based hiring, delivering handpicked, pre‑screened developers across tech stacks—ranging from web, mobile, AI/ML, Web3/blockchain.
- Maintains a bench o vetted candidates, offering fast delivery of interview-ready profiles—often within 24 hours.
- Offers payroll management, handling compliance, tax, attendance, and documentation for both contractors and full-time employees.
2. End-to-End Project Delivery
- Delivers full-stack development solutions: web, mobile, cloud, AI/ML, Blockchain/Web3.
- Manages entire project lifecycle—requirements gathering, design (UI/UX), development, deployment, and ongoing support .
3. Additional Offerings
- Expands into cybersecurity consulting, digital marketing, and cloud platform services (like AWS, GCP, Azure) .
- Provides strategic IT consulting to align technology solutions with business objectives
Similar jobs (10)
Roles & Responsibilities
- Design and develop intelligent AI-based applications using advanced NLP and LLM techniques to solve real-world business challenges in financial services.
- Build and optimize Retrieval-Augmented Generation (RAG) pipelines leveraging structured and unstructured financial data.
- Integrate and orchestrate LLMs/SLMs for question-answering, summarization, semantic search, and document understanding.
- Develop and maintain RESTful APIs (sync and async) to serve NLP models and chatbot interfaces using frameworks like FastAPI, Flask, etc.
- Should have knowledge of advanced prompting techniques.
- Implement semantic search, hybrid search, and text retrieval systems using Elasticsearch and vector databases (e.g., FAISS, Pinecone, Weaviate).
- Perform NLP tasks such as entity recognition, text classification, intent detection, embedding generation, and sentiment analysis where required.
- Monitor and fine-tune LLM/SLM performance with real-world user data to improve relevance, latency, and accuracy.
- Exposure to LLMOps tools for monitoring, evaluation, and versioning of AI models in production.
- Build, train, and evaluate deep learning models for NLP tasks including classification, NER, summarization, and embedding generation.
- Develop traditional machine learning models (e.g., regression, decision trees, clustering) for structured data analysis and prediction tasks.
- Interact with cross-functional teams to understand system issues and follow up with respective teams to get them fixed.
- Understand and identify areas of improvement across businesses and participate in solution identification and implementation.
- Should be able to work as an Individual Contributor on new and existing projects.
- Positive and problem-solving attitude, must work as an independent contributor.
Ideal Candidate
1.Strong Data Scientist / AI Engineer / Generative AI Engineer profile.
2.Mandatory (Experience 1) - Must have 3+ years of hands-on experience in Data Science, Artificial Intelligence, Machine Learning, Deep Learning, NLP, or Generative AI application development.
3.Mandatory (Experience 2) - Must have strong hands-on experience in Python programming, backend development, API development, and production-grade application support
4.Mandatory (Experience 3) - Must have experience working with Machine Learning and Deep Learning frameworks such as PyTorch, TensorFlow, Keras, or Scikit-learn.
5.Mandatory (Experience 4) - Must have hands-on experience in NLP use cases such as text classification, sentiment analysis, entity recognition (NER), semantic search, embeddings, or document understanding.
6.Mandatory (Experience 5) - Must have experience working with Large Language Models (LLMs) such as GPT, LLaMA, Mistral, Phi, Claude, Gemini, or similar models
.7.Mandatory (Experience 6) - Must have hands-on experience building or implementing Retrieval Augmented Generation (RAG) solutions, vector search, semantic search, or knowledge-based AI applications.
8.Mandatory (Experience 7) - Must have experience with Prompt Engineering and Generative AI frameworks such as LangChain, LangGraph, AI Agents, Azure OpenAI, or similar technologies.
9.Mandatory (Experience 8) - Must have experience developing, consuming, or integrating APIs using Python frameworks such as FastAPI, Flask, or similar technologies.
10.Mandatory (CTC) - The CTC breakup offered will be 75% fixed + 25% variable, as per company policy.
11.Preferred (Experience 1) - Experience with LLMOps/MLOps tools for monitoring, evaluation, experimentation, and versioning of AI models.
12.Preferred (Experience 2) - Exposure to Azure OpenAI, Azure Kubernetes Service (AKS), Kubernetes, cloud-native AI deployments, or distributed systems.
13.Preferred (Experience 3) - Experience working with PostgreSQL, MongoDB, Redis, Kafka, or large-scale data platforms.
14.Preferred (Experience 4) - Familiarity with Docker, Kubernetes, cloud platforms, and scalable deployment architecture.
15.Preferred (Company) - Candidates from AI-first startups, product companies, SaaS organizations, fintech, or data-driven technology companies.
16.Mandatory ( Age ) - Candidate Should be Below 28 Years.
17.Mandatory ( Pedigree) - B.TECH / M.TECH from Tier 1 Colleges (IIT's, NIT's, BITS) are Considered.
🚨 Hiring – Data Scientist | Python + Agentic AI
💼 Experience: 5+ Years
Must Have:
• Strong Data Science experience
• Python
• Agentic AI / AI Agents
• Generative AI / LLMs
• RAG / Vector Databases
• LangChain / LangGraph or similar Agent Frameworks
• Machine Learning & NLP
Strong Data Scientist / AI Engineer / Generative AI Engineer profile.
2
Mandatory (Experience 1) - Must have 3+ years of hands-on experience in Data Science, Artificial Intelligence, Machine Learning, Deep Learning, NLP, or Generative AI application development.
3
Mandatory (Experience 2) - Must have strong hands-on experience in Python programming, backend development, API development, and production-grade application support.
4
Mandatory (Experience 3) - Must have experience working with Machine Learning and Deep Learning frameworks such as PyTorch, TensorFlow, Keras, or Scikit-learn.
5
Mandatory (Experience 4) - Must have hands-on experience in NLP use cases such as text classification, sentiment analysis, entity recognition (NER), semantic search, embeddings, or document understanding.
6
Mandatory (Experience 5) - Must have experience working with Large Language Models (LLMs) such as GPT, LLaMA, Mistral, Phi, Claude, Gemini, or similar models.
7
Mandatory (Experience 6) - Must have hands-on experience building or implementing Retrieval Augmented Generation (RAG) solutions, vector search, semantic search, or knowledge-based AI applications.
8
Mandatory (Experience 7) - Must have experience with Prompt Engineering and Generative AI frameworks such as LangChain, LangGraph, AI Agents, Azure OpenAI, or similar technologies.
9
Mandatory (Experience 8) - Must have experience developing, consuming, or integrating APIs using Python frameworks such as FastAPI, Flask, or similar technologies.
10
Mandatory (CTC) - The CTC breakup offered will be 75% fixed + 25% variable, as per company policy.
11
Preferred (Experience 1) - Experience with LLMOps/MLOps tools for monitoring, evaluation, experimentation, and versioning of AI models.
12
Preferred (Experience 2) - Exposure to Azure OpenAI, Azure Kubernetes Service (AKS), Kubernetes, cloud-native AI deployments, or distributed systems.
🚀 We’re Hiring | Data Scientist 🧠📊
Ready to turn data into real-world intelligence? Join us and work on exciting AI/ML & data-driven solutions!
🔹 Experience: 8+ Years
🔹 Must-Have Skills:
🐍 Python | 🤖 Machine Learning | ☁️ Cloud | 🧠 NLP | 📊 Data Visualization
📍 Location: Pune
💼 Work Mode: Work from Office
If you're passionate about Data Science, AI & solving complex business problems, we’d love to hear from you!
📩 Interested? Kindly text
#Hiring #DataScientist #DataScience #MachineLearning #Python #NLP #AI #Cloud #DataVisualization #TechJobs #HiringNow
About the Role:
We are looking for an ideal candidate with 5+ years of experience in Data Science / Machine Learning, with strong hands-on experience in Generative AI, Large Language Models (LLMs), NLP, and AI-powered applications. The candidate should be comfortable working across the complete AI lifecycle—from understanding business requirements and experimenting with models to building, evaluating, deploying, and monitoring production-grade GenAI solutions.
The role requires a combination of strong technical expertise, business understanding, problem-solving ability, and stakeholder management skills.
Key Responsibilities:
Generative AI & LLM
· Design, develop, and deploy Generative AI and LLM-based solutions for enterprise use cases.
· Work with models such as OpenAI, Azure OpenAI, Llama, Mistral, Gemini, or equivalent LLM platforms.
· Develop applications using prompt engineering, structured outputs, function/tool calling, and LLM orchestration.
· Design and implement Retrieval-Augmented Generation (RAG) solutions.
· Work with vector databases and semantic search for enterprise knowledge retrieval.
· Develop and evaluate AI agents and multi-step AI workflows.
· Implement techniques such as prompt optimization, context management, grounding, and hallucination reduction.
· Develop AI solutions for text classification, summarization, information extraction, question answering, document intelligence, and other enterprise use cases.
Machine Learning & Data Science
· Develop and optimize traditional Machine Learning and statistical models where appropriate.
· Perform data exploration, feature engineering, model selection, training, validation, and evaluation.
· Apply appropriate ML and statistical techniques to solve business problems.
· Work with structured, unstructured, and semi-structured data.
· Develop scalable data pipelines to support AI/ML solutions.
· Collaborate with Data Engineers to prepare and manage data for AI applications.
AI Evaluation & Productionization
· Design evaluation frameworks to measure LLM accuracy, relevance, groundedness, toxicity, latency, and cost.
· Implement guardrails and responsible AI practices.
· Monitor model and application performance in production.
· Identify model/data drift and implement appropriate improvement strategies.
· Optimize AI solutions for performance, scalability, reliability, and cost.
· Support deployment and productionization of AI/ML solutions.
· Client & Delivery Responsibilities
· Work closely with the CEO, Delivery team, Solution Architects, Engineering teams, and clients to understand business problems and identify AI opportunities.
· Translate business requirements into practical AI/ML solutions.
· Participate in client discussions, solution presentations, technical workshops, and POCs.
· Develop rapid prototypes and demonstrate the feasibility of GenAI solutions.
· Convert successful POCs into scalable, production-ready applications.
· Provide technical guidance and contribute to AI solution architecture.
· Prepare technical documentation, solution approaches, and project estimates where required.
· Stay current with developments in Generative AI, LLMs, Agentic AI, and AI engineering.
Required Skills:
· 5+ years of hands-on experience in Data Science, Machine Learning, AI, or a related field.
· Strong practical experience in Generative AI and LLM-based applications.
· Strong proficiency in Python.
· Strong understanding of Machine Learning and statistical concepts.
· Hands-on experience with:
o LLMs
o Prompt Engineering
o RAG
o Vector Databases
o Embeddings
o Semantic Search
o LLM Evaluation
o AI Guardrails
· Experience with frameworks/tools such as LangChain, LangGraph, LlamaIndex, or equivalent.
· Experience with APIs and integrating LLMs into enterprise applications.
· Strong SQL and data handling skills.
· Experience working with large and complex datasets.
· Strong understanding of NLP concepts.XX
Technical Skills:
· Experience with OpenAI / Azure OpenAI / AWS Bedrock / Google Vertex AI.
· Experience with vector databases such as Pinecone, Weaviate, Milvus, FAISS, or equivalent.
· Experience with Databricks, Snowflake, or cloud data platforms.
· Experience with Docker and CI/CD.
· Exposure to AWS, Azure, or GCP.
· Experience with ML/AI deployment and MLOps.
· Knowledge of AI security, data privacy, governance, and responsible AI.
· Experience building AI Agents / Agentic AI workflows.
· Experience with multimodal AI is an added advantage
Key Competencies
· Strong analytical and problem-solving ability.
· Ability to translate business problems into practical AI solutions.
· Strong communication and presentation skills.
· Ability to interact confidently with senior stakeholders and clients.
· Strong ownership and delivery mindset.
· Ability to work independently in a fast-paced environment.
- Strong experimentation and innovation mindset.
- Ability to balance technical feasibility, business value, scalability, and cost.
Required Education & Experience:
· Bachelor's or Master's degree in Computer Science, Data Science, Artificial Intelligence, Statistics, Mathematics, Engineering, or a related discipline
Key Responsibilities:
- Develop and deploy machine learning, deep learning, and NLP models for various business use cases.
- Build end-to-end ML pipelines including data preprocessing, feature engineering, training, evaluation, and production deployment.
- Optimize model performance and ensure scalability in production environments.
- Work closely with data scientists, product teams, and engineers to translate business requirements into AI solutions.
- Conduct data analysis to identify trends and insights.
- Implement MLOps practices for versioning, monitoring, and automating ML workflows.
- Research and evaluate new AI/ML techniques, tools, and frameworks.
- Document system architecture, model design, and development processes.
Required Skills:
- Strong programming skills in Python (NumPy, Pandas, Scikit-learn, TensorFlow, PyTorch, Keras).
- Hands-on experience in building and deploying, finetuning ML/DL models in production.
- Good understanding of machine learning algorithms, neural networks, NLP, and computer vision.
- Experience with REST APIs, Docker, Kubernetes, and cloud platforms (AWS/GCP/Azure).
- Working knowledge of MLOps tools such as MLflow, Airflow, DVC, or Kubeflow.
- Familiarity with data pipelines and big data technologies (Spark, Hadoop) is a plus.
- Strong analytical skills and ability to work with large datasets.
- Excellent communication and problem-solving abilities.
- Experience in deploying models using cloud services (AWS Sagemaker, GCP Vertex AI, etc.).
- Experience in LLM fine-tuning or Generative AI, Voice AI, is an added advantage.
Educational Qualification:
- Bachelor’s or Master’s degree in Computer Science, Data Science, AI, Machine Learning, IT, from IIT/NIT colleges strongly preferred
Key Responsibilities
Strong understanding of Machine Learning algorithms (supervised and unsupervised)
Hands-on experience with Deep Learning frameworks (TensorFlow, PyTorch, or similar)
Experience in NLP techniques and libraries (NLTK, spaCy, Hugging Face, etc.)
Solid knowledge of Statistics, probability, and data analysis methods
Proficiency in SQL for querying relational databases
Strong programming skills in Python (or similar languages)
Excellent communication skills with the ability to explain complex concepts simply
Required Qualifications
3+ years of hands-on experience as a Data Scientist or in a similar role.
Strong expertise in classical machine learning and regression modeling.
Solid understanding of statistics, including probability, distributions, hypothesis testing,
and correlation analysis.
Proficiency in Python with libraries such as: scikit-learn pandas NumPy
Experience working with structured/tabular data.
Strong problem-solving and analytical thinking skills.
Ability to clearly explain models and results to non-technical stakeholders.
Must of Skills/Experience
• System Design
• Python
• TensorFlow
• Google ADK or Lang Graph
• Lang Chain , Lang Graph
• Spark
• Agentic AI Design
• ML Ops
• MCP (client and server)
• FastAPI
• Doc Factory
• RAG
• Golang
• LLMs – Gemini, Open AI
• NLP
• Dev Assistant - AI based code - generation
(Qwen or Claude or Copilot)
• CI/CD
• Good in oral and written communication,
collaboration and be a team player
Good to have skills
• DevOps with K8
• Scripting
• Java
• REST API
• UV
• ReACT
• DocFactory
• Unix
We are looking for a Data Science & Machine Learning Senior Associate with 3–5 years of relevant experience in data science, machine learning, and analytics. The candidate will be responsible for developing predictive models, analyzing complex datasets, building scalable ML solutions, and supporting production-grade data science applications on cloud platforms.
Key Responsibilities
- Develop and implement machine learning and predictive analytics models.
- Perform data analysis, statistical modeling, and optimization to solve business problems.
- Build demand forecasting and predictive models using time-series and other advanced techniques.
- Work with large datasets using Python, SQL, PySpark, and cloud-based data platforms.
- Develop and maintain scalable data pipelines for ML model development and deployment.
- Implement MLOps practices including model deployment, monitoring, retraining, and data-drift detection.
- Collaborate with software engineers, product teams, and business stakeholders to convert business requirements into analytical solutions.
- Validate models for accuracy, robustness, bias, and production readiness.
- Create meaningful visualizations and communicate analytical insights to technical and non-technical stakeholders.
Required Skills
- Python
- SQL
- Data Science & Machine Learning
- Predictive Modeling
- Statistical Analysis
- Machine Learning Algorithms
- Optimization Techniques
- Google Cloud Platform (GCP)
- BigQuery
- Dataflow
- Dataproc
- Data Fusion
- Cloud SQL
- Airflow
- PySpark
- PostgreSQL
- Terraform
- Tekton
- APIs
- MLOps
Preferred Skill
- Java
Good to Have
- Demand Forecasting
- Time-Series Analysis
- Neural Networks
- Ensemble Methods
- Support Vector Machines (SVM)
- Regression and Cluster Analysis
- ML Model Testing
- Bias Detection and Data Drift Monitoring
- Production ML Deployment
- QlikSense
- Automotive or Supply Chain Analytics
Education
Bachelor's degree in Computer Science, Data Science, Engineering, Statistics, Mathematics, or a related technical field.Master's degree in a relevant quantitative or technical field is preferred.
Description
We’re seeking a highly skilled, execution-focused Senior Data Scientist with a minimum of 5 years of experience. This role demands hands-on expertise in building, deploying, and optimizing machine learning models at scale, while working with big data technologies and modern cloud platforms. You will be responsible for driving data-driven solutions from experimentation to production, leveraging advanced tools and frameworks across Python, SQL, Spark, and AWS. The role requires strong technical depth, problem-solving ability, and ownership in delivering business impact through data science.
Responsibilities
- Design, build, and deploy scalable machine learning models into production systems.
- Develop advanced analytics and predictive models using Python, SQL, and popular ML/DL frameworks (Pandas, Scikit-learn, TensorFlow, PyTorch).
- Leverage Databricks, Apache Spark, and Hadoop for large-scale data processing and model training.
- Implement workflows and pipelines using Airflow and AWS EMR for automation and orchestration.
- Collaborate with engineering teams to integrate models into cloud-based applications on AWS.
- Optimize query performance, storage usage, and data pipelines for efficiency.
- Conduct end-to-end experiments, including data preprocessing, feature engineering, model training, validation, and deployment.
- Drive initiatives independently with high ownership and accountability.
- Stay up to date with industry best practices in machine learning, big data, and cloud-native deployments.
Requirements
- Minimum 5 years of experience in Data Science or Applied Machine Learning.
- Strong proficiency in Python, SQL, and ML libraries (Pandas, Scikit-learn, TensorFlow, PyTorch).
- Proven expertise in deploying ML models into production systems.
- Experience with big data platforms (Hadoop, Spark) and distributed data processing.
- Hands-on experience with Databricks, Airflow, and AWS EMR.
- Strong knowledge of AWS cloud services (S3, Lambda, SageMaker, EC2, etc.).
- Solid understanding of query optimization, storage systems, and data pipelines.
- Excellent problem-solving skills, with the ability to design scalable solutions.
- Strong communication and collaboration skills to work in cross-functional teams.
Benefits
- Best-in-class salary: We hire strong talent and compensate accordingly.
- Proximity Talks: Meet and learn from designers, engineers, product leaders, and AI practitioners.
- Continuous learning: Work with a world-class team and stay close to the latest in AI, engineering, and product development.
- High-impact work: Build AI-first systems and products used at scale by global clients.
About Us
Proximity is the trusted technology, design, and consulting partner for some of the biggest Sports, Media, and Entertainment companies in the world. We’re headquartered in San Francisco and have offices in Palo Alto, Dubai, Mumbai, and Bangalore.
Since 2019, Proximity has built high-impact, scalable products used by millions of users every day. Today, we are a global team of engineers, designers, product managers, and experts solving complex problems and building cutting-edge technology at scale.













