Data Analyst / ML Engineer at Talent Pro · Mumbai, Pune, Hyderabad, Bengaluru (Bangalore) · 4 - 12 years · ₹20L - ₹40L / yr · Bootstrapped · Posted 23 Apr 2025

What You will do:
● Play the role of Data Analyst / ML Engineer
● Collection, cleanup, exploration and visualization of data
● Perform statistical analysis on data and build ML models
● Implement ML models using some of the popular ML algorithms
● Use Excel to perform analytics on large amounts of data
● Understand, model and build to bring actionable business intelligence out of data that is available in different formats
● Work with data engineers to design, build, test and monitor data pipelines for ongoing business operations
Basic Qualifications:
● Experience: 4+ years.
● Hands-on development experience playing the role of Data Analyst and/or ML Engineer.
● Experience in working with excel for data analytics
● Experience with statistical modelling of large data sets
● Experience with ML models and ML algorithms
● Coding experience in Python
Nice to have Qualifications:
● Experience with wide variety of tools used in ML
● Experience with Deep learning
Benefits:
● Competitive salary.
● Hybrid work model.
● Learning and gaining experience rapidly.
● Reimbursement for basic working set up at home.
● Insurance (including a top up insurance for COVID

Similar jobs (10)
Role Overview
As a Data Scientist, you will work with business stakeholders, AI engineers, and domain experts to transform data into actionable insights and intelligent solutions. You will develop machine learning models, perform statistical analysis, and contribute to AI-driven products that create measurable business impact.
Key Responsibilities
Data Science & Machine Learning
- Analyze structured and unstructured data to identify patterns, trends, and business opportunities.
- Perform exploratory data analysis (EDA), feature engineering, and data preparation.
- Develop, evaluate, and optimize machine learning models for prediction, classification, clustering, and forecasting.
- Apply statistical techniques to solve business problems and validate model performance.
- Design and execute experiments to improve model accuracy and business outcomes.
AI Solution Development
- Collaborate with AI Engineers, Data Engineers, and domain experts to build AI-powered solutions.
- Translate business requirements into scalable data science approaches.
- Contribute to Generative AI and advanced analytics initiatives where applicable.
- Document methodologies, model performance, and key findings.
Required Technical Skills
- Strong programming skills in Python and SQL for data analysis, feature engineering, and machine learning.
- Strong understanding of Statistics, Probability, Linear Algebra, and Calculus as applied to machine learning and data science.
- Experience with Exploratory Data Analysis (EDA), data preprocessing, feature engineering, feature selection, and handling missing or imbalanced data.
- Good understanding of Supervised, Unsupervised, and Ensemble Machine Learning algorithms, including their assumptions, strengths, limitations, and appropriate use cases.
- Strong knowledge of Regression, Classification, Clustering, Time Series Forecasting, Dimensionality Reduction, Recommendation Systems, and Anomaly Detection techniques.
- Experience with Model Evaluation, Cross-Validation, Hyperparameter Optimization, Bias-Variance Trade-off, Feature Importance, Explainable AI (XAI), and Performance Metrics.
- Understanding of Statistical Inference, Hypothesis Testing, Probability Distributions, Sampling Techniques, Confidence Intervals, and A/B Testing.
- Experience translating business problems into analytical approaches and developing scalable, data-driven solutions.
- Working knowledge of Generative AI, Large Language Models (LLMs), Prompt Engineering, and Retrieval-Augmented Generation (RAG) is preferred.
Preferred Qualifications
- Bachelor's or master's degree in computer science, Artificial Intelligence, Data Science, Statistics, Mathematics, Engineering, or a related field.
- 2–4 years of experience developing machine learning or data science solutions.
- Experience working on end-to-end data science projects in a business environment.
Nice to Have
- Exposure to Generative AI, LLMs, RAG, or Agentic AI.
- Experience with Computer Vision or Natural Language Processing (NLP).
- Familiarity with cloud-based AI platforms.
- Knowledge of construction, engineering, manufacturing, or industrial domains.
- Participation in hackathons, research, Kaggle competitions, or open-source projects.
Soft Skills
Strong analytical and problem-solving skills, effective communication and collaboration, ownership mindset, adaptability, continuous learning, and a passion for innovation.
Hiring for Data Scientist / Senior Data Scientist
Exp : 4 - 12 yrs
Edu : BE/B.tech/MCA
Work Location : Pune
Notice Period : Immediate - 15 days
Skills :
4+ years of experience in data engineering, data science, or related domains.
Hands-on experience with SQL, Python, and distributed data systems.
Knowledge of machine learning techniques and statistical analysis.
Experience with cloud data platforms (Azure Data Factory, AWS Glue, GCP BigQuery).
Familiarity with DevOps practices and CI/CD for data pipelines.
Platforms & Operations Experience (Preferred)
- Experience working with Azure, AWS, or Google Cloud data tools.
Operational experience with data orchestration tools (Airflow, ADF, Glue).
Understanding of Kubernetes, Docker, or containerized environments.
Hands-on experience with data warehousing platforms (Snowflake, Redshift, BigQuery).
Experience in monitoring, logging, and alerting operations for data workflows.
At Nineleaps, we work on bleeding-edge technology with class-leading engineering practices on products that touch the lives of millions of users. We endeavor on doing things the right way, while also promoting a culture of excellence.
About the Role:
We are looking for a Data Analyst with strong analytical and problem-solving skills to transform complex data into meaningful, actionable business insights. The role involves working with large datasets, conducting deep-dive analysis, driving automation, and supporting data-driven product and business decisions.
Key Responsibilities:
- Analyse historical and large datasets to understand data sources, identify trends and patterns, and uncover meaningful insights.
- Write complex SQL queries and leverage Python to perform data analysis, ad hoc investigations, and solve business problems.
- Create reports and translate analytical findings into clear, concise, and actionable recommendations for stakeholders.
- Identify opportunities to drive automation and process improvements, improving efficiency and reducing manual effort.
- Communicate data-driven insights effectively to both technical and non-technical stakeholders in a clear and impactful manner.
- Maintain accurate documentation, ensure high-quality deliverables, and consistently meet defined timelines.
Requirements:
- 3–6 years of experience in Data Analytics, Business Intelligence, Data Engineering, or a similar analytical role.
- Strong hands-on expertise in Python and advanced SQL, with the ability to work with and analyse large datasets.
- Experience working with Google Sheets, and implementing automation through data pipelines or workflows.
- Strong analytical and problem-solving skills, with the ability to interpret complex data and derive actionable insights.
- Excellent communication skills with the ability to effectively present methods, results, and recommendations to stakeholders.
- Ability to collaborate effectively with remote and geographically distributed teams across different time zones.
Company Link: https://www.nineleaps.com/
Company LinkedIn: https://www.linkedin.com/company/nineleaps/
Sr.Data Scientist,Python, AI ML
We are looking for a skilled Data Scientist to analyze complex datasets, develop predictive models, and generate actionable insights that support business decisions. The ideal candidate should have strong statistical, analytical, and programming skills, along with hands-on experience in machine learning.
We are seeking a Senior Data Science & ML Associate with 4+ years of applied ML experience to build and ship models end-to-end from data prep and feature engineering to training, evaluation, and deployment driving measurable business impact.
Key Responsibilities
• Build, train, and evaluate ML and deep-learning models
• Engineer features and prepare data at scale
• Deploy models and monitor production performance
• Partner with stakeholders to frame problems and metrics
• Communicate results and drive decisions
• Iterate on models from business feedback
Mandatory Skills
• 4+ years applied machine learning
• Strong Python (Pandas, NumPy, scikit-learn)
• Classical ML and deep learning (TensorFlow/PyTorch)
• Solid statistics and experiment design
• SQL and data wrangling at scale
• Model deployment / MLOps exposure
Nice to Have: NLP or computer vision; cloud ML (SageMaker, Azure ML)
We are looking for a talented and driven Data Scientist to join our growing Analytics team in India. In this role, you will work at the intersection of advanced machine learning, scalable MLOps infrastructure, and domain-specific healthcare analytics. You will collaborate closely with cross-functional teams to build, deploy, and maintain production-grade ML models that drive real-world impact in clinical trials and healthcare operations.
KEY RESPONSIBILITIES
End-to-End ML Development
• Design, build, and optimize predictive models across the full ML lifecycle—from data ingestion to model serving.
• Conduct rigorous Exploratory Data Analysis (EDA) to surface insights and drive feature engineering decisions.
• Validate model performance using appropriate statistical techniques and domain knowledge.
MLOps & Production Deployment
• Deploy, monitor, and maintain production-grade ML models using Databricks MLFlow endpoints and Unity Catalog.
• Implement CI/CD pipelines for model versioning, experiment tracking, and automated retraining.
• Ensure model reliability, observability, and performance in live production environments.
Language Models & LLM Applications
• Apply transformer-based models (BERT, ClinicalBERT, Trial2Vec) for NLP tasks including classification, NER, and information extraction.
• Build and maintain vector similarity search pipelines for semantic retrieval and recommendation use cases.
• Fine-tune pre-trained models for domain-specific applications in clinical and healthcare contexts.
• Support exploratory work around LLM integration and prompt engineering for internal tooling.
Domain-Driven Analytics
• Apply advanced analytics within complex healthcare and clinical trial datasets—including patient records, trial protocols, and adverse event data.
• Translate ambiguous business problems into structured analytical frameworks with measurable outcomes.
• Partner with domain experts, product managers, and engineering teams to deliver data-driven solutions.
REQUIRED QUALIFICATIONS
Education
• Bachelor’s or Master’s degree in Computer Science, Statistics, Mathematics, Bioinformatics, or a closely related field.
Experience
• 2–4 years of hands-on experience in a data science or machine learning role.
• Demonstrable experience deploying ML models in production environments (not just prototyping).
Technical Skills
• Strong proficiency in Python (pandas, NumPy, scikit-learn, PyTorch / TensorFlow).
• Experience with Databricks, MLFlow (experiment tracking, model registry, endpoints), and Unity Catalog.
• Hands-on experience with BERT-family models and Hugging Face Transformers library.
• Familiarity with vector databases (e.g., FAISS, Pinecone, Weaviate) and embedding-based retrieval.
• Solid understanding of SQL and working with large structured/unstructured datasets.
• Exposure to cloud platforms (AWS / GCP / Azure) and distributed computing frameworks (Spark).
GOOD TO HAVE
• Prior experience with clinical trial data standards (CDISC, CDASH, SDTM) or healthcare ontologies (SNOMED, ICD-10).
• Familiarity with Trial2Vec or similar trial-to-vector embedding approaches.
• Experience with LLM fine-tuning, RAG pipelines, or prompt engineering in a production setting.
• Knowledge of regulatory and compliance considerations in healthcare AI (e.g., FDA guidelines, HIPAA).
• Contributions to open-source ML projects or published research.
THIS ROLE IS NOT FOR YOU IF…
• You have strong SQL/BI skills but limited hands-on ML modelling experience — or you’ve built models only in notebooks without ever deploying them to production.
• Your LLM exposure is limited to API calls and prompt engineering — with no experience fine-tuning models, working with embeddings, or building vector search pipelines.
We are looking for a talented and driven Data Scientist to join our growing Analytics team in India. In this role, you will work at the intersection of advanced machine learning, scalable MLOps infrastructure, and domain-specific healthcare analytics. You will collaborate closely with cross-functional teams to build, deploy, and maintain production-grade ML models that drive real-world impact in clinical trials and healthcare operations.
KEY RESPONSIBILITIES
End-to-End ML Development
• Design, build, and optimize predictive models across the full ML lifecycle—from data ingestion to model serving.
• Conduct rigorous Exploratory Data Analysis (EDA) to surface insights and drive feature engineering decisions.
• Validate model performance using appropriate statistical techniques and domain knowledge.
MLOps & Production Deployment
• Deploy, monitor, and maintain production-grade ML models using Databricks MLFlow endpoints and Unity Catalog.
• Implement CI/CD pipelines for model versioning, experiment tracking, and automated retraining.
• Ensure model reliability, observability, and performance in live production environments.
Language Models & LLM Applications
• Apply transformer-based models (BERT, ClinicalBERT, Trial2Vec) for NLP tasks including classification, NER, and information extraction.
• Build and maintain vector similarity search pipelines for semantic retrieval and recommendation use cases.
• Fine-tune pre-trained models for domain-specific applications in clinical and healthcare contexts.
• Support exploratory work around LLM integration and prompt engineering for internal tooling.
Domain-Driven Analytics
• Apply advanced analytics within complex healthcare and clinical trial datasets—including patient records, trial protocols, and adverse event data.
• Translate ambiguous business problems into structured analytical frameworks with measurable outcomes.
• Partner with domain experts, product managers, and engineering teams to deliver data-driven solutions.
REQUIRED QUALIFICATIONS
Education
• Bachelor’s or Master’s degree in Computer Science, Statistics, Mathematics, Bioinformatics, or a closely related field.
Experience
• 2–4 years of hands-on experience in a data science or machine learning role.
• Demonstrable experience deploying ML models in production environments (not just prototyping).
Technical Skills
• Strong proficiency in Python (pandas, NumPy, scikit-learn, PyTorch / TensorFlow).
• Experience with Databricks, MLFlow (experiment tracking, model registry, endpoints), and Unity Catalog.
• Hands-on experience with BERT-family models and Hugging Face Transformers library.
• Familiarity with vector databases (e.g., FAISS, Pinecone, Weaviate) and embedding-based retrieval.
• Solid understanding of SQL and working with large structured/unstructured datasets.
• Exposure to cloud platforms (AWS / GCP / Azure) and distributed computing frameworks (Spark).
GOOD TO HAVE
• Prior experience with clinical trial data standards (CDISC, CDASH, SDTM) or healthcare ontologies (SNOMED, ICD-10).
• Familiarity with Trial2Vec or similar trial-to-vector embedding approaches.
• Experience with LLM fine-tuning, RAG pipelines, or prompt engineering in a production setting.
• Knowledge of regulatory and compliance considerations in healthcare AI (e.g., FDA guidelines, HIPAA).
• Contributions to open-source ML projects or published research.
THIS ROLE IS NOT FOR YOU IF…
• You have strong SQL/BI skills but limited hands-on ML modelling experience — or you’ve built models only in notebooks without ever deploying them to production.
• Your LLM exposure is limited to API calls and prompt engineering — with no experience fine-tuning models, working with embeddings, or building vector search pipelines.
About Us:
The QX Impact was launched with a mission to make A.I accessible and affordable and deliver AI Products/Solutions at scale for the enterprises by bringing the power of Data, AI, and Engineering to drive digital transformation. We believe without insights; businesses will continue to face challenges to better understand their customers and even lose them. Secondly, without insights businesses won't’ be able to deliver differentiated products/services; and finally, without insights, businesses can’t achieve a new level of “Operational Excellence” is crucial to remain competitive, meeting rising customer expectations, expanding markets, and digitalization.
Role Overview
We are seeking a Machine Learning Engineer to lead the end-to-end development of production-grade analytical applications. This is a high-impact role requiring a blend of deep statistical modeling and machine learning. You will be responsible transforming raw consolidated data into high-accuracy forecasts through advanced feature engineering, rigorous model selection, and statistical validation.
This role is for an engineer who thrives in the research-to-code transition, ensuring that every model is mathematically sound, resistant to overfitting, and optimized for high-dimensional manufacturing data.
Responsibilities:
- Feature Engineering & Discovery: Design and build complex feature sets for diverse problem types, including behavioural features for churn, sensor-based lags for maintenance, and seasonal encodings for demand forecasting.
- Model Selection & Optimization: Conduct systematic experimentation across diverse algorithms (e.g., XGBoost, LightGBM, Prophet, or Deep Learning) to identify the best-performing models.
- Model Training & Testing: Develop, train, tune, and test a variety of ML architectures including time-series, classification and regression.
- Statistical Validation & Evaluation: Define and track complex evaluation metrics tailored to manufacturing, such as MAPE, RMSE, etc., while performing deep-dive bias-variance analysis.
- EDA & Research: Perform exploratory data analysis on consolidated "Gold" layer data to uncover hidden drivers of business outcomes and identify correlations between external signals.
- Refinement & Performance Tuning: Address critical modeling challenges including bias-variance tradeoffs, class imbalance, and overfitting to ensure models generalize to real-world production data.
Skills & Requirements:
- 3+ Years of Experience: Proven track record of developing and delivering production-grade ML models across multiple domains (Sales, Finance, Manufacturing, or Supply Chain).
- Mastery of the Python Ecosystem: Expert-level skills in Pandas, NumPy, Scikit-learn, and SciPy.
- Advanced Algorithmic Knowledge: Deep expertise in supervised and unsupervised learning, specifically ensemble methods (Boosting/Bagging) and time-series frameworks.
- Statistical Foundations: Strong grasp of hypothesis testing, probability distributions, and the mathematical principles behind model evaluation and optimization.
- SQL Proficiency: Expert ability to manipulate data within consolidated database layers to create the "Silver" feature sets required for training.
- Education: Bachelor’s or Master’s degree in a quantitative field (e.g., Data Science, Statistics, Mathematics, or Computer Science).
- Cloud Awareness: Experience with Azure Machine Learning or similar cloud modelling environments.
- Engineering Familiarity: Basic understanding of Docker, MLflow, or FastAPI for handing models off to deployment teams.
Personal Attributes:
- Strong problem-solving skills with a passion for data architecture.
- Excellent communication skills with the ability to explain complex data concepts to non-technical stakeholders.
- Highly collaborative, capable of working with cross-functional teams.
- Ability to thrive in a fast-paced, agile environment while managing multiple priorities effectively.
Competencies:
- Tech Savvy - Anticipating and adopting innovations in business-building digital and technology applications.
- Self-Development - Actively seeking new ways to grow and be challenged using both formal and informal development channels.
- Action Oriented - Taking on new opportunities and tough challenges with a sense of urgency, high energy, and enthusiasm.
- Customer Focus - Building strong customer relationships and delivering customer-centric solutions.
- Optimize Work Processes - Knowing the most effective and efficient processes to get things done, with a focus on continuous improvement.
Why Join Us?
- Be part of a collaborative and agile team driving cutting-edge AI and data engineering solutions.
- Work on impactful projects that make a difference across industries.
- Opportunities for professional growth and continuous learning.
- Competitive salary and benefits package.
Application Details
Ready to make an impact? Apply today and become part of the QX Impact team!

🚀 We’re Hiring | Data Scientist 🧠📊
Ready to turn data into real-world intelligence? Join us and work on exciting AI/ML & data-driven solutions!
🔹 Experience: 8+ Years
🔹 Must-Have Skills:
🐍 Python | 🤖 Machine Learning | ☁️ Cloud | 🧠 NLP | 📊 Data Visualization
📍 Location: Pune
💼 Work Mode: Work from Office
If you're passionate about Data Science, AI & solving complex business problems, we’d love to hear from you!
📩 Interested? Kindly text
#Hiring #DataScientist #DataScience #MachineLearning #Python #NLP #AI #Cloud #DataVisualization #TechJobs #HiringNow
Strong Data Analyst Profile with advanced Excel and SQL expertise
2
Mandatory (Experience 1): Must have 4+ years of overall experience as a hands-on Data Analyst
3
Mandatory (Tech skill 1): Must be highly proficient in advanced Excel — complex functions, macros, calculations, and pivots
4
Mandatory (Tech skill 2): Must have strong hands-on SQL and a good understanding of relational database concepts
5
Mandatory (Tech skill 3): Must be able to automate routine tasks using Python (for automation purposes)
6
Mandatory (Skill 1): Must have exceptional analytical, problem-solving, and logical skills, with strong attention to detail and accuracy
7
Mandatory (Skill 2): Must be able to understand complex data and business logic and convert it into a model (the role models complex utility tariffs, rates, and programs)
8
Mandatory (Communication): Must have strong verbal and written communication, able to work independently with India- and US-based team members and articulate problems and solutions over calls and email.
9
Mandatory (Location): Must be based locally in Pune (or the nearby Maharashtra belt — Mumbai, Nagpur), as the final round is in person
10
Preferred (Domain): Experience in the Energy/Utility industry and familiarity with basic utility (electrical/gas) tariff concepts





