Data Engineer at Qube Research · Mumbai · 3 - 8 years · ₹30L - ₹60L / yr · Posted 29 Jan 2025

Data Engineer - Python
Qube Research & Technologies (QRT) is a global quantitative and systematic investment manager, operating in all liquid asset classes across the world. We are a technology and data driven group implementing a scientific approach to investing. Combining data, research, technology and trading expertise has shaped QRT’s collaborative mindset which enables us to solve the most complex challenges. QRT’s culture of innovation continuously drives our ambition to deliver high quality returns for our investors.
Your future role within QRT
- Use Python web frameworks such as Dash or Streamlit to make latency data accessible to end users.
- Leverage OpenAI API to analyse and generate insights by using latency data and provide visualization interface to end users.
- Improve the Latency Data API to provide seamless access to latency data for end users, including researchers, support, and other teams within QRT.
- Optimize and fine-tune Large Language Models (LLMs) for use cases in natural language understanding, data processing, and predictive modelling.
Your present skillset
- 4+ years' experience in a related field
- Advanced Python and OpenAI API proficiency.
- Good with LLM Concepts.
- Experience with large data sets, visualizations, and reports.
- High standards in code quality and development practices.
- Proficiency in managing structured/unstructured data.
- Expertise in scalable data systems.
- Strong AI/ML skills for pattern and anomaly detection.
- Experience with CI/CD processes and DevOps.
- Experience in the financial/trading sector.
- Familiarity with latency-sensitive systems and performance optimization
QRT is an equal opportunity employer. We welcome diversity as essential to our success. QRT empowers employees to work openly and respectfully to achieve collective success. In addition to professional achievement, we are offering initiatives and programs to enable employees achieve a healthy work-life balance.

Similar jobs (10)
Job Description: Python + AI
Company: Wissen Technology
Location: Bangalore, India
Experience: 5+Years
Employment Type: Full-Time
Role: Python + AI / Data Engineer
About the Role
Wissen Technology is looking for experienced Python + AI / Data Engineering professionals to join our technology team in Bangalore. The ideal candidate will have strong hands-on experience in Python, Artificial Intelligence, Generative AI, PySpark, Snowflake, and data pipeline development.
The candidate should be capable of designing and developing scalable data and AI solutions, building robust ETL/ELT pipelines, working with large datasets, and integrating AI/ML capabilities into enterprise applications.
Key Responsibilities
- Design, develop, and maintain scalable data pipelines using Python and PySpark.
- Develop robust ETL/ELT pipelines for processing large volumes of structured and unstructured data.
- Build and optimize data processing solutions using Apache Spark / PySpark.
- Develop data ingestion and transformation pipelines into Snowflake.
- Design and implement scalable Snowflake data models, tables, views, and SQL transformations.
- Work with batch and, where applicable, real-time data processing pipelines.
- Build and integrate AI and Generative AI solutions using Python.
- Develop LLM-based applications, RAG solutions, AI agents, and AI-powered services.
- Integrate AI models with enterprise data platforms and data pipelines.
- Develop REST APIs and microservices using FastAPI, Flask, or Django.
- Perform data cleansing, transformation, validation, and quality checks.
- Optimize PySpark jobs, SQL queries, Snowflake workloads, and data pipelines for performance and scalability.
- Implement data pipeline monitoring, logging, error handling, and alerting.
- Work with cloud platforms such as AWS, Azure, or GCP.
- Collaborate with Data Scientists, Data Engineers, Software Engineers, Architects, and business stakeholders.
- Participate in technical design, architecture, code reviews, and production support.
- Mentor junior engineers and contribute to engineering best practices.
Preferred Qualifications
- Bachelor's or master's degree in computer science, Engineering, Data Science, Artificial Intelligence, or a related field.
- Experience working on enterprise-scale AI and data engineering projects.
- Experience combining Python + PySpark + Snowflake + AI/GenAI in production environments.
- Experience with Databricks is an advantage.
- Experience with AI Agents / Agentic AI and tool/function calling.
- Knowledge of distributed systems and cloud-native architecture.
- Experience leading technical initiatives or mentoring engineering teams.

🚀 We’re Hiring | Data Scientist 🧠📊
Ready to turn data into real-world intelligence? Join us and work on exciting AI/ML & data-driven solutions!
🔹 Experience: 8+ Years
🔹 Must-Have Skills:
🐍 Python | 🤖 Machine Learning | ☁️ Cloud | 🧠 NLP | 📊 Data Visualization
📍 Location: Pune
💼 Work Mode: Work from Office
If you're passionate about Data Science, AI & solving complex business problems, we’d love to hear from you!
📩 Interested? Kindly text
#Hiring #DataScientist #DataScience #MachineLearning #Python #NLP #AI #Cloud #DataVisualization #TechJobs #HiringNow
Strong AI Engineer / Machine Learning Engineer profiles.
2
Mandatory (Experience 1) – Must have minimum 3+ years of hands-on experience in Data Science, Machine Learning, Applied AI, NLP, Deep Learning, or Generative AI solutions.
3
Mandatory (Experience 2) – Must have strong hands-on experience in Python programming, SQL, data analysis, feature engineering, model development, and production-grade ML applications.
4
Mandatory (Experience 3) – Must have experience working with Machine Learning and Deep Learning frameworks such as PyTorch, TensorFlow, Keras, Scikit-learn, or equivalent.
5
Mandatory (Experience 4) – Must have hands-on experience working on NLP, embeddings, semantic search, text classification, document understanding, recommendation systems, or similar AI/ML use cases.
6
Mandatory (Experience 5) – Must have experience working with Large Language Models (LLMs) such as GPT, Llama, Mistral, Claude, Gemini, Phi, or similar foundation models.
7
Mandatory (Experience 6) – Must have hands-on experience building or implementing RAG (Retrieval Augmented Generation) systems, vector search, knowledge retrieval, embeddings, chunking, indexing, or semantic retrieval solutions.
8
Mandatory (Experience 7) – Must have experience working with Git, CI/CD practices, production environments, and scalable AI/ML systems.
9
Mandatory (CTC) – The CTC breakup offered will be 75% fixed + 25% variable, as per company policy.
10
Mandatory (Age) - Candidate's Age should be below 30 Years
11
Preferred (Experience 1) – Experience with MLFlow, Kubeflow, Airflow, Prefect, Feature Stores, Model Registry, or MLOps/LLMOps frameworks.
12
Preferred (Experience 2) – Experience working with Vector Databases, Spark, PySpark, distributed ML pipelines, large-scale data processing, or real-time ML systems..
13
Preferred (Experience 3) – Familiarity with Docker, Kubernetes, Azure, AWS, GCP, cloud-native AI deployments, and scalable ML architecture.
14
Preferred (Company) – Candidates from AI-first startups, Fintech, Banking, Lending, Fraud Analytics, Risk Analytics, Product Companies, SaaS organizations, or data-driven technology companies
15
Mandatory ( Pedigree) - B.TECH / M.TECH from Tier 1 Colleges (IIT's, NIT's, BITS) are Considered.
Role Overview
As a Data Scientist, you will work with business stakeholders, AI engineers, and domain experts to transform data into actionable insights and intelligent solutions. You will develop machine learning models, perform statistical analysis, and contribute to AI-driven products that create measurable business impact.
Key Responsibilities
Data Science & Machine Learning
- Analyze structured and unstructured data to identify patterns, trends, and business opportunities.
- Perform exploratory data analysis (EDA), feature engineering, and data preparation.
- Develop, evaluate, and optimize machine learning models for prediction, classification, clustering, and forecasting.
- Apply statistical techniques to solve business problems and validate model performance.
- Design and execute experiments to improve model accuracy and business outcomes.
AI Solution Development
- Collaborate with AI Engineers, Data Engineers, and domain experts to build AI-powered solutions.
- Translate business requirements into scalable data science approaches.
- Contribute to Generative AI and advanced analytics initiatives where applicable.
- Document methodologies, model performance, and key findings.
Required Technical Skills
- Strong programming skills in Python and SQL for data analysis, feature engineering, and machine learning.
- Strong understanding of Statistics, Probability, Linear Algebra, and Calculus as applied to machine learning and data science.
- Experience with Exploratory Data Analysis (EDA), data preprocessing, feature engineering, feature selection, and handling missing or imbalanced data.
- Good understanding of Supervised, Unsupervised, and Ensemble Machine Learning algorithms, including their assumptions, strengths, limitations, and appropriate use cases.
- Strong knowledge of Regression, Classification, Clustering, Time Series Forecasting, Dimensionality Reduction, Recommendation Systems, and Anomaly Detection techniques.
- Experience with Model Evaluation, Cross-Validation, Hyperparameter Optimization, Bias-Variance Trade-off, Feature Importance, Explainable AI (XAI), and Performance Metrics.
- Understanding of Statistical Inference, Hypothesis Testing, Probability Distributions, Sampling Techniques, Confidence Intervals, and A/B Testing.
- Experience translating business problems into analytical approaches and developing scalable, data-driven solutions.
- Working knowledge of Generative AI, Large Language Models (LLMs), Prompt Engineering, and Retrieval-Augmented Generation (RAG) is preferred.
Preferred Qualifications
- Bachelor's or master's degree in computer science, Artificial Intelligence, Data Science, Statistics, Mathematics, Engineering, or a related field.
- 2–4 years of experience developing machine learning or data science solutions.
- Experience working on end-to-end data science projects in a business environment.
Nice to Have
- Exposure to Generative AI, LLMs, RAG, or Agentic AI.
- Experience with Computer Vision or Natural Language Processing (NLP).
- Familiarity with cloud-based AI platforms.
- Knowledge of construction, engineering, manufacturing, or industrial domains.
- Participation in hackathons, research, Kaggle competitions, or open-source projects.
Soft Skills
Strong analytical and problem-solving skills, effective communication and collaboration, ownership mindset, adaptability, continuous learning, and a passion for innovation.
Hiring for AI Engineer
Exp: 6 - 8 yrs
Edu : BE/B.Tech/MCA
Work Location : Pune
Skill Set:
- Total experience ranging from 6–8 years in software engineering/AI roles
- Min 5 years strong programming experience in Python is a MUST
- Min 3.5 years hands-on experience in AI with LLMs, RAG pipelines, and AI frameworks
- Experience with cloud platforms (AWS/Azure/GCP)
Sr.Data Scientist,Python, AI ML
We are looking for a skilled Data Scientist to analyze complex datasets, develop predictive models, and generate actionable insights that support business decisions. The ideal candidate should have strong statistical, analytical, and programming skills, along with hands-on experience in machine learning.
Strong Data Scientist / AI Engineer / Generative AI Engineer profile.
2
Mandatory (Experience 1) - Must have 3+ years of hands-on experience in Data Science, Artificial Intelligence, Machine Learning, Deep Learning, NLP, or Generative AI application development.
3
Mandatory (Experience 2) - Must have strong hands-on experience in Python programming, backend development, API development, and production-grade application support.
4
Mandatory (Experience 3) - Must have experience working with Machine Learning and Deep Learning frameworks such as PyTorch, TensorFlow, Keras, or Scikit-learn.
5
Mandatory (Experience 4) - Must have hands-on experience in NLP use cases such as text classification, sentiment analysis, entity recognition (NER), semantic search, embeddings, or document understanding.
6
Mandatory (Experience 5) - Must have experience working with Large Language Models (LLMs) such as GPT, LLaMA, Mistral, Phi, Claude, Gemini, or similar models.
7
Mandatory (Experience 6) - Must have hands-on experience building or implementing Retrieval Augmented Generation (RAG) solutions, vector search, semantic search, or knowledge-based AI applications.
8
Mandatory (Experience 7) - Must have experience with Prompt Engineering and Generative AI frameworks such as LangChain, LangGraph, AI Agents, Azure OpenAI, or similar technologies.
9
Mandatory (Experience 8) - Must have experience developing, consuming, or integrating APIs using Python frameworks such as FastAPI, Flask, or similar technologies.
10
Mandatory (CTC) - The CTC breakup offered will be 75% fixed + 25% variable, as per company policy.
11
Preferred (Experience 1) - Experience with LLMOps/MLOps tools for monitoring, evaluation, experimentation, and versioning of AI models.
12
Preferred (Experience 2) - Exposure to Azure OpenAI, Azure Kubernetes Service (AKS), Kubernetes, cloud-native AI deployments, or distributed systems.
Description
We’re seeking a highly skilled, execution-focused Senior Data Scientist with a minimum of 5 years of experience. This role demands hands-on expertise in building, deploying, and optimizing machine learning models at scale, while working with big data technologies and modern cloud platforms. You will be responsible for driving data-driven solutions from experimentation to production, leveraging advanced tools and frameworks across Python, SQL, Spark, and AWS. The role requires strong technical depth, problem-solving ability, and ownership in delivering business impact through data science.
Responsibilities
- Design, build, and deploy scalable machine learning models into production systems.
- Develop advanced analytics and predictive models using Python, SQL, and popular ML/DL frameworks (Pandas, Scikit-learn, TensorFlow, PyTorch).
- Leverage Databricks, Apache Spark, and Hadoop for large-scale data processing and model training.
- Implement workflows and pipelines using Airflow and AWS EMR for automation and orchestration.
- Collaborate with engineering teams to integrate models into cloud-based applications on AWS.
- Optimize query performance, storage usage, and data pipelines for efficiency.
- Conduct end-to-end experiments, including data preprocessing, feature engineering, model training, validation, and deployment.
- Drive initiatives independently with high ownership and accountability.
- Stay up to date with industry best practices in machine learning, big data, and cloud-native deployments.
Requirements
- Minimum 5 years of experience in Data Science or Applied Machine Learning.
- Strong proficiency in Python, SQL, and ML libraries (Pandas, Scikit-learn, TensorFlow, PyTorch).
- Proven expertise in deploying ML models into production systems.
- Experience with big data platforms (Hadoop, Spark) and distributed data processing.
- Hands-on experience with Databricks, Airflow, and AWS EMR.
- Strong knowledge of AWS cloud services (S3, Lambda, SageMaker, EC2, etc.).
- Solid understanding of query optimization, storage systems, and data pipelines.
- Excellent problem-solving skills, with the ability to design scalable solutions.
- Strong communication and collaboration skills to work in cross-functional teams.
Benefits
- Best-in-class salary: We hire strong talent and compensate accordingly.
- Proximity Talks: Meet and learn from designers, engineers, product leaders, and AI practitioners.
- Continuous learning: Work with a world-class team and stay close to the latest in AI, engineering, and product development.
- High-impact work: Build AI-first systems and products used at scale by global clients.
About Us
Proximity is the trusted technology, design, and consulting partner for some of the biggest Sports, Media, and Entertainment companies in the world. We’re headquartered in San Francisco and have offices in Palo Alto, Dubai, Mumbai, and Bangalore.
Since 2019, Proximity has built high-impact, scalable products used by millions of users every day. Today, we are a global team of engineers, designers, product managers, and experts solving complex problems and building cutting-edge technology at scale.
Sr Engineer – Artificial Intelligence
Job Summary
As an AI Engineer at Emerson, you will be responsible for analysing complex data sets to
identify trends, develop predictive models, and provide actionable insights. You will work closely
with cross-functional teams to understand business needs and deliver data-driven solutions that
enhance decision-making and drive business growth.
In This Role, Your Responsibilities Will Be:
Analyze large, complex data sets using statistical methods and machine learning
techniques to extract meaningful insights.
Develop and implement predictive models and algorithms to solve business problems
and improve processes.
Create visualizations and dashboards to effectively communicate findings and insights to
stakeholders.
Work with data engineers, product managers, and other team members to understand
business requirements and deliver solutions.
Clean and preprocess data to ensure accuracy and completeness for analysis.
Prepare and present reports on data analysis, model performance, and key metrics to
stakeholders and management.
Participate in regular Scrum events such as Sprint Planning, Sprint Review, and Sprint
Retrospective
Stay updated with the latest industry trends and advancements in data science and
machine learning techniques.
For This Role, You Will Need:
Bachelor’s degree in computer science, Data Science, Statistics, or a related field or a
master's degree or higher is preferred.
Total 5-7 years of industry experience
More than 3 years of experience in a data science or analytics role, with a strong track
record of building and deploying models.
Proficiency in programming languages such as Python or R, and experience with data
manipulation libraries (e.g., pandas, NumPy).
Excellent understanding of Agentic Frameworks like Microsoft Agent Framework.
Experience with NLP, NLG, and Large Language Models Open Source as well as Cloud
based models.
Experience with SQL and NoSQL databases such as MongoDB, Cassandra, Vector
databases
Experience with Dockers, Asynchronous Data Orchestrators, environments etc.
Strong analytical and problem-solving skills, with the ability to work with complex data
sets and extract actionable insights.
Excellent verbal and written communication skills, with the ability to present complex
technical information to non-technical stakeholders.
Preferred Qualifications that Set You Apart:
Prior experience in engineering domain would be nice to have
Prior experience in working with teams in Scaled Agile Framework (SAFe) is nice to
have
Possession of relevant certification/s in data science from reputed universities
specializing in AI.
Familiarity with cloud platforms, Microsoft Azure is preferred
Ability to work in a fast-paced environment and manage multiple projects simultaneously.
Strong analytical and troubleshooting skills, with the ability to resolve issues related to
model performance and infrastructure.
The recruiter has not been active on this job recently. You may apply but please expect a delayed response.
AI/ML Engineer (Open-Source LLM – BFSI Domain)
Job Title: AI/ML Engineer (Open-Source LLM – BFSI Domain)
Location: Brookefield, Bengaluru
Employment Type: Full-Time
Work Schedule: Monday to Saturday (Alternate Saturdays Off)
Working Hours: 9:00 AM – 6:00 PM (Extendable based on project requirements)
About the Role
We are seeking a talented AI/ML Engineer with hands-on experience in building, fine-tuning, and deploying Open-Source Large Language Models (LLMs) for the Banking, Financial Services, and Insurance (BFSI) domain. The ideal candidate will have practical experience developing production-ready AI solutions and a passion for leveraging Generative AI to solve real-world business challenges.
In this role, you will collaborate with Product, Engineering, and Data Science teams to design intelligent AI solutions that enhance customer experience, automate business processes, and improve operational efficiency.
Key Responsibilities
Design, develop, fine-tune, and deploy Open-Source LLMs for BFSI use cases.
Build AI-powered applications using modern LLM frameworks and orchestration tools.
Collaborate with Product Managers, Data Scientists, and Software Engineers to understand business requirements and deliver scalable AI solutions.
Conduct model experimentation, evaluation, optimization, and performance benchmarking using real-world datasets.
Implement Retrieval-Augmented Generation (RAG), prompt engineering, vector databases, and AI workflows where applicable.
Monitor model performance in production environments and continuously improve model accuracy, latency, and scalability.
Ensure AI solutions comply with enterprise security, privacy, and regulatory standards.
Develop APIs and integrate AI models into enterprise applications.
Maintain technical documentation, architecture diagrams, and deployment procedures.
Stay updated with the latest advancements in Artificial Intelligence, Machine Learning, and Open-Source LLM technologies.
Required Qualifications
Education
Bachelor's degree in Computer Science, Artificial Intelligence, Information Technology, Engineering, or a related technical field.
Equivalent practical experience will also be considered.
Experience
Minimum 2+ years of hands-on AI/ML development experience.
Experience in the BFSI domain is preferred.
Proven experience building and deploying AI/ML solutions in production environments.
Technical Skills
Programming Languages
Python (Mandatory)
Java (Preferred)
AI/ML Frameworks
PyTorch
TensorFlow
Scikit-learn
Open-Source LLM Technologies
Experience with one or more of the following:
LangChain
LangGraph
Ollama
Hugging Face Transformers
Llama
Mistral
DeepSeek
Qwen
vLLM
FastAPI
Generative AI
Prompt Engineering
Fine-Tuning LLMs
Retrieval-Augmented Generation (RAG)
Embeddings
Vector Databases (FAISS, ChromaDB, Milvus, Pinecone, etc.)
Model Evaluation and Optimization
Additional Skills
Docker
Kubernetes (Preferred)
REST APIs
Git
Linux
CI/CD pipelines
Core Competencies
Strong analytical and problem-solving skills.
Excellent understanding of AI/ML concepts and LLM architectures.
Ability to communicate technical concepts to non-technical stakeholders.
Strong interpersonal and collaboration skills.
Self-motivated with a passion for continuous learning and innovation.
Preferred Experience
Experience in any of the following areas will be an added advantage:
Banking & Financial Services applications
Fraud Detection
Credit Risk Assessment
Intelligent Document Processing
Loan Processing Automation
Customer Support Chatbots
Regulatory Compliance
OCR & Document AI
Agentic AI and Multi-Agent Systems
Interview Process
Round 1 – Technical Interview
Conducted by: Senior AI Engineer
Assessment includes:
Python Programming
Machine Learning Fundamentals
Open-Source LLMs
LangChain & RAG
Coding and Problem Sol…






