Data Scientist at One of our Premium Client · Chennai · 3 - 8 years · ₹3L - ₹17L / yr · Posted 28 Oct 2022

Job Description – Data Science
Basic Qualification:
- ME/MS from premier institute with a background in Mechanical/Industrial/Chemical/Materials engineering.
- Strong Analytical skills and application of Statistical techniques to problem solving
- Expertise in algorithms, data structures and performance optimization techniques
- Proven track record of demonstrating end to end ownership involving taking an idea from incubator to market
- Minimum years of experience in data analysis (2+), statistical analysis, data mining, algorithms for optimization.
Responsibilities
The Data Engineer/Analyst will
- Work with stakeholders throughout the organization to identify opportunities for leveraging company data to drive business solutions.
- Clear interaction with Business teams including product planning, sales, marketing, finance for defining the projects, objectives.
- Mine and analyze data from company databases to drive optimization and improvement of product and process development, marketing techniques and business strategies
- Coordinate with different R&D and Business teams to implement models and monitor outcomes.
- Mentor team members towards developing quick solutions for business impact.
- Skilled at all stages of the analysis process including defining key business questions, recommending measures, data sources, methodology and study design, dataset creation, analysis execution, interpretation and presentation and publication of results.
- 4+ years’ experience in MNC environment with projects involving ML, DL and/or DS
- Experience in Machine Learning, Data Mining or Machine Intelligence (Artificial Intelligence)
- Knowledge on Microsoft Azure will be desired.
- Expertise in machine learning such as Classification, Data/Text Mining, NLP, Image Processing, Decision Trees, Random Forest, Neural Networks, Deep Learning Algorithms
- Proficient in Python and its various libraries such as Numpy, MatPlotLib, Pandas
- Superior verbal and written communication skills, ability to convey rigorous mathematical concepts and considerations to Business Teams.
- Experience in infra development / building platforms is highly desired.
- A drive to learn and master new technologies and techniques.

Similar jobs (10)
Sr Engineer – Artificial Intelligence
Job Summary
As an AI Engineer at Emerson, you will be responsible for analysing complex data sets to
identify trends, develop predictive models, and provide actionable insights. You will work closely
with cross-functional teams to understand business needs and deliver data-driven solutions that
enhance decision-making and drive business growth.
In This Role, Your Responsibilities Will Be:
Analyze large, complex data sets using statistical methods and machine learning
techniques to extract meaningful insights.
Develop and implement predictive models and algorithms to solve business problems
and improve processes.
Create visualizations and dashboards to effectively communicate findings and insights to
stakeholders.
Work with data engineers, product managers, and other team members to understand
business requirements and deliver solutions.
Clean and preprocess data to ensure accuracy and completeness for analysis.
Prepare and present reports on data analysis, model performance, and key metrics to
stakeholders and management.
Participate in regular Scrum events such as Sprint Planning, Sprint Review, and Sprint
Retrospective
Stay updated with the latest industry trends and advancements in data science and
machine learning techniques.
For This Role, You Will Need:
Bachelor’s degree in computer science, Data Science, Statistics, or a related field or a
master's degree or higher is preferred.
Total 5-7 years of industry experience
More than 3 years of experience in a data science or analytics role, with a strong track
record of building and deploying models.
Proficiency in programming languages such as Python or R, and experience with data
manipulation libraries (e.g., pandas, NumPy).
Excellent understanding of Agentic Frameworks like Microsoft Agent Framework.
Experience with NLP, NLG, and Large Language Models Open Source as well as Cloud
based models.
Experience with SQL and NoSQL databases such as MongoDB, Cassandra, Vector
databases
Experience with Dockers, Asynchronous Data Orchestrators, environments etc.
Strong analytical and problem-solving skills, with the ability to work with complex data
sets and extract actionable insights.
Excellent verbal and written communication skills, with the ability to present complex
technical information to non-technical stakeholders.
Preferred Qualifications that Set You Apart:
Prior experience in engineering domain would be nice to have
Prior experience in working with teams in Scaled Agile Framework (SAFe) is nice to
have
Possession of relevant certification/s in data science from reputed universities
specializing in AI.
Familiarity with cloud platforms, Microsoft Azure is preferred
Ability to work in a fast-paced environment and manage multiple projects simultaneously.
Strong analytical and troubleshooting skills, with the ability to resolve issues related to
model performance and infrastructure.
Job Summary
We are looking for a skilled and experienced Data Engineer to join our growing data team. The ideal candidate will have strong expertise in Python, PySpark, Data Modeling, and Power BI, with hands-on experience in designing, developing, and optimizing scalable data solutions. The role requires working closely with business stakeholders, data architects, and analytics teams to build robust data pipelines and semantic models that enable data-driven decision-making.
Technical Skills
- Strong hands-on experience in Python and PySpark development.
- Expertise in building and optimizing Data Engineering solutions and ETL pipelines.
- Strong understanding of Data Modeling concepts (Star Schema, Snowflake Schema, Dimensional Modeling).
- Experience with Power BI Data Modeling and Semantic Layer development.
- Proficiency in DAX (Data Analysis Expressions).
- Experience designing and managing Semantic Models in Power BI.
- Strong SQL skills and experience working with large datasets.
- Knowledge of data warehousing concepts and best practices.
Preferred Skills
- Experience with cloud platforms such as Azure, AWS, or GCP.
- Exposure to modern data platforms like Databricks.
- Understanding of data governance and data quality frameworks.
Job Summary
Role Overview
We are looking for an experienced Data Engineer with strong expertise in Python, ETL, Advanced SQL, CI/CD, DevOps, and Data Analytics. The ideal candidate should have hands-on experience designing and developing scalable data pipelines, transforming large datasets, and supporting data-driven applications.
Experience with Google Cloud Platform (GCP) will be an added advantage.
Key Responsibilities
- Design, develop, and maintain scalable ETL/ELT data pipelines using Python and SQL.
- Develop complex and optimized SQL queries, stored procedures, and data transformations.
- Build and maintain reliable data integration workflows across multiple data sources.
- Perform data cleansing, validation, transformation, and quality checks.
- Analyze data and provide insights to support business and technical requirements.
- Implement and maintain CI/CD pipelines for data engineering applications.
- Work with DevOps practices and tools to automate deployments, monitoring, and infrastructure processes.
- Troubleshoot data pipeline failures, performance issues, and production incidents.
- Optimize data processing workflows for performance, scalability, and reliability.
- Collaborate with Data Analysts, Data Scientists, Developers, and other stakeholders.
- Follow best practices for version control, testing, documentation, and deployment.
- Contribute to cloud-based data engineering initiatives, preferably on GCP.
Required Skills
- 5–7 years of hands-on experience in Data Engineering.
- Strong programming skills in Python.
- Strong expertise in Advanced SQL and database concepts.
- Hands-on experience with ETL/ELT processes and data pipelines.
- Good understanding of Data Warehousing and Data Modeling concepts.
- Experience with CI/CD practices and tools.
- Strong understanding of DevOps principles, automation, and deployment processes.
- Strong data analytics and problem-solving skills.
- Experience working with large datasets and performance optimization.
- Good understanding of Git/version control and software development best practices.
Good to Have
- Hands-on experience with Google Cloud Platform (GCP).
- Exposure to GCP data services such as BigQuery, Cloud Storage, Dataflow, Composer, or Pub/Sub.
- Experience with containerization/orchestration technologies such as Docker/Kubernetes.
- Experience with workflow orchestration tools such as Airflow.
- Knowledge of cloud-based data architecture and distributed data processing.
Preferred Candidate Profile
- Strong analytical and problem-solving abilities.
- Good communication and stakeholder management skills.
- Ability to work independently as well as in a collaborative team environment.
- Strong ownership of data pipelines and production systems.
- Candidates who can join at short notice are preferred.
Mandatory Skills
Data Engineer, Python , ETL, GCP, Advanced SQL, Strong Data Analytics skills, CICD, Devops
This role will be permanent with NAM info and deploy to client location Hyderabad & Pune.
Work Mode: WORK FROM OFFICE
Role Descriptions:
- Perform detailed data analysis and support business decision-making
- Gather and document business requirements and translate them into technical specifications
- Work closely with stakeholders to define data needs and reporting requirements
- Create user stories, functional specifications, and support UAT activities
- Ensure alignment between business objectives and data solutions
Required Skills:
- Strong expertise in SQL and data querying
- Proven experience in data analysis, requirement gathering, and stakeholder management
- Ability to translate business requirements into technical solutions and user stories
- Good understanding of data models, reporting, and analytics concepts
Skills: Business Analysis~ORACLE SQL
Locations: ~HYDERABAD~PUNE~
Desire candidate
- Candidate should have valid PF.
We are looking for a dynamic Data Engineer to join our team of technology enthusiasts. You will leverage data to drive strategic decision-making and pioneering solutions, working with complex datasets, collaborating closely with stakeholders, and transforming data into actionable insights to drive innovation.
Qualifications and Skills:
- Minimum 4 years of experience as a Data Engineer
- Hands-on experience with Azure cloud-based data solutions
- Fabric experience is a must – designing, implementing, and managing data workflows and pipelines
- Expertise in database design and management, including SQL databases such as SQL Server
- Proficient in ETL (Extract, Transform, Load) design for data integration and processing
- Strong knowledge of data modeling principles and techniques
- Experience with Azure Data Factory (ADF) for orchestrating data workflows
- Ability to analyze and translate data into actionable insights, reports, and visualizations
- Proficiency in Power BI for reporting and data visualization
Desirable Skills:
- Experience with Power BI Report Builder / Reporting Services
- Knowledge of statistical analysis or Data Science
- Experience within the UK Insurance industry is a plus
- Python or R coding skills
Responsibilities:
- Implement efficient data exchange between internal and external systems to increase efficiency and reduce re-keying and translation errors
- Support the Broking business by developing high-quality information resources, ensuring data availability and accessibility for decision-making
- Engineer data inputs and outputs from core applications and semi-structured remote service data through data syncs between data lake, ODS (SQL database), and leveraging Fabric and ADF
- Perform data engineering tasks including ingestion, cleansing, and collation from a wide range of internal and external sources
- Implement different methods of streaming data and create reconciliations for datasets
- Build analytical models to support reporting and analytics
- Collaborate with an agile delivery team to work on the backlog of specified work
Must-Have Skills
- Minimum 3 years of experience in Data Engineering / Analytics Engineering / Fintech Data roles
- Must have worked on SMS Parsing, intelligent platform, converting RAW customer SMS data into structured actionable financial signals and enabling downstream usage of SMS derived variables
- Must have established a continuous learning cycle to expand parser coverage
- Experience in Lending / NBFC / Fintech domain
- Experience working with Bureau, SMS, Device, or Banking data
- Strong Python and SQL (production level)
- Experience handling unstructured data (SMS, logs, JSON, APIs)
- Experience building data pipelines, schedulers, and cron jobs
- Strong database design and data modelling skills
- Ability to work in a startup environment with high ownership
- Familiarity with modern platforms like AWS, Snowflake, Google BigQuery, Redshift
Good to Have
- Experience in STPL, especially less than 25K ticket size
- Experience with streaming (Kafka/Kinesis) and orchestration (Airflow or Step Functions)
- Experience with feature stores and risk analytics datasets
- Knowledge of regex, NLP basics for SMS parsing
- Experience supporting real-time decision engines/underwriting systems
Role Summary
This role will be responsible for owning the end-to-end data-structuring layer across the organisation. The individual will transform large volumes of raw, unstructured, and semi-structured data (such as SMS, device, bureau, and app data) into clean, standardised, and analysis-ready datasets. These structured datasets will directly power risk analytics, fraud detection, marketing insights, collections strategy, and policy decisioning.
Key Objective of the Role
Ensure all raw lending data (SMS, Bureau, Device, AA, App logs) is captured, parsed, structured, and stored in a clean analytics-ready format inside databases (PostgreSQL, DynamoDB, AWS stack) so that the Risk and Data Science team can directly use it for feature creation, policy building, and portfolio monitoring.
Core Responsibilities
- End-to-End Data Ownership
- Design, build, and maintain end-to-end data pipelines (batch + streaming) using AWS native services (Glue, Lambda, Step Functions, Kinesis, S3, Athena, Redshift, EMR/Spark, etc.): ingestion
→ parsing → structuring → storage
- Work closely with Tech, Product, and Data Science to define what data should be captured
- Maintain data documentation, data dictionaries, and schema governance
- Ensure data quality, consistency, and version control
- Unstructured Data Processing (Highest Priority)
- Parse raw SMS dumps and categorise into salary, EMI, loan apps, collections, credits, debits, OTP, etc.
- Process device fingerprint, behavioural logs, and vendor data (FinBox, AA, Bureau APIs)
- Convert JSON, logs, and raw API responses into structured feature tables
- Build regex/keyword-based parsers for financial SMS classification
- Feature Implementation (From Risk & Data Science Team)
- Implement feature creation logic provided by Risk/Data Science team
- Translate business and policy logic into SQL/Python pipelines
- Create reusable feature layers for underwriting, fraud, collections, and monitoring
- Maintain a feature store for consistent model and policy usage
- Lending Data Understanding (Domain-Specific Requirement)
- Work with Bureau data
- Structure SMS-derived financial variables (income, stress, EMI signals)
- Work with Account Aggregator and bank transaction datasets
- Understand fintech alternate data used in underwriting and fraud detection
- Data Pipelines & Automation
- Build and maintain ETL/ELT pipelines using Python & SQL
- Create cron jobs for automated data ingestion and feature refresh
- Automate vendor data pulls (Bureau, SMS SDK, AA, device data)
- Ensure low-latency pipelines for real-time underwriting use cases
- Database Structuring & Storage Architecture
- Structure clean datasets in PostgreSQL (analytics layer)
- Manage raw data storage in DynamoDB / S3 data lake
- Design normalized and denormalised tables for risk analytics
- Optimise database performance for large-scale query workloads
- Dashboards & Readable Data Layer
- Create analytics-ready datasets, implement & write Metabase queries and convert into dashboards (Metabase / Power BI)
- Enable self-serve data access for Risk, Business, and Founders
- Support ad-hoc analysis requirements from leadership
- Cross-Functional Collaboration (Very Important)
- The role requires close collaboration with data science, tech, product, and business teams to ensure reliable data pipelines, well-defined schemas, API integrations, logging architecture and high data quality, enabling faster and more accurate decision-making across lending workflows.
Tech Stack (Current Environment)
- AWS Services
- PostgreSQL (Primary analytics DB)
- DynamoDB (Raw/NoSQL storage)
- Python (Pandas, NumPy, ETL frameworks)
- Advanced SQL
- APIs, JSON, and Log Data Handling
Job Summary
We are seeking a skilled Data Engineer to design, build, and maintain scalable data pipelines and infrastructure. The ideal candidate should have strong expertise in SQL, Python, Linux, and modern data engineering practices to support data integration, transformation, and analytics.
Key Responsibilities
- Design, develop, and maintain ETL/ELT data pipelines.
- Write efficient and optimized SQL queries for data extraction, transformation, and reporting.
- Develop automation scripts using Python for data processing and workflow optimization.
- Work with Linux environments for deployment, monitoring, and troubleshooting.
- Ensure data quality, integrity, and reliability across data platforms.
- Collaborate with data analysts, software engineers, and business stakeholders to deliver data solutions.
- Monitor, troubleshoot, and optimize data pipelines for performance and scalability.
- Implement best practices for data security, governance, and documentation.
Required Skills
- Strong experience in Data Engineering concepts and ETL/ELT processes.
- Proficiency in SQL, including query optimization and database design.
- Strong programming skills in Python.
- Hands-on experience with Linux commands, shell scripting, and system administration basics.
- Experience with relational databases such as PostgreSQL, MySQL, SQL Server, or Oracle.
- Familiarity with Git/version control.
- Strong analytical and problem-solving skills.
Preferred Skills
- Experience with cloud platforms (AWS, Azure, or GCP).
- Knowledge of Apache Spark, Airflow, Kafka, or similar data engineering tools.
- Experience with data warehousing solutions and big data technologies.
- Understanding of CI/CD pipelines and containerization (Docker/Kubernetes).
Qualifications
- Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field.
- Relevant certifications in cloud or data engineering are an added advantage.
The recruiter has not been active on this job recently. You may apply but please expect a delayed response.
About AuxoAI:
AuxoAI is a global platform-based services firm. We help companies—turn their strategies into practical digital and AI solutions. By understanding how our clients make decisions, we use digital and Artificial Intelligence (AI) technologies to drive growth, enhance their operations, improve customer experiences, and provide clear, actionable insights from their data. What We Do We work across various industries such as healthcare, high-tech, consumer packaged goods (CPG), finance etc., and in sales, marketing, and customer support functions.
We help our clients with accelerating their digital and AI journeys through:
• AI Application Development
• Data, Digital and Cloud acceleration using AI
• AI Native Product Engineering
We are seeking a skilled and experienced Data Engineer to join our dynamic team. The ideal candidate will have 6+ years of prior experience in data engineering, with a strong background in AWS (Amazon Web Services) technologies. This role offers an exciting opportunity to work on diverse projects, collaborating with cross-functional teams to design, build, and optimize data pipelines and infrastructure.
Responsibilities:
* Design, develop, and maintain scalable data pipelines and ETL processes leveraging AWS services such as S3, Glue, EMR, Lambda, and Redshift.
* Collaborate with data scientists and analysts to understand data requirements and implement solutions that support analytics and machine learning initiatives.
* Optimize data storage and retrieval mechanisms to ensure performance, reliability, and cost-effectiveness.
* Implement data governance and security best practices to ensure compliance and data integrity.
* Troubleshoot and debug data pipeline issues, providing timely resolution and proactive monitoring.
* Stay abreast of emerging technologies and industry trends, recommending innovative solutions to enhance data engineering capabilities.
Requirements :
* Bachelor's or Master's degree in Computer Science, Engineering, or a related field.
* 6+ years of prior experience in data engineering, with a focus on designing and building data pipelines.
* Proficiency in AWS services, particularly S3, Glue, EMR, Lambda, and Redshift.
* Strong programming skills in languages such as Python, Java, or Scala.
* Experience with SQL and NoSQL databases, data warehousing concepts, and big data technologies.
* Familiarity with containerization technologies (e.g., Docker, Kubernetes) and orchestration tools (e.g., Apache Airflow) is a plus.
We're hiring a Data Analyst to turn raw data into clear, actionable insight. You'll partner with product, sales, marketing, and operations teams to answer their most important questions with data — building dashboards, running deep-dive analyses, and defining the metrics the business runs on. The ideal candidate is fluent in SQL, comfortable wrangling messy datasets, and just as strong at telling the story behind the numbers as they are at producing them. You'll own the accuracy and trustworthiness of the reporting stakeholders rely on to make decisions.
Key Responsibilities
- Design, build, and maintain dashboards and recurring reports across business functions
- Write and optimize SQL queries to extract, join, and transform data from multiple sources
- Run ad-hoc and deep-dive analyses to answer specific business questions
- Define, document, and standardize metrics and KPIs alongside stakeholders
- Investigate data-quality issues and ensure the numbers people see are accurate
- Translate analysis into clear recommendations and present them to non-technical audiences
- Support experimentation and A/B test analysis where relevant
- Automate repetitive reporting to free up time for higher-value analysis
Requirements
- 3+ years in a data analyst or business-intelligence role
- Strong SQL and hands-on experience with Power BI and/or Tableau
- Working knowledge of Python (pandas/numpy) for analysis
- Solid grounding in statistics and analytical methods
- Advanced Excel and strong data-visualization / storytelling skills
- Ability to work independently with stakeholders across functions
Nice to have
- Experience with cloud data warehouses (BigQuery, Snowflake, Redshift)
- Exposure to dbt or other ETL/transformation tooling
Data Engineer Short Hiring Post
🚨 Hiring: Data Engineer
🔹 Experience: 5–9 Years
🔹 Location: Bangalore / Hyderabad
🔹 Skills: PySpark, Python, SQL, ETL, CI/CD, Data Modeling
🔹 Process: L1 Virtual → L2 F2F Karat Test
🔹 F2F: Bangalore / Hyderabad Location
🔹 Positions: Immediate requirement
⚠️ Note: Candidates must be available for F2F Karat immediately after L1.
#Hiring #DataEngineer #PySpark #Python #SQL #BangaloreJobs #HyderabadJobs #Mphasis #ImmediateJoiners







