Cutshort logo
For Employers
They provide both wholesale and retail funding. PM1 logo
Data Pipeline
They provide both wholesale and retail funding. PM1

Data Pipeline at They provide both wholesale and retail funding. PM1 · Mumbai · 5 - 7 years · ₹20L - ₹25L / yr · Posted 23 Jul 2021

Multi Recruit's logo

Data Pipeline

at They provide both wholesale and retail funding. PM1

Agency job
5 - 7 yrs
₹20L - ₹25L / yr
Mumbai
Skills
ETL
Talend
OLAP
Data governance
SQL
MySQL
Redshift
Snowflake
AWS S3
Algorithm
Data & Structures
Data engineering
Data lake
  • Key responsibility is to design and develop a data pipeline including the architecture, prototyping, and development of data extraction, transformation/processing, cleansing/standardizing, and loading in Data Warehouse at real-time/near the real-time frequency. Source data can be structured, semi-structured, and/or unstructured format.
  • Provide technical expertise to design efficient data ingestion solutions to consolidate data from RDBMS, APIs, Messaging queues, weblogs, images, audios, documents, etc of  Enterprise Applications, SAAS applications, external 3rd party sites or APIs, etc through ETL/ELT, API integrations, Change Data Capture, Robotic Process Automation, Custom Python/Java Coding, etc
  • Development of complex data transformation using Talend (BigData edition), Python/Java transformation in Talend, SQL/Python/Java UDXs, AWS S3, etc to load in OLAP Data Warehouse in Structured/Semi-structured form
  • Development of data model and creating transformation logic to populate models for faster data consumption with simple SQL.
  • Implementing automated Audit & Quality assurance checks in Data Pipeline
  • Document & maintain data lineage to enable data governance
  • Coordination with BIU, IT, and other stakeholders to provide best-in-class data pipeline solutions, exposing data via APIs, loading in down streams, No-SQL Databases, etc

Requirements

  • Programming experience using Python / Java, to create functions / UDX
  • Extensive technical experience with SQL on RDBMS (Oracle/MySQL/Postgresql etc) including code optimization techniques
  • Strong ETL/ELT skillset using Talend BigData Edition. Experience in Talend CDC & MDM functionality will be an advantage.
  • Experience & expertise in implementing complex data pipelines, including semi-structured & unstructured data processing
  • Expertise to design efficient data ingestion solutions to consolidate data from RDBMS, APIs, Messaging queues, weblogs, images, audios, documents, etc of  Enterprise Applications, SAAS applications, external 3rd party sites or APIs, etc through ETL/ELT, API integrations, Change Data Capture, Robotic Process Automation, Custom Python/Java Coding, etc
  • Good understanding & working experience in OLAP Data Warehousing solutions (Redshift, Synapse, Snowflake, Teradata, Vertica, etc) and cloud-native Data Lake (S3, ADLS, BigQuery, etc) solutions
  • Familiarity with AWS tool stack for Storage & Processing. Able to recommend the right tools/solutions available to address a technical problem
  • Good knowledge of database performance and tuning, troubleshooting, query optimization, and tuning
  • Good analytical skills with the ability to synthesize data to design and deliver meaningful information
  • Good knowledge of Design, Development & Performance tuning of 3NF/Flat/Hybrid Data Model
  • Know-how on any No-SQL DB (DynamoDB, MongoDB, CosmosDB, etc) will be an advantage.
  • Ability to understand business functionality, processes, and flows
  • Good combination of technical and interpersonal skills with strong written and verbal communication; detail-oriented with the ability to work independently

Functional knowledge

  • Data Governance & Quality Assurance
  • Distributed computing
  • Linux
  • Data structures and algorithm
  • Unstructured Data Processing
Read more
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos

Similar jobs (10)

VY SYSTEMS PRIVATE LIMITED
Bengaluru (Bangalore)
5 - 7 yrs
₹4L - ₹20L / yr
Data Engineer,
skill iconPython
ETL
DevOps

Job Summary

Role Overview

We are looking for an experienced Data Engineer with strong expertise in Python, ETL, Advanced SQL, CI/CD, DevOps, and Data Analytics. The ideal candidate should have hands-on experience designing and developing scalable data pipelines, transforming large datasets, and supporting data-driven applications.

Experience with Google Cloud Platform (GCP) will be an added advantage.

Key Responsibilities

  • Design, develop, and maintain scalable ETL/ELT data pipelines using Python and SQL.
  • Develop complex and optimized SQL queries, stored procedures, and data transformations.
  • Build and maintain reliable data integration workflows across multiple data sources.
  • Perform data cleansing, validation, transformation, and quality checks.
  • Analyze data and provide insights to support business and technical requirements.
  • Implement and maintain CI/CD pipelines for data engineering applications.
  • Work with DevOps practices and tools to automate deployments, monitoring, and infrastructure processes.
  • Troubleshoot data pipeline failures, performance issues, and production incidents.
  • Optimize data processing workflows for performance, scalability, and reliability.
  • Collaborate with Data Analysts, Data Scientists, Developers, and other stakeholders.
  • Follow best practices for version control, testing, documentation, and deployment.
  • Contribute to cloud-based data engineering initiatives, preferably on GCP.

Required Skills

  • 5–7 years of hands-on experience in Data Engineering.
  • Strong programming skills in Python.
  • Strong expertise in Advanced SQL and database concepts.
  • Hands-on experience with ETL/ELT processes and data pipelines.
  • Good understanding of Data Warehousing and Data Modeling concepts.
  • Experience with CI/CD practices and tools.
  • Strong understanding of DevOps principles, automation, and deployment processes.
  • Strong data analytics and problem-solving skills.
  • Experience working with large datasets and performance optimization.
  • Good understanding of Git/version control and software development best practices.

Good to Have

  • Hands-on experience with Google Cloud Platform (GCP).
  • Exposure to GCP data services such as BigQuery, Cloud Storage, Dataflow, Composer, or Pub/Sub.
  • Experience with containerization/orchestration technologies such as Docker/Kubernetes.
  • Experience with workflow orchestration tools such as Airflow.
  • Knowledge of cloud-based data architecture and distributed data processing.

Preferred Candidate Profile

  • Strong analytical and problem-solving abilities.
  • Good communication and stakeholder management skills.
  • Ability to work independently as well as in a collaborative team environment.
  • Strong ownership of data pipelines and production systems.
  • Candidates who can join at short notice are preferred.

Mandatory Skills

 Data Engineer, Python , ETL, GCP, Advanced SQL, Strong Data Analytics skills, CICD, Devops


Read more
VY SYSTEMS PRIVATE LIMITED
Bengaluru (Bangalore)
6 - 10 yrs
₹5L - ₹20L / yr
Hadoop
PySpark

Dear Candidate,

Thanks for showing interest in the opportunity!!!

As discussed, we have an opening for ETL Developer – Hadoop & PySpark with Mphasis for the Bangalore location – Permanent.


Job Description:

We are looking for an experienced ETL Developer with strong expertise in ETL development, Hadoop, and PySpark. The candidate should have hands-on experience in data processing, building and maintaining ETL pipelines, data transformation, and handling large datasets using Big Data technologies.

Primary & Mandatory Skills:

  • ETL Development
  • Hadoop / HDFS
  • PySpark / Apache Spark
  • Data Extraction, Transformation & Loading
  • SQL and Data Processing

Good to Have:

  • Hive / Spark SQL
  • Data Pipeline Development
  • Big Data Processing
  • Python Programming

Company: Mphasis

Role: ETL Developer – Hadoop & PySpark

Experience: As per requirement

Location: Bangalore

Employment Type: Permanent

Read more
TalentOne HR Consulting LLP
Pune, Nagpur
7 - 12 yrs
₹10L - ₹15L / yr
ADF
skill iconPython
snow flake

Urgent Hiring – Senior Data Engineer

We are hiring for a Senior Data Engineer for a reputed product-based company in Pune.

Location: Pune – Magarpatta / Baner

Experience: 7–10 Years

Work Mode: 5 Days WFO

Notice Period: Immediate to 30 Days preferred

What We're Looking For

  • 7+ years of hands-on experience in Data Engineering
  • 5+ years of experience in Python
  • 4+ years of experience in Snowflake
  • 5+ years of experience in SQL
  • 5+ years of experience with ADF / Fivetran / Matillion or equivalent Data Integration tools
  • Hands-on experience with AWS / Azure
  • Experience with Airflow or equivalent orchestration tools
  • Strong experience in ETL/ELT and Data Pipelines
  • Experience with APIs, JSON, XML and Webhooks
  • Knowledge of CI/CD, Git and automated testing
  • Exposure to dbt or similar transformation tools
  • Experience in data pipeline monitoring, troubleshooting and performance optimization

Key Responsibilities

  • Design, develop and maintain scalable ETL/ELT data pipelines
  • Build batch, real-time and on-demand data processing workflows
  • Integrate data from cloud and on-premise sources
  • Ensure data quality, reliability, performance and SLA adherence
  • Work closely with Data Scientists, Analysts, DevOps and Business teams
  • Implement data engineering best practices, CI/CD and automated testing
  • Troubleshoot pipeline issues and perform root cause analysis
  • Optimize data pipelines and SQL queries for performance and cost efficiency

Why Join?

  • Opportunity to work with a reputed product-based organization
  • Work on modern data engineering and cloud technologies
  • Exposure to large-scale data platforms and business-critical data solutions
  • Collaborative and technically strong environment

Interested candidates can share their updated CV for immediate consideration.


Read more
VY SYSTEMS PRIVATE LIMITED
Dharani S
Posted by Dharani S
Bengaluru (Bangalore)
5 - 9 yrs
₹3L - ₹20L / yr
skill iconPython
DevOps
PySpark

Job Description


We are looking for an experienced Data Engineer with strong expertise in Python, ETL, SQL, CI/CD, and DevOps to design, develop, and maintain scalable data pipelines and data processing solutions.


Key Responsibilities

  • Design, develop, and maintain scalable ETL/ELT data pipelines.
  • Develop data processing solutions using Python.
  • Write complex and optimized SQL queries, stored procedures, and data transformations.
  • Build and maintain data ingestion and integration workflows.
  • Implement data quality, validation, monitoring, and error-handling processes.
  • Develop and maintain CI/CD pipelines for data engineering applications.
  • Work with DevOps tools and practices for automated build, deployment, and infrastructure management.
  • Collaborate with data analysts, data scientists, software engineers, and business teams.
  • Optimize data pipelines for performance, reliability, and scalability.
  • Troubleshoot production data issues and ensure timely resolution.
  • Follow best practices for version control, code quality, testing, and deployment.


Mandatory Skills

  • Python
  • ETL
  • SQL
  • CI/CD
  • DevOps
  • Git / Version Control
  • Strong problem-solving and debugging skills


Read more
VY SYSTEMS PRIVATE LIMITED
Hyderabad, Pune
5 - 9 yrs
₹18L - ₹20L / yr
PySpark
SQL
skill iconPython

Data Engineer Short Hiring Post


🚨 Hiring: Data Engineer

🔹 Experience: 5–9 Years

🔹 Location: Bangalore / Hyderabad

🔹 Skills: PySpark, Python, SQL, ETL, CI/CD, Data Modeling

🔹 Process: L1 Virtual → L2 F2F Karat Test

🔹 F2F: Bangalore / Hyderabad Location

🔹 Positions: Immediate requirement

⚠️ Note: Candidates must be available for F2F Karat immediately after L1.

#Hiring #DataEngineer #PySpark #Python #SQL #BangaloreJobs #HyderabadJobs #Mphasis #ImmediateJoiners

Read more
Wissen Technology
at Wissen Technology
4 recruiters
Meghana Shinde
Posted by Meghana Shinde
Pune
2.5 - 5 yrs
Best in industry
skill iconJava
Spark
Hadoop
Apache HBase
SQL
+1 more

Company Name – Wissen Technology

Group of companies in India – Wissen Technology & Wissen Infotech

Work Location – Whitefield, Bangalore


Website and Company profile:

www.wissen.com


LinkedIn Page:

https://www.linkedin.com/company/wissen-technology/


While you may already know about Wissen and the company history, here is a quick rundown for you.

 

About Wissen Technology:


·    The Wissen Group was founded in the year 2000. Wissen Technology, a part of Wissen Group, was established in the year 2015.

·    Wissen Technology is a specialized technology company that delivers high-end consulting for organizations in the Banking & Finance, Telecom, and Healthcare domains. We help clients build world class products.

·    Our workforce has highly skilled professionals, with leadership and senior management executives who have graduated from Ivy League Universities like Wharton, MIT, IITs, IIMs, and NITs and with rich work experience in some of the biggest companies in the world.

·    Wissen Technology has grown its revenues by 400% in these five years without any external funding or investments.

·    Globally present with offices US, India, UK, Australia, Mexico, and Canada.

·    We offer an array of services including Application Development, Artificial Intelligence & Machine Learning, Big Data & Analytics, Visualization & Business Intelligence, Robotic Process Automation, Cloud, Mobility, Agile & DevOps, Quality Assurance & Test Automation.

·    Wissen Technology has been certified as a Great Place to Work®.

·    Wissen Technology has been voted as the Top 20 AI/ML vendor by CIO Insider in 2020.

·    Over the years, Wissen Group has successfully delivered $650 million worth of projects for more than 20 of the Fortune 500 companies.

·    We have served client across sectors like Banking, Telecom, Healthcare, Manufacturing, and Energy. They include likes of Morgan Stanley, Goldman Sachs, MSCI, StateStreet, Flipkart, Swiggy, Trafigura, GE to name a few.


About Role :


Key Responsibilities

  • Build and maintain data transformation pipelines using java Spark
  • Develop and optimize large-scale/CPU intensive data processing using Apache Spark
  • Orchestrate workflows using Airflow
  • Implement data quality checks, testing, and monitoring for pipeline. Good to have exposer into managing metadata, cataloguing, and lineage
  • Support schema evolution, backfills, and incremental processing
  • Ensure pipelines meet SLAs for freshness, reliability, and performance
  • Expertise/working knowledge in Spark and HBase(semantic layer, virtual datasets, Reflections)


Required Skills & Qualifications

  • Strong hands-on experience with 
  • HBase
  • Apache Spark
  • Experience with HBase or similar lakehouse query engines
  • Airflow 
  • Understanding of data catalogs and lineage (e.g., OpenLineage, DataHub, Apache Polaris , openlineage)
  • Proficiency in Java
  • Experience with Git-based development and CI/CD


Nice-to-Have Skills

  • OpenTable format/Iceberg ,Apache Arrow
  • CDC-based analytics pipelines
  • Cloud platforms (AWS)
  • Kubernetes-based data platforms
Read more
VY SYSTEMS PRIVATE LIMITED
Hyderabad, Bengaluru (Bangalore)
5 - 12 yrs
₹4L - ₹22L / yr
Data engineering
skill iconPython
SQL

Job Summary

We are seeking a skilled Data Engineer to design, build, and maintain scalable data pipelines and infrastructure. The ideal candidate should have strong expertise in SQL, Python, Linux, and modern data engineering practices to support data integration, transformation, and analytics.

Key Responsibilities

  • Design, develop, and maintain ETL/ELT data pipelines.
  • Write efficient and optimized SQL queries for data extraction, transformation, and reporting.
  • Develop automation scripts using Python for data processing and workflow optimization.
  • Work with Linux environments for deployment, monitoring, and troubleshooting.
  • Ensure data quality, integrity, and reliability across data platforms.
  • Collaborate with data analysts, software engineers, and business stakeholders to deliver data solutions.
  • Monitor, troubleshoot, and optimize data pipelines for performance and scalability.
  • Implement best practices for data security, governance, and documentation.

Required Skills

  • Strong experience in Data Engineering concepts and ETL/ELT processes.
  • Proficiency in SQL, including query optimization and database design.
  • Strong programming skills in Python.
  • Hands-on experience with Linux commands, shell scripting, and system administration basics.
  • Experience with relational databases such as PostgreSQL, MySQL, SQL Server, or Oracle.
  • Familiarity with Git/version control.
  • Strong analytical and problem-solving skills.

Preferred Skills

  • Experience with cloud platforms (AWS, Azure, or GCP).
  • Knowledge of Apache Spark, Airflow, Kafka, or similar data engineering tools.
  • Experience with data warehousing solutions and big data technologies.
  • Understanding of CI/CD pipelines and containerization (Docker/Kubernetes).

Qualifications

  • Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field.
  • Relevant certifications in cloud or data engineering are an added advantage.


Read more
Amazech
Preetham SR
Posted by Preetham SR
Hyderabad, Chennai
5 - 8 yrs
₹5L - ₹13L / yr
ETL
Data Warehouse (DWH)
snowflake
skill iconPython
SQL
+1 more

Location: Hyderabad / Chennai

Experience: 5+ years

Employment type: Full-time, permanent

Work Hours: General Shift

website: www.amazech.com

 

 

Qualifications: 

  • B.E./B.Tech/M.E./M.Tech in Computer Science, Information Technology, Data Science, or related disciplines.
  • Strong academic background with relevant industry experience in Data Engineering and Data Warehousing.

 


 

Key Responsibilities:

·       Design, develop, and maintain scalable data warehouse solutions using Snowflake.

·       Write, optimize, troubleshoot, and enhance Snowflake SQL queries with a focus on performance and scalability.

·       Develop and support ETL processes using Talend to ensure reliable and efficient data movement.

·       Collaborate with business, analytics, and application teams to enable reporting, dashboards, metrics, and data exploration capabilities.

·       Perform data analysis and resolve issues across data ingestion, transformation, and reporting pipelines.

·       Debug and troubleshoot Python-based data processing scripts and automation workflows.

·       Implement best practices for data quality, testing, deployment, and code reviews.

·       Work across UI, API, and Data Warehouse layers to support end-to-end data integration and business requirements.

·       Monitor, optimize, and maintain data warehouse performance and operational stability.

·       Create and maintain technical documentation, data models, and process workflows.

 

Required Skills and Experience: 

·       Strong hands-on expertise in Snowflake Data Warehouse.

·       Advanced SQL skills with experience handling large-scale datasets.

·       Strong understanding of Data Warehousing concepts, dimensional modelling, and data architecture.

·       Hands-on experience with Analytical SQL functions, query tuning, and performance optimization.

·       Experience developing and maintaining ETL solutions using Talend.

·       Proficiency in Python for scripting, debugging, automation, and data processing.

·       Experience integrating UI, API, and Data Warehouse workflows.

·       Strong problem-solving and analytical skills.

·       Experience with testing, code reviews, and deployment best practices.

·       Excellent communication and stakeholder management skills.

Read more
Wissen Technology
at Wissen Technology
4 recruiters
Chaitanya Rajadnya
Posted by Chaitanya Rajadnya
Pune
5 - 12 yrs
Best in industry
skill iconJava
Apache Spark
ETL
skill iconAmazon Web Services (AWS)

Key Responsibilities

Build and maintain data transformation pipelines using java Spark

Develop and optimize large-scale/CPU intensive data processing using Apache Spark

Orchestrate workflows using Airflow

Implement data quality checks, testing, and monitoring for pipeline. Good to have exposer into managing metadata, cataloguing, and lineage

Support schema evolution, backfills, and incremental processing

Ensure pipelines meet SLAs for freshness, reliability, and performance

Expertise/working knowledge in Spark and HBase(semantic layer, virtual datasets, Reflections)


Required Skills & Qualifications

Strong hands-on experience with Apache Spark

Experience with HBase/SQL or similar lakehouse query engines

Airflow 

Understanding of data catalogs and lineage (e.g., OpenLineage, DataHub, Apache Polaris , openlineage)

Proficiency in Java

Experience with Git-based development and CI/CD

Read more
Wissen Technology
at Wissen Technology
4 recruiters
Anisha Jindal
Posted by Anisha Jindal
Bengaluru (Bangalore), Mumbai
5 - 14 yrs
Best in industry
Data engineering
skill iconPython
PySpark
DAX
PowerBI

Job Summary

We are looking for a skilled and experienced Data Engineer to join our growing data team. The ideal candidate will have strong expertise in Python, PySpark, Data Modeling, and Power BI, with hands-on experience in designing, developing, and optimizing scalable data solutions. The role requires working closely with business stakeholders, data architects, and analytics teams to build robust data pipelines and semantic models that enable data-driven decision-making.


Technical Skills

  • Strong hands-on experience in Python and PySpark development.
  • Expertise in building and optimizing Data Engineering solutions and ETL pipelines.
  • Strong understanding of Data Modeling concepts (Star Schema, Snowflake Schema, Dimensional Modeling).
  • Experience with Power BI Data Modeling and Semantic Layer development.
  • Proficiency in DAX (Data Analysis Expressions).
  • Experience designing and managing Semantic Models in Power BI.
  • Strong SQL skills and experience working with large datasets.
  • Knowledge of data warehousing concepts and best practices.


Preferred Skills

  • Experience with cloud platforms such as Azure, AWS, or GCP.
  • Exposure to modern data platforms like Databricks.
  • Understanding of data governance and data quality frameworks.
Read more
Why apply to jobs via Cutshort
people_solving_puzzle
Personalized job matches
Stop wasting time. Get matched with jobs that meet your skills, aspirations and preferences.
people_verifying_people
Verified hiring teams
See actual hiring teams, find common social connections or connect with them directly.
ai_chip
Move faster with AI
We use AI to get you faster responses, recommendations and unmatched user experience.
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos