Cutshort logo
For Employers
Product company logo
Python Data Engineer
Product company
Python Data Engineer

Python Data Engineer at Product company · Guindy · 3 - 6 years · ₹5L - ₹12L / yr · Posted 14 May 2024

SangatHR's logo

Python Data Engineer

at Product company

Agency job
3 - 6 yrs
₹5L - ₹12L / yr
Guindy
Skills
skill iconVue.js
skill iconAngularJS (1.x)
skill iconAngular (2+)
skill iconReact.js
skill iconJavascript
skill iconPython
MySQL
skill iconPostgreSQL
SQL server
NOSQL Databases
skill iconMongoDB
Cassandra

Python Data Engineer

Job Description:

• Design, develop, and maintain database scripts and procedures to support

application requirements.

• Collaborate with software developers to integrate database scripts with

application code.

• Troubleshoot and resolve database issues in a timely manner.

• Perform database maintenance tasks, such as backups, restores, and migrations.

• Implement data security measures to protect sensitive information.

• Develop and maintain documentation for database scripts and procedures.

• Stay up-to-date with emerging technologies and best practices in database

management.


Job Requirements:

• Bachelor’s degree in Computer Science, Information Technology, or related field.

• 3+ years of Proven experience as a Database Engineer or similar role with python

• Proficiency in SQL and scripting languages such as Python or Js.

• Strong understanding of database management systems, including relational

databases (e.g., MySQL, PostgreSQL, SQL Server) and NoSQL databases (e.g.,

MongoDB, Cassandra).

• Experience with database design principles and data modelling techniques.

• Knowledge of database optimisation techniques and performance tuning.

• Familiarity with version control systems (e.g., Git) and continuous integration

tools.

• Excellent problem-solving skills and attention to detail.

• Strong communication and collaboration skills.

Read more
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos

Similar jobs (10)

company logo
Sarika Shitole
Posted by Sarika Shitole
Remote only
2 - 6 yrs
Best in industry
Data engineering
skill iconPython
SQL
SQL Azure
Data Warehouse (DWH)

About Us


We believe the future of software development is AI-native — where engineers operate at a higher level of abstraction and quality remains non-negotiable. 

Incubyte is a software craft consultancy where the “how” of building software matters as much as the “what”.  

We partner with companies of all sizes, from helping enterprises build, scale, and modernize to early-stage founders bring their ideas to life. 

Our engineers operate in an AI-native development model, using AI as a collaborator across the SDLC to accelerate development while upholding the discipline of software craftsmanship. Guided by Software Craftsmanship and Extreme Programming practices, we build reliable, maintainable, and scalable systems with speed, without compromising quality. If this way of building software resonates with you, we’d like to talk. 


Our Guiding Principles 

These principles define how we work at Incubyte. They are non-negotiable. 


Relentless Pursuit of Quality with Pragmatism 

  We build high-quality systems without losing sight of delivery. 

Extreme Ownership 

  We take responsibility end-to-end for decisions, execution, and outcomes. 

Proactive Collaboration 

  We collaborate closely, challenge each other, and solve problems together. 

Active Pursuit of Mastery 

  We continuously improve our craft and raise our bar. 

Invite, Give, and Act on Feedback 

We seek, give, and act on feedback to get better every day. 

Ensuring Client Success 

We act as trusted partners and focus on real outcomes, not just output. 


Job Description


This is a remote position.


Experience Level

2+ years of experience in SQL, Python, and Snowflake (or equivalent cloud data warehouse), Azure Cloud services.


Role Overview


If you're a Data Craftsperson who takes pride in clean, well-tested data solutions and believes in the principles of Extreme Programming, we'd love to meet you. At Incubyte, we're a DevOps organization where developers own the entire release cycle — you'll get hands-on experience across data engineering, analytics, cloud infrastructure, and direct client communication. This role sits primarily in data engineering (80%) with a meaningful analytics component (20%), supporting our client's data systems end-to-end.


What You'll Do


  • Design, build, and maintain data pipelines and infrastructure using SQL and Python
  • Work within Snowflake to build and optimize data models supporting business use cases
  • Parse and process structured and semi-structured data (JSON, XML) from varied sources
  • Diagnose issues across raw, intermediate, and summary tables
  • Build SQL queries to support repeatable analytics use cases based on stakeholder requirements
  • Investigate and resolve data quality issues, including time-sensitive or urgent ones
  • Identify opportunities to consolidate models and maintain a single source of truth (SSOT)


Requirements


What We're Looking For


  • 2+ years of experience with SQL and relational databases, with the ability to understand complex data relationships and transformations (required)
  • 2+ years of experience with Python for data engineering tasks (required)
  • Experience with Snowflake or an equivalent cloud data warehouse (required)
  • Experience working with Snowflake Coco or any other AI tools(required) 
  • Experience parsing JSON and XML data (a plus)
  • A strong eye for data quality and attention to detail
  • Knowledge of Git (required)
  • Knowledge of Azure cloud services such as Azure Data Factory, Azure Blob Storage, and Azure SQL Database (required)
  • Knowledge of data infrastructure/modeling tools like DBT, Fivetran (a plus)
  • Experience with BI tools like Power BI(a plus, not core to this role)
  • Knowledge of Docker, Linux, Shell/Bash, and virtualization technologies (a plus)
  • Knowledge of SSIS packages (a plus)
  • Familiarity with CI/CD methodologies



Benefits


Life at Incubyte  


We are a remote-first company with structured flexibility. Teams commit to shared rhythms during core hours, ensuring smooth collaboration while maintaining autonomy. Twice a year, we come together in person for a co-working sprint and once a year for a retreat - with all travel expenses covered. 


Our environment is built for crafters: pairing, refactoring, experimenting with AI, and pushing the boundaries of software excellence. We are all lifelong learners, and our work is our passion. 


Perks


  • Dedicated learning & development budget. 
  • Sponsorship for conference talks. 
  • Comprehensive medical & term insurance.
  • Employee-friendly leave policies. 
  • Home Office fund 
  • Medical Insurance



Read more
company logo
Anisha Jindal
Posted by Anisha Jindal
Bengaluru (Bangalore), Mumbai
5 - 14 yrs
Best in industry
Data engineering
skill iconPython
PySpark
DAX
PowerBI

Job Summary

We are looking for a skilled and experienced Data Engineer to join our growing data team. The ideal candidate will have strong expertise in Python, PySpark, Data Modeling, and Power BI, with hands-on experience in designing, developing, and optimizing scalable data solutions. The role requires working closely with business stakeholders, data architects, and analytics teams to build robust data pipelines and semantic models that enable data-driven decision-making.


Technical Skills

  • Strong hands-on experience in Python and PySpark development.
  • Expertise in building and optimizing Data Engineering solutions and ETL pipelines.
  • Strong understanding of Data Modeling concepts (Star Schema, Snowflake Schema, Dimensional Modeling).
  • Experience with Power BI Data Modeling and Semantic Layer development.
  • Proficiency in DAX (Data Analysis Expressions).
  • Experience designing and managing Semantic Models in Power BI.
  • Strong SQL skills and experience working with large datasets.
  • Knowledge of data warehousing concepts and best practices.


Preferred Skills

  • Experience with cloud platforms such as Azure, AWS, or GCP.
  • Exposure to modern data platforms like Databricks.
  • Understanding of data governance and data quality frameworks.
Read more
company logo
Banu S
Posted by Banu S
Hyderabad, Bengaluru (Bangalore)
5 - 12 yrs
₹4L - ₹22L / yr
Data engineering
skill iconPython
SQL

Job Summary

We are seeking a skilled Data Engineer to design, build, and maintain scalable data pipelines and infrastructure. The ideal candidate should have strong expertise in SQL, Python, Linux, and modern data engineering practices to support data integration, transformation, and analytics.

Key Responsibilities

  • Design, develop, and maintain ETL/ELT data pipelines.
  • Write efficient and optimized SQL queries for data extraction, transformation, and reporting.
  • Develop automation scripts using Python for data processing and workflow optimization.
  • Work with Linux environments for deployment, monitoring, and troubleshooting.
  • Ensure data quality, integrity, and reliability across data platforms.
  • Collaborate with data analysts, software engineers, and business stakeholders to deliver data solutions.
  • Monitor, troubleshoot, and optimize data pipelines for performance and scalability.
  • Implement best practices for data security, governance, and documentation.

Required Skills

  • Strong experience in Data Engineering concepts and ETL/ELT processes.
  • Proficiency in SQL, including query optimization and database design.
  • Strong programming skills in Python.
  • Hands-on experience with Linux commands, shell scripting, and system administration basics.
  • Experience with relational databases such as PostgreSQL, MySQL, SQL Server, or Oracle.
  • Familiarity with Git/version control.
  • Strong analytical and problem-solving skills.

Preferred Skills

  • Experience with cloud platforms (AWS, Azure, or GCP).
  • Knowledge of Apache Spark, Airflow, Kafka, or similar data engineering tools.
  • Experience with data warehousing solutions and big data technologies.
  • Understanding of CI/CD pipelines and containerization (Docker/Kubernetes).

Qualifications

  • Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field.
  • Relevant certifications in cloud or data engineering are an added advantage.


Read more
company logo
Jancy A
Posted by Jancy A
Bengaluru (Bangalore)
5 - 7 yrs
₹4L - ₹20L / yr
Data Engineer,
skill iconPython
ETL
DevOps

Job Summary

Role Overview

We are looking for an experienced Data Engineer with strong expertise in Python, ETL, Advanced SQL, CI/CD, DevOps, and Data Analytics. The ideal candidate should have hands-on experience designing and developing scalable data pipelines, transforming large datasets, and supporting data-driven applications.

Experience with Google Cloud Platform (GCP) will be an added advantage.

Key Responsibilities

  • Design, develop, and maintain scalable ETL/ELT data pipelines using Python and SQL.
  • Develop complex and optimized SQL queries, stored procedures, and data transformations.
  • Build and maintain reliable data integration workflows across multiple data sources.
  • Perform data cleansing, validation, transformation, and quality checks.
  • Analyze data and provide insights to support business and technical requirements.
  • Implement and maintain CI/CD pipelines for data engineering applications.
  • Work with DevOps practices and tools to automate deployments, monitoring, and infrastructure processes.
  • Troubleshoot data pipeline failures, performance issues, and production incidents.
  • Optimize data processing workflows for performance, scalability, and reliability.
  • Collaborate with Data Analysts, Data Scientists, Developers, and other stakeholders.
  • Follow best practices for version control, testing, documentation, and deployment.
  • Contribute to cloud-based data engineering initiatives, preferably on GCP.

Required Skills

  • 5–7 years of hands-on experience in Data Engineering.
  • Strong programming skills in Python.
  • Strong expertise in Advanced SQL and database concepts.
  • Hands-on experience with ETL/ELT processes and data pipelines.
  • Good understanding of Data Warehousing and Data Modeling concepts.
  • Experience with CI/CD practices and tools.
  • Strong understanding of DevOps principles, automation, and deployment processes.
  • Strong data analytics and problem-solving skills.
  • Experience working with large datasets and performance optimization.
  • Good understanding of Git/version control and software development best practices.

Good to Have

  • Hands-on experience with Google Cloud Platform (GCP).
  • Exposure to GCP data services such as BigQuery, Cloud Storage, Dataflow, Composer, or Pub/Sub.
  • Experience with containerization/orchestration technologies such as Docker/Kubernetes.
  • Experience with workflow orchestration tools such as Airflow.
  • Knowledge of cloud-based data architecture and distributed data processing.

Preferred Candidate Profile

  • Strong analytical and problem-solving abilities.
  • Good communication and stakeholder management skills.
  • Ability to work independently as well as in a collaborative team environment.
  • Strong ownership of data pipelines and production systems.
  • Candidates who can join at short notice are preferred.

Mandatory Skills

 Data Engineer, Python , ETL, GCP, Advanced SQL, Strong Data Analytics skills, CICD, Devops


Read more
NBFC for Digital Lending
NBFC for Digital Lending
Agency job
via by Bisman Gill
Mumbai
3yrs+
Upto ₹45L / yr (Varies
)
SQL
Data Structures
skill iconPython
skill iconAmazon Web Services (AWS)
skill iconPostgreSQL

Must-Have Skills

  • Minimum 3 years of experience in Data Engineering / Analytics Engineering / Fintech Data roles
  • Must have worked on SMS Parsing, intelligent platform, converting RAW customer SMS data into structured actionable financial signals and enabling downstream usage of SMS derived variables
  • Must have established a continuous learning cycle to expand parser coverage
  • Experience in Lending / NBFC / Fintech domain
  • Experience working with Bureau, SMS, Device, or Banking data
  • Strong Python and SQL (production level)
  • Experience handling unstructured data (SMS, logs, JSON, APIs)
  • Experience building data pipelines, schedulers, and cron jobs
  • Strong database design and data modelling skills
  • Ability to work in a startup environment with high ownership
  • Familiarity with modern platforms like AWS, Snowflake, Google BigQuery, Redshift


Good to Have

  • Experience in STPL, especially less than 25K ticket size
  • Experience with streaming (Kafka/Kinesis) and orchestration (Airflow or Step Functions)
  • Experience with feature stores and risk analytics datasets
  • Knowledge of regex, NLP basics for SMS parsing
  • Experience supporting real-time decision engines/underwriting systems


Role Summary

This role will be responsible for owning the end-to-end data-structuring layer across the organisation. The individual will transform large volumes of raw, unstructured, and semi-structured data (such as SMS, device, bureau, and app data) into clean, standardised, and analysis-ready datasets. These structured datasets will directly power risk analytics, fraud detection, marketing insights, collections strategy, and policy decisioning.


Key Objective of the Role

Ensure all raw lending data (SMS, Bureau, Device, AA, App logs) is captured, parsed, structured, and stored in a clean analytics-ready format inside databases (PostgreSQL, DynamoDB, AWS stack) so that the Risk and Data Science team can directly use it for feature creation, policy building, and portfolio monitoring.


Core Responsibilities

  1. End-to-End Data Ownership
  • Design, build, and maintain end-to-end data pipelines (batch + streaming) using AWS native services (Glue, Lambda, Step Functions, Kinesis, S3, Athena, Redshift, EMR/Spark, etc.): ingestion

→ parsing → structuring → storage

  • Work closely with Tech, Product, and Data Science to define what data should be captured
  • Maintain data documentation, data dictionaries, and schema governance
  • Ensure data quality, consistency, and version control
  1. Unstructured Data Processing (Highest Priority)
  • Parse raw SMS dumps and categorise into salary, EMI, loan apps, collections, credits, debits, OTP, etc.
  • Process device fingerprint, behavioural logs, and vendor data (FinBox, AA, Bureau APIs)


  • Convert JSON, logs, and raw API responses into structured feature tables
  • Build regex/keyword-based parsers for financial SMS classification
  1. Feature Implementation (From Risk & Data Science Team)
  • Implement feature creation logic provided by Risk/Data Science team
  • Translate business and policy logic into SQL/Python pipelines
  • Create reusable feature layers for underwriting, fraud, collections, and monitoring
  • Maintain a feature store for consistent model and policy usage
  1. Lending Data Understanding (Domain-Specific Requirement)
  • Work with Bureau data
  • Structure SMS-derived financial variables (income, stress, EMI signals)
  • Work with Account Aggregator and bank transaction datasets
  • Understand fintech alternate data used in underwriting and fraud detection
  1. Data Pipelines & Automation
  • Build and maintain ETL/ELT pipelines using Python & SQL
  • Create cron jobs for automated data ingestion and feature refresh
  • Automate vendor data pulls (Bureau, SMS SDK, AA, device data)
  • Ensure low-latency pipelines for real-time underwriting use cases
  1. Database Structuring & Storage Architecture
  • Structure clean datasets in PostgreSQL (analytics layer)
  • Manage raw data storage in DynamoDB / S3 data lake
  • Design normalized and denormalised tables for risk analytics
  • Optimise database performance for large-scale query workloads
  1. Dashboards & Readable Data Layer
  • Create analytics-ready datasets, implement & write Metabase queries and convert into dashboards (Metabase / Power BI)
  • Enable self-serve data access for Risk, Business, and Founders
  • Support ad-hoc analysis requirements from leadership
  1. Cross-Functional Collaboration (Very Important)
  • The role requires close collaboration with data science, tech, product, and business teams to ensure reliable data pipelines, well-defined schemas, API integrations, logging architecture and high data quality, enabling faster and more accurate decision-making across lending workflows.

Tech Stack (Current Environment)

  • AWS Services
  • PostgreSQL (Primary analytics DB)
  • DynamoDB (Raw/NoSQL storage)
  • Python (Pandas, NumPy, ETL frameworks)
  • Advanced SQL
  • APIs, JSON, and Log Data Handling
Read more
company logo
Pavithra E
Posted by Pavithra E
Hyderabad, Pune
5 - 9 yrs
₹18L - ₹20L / yr
PySpark
SQL
skill iconPython

Data Engineer Short Hiring Post


🚨 Hiring: Data Engineer

🔹 Experience: 5–9 Years

🔹 Location: Bangalore / Hyderabad

🔹 Skills: PySpark, Python, SQL, ETL, CI/CD, Data Modeling

🔹 Process: L1 Virtual → L2 F2F Karat Test

🔹 F2F: Bangalore / Hyderabad Location

🔹 Positions: Immediate requirement

⚠️ Note: Candidates must be available for F2F Karat immediately after L1.

#Hiring #DataEngineer #PySpark #Python #SQL #BangaloreJobs #HyderabadJobs #Mphasis #ImmediateJoiners

Read more
company logo
Jancy A
Posted by Jancy A
Bengaluru (Bangalore)
7 - 13 yrs
₹10L - ₹24L / yr
skill iconPython
PySpark
Oracle
PL/SQL
Google Cloud Platform (GCP)
+2 more

Senior Data Engineer – PySpark & Oracle

Experience: 7+ Years

Location: Bangalore

Notice Period: Immediate to 10 Days

Key Skills:

  • Strong expertise in Data Modeling, Data Design & Modernization
  • Primary skills: PySpark, Oracle SQL/PLSQL
  • Secondary skills: Python, ETL & Data Pipelines
  • Experience with Kafka and Hadoop
  • Exposure to AWS / Azure / GCP
  • Good knowledge of Git and JIRA

Roles & Responsibilities:

  • Design, develop, and modernize scalable data models and data architecture.
  • Develop and optimize data processing solutions using PySpark and Oracle SQL/PLSQL.
  • Build and maintain robust ETL workflows and data pipelines.
  • Work with Kafka, Hadoop, and cloud platforms for data processing and integration.
  • Perform data transformation, optimization, and performance tuning.
  • Collaborate with technical teams on data design, development, testing, and deployment.


Read more
company logo
Santhanalakshmi A
Posted by Santhanalakshmi A
Bengaluru (Bangalore), Hyderabad
5 - 12 yrs
₹8L - ₹16L / yr
skill iconPython
PySpark
SQL

1st virtual , 2nd round F2F

Python pyspark, SQL, data engineer

5+yrs

Bang/hyderabad

immediate to 15days.

Read more
company logo
Dharani S
Posted by Dharani S
Bengaluru (Bangalore)
5 - 9 yrs
₹3L - ₹20L / yr
skill iconPython
DevOps
PySpark

Job Description


We are looking for an experienced Data Engineer with strong expertise in Python, ETL, SQL, CI/CD, and DevOps to design, develop, and maintain scalable data pipelines and data processing solutions.


Key Responsibilities

  • Design, develop, and maintain scalable ETL/ELT data pipelines.
  • Develop data processing solutions using Python.
  • Write complex and optimized SQL queries, stored procedures, and data transformations.
  • Build and maintain data ingestion and integration workflows.
  • Implement data quality, validation, monitoring, and error-handling processes.
  • Develop and maintain CI/CD pipelines for data engineering applications.
  • Work with DevOps tools and practices for automated build, deployment, and infrastructure management.
  • Collaborate with data analysts, data scientists, software engineers, and business teams.
  • Optimize data pipelines for performance, reliability, and scalability.
  • Troubleshoot production data issues and ensure timely resolution.
  • Follow best practices for version control, code quality, testing, and deployment.


Mandatory Skills

  • Python
  • ETL
  • SQL
  • CI/CD
  • DevOps
  • Git / Version Control
  • Strong problem-solving and debugging skills


Read more
company logo
Pavithra E
Posted by Pavithra E
Hyderabad, Bengaluru (Bangalore)
5 - 9 yrs
₹9L - ₹20L / yr
PySpark
SQL
skill iconPython
ETL

Data Engineer Hiring Post


🚨 Hiring: Data Engineer | PySpark + Python + SQL

We are looking for experienced Data Engineers to join our team!

🔹 Experience: 5 to 9 Years

🔹 Locations: Bangalore / Hyderabad

🔹 Interview Process:

• 1st Round – Virtual

• 2nd Round – Face-to-Face (Karat Test)

🔑 Key Skills:

✅ PySpark

✅ SQL

✅ Python

✅ ETL

📩 Interested candidates can share their updated resume.

#Hiring #DataEngineer #PySpark #Python #SQL #ETL #BangaloreJobs #HyderabadJobs #TechHiring #ImmediateHiring

Read more
Why apply to jobs via Cutshort
people_solving_puzzle
Personalized job matches
Stop wasting time. Get matched with jobs that meet your skills, aspirations and preferences.
people_verifying_people
Verified hiring teams
See actual hiring teams, find common social connections or connect with them directly.
ai_chip
Move faster with AI
We use AI to get you faster responses, recommendations and unmatched user experience.
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos