Cutshort logo
For Employers
Mitibase logo
Data engineer
Data engineer

Data engineer at Mitibase · Pune · 2 - 4 years · ₹6L - ₹8L / yr · Raised funding · Posted 20 Sep 2023

Mitibase's logo

Data engineer

Vaidehi Ghangurde's profile picture
Posted by Vaidehi Ghangurde
2 - 4 yrs
₹6L - ₹8L / yr
Pune
Skills
skill iconVue.js
skill iconAngularJS (1.x)
skill iconReact.js
skill iconAngular (2+)
skill iconJavascript
skill iconDjango
skill iconFlask
pandas
skill iconData Analytics
PySpark
Bachelor of Computer Science

·      The Objective:

You will play a crucial role in designing, implementing, and maintaining our data infrastructure, run tests and update the systems


·      Job function and requirements

 

o  Expert in Python, Pandas and Numpy with knowledge of Python web Framework such as Django and Flask.

o  Able to integrate multiple data sources and databases into one system.

o  Basic understanding of frontend technologies like HTML, CSS, JavaScript.

o  Able to build data pipelines.

o  Strong unit test and debugging skills.

o  Understanding of fundamental design principles behind a scalable application

o  Good understanding of RDBMS databases among Mysql or Postgresql.

o  Able to analyze and transform raw data.

 

·      About us

Mitibase helps companies find warm prospects every month that are most relevant, and then helps their team to act on those with automation. We do so by automatically tracking key accounts and contacts for job changes and relationships triggers and surfaces them as warm leads in your sales pipeline.

Read more
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos

About Mitibase

Founded :
2019
Type :
Product
Size :
0-20
Stage :
Raised funding

About

Mitibase is a relationship intelligence platform for revenue acceleration. Mitibase enables companies to harness their collective relationships network to find, manage and close higher-quality revenue opportunities more quickly. With Mitibase, revenue and growth teams can focus on engaging high-probability prospects, resulting in high opportunity conversions and competitive advantage.
Read more

Company social profiles

bloglinkedin

Similar jobs (10)

company logo
Pavithra E
Posted by Pavithra E
Hyderabad, Pune
5 - 9 yrs
₹18L - ₹20L / yr
PySpark
SQL
skill iconPython

Data Engineer Short Hiring Post


🚨 Hiring: Data Engineer

🔹 Experience: 5–9 Years

🔹 Location: Bangalore / Hyderabad

🔹 Skills: PySpark, Python, SQL, ETL, CI/CD, Data Modeling

🔹 Process: L1 Virtual → L2 F2F Karat Test

🔹 F2F: Bangalore / Hyderabad Location

🔹 Positions: Immediate requirement

⚠️ Note: Candidates must be available for F2F Karat immediately after L1.

#Hiring #DataEngineer #PySpark #Python #SQL #BangaloreJobs #HyderabadJobs #Mphasis #ImmediateJoiners

Read more
company logo
Remote only
0 - 0 yrs
₹3000 - ₹3500 / mo
SQL

About the Role

We are seeking motivated Data Engineering Interns to join our team remotely for a 3-month internship. This role is designed for students or recent graduates interested in working with data pipelines, ETL processes, and big data tools. You will gain practical experience in building scalable data solutions. While this is an unpaid internship, interns who successfully complete the program will receive a Completion Certificate and a Letter of Recommendation.

Responsibilities

  • Assist in designing and building data pipelines for structured and unstructured data.
  • Support ETL (Extract, Transform, Load) processes to prepare data for analytics.
  • Work with databases (SQL/NoSQL) for data storage and retrieval.
  • Help optimize data workflows for performance and scalability.
  • Collaborate with data scientists and analysts to ensure data quality and consistency.
  • Document workflows, schemas, and technical processes.

Requirements

  • Strong interest in data engineering, databases, and big data systems.
  • Basic knowledge of SQL and relational database concepts.
  • Familiarity with Python, Java, or Scala for data processing.
  • Understanding of ETL concepts and data pipelines.
  • Exposure to cloud platforms (AWS, Azure, or GCP) is a plus.
  • Familiarity with big data frameworks (Hadoop, Spark, Kafka) is an advantage.
  • Good problem-solving skills and ability to work independently in a remote setup.

What You’ll Gain

  • Hands-on experience in data engineering and ETL pipelines.
  • Exposure to real-world data workflows.
  • Mentorship and guidance from experienced engineers.
  • Completion Certificate upon successful completion.
  • Letter of Recommendation based on performance.

Internship Details

  • Duration: 3 months
  • Location: Remote (Work from Home)
  • Stipend: Unpaid
  • Perks: Completion Certificate + Letter of Recommendation


Read more
company logo
Banu S
Posted by Banu S
Hyderabad, Bengaluru (Bangalore)
5 - 12 yrs
₹4L - ₹22L / yr
Data engineering
skill iconPython
SQL

Job Summary

We are seeking a skilled Data Engineer to design, build, and maintain scalable data pipelines and infrastructure. The ideal candidate should have strong expertise in SQL, Python, Linux, and modern data engineering practices to support data integration, transformation, and analytics.

Key Responsibilities

  • Design, develop, and maintain ETL/ELT data pipelines.
  • Write efficient and optimized SQL queries for data extraction, transformation, and reporting.
  • Develop automation scripts using Python for data processing and workflow optimization.
  • Work with Linux environments for deployment, monitoring, and troubleshooting.
  • Ensure data quality, integrity, and reliability across data platforms.
  • Collaborate with data analysts, software engineers, and business stakeholders to deliver data solutions.
  • Monitor, troubleshoot, and optimize data pipelines for performance and scalability.
  • Implement best practices for data security, governance, and documentation.

Required Skills

  • Strong experience in Data Engineering concepts and ETL/ELT processes.
  • Proficiency in SQL, including query optimization and database design.
  • Strong programming skills in Python.
  • Hands-on experience with Linux commands, shell scripting, and system administration basics.
  • Experience with relational databases such as PostgreSQL, MySQL, SQL Server, or Oracle.
  • Familiarity with Git/version control.
  • Strong analytical and problem-solving skills.

Preferred Skills

  • Experience with cloud platforms (AWS, Azure, or GCP).
  • Knowledge of Apache Spark, Airflow, Kafka, or similar data engineering tools.
  • Experience with data warehousing solutions and big data technologies.
  • Understanding of CI/CD pipelines and containerization (Docker/Kubernetes).

Qualifications

  • Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field.
  • Relevant certifications in cloud or data engineering are an added advantage.


Read more
NBFC for Digital Lending
NBFC for Digital Lending
Agency job
via by Bisman Gill
Mumbai
3yrs+
Upto ₹45L / yr (Varies
)
SQL
Data Structures
skill iconPython
skill iconAmazon Web Services (AWS)
skill iconPostgreSQL

Must-Have Skills

  • Minimum 3 years of experience in Data Engineering / Analytics Engineering / Fintech Data roles
  • Must have worked on SMS Parsing, intelligent platform, converting RAW customer SMS data into structured actionable financial signals and enabling downstream usage of SMS derived variables
  • Must have established a continuous learning cycle to expand parser coverage
  • Experience in Lending / NBFC / Fintech domain
  • Experience working with Bureau, SMS, Device, or Banking data
  • Strong Python and SQL (production level)
  • Experience handling unstructured data (SMS, logs, JSON, APIs)
  • Experience building data pipelines, schedulers, and cron jobs
  • Strong database design and data modelling skills
  • Ability to work in a startup environment with high ownership
  • Familiarity with modern platforms like AWS, Snowflake, Google BigQuery, Redshift


Good to Have

  • Experience in STPL, especially less than 25K ticket size
  • Experience with streaming (Kafka/Kinesis) and orchestration (Airflow or Step Functions)
  • Experience with feature stores and risk analytics datasets
  • Knowledge of regex, NLP basics for SMS parsing
  • Experience supporting real-time decision engines/underwriting systems


Role Summary

This role will be responsible for owning the end-to-end data-structuring layer across the organisation. The individual will transform large volumes of raw, unstructured, and semi-structured data (such as SMS, device, bureau, and app data) into clean, standardised, and analysis-ready datasets. These structured datasets will directly power risk analytics, fraud detection, marketing insights, collections strategy, and policy decisioning.


Key Objective of the Role

Ensure all raw lending data (SMS, Bureau, Device, AA, App logs) is captured, parsed, structured, and stored in a clean analytics-ready format inside databases (PostgreSQL, DynamoDB, AWS stack) so that the Risk and Data Science team can directly use it for feature creation, policy building, and portfolio monitoring.


Core Responsibilities

  1. End-to-End Data Ownership
  • Design, build, and maintain end-to-end data pipelines (batch + streaming) using AWS native services (Glue, Lambda, Step Functions, Kinesis, S3, Athena, Redshift, EMR/Spark, etc.): ingestion

→ parsing → structuring → storage

  • Work closely with Tech, Product, and Data Science to define what data should be captured
  • Maintain data documentation, data dictionaries, and schema governance
  • Ensure data quality, consistency, and version control
  1. Unstructured Data Processing (Highest Priority)
  • Parse raw SMS dumps and categorise into salary, EMI, loan apps, collections, credits, debits, OTP, etc.
  • Process device fingerprint, behavioural logs, and vendor data (FinBox, AA, Bureau APIs)


  • Convert JSON, logs, and raw API responses into structured feature tables
  • Build regex/keyword-based parsers for financial SMS classification
  1. Feature Implementation (From Risk & Data Science Team)
  • Implement feature creation logic provided by Risk/Data Science team
  • Translate business and policy logic into SQL/Python pipelines
  • Create reusable feature layers for underwriting, fraud, collections, and monitoring
  • Maintain a feature store for consistent model and policy usage
  1. Lending Data Understanding (Domain-Specific Requirement)
  • Work with Bureau data
  • Structure SMS-derived financial variables (income, stress, EMI signals)
  • Work with Account Aggregator and bank transaction datasets
  • Understand fintech alternate data used in underwriting and fraud detection
  1. Data Pipelines & Automation
  • Build and maintain ETL/ELT pipelines using Python & SQL
  • Create cron jobs for automated data ingestion and feature refresh
  • Automate vendor data pulls (Bureau, SMS SDK, AA, device data)
  • Ensure low-latency pipelines for real-time underwriting use cases
  1. Database Structuring & Storage Architecture
  • Structure clean datasets in PostgreSQL (analytics layer)
  • Manage raw data storage in DynamoDB / S3 data lake
  • Design normalized and denormalised tables for risk analytics
  • Optimise database performance for large-scale query workloads
  1. Dashboards & Readable Data Layer
  • Create analytics-ready datasets, implement & write Metabase queries and convert into dashboards (Metabase / Power BI)
  • Enable self-serve data access for Risk, Business, and Founders
  • Support ad-hoc analysis requirements from leadership
  1. Cross-Functional Collaboration (Very Important)
  • The role requires close collaboration with data science, tech, product, and business teams to ensure reliable data pipelines, well-defined schemas, API integrations, logging architecture and high data quality, enabling faster and more accurate decision-making across lending workflows.

Tech Stack (Current Environment)

  • AWS Services
  • PostgreSQL (Primary analytics DB)
  • DynamoDB (Raw/NoSQL storage)
  • Python (Pandas, NumPy, ETL frameworks)
  • Advanced SQL
  • APIs, JSON, and Log Data Handling
Read more
company logo
Anisha Jindal
Posted by Anisha Jindal
Bengaluru (Bangalore), Mumbai
5 - 14 yrs
Best in industry
Data engineering
skill iconPython
PySpark
DAX
PowerBI

Job Summary

We are looking for a skilled and experienced Data Engineer to join our growing data team. The ideal candidate will have strong expertise in Python, PySpark, Data Modeling, and Power BI, with hands-on experience in designing, developing, and optimizing scalable data solutions. The role requires working closely with business stakeholders, data architects, and analytics teams to build robust data pipelines and semantic models that enable data-driven decision-making.


Technical Skills

  • Strong hands-on experience in Python and PySpark development.
  • Expertise in building and optimizing Data Engineering solutions and ETL pipelines.
  • Strong understanding of Data Modeling concepts (Star Schema, Snowflake Schema, Dimensional Modeling).
  • Experience with Power BI Data Modeling and Semantic Layer development.
  • Proficiency in DAX (Data Analysis Expressions).
  • Experience designing and managing Semantic Models in Power BI.
  • Strong SQL skills and experience working with large datasets.
  • Knowledge of data warehousing concepts and best practices.


Preferred Skills

  • Experience with cloud platforms such as Azure, AWS, or GCP.
  • Exposure to modern data platforms like Databricks.
  • Understanding of data governance and data quality frameworks.
Read more
Service Co
Service Co
Agency job
via by Rishika Teja
Pune
5 - 12 yrs
₹15L - ₹34L / yr
SQL
skill iconPython
skill iconData Science
Spark

Hiring for Data Scientist / Senior Data Scientist


Exp : 4 - 12 yrs

Edu : BE/B.tech/MCA

Work Location : Pune

Notice Period : Immediate - 15 days


Skills :


4+ years of experience in data engineering, data science, or related domains.


Hands-on experience with SQL, Python, and distributed data systems.


Knowledge of machine learning techniques and statistical analysis.


Experience with cloud data platforms (Azure Data Factory, AWS Glue, GCP BigQuery).


Familiarity with DevOps practices and CI/CD for data pipelines.


Platforms & Operations Experience (Preferred)

- Experience working with Azure, AWS, or Google Cloud data tools.


Operational experience with data orchestration tools (Airflow, ADF, Glue).


Understanding of Kubernetes, Docker, or containerized environments.


Hands-on experience with data warehousing platforms (Snowflake, Redshift, BigQuery).


Experience in monitoring, logging, and alerting operations for data workflows.

Read more
company logo
Jancy A
Posted by Jancy A
Bengaluru (Bangalore)
5 - 7 yrs
₹4L - ₹20L / yr
Data Engineer,
skill iconPython
ETL
DevOps

Job Summary

Role Overview

We are looking for an experienced Data Engineer with strong expertise in Python, ETL, Advanced SQL, CI/CD, DevOps, and Data Analytics. The ideal candidate should have hands-on experience designing and developing scalable data pipelines, transforming large datasets, and supporting data-driven applications.

Experience with Google Cloud Platform (GCP) will be an added advantage.

Key Responsibilities

  • Design, develop, and maintain scalable ETL/ELT data pipelines using Python and SQL.
  • Develop complex and optimized SQL queries, stored procedures, and data transformations.
  • Build and maintain reliable data integration workflows across multiple data sources.
  • Perform data cleansing, validation, transformation, and quality checks.
  • Analyze data and provide insights to support business and technical requirements.
  • Implement and maintain CI/CD pipelines for data engineering applications.
  • Work with DevOps practices and tools to automate deployments, monitoring, and infrastructure processes.
  • Troubleshoot data pipeline failures, performance issues, and production incidents.
  • Optimize data processing workflows for performance, scalability, and reliability.
  • Collaborate with Data Analysts, Data Scientists, Developers, and other stakeholders.
  • Follow best practices for version control, testing, documentation, and deployment.
  • Contribute to cloud-based data engineering initiatives, preferably on GCP.

Required Skills

  • 5–7 years of hands-on experience in Data Engineering.
  • Strong programming skills in Python.
  • Strong expertise in Advanced SQL and database concepts.
  • Hands-on experience with ETL/ELT processes and data pipelines.
  • Good understanding of Data Warehousing and Data Modeling concepts.
  • Experience with CI/CD practices and tools.
  • Strong understanding of DevOps principles, automation, and deployment processes.
  • Strong data analytics and problem-solving skills.
  • Experience working with large datasets and performance optimization.
  • Good understanding of Git/version control and software development best practices.

Good to Have

  • Hands-on experience with Google Cloud Platform (GCP).
  • Exposure to GCP data services such as BigQuery, Cloud Storage, Dataflow, Composer, or Pub/Sub.
  • Experience with containerization/orchestration technologies such as Docker/Kubernetes.
  • Experience with workflow orchestration tools such as Airflow.
  • Knowledge of cloud-based data architecture and distributed data processing.

Preferred Candidate Profile

  • Strong analytical and problem-solving abilities.
  • Good communication and stakeholder management skills.
  • Ability to work independently as well as in a collaborative team environment.
  • Strong ownership of data pipelines and production systems.
  • Candidates who can join at short notice are preferred.

Mandatory Skills

 Data Engineer, Python , ETL, GCP, Advanced SQL, Strong Data Analytics skills, CICD, Devops


Read more
company logo
Dharani S
Posted by Dharani S
Bengaluru (Bangalore)
5 - 9 yrs
₹3L - ₹20L / yr
skill iconPython
DevOps
PySpark

Job Description


We are looking for an experienced Data Engineer with strong expertise in Python, ETL, SQL, CI/CD, and DevOps to design, develop, and maintain scalable data pipelines and data processing solutions.


Key Responsibilities

  • Design, develop, and maintain scalable ETL/ELT data pipelines.
  • Develop data processing solutions using Python.
  • Write complex and optimized SQL queries, stored procedures, and data transformations.
  • Build and maintain data ingestion and integration workflows.
  • Implement data quality, validation, monitoring, and error-handling processes.
  • Develop and maintain CI/CD pipelines for data engineering applications.
  • Work with DevOps tools and practices for automated build, deployment, and infrastructure management.
  • Collaborate with data analysts, data scientists, software engineers, and business teams.
  • Optimize data pipelines for performance, reliability, and scalability.
  • Troubleshoot production data issues and ensure timely resolution.
  • Follow best practices for version control, code quality, testing, and deployment.


Mandatory Skills

  • Python
  • ETL
  • SQL
  • CI/CD
  • DevOps
  • Git / Version Control
  • Strong problem-solving and debugging skills


Read more
company logo
Pavithra E
Posted by Pavithra E
Hyderabad, Bengaluru (Bangalore)
5 - 9 yrs
₹9L - ₹20L / yr
PySpark
SQL
skill iconPython
ETL

Data Engineer Hiring Post


🚨 Hiring: Data Engineer | PySpark + Python + SQL

We are looking for experienced Data Engineers to join our team!

🔹 Experience: 5 to 9 Years

🔹 Locations: Bangalore / Hyderabad

🔹 Interview Process:

• 1st Round – Virtual

• 2nd Round – Face-to-Face (Karat Test)

🔑 Key Skills:

✅ PySpark

✅ SQL

✅ Python

✅ ETL

📩 Interested candidates can share their updated resume.

#Hiring #DataEngineer #PySpark #Python #SQL #ETL #BangaloreJobs #HyderabadJobs #TechHiring #ImmediateHiring

Read more
company logo
Bengaluru (Bangalore)
14 - 25 yrs
₹50L - ₹70L / yr
Data engineering
databricks
Apache Spark
PySpark
skill iconPython
+19 more

Job Title : Senior Data Engineer – Databricks

Experience : 14 to 20 Years

Location : HSR Layout, Bangalore

Work Mode : Hybrid – 3 Days WFO

Shift : 11:30 AM – 07:30 PM IST

Positions : 2

Notice Period : Immediate Joiners Only

Interview : 1 Technical Round + 2 Client Rounds


Role Overview :

We are looking for a Senior Data Engineer to build and lead enterprise-scale data platforms for a Switzerland-based commodity client.

The role requires a strong hands-on Data Engineering professional with expertise in Databricks, PySpark, Python, SQL, and AWS, along with technical leadership and stakeholder management experience.


Must-Have Skills :

  • 14 to 20 years of Data Engineering experience
  • Databricks & Apache Spark / PySpark
  • Python & SQL
  • AWS Cloud
  • Lakehouse Architecture
  • ETL / ELT & Distributed Data Processing
  • Batch & Streaming Pipelines
  • Data Pipeline Optimization & Data Modeling
  • CDC & Incremental Processing
  • Git, CI/CD & Testing
  • Data Quality, Monitoring & Observability
  • Technical Leadership & Stakeholder Management


Key Responsibilities :

  • Design and build scalable data pipelines using Databricks, PySpark, Python, SQL, and AWS.
  • Own data products from design through production.
  • Develop batch / streaming pipelines and reusable ETL / ELT frameworks.
  • Optimize pipelines for performance, scalability, reliability, and cost.
  • Design scalable data architectures and data models.
  • Implement data quality, monitoring, lineage, and CI/CD practices.
  • Lead technical discussions and mentor engineering teams.
  • Collaborate with business stakeholders, architects, product owners, and engineering teams.
  • Remain hands-on while providing technical leadership.


Ideal Candidate :

A 14 to 20 years experienced, hands-on Data Engineering leader with strong Databricks + PySpark + AWS expertise, excellent communication, stakeholder management, and experience delivering enterprise-scale data platforms.

🔴 Super Urgent : Only Bangalore-based immediate joiners.

Read more
Why apply to jobs via Cutshort
people_solving_puzzle
Personalized job matches
Stop wasting time. Get matched with jobs that meet your skills, aspirations and preferences.
people_verifying_people
Verified hiring teams
See actual hiring teams, find common social connections or connect with them directly.
ai_chip
Move faster with AI
We use AI to get you faster responses, recommendations and unmatched user experience.
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos