Cutshort logo
For Employers
Numerator logo
Data Engineer

Data Engineer at Numerator · Remote, Pune · 3 - 9 years · ₹5L - ₹20L / yr · Profitable · Remote friendly · Posted 24 Dec 2021

Numerator's logo

Data Engineer

Ketaki Kambale's profile picture
Posted by Ketaki Kambale
3 - 9 yrs
₹5L - ₹20L / yr
Remote, Pune
Skills
Data Warehouse (DWH)
Informatica
ETL
skill iconPython
SQL
Datawarehousing

We’re hiring a talented Data Engineer and Big Data enthusiast to work in our platform to help ensure that our data quality is flawless.  As a company, we have millions of new data points every day that come into our system. You will be working with a passionate team of engineers to solve challenging problems and ensure that we can deliver the best data to our customers, on-time. You will be using the latest cloud data warehouse technology to build robust and reliable data pipelines.

Duties/Responsibilities Include:

  •  Develop expertise in the different upstream data stores and systems across Numerator.
  • Design, develop and maintain data integration pipelines for Numerators growing data sets and product offerings.
  • Build testing and QA plans for data pipelines.
  • Build data validation testing frameworks to ensure high data quality and integrity.
  • Write and maintain documentation on data pipelines and schemas
 

Requirements:

  • BS or MS in Computer Science or related field of study
  • 3 + years of experience in the data warehouse space
  • Expert in SQL, including advanced analytical queries
  • Proficiency in Python (data structures, algorithms, object oriented programming, using API’s)
  • Experience working with a cloud data warehouse (Redshift, Snowflake, Vertica)
  • Experience with a data pipeline scheduling framework (Airflow)
  • Experience with schema design and data modeling

Exceptional candidates will have:

  • Amazon Web Services (EC2, DMS, RDS) experience
  • Terraform and/or ansible (or similar) for infrastructure deployment
  • Airflow -- Experience building and monitoring DAGs, developing custom operators, using script templating solutions.
  • Experience supporting production systems in an on-call environment
Read more
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos

About Numerator

Founded :
2018
Type :
Product
Size :
500-1000
Stage :
Profitable

About

Numerator backed by Vista Equity Partners is a Data-Tech company reinventing Market Research. Headquartered in Chicago, USA, Numerator has more than 1,600 employees worldwide. We blend proprietary data with advanced technology and elite services to create unique insights in a market research industry that has been slow to change. The majority of Fortune 100 companies are Numerator clients.

Read more

Photos

Company featured pictures
Company featured pictures
Company featured pictures
Company featured pictures
Company featured pictures
Company featured pictures

Connect with the team

Profile picture
Ketaki Kambale
Profile picture
Spurthi Mangalwedhe
Profile picture
Chirag Bhayani
Profile picture
Darshit Kanani

Company social profiles

bloglinkedintwitterfacebook

Similar jobs (10)

company logo
Sarika Shitole
Posted by Sarika Shitole
Remote only
2 - 6 yrs
Best in industry
Data engineering
skill iconPython
SQL
SQL Azure
Data Warehouse (DWH)

About Us


We believe the future of software development is AI-native — where engineers operate at a higher level of abstraction and quality remains non-negotiable. 

Incubyte is a software craft consultancy where the “how” of building software matters as much as the “what”.  

We partner with companies of all sizes, from helping enterprises build, scale, and modernize to early-stage founders bring their ideas to life. 

Our engineers operate in an AI-native development model, using AI as a collaborator across the SDLC to accelerate development while upholding the discipline of software craftsmanship. Guided by Software Craftsmanship and Extreme Programming practices, we build reliable, maintainable, and scalable systems with speed, without compromising quality. If this way of building software resonates with you, we’d like to talk. 


Our Guiding Principles 

These principles define how we work at Incubyte. They are non-negotiable. 


Relentless Pursuit of Quality with Pragmatism 

  We build high-quality systems without losing sight of delivery. 

Extreme Ownership 

  We take responsibility end-to-end for decisions, execution, and outcomes. 

Proactive Collaboration 

  We collaborate closely, challenge each other, and solve problems together. 

Active Pursuit of Mastery 

  We continuously improve our craft and raise our bar. 

Invite, Give, and Act on Feedback 

We seek, give, and act on feedback to get better every day. 

Ensuring Client Success 

We act as trusted partners and focus on real outcomes, not just output. 


Job Description


This is a remote position.


Experience Level

2+ years of experience in SQL, Python, and Snowflake (or equivalent cloud data warehouse), Azure Cloud services.


Role Overview


If you're a Data Craftsperson who takes pride in clean, well-tested data solutions and believes in the principles of Extreme Programming, we'd love to meet you. At Incubyte, we're a DevOps organization where developers own the entire release cycle — you'll get hands-on experience across data engineering, analytics, cloud infrastructure, and direct client communication. This role sits primarily in data engineering (80%) with a meaningful analytics component (20%), supporting our client's data systems end-to-end.


What You'll Do


  • Design, build, and maintain data pipelines and infrastructure using SQL and Python
  • Work within Snowflake to build and optimize data models supporting business use cases
  • Parse and process structured and semi-structured data (JSON, XML) from varied sources
  • Diagnose issues across raw, intermediate, and summary tables
  • Build SQL queries to support repeatable analytics use cases based on stakeholder requirements
  • Investigate and resolve data quality issues, including time-sensitive or urgent ones
  • Identify opportunities to consolidate models and maintain a single source of truth (SSOT)


Requirements


What We're Looking For


  • 2+ years of experience with SQL and relational databases, with the ability to understand complex data relationships and transformations (required)
  • 2+ years of experience with Python for data engineering tasks (required)
  • Experience with Snowflake or an equivalent cloud data warehouse (required)
  • Experience working with Snowflake Coco or any other AI tools(required) 
  • Experience parsing JSON and XML data (a plus)
  • A strong eye for data quality and attention to detail
  • Knowledge of Git (required)
  • Knowledge of Azure cloud services such as Azure Data Factory, Azure Blob Storage, and Azure SQL Database (required)
  • Knowledge of data infrastructure/modeling tools like DBT, Fivetran (a plus)
  • Experience with BI tools like Power BI(a plus, not core to this role)
  • Knowledge of Docker, Linux, Shell/Bash, and virtualization technologies (a plus)
  • Knowledge of SSIS packages (a plus)
  • Familiarity with CI/CD methodologies



Benefits


Life at Incubyte  


We are a remote-first company with structured flexibility. Teams commit to shared rhythms during core hours, ensuring smooth collaboration while maintaining autonomy. Twice a year, we come together in person for a co-working sprint and once a year for a retreat - with all travel expenses covered. 


Our environment is built for crafters: pairing, refactoring, experimenting with AI, and pushing the boundaries of software excellence. We are all lifelong learners, and our work is our passion. 


Perks


  • Dedicated learning & development budget. 
  • Sponsorship for conference talks. 
  • Comprehensive medical & term insurance.
  • Employee-friendly leave policies. 
  • Home Office fund 
  • Medical Insurance



Read more
company logo
Jancy A
Posted by Jancy A
Bengaluru (Bangalore)
5 - 7 yrs
₹4L - ₹20L / yr
Data Engineer,
skill iconPython
ETL
DevOps

Job Summary

Role Overview

We are looking for an experienced Data Engineer with strong expertise in Python, ETL, Advanced SQL, CI/CD, DevOps, and Data Analytics. The ideal candidate should have hands-on experience designing and developing scalable data pipelines, transforming large datasets, and supporting data-driven applications.

Experience with Google Cloud Platform (GCP) will be an added advantage.

Key Responsibilities

  • Design, develop, and maintain scalable ETL/ELT data pipelines using Python and SQL.
  • Develop complex and optimized SQL queries, stored procedures, and data transformations.
  • Build and maintain reliable data integration workflows across multiple data sources.
  • Perform data cleansing, validation, transformation, and quality checks.
  • Analyze data and provide insights to support business and technical requirements.
  • Implement and maintain CI/CD pipelines for data engineering applications.
  • Work with DevOps practices and tools to automate deployments, monitoring, and infrastructure processes.
  • Troubleshoot data pipeline failures, performance issues, and production incidents.
  • Optimize data processing workflows for performance, scalability, and reliability.
  • Collaborate with Data Analysts, Data Scientists, Developers, and other stakeholders.
  • Follow best practices for version control, testing, documentation, and deployment.
  • Contribute to cloud-based data engineering initiatives, preferably on GCP.

Required Skills

  • 5–7 years of hands-on experience in Data Engineering.
  • Strong programming skills in Python.
  • Strong expertise in Advanced SQL and database concepts.
  • Hands-on experience with ETL/ELT processes and data pipelines.
  • Good understanding of Data Warehousing and Data Modeling concepts.
  • Experience with CI/CD practices and tools.
  • Strong understanding of DevOps principles, automation, and deployment processes.
  • Strong data analytics and problem-solving skills.
  • Experience working with large datasets and performance optimization.
  • Good understanding of Git/version control and software development best practices.

Good to Have

  • Hands-on experience with Google Cloud Platform (GCP).
  • Exposure to GCP data services such as BigQuery, Cloud Storage, Dataflow, Composer, or Pub/Sub.
  • Experience with containerization/orchestration technologies such as Docker/Kubernetes.
  • Experience with workflow orchestration tools such as Airflow.
  • Knowledge of cloud-based data architecture and distributed data processing.

Preferred Candidate Profile

  • Strong analytical and problem-solving abilities.
  • Good communication and stakeholder management skills.
  • Ability to work independently as well as in a collaborative team environment.
  • Strong ownership of data pipelines and production systems.
  • Candidates who can join at short notice are preferred.

Mandatory Skills

 Data Engineer, Python , ETL, GCP, Advanced SQL, Strong Data Analytics skills, CICD, Devops


Read more
company logo
Banu S
Posted by Banu S
Hyderabad, Bengaluru (Bangalore)
5 - 12 yrs
₹4L - ₹22L / yr
Data engineering
skill iconPython
SQL

Job Summary

We are seeking a skilled Data Engineer to design, build, and maintain scalable data pipelines and infrastructure. The ideal candidate should have strong expertise in SQL, Python, Linux, and modern data engineering practices to support data integration, transformation, and analytics.

Key Responsibilities

  • Design, develop, and maintain ETL/ELT data pipelines.
  • Write efficient and optimized SQL queries for data extraction, transformation, and reporting.
  • Develop automation scripts using Python for data processing and workflow optimization.
  • Work with Linux environments for deployment, monitoring, and troubleshooting.
  • Ensure data quality, integrity, and reliability across data platforms.
  • Collaborate with data analysts, software engineers, and business stakeholders to deliver data solutions.
  • Monitor, troubleshoot, and optimize data pipelines for performance and scalability.
  • Implement best practices for data security, governance, and documentation.

Required Skills

  • Strong experience in Data Engineering concepts and ETL/ELT processes.
  • Proficiency in SQL, including query optimization and database design.
  • Strong programming skills in Python.
  • Hands-on experience with Linux commands, shell scripting, and system administration basics.
  • Experience with relational databases such as PostgreSQL, MySQL, SQL Server, or Oracle.
  • Familiarity with Git/version control.
  • Strong analytical and problem-solving skills.

Preferred Skills

  • Experience with cloud platforms (AWS, Azure, or GCP).
  • Knowledge of Apache Spark, Airflow, Kafka, or similar data engineering tools.
  • Experience with data warehousing solutions and big data technologies.
  • Understanding of CI/CD pipelines and containerization (Docker/Kubernetes).

Qualifications

  • Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field.
  • Relevant certifications in cloud or data engineering are an added advantage.


Read more
company logo
Dharani S
Posted by Dharani S
Bengaluru (Bangalore)
5 - 9 yrs
₹3L - ₹20L / yr
skill iconPython
DevOps
PySpark

Job Description


We are looking for an experienced Data Engineer with strong expertise in Python, ETL, SQL, CI/CD, and DevOps to design, develop, and maintain scalable data pipelines and data processing solutions.


Key Responsibilities

  • Design, develop, and maintain scalable ETL/ELT data pipelines.
  • Develop data processing solutions using Python.
  • Write complex and optimized SQL queries, stored procedures, and data transformations.
  • Build and maintain data ingestion and integration workflows.
  • Implement data quality, validation, monitoring, and error-handling processes.
  • Develop and maintain CI/CD pipelines for data engineering applications.
  • Work with DevOps tools and practices for automated build, deployment, and infrastructure management.
  • Collaborate with data analysts, data scientists, software engineers, and business teams.
  • Optimize data pipelines for performance, reliability, and scalability.
  • Troubleshoot production data issues and ensure timely resolution.
  • Follow best practices for version control, code quality, testing, and deployment.


Mandatory Skills

  • Python
  • ETL
  • SQL
  • CI/CD
  • DevOps
  • Git / Version Control
  • Strong problem-solving and debugging skills


Read more
NBFC for Digital Lending
NBFC for Digital Lending
Agency job
via by Bisman Gill
Mumbai
3yrs+
Upto ₹45L / yr (Varies
)
SQL
Data Structures
skill iconPython
skill iconAmazon Web Services (AWS)
skill iconPostgreSQL

Must-Have Skills

  • Minimum 3 years of experience in Data Engineering / Analytics Engineering / Fintech Data roles
  • Must have worked on SMS Parsing, intelligent platform, converting RAW customer SMS data into structured actionable financial signals and enabling downstream usage of SMS derived variables
  • Must have established a continuous learning cycle to expand parser coverage
  • Experience in Lending / NBFC / Fintech domain
  • Experience working with Bureau, SMS, Device, or Banking data
  • Strong Python and SQL (production level)
  • Experience handling unstructured data (SMS, logs, JSON, APIs)
  • Experience building data pipelines, schedulers, and cron jobs
  • Strong database design and data modelling skills
  • Ability to work in a startup environment with high ownership
  • Familiarity with modern platforms like AWS, Snowflake, Google BigQuery, Redshift


Good to Have

  • Experience in STPL, especially less than 25K ticket size
  • Experience with streaming (Kafka/Kinesis) and orchestration (Airflow or Step Functions)
  • Experience with feature stores and risk analytics datasets
  • Knowledge of regex, NLP basics for SMS parsing
  • Experience supporting real-time decision engines/underwriting systems


Role Summary

This role will be responsible for owning the end-to-end data-structuring layer across the organisation. The individual will transform large volumes of raw, unstructured, and semi-structured data (such as SMS, device, bureau, and app data) into clean, standardised, and analysis-ready datasets. These structured datasets will directly power risk analytics, fraud detection, marketing insights, collections strategy, and policy decisioning.


Key Objective of the Role

Ensure all raw lending data (SMS, Bureau, Device, AA, App logs) is captured, parsed, structured, and stored in a clean analytics-ready format inside databases (PostgreSQL, DynamoDB, AWS stack) so that the Risk and Data Science team can directly use it for feature creation, policy building, and portfolio monitoring.


Core Responsibilities

  1. End-to-End Data Ownership
  • Design, build, and maintain end-to-end data pipelines (batch + streaming) using AWS native services (Glue, Lambda, Step Functions, Kinesis, S3, Athena, Redshift, EMR/Spark, etc.): ingestion

→ parsing → structuring → storage

  • Work closely with Tech, Product, and Data Science to define what data should be captured
  • Maintain data documentation, data dictionaries, and schema governance
  • Ensure data quality, consistency, and version control
  1. Unstructured Data Processing (Highest Priority)
  • Parse raw SMS dumps and categorise into salary, EMI, loan apps, collections, credits, debits, OTP, etc.
  • Process device fingerprint, behavioural logs, and vendor data (FinBox, AA, Bureau APIs)


  • Convert JSON, logs, and raw API responses into structured feature tables
  • Build regex/keyword-based parsers for financial SMS classification
  1. Feature Implementation (From Risk & Data Science Team)
  • Implement feature creation logic provided by Risk/Data Science team
  • Translate business and policy logic into SQL/Python pipelines
  • Create reusable feature layers for underwriting, fraud, collections, and monitoring
  • Maintain a feature store for consistent model and policy usage
  1. Lending Data Understanding (Domain-Specific Requirement)
  • Work with Bureau data
  • Structure SMS-derived financial variables (income, stress, EMI signals)
  • Work with Account Aggregator and bank transaction datasets
  • Understand fintech alternate data used in underwriting and fraud detection
  1. Data Pipelines & Automation
  • Build and maintain ETL/ELT pipelines using Python & SQL
  • Create cron jobs for automated data ingestion and feature refresh
  • Automate vendor data pulls (Bureau, SMS SDK, AA, device data)
  • Ensure low-latency pipelines for real-time underwriting use cases
  1. Database Structuring & Storage Architecture
  • Structure clean datasets in PostgreSQL (analytics layer)
  • Manage raw data storage in DynamoDB / S3 data lake
  • Design normalized and denormalised tables for risk analytics
  • Optimise database performance for large-scale query workloads
  1. Dashboards & Readable Data Layer
  • Create analytics-ready datasets, implement & write Metabase queries and convert into dashboards (Metabase / Power BI)
  • Enable self-serve data access for Risk, Business, and Founders
  • Support ad-hoc analysis requirements from leadership
  1. Cross-Functional Collaboration (Very Important)
  • The role requires close collaboration with data science, tech, product, and business teams to ensure reliable data pipelines, well-defined schemas, API integrations, logging architecture and high data quality, enabling faster and more accurate decision-making across lending workflows.

Tech Stack (Current Environment)

  • AWS Services
  • PostgreSQL (Primary analytics DB)
  • DynamoDB (Raw/NoSQL storage)
  • Python (Pandas, NumPy, ETL frameworks)
  • Advanced SQL
  • APIs, JSON, and Log Data Handling
Read more
company logo
Anisha Jindal
Posted by Anisha Jindal
Bengaluru (Bangalore), Mumbai
5 - 14 yrs
Best in industry
Data engineering
skill iconPython
PySpark
DAX
PowerBI

Job Summary

We are looking for a skilled and experienced Data Engineer to join our growing data team. The ideal candidate will have strong expertise in Python, PySpark, Data Modeling, and Power BI, with hands-on experience in designing, developing, and optimizing scalable data solutions. The role requires working closely with business stakeholders, data architects, and analytics teams to build robust data pipelines and semantic models that enable data-driven decision-making.


Technical Skills

  • Strong hands-on experience in Python and PySpark development.
  • Expertise in building and optimizing Data Engineering solutions and ETL pipelines.
  • Strong understanding of Data Modeling concepts (Star Schema, Snowflake Schema, Dimensional Modeling).
  • Experience with Power BI Data Modeling and Semantic Layer development.
  • Proficiency in DAX (Data Analysis Expressions).
  • Experience designing and managing Semantic Models in Power BI.
  • Strong SQL skills and experience working with large datasets.
  • Knowledge of data warehousing concepts and best practices.


Preferred Skills

  • Experience with cloud platforms such as Azure, AWS, or GCP.
  • Exposure to modern data platforms like Databricks.
  • Understanding of data governance and data quality frameworks.
Read more
company logo
Rachit Gupta
Posted by Rachit Gupta
Bengaluru (Bangalore)
1 - 5 yrs
Best in industry
Snow flake schema
snowflake
Data Warehouse (DWH)
ETL
Data Pipeline
+4 more

We are looking for a Data Engineer with at least 1 year of hands-on experience building solutions on Snowflake. The candidate should be comfortable designing, building, and managing reliable data pipelines that move data from multiple sources into a central data platform.

Responsibilities

  • Build and maintain data pipelines for ingesting, transforming, and loading data into Snowflake
  • Design scalable data models, schemas, tables, and views in Snowflake
  • Develop ETL/ELT workflows using SQL, Python, or data orchestration tools
  • Integrate data from APIs, databases, files, and third-party platforms
  • Monitor pipeline performance, failures, data quality, and freshness
  • Optimize Snowflake queries, warehouses, storage, and compute usage
  • Implement incremental loads, change data capture, and scheduled workflows
  • Work with engineering and business teams to understand data requirements
  • Maintain documentation for pipelines, datasets, and data transformations

Requirements

  • 1+ year of hands-on experience working with Snowflake
  • Strong SQL skills and experience writing complex queries
  • Experience building and managing ETL or ELT data pipelines
  • Knowledge of data warehousing concepts, dimensional modelling, and data quality
  • Experience with Python or another scripting language
  • Familiarity with orchestration tools such as Airflow, Dagster, Prefect, dbt, or similar
  • Understanding of APIs, relational databases, file formats, and cloud storage
  • Ability to troubleshoot pipeline failures and performance issues
  • Strong analytical, problem-solving, and communication skills

Good to Have

  • Experience with dbt and Snowflake Tasks, Streams, Snowpipe, or Dynamic Tables
  • Knowledge of AWS, Azure, or Google Cloud
  • Experience with Kafka or other streaming platforms
  • Familiarity with CI/CD, Git, monitoring, and data governance practices
  • Experience integrating ERP, finance, or operational systems
Read more
company logo
Pavithra E
Posted by Pavithra E
Hyderabad, Pune
5 - 9 yrs
₹18L - ₹20L / yr
PySpark
SQL
skill iconPython

Data Engineer Short Hiring Post


🚨 Hiring: Data Engineer

🔹 Experience: 5–9 Years

🔹 Location: Bangalore / Hyderabad

🔹 Skills: PySpark, Python, SQL, ETL, CI/CD, Data Modeling

🔹 Process: L1 Virtual → L2 F2F Karat Test

🔹 F2F: Bangalore / Hyderabad Location

🔹 Positions: Immediate requirement

⚠️ Note: Candidates must be available for F2F Karat immediately after L1.

#Hiring #DataEngineer #PySpark #Python #SQL #BangaloreJobs #HyderabadJobs #Mphasis #ImmediateJoiners

Read more
company logo
shruthi k
Posted by shruthi k
Remote, Pune
7 - 18 yrs
₹1L - ₹35L / yr (ESOP available)
microsoft fabric,dataengineering,one lake

Data Engineer – Microsoft Fabric

Location: Pune, India

Work Mode: Hybrid

Experience: 6+ Years

Employment Type: Full-time contactor

Compensation: As per market standards, commensurate with experience and expertise

Shift Timings: 2:00 PM – 11:00 PM IST

Notice Period: 0 – 15 days

About the Role

Jade Business Services (JBS) is seeking a Data Engineer – Microsoft Fabric to join our Pune team and work on enterprise-scale data transformation and analytics initiatives.

We are looking for a hands-on Data Engineer with strong experience in Microsoft Fabric, SQL, Python/PySpark and modern data engineering practices. The candidate will be responsible for building scalable data pipelines, implementing Lakehouse and Warehouse solutions, developing data models and supporting governed, reliable and AI-ready data platforms.

The ideal candidate should be comfortable working with architects, engineering teams and client stakeholders to translate business requirements into scalable and production-ready data solutions.

Roles and Responsibilities

  • Design and develop data solutions using Microsoft Fabric, including OneLake, Lakehouse, Warehouse and Data Factory pipelines.
  • Build and maintain scalable ETL/ELT pipelines for batch and incremental data processing.
  • Develop data ingestion and transformation pipelines using Fabric Data Factory, SQL, Python and/or PySpark.
  • Implement Medallion Architecture using Bronze, Silver and Gold layers.
  • Work with Lakehouse and Fabric Warehouse for enterprise data processing and analytics.
  • Develop and maintain data models, tables, views and optimized SQL queries.
  • Build and support semantic models for Power BI and analytical workloads.
  • Implement data quality, validation, monitoring and error-handling mechanisms.
  • Work with metadata, lineage and governance requirements using Microsoft Purview.
  • Implement data security, access controls and role-based permissions across data platforms.
  • Support Data Product and domain-oriented data architecture principles.
  • Follow DataOps practices including CI/CD, deployment, monitoring and production support.
  • Troubleshoot pipeline failures, performance issues and data quality problems.
  • Optimize data pipelines, queries and storage for performance and cost efficiency.
  • Work closely with Data Architects and business stakeholders to understand requirements and implement technical solutions.
  • Participate in technical design discussions, code reviews and architecture reviews.
  • Maintain technical documentation, data flow diagrams and pipeline documentation.
  • Support production deployments, incident resolution and SLA-driven data platform operations.
  • Identify opportunities for automation and AI-assisted improvements across data engineering processes.

Qualifications and Skills

  • 6+ years of experience in Data Engineering, Data Integration or Data Platform development.
  • Strong hands-on experience with Microsoft Fabric.
  • Experience with:
  • Microsoft Fabric Lakehouse
  • Fabric Warehouse
  • OneLake
  • Fabric Data Factory / Pipelines
  • Semantic Models
  • Strong understanding of Lakehouse and Medallion Architecture.
  • Strong SQL development and query optimization skills.
  • Hands-on experience with Python and/or PySpark.
  • Experience developing enterprise ETL/ELT and data integration pipelines.
  • Experience with batch and incremental data processing.
  • Understanding of data modelling concepts including dimensional modelling.
  • Knowledge of data quality, metadata, lineage and data governance.
  • Working knowledge of Microsoft Purview.
  • Understanding of Data Mesh and Data Product concepts.
  • Experience with CI/CD, version control, monitoring and DataOps practices.
  • Understanding of cloud security, access controls and data privacy.
  • Good troubleshooting and problem-solving skills.
  • Strong communication skills and ability to work with distributed and client-facing teams.

Preferred Skills

  • Microsoft Fabric or Azure Data certifications.
  • Experience migrating workloads from Azure Synapse, SQL Server, Databricks or other data platforms to Microsoft Fabric.
  • Experience implementing Medallion Architecture on Microsoft Fabric.
  • Experience with Power BI and semantic modelling.
  • Exposure to AI/ML, Generative AI or Agentic AI use cases on enterprise data platforms.
  • Experience working with Data Products or domain-oriented data solutions.
  • Experience in Energy & Utilities, Healthcare, Financial Services or Insurance.
  • Experience working with US or international enterprise clients.

What We Expect

The ideal candidate should be hands-on first and capable of independently building, troubleshooting and optimizing Fabric data solutions. You should be able to explain the technical decisions behind your implementation and work effectively with architects and engineering teams to deliver production-ready solutions.

 

Read more
company logo
Akshay Patil
Posted by Akshay Patil
Noida, Bengaluru (Bangalore), Pune, Hyderabad, Chennai
6 - 8 yrs
₹6L - ₹12L / yr
Data engineering
databricks
Snow flake schema
skill iconPython
Apache Spark
+8 more

Job Title : Data Engineer – Databricks

Experience : 6+ Years

Location : Noida / Hyderabad / Chennai / Pune / Bengaluru (Hybrid)

Shift : IST (Normal Shift)


Job Summary :

We are seeking an experienced Data Engineer with strong expertise in Databricks, Snowflake, Python, and Spark to build and optimize scalable data pipelines and support AI/ML model deployments. The ideal candidate should have experience working with cloud-based data platforms and preferably possess exposure to the Healthcare domain.


Required Skills :

  • Databricks (Preferred)
  • Snowflake
  • Python
  • Apache Spark
  • SQL
  • Azure Cloud
  • Kubernetes
  • Apache Airflow
  • GitHub & CI/CD Pipelines
  • AI/ML Model Deployment
  • Data Analytics

Preferred :

  • Experience in the Healthcare domain.
  • Strong understanding of scalable data engineering architectures and best practices.
Read more
Why apply to jobs via Cutshort
people_solving_puzzle
Personalized job matches
Stop wasting time. Get matched with jobs that meet your skills, aspirations and preferences.
people_verifying_people
Verified hiring teams
See actual hiring teams, find common social connections or connect with them directly.
ai_chip
Move faster with AI
We use AI to get you faster responses, recommendations and unmatched user experience.
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos