Cutshort logo
For Employers
one-to-one, one-to-many, and many-to-many logo
Data Engineer
one-to-one, one-to-many, and many-to-many
Data Engineer

Data Engineer at one-to-one, one-to-many, and many-to-many · Chennai · 5 - 10 years · ₹1L - ₹15L / yr · Posted 6 Mar 2024

The Hub's logo

Data Engineer

at one-to-one, one-to-many, and many-to-many

Agency job
5 - 10 yrs
₹1L - ₹15L / yr
Chennai
Skills
AWS CloudFormation
skill iconPython
PySpark
AWS Lambda

5-7 years of experience in Data Engineering with solid experience in design, development and implementation of end-to-end data ingestion and data processing system in AWS platform.

2-3 years of experience in AWS Glue, Lambda, Appflow, EventBridge, Python, PySpark, Lake House, S3, Redshift, Postgres, API Gateway, CloudFormation, Kinesis, Athena, KMS, IAM.

Experience in modern data architecture, Lake House, Enterprise Data Lake, Data Warehouse, API interfaces, solution patterns, standards and optimizing data ingestion.

Experience in build of data pipelines from source systems like SAP Concur, Veeva Vault, Azure Cost, various social media platforms or similar source systems.

Expertise in analyzing source data and designing a robust and scalable data ingestion framework and pipelines adhering to client Enterprise Data Architecture guidelines.

Proficient in design and development of solutions for real-time (or near real time) stream data processing as well as batch processing on the AWS platform.

Work closely with business analysts, data architects, data engineers, and data analysts to ensure that the data ingestion solutions meet the needs of the business.

Troubleshoot and provide support for issues related to data quality and data ingestion solutions. This may involve debugging data pipeline processes, optimizing queries, or troubleshooting application performance issues.

Experience in working in Agile/Scrum methodologies, CI/CD tools and practices, coding standards, code reviews, source management (GITHUB), JIRA, JIRA Xray and Confluence.

Experience or exposure to design and development using Full Stack tools.

Strong analytical and problem-solving skills, excellent communication (written and oral), and interpersonal skills.

Bachelor's or master's degree in computer science or related field.

 

 

Read more
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos

Similar jobs (10)

It is an Product Based Company(Domain- EV Charging)
It is an Product Based Company(Domain- EV Charging)
Agency job
via by Mantasha Naaz
Bengaluru (Bangalore)
3 - 5 yrs
₹13L - ₹15L / yr
skill iconAmazon Web Services (AWS)
skill iconPython
PySpark
SQL
ETL
+2 more

Data Engineer

Location: Bengaluru, India (Hybrid)

Employment Type: Full-time

Experience: 3-5 years



Role Overview  

What We’re Looking For:

  • Bachelor’s degree in Computer Science/Engineering or equivalent experience required.
  • Experience designing and shipping cloud services products.
  • Experience driving and managing technical and architectural dependencies on AWS Cloud.
  • A firm understanding of system architecture, cloud computing, PaaS/SaaS design principles, S3, DynamoDB, RDS mandatory.
  • Experience in building or maintaining ETL processes and tools, i.e., AWS Glue or any open-source tool.
  • Proven system-level design contribution to a current “Live” (in production / under daily high load) multi-region SaaS or PaaS offering.
  • Proven experience with S3, DynamoDB, SQL, and AWS RDS services.
  • Proficiency in programming languages such as Python.
  • Strong analytical and problem-solving skills.

Required Skills & Experience

  • Experience with Python, SQL, and data visualization/exploration tools.
  • Familiarity with the AWS ecosystem, specifically S3, DynamoDB, and RDS.
  • Communication skills, especially for explaining technical concepts to nontechnical business leaders.
  • Ability to work on a dynamic, research-oriented team that has concurrent projects.
  • Experience in AWS cost optimization (Savings Plans, Reserved Instances, Spot Instances) and governance frameworks.
  • Experience developing solutions using infrastructure orchestration tools (SSM, automation account, Ansible, etc.).
  • Excellent leadership, stakeholder management, and communication skills.

 

What We Offer

  • Work with some of the brightest minds in the emerging EV industry.
  • Make a tangible impact in reducing carbon emissions and enabling sustainable energy.
  • Freedom to suggest, implement, and innovate on systems, processes, and technologies.
  • Daily ownership in a high-growth, challenging environment.
  • Flexible work environment with hybrid schedules and virtualization options.
  • Competitive pay and benefits including health coverage, innovative PTO program, and performance bonuses.


Read more
company logo
Agency job
via by Akash Bhatt
Bengaluru (Bangalore), Hyderabad, Chennai, Bhubaneswar, Kolkata, Delhi, Gurugram, Noida, Ghaziabad, Faridabad
2 - 12 yrs
₹6L - ₹35L / yr
Amazon Redshift
Snow flake schema
Apache Airflow
Data Transformation Tool (DBT)
ETL
+1 more

We are looking for an AWS Data Engineer to build reliable, scalable data pipelines on AWS.


Responsibilities

  • Build ETL and ELT pipelines with AWS Glue and dbt
  • Model and optimise data warehouses in Redshift and Snowflake
  • Orchestrate workflows with Apache Airflow
  • Build streaming ingestion with Amazon Kinesis
  • Ensure data quality, monitoring and cost efficiency


Requirements

  • 2+ years of data engineering on AWS
  • Hands-on with Glue, Redshift or Snowflake, and Airflow
  • Strong data modelling and warehousing skills
Read more
NBFC for Digital Lending
NBFC for Digital Lending
Agency job
via by Bisman Gill
Mumbai
3yrs+
Upto ₹45L / yr (Varies
)
SQL
Data Structures
skill iconPython
skill iconAmazon Web Services (AWS)
skill iconPostgreSQL

Must-Have Skills

  • Minimum 3 years of experience in Data Engineering / Analytics Engineering / Fintech Data roles
  • Must have worked on SMS Parsing, intelligent platform, converting RAW customer SMS data into structured actionable financial signals and enabling downstream usage of SMS derived variables
  • Must have established a continuous learning cycle to expand parser coverage
  • Experience in Lending / NBFC / Fintech domain
  • Experience working with Bureau, SMS, Device, or Banking data
  • Strong Python and SQL (production level)
  • Experience handling unstructured data (SMS, logs, JSON, APIs)
  • Experience building data pipelines, schedulers, and cron jobs
  • Strong database design and data modelling skills
  • Ability to work in a startup environment with high ownership
  • Familiarity with modern platforms like AWS, Snowflake, Google BigQuery, Redshift


Good to Have

  • Experience in STPL, especially less than 25K ticket size
  • Experience with streaming (Kafka/Kinesis) and orchestration (Airflow or Step Functions)
  • Experience with feature stores and risk analytics datasets
  • Knowledge of regex, NLP basics for SMS parsing
  • Experience supporting real-time decision engines/underwriting systems


Role Summary

This role will be responsible for owning the end-to-end data-structuring layer across the organisation. The individual will transform large volumes of raw, unstructured, and semi-structured data (such as SMS, device, bureau, and app data) into clean, standardised, and analysis-ready datasets. These structured datasets will directly power risk analytics, fraud detection, marketing insights, collections strategy, and policy decisioning.


Key Objective of the Role

Ensure all raw lending data (SMS, Bureau, Device, AA, App logs) is captured, parsed, structured, and stored in a clean analytics-ready format inside databases (PostgreSQL, DynamoDB, AWS stack) so that the Risk and Data Science team can directly use it for feature creation, policy building, and portfolio monitoring.


Core Responsibilities

  1. End-to-End Data Ownership
  • Design, build, and maintain end-to-end data pipelines (batch + streaming) using AWS native services (Glue, Lambda, Step Functions, Kinesis, S3, Athena, Redshift, EMR/Spark, etc.): ingestion

→ parsing → structuring → storage

  • Work closely with Tech, Product, and Data Science to define what data should be captured
  • Maintain data documentation, data dictionaries, and schema governance
  • Ensure data quality, consistency, and version control
  1. Unstructured Data Processing (Highest Priority)
  • Parse raw SMS dumps and categorise into salary, EMI, loan apps, collections, credits, debits, OTP, etc.
  • Process device fingerprint, behavioural logs, and vendor data (FinBox, AA, Bureau APIs)


  • Convert JSON, logs, and raw API responses into structured feature tables
  • Build regex/keyword-based parsers for financial SMS classification
  1. Feature Implementation (From Risk & Data Science Team)
  • Implement feature creation logic provided by Risk/Data Science team
  • Translate business and policy logic into SQL/Python pipelines
  • Create reusable feature layers for underwriting, fraud, collections, and monitoring
  • Maintain a feature store for consistent model and policy usage
  1. Lending Data Understanding (Domain-Specific Requirement)
  • Work with Bureau data
  • Structure SMS-derived financial variables (income, stress, EMI signals)
  • Work with Account Aggregator and bank transaction datasets
  • Understand fintech alternate data used in underwriting and fraud detection
  1. Data Pipelines & Automation
  • Build and maintain ETL/ELT pipelines using Python & SQL
  • Create cron jobs for automated data ingestion and feature refresh
  • Automate vendor data pulls (Bureau, SMS SDK, AA, device data)
  • Ensure low-latency pipelines for real-time underwriting use cases
  1. Database Structuring & Storage Architecture
  • Structure clean datasets in PostgreSQL (analytics layer)
  • Manage raw data storage in DynamoDB / S3 data lake
  • Design normalized and denormalised tables for risk analytics
  • Optimise database performance for large-scale query workloads
  1. Dashboards & Readable Data Layer
  • Create analytics-ready datasets, implement & write Metabase queries and convert into dashboards (Metabase / Power BI)
  • Enable self-serve data access for Risk, Business, and Founders
  • Support ad-hoc analysis requirements from leadership
  1. Cross-Functional Collaboration (Very Important)
  • The role requires close collaboration with data science, tech, product, and business teams to ensure reliable data pipelines, well-defined schemas, API integrations, logging architecture and high data quality, enabling faster and more accurate decision-making across lending workflows.

Tech Stack (Current Environment)

  • AWS Services
  • PostgreSQL (Primary analytics DB)
  • DynamoDB (Raw/NoSQL storage)
  • Python (Pandas, NumPy, ETL frameworks)
  • Advanced SQL
  • APIs, JSON, and Log Data Handling
Read more
company logo
Dharani S
Posted by Dharani S
Bengaluru (Bangalore)
5 - 9 yrs
₹3L - ₹20L / yr
skill iconPython
DevOps
PySpark

Job Description


We are looking for an experienced Data Engineer with strong expertise in Python, ETL, SQL, CI/CD, and DevOps to design, develop, and maintain scalable data pipelines and data processing solutions.


Key Responsibilities

  • Design, develop, and maintain scalable ETL/ELT data pipelines.
  • Develop data processing solutions using Python.
  • Write complex and optimized SQL queries, stored procedures, and data transformations.
  • Build and maintain data ingestion and integration workflows.
  • Implement data quality, validation, monitoring, and error-handling processes.
  • Develop and maintain CI/CD pipelines for data engineering applications.
  • Work with DevOps tools and practices for automated build, deployment, and infrastructure management.
  • Collaborate with data analysts, data scientists, software engineers, and business teams.
  • Optimize data pipelines for performance, reliability, and scalability.
  • Troubleshoot production data issues and ensure timely resolution.
  • Follow best practices for version control, code quality, testing, and deployment.


Mandatory Skills

  • Python
  • ETL
  • SQL
  • CI/CD
  • DevOps
  • Git / Version Control
  • Strong problem-solving and debugging skills


Read more
Bengaluru (Bangalore)
8 - 12 yrs
₹50L - ₹100L / yr
skill iconPython
Software engineering
Data engineering
OAuth
RESTful APIs
+1 more

About LH2 AI LabsLH2 AI Labs is an applied research lab solving data platform and curation challenges for foundation model development. We serve every frontier AI lab with the mission of delivering the best data to power the best models.


Our customers are the ones building the foundation models themselves and our work sits directly in the loop of how those systems improve. This is a rare opportunity to join a company at a defining moment in AI.


About the role:

This is a full-time, hands-on role where you will own the core infrastructure and systems that enable us to discover customers' data landscape, handle sensitive data like names, address securely and to clean and catalog them into training-ready datasets.

That path covers connectors and on-prem agentic discovery components, secrets and PII scrubbing, scale and cost efficient ingestion and clearance for use in model training. You lead a small team of 3-4 engineers and raise the quality bar in your pod.


What you'll own

  • Connectors and extraction - You'll build connectors for common SaaS and databases (Google Workspace, Slack, Postgres/MySQL, Jira, MongoDB) and for specific tools such as Razorpay, GreytHR, Keka, Zoho and LeadSquared. The rule is to build only where no good open-source connector or clean export exists.
  • On-prem scanning - You'll build a CLI or agent that runs at the data owner's end to sample and estimate the value of their data without shipping all of it to us.
  • Clearance pipeline - Secrets detection across full git history, fail-closed. PII redaction for code and conversational data, including Indian identifiers (PAN, Aadhaar, GSTIN, IFSC, UPI). Measurable recall and precision.
  • Lineage - Every output record must trace back to its raw source, and every lot must be revocable.
  • Delivery - Packaging, sampling for buyers, and supporting the supply and BD teams on data questions.


You have

  • 8+ years in backend or data engineering, with at least 2 years leading up to 4 or more engineers.
  • You've built ingestion or ETL pipelines that run in production.
  • Hands-on work with sensitive data such as PII, financial or health data, including redaction, masking or access controls.
  • The judgment to decide what to build and what to adopt.


Nice to have

  • Experience with Presidio, gitleaks/TruffleHog, dlt or Airbyte.
  • Knowledge of DPDP, GDPR, HIPAA or SOC 2.
  • You've shipped software that runs in customers' environments.


This is not a pure people-management role. You'll write code, review everything, and personally own the hardest problems in the pod.

Read more
company logo
Bengaluru (Bangalore)
14 - 25 yrs
₹50L - ₹70L / yr
Data engineering
databricks
Apache Spark
PySpark
skill iconPython
+19 more

Job Title : Senior Data Engineer – Databricks

Experience : 14 to 20 Years

Location : HSR Layout, Bangalore

Work Mode : Hybrid – 3 Days WFO

Shift : 11:30 AM – 07:30 PM IST

Positions : 2

Notice Period : Immediate Joiners Only

Interview : 1 Technical Round + 2 Client Rounds


Role Overview :

We are looking for a Senior Data Engineer to build and lead enterprise-scale data platforms for a Switzerland-based commodity client.

The role requires a strong hands-on Data Engineering professional with expertise in Databricks, PySpark, Python, SQL, and AWS, along with technical leadership and stakeholder management experience.


Must-Have Skills :

  • 14 to 20 years of Data Engineering experience
  • Databricks & Apache Spark / PySpark
  • Python & SQL
  • AWS Cloud
  • Lakehouse Architecture
  • ETL / ELT & Distributed Data Processing
  • Batch & Streaming Pipelines
  • Data Pipeline Optimization & Data Modeling
  • CDC & Incremental Processing
  • Git, CI/CD & Testing
  • Data Quality, Monitoring & Observability
  • Technical Leadership & Stakeholder Management


Key Responsibilities :

  • Design and build scalable data pipelines using Databricks, PySpark, Python, SQL, and AWS.
  • Own data products from design through production.
  • Develop batch / streaming pipelines and reusable ETL / ELT frameworks.
  • Optimize pipelines for performance, scalability, reliability, and cost.
  • Design scalable data architectures and data models.
  • Implement data quality, monitoring, lineage, and CI/CD practices.
  • Lead technical discussions and mentor engineering teams.
  • Collaborate with business stakeholders, architects, product owners, and engineering teams.
  • Remain hands-on while providing technical leadership.


Ideal Candidate :

A 14 to 20 years experienced, hands-on Data Engineering leader with strong Databricks + PySpark + AWS expertise, excellent communication, stakeholder management, and experience delivering enterprise-scale data platforms.

🔴 Super Urgent : Only Bangalore-based immediate joiners.

Read more
company logo
Pavithra E
Posted by Pavithra E
Hyderabad, Pune
5 - 9 yrs
₹18L - ₹20L / yr
PySpark
SQL
skill iconPython

Data Engineer Short Hiring Post


🚨 Hiring: Data Engineer

🔹 Experience: 5–9 Years

🔹 Location: Bangalore / Hyderabad

🔹 Skills: PySpark, Python, SQL, ETL, CI/CD, Data Modeling

🔹 Process: L1 Virtual → L2 F2F Karat Test

🔹 F2F: Bangalore / Hyderabad Location

🔹 Positions: Immediate requirement

⚠️ Note: Candidates must be available for F2F Karat immediately after L1.

#Hiring #DataEngineer #PySpark #Python #SQL #BangaloreJobs #HyderabadJobs #Mphasis #ImmediateJoiners

Read more
company logo
Jancy A
Posted by Jancy A
Bengaluru (Bangalore)
5 - 7 yrs
₹4L - ₹20L / yr
Data Engineer,
skill iconPython
ETL
DevOps

Job Summary

Role Overview

We are looking for an experienced Data Engineer with strong expertise in Python, ETL, Advanced SQL, CI/CD, DevOps, and Data Analytics. The ideal candidate should have hands-on experience designing and developing scalable data pipelines, transforming large datasets, and supporting data-driven applications.

Experience with Google Cloud Platform (GCP) will be an added advantage.

Key Responsibilities

  • Design, develop, and maintain scalable ETL/ELT data pipelines using Python and SQL.
  • Develop complex and optimized SQL queries, stored procedures, and data transformations.
  • Build and maintain reliable data integration workflows across multiple data sources.
  • Perform data cleansing, validation, transformation, and quality checks.
  • Analyze data and provide insights to support business and technical requirements.
  • Implement and maintain CI/CD pipelines for data engineering applications.
  • Work with DevOps practices and tools to automate deployments, monitoring, and infrastructure processes.
  • Troubleshoot data pipeline failures, performance issues, and production incidents.
  • Optimize data processing workflows for performance, scalability, and reliability.
  • Collaborate with Data Analysts, Data Scientists, Developers, and other stakeholders.
  • Follow best practices for version control, testing, documentation, and deployment.
  • Contribute to cloud-based data engineering initiatives, preferably on GCP.

Required Skills

  • 5–7 years of hands-on experience in Data Engineering.
  • Strong programming skills in Python.
  • Strong expertise in Advanced SQL and database concepts.
  • Hands-on experience with ETL/ELT processes and data pipelines.
  • Good understanding of Data Warehousing and Data Modeling concepts.
  • Experience with CI/CD practices and tools.
  • Strong understanding of DevOps principles, automation, and deployment processes.
  • Strong data analytics and problem-solving skills.
  • Experience working with large datasets and performance optimization.
  • Good understanding of Git/version control and software development best practices.

Good to Have

  • Hands-on experience with Google Cloud Platform (GCP).
  • Exposure to GCP data services such as BigQuery, Cloud Storage, Dataflow, Composer, or Pub/Sub.
  • Experience with containerization/orchestration technologies such as Docker/Kubernetes.
  • Experience with workflow orchestration tools such as Airflow.
  • Knowledge of cloud-based data architecture and distributed data processing.

Preferred Candidate Profile

  • Strong analytical and problem-solving abilities.
  • Good communication and stakeholder management skills.
  • Ability to work independently as well as in a collaborative team environment.
  • Strong ownership of data pipelines and production systems.
  • Candidates who can join at short notice are preferred.

Mandatory Skills

 Data Engineer, Python , ETL, GCP, Advanced SQL, Strong Data Analytics skills, CICD, Devops


Read more
company logo
SaiSruthi Nuthanpati
Posted by SaiSruthi Nuthanpati
Tirupati, Chennai
5 - 10 yrs
Best in industry
SQL
skill iconPython
Stored Procedures
skill iconAmazon Web Services (AWS)
Microsoft Windows Azure
+10 more

About Us:

The QX Impact was launched with a mission to make A.I accessible and affordable and deliver AI Products/Solutions at scale for the enterprises by bringing the power of Data, AI, and Engineering to drive digital transformation. We believe without insights; businesses will continue to face challenges to better understand their customers and even lose them. Secondly, without insights businesses won't’ be able to deliver differentiated products/services; and finally, without insights, businesses can’t achieve a new level of “Operational Excellence” is crucial to remain competitive, meeting rising customer expectations, expanding markets, and digitalization.


Job Summary:

We are looking for a Senior Data Engineer who is creative, collaborative, and adaptable to join our agile team of data scientists, engineers, and UX developers. The role focuses on building and maintaining robust data pipelines to support advanced analytics, data science, and BI solutions.

As a Senior Data Engineer, you will work with internal and external data, collaborate with data scientists, and contribute to the design, development, and deployment of innovative solutions.


Key Responsibilities:

  • Design, develop, test, and maintain optimal data pipeline and ETL architectures.
  • Map out data systems and define/design required integrations, ETL, BI, and AI systems/processes.
  • Prepare and optimize data for predictive and prescriptive modeling.
  • Collaborate with teams to integrate ERP data into the enterprise data lake, ensuring seamless flow and quality.
  • Enhance cloud data infrastructure on AWS or Azure for scalability and performance.
  • Utilize big data tools and frameworks to optimize data acquisition and preparation.
  • Build architectures to move data to/from data lakes and data warehouses for advanced analytics.
  • Develop and curate data models for analytics, dashboards, and reports.
  • Conduct code reviews, maintain production-level code, and implement testing approaches.
  • Monitor, troubleshoot, and resolve data ingestion workflows to maintain reliability and uptime.
  • Drive innovation and implement efficient new approaches to data engineering tasks.


Must-Have Skills:

  • Bachelor’s degree in Computer Science, Mathematics, Engineering, or a related field.
  • 5+ years of experience working with enterprise data platforms, including building and managing data lakes.
  • 3–5 years of experience designing and implementing data warehouse solutions.
  • Expertise in SQL, including developing stored procedures (SP) and applying advanced data design concepts.
  • Proficiency in Spark (Python/Scala) and Spark Streaming for real-time data pipelines.
  • Experience with AWS or Azure services (e.g., AWS Glue, Azure Data Factory, Redshift, Snowflake).
  • Familiarity with big data tools such as Apache Kafka, Apache Spark, or Flink.
  • Hands-on experience with orchestration tools (e.g., Apache Airflow, Prefect).
  • Knowledge of CI/CD processes, version control (e.g., Git, Jenkins), and deployment automation.
  • Strong problem-solving, communication, and collaboration skills.


Good-to-Have Skills:

  • Experience in integrating ERP data into data lakes.
  • Experience with traditional ETL tools (e.g., Talend, Pentaho).


Competencies:

  • Tech Savvy - Anticipating and adopting innovations in business-building digital and technology applications.
  • Self-Development - Actively seeking new ways to grow and be challenged using both formal and informal development channels.
  • Action Oriented - Taking on new opportunities and tough challenges with a sense of urgency, high energy, and enthusiasm.
  • Customer Focus - Building strong customer relationships and delivering customer-centric solutions.
  • Optimize Work Processes - Knowing the most effective and efficient processes to get things done, with a focus on continuous improvement.


Why Join Us?

  • Be part of a collaborative and agile team driving cutting-edge AI and data engineering solutions.
  • Work on impactful projects that make a difference across industries.
  • Opportunities for professional growth and continuous learning.
  • Competitive salary and benefits package.


Application Details

Ready to make an impact? Apply today and become part of the QX Impact team!


Read more
company logo
Tushar Vaghela
Posted by Tushar Vaghela
Bengaluru (Bangalore)
5 - 10 yrs
Best in industry
skill iconPython
skill iconScala
Apache Spark
Apache Kafka
databricks
+1 more

Description


We are looking for Senior Data Engineers to join our Data Platform team and build scalable, high-performance data platforms that power data processing, analytics, and downstream applications.

The ideal candidate will have strong experience in distributed data processing, ETL pipelines, and Big Data technologies, with hands-on expertise in Apache Spark and Python Scala.

You will be responsible for designing, developing, and optimizing large-scale data pipelines while collaborating closely with cross-functional engineering teams to build reliable, production-grade data solutions.



Key Responsibilities

  • Design, develop, and maintain scalable ETL and data processing pipelines for large-scale datasets.
  • Build and optimize distributed data applications using Apache Spark and Python Scala.
  • Develop reliable, high-performance data pipelines for batch and streaming workloads.
  • Design and manage data workflows using Apache Airflow.
  • Build and operate data workloads on AWS, with strong usage of Amazon S3 for large-scale data storage.
  • Work with large datasets to ensure data quality, consistency, reliability, and performance.
  • Collaborate with engineering, product, analytics, and other platform teams to deliver robust data solutions.
  • Optimize data workflows for scalability, reliability, performance, and cost efficiency.
  • Troubleshoot production issues, identify bottlenecks, and continuously improve platform performance.



Requirements

Candidates who demonstrate:

  • 5+ years of experience in Data Engineering, Big Data Engineering, or a similar role.
  • Strong hands-on experience with Apache Spark and Scala.
  • Experience designing, building, and maintaining large-scale ETL pipelines.
  • Strong hands-on experience with AWS, particularly Amazon S3.
  • Hands-on experience with Apache Airflow for workflow orchestration and scheduling.
  • Strong SQL skills and a solid understanding of distributed data processing concepts.
  • Experience working with batch and/or streaming data pipelines.
  • Excellent debugging, problem-solving, and performance optimization skills.
  • Strong communication and collaboration skills.


Good to Have

  • Experience with Databricks and the broader Databricks data platform.
  • Familiarity with streaming technologies such as Apache Kafka.
  • Experience working on large-scale data platforms handling high-volume data workloads.
  • Exposure to additional AWS data services and cloud-native data architectures.
Read more
Why apply to jobs via Cutshort
people_solving_puzzle
Personalized job matches
Stop wasting time. Get matched with jobs that meet your skills, aspirations and preferences.
people_verifying_people
Verified hiring teams
See actual hiring teams, find common social connections or connect with them directly.
ai_chip
Move faster with AI
We use AI to get you faster responses, recommendations and unmatched user experience.
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos