Cutshort logo
For Employers
Publicis Sapient logo
Senior Data Engineer (L2)
Senior Data Engineer (L2)

Senior Data Engineer (L2) at Publicis Sapient · Bengaluru (Bangalore), Pune, Hyderabad, Gurugram, Noida · 5 - 11 years · ₹20L - ₹36L / yr · Profitable · Posted 12 Apr 2024

Publicis Sapient's logo

Senior Data Engineer (L2)

Mohit Singh's profile picture
Posted by Mohit Singh
5 - 11 yrs
₹20L - ₹36L / yr
Bengaluru (Bangalore), Pune, Hyderabad, Gurugram, Noida
Skills
PySpark
Data engineering
Big Data
Hadoop
Spark
skill iconPython
skill iconScala
skill iconJava
Google Cloud Platform (GCP)
skill iconAmazon Web Services (AWS)
Microsoft Windows Azure
Apache Spark

Publicis Sapient Overview:

The Senior Associate People Senior Associate L1 in Data Engineering, you will translate client requirements into technical design, and implement components for data engineering solution. Utilize deep understanding of data integration and big data design principles in creating custom solutions or implementing package solutions. You will independently drive design discussions to insure the necessary health of the overall solution 

.

Job Summary:

As Senior Associate L2 in Data Engineering, you will translate client requirements into technical design, and implement components for data engineering solution. Utilize deep understanding of data integration and big data design principles in creating custom solutions or implementing package solutions. You will independently drive design discussions to insure the necessary health of the overall solution

The role requires a hands-on technologist who has strong programming background like Java / Scala / Python, should have experience in Data Ingestion, Integration and data Wrangling, Computation, Analytics pipelines and exposure to Hadoop ecosystem components. You are also required to have hands-on knowledge on at least one of AWS, GCP, Azure cloud platforms.


Role & Responsibilities:

Your role is focused on Design, Development and delivery of solutions involving:

• Data Integration, Processing & Governance

• Data Storage and Computation Frameworks, Performance Optimizations

• Analytics & Visualizations

• Infrastructure & Cloud Computing

• Data Management Platforms

• Implement scalable architectural models for data processing and storage

• Build functionality for data ingestion from multiple heterogeneous sources in batch & real-time mode

• Build functionality for data analytics, search and aggregation

Experience Guidelines:

Mandatory Experience and Competencies:

# Competency

1.Overall 5+ years of IT experience with 3+ years in Data related technologies

2.Minimum 2.5 years of experience in Big Data technologies and working exposure in at least one cloud platform on related data services (AWS / Azure / GCP)

3.Hands-on experience with the Hadoop stack – HDFS, sqoop, kafka, Pulsar, NiFi, Spark, Spark Streaming, Flink, Storm, hive, oozie, airflow and other components required in building end to end data pipeline.

4.Strong experience in at least of the programming language Java, Scala, Python. Java preferable

5.Hands-on working knowledge of NoSQL and MPP data platforms like Hbase, MongoDb, Cassandra, AWS Redshift, Azure SQLDW, GCP BigQuery etc

6.Well-versed and working knowledge with data platform related services on at least 1 cloud platform, IAM and data security


Preferred Experience and Knowledge (Good to Have):

# Competency

1.Good knowledge of traditional ETL tools (Informatica, Talend, etc) and database technologies (Oracle, MySQL, SQL Server, Postgres) with hands on experience

2.Knowledge on data governance processes (security, lineage, catalog) and tools like Collibra, Alation etc

3.Knowledge on distributed messaging frameworks like ActiveMQ / RabbiMQ / Solace, search & indexing and Micro services architectures

4.Performance tuning and optimization of data pipelines

5.CI/CD – Infra provisioning on cloud, auto build & deployment pipelines, code quality

6.Cloud data specialty and other related Big data technology certifications


Personal Attributes:

• Strong written and verbal communication skills

• Articulation skills

• Good team player

• Self-starter who requires minimal oversight

• Ability to prioritize and manage multiple tasks

• Process orientation and the ability to define and set up processes


Read more
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos

About Publicis Sapient

Founded :
1990
Type :
Services
Size :
5000+
Stage :
Profitable

About

Publicis Sapient helps established organizations get to their future, digitally-enabled state, both in the way they work and serve their customers.
Read more

Connect with the team

Profile picture
Yamini Murthy
Profile picture
Prem Sagar
Profile picture
Ankush Khatkar
Profile picture
Nisha Garg
Profile picture
JYoti Kaushik
Profile picture
Shaila Gupta
Profile picture
Aman Nagpal
Profile picture
Simer Arora
Profile picture
Pooja Singh
Profile picture
Neelam Pant

Company social profiles

instagramlinkedintwitterfacebook

Similar jobs (10)

company logo
Banu S
Posted by Banu S
Bengaluru (Bangalore)
5 - 15 yrs
₹3L - ₹30L / yr
skill iconJava
skill iconScala
databricks
snowflake

4 - 10 years of experience in designing and buildingarchitecting highly resilient data platforms 

∙Strong knowledge of data engineering, architecture and data modeling 

∙Experience in platforms like Databricks and Snowflake 

∙Experience on building applications on cloud (AWS or Azure or Google Cloud) 

∙Strong analytical and problem-solving skills 

∙Prior experience in developing data or computation intensive (e.g. grid based) backend applications is an 

advantage 

∙OOP design skills with an understanding or at least personal interest towards the concepts of Functional 

Programming  

∙Willingness to understand and enhance other people’s code, being able to work in an environment where 

developers will oversee and work on wider components also dealing with older “legacy” code 

 


∙Strong programming skills (Java/ Scala / Python) skills with the willingness to pick up the other language if not 

already mastered at a sufficient level is important 

∙Spring knowledge is an advantage, but in general willingness to learn, work with and even enhance in-house 

developed frameworks is a must 

∙Prior experience in working with Git, Bitbucket, Jenkins, working with PR-s, using JIRA, following the Scrum Agile 

methodology is an advantage 

∙Prior knowledge of financial products is an advantage 

∙Bachelors or Masters in any relevant field of IT/Engineering area is an advantage  

Read more
company logo
Bengaluru (Bangalore)
14 - 25 yrs
₹50L - ₹70L / yr
Data engineering
databricks
Apache Spark
PySpark
skill iconPython
+19 more

Job Title : Senior Data Engineer – Databricks

Experience : 14 to 20 Years

Location : HSR Layout, Bangalore

Work Mode : Hybrid – 3 Days WFO

Shift : 11:30 AM – 07:30 PM IST

Positions : 2

Notice Period : Immediate Joiners Only

Interview : 1 Technical Round + 2 Client Rounds


Role Overview :

We are looking for a Senior Data Engineer to build and lead enterprise-scale data platforms for a Switzerland-based commodity client.

The role requires a strong hands-on Data Engineering professional with expertise in Databricks, PySpark, Python, SQL, and AWS, along with technical leadership and stakeholder management experience.


Must-Have Skills :

  • 14 to 20 years of Data Engineering experience
  • Databricks & Apache Spark / PySpark
  • Python & SQL
  • AWS Cloud
  • Lakehouse Architecture
  • ETL / ELT & Distributed Data Processing
  • Batch & Streaming Pipelines
  • Data Pipeline Optimization & Data Modeling
  • CDC & Incremental Processing
  • Git, CI/CD & Testing
  • Data Quality, Monitoring & Observability
  • Technical Leadership & Stakeholder Management


Key Responsibilities :

  • Design and build scalable data pipelines using Databricks, PySpark, Python, SQL, and AWS.
  • Own data products from design through production.
  • Develop batch / streaming pipelines and reusable ETL / ELT frameworks.
  • Optimize pipelines for performance, scalability, reliability, and cost.
  • Design scalable data architectures and data models.
  • Implement data quality, monitoring, lineage, and CI/CD practices.
  • Lead technical discussions and mentor engineering teams.
  • Collaborate with business stakeholders, architects, product owners, and engineering teams.
  • Remain hands-on while providing technical leadership.


Ideal Candidate :

A 14 to 20 years experienced, hands-on Data Engineering leader with strong Databricks + PySpark + AWS expertise, excellent communication, stakeholder management, and experience delivering enterprise-scale data platforms.

🔴 Super Urgent : Only Bangalore-based immediate joiners.

Read more
company logo
Jancy A
Posted by Jancy A
Bengaluru (Bangalore)
5 - 7 yrs
₹4L - ₹20L / yr
Data Engineer,
skill iconPython
ETL
DevOps

Job Summary

Role Overview

We are looking for an experienced Data Engineer with strong expertise in Python, ETL, Advanced SQL, CI/CD, DevOps, and Data Analytics. The ideal candidate should have hands-on experience designing and developing scalable data pipelines, transforming large datasets, and supporting data-driven applications.

Experience with Google Cloud Platform (GCP) will be an added advantage.

Key Responsibilities

  • Design, develop, and maintain scalable ETL/ELT data pipelines using Python and SQL.
  • Develop complex and optimized SQL queries, stored procedures, and data transformations.
  • Build and maintain reliable data integration workflows across multiple data sources.
  • Perform data cleansing, validation, transformation, and quality checks.
  • Analyze data and provide insights to support business and technical requirements.
  • Implement and maintain CI/CD pipelines for data engineering applications.
  • Work with DevOps practices and tools to automate deployments, monitoring, and infrastructure processes.
  • Troubleshoot data pipeline failures, performance issues, and production incidents.
  • Optimize data processing workflows for performance, scalability, and reliability.
  • Collaborate with Data Analysts, Data Scientists, Developers, and other stakeholders.
  • Follow best practices for version control, testing, documentation, and deployment.
  • Contribute to cloud-based data engineering initiatives, preferably on GCP.

Required Skills

  • 5–7 years of hands-on experience in Data Engineering.
  • Strong programming skills in Python.
  • Strong expertise in Advanced SQL and database concepts.
  • Hands-on experience with ETL/ELT processes and data pipelines.
  • Good understanding of Data Warehousing and Data Modeling concepts.
  • Experience with CI/CD practices and tools.
  • Strong understanding of DevOps principles, automation, and deployment processes.
  • Strong data analytics and problem-solving skills.
  • Experience working with large datasets and performance optimization.
  • Good understanding of Git/version control and software development best practices.

Good to Have

  • Hands-on experience with Google Cloud Platform (GCP).
  • Exposure to GCP data services such as BigQuery, Cloud Storage, Dataflow, Composer, or Pub/Sub.
  • Experience with containerization/orchestration technologies such as Docker/Kubernetes.
  • Experience with workflow orchestration tools such as Airflow.
  • Knowledge of cloud-based data architecture and distributed data processing.

Preferred Candidate Profile

  • Strong analytical and problem-solving abilities.
  • Good communication and stakeholder management skills.
  • Ability to work independently as well as in a collaborative team environment.
  • Strong ownership of data pipelines and production systems.
  • Candidates who can join at short notice are preferred.

Mandatory Skills

 Data Engineer, Python , ETL, GCP, Advanced SQL, Strong Data Analytics skills, CICD, Devops


Read more
company logo
Hema Dekonda
Posted by Hema Dekonda
Bengaluru (Bangalore), Mumbai, Hyderabad, Gurugram
6 - 11 yrs
₹15L - ₹50L / yr
Data engineering
skill iconAmazon Web Services (AWS)
ETL
SQL
NOSQL Databases

About AuxoAI:


AuxoAI is a global platform-based services firm. We help companies—turn their strategies into practical digital and AI solutions. By understanding how our clients make decisions, we use digital and Artificial Intelligence (AI) technologies to drive growth, enhance their operations, improve customer experiences, and provide clear, actionable insights from their data. What We Do We work across various industries such as healthcare, high-tech, consumer packaged goods (CPG), finance etc., and in sales, marketing, and customer support functions.

We help our clients with accelerating their digital and AI journeys through:

• AI Application Development

• Data, Digital and Cloud acceleration using AI

• AI Native Product Engineering


We are seeking a skilled and experienced Data Engineer to join our dynamic team. The ideal candidate will have 6+ years of prior experience in data engineering, with a strong background in AWS (Amazon Web Services) technologies. This role offers an exciting opportunity to work on diverse projects, collaborating with cross-functional teams to design, build, and optimize data pipelines and infrastructure.


Responsibilities:

* Design, develop, and maintain scalable data pipelines and ETL processes leveraging AWS services such as S3, Glue, EMR, Lambda, and Redshift.

* Collaborate with data scientists and analysts to understand data requirements and implement solutions that support analytics and machine learning initiatives.

* Optimize data storage and retrieval mechanisms to ensure performance, reliability, and cost-effectiveness.

* Implement data governance and security best practices to ensure compliance and data integrity.

* Troubleshoot and debug data pipeline issues, providing timely resolution and proactive monitoring.

* Stay abreast of emerging technologies and industry trends, recommending innovative solutions to enhance data engineering capabilities.


Requirements :

* Bachelor's or Master's degree in Computer Science, Engineering, or a related field.

* 6+ years of prior experience in data engineering, with a focus on designing and building data pipelines.

* Proficiency in AWS services, particularly S3, Glue, EMR, Lambda, and Redshift.

* Strong programming skills in languages such as Python, Java, or Scala.

* Experience with SQL and NoSQL databases, data warehousing concepts, and big data technologies.

* Familiarity with containerization technologies (e.g., Docker, Kubernetes) and orchestration tools (e.g., Apache Airflow) is a plus.

Read more
company logo
Kanakavalli Kosuri
Posted by Kanakavalli Kosuri
Remote only
10 - 16 yrs
Best in industry
Snow flake schema
Data Transformation Tool (DBT)
fivetran
skill iconRuby on Rails (ROR)
skill iconReact.js
+6 more

At Mitratech, we are a team of technocrats focused on building world-class products that simplify operations in the Legal, Risk, Compliance, and HR functions. We are a close-knit, globally dispersed team that thrives in an ecosystem that supports individual excellence and takes pride in its diverse and inclusive work culture centered around great people practices, learning opportunities, and having fun! Our culture is the ideal blend of entrepreneurial spirit and enterprise investment, enabling the chance to move at a rapid pace with some of the most complex, leading-edge technologies available.


For over 35 years, the experts at Mitratech have been focused on solving the complex needs. Today, we serve 20,000 client companies of all sizes globally, representing 30% of the Fortune 500 and over 500,000 users in over 160 countries.


As we continue to grow, we’re always looking for resourceful, enthusiastic, and fresh perspectives. Join our global team and see what makes Mitratech a truly exceptional place to work!


Job Overview 

Principal Data Engineer

About Engineering at Mitratech Legal Solutions 

Mitratech's engineering organization is a collaborative and dynamic environment where engineers are empowered to drive technical direction and innovation. Our engineers are passionate about delivering high-quality products and solutions that meet the evolving needs of our customers, and we're committed to fostering a culture of continuous learning and growth. 


About the Role 

Mitratech is a fast-paced and dynamic environment, and this role requires someone who is adaptable, resilient, and able to thrive in a rapidly changing landscape. If you’re a seasoned engineer with a passion for technical leadership, innovation, and collaboration — including building the data foundations that power trusted reporting and agentic AI-driven products — we’d love to hear from you. 


What You Will Do

•  Drive technical direction for a significant product domain or platform capability, ensuring alignment with business objectives and customer needs 

•  Design and maintain data pipelines and reporting models that power trusted business metrics and increasingly feed agentic AI systems (e.g., RAG ingestion, embeddings, vector stores, AI agent workflows) 

•  Use AI-assisted and agentic engineering tools (e.g., Claude Code, Copilot, Cursor, AI agents) as part of your own workflow, and help other engineers adopt agentic development practices effectively 

•  Reduce systemic complexity by identifying and leading architectural debt remediation, and developing strategies for ongoing technical debt management 

•  Partner with Product and Engineering leadership to inform multi-quarter roadmap feasibility, and provide technical guidance and oversight to ensure successful implementation 

•  Elevate engineering craft across multiple teams through RFCs, mentorship, and knowledge sharing, and develop training programs to improve engineering skills and knowledge 

•  Represent Mitratech’s technical capabilities externally, including speaking at conferences, contributing to open-source projects, and engaging with industry peers and thought leaders 


What We Are Looking For 

To be successful in this role, you will need: 

•  10+ years of experience in software engineering, with a focus on technical leadership and architecture 

•  Deep understanding of data engineering principles, including data modeling, data warehousing, reporting, and data governance 

•  Strong technical expertise in SQL, PostgreSQL, ETL/ELT pipelines, BI tools, and analytics platforms 

•  Practical experience with AI/LLM-adjacent and agentic AI data work — e.g., RAG ingestion pipelines, embedding generation, vector store management, or building/operating AI agent workflows over data — using AI coding assistants (Claude Code, Copilot, Cursor, or similar) as a regular part of the engineering workflow 

•  Working knowledge of modern cloud platforms such as AWS 

•  Experience with BI, reporting, dashboards, and customer-facing analytics 

•  Experience leading cross-functional initiatives with product, engineering, analytics, and business teams 


Nice to Have 

•  Working knowledge of Ruby on Rails and React 

•  Experience with a semantic or metrics layer (e.g., dbt Semantic Layer, headless BI) 

•  Understanding of CI/CD, Git-based workflows, and infrastructure-as-code 

 

The Stack Context 

•  Modern data stack: Fivetran, Airbyte, dbt, Snowflake, GitHub, Terraform, or similar tools 

•  Application context (nice to have): Ruby on Rails, React, or similar backend/frontend frameworks 

•  Data modeling: SQL, analytics models, documentation, testing, naming standards, and version control 

•  Infrastructure: cloud-based data infrastructure, infrastructure-as-code, CI/CD, monitoring, and cloud storage 

•  Data workflows: ingestion, transformation, orchestration, reporting, deployment, and change management 

•  Reporting focus: trusted metrics, scalable reporting models, dashboards, exports, and data quality 

•  AI surface: data pipelines and quality practices supporting AI/LLM and agentic AI use cases (RAG, embeddings, vector stores, AI agents) alongside traditional BI 


Why This Role 

This role offers a unique opportunity to drive technical direction and innovation at a rapidly growing company, while also mentoring and coaching engineers to improve their craft. As a Principal Data Engineer at Mitratech, you will have the chance to work on complex and challenging problems spanning trusted reporting and agentic AI systems, collaborate with cross-functional teams, and represent the company's technical capabilities externally. If you're looking for a role that offers a mix of technical leadership, data and reporting depth, agentic AI innovation, and collaboration, this could be the perfect fit for you. 

 

We are an equal-opportunity employer that values diversity at all levels. All qualified applicants will receive consideration for employment without regard to race, color, religion, gender, national origin, age, sexual orientation, gender identity, disability, or veteran status.

 

Read more
company logo
Anisha Jindal
Posted by Anisha Jindal
Bengaluru (Bangalore), Mumbai
5 - 14 yrs
Best in industry
Data engineering
skill iconPython
PySpark
DAX
PowerBI

Job Summary

We are looking for a skilled and experienced Data Engineer to join our growing data team. The ideal candidate will have strong expertise in Python, PySpark, Data Modeling, and Power BI, with hands-on experience in designing, developing, and optimizing scalable data solutions. The role requires working closely with business stakeholders, data architects, and analytics teams to build robust data pipelines and semantic models that enable data-driven decision-making.


Technical Skills

  • Strong hands-on experience in Python and PySpark development.
  • Expertise in building and optimizing Data Engineering solutions and ETL pipelines.
  • Strong understanding of Data Modeling concepts (Star Schema, Snowflake Schema, Dimensional Modeling).
  • Experience with Power BI Data Modeling and Semantic Layer development.
  • Proficiency in DAX (Data Analysis Expressions).
  • Experience designing and managing Semantic Models in Power BI.
  • Strong SQL skills and experience working with large datasets.
  • Knowledge of data warehousing concepts and best practices.


Preferred Skills

  • Experience with cloud platforms such as Azure, AWS, or GCP.
  • Exposure to modern data platforms like Databricks.
  • Understanding of data governance and data quality frameworks.
Read more
company logo
shruthi k
Posted by shruthi k
Remote, Pune
7 - 18 yrs
₹1L - ₹35L / yr (ESOP available)
microsoft fabric,dataengineering,one lake

Data Engineer – Microsoft Fabric

Location: Pune, India

Work Mode: Hybrid

Experience: 6+ Years

Employment Type: Full-time contactor

Compensation: As per market standards, commensurate with experience and expertise

Shift Timings: 2:00 PM – 11:00 PM IST

Notice Period: 0 – 15 days

About the Role

Jade Business Services (JBS) is seeking a Data Engineer – Microsoft Fabric to join our Pune team and work on enterprise-scale data transformation and analytics initiatives.

We are looking for a hands-on Data Engineer with strong experience in Microsoft Fabric, SQL, Python/PySpark and modern data engineering practices. The candidate will be responsible for building scalable data pipelines, implementing Lakehouse and Warehouse solutions, developing data models and supporting governed, reliable and AI-ready data platforms.

The ideal candidate should be comfortable working with architects, engineering teams and client stakeholders to translate business requirements into scalable and production-ready data solutions.

Roles and Responsibilities

  • Design and develop data solutions using Microsoft Fabric, including OneLake, Lakehouse, Warehouse and Data Factory pipelines.
  • Build and maintain scalable ETL/ELT pipelines for batch and incremental data processing.
  • Develop data ingestion and transformation pipelines using Fabric Data Factory, SQL, Python and/or PySpark.
  • Implement Medallion Architecture using Bronze, Silver and Gold layers.
  • Work with Lakehouse and Fabric Warehouse for enterprise data processing and analytics.
  • Develop and maintain data models, tables, views and optimized SQL queries.
  • Build and support semantic models for Power BI and analytical workloads.
  • Implement data quality, validation, monitoring and error-handling mechanisms.
  • Work with metadata, lineage and governance requirements using Microsoft Purview.
  • Implement data security, access controls and role-based permissions across data platforms.
  • Support Data Product and domain-oriented data architecture principles.
  • Follow DataOps practices including CI/CD, deployment, monitoring and production support.
  • Troubleshoot pipeline failures, performance issues and data quality problems.
  • Optimize data pipelines, queries and storage for performance and cost efficiency.
  • Work closely with Data Architects and business stakeholders to understand requirements and implement technical solutions.
  • Participate in technical design discussions, code reviews and architecture reviews.
  • Maintain technical documentation, data flow diagrams and pipeline documentation.
  • Support production deployments, incident resolution and SLA-driven data platform operations.
  • Identify opportunities for automation and AI-assisted improvements across data engineering processes.

Qualifications and Skills

  • 6+ years of experience in Data Engineering, Data Integration or Data Platform development.
  • Strong hands-on experience with Microsoft Fabric.
  • Experience with:
  • Microsoft Fabric Lakehouse
  • Fabric Warehouse
  • OneLake
  • Fabric Data Factory / Pipelines
  • Semantic Models
  • Strong understanding of Lakehouse and Medallion Architecture.
  • Strong SQL development and query optimization skills.
  • Hands-on experience with Python and/or PySpark.
  • Experience developing enterprise ETL/ELT and data integration pipelines.
  • Experience with batch and incremental data processing.
  • Understanding of data modelling concepts including dimensional modelling.
  • Knowledge of data quality, metadata, lineage and data governance.
  • Working knowledge of Microsoft Purview.
  • Understanding of Data Mesh and Data Product concepts.
  • Experience with CI/CD, version control, monitoring and DataOps practices.
  • Understanding of cloud security, access controls and data privacy.
  • Good troubleshooting and problem-solving skills.
  • Strong communication skills and ability to work with distributed and client-facing teams.

Preferred Skills

  • Microsoft Fabric or Azure Data certifications.
  • Experience migrating workloads from Azure Synapse, SQL Server, Databricks or other data platforms to Microsoft Fabric.
  • Experience implementing Medallion Architecture on Microsoft Fabric.
  • Experience with Power BI and semantic modelling.
  • Exposure to AI/ML, Generative AI or Agentic AI use cases on enterprise data platforms.
  • Experience working with Data Products or domain-oriented data solutions.
  • Experience in Energy & Utilities, Healthcare, Financial Services or Insurance.
  • Experience working with US or international enterprise clients.

What We Expect

The ideal candidate should be hands-on first and capable of independently building, troubleshooting and optimizing Fabric data solutions. You should be able to explain the technical decisions behind your implementation and work effectively with architects and engineering teams to deliver production-ready solutions.

 

Read more
company logo
Bengaluru (Bangalore), Mumbai, Pune, Hyderabad, Noida, Kolkata
8 - 15 yrs
₹13L - ₹20L / yr
Azure Data Factory
Azure Databricks
PySpark
skill iconPython
SQL
+4 more

Roles & Responsibilities

  • Design, develop, and deliver scalable end-to-end data pipelines using Azure Data Factory, ensuring robust integration

of enterprise-wide data from diverse sources

• Build and optimize data engineering workflows using Databricks and PySpark

• Write efficient, high-performance SQL for data transformation and analysis

• Work with the Azure Cloud platform and associated services, applying strong understanding of data warehousing,

data models, and pipelines

• Provide technical leadership to a team of developers, including code reviews and enforcing best practices across the

development lifecycle

• Oversee CI/CD implementation using Azure DevOps, managing deployments across development, QA, and production

environments with proper change control processes

• Collaborate with cross-functional teams to translate business requirements into scalable data solutions

• Ensure data quality, reliability, and performance across all pipelines and platforms

Ideal Candidate

1Strong Azure Databricks Engineer / Senior Data Engineer Profile

2Mandatory (Experience 1) – Must have minimum 8+ years of overall experience in Data Engineering, Data Development, or related data technology roles, with strong hands-on experience in enterprise data pipeline development.

3Mandatory (Experience 2) – Must have strong hands-on experience with Azure Databricks, including development and optimization of scalable data engineering workflows using Databricks and PySpark.

4Mandatory (Experience 3) – Must have strong hands-on proficiency in PySpark/Python and SQL, with proven experience developing complex data transformations, processing workflows, and performance-optimized queries.

5Mandatory (Experience 4) – Must have hands-on experience with Azure Data Factory (ADF) for designing, developing, and orchestrating end-to-end data pipelines and integrating data from multiple sources.

6Mandatory (Experience 5) – Must have strong experience working on the Azure Cloud platform and associated data services, with solid understanding of data warehousing, data modeling, pipeline architecture, and enterprise data solutions.

7Mandatory (Experience 6) – Must have hands-on experience implementing CI/CD using Azure DevOps, including deployment and release management across development, QA, and production environments.

8Mandatory (Experience 7) – Must have proven technical leadership experience, including code reviews, enforcing development best practices, mentoring developers, and providing technical guidance to a data engineering team.

9Mandatory (Notice Period) – Immediate joiners or candidates who can join within 15 days.

10Mandatory (Note) - The position is open across all Cognizant offices pan India. Candidates must be willing to attend the F2F interview at the nearest Cognizant office location.

Read more
company logo
Remote only
5 - 10 yrs
Best in industry
ETL
Google Cloud Platform (GCP)
skill iconKubernetes
skill iconScala
Apache Spark

Description

We are looking for Senior Data Engineers to join our AdTech team and build scalable, high-performance data platforms that power advertising insights and analytics. The ideal candidate will have strong experience in distributed data processing, ETL pipelines, and Big Data technologies, with hands-on expertise in Spark and Scala.

You will work on designing, developing, and optimizing large-scale data pipelines while collaborating closely with cross-functional engineering teams to build reliable, production-grade data solutions.


Key Responsibilities

  • Design, develop, and maintain scalable ETL pipelines for large-scale data processing.
  • Build and optimize distributed data applications using Spark and Scala.
  • Develop reliable, high-performance data pipelines for batch and streaming workloads.
  • Work with large datasets to ensure data quality, consistency, and performance.
  • Collaborate with engineering, product, and analytics teams to deliver robust data solutions.
  • Optimize data workflows for scalability, reliability, and cost efficiency.
  • Deploy and manage data workloads in cloud and containerized environments.
  • Troubleshoot production issues and continuously improve platform performance.


Requirements

Candidates who demonstrate:

  • 5+ years of experience in Data Engineering or Big Data Engineering.
  • Strong hands-on experience with Apache Spark and Scala.
  • Experience building and maintaining ETL pipelines.
  • Familiarity with Google Cloud Storage (GCS).
  • Experience with Kubernetes (K8s).
  • Strong SQL skills and understanding of distributed data processing.
  • Excellent debugging, problem-solving, and performance optimization skills.
  • Strong communication and collaboration skills.



Good to Have

  • Experience with AWS and cloud-native data services.
  • Familiarity with streaming technologies such as Kafka.
  • Experience working on large-scale data platforms or AdTech systems.
  • Exposure to orchestration tools such as Airflow.


Benefits

  • Best-in-class salary: We hire strong talent and compensate accordingly.
  • Proximity Talks: Meet and learn from designers, engineers, product leaders, and AI practitioners.
  • Continuous learning: Work with a world-class team and stay close to the latest in AI, engineering, and product development.
  • High-impact work: Build AI-first systems and products used at scale by global clients.



About Us

Proximity is the trusted technology, design, and consulting partner for some of the biggest Sports, Media, and Entertainment companies in the world. We’re headquartered in San Francisco and have offices in Palo Alto, Dubai, Mumbai, and Bangalore.

Since 2019, Proximity has built high-impact, scalable products used by millions of users every day. Today, we are a global team of engineers, designers, product managers, and experts solving complex problems and building cutting-edge technology at scale.

Read more
company logo
Akshay Patil
Posted by Akshay Patil
Noida, Bengaluru (Bangalore), Pune, Hyderabad, Chennai
6 - 8 yrs
₹6L - ₹12L / yr
Data engineering
databricks
Snow flake schema
skill iconPython
Apache Spark
+8 more

Job Title : Data Engineer – Databricks

Experience : 6+ Years

Location : Noida / Hyderabad / Chennai / Pune / Bengaluru (Hybrid)

Shift : IST (Normal Shift)


Job Summary :

We are seeking an experienced Data Engineer with strong expertise in Databricks, Snowflake, Python, and Spark to build and optimize scalable data pipelines and support AI/ML model deployments. The ideal candidate should have experience working with cloud-based data platforms and preferably possess exposure to the Healthcare domain.


Required Skills :

  • Databricks (Preferred)
  • Snowflake
  • Python
  • Apache Spark
  • SQL
  • Azure Cloud
  • Kubernetes
  • Apache Airflow
  • GitHub & CI/CD Pipelines
  • AI/ML Model Deployment
  • Data Analytics

Preferred :

  • Experience in the Healthcare domain.
  • Strong understanding of scalable data engineering architectures and best practices.
Read more
Why apply to jobs via Cutshort
people_solving_puzzle
Personalized job matches
Stop wasting time. Get matched with jobs that meet your skills, aspirations and preferences.
people_verifying_people
Verified hiring teams
See actual hiring teams, find common social connections or connect with them directly.
ai_chip
Move faster with AI
We use AI to get you faster responses, recommendations and unmatched user experience.
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos