Cutshort logo
For Employers
Xebia IT Architects logo
GCP Senior Data Engineer
GCP Senior Data Engineer

GCP Senior Data Engineer at Xebia IT Architects · Bengaluru (Bangalore), Gurugram, Pune, Hyderabad, Chennai, Bhopal, Jaipur · 10 - 15 years · ₹30L - ₹40L / yr · Profitable · Posted 17 May 2025

Xebia IT Architects's logo

GCP Senior Data Engineer

Vijay S's profile picture
Posted by Vijay S
10 - 15 yrs
₹30L - ₹40L / yr
Bengaluru (Bangalore), Gurugram, Pune, Hyderabad, Chennai, Bhopal, Jaipur
Skills
Spark
Google Cloud Platform (GCP)
skill iconPython
Apache Airflow
PySpark
SQL

We are looking for a Senior Data Engineer with strong expertise in GCP, Databricks, and Airflow to design and implement a GCP Cloud Native Data Processing Framework. The ideal candidate will work on building scalable data pipelines and help migrate existing workloads to a modern framework.


  • Shift: 2 PM 11 PM
  • Work Mode: Hybrid (3 days a week) across Xebia locations
  • Notice Period: Immediate joiners or those with a notice period of up to 30 days


Key Responsibilities:

  • Design and implement a GCP Native Data Processing Framework leveraging Spark and GCP Cloud Services.
  • Develop and maintain data pipelines using Databricks and Airflow for transforming Raw → Silver → Gold data layers.
  • Ensure data integrity, consistency, and availability across all systems.
  • Collaborate with data engineers, analysts, and stakeholders to optimize performance.
  • Document standards and best practices for data engineering workflows.

Required Experience:


  • 7-8 years of experience in data engineering, architecture, and pipeline development.
  • Strong knowledge of GCP, Databricks, PySpark, and BigQuery.
  • Experience with Orchestration tools like Airflow, Dagster, or GCP equivalents.
  • Understanding of Data Lake table formats (Delta, Iceberg, etc.).
  • Proficiency in Python for scripting and automation.
  • Strong problem-solving skills and collaborative mindset.


⚠️ Please apply only if you have not applied recently or are not currently in the interview process for any open roles at Xebia.


Looking forward to your response!


Best regards,

Vijay S

Assistant Manager - TAG

https://www.linkedin.com/in/vijay-selvarajan/

Read more
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos

About Xebia IT Architects

Founded :
2001
Type :
Products & Services
Size :
100-1000
Stage :
Profitable

About

Xebia explores and creates new frontiers in IT. We provide innovative products and services and strive to stay one step ahead of our customers’ needs.We turn new technology trends into business advantages. As mainstream frontrunners, we create new IT solutions and build the future with our customers.Customers choose Xebia for our innovative solutions, technological depth, and craftsmanship.True knowledge workers find Xebia to be an inspiring place to work where they are challenged by peers.It’s your business, we accelerate it.We are a group of highly ambitious craftsmen. From digital strategy to technology implementation. As such we are a one stop shop for full stack digital transformation. We provide innovative solutions and services to help your organization become a digital winner.We are organized in specialized centers of excellence all over the world. With like-minded individuals aiming for authority in their respective fields.Xebia Group consists of six specialized companies: Xebia Services, Xebia Academy, XebiaLabs, StackState, GoDataDriven, Instruqt, and Xpirit.
Read more

Connect with the team

Profile picture
Amrutesh Iyer
Profile picture
Garima Bhardwaj

Company social profiles

bloglinkedintwitterfacebook

Similar jobs (10)

company logo
Robin Silverster
Posted by Robin Silverster
Bengaluru (Bangalore)
7 - 10 yrs
₹15L - ₹40L / yr
Data engineering
skill iconPython
PySpark
Data Transformation Tool (DBT)
Apache Airflow
+2 more

About the Role

We are looking for a Senior Data Engineer with strong hands-on expertise in Databricks, Python, PySpark, and SQL to build scalable, high-performance data engineering solutions. You’ll architect and develop large scale, high-performance data pipelines capable of handling massive real-time and batch data volumes across multiple business systems. Databricks is the core enterprise data and processing platform for this role. You will also use Apache Airflow for workflow orchestration and dbt for ELT transformations, and will contribute to designing reliable, secure, and governed data platforms that enable analytics, reporting, and AI-driven use cases.

Key Responsibilities

  • Design and implement large-scale data pipelines using Python/PySpark, Databricks, and Microsoft Fabric.
  • Develop and optimize data processing workloads in Databricks using PySpark and Spark SQL, with a strong focus on scalability, reliability, performance, and maintainability.
  • Develop and maintain dbt models including layered architecture, incremental models, snapshots, macros, testing, and documentation.
  • Design, develop, and maintain Apache Airflow DAGs for orchestrating reliable, scalable, and observable data pipelines.
  • Design and implement data quality, observability, and governance frameworks, including automated testing, monitoring, lineage, access control, and data privacy standards.
  • Partner with analytics, product, and business stakeholders to turn requirements into trustworthy datasets, and raise the engineering bar through design discussions, code reviews, and mentoring junior engineers.

Required Skills

  • Strong expertise in Python for developing scalable, modular, and production-ready data engineering applications.
  • Strong expertise in PySpark, including DataFrame API, Spark SQL, Structured Streaming, partitioning strategies, joins, caching, handling data skew, and Spark performance optimization.
  • Strong hands-on experience with Databricks for data ingestion, transformation, processing, and optimization, including Delta Lake, Unity Catalog, Databricks Workflows, notebooks, jobs, and Databricks-native data engineering capabilities.
  • Strong experience in Databricks/Spark performance tuning, including query and job optimization, partitioning, file sizing, caching, join optimization, handling data skew, and efficient use of compute resources.
  • Hands-on experience with Delta Lake, including transactional data processing, schema management, incremental data processing, and reliable batch and streaming data pipelines.
  • Hands-on experience in developing dbt projects using layered architecture, incremental models, snapshots, macros/Jinja, testing, documentation, and deployment best practices.
  • Expertise in advanced SQL and data modelling — dimensional modeling, slowly changing dimensions, schema evolution, and query optimization.
  • Hands-on experience in developing and managing Apache Airflow DAGs, scheduling workflows, dependency management, retries, backfills, and operational monitoring.
  • Hands-on experience with at least one major cloud platform (AWS, Azure or GCP).
  • Strong problem-solving skills and the ability to work independently with business and analytics stakeholders.

Nice to Have

  • Hands-on exposure to Microsoft Fabric for data integration and analytics.
  • Experience using AI coding assistants (e.g. Claude Code, GitHub Copilot) as part of a development workflow.
  • Familiarity with modern DevOps practices, including CI/CD pipelines, Infrastructure as Code (IaC), and containerization (Docker/Kubernetes).
  • Domain expertise in financial services.


Read more
company logo
Noora J
Posted by Noora J
Chennai
5 - 12 yrs
₹8L - ₹30L / yr
Google Vertex AI
Google Cloud Platform (GCP)
skill iconPython
Google BigQuery

Experience: 5+ Years

Employment Type: Full-Time


Role Overview

We are looking for an experienced GCP Data Engineer with 5+ years of experience in data engineering and strong hands-on expertise in Google BigQuery, Google Cloud Storage (GCS), Airflow/Cloud Composer, Python, and Vertex AI. The candidate should be capable of designing, developing, and maintaining scalable data pipelines and cloud-based data solutions on Google Cloud Platform.


Key Skills – Mandatory

  • BigQuery – Strong hands-on experience in data warehousing, SQL, optimization, and performance tuning.
  • Google Cloud Storage (GCS) – Experience with data storage, file management, and integration with data pipelines.
  • Airflow / Cloud Composer – Experience in developing, scheduling, monitoring, and managing data workflows.
  • Python – Strong programming skills for data engineering, ETL/ELT development, automation, and pipeline implementation.
  • Vertex AI – Experience working with ML/AI workflows, model integration, or data pipelines supporting AI/ML solutions.

Good to Have / Added Advantage

  • Dataproc – Experience with distributed data processing and Spark-based workloads.
  • Cloud Data Fusion – Experience in building and managing data integration pipelines.
  • Cloud Run – Understanding of deploying and running containerized applications/services on GCP.
  • Experience with ETL/ELT processes and data pipeline development.
  • Knowledge of GCP data architecture and cloud-native services.
  • Experience in data quality, validation, monitoring, and troubleshooting.

Responsibilities

  • Design, develop, and maintain scalable GCP-based data pipelines.
  • Build and optimize data solutions using BigQuery and Cloud Storage.
  • Develop and manage workflows using Airflow / Cloud Composer.
  • Write efficient and reusable Python code for data processing and automation.
  • Support Vertex AI integrations and AI/ML data workflows.
  • Monitor pipeline performance and troubleshoot data processing issues.
  • Work with cross-functional teams to understand data requirements and deliver reliable solutions.
  • Implement best practices for data security, quality, scalability, and performance.

You must have :

  • 5+ years of overall experience in Data Engineering.
  • Strong hands-on experience with BigQuery, GCS, Airflow/Cloud Composer, Python, and Vertex AI.
  • Strong understanding of data engineering concepts, ETL/ELT, data pipelines, and cloud technologies.
  • Dataproc, Data Fusion, and Cloud Run experience will be an added advantage.


Read more
company logo
Sanikha M
Posted by Sanikha M
Chennai
5 - 8 yrs
₹8L - ₹12L / yr
Google BigQuery
CI/CD
SQL
Google Cloud Platform (GCP)

What you'll need

  • Bachelor's degree in Computer Science, Engineering, Information Systems, or a related technical field.
  • 5+ years of professional data engineering experience.
  • Experience designing and building cloud-native data solutions.
  • Strong expertise with Google Cloud Platform, including BigQuery. Experience developing transformation frameworks using dbt.
  • Strong SQL and Python programming skills.
  • Experience with PostgreSQL or other relational databases.
  • Experience orchestrating workflows using Apache Airflow or Cloud Composer.
  • Experience implementing Infrastructure as Code using Terraform. Experience building CI/CD pipelines using GitHub Actions, GitLab CI, Jenkins, or similar platforms.
  • Experience developing scalable batch and streaming data pipelines. Strong problem-solving skills with the ability to balance scalability, reliability, and cloud cost optimization.


Preferred Qualifications

  • Experience with Pub/Sub, Datastream, Dataflow, Cloud Storage, Cloud Functions, or Cloud Run.
  • Experience building multi-tenant SaaS platforms.
  • Experience implementing metadata-driven governance, lineage, and data quality frameworks.
  • Experience supporting AI, machine learning, or customer-facing analytics platforms.
  • Experience with Kubernetes and Docker.
Read more
company logo
Banu S
Posted by Banu S
Hyderabad
5 - 12 yrs
₹4L - ₹18L / yr
Google Cloud Platform (GCP)
PySpark

Job Summary

We are seeking a highly skilled GCP Data Engineer with strong expertise in Google Cloud Platform (GCP), Python, ETL, and modern data engineering technologies. The ideal candidate should have hands-on experience designing and building scalable data pipelines using BigQuery, Dataflow, Pub/Sub, Airflow, and modern data lake technologies such as Apache Iceberg or Delta Lake.

Key Responsibilities

  • Design, develop, and maintain scalable ETL/ELT data pipelines on Google Cloud Platform.
  • Build and optimize data processing workflows using Python and Google Cloud Dataflow (Apache Beam).
  • Develop and manage large-scale analytical data models in BigQuery.
  • Implement event-driven data ingestion using Google Cloud Pub/Sub.
  • Create, schedule, and monitor workflows using Apache Airflow and Autosys.
  • Design and implement modern data lake architectures using Apache Iceberg or Delta Lake.
  • Optimize query performance, storage, and compute costs in GCP.
  • Ensure data quality, governance, security, and compliance across data platforms.
  • Collaborate with Data Scientists, Analysts, and Application teams to deliver scalable data solutions.
  • Troubleshoot production issues and continuously improve pipeline reliability and performance.

Mandatory Skills

  • Strong hands-on experience with Google Cloud Platform (GCP).
  • Proficiency in Python programming.
  • Experience in designing and implementing ETL/ELT pipelines.
  • Strong knowledge of BigQuery.
  • Experience with Google Cloud Dataflow (Apache Beam).
  • Experience with Google Cloud Pub/Sub.
  • Hands-on experience with Apache Airflow.
  • Experience in job scheduling using Autosys.
  • Experience with modern table formats such as Apache Iceberg or Delta Lake.
  • Strong SQL and data modeling skills.

Preferred Skills

  • Experience with Cloud Storage, Dataproc, Cloud Composer, and Cloud Functions.
  • Knowledge of CI/CD pipelines and DevOps practices.
  • Experience with Docker and Kubernetes.
  • Familiarity with Git and Agile/Scrum methodologies.
  • Knowledge of data warehousing and dimensional modeling.
  • Exposure to streaming and real-time data processing.

Qualifications

  • Bachelor's or Master's degree in Computer Science, Information Technology, Engineering, or a related field.
  • 4–8+ years of experience in Data Engineering with hands-on expertise in GCP technologies.

Required Experience

  • Strong experience in developing enterprise-grade data pipelines using Python and GCP.
  • Hands-on experience with BigQuery, Dataflow, Pub/Sub, and Airflow.
  • Experience scheduling and monitoring batch workflows using Autosys.
  • Experience implementing modern data lake architectures using Apache Iceberg or Delta Lake.
  • Strong understanding of ETL best practices, performance tuning, and data optimization.
  • Excellent analytical, troubleshooting, and problem-solving skills.

Mandatory Skills

  • Google Cloud Platform (GCP)
  • Python
  • ETL
  • BigQuery
  • Autosys
  • Apache Airflow
  • Google Cloud Pub/Sub
  • Google Cloud Dataflow (Apache Beam)
  • Apache Iceberg / Delta Lake
  • SQL & Data Modeling
Read more
company logo
Banu S
Posted by Banu S
Bengaluru (Bangalore), Hyderabad
5 - 18 yrs
₹4L - ₹35L / yr
Google Cloud Platform (GCP)
GCP
Spark
Ab Initio

Senior Data Engineer – Ab Initio | GCP | Spark | Agentic AI

Location: Bangalore

Experience: 5+ Years

Role: Senior Data Engineer

Work Mode: Bangalore

Job Summary

We are looking for an experienced Senior Data Engineer with strong expertise in Ab Initio, GCP, Apache Spark, and Agentic AI. The ideal candidate will have hands-on experience designing and developing scalable data engineering solutions, building data pipelines, and working with modern cloud and AI technologies.

The candidate should be comfortable working across traditional enterprise data platforms and emerging Generative AI / Agentic AI solutions.

Key Responsibilities

  • Design, develop, and maintain scalable and high-performance data pipelines using Ab Initio, Spark, and GCP services.
  • Develop and optimize complex ETL/ELT workflows using Ab Initio.
  • Build and maintain data processing solutions using Apache Spark / PySpark.
  • Develop cloud-based data solutions on Google Cloud Platform (GCP).
  • Work with GCP data services such as BigQuery, Cloud Storage, Dataflow, Dataproc, Pub/Sub, or equivalent services.
  • Perform data integration, transformation, cleansing, and validation.
  • Optimize data pipelines for performance, scalability, reliability, and cost.
  • Collaborate with data architects,

Read more
company logo
Vineeth Kumar
Posted by Vineeth Kumar
Remote only
6 - 14 yrs
Best in industry
SQL
Google Cloud Platform (GCP)
skill iconPython
Google BigQuery

Experience: 6+ years overall Data Engineering experience.

Must-have — candidates should have hands-on experience in ALL of these:

  1. GCP (Google Cloud Platform) – strong hands-on experience
  2. Python – data engineering/ETL development
  3. SQL – advanced SQL, query optimization, data transformation
  4. BigQuery – strong hands-on experience with development, optimization and data warehousing
  5. Data Engineering / ETL – building and maintaining data pipelines
  6. GCP data services – preferably Cloud Storage, Dataflow, Pub/Sub, Composer/Airflow, etc.
  7. Data warehousing / dimensional modeling


Read more
company logo
Resume TGS
Posted by Resume TGS
Hyderabad
7 - 9 yrs
₹12L - ₹24L / yr
ELT
Google BigQuery
skill iconPython
Snow flake schema
SQL
+2 more
  • Design, build, and maintain scalable ETL/ELT pipelines for batch and real-time data ingestion and transformation.
  • Develop and optimize data lake and data warehouse architectures (e.g., Snowflake, BigQuery, Redshift).
  • Work with cloud platforms GCP, Azure to manage data infrastructure.
  • GCP as mandatory skills
  • Collaborate with analytics and product teams to understand data needs and deliver solutions.
  • Ensure data quality, reliability, security, and compliance across all data systems.
  • Mentor junior data engineers and contribute to best practices and code reviews.
  • Monitor and troubleshoot data pipeline performance and resolve data-related issues.
  • Automate data validation, monitoring, and alerting processes.
  • 8+ years of experience in data engineering or software engineering with a data focus.
  • Proficient in SQL and at least one programming language (e.g., Python, Scala, Java).
  • Experience with modern data warehousing tools (e.g., Snowflake, Redshift, BigQuery).
  • Strong understanding of data modeling, data lakes, and ETL/ELT design.
  • Hands-on experience with orchestration tools like Airflow, dbt, or similar.
  • Solid experience with cloud data platforms (AWS/GCP/Azure).
  • Familiarity with CI/CD pipelines, containerization (Docker/Kubernetes), and version control (Git).
  • Experience working in a DevOps or DataOps environment.
  • Knowledge of data governance, lineage, and cataloging tools (e.g., Collibra, Alation).
  • Familiarity with streaming technologies (Kafka, Spark Streaming, Flink).
  • Experience supporting machine learning workflows and data science initiatives.
Read more
Leading US based Internet service provider
Leading US based Internet service provider
Agency job
via by Jason Pinto
Remote only
8 - 12 yrs
₹18L - ₹25L / yr
Google Cloud Platform (GCP)
skill iconPython
SQL

Role Overview

We are looking for a GCP Data Engineer with 10+ years of experience to design, develop, and optimize scalable cloud-based data solutions. The ideal candidate will have strong hands-on expertise in GCP, BigQuery, and advanced SQL, with experience building data pipelines and working with large-scale datasets.


Key Responsibilities

  • Design and develop scalable data pipelines and ETL/ELT processes on GCP.
  • Build, optimize, and maintain data solutions using Google BigQuery.
  • Develop complex SQL queries for data transformation, aggregation, and analysis.
  • Design efficient data models and optimize pipelines for performance, scalability, and cost.
  • Integrate data from multiple sources and ensure data quality, reliability, and availability.
  • Troubleshoot pipeline and data issues and drive continuous improvement.
  • Collaborate with data architects, analysts, application teams, and business stakeholders.
  • Follow best practices for cloud security, data governance, testing, and documentation.


Required Skills

  • 8+ years of Data Engineering experience
  • Strong hands-on experience with GCP, Django, and MongoDB
  • Extensive experience with BigQuery
  • Advanced SQL skills
  • Strong understanding of ETL/ELT and data pipeline development
  • Data modeling and data warehousing experience
  • Experience handling large-scale datasets and performance optimization
  • Strong problem-solving and communication skills


Good to Have

  • GCP services such as Cloud Storage, Dataflow, Pub/Sub, Cloud Composer, or Cloud Functions
  • Python or other data engineering languages
  • Experience with data governance and security
  • Agile development experience
Read more
company logo
Bengaluru (Bangalore)
14 - 25 yrs
₹50L - ₹70L / yr
Data engineering
databricks
Apache Spark
PySpark
skill iconPython
+19 more

Job Title : Senior Data Engineer – Databricks

Experience : 14 to 20 Years

Location : HSR Layout, Bangalore

Work Mode : Hybrid – 3 Days WFO

Shift : 11:30 AM – 07:30 PM IST

Positions : 2

Notice Period : Immediate Joiners Only

Interview : 1 Technical Round + 2 Client Rounds


Role Overview :

We are looking for a Senior Data Engineer to build and lead enterprise-scale data platforms for a Switzerland-based commodity client.

The role requires a strong hands-on Data Engineering professional with expertise in Databricks, PySpark, Python, SQL, and AWS, along with technical leadership and stakeholder management experience.


Must-Have Skills :

  • 14 to 20 years of Data Engineering experience
  • Databricks & Apache Spark / PySpark
  • Python & SQL
  • AWS Cloud
  • Lakehouse Architecture
  • ETL / ELT & Distributed Data Processing
  • Batch & Streaming Pipelines
  • Data Pipeline Optimization & Data Modeling
  • CDC & Incremental Processing
  • Git, CI/CD & Testing
  • Data Quality, Monitoring & Observability
  • Technical Leadership & Stakeholder Management


Key Responsibilities :

  • Design and build scalable data pipelines using Databricks, PySpark, Python, SQL, and AWS.
  • Own data products from design through production.
  • Develop batch / streaming pipelines and reusable ETL / ELT frameworks.
  • Optimize pipelines for performance, scalability, reliability, and cost.
  • Design scalable data architectures and data models.
  • Implement data quality, monitoring, lineage, and CI/CD practices.
  • Lead technical discussions and mentor engineering teams.
  • Collaborate with business stakeholders, architects, product owners, and engineering teams.
  • Remain hands-on while providing technical leadership.


Ideal Candidate :

A 14 to 20 years experienced, hands-on Data Engineering leader with strong Databricks + PySpark + AWS expertise, excellent communication, stakeholder management, and experience delivering enterprise-scale data platforms.

🔴 Super Urgent : Only Bangalore-based immediate joiners.

Read more
company logo
Agency job
via by Ajeethkumar s
Hyderabad, Bengaluru (Bangalore)
5 - 10 yrs
₹4L - ₹16L / yr
skill iconPython
ETL
PySpark
Data engineering
skill iconAmazon Web Services (AWS)
+2 more

Skills Referential (Required knowledge, skills and abilities)

Technical Skills:

Python

Pyspark

SQL

ETL Aws, Azure, gcp

Read more
Why apply to jobs via Cutshort
people_solving_puzzle
Personalized job matches
Stop wasting time. Get matched with jobs that meet your skills, aspirations and preferences.
people_verifying_people
Verified hiring teams
See actual hiring teams, find common social connections or connect with them directly.
ai_chip
Move faster with AI
We use AI to get you faster responses, recommendations and unmatched user experience.
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos