BACKUP DATA ETL DEVELOPER at MNC · Mumbai · 6 - 9 years · ₹8L - ₹15L / yr · Posted 6 Oct 2026

Position: Backup Data ETL Developer
Experience: 6–9 Years
Location: Mumbai
Employment Type: Full-time
Job Description:
We are looking for a Backup Data ETL Developer with 6–9 years of experience in designing, developing, and supporting data extraction, transformation, and loading (ETL) processes.
The candidate should have strong hands-on experience with Hadoop, Spark, Informatica, and Snowflake, along with expertise in large-scale data processing and data integration technologies.
Key Skills:
Hadoop
Spark
Informatica
Snowflake
ETL Development
Data Integration
Data Transformation
Large-Scale Data Processing
ETL Workflows and Data Pipelines
Troubleshooting and Problem-Solving
Strong analytical and collaboration skills

Similar jobs (10)
Must have experience in Java
Must have experience in Spark
Must have experience in ETL coding
Strong expertise in coding
Data Engineer Hiring Post
🚨 Hiring: Data Engineer | PySpark + Python + SQL
We are looking for experienced Data Engineers to join our team!
🔹 Experience: 5 to 9 Years
🔹 Locations: Bangalore / Hyderabad
🔹 Interview Process:
• 1st Round – Virtual
• 2nd Round – Face-to-Face (Karat Test)
🔑 Key Skills:
✅ PySpark
✅ SQL
✅ Python
✅ ETL
📩 Interested candidates can share their updated resume.
#Hiring #DataEngineer #PySpark #Python #SQL #ETL #BangaloreJobs #HyderabadJobs #TechHiring #ImmediateHiring
Data Engineer Short Hiring Post
🚨 Hiring: Data Engineer
🔹 Experience: 5–9 Years
🔹 Location: Bangalore / Hyderabad
🔹 Skills: PySpark, Python, SQL, ETL, CI/CD, Data Modeling
🔹 Process: L1 Virtual → L2 F2F Karat Test
🔹 F2F: Bangalore / Hyderabad Location
🔹 Positions: Immediate requirement
⚠️ Note: Candidates must be available for F2F Karat immediately after L1.
#Hiring #DataEngineer #PySpark #Python #SQL #BangaloreJobs #HyderabadJobs #Mphasis #ImmediateJoiners
Description
We are looking for Senior Data Engineers to join our Data Platform team and build scalable, high-performance data platforms that power data processing, analytics, and downstream applications.
The ideal candidate will have strong experience in distributed data processing, ETL pipelines, and Big Data technologies, with hands-on expertise in Apache Spark and Python Scala.
You will be responsible for designing, developing, and optimizing large-scale data pipelines while collaborating closely with cross-functional engineering teams to build reliable, production-grade data solutions.
Key Responsibilities
- Design, develop, and maintain scalable ETL and data processing pipelines for large-scale datasets.
- Build and optimize distributed data applications using Apache Spark and Python Scala.
- Develop reliable, high-performance data pipelines for batch and streaming workloads.
- Design and manage data workflows using Apache Airflow.
- Build and operate data workloads on AWS, with strong usage of Amazon S3 for large-scale data storage.
- Work with large datasets to ensure data quality, consistency, reliability, and performance.
- Collaborate with engineering, product, analytics, and other platform teams to deliver robust data solutions.
- Optimize data workflows for scalability, reliability, performance, and cost efficiency.
- Troubleshoot production issues, identify bottlenecks, and continuously improve platform performance.
Requirements
Candidates who demonstrate:
- 5+ years of experience in Data Engineering, Big Data Engineering, or a similar role.
- Strong hands-on experience with Apache Spark and Scala.
- Experience designing, building, and maintaining large-scale ETL pipelines.
- Strong hands-on experience with AWS, particularly Amazon S3.
- Hands-on experience with Apache Airflow for workflow orchestration and scheduling.
- Strong SQL skills and a solid understanding of distributed data processing concepts.
- Experience working with batch and/or streaming data pipelines.
- Excellent debugging, problem-solving, and performance optimization skills.
- Strong communication and collaboration skills.
Good to Have
- Experience with Databricks and the broader Databricks data platform.
- Familiarity with streaming technologies such as Apache Kafka.
- Experience working on large-scale data platforms handling high-volume data workloads.
- Exposure to additional AWS data services and cloud-native data architectures.
Description
We are looking for Senior Data Engineers to join our AdTech team and build scalable, high-performance data platforms that power advertising insights and analytics. The ideal candidate will have strong experience in distributed data processing, ETL pipelines, and Big Data technologies, with hands-on expertise in Spark and Scala.
You will work on designing, developing, and optimizing large-scale data pipelines while collaborating closely with cross-functional engineering teams to build reliable, production-grade data solutions.
Key Responsibilities
- Design, develop, and maintain scalable ETL pipelines for large-scale data processing.
- Build and optimize distributed data applications using Spark and Scala.
- Develop reliable, high-performance data pipelines for batch and streaming workloads.
- Work with large datasets to ensure data quality, consistency, and performance.
- Collaborate with engineering, product, and analytics teams to deliver robust data solutions.
- Optimize data workflows for scalability, reliability, and cost efficiency.
- Deploy and manage data workloads in cloud and containerized environments.
- Troubleshoot production issues and continuously improve platform performance.
Requirements
Candidates who demonstrate:
- 5+ years of experience in Data Engineering or Big Data Engineering.
- Strong hands-on experience with Apache Spark and Scala.
- Experience building and maintaining ETL pipelines.
- Familiarity with Google Cloud Storage (GCS).
- Experience with Kubernetes (K8s).
- Strong SQL skills and understanding of distributed data processing.
- Excellent debugging, problem-solving, and performance optimization skills.
- Strong communication and collaboration skills.
Good to Have
- Experience with AWS and cloud-native data services.
- Familiarity with streaming technologies such as Kafka.
- Experience working on large-scale data platforms or AdTech systems.
- Exposure to orchestration tools such as Airflow.
Benefits
- Best-in-class salary: We hire strong talent and compensate accordingly.
- Proximity Talks: Meet and learn from designers, engineers, product leaders, and AI practitioners.
- Continuous learning: Work with a world-class team and stay close to the latest in AI, engineering, and product development.
- High-impact work: Build AI-first systems and products used at scale by global clients.
About Us
Proximity is the trusted technology, design, and consulting partner for some of the biggest Sports, Media, and Entertainment companies in the world. We’re headquartered in San Francisco and have offices in Palo Alto, Dubai, Mumbai, and Bangalore.
Since 2019, Proximity has built high-impact, scalable products used by millions of users every day. Today, we are a global team of engineers, designers, product managers, and experts solving complex problems and building cutting-edge technology at scale.
Key Responsibilities
Build and maintain data transformation pipelines using java Spark
Develop and optimize large-scale/CPU intensive data processing using Apache Spark
Orchestrate workflows using Airflow
Implement data quality checks, testing, and monitoring for pipeline. Good to have exposer into managing metadata, cataloguing, and lineage
Support schema evolution, backfills, and incremental processing
Ensure pipelines meet SLAs for freshness, reliability, and performance
Expertise/working knowledge in Spark and HBase(semantic layer, virtual datasets, Reflections)
Required Skills & Qualifications
Strong hands-on experience with Apache Spark
Experience with HBase/SQL or similar lakehouse query engines
Airflow
Understanding of data catalogs and lineage (e.g., OpenLineage, DataHub, Apache Polaris , openlineage)
Proficiency in Java
Experience with Git-based development and CI/CD
Location: Hyderabad / Chennai
Experience: 5+ years
Employment type: Full-time, permanent
Work Hours: General Shift
website: www.amazech.com
Qualifications:
- B.E./B.Tech/M.E./M.Tech in Computer Science, Information Technology, Data Science, or related disciplines.
- Strong academic background with relevant industry experience in Data Engineering and Data Warehousing.
Key Responsibilities:
· Design, develop, and maintain scalable data warehouse solutions using Snowflake.
· Write, optimize, troubleshoot, and enhance Snowflake SQL queries with a focus on performance and scalability.
· Develop and support ETL processes using Talend to ensure reliable and efficient data movement.
· Collaborate with business, analytics, and application teams to enable reporting, dashboards, metrics, and data exploration capabilities.
· Perform data analysis and resolve issues across data ingestion, transformation, and reporting pipelines.
· Debug and troubleshoot Python-based data processing scripts and automation workflows.
· Implement best practices for data quality, testing, deployment, and code reviews.
· Work across UI, API, and Data Warehouse layers to support end-to-end data integration and business requirements.
· Monitor, optimize, and maintain data warehouse performance and operational stability.
· Create and maintain technical documentation, data models, and process workflows.
Required Skills and Experience:
· Strong hands-on expertise in Snowflake Data Warehouse.
· Advanced SQL skills with experience handling large-scale datasets.
· Strong understanding of Data Warehousing concepts, dimensional modelling, and data architecture.
· Hands-on experience with Analytical SQL functions, query tuning, and performance optimization.
· Experience developing and maintaining ETL solutions using Talend.
· Proficiency in Python for scripting, debugging, automation, and data processing.
· Experience integrating UI, API, and Data Warehouse workflows.
· Strong problem-solving and analytical skills.
· Experience with testing, code reviews, and deployment best practices.
· Excellent communication and stakeholder management skills.
We are hiring an Informatica ETL Developer to build and maintain enterprise data integration pipelines.
Responsibilities
- Develop ETL mappings and workflows in Informatica PowerCenter and IICS
- Build data warehouse loads and transformations
- Tune ETL performance and troubleshoot failures
- Migrate workloads from PowerCenter to IICS
- Document data flows and lineage
Requirements
- 2+ years of Informatica ETL development
- Strong data warehousing concepts
- Exposure to Talend is a plus
Strong Senior Developer – PL/SQL, SQL & ETL (Microsoft SSIS) Profile
2
Mandatory (Experience 1) – Must have minimum 5+ years of strong hands-on experience in PL/SQL and SQL development, including complex stored procedures, functions, queries, joins, data manipulation, and query/performance optimization.
3
Mandatory (Experience 2) – Must have strong hands-on experience in ETL development using Microsoft SSIS, including building, maintaining, optimizing, and troubleshooting SSIS packages for large-volume data movement and transformation.
4
Mandatory (Experience 3) – Must have solid experience working with Data Warehousing concepts and architectures, including data models, fact/dimension structures, ETL data flows, and enterprise reporting/data warehouse environments.
5
Mandatory (Experience 4) – Must have experience managing batch jobs, scheduling, and data pipelines, ensuring timely and reliable execution of enterprise ETL workflows.
6
Mandatory (Experience 5) – Must have hands-on experience in production support for SSIS/ETL and data warehouse jobs, including monitoring job execution, troubleshooting failures, performing root cause analysis, and implementing preventive fixes.
7
Mandatory (Experience 6) – Must have experience with data quality, validation, and troubleshooting, including identifying and resolving data discrepancies/issues affecting downstream reports, dashboards, and analytics.
8
Mandatory (Experience 7) – Must have experience with unit, integration, and regression testing of SQL, PL/SQL, and ETL components, along with strong documentation of technical designs, data mappings, data flows, and deployment processes.
9
Mandatory (Location) – Must be willing to work in a hybrid model from a city where Cognizant has an office.
10
Mandatory (Notice Period) – Immediate joiners or candidates who can join within 2–4 weeks.
Responsibilities and JD
Job Description: We are looking for a Senior Developer with strong expertise in PySpark, Databricks, and Snowflake to build scalable data engineering solutions and enterprise data platforms.
Key Responsibilities:
- Design, develop, and maintain ETL/ELT pipelines using PySpark, Databricks, and Snowflake.
- Develop batch and real-time data processing solutions for structured and semi-structured data.
- Build and optimize Databricks notebooks, workflows, and Delta Lake solutions.
- Design and implement Snowflake databases, schemas, views, stored procedures, tasks, and streams.
- Develop scalable data models, data marts, and data warehouse solutions.
- Optimize PySpark jobs, Databricks workloads, and Snowflake queries for performance and cost efficiency.
- Implement data quality, validation, governance, and security controls.
- Collaborate with business stakeholders, architects, and cross-functional teams to deliver data solutions.
- Manage source control and CI/CD deployments using Git and Azure DevOps.
- Troubleshoot production issues, perform root cause analysis, and ensure pipeline reliability.
- Mentor junior team members and participate in code reviews and technical design discussions.
Required Skills: PySpark, Databricks, Snowflake, Python, SQL.
Experience: 5+ years of Data Engineering experience with strong hands-on expertise in PySpark, Databricks, and Snowflake.







