Cutshort logo
For Employers
A fast growing Big Data company  logo
AWS Glue Developer
A fast growing Big Data company
AWS Glue Developer

AWS Glue Developer at A fast growing Big Data company · Noida, Bengaluru (Bangalore), Chennai, Hyderabad · 6 - 8 years · ₹10L - ₹15L / yr · Posted 27 Oct 2023

Careerconnects's logo

AWS Glue Developer

at A fast growing Big Data company

Agency job
6 - 8 yrs
₹10L - ₹15L / yr
Noida, Bengaluru (Bangalore), Chennai, Hyderabad
Skills
AWS Glue
SQL
skill iconPython
PySpark
Data engineering
Big Data
Hadoop
Spark
DMS
Data integration
Data Ops

AWS Glue Developer 

Work Experience: 6 to 8 Years

Work Location:  Noida, Bangalore, Chennai & Hyderabad

Must Have Skills: AWS Glue, DMS, SQL, Python, PySpark, Data integrations and Data Ops, 

Job Reference ID:BT/F21/IND


Job Description:

Design, build and configure applications to meet business process and application requirements.


Responsibilities:

7 years of work experience with ETL, Data Modelling, and Data Architecture Proficient in ETL optimization, designing, coding, and tuning big data processes using Pyspark Extensive experience to build data platforms on AWS using core AWS services Step function, EMR, Lambda, Glue and Athena, Redshift, Postgres, RDS etc and design/develop data engineering solutions. Orchestrate using Airflow.


Technical Experience:

Hands-on experience on developing Data platform and its components Data Lake, cloud Datawarehouse, APIs, Batch and streaming data pipeline Experience with building data pipelines and applications to stream and process large datasets at low latencies.


➢ Enhancements, new development, defect resolution and production support of Big data ETL development using AWS native services.

➢ Create data pipeline architecture by designing and implementing data ingestion solutions.

➢ Integrate data sets using AWS services such as Glue, Lambda functions/ Airflow.

➢ Design and optimize data models on AWS Cloud using AWS data stores such as Redshift, RDS, S3, Athena.

➢ Author ETL processes using Python, Pyspark.

➢ Build Redshift Spectrum direct transformations and data modelling using data in S3.

➢ ETL process monitoring using CloudWatch events.

➢ You will be working in collaboration with other teams. Good communication must.

➢ Must have experience in using AWS services API, AWS CLI and SDK


Professional Attributes:

➢ Experience operating very large data warehouses or data lakes Expert-level skills in writing and optimizing SQL Extensive, real-world experience designing technology components for enterprise solutions and defining solution architectures and reference architectures with a focus on cloud technology.

➢ Must have 6+ years of big data ETL experience using Python, S3, Lambda, Dynamo DB, Athena, Glue in AWS environment.

➢ Expertise in S3, RDS, Redshift, Kinesis, EC2 clusters highly desired.


Qualification:

➢ Degree in Computer Science, Computer Engineering or equivalent.


Salary: Commensurate with experience and demonstrated competence

Read more
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos

Similar jobs (10)

company logo
Shelly Singh
Posted by Shelly Singh
Bengaluru (Bangalore)
8 - 18 yrs
₹5L - ₹18L / yr
ELT
SQL
PySpark
skill iconAmazon Web Services (AWS)
NOSQL Databases

Design, develop, and maintain ETL pipelines involving large-scale data.

Develop data processing and analytics applications primarily using PySpark and Python.

Build scalable and distributed data processing solutions using Apache Spark.

Develop and deploy data applications on AWS cloud.

Work with AWS services related to storage, compute, ETL, data warehousing, analytics, and streaming.

Implement distributed storage and processing solutions capable of handling high-volume datasets.

Design data processing applications with a focus on performance, scalability, reliability, and optimization.

Work with both SQL and NoSQL databases for data storage, processing, and analytics.

Write, optimize, and analyze SQL, HQL, and NoSQL queries.

Troubleshoot data pipeline and processing issues and ensure data quality and reliability.

Collaborate with data engineers, analysts, architects, and other technical teams to deliver data-driven solutions.

Read more
company logo
Hema Dekonda
Posted by Hema Dekonda
Bengaluru (Bangalore), Mumbai, Hyderabad, Gurugram
6 - 11 yrs
₹15L - ₹50L / yr
Data engineering
skill iconAmazon Web Services (AWS)
ETL
SQL
NOSQL Databases

About AuxoAI:


AuxoAI is a global platform-based services firm. We help companies—turn their strategies into practical digital and AI solutions. By understanding how our clients make decisions, we use digital and Artificial Intelligence (AI) technologies to drive growth, enhance their operations, improve customer experiences, and provide clear, actionable insights from their data. What We Do We work across various industries such as healthcare, high-tech, consumer packaged goods (CPG), finance etc., and in sales, marketing, and customer support functions.

We help our clients with accelerating their digital and AI journeys through:

• AI Application Development

• Data, Digital and Cloud acceleration using AI

• AI Native Product Engineering


We are seeking a skilled and experienced Data Engineer to join our dynamic team. The ideal candidate will have 6+ years of prior experience in data engineering, with a strong background in AWS (Amazon Web Services) technologies. This role offers an exciting opportunity to work on diverse projects, collaborating with cross-functional teams to design, build, and optimize data pipelines and infrastructure.


Responsibilities:

* Design, develop, and maintain scalable data pipelines and ETL processes leveraging AWS services such as S3, Glue, EMR, Lambda, and Redshift.

* Collaborate with data scientists and analysts to understand data requirements and implement solutions that support analytics and machine learning initiatives.

* Optimize data storage and retrieval mechanisms to ensure performance, reliability, and cost-effectiveness.

* Implement data governance and security best practices to ensure compliance and data integrity.

* Troubleshoot and debug data pipeline issues, providing timely resolution and proactive monitoring.

* Stay abreast of emerging technologies and industry trends, recommending innovative solutions to enhance data engineering capabilities.


Requirements :

* Bachelor's or Master's degree in Computer Science, Engineering, or a related field.

* 6+ years of prior experience in data engineering, with a focus on designing and building data pipelines.

* Proficiency in AWS services, particularly S3, Glue, EMR, Lambda, and Redshift.

* Strong programming skills in languages such as Python, Java, or Scala.

* Experience with SQL and NoSQL databases, data warehousing concepts, and big data technologies.

* Familiarity with containerization technologies (e.g., Docker, Kubernetes) and orchestration tools (e.g., Apache Airflow) is a plus.

Read more
company logo
Bengaluru (Bangalore)
14 - 25 yrs
₹50L - ₹70L / yr
Data engineering
databricks
Apache Spark
PySpark
skill iconPython
+19 more

Job Title : Senior Data Engineer – Databricks

Experience : 14 to 20 Years

Location : HSR Layout, Bangalore

Work Mode : Hybrid – 3 Days WFO

Shift : 11:30 AM – 07:30 PM IST

Positions : 2

Notice Period : Immediate Joiners Only

Interview : 1 Technical Round + 2 Client Rounds


Role Overview :

We are looking for a Senior Data Engineer to build and lead enterprise-scale data platforms for a Switzerland-based commodity client.

The role requires a strong hands-on Data Engineering professional with expertise in Databricks, PySpark, Python, SQL, and AWS, along with technical leadership and stakeholder management experience.


Must-Have Skills :

  • 14 to 20 years of Data Engineering experience
  • Databricks & Apache Spark / PySpark
  • Python & SQL
  • AWS Cloud
  • Lakehouse Architecture
  • ETL / ELT & Distributed Data Processing
  • Batch & Streaming Pipelines
  • Data Pipeline Optimization & Data Modeling
  • CDC & Incremental Processing
  • Git, CI/CD & Testing
  • Data Quality, Monitoring & Observability
  • Technical Leadership & Stakeholder Management


Key Responsibilities :

  • Design and build scalable data pipelines using Databricks, PySpark, Python, SQL, and AWS.
  • Own data products from design through production.
  • Develop batch / streaming pipelines and reusable ETL / ELT frameworks.
  • Optimize pipelines for performance, scalability, reliability, and cost.
  • Design scalable data architectures and data models.
  • Implement data quality, monitoring, lineage, and CI/CD practices.
  • Lead technical discussions and mentor engineering teams.
  • Collaborate with business stakeholders, architects, product owners, and engineering teams.
  • Remain hands-on while providing technical leadership.


Ideal Candidate :

A 14 to 20 years experienced, hands-on Data Engineering leader with strong Databricks + PySpark + AWS expertise, excellent communication, stakeholder management, and experience delivering enterprise-scale data platforms.

🔴 Super Urgent : Only Bangalore-based immediate joiners.

Read more
company logo
Anisha Jindal
Posted by Anisha Jindal
Bengaluru (Bangalore), Mumbai
5 - 14 yrs
Best in industry
Data engineering
skill iconPython
PySpark
DAX
PowerBI

Job Summary

We are looking for a skilled and experienced Data Engineer to join our growing data team. The ideal candidate will have strong expertise in Python, PySpark, Data Modeling, and Power BI, with hands-on experience in designing, developing, and optimizing scalable data solutions. The role requires working closely with business stakeholders, data architects, and analytics teams to build robust data pipelines and semantic models that enable data-driven decision-making.


Technical Skills

  • Strong hands-on experience in Python and PySpark development.
  • Expertise in building and optimizing Data Engineering solutions and ETL pipelines.
  • Strong understanding of Data Modeling concepts (Star Schema, Snowflake Schema, Dimensional Modeling).
  • Experience with Power BI Data Modeling and Semantic Layer development.
  • Proficiency in DAX (Data Analysis Expressions).
  • Experience designing and managing Semantic Models in Power BI.
  • Strong SQL skills and experience working with large datasets.
  • Knowledge of data warehousing concepts and best practices.


Preferred Skills

  • Experience with cloud platforms such as Azure, AWS, or GCP.
  • Exposure to modern data platforms like Databricks.
  • Understanding of data governance and data quality frameworks.
Read more
company logo
Tushar Vaghela
Posted by Tushar Vaghela
Bengaluru (Bangalore)
5 - 10 yrs
Best in industry
skill iconPython
skill iconScala
Apache Spark
Apache Kafka
databricks
+1 more

Description


We are looking for Senior Data Engineers to join our Data Platform team and build scalable, high-performance data platforms that power data processing, analytics, and downstream applications.

The ideal candidate will have strong experience in distributed data processing, ETL pipelines, and Big Data technologies, with hands-on expertise in Apache Spark and Python Scala.

You will be responsible for designing, developing, and optimizing large-scale data pipelines while collaborating closely with cross-functional engineering teams to build reliable, production-grade data solutions.



Key Responsibilities

  • Design, develop, and maintain scalable ETL and data processing pipelines for large-scale datasets.
  • Build and optimize distributed data applications using Apache Spark and Python Scala.
  • Develop reliable, high-performance data pipelines for batch and streaming workloads.
  • Design and manage data workflows using Apache Airflow.
  • Build and operate data workloads on AWS, with strong usage of Amazon S3 for large-scale data storage.
  • Work with large datasets to ensure data quality, consistency, reliability, and performance.
  • Collaborate with engineering, product, analytics, and other platform teams to deliver robust data solutions.
  • Optimize data workflows for scalability, reliability, performance, and cost efficiency.
  • Troubleshoot production issues, identify bottlenecks, and continuously improve platform performance.



Requirements

Candidates who demonstrate:

  • 5+ years of experience in Data Engineering, Big Data Engineering, or a similar role.
  • Strong hands-on experience with Apache Spark and Scala.
  • Experience designing, building, and maintaining large-scale ETL pipelines.
  • Strong hands-on experience with AWS, particularly Amazon S3.
  • Hands-on experience with Apache Airflow for workflow orchestration and scheduling.
  • Strong SQL skills and a solid understanding of distributed data processing concepts.
  • Experience working with batch and/or streaming data pipelines.
  • Excellent debugging, problem-solving, and performance optimization skills.
  • Strong communication and collaboration skills.


Good to Have

  • Experience with Databricks and the broader Databricks data platform.
  • Familiarity with streaming technologies such as Apache Kafka.
  • Experience working on large-scale data platforms handling high-volume data workloads.
  • Exposure to additional AWS data services and cloud-native data architectures.
Read more
MNC
MNC
Agency job
via by aafia parveen
Hyderabad
5 - 8 yrs
₹2L - ₹20L / yr
Data engineering
Google Cloud Platform (GCP)
Oracle
PySpark
ETL

Job Title: Data Engineer – PySpark | Oracle | GCP


Experience: 5–7 Years

Location: Hyderabad

Notice Period: Immediate Joiners Preferred


Job Summary

We are seeking an experienced Data Engineer with strong expertise in PySpark, Oracle, and Google Cloud Platform (GCP) to design, develop, and optimize scalable data pipelines. The ideal candidate should have hands-on experience in ETL development, data integration, and cloud-based data engineering solutions.

Key Responsibilities


  • Design, develop, and maintain scalable ETL/data pipelines using PySpark.
  • Extract, transform, and load data from Oracle databases into GCP environments.
  • Build and optimize batch data processing workflows for high performance and reliability.
  • Develop data engineering solutions using GCP services.
  • Ensure data quality through validation, monitoring, and troubleshooting.
  • Optimize SQL queries and ETL jobs for performance and scalability.


Required Skills

  • 5–7 years of experience as a Data Engineer.
  • Strong hands-on experience with PySpark.
  • Solid experience with Oracle Database and advanced SQL.
  • Hands-on experience with Google Cloud Platform (GCP).
  • Strong understanding of ETL processes and data warehousing concepts.


Work Location: Hyderabad

Notice Period: Immediate Joiners Preferred

Read more
It is an Product Based Company(Domain- EV Charging)
It is an Product Based Company(Domain- EV Charging)
Agency job
via by Mantasha Naaz
Bengaluru (Bangalore)
3 - 5 yrs
₹13L - ₹15L / yr
skill iconAmazon Web Services (AWS)
skill iconPython
PySpark
SQL
ETL
+2 more

Data Engineer

Location: Bengaluru, India (Hybrid)

Employment Type: Full-time

Experience: 3-5 years



Role Overview  

What We’re Looking For:

  • Bachelor’s degree in Computer Science/Engineering or equivalent experience required.
  • Experience designing and shipping cloud services products.
  • Experience driving and managing technical and architectural dependencies on AWS Cloud.
  • A firm understanding of system architecture, cloud computing, PaaS/SaaS design principles, S3, DynamoDB, RDS mandatory.
  • Experience in building or maintaining ETL processes and tools, i.e., AWS Glue or any open-source tool.
  • Proven system-level design contribution to a current “Live” (in production / under daily high load) multi-region SaaS or PaaS offering.
  • Proven experience with S3, DynamoDB, SQL, and AWS RDS services.
  • Proficiency in programming languages such as Python.
  • Strong analytical and problem-solving skills.

Required Skills & Experience

  • Experience with Python, SQL, and data visualization/exploration tools.
  • Familiarity with the AWS ecosystem, specifically S3, DynamoDB, and RDS.
  • Communication skills, especially for explaining technical concepts to nontechnical business leaders.
  • Ability to work on a dynamic, research-oriented team that has concurrent projects.
  • Experience in AWS cost optimization (Savings Plans, Reserved Instances, Spot Instances) and governance frameworks.
  • Experience developing solutions using infrastructure orchestration tools (SSM, automation account, Ansible, etc.).
  • Excellent leadership, stakeholder management, and communication skills.

 

What We Offer

  • Work with some of the brightest minds in the emerging EV industry.
  • Make a tangible impact in reducing carbon emissions and enabling sustainable energy.
  • Freedom to suggest, implement, and innovate on systems, processes, and technologies.
  • Daily ownership in a high-growth, challenging environment.
  • Flexible work environment with hybrid schedules and virtualization options.
  • Competitive pay and benefits including health coverage, innovative PTO program, and performance bonuses.


Read more
company logo
Agency job
via by Ajeethkumar s
Hyderabad, Bengaluru (Bangalore)
5 - 10 yrs
₹4L - ₹16L / yr
skill iconPython
ETL
PySpark
Data engineering
skill iconAmazon Web Services (AWS)
+2 more

Skills Referential (Required knowledge, skills and abilities)

Technical Skills:

Python

Pyspark

SQL

ETL Aws, Azure, gcp

Read more
company logo
Pavithra E
Posted by Pavithra E
Hyderabad, Pune
5 - 9 yrs
₹18L - ₹20L / yr
PySpark
SQL
skill iconPython

Data Engineer Short Hiring Post


🚨 Hiring: Data Engineer

🔹 Experience: 5–9 Years

🔹 Location: Bangalore / Hyderabad

🔹 Skills: PySpark, Python, SQL, ETL, CI/CD, Data Modeling

🔹 Process: L1 Virtual → L2 F2F Karat Test

🔹 F2F: Bangalore / Hyderabad Location

🔹 Positions: Immediate requirement

⚠️ Note: Candidates must be available for F2F Karat immediately after L1.

#Hiring #DataEngineer #PySpark #Python #SQL #BangaloreJobs #HyderabadJobs #Mphasis #ImmediateJoiners

Read more
company logo
Robin Silverster
Posted by Robin Silverster
Bengaluru (Bangalore)
7 - 10 yrs
₹15L - ₹40L / yr
Data engineering
skill iconPython
PySpark
Data Transformation Tool (DBT)
Apache Airflow
+2 more

About the Role

We are looking for a Senior Data Engineer with strong hands-on expertise in Databricks, Python, PySpark, and SQL to build scalable, high-performance data engineering solutions. You’ll architect and develop large scale, high-performance data pipelines capable of handling massive real-time and batch data volumes across multiple business systems. Databricks is the core enterprise data and processing platform for this role. You will also use Apache Airflow for workflow orchestration and dbt for ELT transformations, and will contribute to designing reliable, secure, and governed data platforms that enable analytics, reporting, and AI-driven use cases.

Key Responsibilities

  • Design and implement large-scale data pipelines using Python/PySpark, Databricks, and Microsoft Fabric.
  • Develop and optimize data processing workloads in Databricks using PySpark and Spark SQL, with a strong focus on scalability, reliability, performance, and maintainability.
  • Develop and maintain dbt models including layered architecture, incremental models, snapshots, macros, testing, and documentation.
  • Design, develop, and maintain Apache Airflow DAGs for orchestrating reliable, scalable, and observable data pipelines.
  • Design and implement data quality, observability, and governance frameworks, including automated testing, monitoring, lineage, access control, and data privacy standards.
  • Partner with analytics, product, and business stakeholders to turn requirements into trustworthy datasets, and raise the engineering bar through design discussions, code reviews, and mentoring junior engineers.

Required Skills

  • Strong expertise in Python for developing scalable, modular, and production-ready data engineering applications.
  • Strong expertise in PySpark, including DataFrame API, Spark SQL, Structured Streaming, partitioning strategies, joins, caching, handling data skew, and Spark performance optimization.
  • Strong hands-on experience with Databricks for data ingestion, transformation, processing, and optimization, including Delta Lake, Unity Catalog, Databricks Workflows, notebooks, jobs, and Databricks-native data engineering capabilities.
  • Strong experience in Databricks/Spark performance tuning, including query and job optimization, partitioning, file sizing, caching, join optimization, handling data skew, and efficient use of compute resources.
  • Hands-on experience with Delta Lake, including transactional data processing, schema management, incremental data processing, and reliable batch and streaming data pipelines.
  • Hands-on experience in developing dbt projects using layered architecture, incremental models, snapshots, macros/Jinja, testing, documentation, and deployment best practices.
  • Expertise in advanced SQL and data modelling — dimensional modeling, slowly changing dimensions, schema evolution, and query optimization.
  • Hands-on experience in developing and managing Apache Airflow DAGs, scheduling workflows, dependency management, retries, backfills, and operational monitoring.
  • Hands-on experience with at least one major cloud platform (AWS, Azure or GCP).
  • Strong problem-solving skills and the ability to work independently with business and analytics stakeholders.

Nice to Have

  • Hands-on exposure to Microsoft Fabric for data integration and analytics.
  • Experience using AI coding assistants (e.g. Claude Code, GitHub Copilot) as part of a development workflow.
  • Familiarity with modern DevOps practices, including CI/CD pipelines, Infrastructure as Code (IaC), and containerization (Docker/Kubernetes).
  • Domain expertise in financial services.


Read more
Why apply to jobs via Cutshort
people_solving_puzzle
Personalized job matches
Stop wasting time. Get matched with jobs that meet your skills, aspirations and preferences.
people_verifying_people
Verified hiring teams
See actual hiring teams, find common social connections or connect with them directly.
ai_chip
Move faster with AI
We use AI to get you faster responses, recommendations and unmatched user experience.
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos