Cutshort logo
For Employers
Apache Spark Jobs in Chennai

10+ Apache Spark Jobs in Chennai | Apache Spark Job openings in Chennai

Apply to 10+ Apache Spark Jobs in Chennai on CutShort.io. Explore the latest Apache Spark Job opportunities across top companies like Google, Amazon & Adobe.

icon
QuaXigma IT solutions Private Limited
Tirupati, Chennai
5 - 10 yrs
Best in industry
SQL
skill iconPython
Stored Procedures
skill iconAmazon Web Services (AWS)
Microsoft Windows Azure
+10 more

About Us:

The QX Impact was launched with a mission to make A.I accessible and affordable and deliver AI Products/Solutions at scale for the enterprises by bringing the power of Data, AI, and Engineering to drive digital transformation. We believe without insights; businesses will continue to face challenges to better understand their customers and even lose them. Secondly, without insights businesses won't’ be able to deliver differentiated products/services; and finally, without insights, businesses can’t achieve a new level of “Operational Excellence” is crucial to remain competitive, meeting rising customer expectations, expanding markets, and digitalization.


Job Summary:

We are looking for a Senior Data Engineer who is creative, collaborative, and adaptable to join our agile team of data scientists, engineers, and UX developers. The role focuses on building and maintaining robust data pipelines to support advanced analytics, data science, and BI solutions.

As a Senior Data Engineer, you will work with internal and external data, collaborate with data scientists, and contribute to the design, development, and deployment of innovative solutions.


Key Responsibilities:

  • Design, develop, test, and maintain optimal data pipeline and ETL architectures.
  • Map out data systems and define/design required integrations, ETL, BI, and AI systems/processes.
  • Prepare and optimize data for predictive and prescriptive modeling.
  • Collaborate with teams to integrate ERP data into the enterprise data lake, ensuring seamless flow and quality.
  • Enhance cloud data infrastructure on AWS or Azure for scalability and performance.
  • Utilize big data tools and frameworks to optimize data acquisition and preparation.
  • Build architectures to move data to/from data lakes and data warehouses for advanced analytics.
  • Develop and curate data models for analytics, dashboards, and reports.
  • Conduct code reviews, maintain production-level code, and implement testing approaches.
  • Monitor, troubleshoot, and resolve data ingestion workflows to maintain reliability and uptime.
  • Drive innovation and implement efficient new approaches to data engineering tasks.


Must-Have Skills:

  • Bachelor’s degree in Computer Science, Mathematics, Engineering, or a related field.
  • 5+ years of experience working with enterprise data platforms, including building and managing data lakes.
  • 3–5 years of experience designing and implementing data warehouse solutions.
  • Expertise in SQL, including developing stored procedures (SP) and applying advanced data design concepts.
  • Proficiency in Spark (Python/Scala) and Spark Streaming for real-time data pipelines.
  • Experience with AWS or Azure services (e.g., AWS Glue, Azure Data Factory, Redshift, Snowflake).
  • Familiarity with big data tools such as Apache Kafka, Apache Spark, or Flink.
  • Hands-on experience with orchestration tools (e.g., Apache Airflow, Prefect).
  • Knowledge of CI/CD processes, version control (e.g., Git, Jenkins), and deployment automation.
  • Strong problem-solving, communication, and collaboration skills.


Good-to-Have Skills:

  • Experience in integrating ERP data into data lakes.
  • Experience with traditional ETL tools (e.g., Talend, Pentaho).


Competencies:

  • Tech Savvy - Anticipating and adopting innovations in business-building digital and technology applications.
  • Self-Development - Actively seeking new ways to grow and be challenged using both formal and informal development channels.
  • Action Oriented - Taking on new opportunities and tough challenges with a sense of urgency, high energy, and enthusiasm.
  • Customer Focus - Building strong customer relationships and delivering customer-centric solutions.
  • Optimize Work Processes - Knowing the most effective and efficient processes to get things done, with a focus on continuous improvement.


Why Join Us?

  • Be part of a collaborative and agile team driving cutting-edge AI and data engineering solutions.
  • Work on impactful projects that make a difference across industries.
  • Opportunities for professional growth and continuous learning.
  • Competitive salary and benefits package.


Application Details

Ready to make an impact? Apply today and become part of the QX Impact team!


Read more
Staffnixcom
Mayank Choudhary
Posted by Mayank Choudhary
Chennai
10 - 15 yrs
₹27L - ₹32L / yr
Data engineering
databricks
skill iconPython
SQL
Apache Spark

Strong Databricks Architect Profile with end-to-end Lakehouse ownership

2

Mandatory (Experience 1) – Must have 10+ years of software engineering experience with atleast 5+ years in Data Engineering with hands on exposure to Databricks and strong ownership of end-to-end data pipeline development.

3

Mandatory (Experience 2) – Must have atleast 5+ years of expertise across the Databricks ecosystem — Delta Lake, Delta Live Tables, Autoloader, Structured Streaming, Workflows, Unity Catalog

4

Mandatory (Tech skill 1) – Must have worked at architecture level, owning end-to-end design through deployment

5

Mandatory (Tech skill 2) – Must have strong experience with Python and SQL for data processing and Apache Spark for performance tuning & scalability

6

Mandatory (Tech skill 3) – Must have experience in large-scale data warehousing & advanced data modeling (3NF and dimensional) across batch and real-time systems

7

Mandatory (AI Exposure) – Must have at least a basic working understanding of how AI services or tools work

8

Mandatory (Communication) – Must have strong stakeholder management & requirement-gathering experience with US or UK clients

9

Mandatory (Company) – Must come from a B2B IT services or IT consulting background

10

Mandatory (Note) – CTC is inclusive of 5% variable

11

Preferred (Tech skill 1) – Azure Databricks or Azure data services experience (project runs on Azure DevOps)

12

Preferred (Tech skill 2) – MLflow or MLOps practices and AI use cases (RAG, AI/BI)

13

Preferred (Tech skill 3) – CI/CD, Databricks Asset Bundles (DABs) or equivalent packaging, Terraform or IaC, reusable deployment templates

14

Preferred (Integrations) – ServiceNow or enterprise system integrations

15

Preferred (Certifications) – Databricks (Data Engineer Associate or Professional, ML or GenAI tracks), Azure or AWS cloud certifications

Read more
NeoGenCode Technologies Pvt Ltd
Akshay Patil
Posted by Akshay Patil
Noida, Bengaluru (Bangalore), Pune, Hyderabad, Chennai
6 - 8 yrs
₹6L - ₹12L / yr
Data engineering
databricks
Snow flake schema
skill iconPython
Apache Spark
+8 more

Job Title : Data Engineer – Databricks

Experience : 6+ Years

Location : Noida / Hyderabad / Chennai / Pune / Bengaluru (Hybrid)

Shift : IST (Normal Shift)


Job Summary :

We are seeking an experienced Data Engineer with strong expertise in Databricks, Snowflake, Python, and Spark to build and optimize scalable data pipelines and support AI/ML model deployments. The ideal candidate should have experience working with cloud-based data platforms and preferably possess exposure to the Healthcare domain.


Required Skills :

  • Databricks (Preferred)
  • Snowflake
  • Python
  • Apache Spark
  • SQL
  • Azure Cloud
  • Kubernetes
  • Apache Airflow
  • GitHub & CI/CD Pipelines
  • AI/ML Model Deployment
  • Data Analytics

Preferred :

  • Experience in the Healthcare domain.
  • Strong understanding of scalable data engineering architectures and best practices.
Read more
Remote, Bengaluru (Bangalore), Pune, Chennai, Nagpur
5 - 15 yrs
₹20L - ₹30L / yr
databricks
PySpark
Apache Spark
CI/CD
Data engineering


Technical Architect (Databricks)

  • 10+ Years Data Engineering Experience with expertise in Databricks
  • 3+ years of consulting experience
  • Completed Data Engineering Professional certification & required classes
  • Minimum 2-3 projects delivered with hands-on experience in Databricks
  • Completed Apache Spark Programming with Databricks, Data Engineering with Databricks, Optimizing Apache Spark™ on Databricks
  • Experience in Spark and/or Hadoop, Flink, Presto, other popular big data engines
  • Familiarity with Databricks multi-hop pipeline architecture

 

 

Sr. Data Engineer (Databricks)

 

  • 5+ Years Data Engineering Experience with expertise in Databricks
  • Completed Data Engineering Associate certification & required classes
  • Minimum 1 project delivered with hands-on experience in development on Databricks
  • Completed Apache Spark Programming with Databricks, Data Engineering with Databricks, Optimizing Apache Spark™ on Databricks
  • SQL delivery experience, and familiarity with Bigquery, Synapse or Redshift
  • Proficient in Python, knowledge of additional databricks programming languages (Scala)


Read more
top MNC

top MNC

Agency job
via Vy Systems by thirega thanasekaran
Bengaluru (Bangalore), Chennai, Hyderabad, Coimbatore, Kochi (Cochin), Thrissur, Thiruvananthapuram, Kozhikode (Calicut), Kasaragod
5 - 12 yrs
₹5L - ₹9L / yr
Data engineering
databricks
Apache Synapse
Apache Spark

Job Summary:


Seeking an experienced Senior Data Engineer to lead data ingestion, transformation, and optimization initiatives using the modern Apache and Azure data stack. The role involves working on scalable pipelines, large-scale distributed systems, and data lake management.

Core Responsibilities:

· Build and manage high-volume data pipelines using Spark/Databricks.

· Implement ELT frameworks using Azure Data Factory/Synapse Pipelines.

· Optimize large-scale datasets in Delta/Iceberg formats.

· Implement robust data quality, monitoring, and governance layers.

· Collaborate with Data Scientists, Analysts, and Business stakeholders.

Technical Stack:

· Big Data: Apache Spark, Kafka, Hive, Airflow, Hudi/Iceberg

· Cloud: Azure (Synapse, ADF, ADLS Gen2), Databricks, AWS (Glue/S3)

· Languages: Python, Scala, SQL

· Storage Formats: Delta Lake, Iceberg, Parquet, ORC

· CI/CD: Azure DevOps, Terraform (infra as code), Git

Senior Data Engineer (Apache Stack + Databricks/Synapse)


Share cv to

Thirega@ vysystems dot com - WhatsApp - 91Five0033Five2Three

Read more
Cubera Tech India Pvt Ltd
Bengaluru (Bangalore), Chennai
5 - 8 yrs
Best in industry
Data engineering
Big Data
skill iconJava
skill iconPython
Hibernate (Java)
+10 more

Data Engineer- Senior

Cubera is a data company revolutionizing big data analytics and Adtech through data share value principles wherein the users entrust their data to us. We refine the art of understanding, processing, extracting, and evaluating the data that is entrusted to us. We are a gateway for brands to increase their lead efficiency as the world moves towards web3.

What are you going to do?

Design & Develop high performance and scalable solutions that meet the needs of our customers.

Closely work with the Product Management, Architects and cross functional teams.

Build and deploy large-scale systems in Java/Python.

Identify, design, and implement internal process improvements: automating manual processes, optimizing data delivery, re-designing infrastructure for greater scalability, etc.

Create data tools for analytics and data scientist team members that assist them in building and optimizing their algorithms.

Follow best practices that can be adopted in Bigdata stack.

Use your engineering experience and technical skills to drive the features and mentor the engineers.

What are we looking for ( Competencies) :

Bachelor’s degree in computer science, computer engineering, or related technical discipline.

Overall 5 to 8 years of programming experience in Java, Python including object-oriented design.

Data handling frameworks: Should have a working knowledge of one or more data handling frameworks like- Hive, Spark, Storm, Flink, Beam, Airflow, Nifi etc.

Data Infrastructure: Should have experience in building, deploying and maintaining applications on popular cloud infrastructure like AWS, GCP etc.

Data Store: Must have expertise in one of general-purpose No-SQL data stores like Elasticsearch, MongoDB, Redis, RedShift, etc.

Strong sense of ownership, focus on quality, responsiveness, efficiency, and innovation.

Ability to work with distributed teams in a collaborative and productive manner.

Benefits:

Competitive Salary Packages and benefits.

Collaborative, lively and an upbeat work environment with young professionals.

Job Category: Development

Job Type: Full Time

Job Location: Bangalore

 

Read more
Leading Manufacturing Company

Leading Manufacturing Company

Agency job
Chennai
3 - 6 yrs
₹3L - ₹8L / yr
skill iconMachine Learning (ML)
skill iconData Science
Natural Language Processing (NLP)
Data modeling
skill iconData Analytics
+2 more

Location:  Chennai
Education: BE/BTech
Experience: Minimum 3+ years of experience as a Data Scientist/Data Engineer

Domain knowledge: Data cleaning, modelling, analytics, statistics, machine learning, AI

Requirements:

  • To be part of Digital Manufacturing and Industrie 4.0 projects across client group of companies
  • Design and develop AI//ML models to be deployed across factories
  • Knowledge on Hadoop, Apache Spark, MapReduce, Scala, Python programming, SQL and NoSQL databases is required
  • Should be strong in statistics, data analysis, data modelling, machine learning techniques and Neural Networks
  • Prior experience in developing AI and ML models is required
  • Experience with data from the Manufacturing Industry would be a plus

Roles and Responsibilities:

  • Develop AI and ML models for the Manufacturing Industry with a focus on Energy, Asset Performance Optimization and Logistics
  • Multitasking, good communication necessary
  • Entrepreneurial attitude

Additional Information:

  • Travel:                                  Must be willing to travel on shorter duration within India and abroad
  • Job Location:                      Chennai
  • Reporting to:                      Team Leader, Energy Management System
Read more
American Multinational Retail Corp

American Multinational Retail Corp

Agency job
via Hunt & Badge Consulting Pvt Ltd by Chandramohan Subramanian
Chennai
2 - 5 yrs
₹5L - ₹15L / yr
skill iconScala
Spark
Apache Spark

Should have Passion to learn and adapt new technologies, understanding,

solving/troubleshooting issues and risks, able to make informed decisions and ability to

lead the projects.

 

Your Qualifications

 

  • 2-5 Years’ Experience with functional programming
  • Experience with functional programming using Scala with Spark framework.
  • Strong understanding of Object-oriented programming, data structures and algorithms
  • Good experience in any of the cloud platforms (Azure, AWS, GCP) etc.,
  • Experience with distributed (multi-tiered) systems, relational databases and NoSql storage solutions
  • Desire to learn new technologies and languages
  • Participation in software design, development, and code reviews
  • High level of proficiency with Computer Science/Software Engineering knowledge and contribution to the technical skills growth of other team members


Your Responsibility

 

  • Design, build and configure applications to meet business process and application requirements
  • Proactively identify and communicate potential issues and concerns and recommend/implement alternative solutions as appropriate.
  • Troubleshooting & Optimization of existing solution

 

Provide advice on technical design to ensure solutions are forward looking and flexible for potential future requirements and business needs.
Read more
Data & Cloud Technology serviced based company.

Data & Cloud Technology serviced based company.

Agency job
via Multi Recruit by Ragul Ragul
Chennai, Coimbatore, Madurai
5 - 10 yrs
₹12L - ₹19L / yr
Apache Spark
HiveQL
skill iconAmazon Web Services (AWS)
Data engineering
JSON
+2 more
  • Must have the experience of leading teams and drive customer interactions
  • Must have multiple successful deployments user stories
  • Extensive hands on experience in Apache Spark along with HiveQL
  • Sound knowledge in Amazon Web Services or any other Cloud environment.
  • Experienced in data flow orchestration using Apache Airflow
  • JSON, XML, CSV, Parquet file formats with snappy compression.
  • File movements between HDFS and AWS S3
  • Experience in shell scripting and scripting to automate report generation and migration of reports to AWS S3
  • Worked in building a data pipeline using Pandas and Flask FrameworkGood Familiarity with Anaconda and Jupyternotebook
Read more
Lymbyc

at Lymbyc

1 video
2 recruiters
Venky Thiriveedhi
Posted by Venky Thiriveedhi
Bengaluru (Bangalore), Chennai
4 - 8 yrs
₹9L - ₹14L / yr
Apache Spark
Apache Kafka
Druid Database
Big Data
Apache Sqoop
+5 more
Key skill set : Apache NiFi, Kafka Connect (Confluent), Sqoop, Kylo, Spark, Druid, Presto, RESTful services, Lambda / Kappa architectures Responsibilities : - Build a scalable, reliable, operable and performant big data platform for both streaming and batch analytics - Design and implement data aggregation, cleansing and transformation layers Skills : - Around 4+ years of hands-on experience designing and operating large data platforms - Experience in Big data Ingestion, Transformation and stream/batch processing technologies using Apache NiFi, Apache Kafka, Kafka Connect (Confluent), Sqoop, Spark, Storm, Hive etc; - Experience in designing and building streaming data platforms in Lambda, Kappa architectures - Should have working experience in one of NoSQL, OLAP data stores like Druid, Cassandra, Elasticsearch, Pinot etc; - Experience in one of data warehousing tools like RedShift, BigQuery, Azure SQL Data Warehouse - Exposure to other Data Ingestion, Data Lake and querying frameworks like Marmaray, Kylo, Drill, Presto - Experience in designing and consuming microservices - Exposure to security and governance tools like Apache Ranger, Apache Atlas - Any contributions to open source projects a plus - Experience in performance benchmarks will be a plus
Read more
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Why apply via Cutshort?
Connect with actual hiring teams and get their fast response. No spam.
Find more jobs
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort