Cutshort logo
For Employers

AWS Data Engineer at Mactores Cognition Private Limited · Remote only · 2 - 15 years · ₹6L - ₹40L / yr · Profitable · Remote only · Posted 27 Feb 2023

Mactores Cognition Private Limited's logo

AWS Data Engineer

Disha Thakar's profile picture
Posted by Disha Thakar
2 - 15 yrs
₹6L - ₹40L / yr
Remote only
Skills
skill iconAmazon Web Services (AWS)
PySpark
athena
Data engineering

As AWS Data Engineer, you are a full-stack data engineer that loves solving business problems. You work with business leads, analysts, and data scientists to understand the business domain and engage with fellow engineers to build data products that empower better decision-making. You are passionate about the data quality of our business metrics and the flexibility of your solution that scales to respond to broader business questions. 


If you love to solve problems using your skills, then come join the Team Mactores. We have a casual and fun office environment that actively steers clear of rigid "corporate" culture, focuses on productivity and creativity, and allows you to be part of a world-class team while still being yourself.

What you will do?

  • Write efficient code in - PySpark, Amazon Glue
  • Write SQL Queries in - Amazon Athena, Amazon Redshift
  • Explore new technologies and learn new techniques to solve business problems creatively
  • Collaborate with many teams - engineering and business, to build better data products and services 
  • Deliver the projects along with the team collaboratively and manage updates to customers on time


What are we looking for?

  • 1 to 3 years of experience in Apache Spark, PySpark, Amazon Glue
  • 2+ years of experience in writing ETL jobs using pySpark, and SparkSQL
  • 2+ years of experience in SQL queries and stored procedures
  • Have a deep understanding of all the Dataframe API with all the transformation functions supported by Spark 2.7+


You will be preferred if you have

  • Prior experience in working on AWS EMR, Apache Airflow
  • Certifications AWS Certified Big Data – Specialty OR Cloudera Certified Big Data Engineer OR Hortonworks Certified Big Data Engineer
  • Understanding of DataOps Engineering


Life at Mactores


We care about creating a culture that makes a real difference in the lives of every Mactorian. Our 10 Core Leadership Principles that honor Decision-making, Leadership, Collaboration, and Curiosity drive how we work.


1. Be one step ahead

2. Deliver the best

3. Be bold

4. Pay attention to the detail

5. Enjoy the challenge

6. Be curious and take action

7. Take leadership

8. Own it

9. Deliver value

10. Be collaborative


We would like you to read more details about the work culture on https://mactores.com/careers 


The Path to Joining the Mactores Team

At Mactores, our recruitment process is structured around three distinct stages:


Pre-Employment Assessment: 

You will be invited to participate in a series of pre-employment evaluations to assess your technical proficiency and suitability for the role.


Managerial Interview: The hiring manager will engage with you in multiple discussions, lasting anywhere from 30 minutes to an hour, to assess your technical skills, hands-on experience, leadership potential, and communication abilities.


HR Discussion: During this 30-minute session, you'll have the opportunity to discuss the offer and next steps with a member of the HR team.


At Mactores, we are committed to providing equal opportunities in all of our employment practices, and we do not discriminate based on race, religion, gender, national origin, age, disability, marital status, military status, genetic information, or any other category protected by federal, state, and local laws. This policy extends to all aspects of the employment relationship, including recruitment, compensation, promotions, transfers, disciplinary action, layoff, training, and social and recreational programs. All employment decisions will be made in compliance with these principles.

Read more
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos

About Mactores Cognition Private Limited

Founded :
2008
Type :
Services
Size :
20-100
Stage :
Profitable

About

Mactores is a global technology consulting and product company with focus on delivering solutions on Cloud, Big Data, Deep Analytics, DevOps, IoT & AI

Read more

Company social profiles

bloginstagramlinkedintwitterfacebook

Similar jobs (10)

company logo
Hema Dekonda
Posted by Hema Dekonda
Bengaluru (Bangalore), Mumbai, Hyderabad, Gurugram
6 - 11 yrs
₹15L - ₹50L / yr
Data engineering
skill iconAmazon Web Services (AWS)
ETL
SQL
NOSQL Databases

About AuxoAI:


AuxoAI is a global platform-based services firm. We help companies—turn their strategies into practical digital and AI solutions. By understanding how our clients make decisions, we use digital and Artificial Intelligence (AI) technologies to drive growth, enhance their operations, improve customer experiences, and provide clear, actionable insights from their data. What We Do We work across various industries such as healthcare, high-tech, consumer packaged goods (CPG), finance etc., and in sales, marketing, and customer support functions.

We help our clients with accelerating their digital and AI journeys through:

• AI Application Development

• Data, Digital and Cloud acceleration using AI

• AI Native Product Engineering


We are seeking a skilled and experienced Data Engineer to join our dynamic team. The ideal candidate will have 6+ years of prior experience in data engineering, with a strong background in AWS (Amazon Web Services) technologies. This role offers an exciting opportunity to work on diverse projects, collaborating with cross-functional teams to design, build, and optimize data pipelines and infrastructure.


Responsibilities:

* Design, develop, and maintain scalable data pipelines and ETL processes leveraging AWS services such as S3, Glue, EMR, Lambda, and Redshift.

* Collaborate with data scientists and analysts to understand data requirements and implement solutions that support analytics and machine learning initiatives.

* Optimize data storage and retrieval mechanisms to ensure performance, reliability, and cost-effectiveness.

* Implement data governance and security best practices to ensure compliance and data integrity.

* Troubleshoot and debug data pipeline issues, providing timely resolution and proactive monitoring.

* Stay abreast of emerging technologies and industry trends, recommending innovative solutions to enhance data engineering capabilities.


Requirements :

* Bachelor's or Master's degree in Computer Science, Engineering, or a related field.

* 6+ years of prior experience in data engineering, with a focus on designing and building data pipelines.

* Proficiency in AWS services, particularly S3, Glue, EMR, Lambda, and Redshift.

* Strong programming skills in languages such as Python, Java, or Scala.

* Experience with SQL and NoSQL databases, data warehousing concepts, and big data technologies.

* Familiarity with containerization technologies (e.g., Docker, Kubernetes) and orchestration tools (e.g., Apache Airflow) is a plus.

Read more
company logo
Shelly Singh
Posted by Shelly Singh
Bengaluru (Bangalore)
8 - 18 yrs
₹5L - ₹18L / yr
ELT
SQL
PySpark
skill iconAmazon Web Services (AWS)
NOSQL Databases

Design, develop, and maintain ETL pipelines involving large-scale data.

Develop data processing and analytics applications primarily using PySpark and Python.

Build scalable and distributed data processing solutions using Apache Spark.

Develop and deploy data applications on AWS cloud.

Work with AWS services related to storage, compute, ETL, data warehousing, analytics, and streaming.

Implement distributed storage and processing solutions capable of handling high-volume datasets.

Design data processing applications with a focus on performance, scalability, reliability, and optimization.

Work with both SQL and NoSQL databases for data storage, processing, and analytics.

Write, optimize, and analyze SQL, HQL, and NoSQL queries.

Troubleshoot data pipeline and processing issues and ensure data quality and reliability.

Collaborate with data engineers, analysts, architects, and other technical teams to deliver data-driven solutions.

Read more
It is an Product Based Company(Domain- EV Charging)
It is an Product Based Company(Domain- EV Charging)
Agency job
via by Mantasha Naaz
Bengaluru (Bangalore)
3 - 5 yrs
₹13L - ₹15L / yr
skill iconAmazon Web Services (AWS)
skill iconPython
PySpark
SQL
ETL
+2 more

Data Engineer

Location: Bengaluru, India (Hybrid)

Employment Type: Full-time

Experience: 3-5 years



Role Overview  

What We’re Looking For:

  • Bachelor’s degree in Computer Science/Engineering or equivalent experience required.
  • Experience designing and shipping cloud services products.
  • Experience driving and managing technical and architectural dependencies on AWS Cloud.
  • A firm understanding of system architecture, cloud computing, PaaS/SaaS design principles, S3, DynamoDB, RDS mandatory.
  • Experience in building or maintaining ETL processes and tools, i.e., AWS Glue or any open-source tool.
  • Proven system-level design contribution to a current “Live” (in production / under daily high load) multi-region SaaS or PaaS offering.
  • Proven experience with S3, DynamoDB, SQL, and AWS RDS services.
  • Proficiency in programming languages such as Python.
  • Strong analytical and problem-solving skills.

Required Skills & Experience

  • Experience with Python, SQL, and data visualization/exploration tools.
  • Familiarity with the AWS ecosystem, specifically S3, DynamoDB, and RDS.
  • Communication skills, especially for explaining technical concepts to nontechnical business leaders.
  • Ability to work on a dynamic, research-oriented team that has concurrent projects.
  • Experience in AWS cost optimization (Savings Plans, Reserved Instances, Spot Instances) and governance frameworks.
  • Experience developing solutions using infrastructure orchestration tools (SSM, automation account, Ansible, etc.).
  • Excellent leadership, stakeholder management, and communication skills.

 

What We Offer

  • Work with some of the brightest minds in the emerging EV industry.
  • Make a tangible impact in reducing carbon emissions and enabling sustainable energy.
  • Freedom to suggest, implement, and innovate on systems, processes, and technologies.
  • Daily ownership in a high-growth, challenging environment.
  • Flexible work environment with hybrid schedules and virtualization options.
  • Competitive pay and benefits including health coverage, innovative PTO program, and performance bonuses.


Read more
company logo
Bengaluru (Bangalore)
14 - 25 yrs
₹50L - ₹70L / yr
Data engineering
databricks
Apache Spark
PySpark
skill iconPython
+19 more

Job Title : Senior Data Engineer – Databricks

Experience : 14 to 20 Years

Location : HSR Layout, Bangalore

Work Mode : Hybrid – 3 Days WFO

Shift : 11:30 AM – 07:30 PM IST

Positions : 2

Notice Period : Immediate Joiners Only

Interview : 1 Technical Round + 2 Client Rounds


Role Overview :

We are looking for a Senior Data Engineer to build and lead enterprise-scale data platforms for a Switzerland-based commodity client.

The role requires a strong hands-on Data Engineering professional with expertise in Databricks, PySpark, Python, SQL, and AWS, along with technical leadership and stakeholder management experience.


Must-Have Skills :

  • 14 to 20 years of Data Engineering experience
  • Databricks & Apache Spark / PySpark
  • Python & SQL
  • AWS Cloud
  • Lakehouse Architecture
  • ETL / ELT & Distributed Data Processing
  • Batch & Streaming Pipelines
  • Data Pipeline Optimization & Data Modeling
  • CDC & Incremental Processing
  • Git, CI/CD & Testing
  • Data Quality, Monitoring & Observability
  • Technical Leadership & Stakeholder Management


Key Responsibilities :

  • Design and build scalable data pipelines using Databricks, PySpark, Python, SQL, and AWS.
  • Own data products from design through production.
  • Develop batch / streaming pipelines and reusable ETL / ELT frameworks.
  • Optimize pipelines for performance, scalability, reliability, and cost.
  • Design scalable data architectures and data models.
  • Implement data quality, monitoring, lineage, and CI/CD practices.
  • Lead technical discussions and mentor engineering teams.
  • Collaborate with business stakeholders, architects, product owners, and engineering teams.
  • Remain hands-on while providing technical leadership.


Ideal Candidate :

A 14 to 20 years experienced, hands-on Data Engineering leader with strong Databricks + PySpark + AWS expertise, excellent communication, stakeholder management, and experience delivering enterprise-scale data platforms.

🔴 Super Urgent : Only Bangalore-based immediate joiners.

Read more
company logo
Anisha Jindal
Posted by Anisha Jindal
Bengaluru (Bangalore), Mumbai
5 - 14 yrs
Best in industry
Data engineering
skill iconPython
PySpark
DAX
PowerBI

Job Summary

We are looking for a skilled and experienced Data Engineer to join our growing data team. The ideal candidate will have strong expertise in Python, PySpark, Data Modeling, and Power BI, with hands-on experience in designing, developing, and optimizing scalable data solutions. The role requires working closely with business stakeholders, data architects, and analytics teams to build robust data pipelines and semantic models that enable data-driven decision-making.


Technical Skills

  • Strong hands-on experience in Python and PySpark development.
  • Expertise in building and optimizing Data Engineering solutions and ETL pipelines.
  • Strong understanding of Data Modeling concepts (Star Schema, Snowflake Schema, Dimensional Modeling).
  • Experience with Power BI Data Modeling and Semantic Layer development.
  • Proficiency in DAX (Data Analysis Expressions).
  • Experience designing and managing Semantic Models in Power BI.
  • Strong SQL skills and experience working with large datasets.
  • Knowledge of data warehousing concepts and best practices.


Preferred Skills

  • Experience with cloud platforms such as Azure, AWS, or GCP.
  • Exposure to modern data platforms like Databricks.
  • Understanding of data governance and data quality frameworks.
Read more
MNC
MNC
Agency job
via by aafia parveen
Hyderabad
5 - 8 yrs
₹2L - ₹20L / yr
Data engineering
Google Cloud Platform (GCP)
Oracle
PySpark
ETL

Job Title: Data Engineer – PySpark | Oracle | GCP


Experience: 5–7 Years

Location: Hyderabad

Notice Period: Immediate Joiners Preferred


Job Summary

We are seeking an experienced Data Engineer with strong expertise in PySpark, Oracle, and Google Cloud Platform (GCP) to design, develop, and optimize scalable data pipelines. The ideal candidate should have hands-on experience in ETL development, data integration, and cloud-based data engineering solutions.

Key Responsibilities


  • Design, develop, and maintain scalable ETL/data pipelines using PySpark.
  • Extract, transform, and load data from Oracle databases into GCP environments.
  • Build and optimize batch data processing workflows for high performance and reliability.
  • Develop data engineering solutions using GCP services.
  • Ensure data quality through validation, monitoring, and troubleshooting.
  • Optimize SQL queries and ETL jobs for performance and scalability.


Required Skills

  • 5–7 years of experience as a Data Engineer.
  • Strong hands-on experience with PySpark.
  • Solid experience with Oracle Database and advanced SQL.
  • Hands-on experience with Google Cloud Platform (GCP).
  • Strong understanding of ETL processes and data warehousing concepts.


Work Location: Hyderabad

Notice Period: Immediate Joiners Preferred

Read more
company logo
Shikha Nagar
Posted by Shikha Nagar
icon

The recruiter has not been active on this job recently. You may apply but please expect a delayed response.

Pune
7 - 20 yrs
Best in industry
skill iconAmazon Web Services (AWS)
Apache HBase
databricks
skill iconJava
skill iconPython

Job Description – Lead Data Engineer (AWS + Big Data)

Location: Pune - Hybrid

Experience: 8+ Years

Role Overview

We are looking for a Lead Data Engineer with strong hands-on expertise in AWS-based Big Data platforms. The ideal candidate should have extensive implementation experience in designing and building scalable data pipelines, mentoring engineering teams, and driving technical delivery. Databricks exposure is mandatory, while deep implementation experience in Databricks is not essential. Candidates with HBase experience will be preferred.

Key Responsibilities

  • Design, develop, and maintain scalable data pipelines and ETL/ELT solutions on AWS.
  • Build and optimize Big Data applications using Spark/PySpark, HBase, Hive, Kafka, and related technologies.
  • Work with AWS services such as S3, Glue, Lambda, IAM, and CloudWatch for cloud-native data engineering.
  • Lead technical implementation, mentor engineers, conduct code reviews, and drive engineering best practices.
  • Ensure data quality, performance optimization, CI/CD adoption, and production support.

Required Skills

  • Strong hands-on experience with AWS (S3, Glue, Lambda, IAM, CloudWatch)
  • Apache Spark (PySpark/Scala) and Big Data ecosystem
  • Databricks exposure (mandatory)
  • HBase (strongly preferred)
  • Python or Java, Advanced SQL
  • ETL/ELT development and Data Warehousing concepts
  • Apache Airflow or similar orchestration tools
  • Git, CI/CD, Performance Tuning, and Data Quality

Good to Have

  • Kafka / Spark Structured Streaming
  • Hive, Impala, Hadoop ecosystem
  • Delta Lake / Lakehouse concepts
  • Snowflake or other modern cloud data platforms

Experience Required

  • 8–14 years of Data Engineering experience with strong AWS implementation expertise.
  • Proven experience leading technical delivery and mentoring engineering teams.
  • Strong understanding of enterprise-scale data platforms and Big Data architectures.
  • Ability to collaborate with architects, stakeholders, and cross-functional teams to deliver scalable solutions.


NOTE: Final Technical round is mandatory to be taken F2F from Pune, office.


 

Read more
company logo
Tushar Vaghela
Posted by Tushar Vaghela
Bengaluru (Bangalore)
5 - 10 yrs
Best in industry
skill iconPython
skill iconScala
Apache Spark
Apache Kafka
databricks
+1 more

Description


We are looking for Senior Data Engineers to join our Data Platform team and build scalable, high-performance data platforms that power data processing, analytics, and downstream applications.

The ideal candidate will have strong experience in distributed data processing, ETL pipelines, and Big Data technologies, with hands-on expertise in Apache Spark and Python Scala.

You will be responsible for designing, developing, and optimizing large-scale data pipelines while collaborating closely with cross-functional engineering teams to build reliable, production-grade data solutions.



Key Responsibilities

  • Design, develop, and maintain scalable ETL and data processing pipelines for large-scale datasets.
  • Build and optimize distributed data applications using Apache Spark and Python Scala.
  • Develop reliable, high-performance data pipelines for batch and streaming workloads.
  • Design and manage data workflows using Apache Airflow.
  • Build and operate data workloads on AWS, with strong usage of Amazon S3 for large-scale data storage.
  • Work with large datasets to ensure data quality, consistency, reliability, and performance.
  • Collaborate with engineering, product, analytics, and other platform teams to deliver robust data solutions.
  • Optimize data workflows for scalability, reliability, performance, and cost efficiency.
  • Troubleshoot production issues, identify bottlenecks, and continuously improve platform performance.



Requirements

Candidates who demonstrate:

  • 5+ years of experience in Data Engineering, Big Data Engineering, or a similar role.
  • Strong hands-on experience with Apache Spark and Scala.
  • Experience designing, building, and maintaining large-scale ETL pipelines.
  • Strong hands-on experience with AWS, particularly Amazon S3.
  • Hands-on experience with Apache Airflow for workflow orchestration and scheduling.
  • Strong SQL skills and a solid understanding of distributed data processing concepts.
  • Experience working with batch and/or streaming data pipelines.
  • Excellent debugging, problem-solving, and performance optimization skills.
  • Strong communication and collaboration skills.


Good to Have

  • Experience with Databricks and the broader Databricks data platform.
  • Familiarity with streaming technologies such as Apache Kafka.
  • Experience working on large-scale data platforms handling high-volume data workloads.
  • Exposure to additional AWS data services and cloud-native data architectures.
Read more
company logo
Hema V
Posted by Hema V
Remote, Bengaluru (Bangalore), Noida, Chennai
3 - 10 yrs
Best in industry
Generative AI
LangGraph
ETL
databricks
Retrieval Augmented Generation (RAG)
+1 more

Role Summary

We are hiring a Data Engineer / ML Data Pipeline Engineer to build and operate the data backbone of the Enterprise AI platform:

 

What You'll Own

  • Ingestion & ETL/ELT pipelines for heterogeneous project folders (PDF drawings, SVG files, IFC models, BBS.json bar-bending-schedule data, Excel exports, and AI agent output JSON).
  • AWS-based data architecture: S3 raw/staging/curated/outputs structuring, partitioning, versioning, and lifecycle management; querying via Athena/Glue and warehousing via Redshift or Snowflake as needed.
  • Data validation frameworks: GUID cross-referencing between SVG and BBS data, schema enforcement, duplicate/orphan detection, reference integrity checks, and structured validation reporting.
  • Agent run logging & observability: designing the database schema and pipelines that track every AI agent run (inputs, outputs, status, errors, cost, retries, reviewer feedback).
  • AI Factory monitoring dashboards: operational dashboards (failure rates, retries, latency, data quality) and business dashboards (throughput, cost per run, rework rate) for Power BI/QuickSight or equivalent.
  • ML data pipeline support: dataset preparation, labeling/annotation workflows, human-in-the-loop review tooling, and dataset versioning for models that classify or QC drawing issues.
  • APIs: designing and building FastAPI/Flask endpoints to trigger validation runs and expose agent processing status to internal tools.
  • Data quality & testing discipline: idempotent pipelines, quarantine/reject handling, regression and reconciliation testing, and root-cause debugging when pipelines or query performance degrade in production.

Key Skills — Non-Negotiable (Must-Have, Strong Level)

  • Python — production-grade scripting: file/folder handling, JSON/schema processing, clean error handling, not just notebook-level scripting.
  • SQL — strong hands-on ability, including GROUP BY/HAVING for duplicate detection, window functions, and daily aggregate/rate calculations (e.g., success-rate queries).
  • AWS S3 data handling — practical experience structuring buckets for raw/staging/curated data, versioning, and avoiding overwrite issues at scale.
  • Data validation — demonstrable experience building validation logic (set comparisons, duplicate/missing detection, structured pass/fail reporting), not just "I write assertions."
  • ETL/ELT pipeline design — end-to-end ownership of at least one pipeline: source → transform → storage → validation → monitoring → business outcome, with clear articulation of what they personally built.
  • Query/warehouse engine judgment — working knowledge of when to use Athena vs. Redshift vs. Snowflake (or equivalent), partitioning, clustering, sort/distribution keys, and storage format trade-offs (Parquet vs. JSON vs. CSV).

Key Skills — Good to Have

  • Dashboarding — Power BI / QuickSight (or equivalent) fact/dimension table design, KPI cards, drill-downs; medium-to-strong level is a plus but trainable.
  • FastAPI / Flask — building real endpoints with request/response schemas and basic error handling; especially valuable for validation-trigger and agent-status APIs.
  • ML data pipeline experience — dataset labeling, annotation platform design, train/test/validation splitting, dataset versioning; strong on the pipeline/data side rather than model training itself.
  • Human-in-the-loop / review tooling — experience building or contributing to browser-based labeling/review platforms (session persistence, label schema, export formats).
  • Large-scale metadata querying — experience making file discovery fast across large volumes (1,000+ projects, thousands of files each) via metadata index tables, event-based ingestion, or catalog tools like AWS Glue.
Read more
company logo
Agency job
via by Ajeethkumar s
Hyderabad, Bengaluru (Bangalore)
5 - 10 yrs
₹4L - ₹16L / yr
skill iconPython
ETL
PySpark
Data engineering
skill iconAmazon Web Services (AWS)
+2 more

Skills Referential (Required knowledge, skills and abilities)

Technical Skills:

Python

Pyspark

SQL

ETL Aws, Azure, gcp

Read more
Why apply to jobs via Cutshort
people_solving_puzzle
Personalized job matches
Stop wasting time. Get matched with jobs that meet your skills, aspirations and preferences.
people_verifying_people
Verified hiring teams
See actual hiring teams, find common social connections or connect with them directly.
ai_chip
Move faster with AI
We use AI to get you faster responses, recommendations and unmatched user experience.
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos