ETL+ SQL Developer at ScatterPie Analytics · Bengaluru (Bangalore) · 4 - 6 years · ₹10L - ₹16L / yr · Profitable · Posted 17 Dec 2024

Skills: ETL+ SQL
· Experience with SQL and data querying languages.
· Knowledge of data governance frameworks and best practices.
· Familiarity with programming/scripting languages (e.g., SparkSQL)
· Strong understanding of data integration techniques and ETL processes.
· Experience with data quality tools and methodologies.
· Strong communication and problem-solving skills
Detailed JD: Data Integration: Manage the seamless integration various data lake, ensuring that jobs are running as expected, validate the data ingested , track the DQ checks , rerun/reprocess the jobs in case of failures post figuring out the RCAs
Data Quality Assurance: Monitor and validate data quality during and after the migration process, implementing checks and corrective actions as needed.
Documentation: Maintain comprehensive documentation related to data issues encountered during the weekly/monthly processing and operational procedures.
Continuous Improvement: Recommend and implement improvements to data processing, tools, and technologies to enhance efficiency and effectiveness.

About ScatterPie Analytics
About
Similar jobs (10)
Data Engineer Short Hiring Post
🚨 Hiring: Data Engineer
🔹 Experience: 5–9 Years
🔹 Location: Bangalore / Hyderabad
🔹 Skills: PySpark, Python, SQL, ETL, CI/CD, Data Modeling
🔹 Process: L1 Virtual → L2 F2F Karat Test
🔹 F2F: Bangalore / Hyderabad Location
🔹 Positions: Immediate requirement
⚠️ Note: Candidates must be available for F2F Karat immediately after L1.
#Hiring #DataEngineer #PySpark #Python #SQL #BangaloreJobs #HyderabadJobs #Mphasis #ImmediateJoiners
Data Engineer Hiring Post
🚨 Hiring: Data Engineer | PySpark + Python + SQL
We are looking for experienced Data Engineers to join our team!
🔹 Experience: 5 to 9 Years
🔹 Locations: Bangalore / Hyderabad
🔹 Interview Process:
• 1st Round – Virtual
• 2nd Round – Face-to-Face (Karat Test)
🔑 Key Skills:
✅ PySpark
✅ SQL
✅ Python
✅ ETL
📩 Interested candidates can share their updated resume.
#Hiring #DataEngineer #PySpark #Python #SQL #ETL #BangaloreJobs #HyderabadJobs #TechHiring #ImmediateHiring
Job Summary
We are seeking a skilled Data Engineer to design, build, and maintain scalable data pipelines and infrastructure. The ideal candidate should have strong expertise in SQL, Python, Linux, and modern data engineering practices to support data integration, transformation, and analytics.
Key Responsibilities
- Design, develop, and maintain ETL/ELT data pipelines.
- Write efficient and optimized SQL queries for data extraction, transformation, and reporting.
- Develop automation scripts using Python for data processing and workflow optimization.
- Work with Linux environments for deployment, monitoring, and troubleshooting.
- Ensure data quality, integrity, and reliability across data platforms.
- Collaborate with data analysts, software engineers, and business stakeholders to deliver data solutions.
- Monitor, troubleshoot, and optimize data pipelines for performance and scalability.
- Implement best practices for data security, governance, and documentation.
Required Skills
- Strong experience in Data Engineering concepts and ETL/ELT processes.
- Proficiency in SQL, including query optimization and database design.
- Strong programming skills in Python.
- Hands-on experience with Linux commands, shell scripting, and system administration basics.
- Experience with relational databases such as PostgreSQL, MySQL, SQL Server, or Oracle.
- Familiarity with Git/version control.
- Strong analytical and problem-solving skills.
Preferred Skills
- Experience with cloud platforms (AWS, Azure, or GCP).
- Knowledge of Apache Spark, Airflow, Kafka, or similar data engineering tools.
- Experience with data warehousing solutions and big data technologies.
- Understanding of CI/CD pipelines and containerization (Docker/Kubernetes).
Qualifications
- Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field.
- Relevant certifications in cloud or data engineering are an added advantage.
The recruiter has not been active on this job recently. You may apply but please expect a delayed response.
Must have experience in Java
Must have experience in Spark
Must have experience in ETL coding
Strong expertise in coding
Location: Hyderabad / Chennai
Experience: 5+ years
Employment type: Full-time, permanent
Work Hours: General Shift
website: www.amazech.com
Qualifications:
- B.E./B.Tech/M.E./M.Tech in Computer Science, Information Technology, Data Science, or related disciplines.
- Strong academic background with relevant industry experience in Data Engineering and Data Warehousing.
Key Responsibilities:
· Design, develop, and maintain scalable data warehouse solutions using Snowflake.
· Write, optimize, troubleshoot, and enhance Snowflake SQL queries with a focus on performance and scalability.
· Develop and support ETL processes using Talend to ensure reliable and efficient data movement.
· Collaborate with business, analytics, and application teams to enable reporting, dashboards, metrics, and data exploration capabilities.
· Perform data analysis and resolve issues across data ingestion, transformation, and reporting pipelines.
· Debug and troubleshoot Python-based data processing scripts and automation workflows.
· Implement best practices for data quality, testing, deployment, and code reviews.
· Work across UI, API, and Data Warehouse layers to support end-to-end data integration and business requirements.
· Monitor, optimize, and maintain data warehouse performance and operational stability.
· Create and maintain technical documentation, data models, and process workflows.
Required Skills and Experience:
· Strong hands-on expertise in Snowflake Data Warehouse.
· Advanced SQL skills with experience handling large-scale datasets.
· Strong understanding of Data Warehousing concepts, dimensional modelling, and data architecture.
· Hands-on experience with Analytical SQL functions, query tuning, and performance optimization.
· Experience developing and maintaining ETL solutions using Talend.
· Proficiency in Python for scripting, debugging, automation, and data processing.
· Experience integrating UI, API, and Data Warehouse workflows.
· Strong problem-solving and analytical skills.
· Experience with testing, code reviews, and deployment best practices.
· Excellent communication and stakeholder management skills.
Position: Backup Data ETL Developer
Experience: 6–9 Years
Location: Mumbai
Employment Type: Full-time
Job Description:
We are looking for a Backup Data ETL Developer with 6–9 years of experience in designing, developing, and supporting data extraction, transformation, and loading (ETL) processes.
The candidate should have strong hands-on experience with Hadoop, Spark, Informatica, and Snowflake, along with expertise in large-scale data processing and data integration technologies.
Key Skills:
Hadoop
Spark
Informatica
Snowflake
ETL Development
Data Integration
Data Transformation
Large-Scale Data Processing
ETL Workflows and Data Pipelines
Troubleshooting and Problem-Solving
Strong analytical and collaboration skills
The recruiter has not been active on this job recently. You may apply but please expect a delayed response.
Strong Senior Developer – PL/SQL, SQL & ETL (Microsoft SSIS) Profile
2
Mandatory (Experience 1) – Must have minimum 5+ years of strong hands-on experience in PL/SQL and SQL development, including complex stored procedures, functions, queries, joins, data manipulation, and query/performance optimization.
3
Mandatory (Experience 2) – Must have strong hands-on experience in ETL development using Microsoft SSIS, including building, maintaining, optimizing, and troubleshooting SSIS packages for large-volume data movement and transformation.
4
Mandatory (Experience 3) – Must have solid experience working with Data Warehousing concepts and architectures, including data models, fact/dimension structures, ETL data flows, and enterprise reporting/data warehouse environments.
5
Mandatory (Experience 4) – Must have experience managing batch jobs, scheduling, and data pipelines, ensuring timely and reliable execution of enterprise ETL workflows.
6
Mandatory (Experience 5) – Must have hands-on experience in production support for SSIS/ETL and data warehouse jobs, including monitoring job execution, troubleshooting failures, performing root cause analysis, and implementing preventive fixes.
7
Mandatory (Experience 6) – Must have experience with data quality, validation, and troubleshooting, including identifying and resolving data discrepancies/issues affecting downstream reports, dashboards, and analytics.
8
Mandatory (Experience 7) – Must have experience with unit, integration, and regression testing of SQL, PL/SQL, and ETL components, along with strong documentation of technical designs, data mappings, data flows, and deployment processes.
9
Mandatory (Location) – Must be willing to work in a hybrid model from a city where Cognizant has an office.
10
Mandatory (Notice Period) – Immediate joiners or candidates who can join within 2–4 weeks.
Data Engineer - Remote
Nearshore engineer on a team converting SAS code to Python and SQL on Databricks using generative AI. You will work with the existing accelerators, own deliverables end to end, and communicate directly with client and partner stakeholders.
Required for both:
4+ years of professional data or software engineering experience
Strong Python and SQL
Hands-on Databricks (Unity Catalog, Workflows, Databricks Asset Bundles)
GitLab CI/CD: pipelines, merge request workflows, automated testing
Git branching and code review discipline
Clear written and spoken English with client-facing partners
Demonstrated ownership: scoping, delivering, and flagging risk without prompting
Focus: Pipeline reliability, validation, and delivery of converted code.
Responsibilities:
Build and run the pipelines that process SAS inventories and converted outputs
Validate converted code for parity against SAS outputs (row counts, checksums, schema, data types)
Own deployment through DABs and GitLab CI/CD
Manage Unity Catalog objects, permissions, and environment promotion
Troubleshoot job failures and performance issues
Required:
Spark and Delta Lake performance tuning
Data validation and reconciliation experience
Infrastructure as code or DAB-based deployment experience
Nice to have:
SAS reading ability, healthcare data exposure, Azure.
Job Description
We are looking for an experienced Data Engineer with strong expertise in Python, ETL, SQL, CI/CD, and DevOps to design, develop, and maintain scalable data pipelines and data processing solutions.
Key Responsibilities
- Design, develop, and maintain scalable ETL/ELT data pipelines.
- Develop data processing solutions using Python.
- Write complex and optimized SQL queries, stored procedures, and data transformations.
- Build and maintain data ingestion and integration workflows.
- Implement data quality, validation, monitoring, and error-handling processes.
- Develop and maintain CI/CD pipelines for data engineering applications.
- Work with DevOps tools and practices for automated build, deployment, and infrastructure management.
- Collaborate with data analysts, data scientists, software engineers, and business teams.
- Optimize data pipelines for performance, reliability, and scalability.
- Troubleshoot production data issues and ensure timely resolution.
- Follow best practices for version control, code quality, testing, and deployment.
Mandatory Skills
- Python
- ETL
- SQL
- CI/CD
- DevOps
- Git / Version Control
- Strong problem-solving and debugging skills
Description
We are looking for Senior Data Engineers to join our Data Platform team and build scalable, high-performance data platforms that power data processing, analytics, and downstream applications.
The ideal candidate will have strong experience in distributed data processing, ETL pipelines, and Big Data technologies, with hands-on expertise in Apache Spark and Python Scala.
You will be responsible for designing, developing, and optimizing large-scale data pipelines while collaborating closely with cross-functional engineering teams to build reliable, production-grade data solutions.
Key Responsibilities
- Design, develop, and maintain scalable ETL and data processing pipelines for large-scale datasets.
- Build and optimize distributed data applications using Apache Spark and Python Scala.
- Develop reliable, high-performance data pipelines for batch and streaming workloads.
- Design and manage data workflows using Apache Airflow.
- Build and operate data workloads on AWS, with strong usage of Amazon S3 for large-scale data storage.
- Work with large datasets to ensure data quality, consistency, reliability, and performance.
- Collaborate with engineering, product, analytics, and other platform teams to deliver robust data solutions.
- Optimize data workflows for scalability, reliability, performance, and cost efficiency.
- Troubleshoot production issues, identify bottlenecks, and continuously improve platform performance.
Requirements
Candidates who demonstrate:
- 5+ years of experience in Data Engineering, Big Data Engineering, or a similar role.
- Strong hands-on experience with Apache Spark and Scala.
- Experience designing, building, and maintaining large-scale ETL pipelines.
- Strong hands-on experience with AWS, particularly Amazon S3.
- Hands-on experience with Apache Airflow for workflow orchestration and scheduling.
- Strong SQL skills and a solid understanding of distributed data processing concepts.
- Experience working with batch and/or streaming data pipelines.
- Excellent debugging, problem-solving, and performance optimization skills.
- Strong communication and collaboration skills.
Good to Have
- Experience with Databricks and the broader Databricks data platform.
- Familiarity with streaming technologies such as Apache Kafka.
- Experience working on large-scale data platforms handling high-volume data workloads.
- Exposure to additional AWS data services and cloud-native data architectures.






