Azure Data Engineer at VY SYSTEMS PRIVATE LIMITED · Bengaluru (Bangalore) · 9 - 16 years · ₹4L - ₹17L / yr · Profitable · Posted 6 Aug 2026

Job Summary
We are seeking a skilled Azure Data Engineer with hands-on experience in Azure Data Services, Azure Databricks, Python, PySpark, and SQL. The ideal candidate will be responsible for designing, developing, and optimizing scalable data pipelines and data processing solutions to support business intelligence, analytics, and reporting requirements.
Key Responsibilities
- Design, develop, and maintain scalable ETL/ELT pipelines using Azure Databricks and PySpark.
- Build and optimize data processing workflows using Python and SQL.
- Develop and manage data ingestion pipelines from multiple structured and unstructured data sources.
- Work with Azure Data Factory (ADF) to orchestrate and schedule data pipelines.
- Implement data transformation and cleansing logic using PySpark.
- Optimize SQL queries and Spark jobs for performance and scalability.
- Collaborate with data architects, analysts, and business stakeholders to understand data requirements.
- Ensure data quality, integrity, and governance across the data platform.
- Monitor, troubleshoot, and resolve production data pipeline issues.
- Follow coding standards, version control, and CI/CD best practices.
Required Skills
- Strong experience with Microsoft Azure cloud services.
- Hands-on experience with Azure Databricks.
- Strong programming skills in Python.
- Expertise in PySpark for large-scale data processing.
- Strong SQL skills, including query optimization and performance tuning.
- Experience with Azure Data Factory (ADF).
- Knowledge of Delta Lake, Spark SQL, and Databricks notebooks.
- Experience with Git or Azure DevOps for source code management.
- Understanding of data warehousing concepts and ETL/ELT processes.
Preferred Skills
- Experience with Azure Synapse Analytics.
- Knowledge of Delta Live Tables (DLT).
- Experience with Azure Data Lake Storage (ADLS Gen2).
- Familiarity with Unity Catalog and data governance.
- Exposure to CI/CD pipelines and infrastructure-as-code.
- Experience working in Agile/Scrum environments.

About VY SYSTEMS PRIVATE LIMITED
About
Vy Systems is a Global Technology consulting, Solutions, and Managed Technology Services company. We service our customers with ‘RESPONSIVENESS’ as a key factor and we believe that timely response to any transaction increases the operational efficiency and accelerates the revenue and profitability to our customers.
The Company is founded and managed by a team of professionals having more than two+ decades of global experience in the business of Technology Consulting and Services.
Tech stack
Similar jobs (10)
Job Description – Azure Data Engineer
Role: Azure Data Engineer
Experience: 9+ Years
Location: Bangalore / Hyderabad
Notice Period: Immediate to 15 Days
Interview Process: 1st Round – Virtual | 2nd Round – F2F
Mandatory Skills
- Python
- PySpark
- SQL
- Azure Data Engineering
Job Description
We are looking for an experienced Azure Data Engineer with 9+ years of experience and strong hands-on expertise in Python, PySpark, SQL, and Azure Data Engineering.
Key Responsibilities
- Develop and maintain scalable data engineering solutions using Azure.
- Build and optimize data processing pipelines using PySpark and Python.
- Write complex SQL queries for data extraction and transformation.
- Work with Azure data services and cloud-based data platforms.
- Perform data processing, transformation, and integration.
- Troubleshoot data pipeline and production issues.
- Collaborate with technical and business teams to deliver data solutions.
Preferred: Immediate to 15 Days joiners.
We are hiring a Databricks Data Engineer to build scalable data pipelines on the lakehouse.
Responsibilities
- Build batch and streaming pipelines in Databricks with PySpark
- Model data in Delta Lake
- Optimise Spark jobs for cost and performance
- Set up data quality checks and monitoring
Requirements
- 2+ years of data engineering with Databricks
- Strong PySpark and Spark SQL skills
- Cloud experience on Azure, AWS or GCP
Location: Bangalore Experience: 5 to 7 years Employment type: Full-time, permanent Work Hours: General Shift (10.00 AM to 7.00 PM) website: www.amazech.com Qualifications: Minimum B.E./B.Tech, or higher in Computer Science, Information Technology, Data Engineering, or a related field, with a good academic background. Key Responsibilities: • Design, develop, and maintain end-to-end data pipelines. • Build and optimize ETL/ELT processes for large-scale data processing. • Implement Azure-based data solutions ensuring scalability and performance. • Collaborate with stakeholders to understand business data requirements. • Perform performance tuning and optimization across data platforms. • Support production environments, conduct root cause analysis, and resolve data-related issues. • Maintain comprehensive technical documentation and structured knowledge transfer documents. • Work effectively within Agile/Scrum frameworks and contribute to sprint planning and delivery. • Ensure secure and compliant data handling using Azure best practices. Required Skills & Experience • End-to-end ETL/ELT pipeline development, integration, and performance optimization • Azure Synapse, Azure Logic Apps, Azure SQL, and Azure Databricks, with a strong focus on performance tuning • Microsoft Azure services, including Storage Accounts, Key Vault, and Cognitive Services • Advanced proficiency in Python development and modern productivity tools such as GitHub Copilot and Cursor • Strong documentation discipline, including technical design documentation, and structured knowledge transfer documents • Experience operating within Agile/Scrum frameworks, including production support, root cause analysis, and effective stakeholder collaboration
Job Description
• Design and Implement Data Solutions: Lead the design, development, and implementation of scalable and secure Azure-
based data platforms, ensuring integration with various data sources and business systems. Deliver at least 2 major
projects every year with a focus on data engineering best practices.
• Optimize Data Pipelines: Build and optimize end-to-end data pipelines using Azure Data Factory, Azure Databricks, and
Azure Synapse, with an emphasis on automating data workflows. Achieve a 20% reduction in pipeline execution times
within the first 6 months.
• Cloud Infrastructure Management: Manage and maintain the Azure data environment, ensuring high availability, disaster
recovery, and cost optimization. Track and improve system uptime to exceed 99.9% reliability.
• Collaborate with Cross-Functional Teams: Partner with data scientists, data analysts, and business stakeholders to translate
business requirements into efficient data solutions. Facilitate at least 3 collaborative sessions per quarter to address key
business use cases.
• Ensure Data Security & Compliance: Implement data security measures, ensuring compliance with industry regulations
(GDPR, HIPAA, etc.) and company policies. Achieve and maintain full compliance in all data environments within the first
quarter of onboarding.
• Continuous Learning & Knowledge Sharing: Stay up-to-date with emerging Azure technologies, and mentor junior
engineers to promote knowledge sharing. Complete 1 Azure certification annually and conduct at least 2 internal
knowledge-sharing sessions per year.
• Sound knowledge of data governance practices, data quality management, and data security principles.
• Play a pivotal role in shaping our organization's data-driven journey, driving innovation through data analytics and insights.
• Optimize data storage, processing and retrieval mechanisms for performance, cost, and scalability using data storage
services (such as Azure Data Lake Storage, Azure SQL Database, etc.), data processing services (such as Azure Data Bricks,
Azure Synapse, etc.) and data visualization (PowerBI, Qlik, etc.) & integration services (Data API builder, logic apps, etc.)
• Monitor and troubleshoot data platform performance, identify and resolve issues, and provide recommendations for
continuous improvement.
• Collaborate with DevOps teams to automate deployment, configuration, and monitoring processes using Azure DevOps,
PowerShell, or other relevant tools.
• Stay up to date with the latest trends and advancements in cloud data services and provide recommendations on adopting
new technologies or features to enhance the data platform.
• Document technical designs, procedures, and guidelines for data platform engineering and operations
Knowledge, Skills & Experience
Job Experience • Bachelor's degree in Computer Science, Engineering, or a related field. Advanced
degree preferred.
• Proven 6-10 years experience in playing platform engineer or admin role
• Experience with big data technologies such as Apache Spark, Hadoop, or similar
frameworks.
• Solid understanding of cloud computing concepts and experience with cloud
infrastructure management and provisioning.
• Solid understanding of network security concepts and technologies (such as
firewalls, VPNs, intrusion detection/prevention systems, etc.) and data security
concepts and technologies (such as access controls, encryption, observability,
privacy laws/regulations, etc.)
• Experience in a Retail setup is preferred.
Required Skills The position will require someone with the following:
• Strategic Planning
Public
• Communication and Collaboration
• Problem Solving Skills A/B testing & experimentation
• SQL, BI tools, and storytelling with data
About the Role
We are looking for a Senior Data Engineer with strong hands-on expertise in Databricks, Python, PySpark, and SQL to build scalable, high-performance data engineering solutions. You’ll architect and develop large scale, high-performance data pipelines capable of handling massive real-time and batch data volumes across multiple business systems. Databricks is the core enterprise data and processing platform for this role. You will also use Apache Airflow for workflow orchestration and dbt for ELT transformations, and will contribute to designing reliable, secure, and governed data platforms that enable analytics, reporting, and AI-driven use cases.
Key Responsibilities
- Design and implement large-scale data pipelines using Python/PySpark, Databricks, and Microsoft Fabric.
- Develop and optimize data processing workloads in Databricks using PySpark and Spark SQL, with a strong focus on scalability, reliability, performance, and maintainability.
- Develop and maintain dbt models including layered architecture, incremental models, snapshots, macros, testing, and documentation.
- Design, develop, and maintain Apache Airflow DAGs for orchestrating reliable, scalable, and observable data pipelines.
- Design and implement data quality, observability, and governance frameworks, including automated testing, monitoring, lineage, access control, and data privacy standards.
- Partner with analytics, product, and business stakeholders to turn requirements into trustworthy datasets, and raise the engineering bar through design discussions, code reviews, and mentoring junior engineers.
Required Skills
- Strong expertise in Python for developing scalable, modular, and production-ready data engineering applications.
- Strong expertise in PySpark, including DataFrame API, Spark SQL, Structured Streaming, partitioning strategies, joins, caching, handling data skew, and Spark performance optimization.
- Strong hands-on experience with Databricks for data ingestion, transformation, processing, and optimization, including Delta Lake, Unity Catalog, Databricks Workflows, notebooks, jobs, and Databricks-native data engineering capabilities.
- Strong experience in Databricks/Spark performance tuning, including query and job optimization, partitioning, file sizing, caching, join optimization, handling data skew, and efficient use of compute resources.
- Hands-on experience with Delta Lake, including transactional data processing, schema management, incremental data processing, and reliable batch and streaming data pipelines.
- Hands-on experience in developing dbt projects using layered architecture, incremental models, snapshots, macros/Jinja, testing, documentation, and deployment best practices.
- Expertise in advanced SQL and data modelling — dimensional modeling, slowly changing dimensions, schema evolution, and query optimization.
- Hands-on experience in developing and managing Apache Airflow DAGs, scheduling workflows, dependency management, retries, backfills, and operational monitoring.
- Hands-on experience with at least one major cloud platform (AWS, Azure or GCP).
- Strong problem-solving skills and the ability to work independently with business and analytics stakeholders.
Nice to Have
- Hands-on exposure to Microsoft Fabric for data integration and analytics.
- Experience using AI coding assistants (e.g. Claude Code, GitHub Copilot) as part of a development workflow.
- Familiarity with modern DevOps practices, including CI/CD pipelines, Infrastructure as Code (IaC), and containerization (Docker/Kubernetes).
- Domain expertise in financial services.
Dear Candidate,
Greeting from NAM Info Pvt Ltd.
We have a role for Data Engineer position with NAM Info.
This role will be permanent with NAM info and deploy to client
location NEW DELHI~CHENNAI~HYDERABAD~PUNE~KOLKATA.
Work Mode: WORK FROM OFFICE
A decent hike can be provided based on current CTC
Interview Mode: Virtual
Role Descriptions:
Exp Range: 7 - 10 years
City Locations: NEW DELHI~CHENNAI~HYDERABAD~PUNE~KOLKATA
Key Responsibilities*
Role: Data Engineer
Location: ~NEW DELHI~CHENNAI~HYDERABAD~PUNE~KOLKATA
Skills: Digital: Databricks, Azure Data Factory
Experience Required: 8-10
Descriptions:
Good information and sound knowledge in Azure Synapse Analytics Azure Data Factory (ADF)Big Data technologies and data processing frameworks Azure Data Warehouse and associated Azure data platform services Data integration| data modelling| and performance optimization
Desire candidate
- Candidate should have valid PF.
Regards,
NAM Info
Job Summary
We are looking for a skilled and experienced Data Engineer to join our growing data team. The ideal candidate will have strong expertise in Python, PySpark, Data Modeling, and Power BI, with hands-on experience in designing, developing, and optimizing scalable data solutions. The role requires working closely with business stakeholders, data architects, and analytics teams to build robust data pipelines and semantic models that enable data-driven decision-making.
Technical Skills
- Strong hands-on experience in Python and PySpark development.
- Expertise in building and optimizing Data Engineering solutions and ETL pipelines.
- Strong understanding of Data Modeling concepts (Star Schema, Snowflake Schema, Dimensional Modeling).
- Experience with Power BI Data Modeling and Semantic Layer development.
- Proficiency in DAX (Data Analysis Expressions).
- Experience designing and managing Semantic Models in Power BI.
- Strong SQL skills and experience working with large datasets.
- Knowledge of data warehousing concepts and best practices.
Preferred Skills
- Experience with cloud platforms such as Azure, AWS, or GCP.
- Exposure to modern data platforms like Databricks.
- Understanding of data governance and data quality frameworks.
Responsibilities and JD
Job Description: We are looking for a Senior Developer with strong expertise in PySpark, Databricks, and Snowflake to build scalable data engineering solutions and enterprise data platforms.
Key Responsibilities:
- Design, develop, and maintain ETL/ELT pipelines using PySpark, Databricks, and Snowflake.
- Develop batch and real-time data processing solutions for structured and semi-structured data.
- Build and optimize Databricks notebooks, workflows, and Delta Lake solutions.
- Design and implement Snowflake databases, schemas, views, stored procedures, tasks, and streams.
- Develop scalable data models, data marts, and data warehouse solutions.
- Optimize PySpark jobs, Databricks workloads, and Snowflake queries for performance and cost efficiency.
- Implement data quality, validation, governance, and security controls.
- Collaborate with business stakeholders, architects, and cross-functional teams to deliver data solutions.
- Manage source control and CI/CD deployments using Git and Azure DevOps.
- Troubleshoot production issues, perform root cause analysis, and ensure pipeline reliability.
- Mentor junior team members and participate in code reviews and technical design discussions.
Required Skills: PySpark, Databricks, Snowflake, Python, SQL.
Experience: 5+ years of Data Engineering experience with strong hands-on expertise in PySpark, Databricks, and Snowflake.
Data Engineer Hiring Post
🚨 Hiring: Data Engineer | PySpark + Python + SQL
We are looking for experienced Data Engineers to join our team!
🔹 Experience: 5 to 9 Years
🔹 Locations: Bangalore / Hyderabad
🔹 Interview Process:
• 1st Round – Virtual
• 2nd Round – Face-to-Face (Karat Test)
🔑 Key Skills:
✅ PySpark
✅ SQL
✅ Python
✅ ETL
📩 Interested candidates can share their updated resume.
#Hiring #DataEngineer #PySpark #Python #SQL #ETL #BangaloreJobs #HyderabadJobs #TechHiring #ImmediateHiring
Data Engineer – Microsoft Fabric
Location: Pune, India
Work Mode: Hybrid
Experience: 6+ Years
Employment Type: Full-time contactor
Compensation: As per market standards, commensurate with experience and expertise
Shift Timings: 2:00 PM – 11:00 PM IST
Notice Period: 0 – 15 days
About the Role
Jade Business Services (JBS) is seeking a Data Engineer – Microsoft Fabric to join our Pune team and work on enterprise-scale data transformation and analytics initiatives.
We are looking for a hands-on Data Engineer with strong experience in Microsoft Fabric, SQL, Python/PySpark and modern data engineering practices. The candidate will be responsible for building scalable data pipelines, implementing Lakehouse and Warehouse solutions, developing data models and supporting governed, reliable and AI-ready data platforms.
The ideal candidate should be comfortable working with architects, engineering teams and client stakeholders to translate business requirements into scalable and production-ready data solutions.
Roles and Responsibilities
- Design and develop data solutions using Microsoft Fabric, including OneLake, Lakehouse, Warehouse and Data Factory pipelines.
- Build and maintain scalable ETL/ELT pipelines for batch and incremental data processing.
- Develop data ingestion and transformation pipelines using Fabric Data Factory, SQL, Python and/or PySpark.
- Implement Medallion Architecture using Bronze, Silver and Gold layers.
- Work with Lakehouse and Fabric Warehouse for enterprise data processing and analytics.
- Develop and maintain data models, tables, views and optimized SQL queries.
- Build and support semantic models for Power BI and analytical workloads.
- Implement data quality, validation, monitoring and error-handling mechanisms.
- Work with metadata, lineage and governance requirements using Microsoft Purview.
- Implement data security, access controls and role-based permissions across data platforms.
- Support Data Product and domain-oriented data architecture principles.
- Follow DataOps practices including CI/CD, deployment, monitoring and production support.
- Troubleshoot pipeline failures, performance issues and data quality problems.
- Optimize data pipelines, queries and storage for performance and cost efficiency.
- Work closely with Data Architects and business stakeholders to understand requirements and implement technical solutions.
- Participate in technical design discussions, code reviews and architecture reviews.
- Maintain technical documentation, data flow diagrams and pipeline documentation.
- Support production deployments, incident resolution and SLA-driven data platform operations.
- Identify opportunities for automation and AI-assisted improvements across data engineering processes.
Qualifications and Skills
- 6+ years of experience in Data Engineering, Data Integration or Data Platform development.
- Strong hands-on experience with Microsoft Fabric.
- Experience with:
- Microsoft Fabric Lakehouse
- Fabric Warehouse
- OneLake
- Fabric Data Factory / Pipelines
- Semantic Models
- Strong understanding of Lakehouse and Medallion Architecture.
- Strong SQL development and query optimization skills.
- Hands-on experience with Python and/or PySpark.
- Experience developing enterprise ETL/ELT and data integration pipelines.
- Experience with batch and incremental data processing.
- Understanding of data modelling concepts including dimensional modelling.
- Knowledge of data quality, metadata, lineage and data governance.
- Working knowledge of Microsoft Purview.
- Understanding of Data Mesh and Data Product concepts.
- Experience with CI/CD, version control, monitoring and DataOps practices.
- Understanding of cloud security, access controls and data privacy.
- Good troubleshooting and problem-solving skills.
- Strong communication skills and ability to work with distributed and client-facing teams.
Preferred Skills
- Microsoft Fabric or Azure Data certifications.
- Experience migrating workloads from Azure Synapse, SQL Server, Databricks or other data platforms to Microsoft Fabric.
- Experience implementing Medallion Architecture on Microsoft Fabric.
- Experience with Power BI and semantic modelling.
- Exposure to AI/ML, Generative AI or Agentic AI use cases on enterprise data platforms.
- Experience working with Data Products or domain-oriented data solutions.
- Experience in Energy & Utilities, Healthcare, Financial Services or Insurance.
- Experience working with US or international enterprise clients.
What We Expect
The ideal candidate should be hands-on first and capable of independently building, troubleshooting and optimizing Fabric data solutions. You should be able to explain the technical decisions behind your implementation and work effectively with architects and engineering teams to deliver production-ready solutions.






