Data Architect (AWS, Medalliion Architecture) at TalentXO · Bengaluru (Bangalore), Mumbai, Hyderabad, Gurugram · 8 - 12 years · ₹24L - ₹35L / yr · Profitable · Posted 14 May 2026

Role & Responsibilities
- Own end-to-end data architecture across Medallion layers (Bronze/Silver/Gold) on AWS
- Design atomic star schema data models (fact and dimension tables) across multiple business domains
- Define and implement aggregation and KPI layers aligned to business reporting requirements
- Establish data modelling standards, naming conventions, and design patterns across teams
- Ensure alignment with enterprise data strategy and roadmap
- Translate business processes into scalable and reusable data models
- Write and review SQL, Python, and Spark code for pipelines and transformations
- Build and validate data models in Snowflake or Databricks
- Deliver pipelines using AWS services (S3, Glue, Athena)
- Perform data reconciliation and validation across layers (source → curated → BI)
- Ensure data quality, consistency, and integrity across the pipeline
- Own end-to-end SDLC (requirements → design → build → test → release → operate)
- Lead and unblock data engineers and analysts on day-to-day delivery
- Identify blockers, dependencies, and risks early, and propose solutions
- Align delivery with programme roadmap and milestones
- Track execution using Jira and maintain documentation in Confluence
- Gather and validate requirements across multiple business domains
- Translate business requirements into technical designs and data models
- Communicate architecture decisions clearly to non-technical stakeholders
- Drive cross-team alignment to ensure consistent, reliable, and trusted outputs
Ideal Candidate
- Strong Data Architect Profile (AWS / Medallion Architecture / Analytics Transformation)
- Mandatory (Experience 1) – Must have 8+ years of experience in Data Engineering / Data Architecture with strong exposure to large-scale analytics transformation programmes
- Mandatory (Experience 2) – Strong hands-on experience designing and implementing Medallion Architecture (Bronze / Silver / Gold layers) on AWS-based data platforms
- Mandatory (Experience 3) – Strong hands-on experience with Snowflake or Databricks, including data modelling, performance optimization, transformation pipelines, and scalable data warehouse implementations
- Mandatory (Experience 4) – Must have strong expertise in advanced SQL including complex joins, CTEs, window functions, query optimization, and performance tuning
- Mandatory (Experience 5) – Must have Hands-on development experience with Python and Apache Spark/PySpark for large-scale data transformation and processing pipelines
- Mandatory (Experience 6) – Strong experience working with AWS data services including S3, Glue, Athena, and cloud-native analytics/data lake architectures
- Mandatory (Experience 7) – Strong experience in end-to-end ETL/ELT pipeline development including ingestion, transformation, reconciliation, validation, testing, deployment, and production support
- Mandatory (Experience 8) – Experience working across transaction-heavy enterprise domains such as Finance, Supply Chain, HR, Customer, or Operations datasets
- Mandatory (Note) – Only immediate joiners or candidates who can join within 15 days will be considered
- Preferred (Experience) – Experience working with ERP / enterprise systems such as SAP, Oracle, Salesforce, or similar enterprise platforms
- Preferred (Frameworks) – Familiarity with APQC, SCOR, or enterprise process modelling frameworks is an added advantage

About TalentXO
About
Company social profiles
Similar jobs (10)
The recruiter has not been active on this job recently. You may apply but please expect a delayed response.
Example:
We are looking for an experienced Data Architect to design, develop, and manage the organization's enterprise data architecture. The candidate will be responsible for building scalable data platforms, ensuring data quality and governance, and supporting business analytics through modern data solutions.
Experience Required
Mention the minimum years of experience.
Example:
- Minimum 12 years of experience in Data Architecture, Data Engineering, or related fields.
- 5+ years of experience in the Banking/Financial Services domain is preferred.
Educational Qualification
Mention the required degree.
Example:
- BE/BTech in Computer Science, Information Technology, Software Engineering, Electronics & Communication Engineering, or equivalent.
- OR MCA/MTech/MSc in Computer Science, IT, or related disciplines.
- MBA is preferred.
Technical Skills
List the skills the candidate must have.
Example:
- AWS, Azure, or GCP
- Data Warehousing (DWH)
- ETL/ELT
- Database Management
- Data Modeling
- Data Analytics
- Data Lakes
- Data Governance
Key Responsibilities
Convert the points you received into simple action statements.
Example:
- Design and maintain enterprise data architecture.
- Develop data warehouses and data lakes.
- Define data standards and governance policies.
- Ensure data quality and security.
- Design ETL/ELT processes.
- Integrate data from multiple systems.
- Plan and execute data migration projects.
- Review existing data architecture and recommend improvements.
- Provide technical guidance to project teams.
- Evaluate new data technologies and tools.
Preferred Skills
These are not mandatory but are an advantage.
Example:
- Banking domain experience
- Strong analytical and problem-solving skills
- Good communication skills
- Leadership and stakeholder management
- Experience mentoring technical teams
Solution Architect – AZURE Data Engineering
Job Overview
We are looking for an experienced Solution Architect – Data Engineering with strong expertise in designing data solutions and hands-on experience with Azure, Synapse, PySpark, Data Warehousing, and Data Lakes. The ideal candidate should have strong architectural and data engineering knowledge.
Key Responsibilities
- Design and implement scalable data architecture and solutions.
- Develop and manage Data Warehouse and Data Lake architectures.
- Design data platforms using Medallion Architecture.
- Lead Data Engineering and ETL activities.
- Work with Azure Synapse Analytics for data processing and analytics.
- Develop data solutions using PySpark / Apache Spark.
- Define and implement data validation and data quality processes.
- Collaborate with business, data, and technology teams to deliver effective data solutions.
Required Skills
- Strong experience in Solution Architecture / Data Architecture.
- Strong knowledge of Data Warehouse and Data Lake architecture.
- Good understanding of Medallion Architecture.
- Strong experience in Data Engineering and ETL.
- Hands-on experience with Microsoft Azure and Azure Synapse Analytics.
- Strong knowledge of PySpark / Apache Spark.
- Experience with Data Validation and Data Quality.
- Good communication and stakeholder management skills.
Experience
8+ Years

About the Role
You'll be at the forefront of designing and implementing robust data platform solutions that power advanced analytics, AI, and machine learning. Working with modern cloud technologies, you'll build scalable data foundations that enable clients to make smarter, data-driven decisions.
Key Responsibilities
- Build scalable data pipelines using Snowflake, AWS, GCP, and Databricks.
- Design and optimize data models for AI and machine learning workloads.
- Develop reliable data foundations for MLOps, governance, and data lineage.
- Integrate data from multiple sources into modern data platforms.
- Leverage Snowpark ML and Snowflake's native AI capabilities.
- Ensure data platforms are secure, scalable, and high-performing.
What We're Looking For
- 5+ years of hands-on experience with Snowflake.
- Strong proficiency in SQL and Python.
- Experience with AWS, Azure, or GCP.
- Knowledge of cloud storage services such as S3, ADLS, or GCS.
- Strong understanding of Dimensional Modeling and Data Vault.
- Experience with Scala or Java is a plus.
Tech Stack
- Data Warehouse: Snowflake
- Programming: SQL, Python, Scala (Good to Have), Java (Good to Have)
- Cloud: AWS, Azure, GCP
- Storage: S3, ADLS, GCS
- AI/ML: Snowpark ML, MLOps
Perks & Benefits
- Public Speaking & Communication Program
- Mentoring Program with Senior Support Leads
- 360° Progress Reviews
- Weekly Learning Sessions & Guilds
- Paid Certifications
- Hackathons & Innovation Days
- Recognition & Rewards Programs
- Team Socials & Annual Offsites
- Employee Assistance Program (24/7 Wellbeing Support)
The Data People Shaping Tomorrow
Our client helps organizations unlock the power of data through modern cloud, analytics, and AI solutions. We believe in creating an environment where talented technologists can learn, innovate, and make a real impact while building cutting-edge data platforms for global clients. If you're passionate about data engineering and want to work with the latest technologies in AI, cloud, and analytics, we'd love to hear from you.
Role Summary:
We are looking for an experienced Snowflake Lead to lead the design, development, migration, and optimization of enterprise data platforms using Snowflake. The candidate will provide technical leadership to data engineering teams and work closely with architects, business stakeholders, and application teams.
Key Responsibilities
- Lead the architecture and development of scalable Snowflake data warehouse solutions.
- Design and develop scalable Azure Data Factory (ADF) pipelines for API-based and batch data ingestion, implementing parameterized workflows, scheduling, and error handling.
- Build and optimize enterprise Snowflake data warehouse solutions using Snowflake SQL, Streams, Tasks, Stored Procedures, VARIANT data type, and LATERAL FLATTEN for semi-structured JSON processing.
- Integrate GraphQL and REST APIs using OAuth 2.0, implementing secure API authentication, JSON parsing, and API validation using Postman.
- Develop cloud-based data ingestion solutions using Azure Data Lake Storage Gen2 (ADLS) as the landing layer and Azure Key Vault for secure credential management.
- Design metadata-driven ELT frameworks with incremental loading, audit logging, watermark processing, and automated data orchestration.
- Optimize Snowflake performance through warehouse sizing, query tuning, clustering strategies, Time Travel, Cloning, and warehouse management best practices.
- Collaborate with DevOps teams using Azure DevOps for source control, CI/CD deployment, release management, and Agile delivery.
- Design dimensional data models, build curated data marts, and support enterprise reporting and analytics requirements.
- Develop and maintain Power BI semantic models, datasets, dashboards, and reports; knowledge of DAX, Power Query, and data visualization best practices is preferred.
- Work closely with business stakeholders, solution architects, and cross-functional teams to deliver secure, scalable, and high-performance cloud data platform solutions.
Thanks,
Mounika P
About Us:
The QX Impact was launched with a mission to make A.I accessible and affordable and deliver AI Products/Solutions at scale for the enterprises by bringing the power of Data, AI, and Engineering to drive digital transformation. We believe without insights; businesses will continue to face challenges to better understand their customers and even lose them. Secondly, without insights businesses won't’ be able to deliver differentiated products/services; and finally, without insights, businesses can’t achieve a new level of “Operational Excellence” is crucial to remain competitive, meeting rising customer expectations, expanding markets, and digitalization.
Job Summary:
We are looking for a Senior Data Engineer who is creative, collaborative, and adaptable to join our agile team of data scientists, engineers, and UX developers. The role focuses on building and maintaining robust data pipelines to support advanced analytics, data science, and BI solutions.
As a Senior Data Engineer, you will work with internal and external data, collaborate with data scientists, and contribute to the design, development, and deployment of innovative solutions.
Key Responsibilities:
- Design, develop, test, and maintain optimal data pipeline and ETL architectures.
- Map out data systems and define/design required integrations, ETL, BI, and AI systems/processes.
- Prepare and optimize data for predictive and prescriptive modeling.
- Collaborate with teams to integrate ERP data into the enterprise data lake, ensuring seamless flow and quality.
- Enhance cloud data infrastructure on AWS or Azure for scalability and performance.
- Utilize big data tools and frameworks to optimize data acquisition and preparation.
- Build architectures to move data to/from data lakes and data warehouses for advanced analytics.
- Develop and curate data models for analytics, dashboards, and reports.
- Conduct code reviews, maintain production-level code, and implement testing approaches.
- Monitor, troubleshoot, and resolve data ingestion workflows to maintain reliability and uptime.
- Drive innovation and implement efficient new approaches to data engineering tasks.
Must-Have Skills:
- Bachelor’s degree in Computer Science, Mathematics, Engineering, or a related field.
- 5+ years of experience working with enterprise data platforms, including building and managing data lakes.
- 3–5 years of experience designing and implementing data warehouse solutions.
- Expertise in SQL, including developing stored procedures (SP) and applying advanced data design concepts.
- Proficiency in Spark (Python/Scala) and Spark Streaming for real-time data pipelines.
- Experience with AWS or Azure services (e.g., AWS Glue, Azure Data Factory, Redshift, Snowflake).
- Familiarity with big data tools such as Apache Kafka, Apache Spark, or Flink.
- Hands-on experience with orchestration tools (e.g., Apache Airflow, Prefect).
- Knowledge of CI/CD processes, version control (e.g., Git, Jenkins), and deployment automation.
- Strong problem-solving, communication, and collaboration skills.
Good-to-Have Skills:
- Experience in integrating ERP data into data lakes.
- Experience with traditional ETL tools (e.g., Talend, Pentaho).
Competencies:
- Tech Savvy - Anticipating and adopting innovations in business-building digital and technology applications.
- Self-Development - Actively seeking new ways to grow and be challenged using both formal and informal development channels.
- Action Oriented - Taking on new opportunities and tough challenges with a sense of urgency, high energy, and enthusiasm.
- Customer Focus - Building strong customer relationships and delivering customer-centric solutions.
- Optimize Work Processes - Knowing the most effective and efficient processes to get things done, with a focus on continuous improvement.
Why Join Us?
- Be part of a collaborative and agile team driving cutting-edge AI and data engineering solutions.
- Work on impactful projects that make a difference across industries.
- Opportunities for professional growth and continuous learning.
- Competitive salary and benefits package.
Application Details
Ready to make an impact? Apply today and become part of the QX Impact team!
4 - 10 years of experience in designing and buildingarchitecting highly resilient data platforms
∙Strong knowledge of data engineering, architecture and data modeling
∙Experience in platforms like Databricks and Snowflake
∙Experience on building applications on cloud (AWS or Azure or Google Cloud)
∙Strong analytical and problem-solving skills
∙Prior experience in developing data or computation intensive (e.g. grid based) backend applications is an
advantage
∙OOP design skills with an understanding or at least personal interest towards the concepts of Functional
Programming
∙Willingness to understand and enhance other people’s code, being able to work in an environment where
developers will oversee and work on wider components also dealing with older “legacy” code
∙Strong programming skills (Java/ Scala / Python) skills with the willingness to pick up the other language if not
already mastered at a sufficient level is important
∙Spring knowledge is an advantage, but in general willingness to learn, work with and even enhance in-house
developed frameworks is a must
∙Prior experience in working with Git, Bitbucket, Jenkins, working with PR-s, using JIRA, following the Scrum Agile
methodology is an advantage
∙Prior knowledge of financial products is an advantage
∙Bachelors or Masters in any relevant field of IT/Engineering area is an advantage
At Mitratech, we are a team of technocrats focused on building world-class products that simplify operations in the Legal, Risk, Compliance, and HR functions. We are a close-knit, globally dispersed team that thrives in an ecosystem that supports individual excellence and takes pride in its diverse and inclusive work culture centered around great people practices, learning opportunities, and having fun! Our culture is the ideal blend of entrepreneurial spirit and enterprise investment, enabling the chance to move at a rapid pace with some of the most complex, leading-edge technologies available.
For over 35 years, the experts at Mitratech have been focused on solving the complex needs. Today, we serve 20,000 client companies of all sizes globally, representing 30% of the Fortune 500 and over 500,000 users in over 160 countries.
As we continue to grow, we’re always looking for resourceful, enthusiastic, and fresh perspectives. Join our global team and see what makes Mitratech a truly exceptional place to work!
Job Overview
Principal Data Engineer
About Engineering at Mitratech Legal Solutions
Mitratech's engineering organization is a collaborative and dynamic environment where engineers are empowered to drive technical direction and innovation. Our engineers are passionate about delivering high-quality products and solutions that meet the evolving needs of our customers, and we're committed to fostering a culture of continuous learning and growth.
About the Role
Mitratech is a fast-paced and dynamic environment, and this role requires someone who is adaptable, resilient, and able to thrive in a rapidly changing landscape. If you’re a seasoned engineer with a passion for technical leadership, innovation, and collaboration — including building the data foundations that power trusted reporting and agentic AI-driven products — we’d love to hear from you.
What You Will Do
• Drive technical direction for a significant product domain or platform capability, ensuring alignment with business objectives and customer needs
• Design and maintain data pipelines and reporting models that power trusted business metrics and increasingly feed agentic AI systems (e.g., RAG ingestion, embeddings, vector stores, AI agent workflows)
• Use AI-assisted and agentic engineering tools (e.g., Claude Code, Copilot, Cursor, AI agents) as part of your own workflow, and help other engineers adopt agentic development practices effectively
• Reduce systemic complexity by identifying and leading architectural debt remediation, and developing strategies for ongoing technical debt management
• Partner with Product and Engineering leadership to inform multi-quarter roadmap feasibility, and provide technical guidance and oversight to ensure successful implementation
• Elevate engineering craft across multiple teams through RFCs, mentorship, and knowledge sharing, and develop training programs to improve engineering skills and knowledge
• Represent Mitratech’s technical capabilities externally, including speaking at conferences, contributing to open-source projects, and engaging with industry peers and thought leaders
What We Are Looking For
To be successful in this role, you will need:
• 10+ years of experience in software engineering, with a focus on technical leadership and architecture
• Deep understanding of data engineering principles, including data modeling, data warehousing, reporting, and data governance
• Strong technical expertise in SQL, PostgreSQL, ETL/ELT pipelines, BI tools, and analytics platforms
• Practical experience with AI/LLM-adjacent and agentic AI data work — e.g., RAG ingestion pipelines, embedding generation, vector store management, or building/operating AI agent workflows over data — using AI coding assistants (Claude Code, Copilot, Cursor, or similar) as a regular part of the engineering workflow
• Working knowledge of modern cloud platforms such as AWS
• Experience with BI, reporting, dashboards, and customer-facing analytics
• Experience leading cross-functional initiatives with product, engineering, analytics, and business teams
Nice to Have
• Working knowledge of Ruby on Rails and React
• Experience with a semantic or metrics layer (e.g., dbt Semantic Layer, headless BI)
• Understanding of CI/CD, Git-based workflows, and infrastructure-as-code
The Stack Context
• Modern data stack: Fivetran, Airbyte, dbt, Snowflake, GitHub, Terraform, or similar tools
• Application context (nice to have): Ruby on Rails, React, or similar backend/frontend frameworks
• Data modeling: SQL, analytics models, documentation, testing, naming standards, and version control
• Infrastructure: cloud-based data infrastructure, infrastructure-as-code, CI/CD, monitoring, and cloud storage
• Data workflows: ingestion, transformation, orchestration, reporting, deployment, and change management
• Reporting focus: trusted metrics, scalable reporting models, dashboards, exports, and data quality
• AI surface: data pipelines and quality practices supporting AI/LLM and agentic AI use cases (RAG, embeddings, vector stores, AI agents) alongside traditional BI
Why This Role
This role offers a unique opportunity to drive technical direction and innovation at a rapidly growing company, while also mentoring and coaching engineers to improve their craft. As a Principal Data Engineer at Mitratech, you will have the chance to work on complex and challenging problems spanning trusted reporting and agentic AI systems, collaborate with cross-functional teams, and represent the company's technical capabilities externally. If you're looking for a role that offers a mix of technical leadership, data and reporting depth, agentic AI innovation, and collaboration, this could be the perfect fit for you.
We are an equal-opportunity employer that values diversity at all levels. All qualified applicants will receive consideration for employment without regard to race, color, religion, gender, national origin, age, sexual orientation, gender identity, disability, or veteran status.
Data Engineer – Microsoft Fabric
Location: Pune, India
Work Mode: Hybrid
Experience: 6+ Years
Employment Type: Full-time contactor
Compensation: As per market standards, commensurate with experience and expertise
Shift Timings: 2:00 PM – 11:00 PM IST
Notice Period: 0 – 15 days
About the Role
Jade Business Services (JBS) is seeking a Data Engineer – Microsoft Fabric to join our Pune team and work on enterprise-scale data transformation and analytics initiatives.
We are looking for a hands-on Data Engineer with strong experience in Microsoft Fabric, SQL, Python/PySpark and modern data engineering practices. The candidate will be responsible for building scalable data pipelines, implementing Lakehouse and Warehouse solutions, developing data models and supporting governed, reliable and AI-ready data platforms.
The ideal candidate should be comfortable working with architects, engineering teams and client stakeholders to translate business requirements into scalable and production-ready data solutions.
Roles and Responsibilities
- Design and develop data solutions using Microsoft Fabric, including OneLake, Lakehouse, Warehouse and Data Factory pipelines.
- Build and maintain scalable ETL/ELT pipelines for batch and incremental data processing.
- Develop data ingestion and transformation pipelines using Fabric Data Factory, SQL, Python and/or PySpark.
- Implement Medallion Architecture using Bronze, Silver and Gold layers.
- Work with Lakehouse and Fabric Warehouse for enterprise data processing and analytics.
- Develop and maintain data models, tables, views and optimized SQL queries.
- Build and support semantic models for Power BI and analytical workloads.
- Implement data quality, validation, monitoring and error-handling mechanisms.
- Work with metadata, lineage and governance requirements using Microsoft Purview.
- Implement data security, access controls and role-based permissions across data platforms.
- Support Data Product and domain-oriented data architecture principles.
- Follow DataOps practices including CI/CD, deployment, monitoring and production support.
- Troubleshoot pipeline failures, performance issues and data quality problems.
- Optimize data pipelines, queries and storage for performance and cost efficiency.
- Work closely with Data Architects and business stakeholders to understand requirements and implement technical solutions.
- Participate in technical design discussions, code reviews and architecture reviews.
- Maintain technical documentation, data flow diagrams and pipeline documentation.
- Support production deployments, incident resolution and SLA-driven data platform operations.
- Identify opportunities for automation and AI-assisted improvements across data engineering processes.
Qualifications and Skills
- 6+ years of experience in Data Engineering, Data Integration or Data Platform development.
- Strong hands-on experience with Microsoft Fabric.
- Experience with:
- Microsoft Fabric Lakehouse
- Fabric Warehouse
- OneLake
- Fabric Data Factory / Pipelines
- Semantic Models
- Strong understanding of Lakehouse and Medallion Architecture.
- Strong SQL development and query optimization skills.
- Hands-on experience with Python and/or PySpark.
- Experience developing enterprise ETL/ELT and data integration pipelines.
- Experience with batch and incremental data processing.
- Understanding of data modelling concepts including dimensional modelling.
- Knowledge of data quality, metadata, lineage and data governance.
- Working knowledge of Microsoft Purview.
- Understanding of Data Mesh and Data Product concepts.
- Experience with CI/CD, version control, monitoring and DataOps practices.
- Understanding of cloud security, access controls and data privacy.
- Good troubleshooting and problem-solving skills.
- Strong communication skills and ability to work with distributed and client-facing teams.
Preferred Skills
- Microsoft Fabric or Azure Data certifications.
- Experience migrating workloads from Azure Synapse, SQL Server, Databricks or other data platforms to Microsoft Fabric.
- Experience implementing Medallion Architecture on Microsoft Fabric.
- Experience with Power BI and semantic modelling.
- Exposure to AI/ML, Generative AI or Agentic AI use cases on enterprise data platforms.
- Experience working with Data Products or domain-oriented data solutions.
- Experience in Energy & Utilities, Healthcare, Financial Services or Insurance.
- Experience working with US or international enterprise clients.
What We Expect
The ideal candidate should be hands-on first and capable of independently building, troubleshooting and optimizing Fabric data solutions. You should be able to explain the technical decisions behind your implementation and work effectively with architects and engineering teams to deliver production-ready solutions.
We are looking for a Lead Data Architect to design, build, and scale our data pipelines and entity resolution systems. This role combines deep technical expertise in data engineering with hands-on experience in AI-assisted tooling, entity matching, and data integration from diverse sources. You will lead architectural decisions for our data platform, mentor engineers, and ensure our pipelines are reliable, scalable, and production-grade.
Key Responsibilities
Architect, build, and maintain robust, scalable data pipelines that ingest, transform, and serve data from multiple internal and external sources.
Own the end-to-end orchestration of data workflows using tools like Dagster, ensuring observability, reliability, and maintainability of pipelines.
Design and implement entity resolution workflows — including matching, merging, and survivorship logic — using tools such as Splink, to produce clean, deduplicated, golden records.
Build and maintain web scrapers to source data from external providers, ensuring resilience to source changes, rate limits, and data quality issues.
Integrate and reconcile data coming from multiple, often inconsistent, sources into unified, trustworthy datasets.
Design and maintain data models and schemas across transactional and analytical systems, ensuring consistency, scalability, and performance.
Leverage AI/LLM-based tools and techniques to enhance data pipeline capabilities — e.g., intelligent data extraction, automated data quality checks, or AI-assisted entity matching.
Define and enforce best practices around pipeline design, testing, monitoring, and documentation.
Collaborate closely with data engineers, product managers, and other stakeholders to translate business requirements into scalable data architecture.
Provide technical leadership and mentorship to the data engineering team.
Required Skills & Experience
Strong hands-on experience building and maintaining production-grade data pipelines at scale.
Practical experience with Dagster (or similar orchestration tools like Airflow/Prefect) for pipeline orchestration.
Experience with Splink or similar probabilistic/deterministic record linkage tools for entity matching, merging, and survivorship.
Strong proficiency in Python, including experience writing and maintaining web scrapers.
Proven experience integrating and maintaining data pipelines that pull from multiple, heterogeneous data sources.
Experience applying AI/ML tools within data engineering workflows (e.g., LLM-assisted data cleaning, extraction, or matching).
Hands-on experience with relational and distributed databases such as PostgreSQL and Google Cloud Spanner.
Strong understanding of data modeling principles (normalization, dimensional modeling, schema design) across OLTP and OLAP systems.
Experience with cloud data warehousing platforms such as BigQuery, Redshift, and cloud platforms (GCP/AWS/Azure).
Strong communication skills and experience working cross-functionally with engineering and product teams.
Experience with distributed data processing frameworks (e.g., Spark, Dask).
Familiarity with data governance, lineage, and cataloging tools.
Prior experience in a lead or architect-level role guiding a data engineering team.
What We're Looking For
A technically strong, hands-on leader who can balance architectural thinking with the practical grit of debugging a flaky scraper or tuning a matching algorithm — someone who's comfortable owning both the big picture and the messy details of real-world data.
Next
The recruiter has not been active on this job recently. You may apply but please expect a delayed response.
Job Title : Senior Data Engineer – Databricks
Experience : 14 to 20 Years
Location : HSR Layout, Bangalore
Work Mode : Hybrid – 3 Days WFO
Shift : 11:30 AM – 07:30 PM IST
Positions : 2
Notice Period : Immediate Joiners Only
Interview : 1 Technical Round + 2 Client Rounds
Role Overview :
We are looking for a Senior Data Engineer to build and lead enterprise-scale data platforms for a Switzerland-based commodity client.
The role requires a strong hands-on Data Engineering professional with expertise in Databricks, PySpark, Python, SQL, and AWS, along with technical leadership and stakeholder management experience.
Must-Have Skills :
- 14 to 20 years of Data Engineering experience
- Databricks & Apache Spark / PySpark
- Python & SQL
- AWS Cloud
- Lakehouse Architecture
- ETL / ELT & Distributed Data Processing
- Batch & Streaming Pipelines
- Data Pipeline Optimization & Data Modeling
- CDC & Incremental Processing
- Git, CI/CD & Testing
- Data Quality, Monitoring & Observability
- Technical Leadership & Stakeholder Management
Key Responsibilities :
- Design and build scalable data pipelines using Databricks, PySpark, Python, SQL, and AWS.
- Own data products from design through production.
- Develop batch / streaming pipelines and reusable ETL / ELT frameworks.
- Optimize pipelines for performance, scalability, reliability, and cost.
- Design scalable data architectures and data models.
- Implement data quality, monitoring, lineage, and CI/CD practices.
- Lead technical discussions and mentor engineering teams.
- Collaborate with business stakeholders, architects, product owners, and engineering teams.
- Remain hands-on while providing technical leadership.
Ideal Candidate :
A 14 to 20 years experienced, hands-on Data Engineering leader with strong Databricks + PySpark + AWS expertise, excellent communication, stakeholder management, and experience delivering enterprise-scale data platforms.
🔴 Super Urgent : Only Bangalore-based immediate joiners.





