Data Engineer at MindCrew Technologies · Pune · 8 - 12 years · ₹10L - ₹15L / yr · Profitable · Posted 29 Aug 2025
Job Title: Lead Data Engineer
📍 Location: Pune
🧾 Experience: 10+ Years
💰 Budget: Up to 1.7 LPM
Responsibilities
- Collaborate with Data & ETL teams to review, optimize, and scale data architectures within Snowflake.
- Design, develop, and maintain efficient ETL/ELT pipelines and robust data models.
- Optimize SQL queries for performance and cost efficiency.
- Ensure data quality, reliability, and security across pipelines and datasets.
- Implement Snowflake best practices for performance, scaling, and governance.
- Participate in code reviews, knowledge sharing, and mentoring within the data engineering team.
- Support BI and analytics initiatives by enabling high-quality, well-modeled datasets.

About MindCrew Technologies
About
Connect with the team
Company social profiles
Similar jobs (10)
Role Summary:
We are looking for an experienced Snowflake Lead to lead the design, development, migration, and optimization of enterprise data platforms using Snowflake. The candidate will provide technical leadership to data engineering teams and work closely with architects, business stakeholders, and application teams.
Key Responsibilities
- Lead the architecture and development of scalable Snowflake data warehouse solutions.
- Design and develop scalable Azure Data Factory (ADF) pipelines for API-based and batch data ingestion, implementing parameterized workflows, scheduling, and error handling.
- Build and optimize enterprise Snowflake data warehouse solutions using Snowflake SQL, Streams, Tasks, Stored Procedures, VARIANT data type, and LATERAL FLATTEN for semi-structured JSON processing.
- Integrate GraphQL and REST APIs using OAuth 2.0, implementing secure API authentication, JSON parsing, and API validation using Postman.
- Develop cloud-based data ingestion solutions using Azure Data Lake Storage Gen2 (ADLS) as the landing layer and Azure Key Vault for secure credential management.
- Design metadata-driven ELT frameworks with incremental loading, audit logging, watermark processing, and automated data orchestration.
- Optimize Snowflake performance through warehouse sizing, query tuning, clustering strategies, Time Travel, Cloning, and warehouse management best practices.
- Collaborate with DevOps teams using Azure DevOps for source control, CI/CD deployment, release management, and Agile delivery.
- Design dimensional data models, build curated data marts, and support enterprise reporting and analytics requirements.
- Develop and maintain Power BI semantic models, datasets, dashboards, and reports; knowledge of DAX, Power Query, and data visualization best practices is preferred.
- Work closely with business stakeholders, solution architects, and cross-functional teams to deliver secure, scalable, and high-performance cloud data platform solutions.
Thanks,
Mounika P
At Mitratech, we are a team of technocrats focused on building world-class products that simplify operations in the Legal, Risk, Compliance, and HR functions. We are a close-knit, globally dispersed team that thrives in an ecosystem that supports individual excellence and takes pride in its diverse and inclusive work culture centered around great people practices, learning opportunities, and having fun! Our culture is the ideal blend of entrepreneurial spirit and enterprise investment, enabling the chance to move at a rapid pace with some of the most complex, leading-edge technologies available.
For over 35 years, the experts at Mitratech have been focused on solving the complex needs. Today, we serve 20,000 client companies of all sizes globally, representing 30% of the Fortune 500 and over 500,000 users in over 160 countries.
As we continue to grow, we’re always looking for resourceful, enthusiastic, and fresh perspectives. Join our global team and see what makes Mitratech a truly exceptional place to work!
Job Overview
Principal Data Engineer
About Engineering at Mitratech Legal Solutions
Mitratech's engineering organization is a collaborative and dynamic environment where engineers are empowered to drive technical direction and innovation. Our engineers are passionate about delivering high-quality products and solutions that meet the evolving needs of our customers, and we're committed to fostering a culture of continuous learning and growth.
About the Role
Mitratech is a fast-paced and dynamic environment, and this role requires someone who is adaptable, resilient, and able to thrive in a rapidly changing landscape. If you’re a seasoned engineer with a passion for technical leadership, innovation, and collaboration — including building the data foundations that power trusted reporting and agentic AI-driven products — we’d love to hear from you.
What You Will Do
• Drive technical direction for a significant product domain or platform capability, ensuring alignment with business objectives and customer needs
• Design and maintain data pipelines and reporting models that power trusted business metrics and increasingly feed agentic AI systems (e.g., RAG ingestion, embeddings, vector stores, AI agent workflows)
• Use AI-assisted and agentic engineering tools (e.g., Claude Code, Copilot, Cursor, AI agents) as part of your own workflow, and help other engineers adopt agentic development practices effectively
• Reduce systemic complexity by identifying and leading architectural debt remediation, and developing strategies for ongoing technical debt management
• Partner with Product and Engineering leadership to inform multi-quarter roadmap feasibility, and provide technical guidance and oversight to ensure successful implementation
• Elevate engineering craft across multiple teams through RFCs, mentorship, and knowledge sharing, and develop training programs to improve engineering skills and knowledge
• Represent Mitratech’s technical capabilities externally, including speaking at conferences, contributing to open-source projects, and engaging with industry peers and thought leaders
What We Are Looking For
To be successful in this role, you will need:
• 10+ years of experience in software engineering, with a focus on technical leadership and architecture
• Deep understanding of data engineering principles, including data modeling, data warehousing, reporting, and data governance
• Strong technical expertise in SQL, PostgreSQL, ETL/ELT pipelines, BI tools, and analytics platforms
• Practical experience with AI/LLM-adjacent and agentic AI data work — e.g., RAG ingestion pipelines, embedding generation, vector store management, or building/operating AI agent workflows over data — using AI coding assistants (Claude Code, Copilot, Cursor, or similar) as a regular part of the engineering workflow
• Working knowledge of modern cloud platforms such as AWS
• Experience with BI, reporting, dashboards, and customer-facing analytics
• Experience leading cross-functional initiatives with product, engineering, analytics, and business teams
Nice to Have
• Working knowledge of Ruby on Rails and React
• Experience with a semantic or metrics layer (e.g., dbt Semantic Layer, headless BI)
• Understanding of CI/CD, Git-based workflows, and infrastructure-as-code
The Stack Context
• Modern data stack: Fivetran, Airbyte, dbt, Snowflake, GitHub, Terraform, or similar tools
• Application context (nice to have): Ruby on Rails, React, or similar backend/frontend frameworks
• Data modeling: SQL, analytics models, documentation, testing, naming standards, and version control
• Infrastructure: cloud-based data infrastructure, infrastructure-as-code, CI/CD, monitoring, and cloud storage
• Data workflows: ingestion, transformation, orchestration, reporting, deployment, and change management
• Reporting focus: trusted metrics, scalable reporting models, dashboards, exports, and data quality
• AI surface: data pipelines and quality practices supporting AI/LLM and agentic AI use cases (RAG, embeddings, vector stores, AI agents) alongside traditional BI
Why This Role
This role offers a unique opportunity to drive technical direction and innovation at a rapidly growing company, while also mentoring and coaching engineers to improve their craft. As a Principal Data Engineer at Mitratech, you will have the chance to work on complex and challenging problems spanning trusted reporting and agentic AI systems, collaborate with cross-functional teams, and represent the company's technical capabilities externally. If you're looking for a role that offers a mix of technical leadership, data and reporting depth, agentic AI innovation, and collaboration, this could be the perfect fit for you.
We are an equal-opportunity employer that values diversity at all levels. All qualified applicants will receive consideration for employment without regard to race, color, religion, gender, national origin, age, sexual orientation, gender identity, disability, or veteran status.
- Design, build, and maintain scalable ETL/ELT pipelines for batch and real-time data ingestion and transformation.
- Develop and optimize data lake and data warehouse architectures (e.g., Snowflake, BigQuery, Redshift).
- Work with cloud platforms GCP, Azure to manage data infrastructure.
- GCP as mandatory skills
- Collaborate with analytics and product teams to understand data needs and deliver solutions.
- Ensure data quality, reliability, security, and compliance across all data systems.
- Mentor junior data engineers and contribute to best practices and code reviews.
- Monitor and troubleshoot data pipeline performance and resolve data-related issues.
- Automate data validation, monitoring, and alerting processes.
- 8+ years of experience in data engineering or software engineering with a data focus.
- Proficient in SQL and at least one programming language (e.g., Python, Scala, Java).
- Experience with modern data warehousing tools (e.g., Snowflake, Redshift, BigQuery).
- Strong understanding of data modeling, data lakes, and ETL/ELT design.
- Hands-on experience with orchestration tools like Airflow, dbt, or similar.
- Solid experience with cloud data platforms (AWS/GCP/Azure).
- Familiarity with CI/CD pipelines, containerization (Docker/Kubernetes), and version control (Git).
- Experience working in a DevOps or DataOps environment.
- Knowledge of data governance, lineage, and cataloging tools (e.g., Collibra, Alation).
- Familiarity with streaming technologies (Kafka, Spark Streaming, Flink).
- Experience supporting machine learning workflows and data science initiatives.
We are looking for a Data Engineer with at least 1 year of hands-on experience building solutions on Snowflake. The candidate should be comfortable designing, building, and managing reliable data pipelines that move data from multiple sources into a central data platform.
Responsibilities
- Build and maintain data pipelines for ingesting, transforming, and loading data into Snowflake
- Design scalable data models, schemas, tables, and views in Snowflake
- Develop ETL/ELT workflows using SQL, Python, or data orchestration tools
- Integrate data from APIs, databases, files, and third-party platforms
- Monitor pipeline performance, failures, data quality, and freshness
- Optimize Snowflake queries, warehouses, storage, and compute usage
- Implement incremental loads, change data capture, and scheduled workflows
- Work with engineering and business teams to understand data requirements
- Maintain documentation for pipelines, datasets, and data transformations
Requirements
- 1+ year of hands-on experience working with Snowflake
- Strong SQL skills and experience writing complex queries
- Experience building and managing ETL or ELT data pipelines
- Knowledge of data warehousing concepts, dimensional modelling, and data quality
- Experience with Python or another scripting language
- Familiarity with orchestration tools such as Airflow, Dagster, Prefect, dbt, or similar
- Understanding of APIs, relational databases, file formats, and cloud storage
- Ability to troubleshoot pipeline failures and performance issues
- Strong analytical, problem-solving, and communication skills
Good to Have
- Experience with dbt and Snowflake Tasks, Streams, Snowpipe, or Dynamic Tables
- Knowledge of AWS, Azure, or Google Cloud
- Experience with Kafka or other streaming platforms
- Familiarity with CI/CD, Git, monitoring, and data governance practices
- Experience integrating ERP, finance, or operational systems
Design, build, and maintain end-to-end data pipelines to ingest, process, and transform data from files, streams,
APIs, and relational/non-relational databases into Snowflake. Develop and optimize ELT/ETL pipelines using Snowflake SQL,
Snowpipe, Streams & Tasks, and cloud-native orchestration tools. Implement scalable data models and schemas (staging, curated, and consumption layers) to support analytics and reporting use cases. Develop transformations and business logic using SQL and Python, including Snowflake UDFs and stored procedures. Optimize Snowflake performance and cost through query tuning, warehouse sizing, clustering, and resource management. Integrate Snowflake with cloud storage and services across AWS and Azure (e.g., object storage, data integration, and mess
Location: Hyderabad / Chennai
Experience: 5+ years
Employment type: Full-time, permanent
Work Hours: General Shift
website: www.amazech.com
Qualifications:
- B.E./B.Tech/M.E./M.Tech in Computer Science, Information Technology, Data Science, or related disciplines.
- Strong academic background with relevant industry experience in Data Engineering and Data Warehousing.
Key Responsibilities:
· Design, develop, and maintain scalable data warehouse solutions using Snowflake.
· Write, optimize, troubleshoot, and enhance Snowflake SQL queries with a focus on performance and scalability.
· Develop and support ETL processes using Talend to ensure reliable and efficient data movement.
· Collaborate with business, analytics, and application teams to enable reporting, dashboards, metrics, and data exploration capabilities.
· Perform data analysis and resolve issues across data ingestion, transformation, and reporting pipelines.
· Debug and troubleshoot Python-based data processing scripts and automation workflows.
· Implement best practices for data quality, testing, deployment, and code reviews.
· Work across UI, API, and Data Warehouse layers to support end-to-end data integration and business requirements.
· Monitor, optimize, and maintain data warehouse performance and operational stability.
· Create and maintain technical documentation, data models, and process workflows.
Required Skills and Experience:
· Strong hands-on expertise in Snowflake Data Warehouse.
· Advanced SQL skills with experience handling large-scale datasets.
· Strong understanding of Data Warehousing concepts, dimensional modelling, and data architecture.
· Hands-on experience with Analytical SQL functions, query tuning, and performance optimization.
· Experience developing and maintaining ETL solutions using Talend.
· Proficiency in Python for scripting, debugging, automation, and data processing.
· Experience integrating UI, API, and Data Warehouse workflows.
· Strong problem-solving and analytical skills.
· Experience with testing, code reviews, and deployment best practices.
· Excellent communication and stakeholder management skills.
Data Engineer – Microsoft Fabric
Location: Pune, India
Work Mode: Hybrid
Experience: 6+ Years
Employment Type: Full-time contactor
Compensation: As per market standards, commensurate with experience and expertise
Shift Timings: 2:00 PM – 11:00 PM IST
Notice Period: 0 – 15 days
About the Role
Jade Business Services (JBS) is seeking a Data Engineer – Microsoft Fabric to join our Pune team and work on enterprise-scale data transformation and analytics initiatives.
We are looking for a hands-on Data Engineer with strong experience in Microsoft Fabric, SQL, Python/PySpark and modern data engineering practices. The candidate will be responsible for building scalable data pipelines, implementing Lakehouse and Warehouse solutions, developing data models and supporting governed, reliable and AI-ready data platforms.
The ideal candidate should be comfortable working with architects, engineering teams and client stakeholders to translate business requirements into scalable and production-ready data solutions.
Roles and Responsibilities
- Design and develop data solutions using Microsoft Fabric, including OneLake, Lakehouse, Warehouse and Data Factory pipelines.
- Build and maintain scalable ETL/ELT pipelines for batch and incremental data processing.
- Develop data ingestion and transformation pipelines using Fabric Data Factory, SQL, Python and/or PySpark.
- Implement Medallion Architecture using Bronze, Silver and Gold layers.
- Work with Lakehouse and Fabric Warehouse for enterprise data processing and analytics.
- Develop and maintain data models, tables, views and optimized SQL queries.
- Build and support semantic models for Power BI and analytical workloads.
- Implement data quality, validation, monitoring and error-handling mechanisms.
- Work with metadata, lineage and governance requirements using Microsoft Purview.
- Implement data security, access controls and role-based permissions across data platforms.
- Support Data Product and domain-oriented data architecture principles.
- Follow DataOps practices including CI/CD, deployment, monitoring and production support.
- Troubleshoot pipeline failures, performance issues and data quality problems.
- Optimize data pipelines, queries and storage for performance and cost efficiency.
- Work closely with Data Architects and business stakeholders to understand requirements and implement technical solutions.
- Participate in technical design discussions, code reviews and architecture reviews.
- Maintain technical documentation, data flow diagrams and pipeline documentation.
- Support production deployments, incident resolution and SLA-driven data platform operations.
- Identify opportunities for automation and AI-assisted improvements across data engineering processes.
Qualifications and Skills
- 6+ years of experience in Data Engineering, Data Integration or Data Platform development.
- Strong hands-on experience with Microsoft Fabric.
- Experience with:
- Microsoft Fabric Lakehouse
- Fabric Warehouse
- OneLake
- Fabric Data Factory / Pipelines
- Semantic Models
- Strong understanding of Lakehouse and Medallion Architecture.
- Strong SQL development and query optimization skills.
- Hands-on experience with Python and/or PySpark.
- Experience developing enterprise ETL/ELT and data integration pipelines.
- Experience with batch and incremental data processing.
- Understanding of data modelling concepts including dimensional modelling.
- Knowledge of data quality, metadata, lineage and data governance.
- Working knowledge of Microsoft Purview.
- Understanding of Data Mesh and Data Product concepts.
- Experience with CI/CD, version control, monitoring and DataOps practices.
- Understanding of cloud security, access controls and data privacy.
- Good troubleshooting and problem-solving skills.
- Strong communication skills and ability to work with distributed and client-facing teams.
Preferred Skills
- Microsoft Fabric or Azure Data certifications.
- Experience migrating workloads from Azure Synapse, SQL Server, Databricks or other data platforms to Microsoft Fabric.
- Experience implementing Medallion Architecture on Microsoft Fabric.
- Experience with Power BI and semantic modelling.
- Exposure to AI/ML, Generative AI or Agentic AI use cases on enterprise data platforms.
- Experience working with Data Products or domain-oriented data solutions.
- Experience in Energy & Utilities, Healthcare, Financial Services or Insurance.
- Experience working with US or international enterprise clients.
What We Expect
The ideal candidate should be hands-on first and capable of independently building, troubleshooting and optimizing Fabric data solutions. You should be able to explain the technical decisions behind your implementation and work effectively with architects and engineering teams to deliver production-ready solutions.

About the Role
You'll be at the forefront of designing and implementing robust data platform solutions that power advanced analytics, AI, and machine learning. Working with modern cloud technologies, you'll build scalable data foundations that enable clients to make smarter, data-driven decisions.
Key Responsibilities
- Build scalable data pipelines using Snowflake, AWS, GCP, and Databricks.
- Design and optimize data models for AI and machine learning workloads.
- Develop reliable data foundations for MLOps, governance, and data lineage.
- Integrate data from multiple sources into modern data platforms.
- Leverage Snowpark ML and Snowflake's native AI capabilities.
- Ensure data platforms are secure, scalable, and high-performing.
What We're Looking For
- 5+ years of hands-on experience with Snowflake.
- Strong proficiency in SQL and Python.
- Experience with AWS, Azure, or GCP.
- Knowledge of cloud storage services such as S3, ADLS, or GCS.
- Strong understanding of Dimensional Modeling and Data Vault.
- Experience with Scala or Java is a plus.
Tech Stack
- Data Warehouse: Snowflake
- Programming: SQL, Python, Scala (Good to Have), Java (Good to Have)
- Cloud: AWS, Azure, GCP
- Storage: S3, ADLS, GCS
- AI/ML: Snowpark ML, MLOps
Perks & Benefits
- Public Speaking & Communication Program
- Mentoring Program with Senior Support Leads
- 360° Progress Reviews
- Weekly Learning Sessions & Guilds
- Paid Certifications
- Hackathons & Innovation Days
- Recognition & Rewards Programs
- Team Socials & Annual Offsites
- Employee Assistance Program (24/7 Wellbeing Support)
The Data People Shaping Tomorrow
Our client helps organizations unlock the power of data through modern cloud, analytics, and AI solutions. We believe in creating an environment where talented technologists can learn, innovate, and make a real impact while building cutting-edge data platforms for global clients. If you're passionate about data engineering and want to work with the latest technologies in AI, cloud, and analytics, we'd love to hear from you.
Location – Hyderabad (Hybrid)
Work Experience – 5 to 7 years
CTC – upto 20 LPA
Roles & Responsibilities:
· We are looking for a Senior Data Engineering who will be majorly responsible for designing, building and maintaining ETL/ ELT pipelines.
· Integration of data from multiple sources or vendors to provide the holistic insights from data.
· You are expected to build and manage Data warehouse solutions, designing data models, creating ETL processes, implementing data quality mechanisms etc.
· Performs EDA (exploratory data analysis) required to troubleshoot data related issues and assist in the resolution of data issues.
· Should have experience in client interaction.
· Experience in mentoring juniors and providing required guidance.
Required Technical Skills
· Extensive hands on experience in Python, Pyspark, SQL, Dataiku.
· Strong experience in Data Warehouse, ETL, Data Modelling, building ETL Pipelines, Snowflake database.
· Working knowledge in Databricks, Redshift, ADF etc.
· Hands-on experience in cloud services like Azure, AWS- S3, Glue, Lambda, CloudWatch, Athena.
· Sound knowledge in end-to-end Data management, Data ops, quality and data governance.
· Familiar with SFDC, Waterfall/ Agile methodology.
· Strong domain knowledge in Pharma domain/ life sciences commercial data operations.
Qualifications
· Bachelor’s or master’s Engineering/ MCA or equivalent degree.
· 5-7 years of relevant industry experience as Data Engineer.
· Experience working on Pharma syndicated data such as IQVIA, Veeva, Symphony; Claims, CRM, Sales etc.
· High motivation, good work ethic, maturity, self-organized and personal initiative.
· Ability to work collaboratively and providing the support to the team.
· Excellent written and verbal communication skills.
· Strong analytical and problem-solving skills.
Job Summary
Role Overview
We are looking for an experienced Data Engineer with strong expertise in Python, ETL, Advanced SQL, CI/CD, DevOps, and Data Analytics. The ideal candidate should have hands-on experience designing and developing scalable data pipelines, transforming large datasets, and supporting data-driven applications.
Experience with Google Cloud Platform (GCP) will be an added advantage.
Key Responsibilities
- Design, develop, and maintain scalable ETL/ELT data pipelines using Python and SQL.
- Develop complex and optimized SQL queries, stored procedures, and data transformations.
- Build and maintain reliable data integration workflows across multiple data sources.
- Perform data cleansing, validation, transformation, and quality checks.
- Analyze data and provide insights to support business and technical requirements.
- Implement and maintain CI/CD pipelines for data engineering applications.
- Work with DevOps practices and tools to automate deployments, monitoring, and infrastructure processes.
- Troubleshoot data pipeline failures, performance issues, and production incidents.
- Optimize data processing workflows for performance, scalability, and reliability.
- Collaborate with Data Analysts, Data Scientists, Developers, and other stakeholders.
- Follow best practices for version control, testing, documentation, and deployment.
- Contribute to cloud-based data engineering initiatives, preferably on GCP.
Required Skills
- 5–7 years of hands-on experience in Data Engineering.
- Strong programming skills in Python.
- Strong expertise in Advanced SQL and database concepts.
- Hands-on experience with ETL/ELT processes and data pipelines.
- Good understanding of Data Warehousing and Data Modeling concepts.
- Experience with CI/CD practices and tools.
- Strong understanding of DevOps principles, automation, and deployment processes.
- Strong data analytics and problem-solving skills.
- Experience working with large datasets and performance optimization.
- Good understanding of Git/version control and software development best practices.
Good to Have
- Hands-on experience with Google Cloud Platform (GCP).
- Exposure to GCP data services such as BigQuery, Cloud Storage, Dataflow, Composer, or Pub/Sub.
- Experience with containerization/orchestration technologies such as Docker/Kubernetes.
- Experience with workflow orchestration tools such as Airflow.
- Knowledge of cloud-based data architecture and distributed data processing.
Preferred Candidate Profile
- Strong analytical and problem-solving abilities.
- Good communication and stakeholder management skills.
- Ability to work independently as well as in a collaborative team environment.
- Strong ownership of data pipelines and production systems.
- Candidates who can join at short notice are preferred.
Mandatory Skills
Data Engineer, Python , ETL, GCP, Advanced SQL, Strong Data Analytics skills, CICD, Devops






