Data Architect at Convosight · Remote only · 5 - 10 years · ₹10L - ₹40L / yr · Raised funding · Remote only · Posted 23 Jul 2026

Requirements:
- 5+ years in data architecture/data engineering, with at least 2+ years in an architect or lead capacity.
- Strong SQL: advanced query optimisation, indexing, partitioning strategies.
- Data modeling dimensional modelling (star/snowflake schema), normalization/denormalization tradeoffs, entity relationship design.
- Cloud data platforms: hands-on with AWS (Redshift, S3 Glue), Azure (Synapse, Data Factory), or GCP (BigQuery, Dataflow).
- Big data ecosystems: Spark, Hadoop, or Kafka for large-scale/streaming data.
- Data warehousing Snowflake, Redshift, BigQuery, or Databricks.
- ETL/ELT pipeline design: Airflow, dbt, Fivetran, or similar orchestration tools.
- Data governance & security: data lineage, access control, compliance (GDPR/SOC2), master data management.
Strongly Preferred:
- Experience architecting systems supporting ML/AI pipelines (feature stores, vector DBs, real-time inference data flows).
- Programming in Python or Scala for pipeline development.
- API/microservices architecture exposure, understanding how data systems integrate with application layers.
- Experience with data mesh/data lake house architectures.
- Prior experience presenting architecture decisions to leadership/stakeholders.
Nice-to-Have (Differentiators):
- Certifications: AWS/GCP/Azure data architecture certs.
- Experience in a high-growth startup (built systems from scratch, not just maintained legacy).
- Exposure to real-time/streaming architecture (Kafka, Kinesis, Flink).

About Convosight
About
Connect with the team
Similar jobs (10)
Data Architect – Databricks & AWS
- Strong experience in Data Architecture, Data Engineering, Databricks, Apache Spark/PySpark, Python, and Advanced SQL.
- Design and implement scalable ETL/ELT pipelines, data platforms, and Lakehouse architectures using Medallion Architecture.
- Experience with Databricks, Databricks Workflows, Unity Catalog OR Databricks Jobs
- Strong knowledge of Delta Lake, Databricks Workflows, Delta Live Tables (DLT), and dimensional data modeling.
- Hands-on experience with AWS services such as S3, Glue, IAM, Lambda, and CloudWatch.
- Experience with Apache Airflow, Data Warehouse concepts, Git/CI-CD, performance optimization, and data quality.
- Good to have exposure to Kafka/Structured Streaming, Unity Catalog, and modern data governance.
Position: Technical Architect – Data Engineering
Job Summary:
- We are looking for an experienced Technical Architect to lead the design and implementation of modern cloud-based data platforms.
- The ideal candidate should have strong expertise in Azure and/or AWS, Databricks, Snowflake, modern data architecture, and large-scale data engineering.
- The candidate will work closely with business stakeholders, architects, and engineering teams to deliver scalable, secure, and high-performance data solutions.
Key Responsibilities
• Design enterprise-scale data lakehouse and data warehouse architectures.
• Define data ingestion, transformation, and serving architecture.
• Lead architecture discussions and technical governance.
• Design scalable ETL/ELT frameworks using Databricks and Snowflake.
• Define best practices for security, performance, CI/CD, and DevOps.
• Guide engineering teams on implementation.
• Collaborate with business, product, and data governance teams.
Must Have Skills :
• Azure or AWS
• Databricks
• Snowflake
• Data Lakehouse Architecture
• PySpark
• SQL
• Data Modelling
• ETL/ELT Architecture
• Performance Optimization
• CI/CD
• Terraform or Infrastructure as Code (preferred)
Preferred
• Insurance domain
• dbt
• Unity Catalog
• Delta Lake
• Iceberg
• Azure Data Factory / AWS Glue
About Us:
The QX Impact was launched with a mission to make A.I accessible and affordable and deliver AI Products/Solutions at scale for the enterprises by bringing the power of Data, AI, and Engineering to drive digital transformation. We believe without insights; businesses will continue to face challenges to better understand their customers and even lose them. Secondly, without insights businesses won't’ be able to deliver differentiated products/services; and finally, without insights, businesses can’t achieve a new level of “Operational Excellence” is crucial to remain competitive, meeting rising customer expectations, expanding markets, and digitalization.
Job Summary:
We are looking for a Senior Data Engineer who is creative, collaborative, and adaptable to join our agile team of data scientists, engineers, and UX developers. The role focuses on building and maintaining robust data pipelines to support advanced analytics, data science, and BI solutions.
As a Senior Data Engineer, you will work with internal and external data, collaborate with data scientists, and contribute to the design, development, and deployment of innovative solutions.
Key Responsibilities:
- Design, develop, test, and maintain optimal data pipeline and ETL architectures.
- Map out data systems and define/design required integrations, ETL, BI, and AI systems/processes.
- Prepare and optimize data for predictive and prescriptive modeling.
- Collaborate with teams to integrate ERP data into the enterprise data lake, ensuring seamless flow and quality.
- Enhance cloud data infrastructure on AWS or Azure for scalability and performance.
- Utilize big data tools and frameworks to optimize data acquisition and preparation.
- Build architectures to move data to/from data lakes and data warehouses for advanced analytics.
- Develop and curate data models for analytics, dashboards, and reports.
- Conduct code reviews, maintain production-level code, and implement testing approaches.
- Monitor, troubleshoot, and resolve data ingestion workflows to maintain reliability and uptime.
- Drive innovation and implement efficient new approaches to data engineering tasks.
Must-Have Skills:
- Bachelor’s degree in Computer Science, Mathematics, Engineering, or a related field.
- 5+ years of experience working with enterprise data platforms, including building and managing data lakes.
- 3–5 years of experience designing and implementing data warehouse solutions.
- Expertise in SQL, including developing stored procedures (SP) and applying advanced data design concepts.
- Proficiency in Spark (Python/Scala) and Spark Streaming for real-time data pipelines.
- Experience with AWS or Azure services (e.g., AWS Glue, Azure Data Factory, Redshift, Snowflake).
- Familiarity with big data tools such as Apache Kafka, Apache Spark, or Flink.
- Hands-on experience with orchestration tools (e.g., Apache Airflow, Prefect).
- Knowledge of CI/CD processes, version control (e.g., Git, Jenkins), and deployment automation.
- Strong problem-solving, communication, and collaboration skills.
Good-to-Have Skills:
- Experience in integrating ERP data into data lakes.
- Experience with traditional ETL tools (e.g., Talend, Pentaho).
Competencies:
- Tech Savvy - Anticipating and adopting innovations in business-building digital and technology applications.
- Self-Development - Actively seeking new ways to grow and be challenged using both formal and informal development channels.
- Action Oriented - Taking on new opportunities and tough challenges with a sense of urgency, high energy, and enthusiasm.
- Customer Focus - Building strong customer relationships and delivering customer-centric solutions.
- Optimize Work Processes - Knowing the most effective and efficient processes to get things done, with a focus on continuous improvement.
Why Join Us?
- Be part of a collaborative and agile team driving cutting-edge AI and data engineering solutions.
- Work on impactful projects that make a difference across industries.
- Opportunities for professional growth and continuous learning.
- Competitive salary and benefits package.
Application Details
Ready to make an impact? Apply today and become part of the QX Impact team!
Solution Architect – AZURE Data Engineering
Job Overview
We are looking for an experienced Solution Architect – Data Engineering with strong expertise in designing data solutions and hands-on experience with Azure, Synapse, PySpark, Data Warehousing, and Data Lakes. The ideal candidate should have strong architectural and data engineering knowledge.
Key Responsibilities
- Design and implement scalable data architecture and solutions.
- Develop and manage Data Warehouse and Data Lake architectures.
- Design data platforms using Medallion Architecture.
- Lead Data Engineering and ETL activities.
- Work with Azure Synapse Analytics for data processing and analytics.
- Develop data solutions using PySpark / Apache Spark.
- Define and implement data validation and data quality processes.
- Collaborate with business, data, and technology teams to deliver effective data solutions.
Required Skills
- Strong experience in Solution Architecture / Data Architecture.
- Strong knowledge of Data Warehouse and Data Lake architecture.
- Good understanding of Medallion Architecture.
- Strong experience in Data Engineering and ETL.
- Hands-on experience with Microsoft Azure and Azure Synapse Analytics.
- Strong knowledge of PySpark / Apache Spark.
- Experience with Data Validation and Data Quality.
- Good communication and stakeholder management skills.
Experience
8+ Years
Data Engineer – Microsoft Fabric
Location: Pune, India
Work Mode: Hybrid
Experience: 6+ Years
Employment Type: Full-time contactor
Compensation: As per market standards, commensurate with experience and expertise
Shift Timings: 2:00 PM – 11:00 PM IST
Notice Period: 0 – 15 days
About the Role
Jade Business Services (JBS) is seeking a Data Engineer – Microsoft Fabric to join our Pune team and work on enterprise-scale data transformation and analytics initiatives.
We are looking for a hands-on Data Engineer with strong experience in Microsoft Fabric, SQL, Python/PySpark and modern data engineering practices. The candidate will be responsible for building scalable data pipelines, implementing Lakehouse and Warehouse solutions, developing data models and supporting governed, reliable and AI-ready data platforms.
The ideal candidate should be comfortable working with architects, engineering teams and client stakeholders to translate business requirements into scalable and production-ready data solutions.
Roles and Responsibilities
- Design and develop data solutions using Microsoft Fabric, including OneLake, Lakehouse, Warehouse and Data Factory pipelines.
- Build and maintain scalable ETL/ELT pipelines for batch and incremental data processing.
- Develop data ingestion and transformation pipelines using Fabric Data Factory, SQL, Python and/or PySpark.
- Implement Medallion Architecture using Bronze, Silver and Gold layers.
- Work with Lakehouse and Fabric Warehouse for enterprise data processing and analytics.
- Develop and maintain data models, tables, views and optimized SQL queries.
- Build and support semantic models for Power BI and analytical workloads.
- Implement data quality, validation, monitoring and error-handling mechanisms.
- Work with metadata, lineage and governance requirements using Microsoft Purview.
- Implement data security, access controls and role-based permissions across data platforms.
- Support Data Product and domain-oriented data architecture principles.
- Follow DataOps practices including CI/CD, deployment, monitoring and production support.
- Troubleshoot pipeline failures, performance issues and data quality problems.
- Optimize data pipelines, queries and storage for performance and cost efficiency.
- Work closely with Data Architects and business stakeholders to understand requirements and implement technical solutions.
- Participate in technical design discussions, code reviews and architecture reviews.
- Maintain technical documentation, data flow diagrams and pipeline documentation.
- Support production deployments, incident resolution and SLA-driven data platform operations.
- Identify opportunities for automation and AI-assisted improvements across data engineering processes.
Qualifications and Skills
- 6+ years of experience in Data Engineering, Data Integration or Data Platform development.
- Strong hands-on experience with Microsoft Fabric.
- Experience with:
- Microsoft Fabric Lakehouse
- Fabric Warehouse
- OneLake
- Fabric Data Factory / Pipelines
- Semantic Models
- Strong understanding of Lakehouse and Medallion Architecture.
- Strong SQL development and query optimization skills.
- Hands-on experience with Python and/or PySpark.
- Experience developing enterprise ETL/ELT and data integration pipelines.
- Experience with batch and incremental data processing.
- Understanding of data modelling concepts including dimensional modelling.
- Knowledge of data quality, metadata, lineage and data governance.
- Working knowledge of Microsoft Purview.
- Understanding of Data Mesh and Data Product concepts.
- Experience with CI/CD, version control, monitoring and DataOps practices.
- Understanding of cloud security, access controls and data privacy.
- Good troubleshooting and problem-solving skills.
- Strong communication skills and ability to work with distributed and client-facing teams.
Preferred Skills
- Microsoft Fabric or Azure Data certifications.
- Experience migrating workloads from Azure Synapse, SQL Server, Databricks or other data platforms to Microsoft Fabric.
- Experience implementing Medallion Architecture on Microsoft Fabric.
- Experience with Power BI and semantic modelling.
- Exposure to AI/ML, Generative AI or Agentic AI use cases on enterprise data platforms.
- Experience working with Data Products or domain-oriented data solutions.
- Experience in Energy & Utilities, Healthcare, Financial Services or Insurance.
- Experience working with US or international enterprise clients.
What We Expect
The ideal candidate should be hands-on first and capable of independently building, troubleshooting and optimizing Fabric data solutions. You should be able to explain the technical decisions behind your implementation and work effectively with architects and engineering teams to deliver production-ready solutions.
The recruiter has not been active on this job recently. You may apply but please expect a delayed response.
Dear Candidate,
Greeting from NAM Info Pvt Ltd.
We have a role for Data Engineer position with NAM Info.
This role will be permanent with NAM info and deploy to client
location NEW DELHI~CHENNAI~HYDERABAD~PUNE~KOLKATA.
Work Mode: WORK FROM OFFICE
A decent hike can be provided based on current CTC
Interview Mode: Virtual
Role Descriptions:
Exp Range: 7 - 10 years
City Locations: NEW DELHI~CHENNAI~HYDERABAD~PUNE~KOLKATA
Key Responsibilities*
Role: Data Engineer
Location: ~NEW DELHI~CHENNAI~HYDERABAD~PUNE~KOLKATA
Skills: Digital: Databricks, Azure Data Factory
Experience Required: 8-10
Descriptions:
Good information and sound knowledge in Azure Synapse Analytics Azure Data Factory (ADF)Big Data technologies and data processing frameworks Azure Data Warehouse and associated Azure data platform services Data integration| data modelling| and performance optimization
Desire candidate
- Candidate should have valid PF.
Regards,
NAM Info
About the Role
You'll be at the forefront of designing and implementing robust data platform solutions that power advanced analytics, AI, and machine learning. Working with modern cloud technologies, you'll build scalable data foundations that enable clients to make smarter, data-driven decisions.
Key Responsibilities
- Build scalable data pipelines using Snowflake, AWS, GCP, and Databricks.
- Design and optimize data models for AI and machine learning workloads.
- Develop reliable data foundations for MLOps, governance, and data lineage.
- Integrate data from multiple sources into modern data platforms.
- Leverage Snowpark ML and Snowflake's native AI capabilities.
- Ensure data platforms are secure, scalable, and high-performing.
What We're Looking For
- 5+ years of hands-on experience with Snowflake.
- Strong proficiency in SQL and Python.
- Experience with AWS, Azure, or GCP.
- Knowledge of cloud storage services such as S3, ADLS, or GCS.
- Strong understanding of Dimensional Modeling and Data Vault.
- Experience with Scala or Java is a plus.
Tech Stack
- Data Warehouse: Snowflake
- Programming: SQL, Python, Scala (Good to Have), Java (Good to Have)
- Cloud: AWS, Azure, GCP
- Storage: S3, ADLS, GCS
- AI/ML: Snowpark ML, MLOps
Perks & Benefits
- Public Speaking & Communication Program
- Mentoring Program with Senior Support Leads
- 360° Progress Reviews
- Weekly Learning Sessions & Guilds
- Paid Certifications
- Hackathons & Innovation Days
- Recognition & Rewards Programs
- Team Socials & Annual Offsites
- Employee Assistance Program (24/7 Wellbeing Support)
The Data People Shaping Tomorrow
Our client helps organizations unlock the power of data through modern cloud, analytics, and AI solutions. We believe in creating an environment where talented technologists can learn, innovate, and make a real impact while building cutting-edge data platforms for global clients. If you're passionate about data engineering and want to work with the latest technologies in AI, cloud, and analytics, we'd love to hear from you.
Job Description:
We are seeking a skilled Senior Data Engineer with expertise in Databricks to join our dynamic data team. The ideal candidate will design, build, and maintain scalable data pipelines and architectures to support our organization's data-driven initiatives and leverage Databricks to process large-scale datasets, optimize data workflows, enable advanced analytics and machine learning, and integrate Power BI for data visualization and reporting.
Key Responsibilities
· 5-8 years of professional work experience in a relevant field
· Proficient in Microsoft Fabric platform, Azure Databricks, ADF, Delta Lake, SQL Data Warehouse, Unity Catalog.
· Good Experience on Microsoft Dynamics 365
· Experience/ prior knowledge on semi structure data and Structured Streaming, Azure synapse, data lake, data warehouse.
· Proficient in creating Azure Data Factory pipelines for ETL/ELT processing; copy activity, custom Azure development etc.
· Good knowledge of SQL and Python for data manipulation, transformation, and analysis
· Understand business requirements to set functional specifications for reporting applications
- Data Pipeline Development: Design, develop, and maintain robust, scalable data pipelines using Databricks, Apache Spark, and other cloud-based technologies.
- Data Integration: Ingest, transform, and integrate data from diverse sources, including APIs, databases, streaming platforms, and third-party systems, into Databricks for analytics and reporting.
- Power BI Integration: Develop and optimize data models and datasets in Databricks for use in Power BI, ensuring efficient data connections and high-quality visualizations.
- Performance Optimization: Optimize data workflows, API calls, and queries on Databricks and Power BI for performance, cost-efficiency, and scalability.
- Data Modeling: Build and maintain data models to support business requirements, ensuring data quality, consistency, and accessibility for analytics and reporting.
- Cloud Integration: Implement data solutions on cloud platforms (e.g., AWS, Azure, GCP) integrated with Databricks, Power BI, and API ecosystems.
- Security & Compliance: Implement data governance, security, and compliance best practices within Databricks, Power BI, and API environments.
- Technical Skills:
- Proficiency in Databricks, including Delta Lake, Spark SQL.
- Strong programming skills in Python, Scala, or Java.
- Experience with Apache Spark for big data processing.
- Knowledge of SQL for querying and transforming data.
- Proficiency in Power BI for creating data models, DAX queries, and interactive dashboards.
Preferred Qualifications
- Databricks certification (e.g., Databricks Certified Data Engineer Associate/Professional).
Experience with real-time data processing and streaming
Job Description:
Experience: 10+ Years
Job Summary
We are looking for an experienced Azure Fabric Data Architect to lead the design and implementation of an enterprise data platform on Microsoft Fabric. The role involves architecting scalable data solutions, defining data governance, and enabling AI-driven analytics for a global financial services client.
Key Responsibilities
- Design end-to-end data architecture using Microsoft Fabric.
- Build enterprise Lakehouse, Data Warehouse, and OneLake solutions.
- Define data ingestion, ETL/ELT, governance, security, and performance strategies.
- Lead architecture for AI-powered analytics, AI Agents, and enterprise chatbots using Azure AI services.
- Work with business stakeholders to translate requirements into technical solutions.
- Mentor engineering teams and provide technical leadership.
Required Skills
- Microsoft Fabric (Data Factory, Lakehouse, Data Warehouse, OneLake)
- Azure Data Engineering
- Power BI
- Azure AI Services / Azure OpenAI
- Data Architecture & Data Modeling
- SQL, Python
- Azure DevOps, CI/CD
- Strong stakeholder management and solution design experience
Preferred: Experience in Capital Markets or Financial Services and Microsoft Azure/Fabric certifications.
NOTE: One technical round is mandatory to be taken F2F from office.
Job Title : Tech Lead – Data Lake Platform
Number of Positions : 2
Experience : 7+ Years
Role Type : Technical Lead / Data Platform Lead
Domain : Data Engineering / Data Platform / AWS
About the Role :
We are looking for an experienced Tech Lead – Data Lake Platform to lead the design, development, and operations of an enterprise-scale AWS-based Data Lake Platform.
The platform will ingest data from multiple business systems, process it through structured data layers, and serve data for analytics, reporting, APIs, and operational applications.
As a Tech Lead, you will be responsible for setting the technical direction, leading data engineering teams, driving platform reliability and performance, and owning the platform's delivery, governance, and production support end to end.
Core Tech Stack :
AWS | S3 | EMR | Glue | Athena | Redshift | DMS | Lambda | RDS | IAM | Airflow | Spark/PySpark | SQL | Data Modeling | Hasura | GraphQL | DBT | PostgreSQL/Aurora | Kafka
Key Responsibilities :
- Own the overall Data Lake Platform architecture, covering data ingestion, staging, curated layers, consumption, analytics, and API/data serving.
- Lead the design and development of production-grade data pipelines using AWS Glue, EMR/Spark, Airflow, DBT, Athena, and Redshift.
- Design scalable batch and analytics pipelines with a focus on reliability, performance, data quality, and maintainability.
- Own the API/data-serving architecture from consumption data → RDS/PostgreSQL → Hasura GraphQL → Lambda/API Gateway.
- Drive improvements in platform stability, including orchestration failures, cluster sizing, pipeline SLAs, query performance, and production reliability.
- Design and implement appropriate AWS security, access control, IAM, PII handling, and data governance practices.
- Lead technical discussions, architecture decisions, code reviews, and engineering best practices.
- Mentor and guide data engineers while ensuring high-quality and scalable engineering delivery.
- Own production support, incident management, troubleshooting, and root-cause analysis (RCA) for critical data platform issues.
- Develop and maintain runbooks, operational procedures, monitoring, and incident response practices.
- Collaborate with business, product, application, and analytics teams to onboard new datasets and support reporting, API, and data consumption requirements.
- Ensure the platform meets defined availability, performance, security, data quality, and compliance requirements.
Must-Have Skills :
- 7+ years of experience in Data Engineering, Data Platform Engineering, or related areas, including experience in technical leadership or architecture.
- Strong hands-on experience with AWS Data Services, including:
- Amazon S3
- AWS EMR
- AWS Glue
- Amazon Athena
- Amazon Redshift
- AWS DMS
- AWS Lambda
- Amazon RDS
- AWS IAM
- Strong production experience with Apache Airflow.
- Strong hands-on experience with Apache Spark / PySpark.
- Strong SQL skills and experience with data modeling, including layered data architecture, data marts, and enterprise data models.
- Experience with Hasura or a similar GraphQL layer for PostgreSQL-based data/API serving.
- Strong understanding of data lake architecture and enterprise data platforms.
- Proven experience leading engineers and driving technical decisions.
- Hands-on experience with production support, troubleshooting, incident management, and RCA.
- Strong understanding of data platform performance, scalability, reliability, and SLA management.
Good-to-Have Skills :
- Experience with DBT and modern data transformation practices.
- Experience with lakehouse table formats on Amazon S3, such as Apache Iceberg or similar technologies.
- Strong knowledge of Amazon Redshift workload optimization, including :
- Distribution keys
- Sort keys
- Spectrum
- External tables
- Experience with Kafka or other streaming/data ingestion technologies.
- Experience with PostgreSQL / Amazon Aurora as a data-serving layer.
- Experience with Lambda and API Gateway for API-based data serving.
- Experience with enterprise data governance, data quality, security, and compliance.
- Experience in BFSI / Banking / Financial Services / Insurance domain.
Ideal Candidate :
The ideal candidate is a hands-on Data Platform / Data Engineering Lead who can operate at both the architecture and implementation level. You should be comfortable designing an AWS Data Lake from end to end, leading engineers, troubleshooting production issues, and working closely with business and application teams to deliver reliable data products.












