AI Architect at The industry’s only Manufacturing Operating System · Hyderabad · 10 - 15 years · Posted 30 Sep 2026

We’re on hunt for AI Architect
Responsibilities:
- 10–15+ years overall experience, with recent hands-on AI/GenAI architecture ownership.
- Must have architected enterprise AI platforms/solutions end-to-end, not just individual ML models or PoCs.
- Strong GenAI/LLM production experience: RAG, embeddings, vector DBs, hybrid search, reranking, evaluation, guardrails.
- Strong Agentic AI understanding: agents, tool calling, workflows, orchestration, human-in-the-loop.
- Experience taking AI solutions from architecture → production → scale, ideally across multiple business teams/use cases.
- Strong cloud architecture — Azure/AWS preferred; hybrid/on-prem experience is a plus.
- Must understand enterprise security, governance, Responsible AI, observability and LLMOps/MLOps.
- Should be able to articulate build-vs-buy, MVP-vs-target architecture, cost/performance/security tradeoffs.
- Strong stakeholder-facing / consulting ability — can work with business leaders, engineering, security and data teams and influence without authority.
There is scope to move to the US for this role if you are aligned for the same, else this will be a WFO role from Hyderabad location

Similar jobs (10)
Senior Generative AI Engineer
Employment Type: Permanent with VDart Digital
Work Location: Marathalli, Bengaluru
Job Description
We are seeking a highly skilled Senior Generative AI Engineer with strong expertise in designing, developing, and deploying enterprise-scale AI solutions using Large Language Models (LLMs) and modern Generative AI frameworks. The ideal candidate should have hands-on production experience building scalable GenAI applications, AI agents, autonomous workflows, and Retrieval-Augmented Generation (RAG) systems in cloud-native environments.
This role requires deep technical expertise in LLM orchestration, AI application architecture, prompt engineering, vector databases, MLOps, and production deployment of AI systems. Candidates should have proven experience delivering real-world AI solutions in enterprise environments with strong exposure to cloud platforms and DevOps practices.
Key Responsibilities
- Design, build, and deploy enterprise-grade Generative AI applications using Large Language Models (LLMs).
- Develop intelligent AI agents and autonomous workflows using frameworks such as LangChain, CrewAI, LangGraph, AutoGen, or similar agentic AI frameworks.
- Implement and optimize Retrieval-Augmented Generation (RAG) pipelines using vector databases and semantic search technologies.
- Work extensively on prompt engineering, tool calling, memory management, agent orchestration, and multi-agent systems.
- Integrate and manage LLMs such as OpenAI, Azure OpenAI, Claude, Gemini, Llama, Mistral, or similar foundation models.
- Develop scalable AI services and APIs using Python and FastAPI.
- Build production-ready AI solutions with high availability, scalability, monitoring, and observability.
- Deploy and manage AI applications in cloud-native environments using Docker and Kubernetes.
- Collaborate with Data Science, ML Engineering, and DevOps teams to operationalize AI solutions.
- Implement CI/CD pipelines and automated deployment processes for AI workloads.
- Monitor model performance, latency, reliability, and operational efficiency in production environments.
- Ensure AI solutions follow enterprise security, governance, and responsible AI standards.
- Evaluate and adopt emerging Generative AI tools, frameworks, and models.
Required Skills
Generative AI & LLM Expertise
- Strong hands-on experience with Generative AI and Large Language Models (LLMs).
- Production-level experience building and deploying GenAI applications.
- Expertise in LangChain, CrewAI, LangGraph, AutoGen, or similar frameworks.
- Experience with AI agents, autonomous workflows, and multi-agent architectures.
- Strong understanding of prompt engineering, embeddings, model evaluation, and LLM orchestration.
- Experience integrating OpenAI, Azure OpenAI, Claude, Gemini, Llama, Mistral, or similar models.
RAG & Vector Databases
- Strong experience implementing RAG pipelines and semantic retrieval systems.
- Experience with vector databases such as Pinecone, Weaviate, ChromaDB, FAISS, or Milvus.
- Understanding of chunking strategies, embeddings, indexing, reranking, and retrieval optimization.
Python & AI Development
- Strong proficiency in Python.
- Experience with FastAPI for AI service and API development.
- Experience with AI/ML libraries and data processing tools such as Pandas and NumPy.
Cloud & Production Deployment
- Mandatory production experience on at least one cloud platform:
- Microsoft Azure
- Experience deploying scalable AI applications in enterprise production environments.
- Hands-on experience with Docker, Kubernetes, Jenkins, Terraform, and CI/CD pipelines.
- Strong understanding of MLOps, AI deployment lifecycle, monitoring, and observability.
Engineering & Operational Excellence
- Strong understanding of software engineering best practices.
- Experience with Git, version control, automated testing, and release management.
- Experience building secure, scalable, and high-performance AI solutions.
- Ability to troubleshoot production AI systems and optimize performance.
Preferred Skills
- Experience with AI observability and evaluation frameworks.
- Exposure to fine-tuning, PEFT, LoRA, or model optimization techniques.
- Experience with enterprise AI governance and responsible AI practices.
- Knowledge of distributed AI systems and scalable inference architectures.
- Familiarity with AI security and compliance standards.
Qualifications
- Bachelor’s or Master’s degree in Computer Science, Artificial Intelligence, Data Science, or a related field.
- 3–8 years of overall software engineering experience.
- Minimum 3+ years of hands-on experience in Generative AI and LLM-based application development,
- Proven track record of delivering enterprise-scale AI solutions in production environments.
- Strong communication and stakeholder management skills.
AI Developer
Primary Skill-set (Must have)
- Generative AI Expertise: 2-3 years of experience in designing and implementing generative AI solutions, including knowledge of various generative and autoregressive models. Ability to apply generative AI techniques to diverse use cases such as image generation, text generation, and creative content synthesis.
• 2 years of experience in prompt engineering, fine tuning, agentic framework, GenAI SDK’s
• 1-2 years of experience in Agentic AI frameworks like Autogen, Lanngraph, MS Agent SDK, A2A, MCP, A2P, memory concepts, multi agent orchestration
• 7+ years of experience in Python
• 5+ years of experience in software development
• Azure Proficiency: 3-5 years of experience with Azure cloud services relevant to AI, including Azure Machine Learning, Azure Cognitive Services, Azure Databricks, and Azure Kubernetes Service (AKS). 2+ years of experience in Azure's capabilities to architect end-to-end AI solutions and optimize performance.
• Architecture Design: 3-5 years of skills with the ability to design scalable, reliable, and cost-effective architectures for AI solutions. Proficiency in designing distributed systems, microservices architectures, and containerized solutions using technologies such as Docker and Kubernetes.
Secondary Skills (Good to have)
• Security and Compliance: Understanding of security principles and best practices in AI development, with the ability to implement security controls, encryption mechanisms, and access management policies to protect AI models and sensitive data.
• Integration and Deployment: Proficiency in implementing CI/CD pipelines, automation scripts, and infrastructure as code (IaC) using tools such as Azure DevOps, Terraform, or Ansible. Experience in containerization and orchestration of AI workloads using Docker and Kubernetes.
• Software Development: Strong programming skills in languages such as Python, with experience in developing AI applications, RESTful APIs, and microservices architectures. Familiarity with software development methodologies such as Agile or Scrum.
• Communication and Presentation: Excellent communication skills with the ability to convey complex technical concepts to non-technical stakeholders. Experience in preparing and delivering technical presentations, architecture diagrams, and documentation to communicate architectural decisions and design rationale effectively.
Job Summary:
Wissen Technology is hiring an AI Implementation Engineer to build, deploy, and scale enterprise-grade Generative AI solutions across business-critical applications. The role involves developing production-ready AI systems using Azure AI services, Large Language Models (LLMs), RAG architectures, and agent-based frameworks while collaborating closely with engineering teams to drive AI adoption and innovation.
Experience
6-12 years
Location
Mumbai / Bangalore
Mode of Work
Hybrid
Mandatory Skills (Must Have)
• Python programming (6+ years) including asynchronous programming and backend application development
• Java and Spring Framework (3+ years) for enterprise-scale application integration
• Azure OpenAI Service, Azure AI Foundry, and Azure AI Search for production GenAI applications
• Retrieval Augmented Generation (RAG) architecture including embeddings, chunking, vector databases, reranking, and grounding techniques
• Agent Frameworks such as Microsoft Agent Framework, Semantic Kernel, AutoGen, LangChain, or LangGraph
• Snowflake and Cortex AI (Cortex Search, LLM Functions) with strong SQL expertise
• Prompt Engineering, LLM evaluation frameworks, testing, and model performance optimization
• DevOps and Cloud Deployment using Azure DevOps, GitHub Actions, Docker, AKS, Azure Functions, and observability tools
Optional Skills (Good to Have)
• React.js for AI-powered user interfaces and conversational applications
• Azure AI Content Safety and Responsible AI implementation experience
• Financial Services, Banking, or other regulated industry domain experience
• Real-time streaming applications and token-level LLM operations
• Performance optimization, caching strategies, and cost optimization for AI workloads
• Microsoft Azure AI Engineer Associate Certification
Location: Hyderabad, India. Based at the KnackLabs headquarters, with occasional travel to client locations for workshops and reviews. This role does not involve extended onsite deployments.
About the Role
You will work as an AI Architect who designs the systems behind our client engagements: AI agents, RAG systems, automation platforms, and the conventional backend systems around them.
This is a hands-on design role, not a slideware role. You will scope architectures with clients, make the hard technical decisions, defend them in review, and stay accountable for how the systems perform in production.
You will work directly with clients. Everyone at KnackLabs does. You will sit in design discussions with client engineering teams, present architecture decisions to technical and business stakeholders, and answer for the choices you make.
A full KnackLabs engineering team in Hyderabad builds with you. You own the technical design and the quality of what ships.
What you'll own
- Architecture - Design AI agents, RAG systems, integrations, and the scalable backend systems around them, for multiple client engagements.
- Technical scoping - Work directly with clients to turn a business problem into a system design, with clear trade-offs and clear reasons.
- Scale and reliability - Make sure what we build handles real load: data stores, queues, caching, horizontal scaling, and fault tolerance.
- Design reviews - Review designs and builds across engagements. Set the technical bar and hold it.
- Evaluation strategy - Define how we measure accuracy, safety, latency, and cost for the AI systems we ship.
- Guiding engineers - Raise the level of the engineers building with you, through reviews and direct pairing.
- Feedback to the platform - Feed what you learn across engagements back into our platform and internal tools.
What we are looking for
- Around 7 or more years of software engineering experience, including direct work with customers on design or delivery.
- Full-stack development experience with strength in backend technologies.
- Experience designing and building scalable applications. You understand how large-scale distributed systems work: data partitioning, queues, caching, horizontal scaling, and fault tolerance.
- At least 2 years of strong, hands-on AI experience with large language models in production.
- You build with AI coding tools like Claude Code or Codex as your default way of working. You understand Claude Skills, have written skills yourself, use them actively, and have contributed to them.
- Hands-on experience building retrieval-augmented generation (RAG) systems: chunking, embeddings, vector databases, retrieval, and reranking.
- Hands-on experience building AI agents.
- Strong programming skills in Python. Working knowledge of TypeScript or JavaScript.
- Experience with at least one cloud platform (AWS, Azure, or GCP).
- Clear communication. You can explain an architecture decision to an engineer and to a business leader, and defend it under questioning.
- High ownership and comfort with ambiguity. You can take an unclear problem and turn it into a design.
Nice to have
- Experience building evaluations to measure accuracy, safety, latency, and cost.
- Experience with observability and tracing tools such as LangSmith or Braintrust.
- Experience with on-premises or private cloud (VPC) deployments.
- Experience deploying AI systems in regulated industries such as insurance, banking, or the public sector.
- Experience with data engineering and pipelines.
- A history of side projects, open source contributions, or products you shipped end-to-end.
- Experience working at a consulting or professional services firm in a client-facing delivery role.
Stack and tools
- Languages: Python and TypeScript.
- Models: Claude and other frontier or open-source models, chosen to fit the customer.
- AI patterns: RAG, agents, prompt engineering, skills, and evaluations.
- Vector and retrieval: vector databases and retrieval pipelines.
- Cloud: AWS, Azure, or GCP, on public or private cloud.
- Integration: REST APIs and enterprise system connectors.
Location: Pune / Gurgaon
Position: AI Engineer
work mode: WFO
Job Description.
Job responsibilities:
- Responsibility for design, implementation and deployment of Generative AI, Agentic frameworks at scale
- Strong in programming - Python a
- Previous experience of working on Computer Vision projects and VLM /VLAM models.
- In depth awareness of Transformer architectures and End to End Deep neural networks
- Full stack AI / ML development experience
- Design, build & maintain efficient and reliable Agentic / Generative AI code leveraging pipelines
- Hosting and deployment knowledge in GCP or AWS or Azure along with advanced engineering concepts to build user friendly UI interface for easy adoption.
Requirements:
· 4 to 8 years overall years of experience (Agentic AI, Generative AI, VLM, VLAM and LLM) with significant exposure in Development, Architecture design, scaling and hosting in cloud.
Must Have –
· Architecting and solutioning experience with Python and FAST API, Agentic Ai frameworks, VLMs, VLAMs, Open source LLM’s and Code based LLM models at scale with - Langchain / Ollama, embeddings, Memory Management etc.,
· Practical experience in implementing Explainable and ethical AI models Practical experience in implementing frameworks like RAG/ CAG/ Self-reflective RAG etc.,
· Experience in cloud hosting either AWS or Azure or GCP.
· Experience in ML-OPS - Implement a feedback mechanism to continually improve the model over time through feedback loop and monitoring KPI’s in production.
· Experience with Quantization and Kubernetes or docker
Good to have
· gRPC implementation to expose the API’s on a server for easy usage and good user interface
· Streamlit front end creation
· Experience with SAFe framework deliveries.
Key Responsibilities:
· Architectural Leadership: Design and lead the development of robust, scalable AI architectures, ensuring high performance, reliability, and security.
· Applied Mathematics & Statistics: Apply statistical analysis, numerical computation, and mathematical modeling to derive insights from large-scale data and optimize model performance.
· Deep Learning Development: Design, train, and deploy advanced Deep Learning (DL) models.
· Technical Mentorship: Mentor engineering teams on best practices for AI/ML, coding standards, and architectural design.
· Model Optimization: Optimize models for speed, efficiency, and accuracy using techniques like pruning, quantization, or GPU acceleration.
· Strategy & Innovation: Evaluate and select appropriate AI frameworks, tools, and platforms, staying abreast of cutting-edge research and industry trends.
Qualifications:
Required:
· Education: Master's or PhD in Computer Science, Applied Mathematics, Statistics, Physics, or a related quantitative field.
· Experience: 10+ years of experience in software development, with at least 3-5 years in a Applied Mathematics and Deep learning.
· AI/ML Expertise: Proven experience designing and deploying deep learning models in production using frameworks.
· Mathematics/Statistics: Strong proficiency in linear algebra, calculus, probability, and statistical methods.
· Programming Skills: Expert-level coding skills in Python (NumPy, Pandas, Scikit-learn) and experience with languages like Java or C++.
Key Competencies:
- Strategic mindset with deep operational awareness.
- Excellent communication and stakeholder management skills.
- Ability to simplify complex technical concepts for executive reporting.
- Strong leadership, people development, and cross-functional influencing skills.
Bias for action and a relentless focus on continuous improvement.
Hiring for AI Engineer
Exp: 5 - 10 yrs
Edu : BE/B.Tech/MCA
Work Location : Pune / Mumbai
Skill Set:
Total experience ranging from 5–10 years in software engineering/AI roles
Min 5 years strong programming experience in Python is a MUST
Min 3.5 years hands-on experience in AI with LLMs, RAG pipelines, and AI frameworks
2+ years shipping LLM systems in production
Experience with cloud platforms (AWS/Azure/GCP)
Role: AI Developer
Experience: 3–4 Years
Employment Type: Full-Time
Location: Goregaon, Mumbai
About the Role
We are looking for an experienced AI Developer with 3–4 years of software development experience and strong hands-on exposure to Generative AI, AI Agents, Copilots, and AI-powered application development.
The candidate will be responsible for building production-ready AI solutions, developing agentic workflows, modernizing legacy applications, and integrating LLM capabilities into enterprise applications.
Key Responsibilities
- Design, develop, and deploy AI Agents and agentic workflows for enterprise use cases.
- Build AI Copilots and LLM-powered applications using modern AI frameworks and APIs.
- Develop RAG-based applications using embeddings, vector databases, and enterprise data.
- Work on legacy application migration and modernization, leveraging AI-assisted development and code transformation techniques.
- Analyze legacy codebases and design strategies for AI-driven migration, refactoring, and modernization.
- Integrate LLMs with enterprise applications, APIs, databases, and third-party systems.
- Implement tool calling, function calling, multi-agent workflows, and workflow automation.
- Perform prompt engineering, context optimization, model evaluation, and AI application testing.
- Take ownership of AI solutions from POC and prototyping through production deployment.
- Collaborate with product managers, architects, and engineering teams to convert business requirements into scalable AI solutions.
- Stay updated with emerging technologies in Generative AI, Agentic AI, LLMs, and AI-assisted software development.
Required Skills
- 3–4 years of professional software development experience.
- Strong proficiency in Python and/or JavaScript/TypeScript.
- Hands-on experience developing Generative AI / LLM-based applications.
- Strong understanding of AI Agents, RAG, Prompt Engineering, LLM APIs, and embeddings.
- Experience with frameworks such as LangChain, LangGraph, Semantic Kernel, AutoGen, or equivalent.
- Experience working with REST APIs, databases, Git, and cloud environments.
- Hands-on experience with vector databases such as Pinecone, Weaviate, Chroma, FAISS, or equivalent.
- Good understanding of software architecture, debugging, testing, and deployment practices.
Good to Have
- Experience with Microsoft Copilot / Copilot Studio.
- Experience working with Claude, OpenAI, Gemini, Azure OpenAI, or open-source LLMs.
- Experience in legacy application migration, modernization, or code conversion.
- Knowledge of Azure AI / AWS / Google Cloud AI services.
- Experience with MCP, multi-agent systems, tool calling, and AI orchestration.
- Experience building enterprise-grade AI solutions with focus on security, scalability, and performance.
PRINCIPAL AI ENGINEER @ METADOME.AI
Company Description
Metadome.ai builds frontier AI models that transform text, drawings, and CAD into production-ready, interactive 3D experiences. The company advances a full generative pipeline—text-to-CAD, 2D-to-3D
reconstruction, CAD completion and harmonization, and real-time interactive rendering—engineered for the precision required in the physical world. Its technology currently powers the modernization of
OEM aftersales for more than 30 automotive and heavy-equipment manufacturers worldwide, delivering accurate, scalable, and fast 3D solutions. Metadome.ai’s platform enables shoppable 3D parts, step-by-step repair animations, and a headless API that feeds consistent 3D assets into commerce, dealer, training, and service systems. The broader mission is to allow anyone to move from an idea, drawing, or specification to a production-grade 3D model and beyond in seconds.
Role Description
As a Principal AI Engineer — Generative CAD & 3D, you will lead the design, development, and deployment of advanced AI models that convert text, 2D drawings, and CAD files into engineering-grade 3D content. You will architect end-to-end generative pipelines, including
text-to-CAD, 2D-to-3D reconstruction, CAD completion, and real-time rendering, collaborating closely with product, design, and engineering teams to ship robust production systems. Day-to-day, you will experiment with novel neural network architectures, optimize model performance on large-scale CAD datasets, write high-quality production code, and guide the integration of AI services into customer-facing platforms. You will mentor other engineers, establish best practices for AI development, and contribute to technical strategy and roadmap. This is a full-time, hybrid role based in Bengaluru, with a mix of on-site collaboration and work-from-home flexibility.
Qualifications
- Strong foundation in Computer Science and Software Development, including data structures, algorithms, system design, and production-grade coding in languages such as Python, C++, or similar.
- Deep expertise in Neural Networks and Pattern Recognition, with hands-on experience designing, training, and deploying modern deep learning architectures for complex, high-dimensional data.
- Experience with Natural Language Processing (NLP), including working with text encoders, multimodal models, and integrating language understanding into generative workflows.
- Advanced degree (Master’s or PhD) in Computer Science, Electrical Engineering, Applied Mathematics, or a related field, or equivalent practical experience in AI/ML research and engineering.
- Background in 3D geometry, CAD, computer graphics, or related domains, with familiarity in 3D representations, mesh processing, and rendering pipelines.
About the Role:
We are looking for an ideal candidate with 5+ years of experience in Data Science / Machine Learning, with strong hands-on experience in Generative AI, Large Language Models (LLMs), NLP, and AI-powered applications. The candidate should be comfortable working across the complete AI lifecycle—from understanding business requirements and experimenting with models to building, evaluating, deploying, and monitoring production-grade GenAI solutions.
The role requires a combination of strong technical expertise, business understanding, problem-solving ability, and stakeholder management skills.
Key Responsibilities:
Generative AI & LLM
· Design, develop, and deploy Generative AI and LLM-based solutions for enterprise use cases.
· Work with models such as OpenAI, Azure OpenAI, Llama, Mistral, Gemini, or equivalent LLM platforms.
· Develop applications using prompt engineering, structured outputs, function/tool calling, and LLM orchestration.
· Design and implement Retrieval-Augmented Generation (RAG) solutions.
· Work with vector databases and semantic search for enterprise knowledge retrieval.
· Develop and evaluate AI agents and multi-step AI workflows.
· Implement techniques such as prompt optimization, context management, grounding, and hallucination reduction.
· Develop AI solutions for text classification, summarization, information extraction, question answering, document intelligence, and other enterprise use cases.
Machine Learning & Data Science
· Develop and optimize traditional Machine Learning and statistical models where appropriate.
· Perform data exploration, feature engineering, model selection, training, validation, and evaluation.
· Apply appropriate ML and statistical techniques to solve business problems.
· Work with structured, unstructured, and semi-structured data.
· Develop scalable data pipelines to support AI/ML solutions.
· Collaborate with Data Engineers to prepare and manage data for AI applications.
AI Evaluation & Productionization
· Design evaluation frameworks to measure LLM accuracy, relevance, groundedness, toxicity, latency, and cost.
· Implement guardrails and responsible AI practices.
· Monitor model and application performance in production.
· Identify model/data drift and implement appropriate improvement strategies.
· Optimize AI solutions for performance, scalability, reliability, and cost.
· Support deployment and productionization of AI/ML solutions.
· Client & Delivery Responsibilities
· Work closely with the CEO, Delivery team, Solution Architects, Engineering teams, and clients to understand business problems and identify AI opportunities.
· Translate business requirements into practical AI/ML solutions.
· Participate in client discussions, solution presentations, technical workshops, and POCs.
· Develop rapid prototypes and demonstrate the feasibility of GenAI solutions.
· Convert successful POCs into scalable, production-ready applications.
· Provide technical guidance and contribute to AI solution architecture.
· Prepare technical documentation, solution approaches, and project estimates where required.
· Stay current with developments in Generative AI, LLMs, Agentic AI, and AI engineering.
Required Skills:
· 5+ years of hands-on experience in Data Science, Machine Learning, AI, or a related field.
· Strong practical experience in Generative AI and LLM-based applications.
· Strong proficiency in Python.
· Strong understanding of Machine Learning and statistical concepts.
· Hands-on experience with:
o LLMs
o Prompt Engineering
o RAG
o Vector Databases
o Embeddings
o Semantic Search
o LLM Evaluation
o AI Guardrails
· Experience with frameworks/tools such as LangChain, LangGraph, LlamaIndex, or equivalent.
· Experience with APIs and integrating LLMs into enterprise applications.
· Strong SQL and data handling skills.
· Experience working with large and complex datasets.
· Strong understanding of NLP concepts.XX
Technical Skills:
· Experience with OpenAI / Azure OpenAI / AWS Bedrock / Google Vertex AI.
· Experience with vector databases such as Pinecone, Weaviate, Milvus, FAISS, or equivalent.
· Experience with Databricks, Snowflake, or cloud data platforms.
· Experience with Docker and CI/CD.
· Exposure to AWS, Azure, or GCP.
· Experience with ML/AI deployment and MLOps.
· Knowledge of AI security, data privacy, governance, and responsible AI.
· Experience building AI Agents / Agentic AI workflows.
· Experience with multimodal AI is an added advantage
Key Competencies
· Strong analytical and problem-solving ability.
· Ability to translate business problems into practical AI solutions.
· Strong communication and presentation skills.
· Ability to interact confidently with senior stakeholders and clients.
· Strong ownership and delivery mindset.
· Ability to work independently in a fast-paced environment.
- Strong experimentation and innovation mindset.
- Ability to balance technical feasibility, business value, scalability, and cost.
Required Education & Experience:
· Bachelor's or Master's degree in Computer Science, Data Science, Artificial Intelligence, Statistics, Mathematics, Engineering, or a related discipline







