Cutshort logo
For Employers
Metadome.ai logo
Principal AI Engineer
Principal AI Engineer

Principal AI Engineer at Metadome.ai · Bengaluru (Bangalore) · 10 - 15 years · ₹45L - ₹55L / yr · Raised funding · Posted 23 Sep 2026

Metadome.ai's logo

Principal AI Engineer

Ananya  Arenavaru's profile picture
Posted by Ananya Arenavaru
10 - 15 yrs
₹45L - ₹55L / yr
Bengaluru (Bangalore)
Skills
Fine-tuning LLMs
post-training SFT, RLHF, DPO
PyTorch
Python
Model deployment
LangChain
Vector database
MLOps
Data pipelines
Retrieval Augmented Generation (RAG)
CUDA
Machine Learning (ML)

PRINCIPAL AI ENGINEER @ METADOME.AI

Company Description

Metadome.ai builds frontier AI models that transform text, drawings, and CAD into production-ready, interactive 3D experiences. The company advances a full generative pipeline—text-to-CAD, 2D-to-3D

reconstruction, CAD completion and harmonization, and real-time interactive rendering—engineered for the precision required in the physical world. Its technology currently powers the modernization of

OEM aftersales for more than 30 automotive and heavy-equipment manufacturers worldwide, delivering accurate, scalable, and fast 3D solutions. Metadome.ai’s platform enables shoppable 3D parts, step-by-step repair animations, and a headless API that feeds consistent 3D assets into commerce, dealer, training, and service systems. The broader mission is to allow anyone to move from an idea, drawing, or specification to a production-grade 3D model and beyond in seconds.


Role Description

As a Principal AI Engineer — Generative CAD & 3D, you will lead the design, development, and deployment of advanced AI models that convert text, 2D drawings, and CAD files into engineering-grade 3D content. You will architect end-to-end generative pipelines, including

text-to-CAD, 2D-to-3D reconstruction, CAD completion, and real-time rendering, collaborating closely with product, design, and engineering teams to ship robust production systems. Day-to-day, you will experiment with novel neural network architectures, optimize model performance on large-scale CAD datasets, write high-quality production code, and guide the integration of AI services into customer-facing platforms. You will mentor other engineers, establish best practices for AI development, and contribute to technical strategy and roadmap. This is a full-time, hybrid role based in Bengaluru, with a mix of on-site collaboration and work-from-home flexibility.


Qualifications

  • Strong foundation in Computer Science and Software Development, including data structures, algorithms, system design, and production-grade coding in languages such as Python, C++, or similar.
  • Deep expertise in Neural Networks and Pattern Recognition, with hands-on experience designing, training, and deploying modern deep learning architectures for complex, high-dimensional data.
  • Experience with Natural Language Processing (NLP), including working with text encoders, multimodal models, and integrating language understanding into generative workflows.
  • Advanced degree (Master’s or PhD) in Computer Science, Electrical Engineering, Applied Mathematics, or a related field, or equivalent practical experience in AI/ML research and engineering.
  • Background in 3D geometry, CAD, computer graphics, or related domains, with familiarity in 3D representations, mesh processing, and rendering pipelines.
Read more
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos

About Metadome.ai

Founded :
2016
Type :
Product
Size :
20-100
Stage :
Raised funding

About

Metadome.ai is a 3D & XR technology company that is enabling agencies, enterprises, and game developers to deliver their 3D, AR, or VR

experiences across millions of mobile, desktop, and XR devices with its proprietary XR streaming platform. The platform is the only one that

provides support for unreal, unity & webGL, and enables the streaming of high-fidelity content consistently across all devices at an industry-best

cost and with low latency.

The company’s in-house team and network of agencies, industry experts, consultants, and creators help global enterprises & fortune 500

companies with the end-to-end conception, creation, and deployment of immersive customer experiences. These use cases span across 3D, AR,

VR, and Metaverse and cover multiple industries like Automotive, Home Decor, Fashion, Accessories, and Cosmetics.

The company is a pioneer in the automotive industry and its automotive arm, Autodome, is the one-stop solution for Automakers & Dealers for

enabling immersive customer experiences across their online & offline channels to drive virtual vehicle sales and the brand's metaverse adoption.

With office presence across India & US, the company is a pioneer in XR with 6+ years of expertise and a go-to option for global brands for driving

real business outcomes with enhanced engagement & conversions including Stellantis, Lexus, HUL, Mahindra, Asian Paints, MG Motor, and

Titan to name a few.

Read more

Candid answers by the company

What does the company do?
What is the location preference of jobs?

Metadome.ai is a 3D & XR technology company that is enabling agencies, enterprises, and game developers to deliver their 3D, AR, or VR

experiences across millions of mobile, desktop, and XR devices with its proprietary XR streaming platform. The platform is the only one that

provides support for unreal, unity & webGL, and enables the streaming of high-fidelity content consistently across all devices at an industry-best

cost and with low latency.

The company’s in-house team and network of agencies, industry experts, consultants, and creators help global enterprises & fortune 500

companies with the end-to-end conception, creation, and deployment of immersive customer experiences. These use cases span across 3D, AR,

VR, and Metaverse and cover multiple industries like Automotive, Home Decor, Fashion, Accessories, and Cosmetics.

The company is a pioneer in the automotive industry and its automotive arm, Autodome, is the one-stop solution for Automakers & Dealers for

enabling immersive customer experiences across their online & offline channels to drive virtual vehicle sales and the brand's metaverse adoption.

With office presence across India & US, the company is a pioneer in XR with 6+ years of expertise and a go-to option for global brands for driving

real business outcomes with enhanced engagement & conversions including Stellantis, Lexus, HUL, Mahindra, Asian Paints, MG Motor, and

Titan to name a few.

Company social profiles

bloginstagramlinkedintwitter

Similar jobs (10)

company logo
Prithisha Kathiresan
Posted by Prithisha Kathiresan
Bengaluru (Bangalore)
3 - 8 yrs
Best in industry
Generative AI (GenAI)
Large Language Models (LLM) tuning
Retrieval Augmented Generation (RAG)
Azure OpenAI
Amazon Web Services (AWS)
+2 more

Senior Generative AI Engineer

Employment Type: Permanent with VDart Digital

Work Location: Marathalli, Bengaluru

Job Description

We are seeking a highly skilled Senior Generative AI Engineer with strong expertise in designing, developing, and deploying enterprise-scale AI solutions using Large Language Models (LLMs) and modern Generative AI frameworks. The ideal candidate should have hands-on production experience building scalable GenAI applications, AI agents, autonomous workflows, and Retrieval-Augmented Generation (RAG) systems in cloud-native environments.

This role requires deep technical expertise in LLM orchestration, AI application architecture, prompt engineering, vector databases, MLOps, and production deployment of AI systems. Candidates should have proven experience delivering real-world AI solutions in enterprise environments with strong exposure to cloud platforms and DevOps practices.

Key Responsibilities

  • Design, build, and deploy enterprise-grade Generative AI applications using Large Language Models (LLMs).
  • Develop intelligent AI agents and autonomous workflows using frameworks such as LangChain, CrewAI, LangGraph, AutoGen, or similar agentic AI frameworks.
  • Implement and optimize Retrieval-Augmented Generation (RAG) pipelines using vector databases and semantic search technologies.
  • Work extensively on prompt engineering, tool calling, memory management, agent orchestration, and multi-agent systems.
  • Integrate and manage LLMs such as OpenAI, Azure OpenAI, Claude, Gemini, Llama, Mistral, or similar foundation models.
  • Develop scalable AI services and APIs using Python and FastAPI.
  • Build production-ready AI solutions with high availability, scalability, monitoring, and observability.
  • Deploy and manage AI applications in cloud-native environments using Docker and Kubernetes.
  • Collaborate with Data Science, ML Engineering, and DevOps teams to operationalize AI solutions.
  • Implement CI/CD pipelines and automated deployment processes for AI workloads.
  • Monitor model performance, latency, reliability, and operational efficiency in production environments.
  • Ensure AI solutions follow enterprise security, governance, and responsible AI standards.
  • Evaluate and adopt emerging Generative AI tools, frameworks, and models.

Required Skills

Generative AI & LLM Expertise

  • Strong hands-on experience with Generative AI and Large Language Models (LLMs).
  • Production-level experience building and deploying GenAI applications.
  • Expertise in LangChain, CrewAI, LangGraph, AutoGen, or similar frameworks.
  • Experience with AI agents, autonomous workflows, and multi-agent architectures.
  • Strong understanding of prompt engineering, embeddings, model evaluation, and LLM orchestration.
  • Experience integrating OpenAI, Azure OpenAI, Claude, Gemini, Llama, Mistral, or similar models.

RAG & Vector Databases

  • Strong experience implementing RAG pipelines and semantic retrieval systems.
  • Experience with vector databases such as Pinecone, Weaviate, ChromaDB, FAISS, or Milvus.
  • Understanding of chunking strategies, embeddings, indexing, reranking, and retrieval optimization.

Python & AI Development

  • Strong proficiency in Python.
  • Experience with FastAPI for AI service and API development.
  • Experience with AI/ML libraries and data processing tools such as Pandas and NumPy.

Cloud & Production Deployment

  • Mandatory production experience on at least one cloud platform:
  • Microsoft Azure
  • Experience deploying scalable AI applications in enterprise production environments.
  • Hands-on experience with Docker, Kubernetes, Jenkins, Terraform, and CI/CD pipelines.
  • Strong understanding of MLOps, AI deployment lifecycle, monitoring, and observability.

Engineering & Operational Excellence

  • Strong understanding of software engineering best practices.
  • Experience with Git, version control, automated testing, and release management.
  • Experience building secure, scalable, and high-performance AI solutions.
  • Ability to troubleshoot production AI systems and optimize performance.

Preferred Skills

  • Experience with AI observability and evaluation frameworks.
  • Exposure to fine-tuning, PEFT, LoRA, or model optimization techniques.
  • Experience with enterprise AI governance and responsible AI practices.
  • Knowledge of distributed AI systems and scalable inference architectures.
  • Familiarity with AI security and compliance standards.

Qualifications

  • Bachelor’s or Master’s degree in Computer Science, Artificial Intelligence, Data Science, or a related field.
  • 3–8 years of overall software engineering experience.
  • Minimum 3+ years of hands-on experience in Generative AI and LLM-based application development,
  • Proven track record of delivering enterprise-scale AI solutions in production environments.
  • Strong communication and stakeholder management skills.
Read more
company logo
Anishka Burde
Posted by Anishka Burde
Mumbai
6 - 13 yrs
Best in industry
Java
Python
LangGraph
LangChain
Retrieval Augmented Generation (RAG)
+2 more

Job Summary:

Wissen Technology is hiring an AI Implementation Engineer to build, deploy, and scale enterprise-grade Generative AI solutions across business-critical applications. The role involves developing production-ready AI systems using Azure AI services, Large Language Models (LLMs), RAG architectures, and agent-based frameworks while collaborating closely with engineering teams to drive AI adoption and innovation.


Experience

6-12 years


Location

Mumbai / Bangalore


Mode of Work

Hybrid


Mandatory Skills (Must Have)

• Python programming (6+ years) including asynchronous programming and backend application development

• Java and Spring Framework (3+ years) for enterprise-scale application integration

• Azure OpenAI Service, Azure AI Foundry, and Azure AI Search for production GenAI applications

• Retrieval Augmented Generation (RAG) architecture including embeddings, chunking, vector databases, reranking, and grounding techniques

• Agent Frameworks such as Microsoft Agent Framework, Semantic Kernel, AutoGen, LangChain, or LangGraph

• Snowflake and Cortex AI (Cortex Search, LLM Functions) with strong SQL expertise

• Prompt Engineering, LLM evaluation frameworks, testing, and model performance optimization

• DevOps and Cloud Deployment using Azure DevOps, GitHub Actions, Docker, AKS, Azure Functions, and observability tools


Optional Skills (Good to Have)

• React.js for AI-powered user interfaces and conversational applications

• Azure AI Content Safety and Responsible AI implementation experience

• Financial Services, Banking, or other regulated industry domain experience

• Real-time streaming applications and token-level LLM operations

• Performance optimization, caching strategies, and cost optimization for AI workloads

• Microsoft Azure AI Engineer Associate Certification

Read more
company logo
Shruti mujbaile
Posted by Shruti mujbaile
Gurugram, Pune
6 - 12 yrs
₹8L - ₹25L / yr
Generative AI
Agentic AI
Python
Machine Learning (ML)

Location: Pune / Gurgaon

Position: AI Engineer

work mode: WFO


  Job Description.

 Job responsibilities:

  • Responsibility for design, implementation and deployment of Generative AI, Agentic frameworks at scale
  • Strong in programming - Python a
  • Previous experience of working on Computer Vision projects and VLM /VLAM models.
  • In depth awareness of Transformer architectures and End to End Deep neural networks
  • Full stack AI / ML development experience
  • Design, build & maintain efficient and reliable Agentic / Generative AI code leveraging pipelines
  • Hosting and deployment knowledge in GCP or AWS or Azure along with advanced engineering concepts to build user friendly UI interface for easy adoption.


    Requirements:

 ·      4 to 8 years overall years of experience (Agentic AI, Generative AI, VLM, VLAM and LLM) with significant exposure in Development, Architecture design, scaling and hosting in cloud.


    Must Have –

 ·      Architecting and solutioning experience with Python and FAST API, Agentic Ai frameworks, VLMs, VLAMs, Open source LLM’s and Code based LLM models at scale with - Langchain /      Ollama, embeddings, Memory      Management etc.,

·      Practical experience in implementing Explainable and ethical AI models  Practical experience in implementing frameworks like RAG/ CAG/ Self-reflective RAG etc.,

·      Experience in cloud hosting either AWS or Azure or GCP.

·      Experience in ML-OPS - Implement a feedback mechanism to continually improve the model over time through feedback loop and monitoring KPI’s in production.

·      Experience with Quantization and Kubernetes or docker


    Good to have

·      gRPC implementation to expose the API’s on a server for easy usage and good user interface

·      Streamlit front end creation

·      Experience with SAFe framework deliveries.


Read more
BASF
Remote only
3 - 15 yrs
₹1.2L - ₹2.5L / yr
Python
API
Retrieval Augmented Generation (RAG)
Artificial Intelligence (AI)

POSITION OVERVIEW

We are seeking an experienced Senior Data Scientist & Generative AI Specialist on a contractual basis to support a premier Germany-based chemical manufacturing enterprise. In this role, you will lead the end-to-end design, development, and deployment of production-grade GenAI applications, multi-modal LLM workflows, and advanced retrieval platforms tailored to complex industrial and enterprise data ecosystems.

Working closely with cross-functional global teams, you will build robust backend microservices, implement state- of-the-art RAG/GraphRAG architectures, and leverage cloud-native AI infrastructure (Azure, Vector DBs, Knowledge Graphs) to drive operational efficiency and data-driven innovation.

KEY RESPONSIBILITIES

  • GenAI & LLM System Engineering: Design, build, and deploy production-grade multi-modal GenAI applications processing text, structured technical documentation, images, and telemetry data.
  • Advanced RAG & Graph Architecture: Implement cutting-edge Retrieval-Augmented Generation (RAG) and GraphRAG pipelines using document parsing frameworks, custom embeddings, vector databases, and knowledge graphs to capture complex domain relationships.
  • Scalable Backend Development: Architect high-throughput, low-latency microservice APIs using Python, FastAPI, and Flask, leveraging asynchronous programming (asyncio) and strict type validation (Pydantic) for long-running LLM processes.
  • Agentic Systems & Azure Ecosystem: Build autonomous agent systems using modern frameworks (MCP, A2A) and orchestrate enterprise workflows across the Microsoft Azure AI ecosystem (Azure AI Foundry, AI Search, Document Intelligence, Databricks).
  • Model Optimization & Evaluation: Execute systematic LLM fine-tuning, prompt optimization, and rigorous evaluation frameworks to assess AI output accuracy, reliability, and business impact against industrial requirements.
  • Data Layer Management: Architect and maintain enterprise database layers combining SQL (PostgreSQL) for structured transactional data with specialized vector search engines and graph stores.
  • Rapid Prototyping: Utilize AI-assisted development tools (Copilot, Claude Code) to accelerate delivery timelines and rapidly build functional UI prototypes for client feedback.

TECHNICAL QUALIFICATIONS

Core Development & Backend:

• Python Mastery: Deep expertise in writing clean, production-ready Python using asynchronous programming (asyncio), strict type-hinting (Pydantic), and automated testing patterns.

• Backend Microservices: Hands-on experience building microservices with FastAPI and Flask structured to handle asynchronous, long-running AI background tasks.

• Database Engineering: Strong command of PostgreSQL, relational schema design, vector indexing, and knowledge graph paradigms.

Machine Learning & AI Infrastructure:

• Model Expertise: Hands-on experience with leading multi-modal LLM architectures (OpenAI, Anthropic, Google) and domain-specific AI workflows.

• Retrieval & Parsing: Proven track record with document extraction frameworks, embedding models, vector search engines, and GraphRAG architectures.

• Cloud Infrastructure: Strong proficiency with Azure AI infrastructure (Foundry, Databricks, AI Search, Document Intelligence).

• Agentic Frameworks: Practical experience with open-source agent protocols (MCP, A2A), parameter-efficient fine-tuning (PEFT/LoRA), and model evaluation methodology.

CONTRACT & REMOTE REQUIREMENTS

• Contract Engagement: Contractual structure tailored to project milestones and deliverables.

• 100% Remote Setup: Fully equipped home office with high-speed, secure internet infrastructure.

• Timezone Overlap: Guaranteed 4-hour daily overlap with Central European Time (CET/CEST - Germany) to ensure smooth collaboration with enterprise stakeholders.

• Communication: Fluent professional English communication skills (written and spoken) for asynchronous and real-time technical coordination. 

Read more
company logo
Faisal AshrafNomani
Posted by Faisal AshrafNomani
Remote only
4 - 15 yrs
Best in industry
Generative AI
Large Language Models (LLM) tuning
Agentic AI
AI Agents
Retrieval Augmented Generation (RAG)

About the Role:

We are looking for an ideal candidate with 5+ years of experience in Data Science / Machine Learning, with strong hands-on experience in Generative AI, Large Language Models (LLMs), NLP, and AI-powered applications. The candidate should be comfortable working across the complete AI lifecycle—from understanding business requirements and experimenting with models to building, evaluating, deploying, and monitoring production-grade GenAI solutions.

The role requires a combination of strong technical expertise, business understanding, problem-solving ability, and stakeholder management skills.



Key Responsibilities:

 

Generative AI & LLM

·      Design, develop, and deploy Generative AI and LLM-based solutions for enterprise use cases.

·      Work with models such as OpenAI, Azure OpenAI, Llama, Mistral, Gemini, or equivalent LLM platforms.

·      Develop applications using prompt engineering, structured outputs, function/tool calling, and LLM orchestration.

·      Design and implement Retrieval-Augmented Generation (RAG) solutions.

·      Work with vector databases and semantic search for enterprise knowledge retrieval.

·      Develop and evaluate AI agents and multi-step AI workflows.

·      Implement techniques such as prompt optimization, context management, grounding, and hallucination reduction.

·      Develop AI solutions for text classification, summarization, information extraction, question answering, document intelligence, and other enterprise use cases.


Machine Learning & Data Science

·      Develop and optimize traditional Machine Learning and statistical models where appropriate.

·      Perform data exploration, feature engineering, model selection, training, validation, and evaluation.

·      Apply appropriate ML and statistical techniques to solve business problems.

·      Work with structured, unstructured, and semi-structured data.

·      Develop scalable data pipelines to support AI/ML solutions.

·      Collaborate with Data Engineers to prepare and manage data for AI applications.


AI Evaluation & Productionization

·      Design evaluation frameworks to measure LLM accuracy, relevance, groundedness, toxicity, latency, and cost.

·      Implement guardrails and responsible AI practices.

·      Monitor model and application performance in production.

·      Identify model/data drift and implement appropriate improvement strategies.

·      Optimize AI solutions for performance, scalability, reliability, and cost.

·      Support deployment and productionization of AI/ML solutions.

·      Client & Delivery Responsibilities

·      Work closely with the CEO, Delivery team, Solution Architects, Engineering teams, and clients to understand business problems and identify AI opportunities.

·      Translate business requirements into practical AI/ML solutions.

·      Participate in client discussions, solution presentations, technical workshops, and POCs.

·      Develop rapid prototypes and demonstrate the feasibility of GenAI solutions.



·      Convert successful POCs into scalable, production-ready applications.

·      Provide technical guidance and contribute to AI solution architecture.

·      Prepare technical documentation, solution approaches, and project estimates where required.

·      Stay current with developments in Generative AI, LLMs, Agentic AI, and AI engineering.

Required Skills:

·       5+ years of hands-on experience in Data Science, Machine Learning, AI, or a related field.

·      Strong practical experience in Generative AI and LLM-based applications.

·      Strong proficiency in Python.

·      Strong understanding of Machine Learning and statistical concepts.

·      Hands-on experience with:

o       LLMs

o       Prompt Engineering

o       RAG

o       Vector Databases

o       Embeddings

o       Semantic Search

o       LLM Evaluation

o       AI Guardrails

·      Experience with frameworks/tools such as LangChain, LangGraph, LlamaIndex, or equivalent.

·      Experience with APIs and integrating LLMs into enterprise applications.

·      Strong SQL and data handling skills.

·      Experience working with large and complex datasets.

·      Strong understanding of NLP concepts.XX



Technical Skills:

·      Experience with OpenAI / Azure OpenAI / AWS Bedrock / Google Vertex AI.

·      Experience with vector databases such as Pinecone, Weaviate, Milvus, FAISS, or equivalent.

·      Experience with Databricks, Snowflake, or cloud data platforms.

·      Experience with Docker and CI/CD.

·      Exposure to AWS, Azure, or GCP.

·      Experience with ML/AI deployment and MLOps.

·       Knowledge of AI security, data privacy, governance, and responsible AI.

·      Experience building AI Agents / Agentic AI workflows.

·      Experience with multimodal AI is an added advantage

Key Competencies

·      Strong analytical and problem-solving ability.

·      Ability to translate business problems into practical AI solutions.

·      Strong communication and presentation skills.

·      Ability to interact confidently with senior stakeholders and clients.

·      Strong ownership and delivery mindset.

·      Ability to work independently in a fast-paced environment.

  • Strong experimentation and innovation mindset.
  • Ability to balance technical feasibility, business value, scalability, and cost.

Required Education & Experience:

·      Bachelor's or Master's degree in Computer Science, Data Science, Artificial Intelligence, Statistics, Mathematics, Engineering, or a related discipline


Read more
company logo
Sandeep C
Posted by Sandeep C
Bengaluru (Bangalore)
8 - 16 yrs
₹1L - ₹2L / yr (ESOP available)
Large Language Models (LLM)
Agentic AI
Applied mathematics

Key Responsibilities:

·      Architectural Leadership: Design and lead the development of robust, scalable AI architectures, ensuring high performance, reliability, and security.

·      Applied Mathematics & Statistics: Apply statistical analysis, numerical computation, and mathematical modeling to derive insights from large-scale data and optimize model performance.

·      Deep Learning Development: Design, train, and deploy advanced Deep Learning (DL) models.

·      Technical Mentorship: Mentor engineering teams on best practices for AI/ML, coding standards, and architectural design.

·      Model Optimization: Optimize models for speed, efficiency, and accuracy using techniques like pruning, quantization, or GPU acceleration.

·      Strategy & Innovation: Evaluate and select appropriate AI frameworks, tools, and platforms, staying abreast of cutting-edge research and industry trends.

Qualifications:

Required:

·      Education: Master's or PhD in Computer Science, Applied Mathematics, Statistics, Physics, or a related quantitative field.

·      Experience: 10+ years of experience in software development, with at least 3-5 years in a Applied Mathematics and Deep learning.

·      AI/ML Expertise: Proven experience designing and deploying deep learning models in production using frameworks.

·      Mathematics/Statistics: Strong proficiency in linear algebra, calculus, probability, and statistical methods.

·      Programming Skills: Expert-level coding skills in Python (NumPy, Pandas, Scikit-learn) and experience with languages like Java or C++.

Key Competencies:

  • Strategic mindset with deep operational awareness.
  • Excellent communication and stakeholder management skills.
  • Ability to simplify complex technical concepts for executive reporting.
  • Strong leadership, people development, and cross-functional influencing skills.

Bias for action and a relentless focus on continuous improvement.

Read more
company logo
Bhawna Khemani
Posted by Bhawna Khemani
Bengaluru (Bangalore), Delhi, Gurugram, Noida, Ghaziabad, Faridabad
4 - 13 yrs
₹11L - ₹35L / yr
Generative AI
Large Language Models (LLM)
Retrieval Augmented Generation (RAG)
Python
LlamaIndex
+4 more

Generative AI Engineer 

Role Overview:

You will be responsible for the hands-on development, coding, and deployment of AI-powered features. Your focus is on writing clean, efficient code to integrate LLMs into our existing tech stack, building robust data pipelines for RAG, and ensuring the reliability of model outputs through rigorous testing and optimization.

Key Responsibilities

  • Application Implementation: Code and integrate LLM APIs (OpenAI, Anthropic, etc.) or local models into backend services using Python, FastAPI, etc.,
  • MCP Server Development: Design and implement custom MCP servers using the official SDKs (Python/TypeScript) to expose internal databases, APIs, and file systems to AI agents.
  • RAG Implementation: Build and maintain the "plumbing" for Retrieval-Augmented Generation—specifically coding the data ingestion scripts, text chunking logic, and metadata filtering.
  • Vector DB Management: Perform day-to-day operations on vector databases (Pinecone, Milvus, etc.), including indexing, querying, and optimizing search retrieval.
  • Prompt Programming: Develop, version-control, and refine complex prompt templates (using Jinja2 or similar) to ensure consistent structured outputs (JSON/YAML).
  • Agent Development: Implement multi-step workflows using LangChain, LangGraph, CrewAI etc.,, focusing on tool-calling logic and error handling.
  • Evaluation & Testing: Build automated test suites to detect "hallucinations" and measure accuracy using frameworks.
  • Performance Tuning: Implement caching layers and streaming responses to reduce latency and improve the end-user experience; Token optimization.
  • Data Pre-processing: Clean and tokenize datasets for model fine-tuning or high-quality context retrieval.

Technical Skills (The "Execution" Stack)

  • Language: Advanced Python (Asyncio, Pydantic) and optional TypeScript/Node.js (for full-stack integration).
  • AI Frameworks: Hands-on experience with any of LangChain, LlamaIndex, and Hugging Face Transformers. RAG and Vector search concepts.
  • Data Handling: Proficiency in SQL and handling unstructured data formats (PDFs, Markdown, JSON).
  • Deployment: Practical experience with Docker, GitHub Actions (CI/CD), and experience with OpenTelemetry, LangSmith, Weights & Biases etc., Understanding of evaluation/guardrails.
  • MCP/API Proficiency: Deep understanding of RESTful APIs, Streaming HTTP, MCP server vs client, JSONRPC
Read more
The industry’s only Manufacturing Operating System
The industry’s only Manufacturing Operating System
Agency job
via by Ariba Khan
Hyderabad
10 - 15 yrs
Best in industry
Artificial Intelligence (AI)
Generative AI (GenAI)
Retrieval Augmented Generation (RAG)
Large Language Models (LLM)

We’re on hunt for AI Architect


Responsibilities:

  • 10–15+ years overall experience, with recent hands-on AI/GenAI architecture ownership.
  • Must have architected enterprise AI platforms/solutions end-to-end, not just individual ML models or PoCs.
  • Strong GenAI/LLM production experience: RAG, embeddings, vector DBs, hybrid search, reranking, evaluation, guardrails.
  • Strong Agentic AI understanding: agents, tool calling, workflows, orchestration, human-in-the-loop.
  • Experience taking AI solutions from architecture → production → scale, ideally across multiple business teams/use cases.
  • Strong cloud architecture — Azure/AWS preferred; hybrid/on-prem experience is a plus.
  • Must understand enterprise security, governance, Responsible AI, observability and LLMOps/MLOps.
  • Should be able to articulate build-vs-buy, MVP-vs-target architecture, cost/performance/security tradeoffs.
  • Strong stakeholder-facing / consulting ability — can work with business leaders, engineering, security and data teams and influence without authority.


There is scope to move to the US for this role if you are aligned for the same, else this will be a WFO role from Hyderabad location

Read more
company logo
Nirmala Lama
Posted by Nirmala Lama
Mumbai
1 - 5 yrs
₹18L - ₹23L / yr
Generative AI
Retrieval Augmented Generation (RAG)
LangGraph
LangChain
Large Language Models (LLM)
+2 more

About the role

We are seeking an AI Engineer to build and implement AI systems for content production at scale. You'll work at the intersection of engineering and content designing prompt pipelines, integrating generative models, and building the tooling that turns source material into finished creative output. The ideal candidate is technically strong but also has taste: someone who understands story and craft, and can tell the difference between output that's technically correct and output that's actually good.


Responsibilities

  • Build and iterate on prompt pipelines and multi-agent workflow components
  • Design and integrate agentic workflows orchestrate multi-step, tool-using agents that plan, call models, and hand off between stages in production
  • Deploy and serve open-source models set up inference endpoints, manage GPU compute, and optimize for latency and cost
  • Write evals compare outputs against references, quantify quality, and feed results back into the pipeline
  • Work on data pipelines: structured extraction from messy source text, localization, similarity/dedup
  • Debug and maintain pipeline stages in production


What you bring:  

  • (1+/3+) years of engineering experience, or a strong portfolio of shipped projects
  • Solid Python fundamentals clean, working, readable code
  • Hands-on experience with LLM APIs and prompt engineering (personal projects count)
  • Comfort with Git, REST APIs, and working in a Linux environment
  • A feel for content and narrative you can judge whether generated output is actually good, not just valid
  • Curiosity and clear communication you ask good questions and don't stay stuck silently


 Preferred

  • Exposure to agent/orchestration frameworks (LangGraph, LangChain, CrewAI)
  • Familiarity with vector databases, embeddings, or RAG (Qdrant, pgvector)
  • Hands-on work with open-source generative media models Flux, LTX, Wan, or similar
  • Experience deploying open-source models for inference (vLLM, ComfyUI, Replicate/Cog, Docker + GPU)
  • Experience writing evals or LLM-as-judge scoring
  • Node.js and Fastapi familiarity, or experience deploying on AWS


Read more
company logo
Waseem Shariff
Posted by Waseem Shariff
Mumbai
1 - 3 yrs
₹15L - ₹24L / yr
Generative AI
Retrieval Augmented Generation (RAG)
LangGraph
LangChain
Large Language Models (LLM) tuning
+3 more

About the role

We are seeking an AI Engineer to build and implement AI systems for content production at scale. You'll work at the intersection of engineering and content designing prompt pipelines, integrating generative models, and building the tooling that turns source material into finished creative output. The ideal candidate is technically strong but also has taste: someone who understands story and craft, and can tell the difference between output that's technically correct and output that's actually good.


Responsibilities

  • Build and iterate on prompt pipelines and multi-agent workflow components
  • Design and integrate agentic workflows orchestrate multi-step, tool-using agents that plan, call models, and hand off between stages in production
  • Deploy and serve open-source models set up inference endpoints, manage GPU compute, and optimize for latency and cost
  • Write evals compare outputs against references, quantify quality, and feed results back into the pipeline
  • Work on data pipelines: structured extraction from messy source text, localization, similarity/dedup
  • Debug and maintain pipeline stages in production


What you bring:  

  • (1+/3+) years of engineering experience, or a strong portfolio of shipped projects
  • Solid Python fundamentals clean, working, readable code
  • Hands-on experience with LLM APIs and prompt engineering (personal projects count)
  • Comfort with Git, REST APIs, and working in a Linux environment
  • A feel for content and narrative you can judge whether generated output is actually good, not just valid
  • Curiosity and clear communication you ask good questions and don't stay stuck silently


 Preferred

  • Exposure to agent/orchestration frameworks (LangGraph, LangChain, CrewAI)
  • Familiarity with vector databases, embeddings, or RAG (Qdrant, pgvector)
  • Hands-on work with open-source generative media models Flux, LTX, Wan, or similar
  • Experience deploying open-source models for inference (vLLM, ComfyUI, Replicate/Cog, Docker + GPU)
  • Experience writing evals or LLM-as-judge scoring
  • Node.js and Fastapi familiarity, or experience deploying on AWS


Read more
Why apply to jobs via Cutshort
people_solving_puzzle
Personalized job matches
Stop wasting time. Get matched with jobs that meet your skills, aspirations and preferences.
people_verifying_people
Verified hiring teams
See actual hiring teams, find common social connections or connect with them directly.
ai_chip
Move faster with AI
We use AI to get you faster responses, recommendations and unmatched user experience.
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos