Senior GenAI Engineer — RAG (Full-Stack) at Ampera Technologies · Chennai, Bengaluru (Bangalore), Delhi, Gurugram, Noida, Ghaziabad, Faridabad, Mumbai, Pune, Hyderabad, Kolkata · 5 - 15 years · Profitable · Posted 24 Sep 2026

Title : Senior GenAI Engineer — RAG (Full-Stack)
Experience : 5+ years
Location : Remote
Work type : Chennai - Work from Office/ Remote – Other locations
Employment Type : Full Time
Notice Period : Immediate
Work Day :Mon to Fri
Key Responsibilities:
- RAG pipeline end to end: ingestion integration, hybrid retrieval with reranking, prompt/context strategy, citation resolution, refusal behavior
- Permission-aware retrieval: source ACL mapping (SharePoint/Entra, Confluence) to fail-closed retrieval filters; zero-leakage test suite partnership with QA
- Vector database design and operations (Milvus or pgvector): schema, metadata filters, sync, performance
- Full-stack product build: React/TypeScript chat and citation experience, Python/FastAPI services, REST APIs, SSO/OIDC integration, admin configuration UI
- Evaluation-driven development: retrieval precision, faithfulness, and citation-accuracy metrics as the daily working loop; A/B testing of retrieval and prompt variants
- Latency engineering to the 3–5s first-token / ~15s complete-answer targets at concurrency
Technical Skills:
- 5+ years software engineering with 2+ years building RAG/LLM applications in production — with real users and real quality metrics, not notebooks
- Deep retrieval craft: chunking strategy, embeddings, hybrid search, rerankers; you can explain why retrieval fails and how you measured the fix
- Genuine full-stack evidence: shipped React/TypeScript front ends AND Python back-end services in production; API design; OIDC/SAML integration
- Vector database production experience (Milvus, pgvector, Weaviate, or equivalent) including permission/metadata filtering
- Evaluation fluency: has built or operated a retrieval/answer quality harness with numeric thresholds
Strongly Preferred:
- Permission-aware/multi-tenant retrieval specifically; Microsoft Graph API; NIM/OpenAI-compatible serving endpoints; streaming UX; enterprise design systems; banking content domains
About Ampera:
Ampera Technologies, a purpose driven Digital IT Services with primary focus on supporting our client with their Data, AI / ML, Accessibility and other Digital IT needs. We also ensure that equal opportunities are provided to Persons with Disabilities Talent. Ampera Technologies has its Global Headquarters in Chicago, USA and its Global Delivery Center is based out of Chennai, India. We are actively expanding our Tech Delivery team in Chennai and across India. We offer exciting benefits for our teams, such as 1) Hybrid and Remote work options available, 2) Opportunity to work directly with our Global Enterprise Clients, 3) Opportunity to learn and implement evolving Technologies, 4) Comprehensive healthcare, and 5) Conducive environment for Persons with Disability Talent meeting Physical and Digital Accessibility standards

About Ampera Technologies
About
At Ampera Technologies, we empower businesses with cutting-edge data analytics, quality assurance, and data engineering solutions
Company social profiles
Similar jobs (10)
The recruiter has not been active on this job recently. You may apply but please expect a delayed response.
Location: Jaipur (Work From Office)
Employment Type: Full-Time
We're looking for a GenAI Engineer (LLM Engineer) to build scalable AI-powered SaaS applications using Large Language Models (LLMs). You'll develop intelligent AI workflows, integrate LLMs into production systems, and build secure, high-performance AI solutions.
Key Responsibilities
- Integrate LLM APIs (OpenAI, Claude, Hugging Face) into production applications.
- Design and optimize RAG pipelines and prompt engineering workflows.
- Build and manage Vector Databases (Pinecone, Weaviate, pgvector).
- Optimize AI performance, latency, and operational cost.
- Ensure secure, scalable AI architecture.
- Collaborate with Product and Engineering teams to deliver AI-powered features.
Requirements
- 3+ years of backend development using Python, Go, or Node.js.
- Hands-on experience with LLMs, LangChain or LlamaIndex.
- Strong understanding of RAG, Prompt Engineering, and Vector Databases.
- Experience with AWS, GCP, or Azure.
- Knowledge of APIs, Microservices, and AI application development.
Preferred: Experience in SaaS/FinTech, LLMOps, or Model Fine-tuning.
Education: B.Tech, BCA, or equivalent technical qualification.
Apply Now
Application Form: https://zfrmz.com/pAKb2ynfomIsuNwRfRbV?utm_source=cutshort
POSITION OVERVIEW
We are seeking an experienced Senior Data Scientist & Generative AI Specialist on a contractual basis to support a premier Germany-based chemical manufacturing enterprise. In this role, you will lead the end-to-end design, development, and deployment of production-grade GenAI applications, multi-modal LLM workflows, and advanced retrieval platforms tailored to complex industrial and enterprise data ecosystems.
Working closely with cross-functional global teams, you will build robust backend microservices, implement state- of-the-art RAG/GraphRAG architectures, and leverage cloud-native AI infrastructure (Azure, Vector DBs, Knowledge Graphs) to drive operational efficiency and data-driven innovation.
KEY RESPONSIBILITIES
- GenAI & LLM System Engineering: Design, build, and deploy production-grade multi-modal GenAI applications processing text, structured technical documentation, images, and telemetry data.
- Advanced RAG & Graph Architecture: Implement cutting-edge Retrieval-Augmented Generation (RAG) and GraphRAG pipelines using document parsing frameworks, custom embeddings, vector databases, and knowledge graphs to capture complex domain relationships.
- Scalable Backend Development: Architect high-throughput, low-latency microservice APIs using Python, FastAPI, and Flask, leveraging asynchronous programming (asyncio) and strict type validation (Pydantic) for long-running LLM processes.
- Agentic Systems & Azure Ecosystem: Build autonomous agent systems using modern frameworks (MCP, A2A) and orchestrate enterprise workflows across the Microsoft Azure AI ecosystem (Azure AI Foundry, AI Search, Document Intelligence, Databricks).
- Model Optimization & Evaluation: Execute systematic LLM fine-tuning, prompt optimization, and rigorous evaluation frameworks to assess AI output accuracy, reliability, and business impact against industrial requirements.
- Data Layer Management: Architect and maintain enterprise database layers combining SQL (PostgreSQL) for structured transactional data with specialized vector search engines and graph stores.
- Rapid Prototyping: Utilize AI-assisted development tools (Copilot, Claude Code) to accelerate delivery timelines and rapidly build functional UI prototypes for client feedback.
TECHNICAL QUALIFICATIONS
Core Development & Backend:
• Python Mastery: Deep expertise in writing clean, production-ready Python using asynchronous programming (asyncio), strict type-hinting (Pydantic), and automated testing patterns.
• Backend Microservices: Hands-on experience building microservices with FastAPI and Flask structured to handle asynchronous, long-running AI background tasks.
• Database Engineering: Strong command of PostgreSQL, relational schema design, vector indexing, and knowledge graph paradigms.
Machine Learning & AI Infrastructure:
• Model Expertise: Hands-on experience with leading multi-modal LLM architectures (OpenAI, Anthropic, Google) and domain-specific AI workflows.
• Retrieval & Parsing: Proven track record with document extraction frameworks, embedding models, vector search engines, and GraphRAG architectures.
• Cloud Infrastructure: Strong proficiency with Azure AI infrastructure (Foundry, Databricks, AI Search, Document Intelligence).
• Agentic Frameworks: Practical experience with open-source agent protocols (MCP, A2A), parameter-efficient fine-tuning (PEFT/LoRA), and model evaluation methodology.
CONTRACT & REMOTE REQUIREMENTS
• Contract Engagement: Contractual structure tailored to project milestones and deliverables.
• 100% Remote Setup: Fully equipped home office with high-speed, secure internet infrastructure.
• Timezone Overlap: Guaranteed 4-hour daily overlap with Central European Time (CET/CEST - Germany) to ensure smooth collaboration with enterprise stakeholders.
• Communication: Fluent professional English communication skills (written and spoken) for asynchronous and real-time technical coordination.
Strong AI/ML Engineer Profile
Mandatory (Experience) : Must have 3+ years of experience in software engineering with atleast 1+ years in GenAI application development and production deployment
Mandatory (GenAI Application Development): Must have proven experience building GenAI applications covering RAG pipelines, multi-agent systems, Text2SQL, and fine-tuning
Mandatory (Production GenAI Deployment): Must have expertise deploying production-grade GenAI applications including model evaluation, optimisation, and ownership of full production rollouts
Mandatory (ML & Data Science Tooling): Must have strong hands-on experience with core ML and data science tools including pandas, scikit-learn, and PyTorch
Mandatory (Cloud ML Infrastructure): Must have experience building and deploying production-grade ML workloads on at least one of AWS, Azure, or GCP
Mandatory (Communication): Must have strong English communication skills with the ability to work across time zones and collaborate cross-functionally with product, engineering, and business stakeholders
Mandatory (Note 1) : Role is Hybrid, WFH flexibility as well upto 6 days a month
Mandatory (Note 2) : CTC is inclusive of 10% variable
Mandatory (Note 3): Candidates should be available to join within May 31st or June first week max
Generative AI Engineer
Role Overview:
You will be responsible for the hands-on development, coding, and deployment of AI-powered features. Your focus is on writing clean, efficient code to integrate LLMs into our existing tech stack, building robust data pipelines for RAG, and ensuring the reliability of model outputs through rigorous testing and optimization.
Key Responsibilities
- Application Implementation: Code and integrate LLM APIs (OpenAI, Anthropic, etc.) or local models into backend services using Python, FastAPI, etc.,
- MCP Server Development: Design and implement custom MCP servers using the official SDKs (Python/TypeScript) to expose internal databases, APIs, and file systems to AI agents.
- RAG Implementation: Build and maintain the "plumbing" for Retrieval-Augmented Generation—specifically coding the data ingestion scripts, text chunking logic, and metadata filtering.
- Vector DB Management: Perform day-to-day operations on vector databases (Pinecone, Milvus, etc.), including indexing, querying, and optimizing search retrieval.
- Prompt Programming: Develop, version-control, and refine complex prompt templates (using Jinja2 or similar) to ensure consistent structured outputs (JSON/YAML).
- Agent Development: Implement multi-step workflows using LangChain, LangGraph, CrewAI etc.,, focusing on tool-calling logic and error handling.
- Evaluation & Testing: Build automated test suites to detect "hallucinations" and measure accuracy using frameworks.
- Performance Tuning: Implement caching layers and streaming responses to reduce latency and improve the end-user experience; Token optimization.
- Data Pre-processing: Clean and tokenize datasets for model fine-tuning or high-quality context retrieval.
Technical Skills (The "Execution" Stack)
- Language: Advanced Python (Asyncio, Pydantic) and optional TypeScript/Node.js (for full-stack integration).
- AI Frameworks: Hands-on experience with any of LangChain, LlamaIndex, and Hugging Face Transformers. RAG and Vector search concepts.
- Data Handling: Proficiency in SQL and handling unstructured data formats (PDFs, Markdown, JSON).
- Deployment: Practical experience with Docker, GitHub Actions (CI/CD), and experience with OpenTelemetry, LangSmith, Weights & Biases etc., Understanding of evaluation/guardrails.
- MCP/API Proficiency: Deep understanding of RESTful APIs, Streaming HTTP, MCP server vs client, JSONRPC
Strong AI Engineer / Machine Learning Engineer profiles.
2
Mandatory (Experience 1) – Must have minimum 3+ years of hands-on experience in Data Science, Machine Learning, Applied AI, NLP, Deep Learning, or Generative AI solutions.
3
Mandatory (Experience 2) – Must have strong hands-on experience in Python programming, SQL, data analysis, feature engineering, model development, and production-grade ML applications.
4
Mandatory (Experience 3) – Must have experience working with Machine Learning and Deep Learning frameworks such as PyTorch, TensorFlow, Keras, Scikit-learn, or equivalent.
5
Mandatory (Experience 4) – Must have hands-on experience working on NLP, embeddings, semantic search, text classification, document understanding, recommendation systems, or similar AI/ML use cases.
6
Mandatory (Experience 5) – Must have experience working with Large Language Models (LLMs) such as GPT, Llama, Mistral, Claude, Gemini, Phi, or similar foundation models.
7
Mandatory (Experience 6) – Must have hands-on experience building or implementing RAG (Retrieval Augmented Generation) systems, vector search, knowledge retrieval, embeddings, chunking, indexing, or semantic retrieval solutions.
8
Mandatory (Experience 7) – Must have experience working with Git, CI/CD practices, production environments, and scalable AI/ML systems.
9
Mandatory (CTC) – The CTC breakup offered will be 75% fixed + 25% variable, as per company policy.
10
Mandatory (Age) - Candidate's Age should be below 28 Years
Strong AI Engineer / Machine Learning Engineer profiles.
2
Mandatory (Experience 1) – Must have minimum 5+ years of hands-on experience in Data Science, Machine Learning, Applied AI, NLP, Deep Learning, or Generative AI solutions.
3
Mandatory (Experience 2) – Must have strong hands-on experience in Python programming, SQL, data analysis, feature engineering, model development, and production-grade ML applications.
4
Mandatory (Experience 3) – Must have experience working with Machine Learning and Deep Learning frameworks such as PyTorch, TensorFlow, Keras, Scikit-learn, or equivalent.
5
Mandatory (Experience 4) – Must have hands-on experience working on NLP, embeddings, semantic search, text classification, document understanding, recommendation systems, or similar AI/ML use cases.
6
Mandatory (Experience 5) – Must have experience working with Large Language Models (LLMs) such as GPT, Llama, Mistral, Claude, Gemini, Phi, or similar foundation models.
7
Mandatory (Experience 6) – Must have hands-on experience building or implementing RAG (Retrieval Augmented Generation) systems, vector search, knowledge retrieval, embeddings, chunking, indexing, or semantic retrieval solutions.
8
Mandatory (Experience 7) – Must have experience working with Git, CI/CD practices, production environments, and scalable AI/ML systems.
9
Mandatory (CTC) – The CTC breakup offered will be 75% fixed + 25% variable, as per company policy.
10
Mandatory (Age) - Candidate's Age should be below 30 Years
11
Preferred (Experience 1) – Experience with MLFlow, Kubeflow, Airflow, Prefect, Feature Stores, Model Registry, or MLOps/LLMOps frameworks.
12
Preferred (Experience 2) – Experience working with Vector Databases, Spark, PySpark, distributed ML pipelines, large-scale data processing, or real-time ML systems..
13
Preferred (Experience 3) – Familiarity with Docker, Kubernetes, Azure, AWS, GCP, cloud-native AI deployments, and scalable ML architecture.
14
Preferred (Company) – Candidates from AI-first startups, Fintech, Banking, Lending, Fraud Analytics, Risk Analytics, Product Companies, SaaS organizations, or data-driven technology companies
15
Mandatory ( Pedigree) - B.TECH / M.TECH from Tier 1 Colleges (IIT's, NIT's, BITS) are Considered.
Strong Data Scientist / AI Engineer / Generative AI Engineer profile.
2
Mandatory (Experience 1) - Must have 3+ years of hands-on experience in Data Science, Artificial Intelligence, Machine Learning, Deep Learning, NLP, or Generative AI application development.
3
Mandatory (Experience 2) - Must have strong hands-on experience in Python programming, backend development, API development, and production-grade application support.
4
Mandatory (Experience 3) - Must have experience working with Machine Learning and Deep Learning frameworks such as PyTorch, TensorFlow, Keras, or Scikit-learn.
5
Mandatory (Experience 4) - Must have hands-on experience in NLP use cases such as text classification, sentiment analysis, entity recognition (NER), semantic search, embeddings, or document understanding.
6
Mandatory (Experience 5) - Must have experience working with Large Language Models (LLMs) such as GPT, LLaMA, Mistral, Phi, Claude, Gemini, or similar models.
7
Mandatory (Experience 6) - Must have hands-on experience building or implementing Retrieval Augmented Generation (RAG) solutions, vector search, semantic search, or knowledge-based AI applications.
8
Mandatory (Experience 7) - Must have experience with Prompt Engineering and Generative AI frameworks such as LangChain, LangGraph, AI Agents, Azure OpenAI, or similar technologies.
9
Mandatory (Experience 8) - Must have experience developing, consuming, or integrating APIs using Python frameworks such as FastAPI, Flask, or similar technologies.
10
Mandatory (CTC) - The CTC breakup offered will be 75% fixed + 25% variable, as per company policy.
11
Preferred (Experience 1) - Experience with LLMOps/MLOps tools for monitoring, evaluation, experimentation, and versioning of AI models.
12
Preferred (Experience 2) - Exposure to Azure OpenAI, Azure Kubernetes Service (AKS), Kubernetes, cloud-native AI deployments, or distributed systems.
Location: Hyderabad, India (home base), deployed at client sites in India. Occasional Middle East exposure possible.
About the Role
You will work as a senior AI engineer who embeds inside a customer's business. Your job is to learn how the business makes money, find the highest value problem, and build a working system that solves it.
Four behaviors define this role:
- Go where the work happens. You work onsite with the customer, in the room where decisions are made.
- Show working software early. You build a prototype in days, not a document in weeks.
- One person owns the outcome. You are the single point of accountability for the result.
- Stay after go-live. You keep running and improving the system after launch.
You are the single point of accountability. You are not a solo builder. A full KnackLabs engineering team in Hyderabad builds and runs the production systems behind you.
This role involves extended onsite deployments at client locations in other cities, sometimes up to six months at a stretch. Please apply only if you are ready for this way of working.
What you'll own
- Discovery - Learn how the customer makes money. Find the highest value problem to solve first.
- The prototype - Build a working prototype fast, using real or sample data, to prove the idea.
- The roadmap - Decide what to build, in what order, and set clear success measures tied to business outcomes.
- The build - Design and ship the production system with the Hyderabad engineering team. This includes data integration, agents, retrieval, and evaluations.
- The client relationship - Be the trusted technical contact for the customer, from engineers to senior leaders.
- Go live and after - Deploy the system, watch how it performs, fix problems, and improve it over time.
- Feedback to the product - Share what you learn in the field so the vendor's product and our internal tools get better.
What we are looking for
- Around 4 or more years of software engineering experience, including customer-facing or client delivery work.
- Strong programming skills in Python. Working knowledge of TypeScript or JavaScript.
- A full-stack development experience with strength in backend technologies.
- Production experience with large language models, including prompt engineering and agent development.
- You build with AI coding tools like Claude Code or Codex as your default way of working, and you have shipped real apps or agents this way.
- Experience building retrieval-augmented generation (RAG) systems: chunking, embeddings, vector databases, retrieval, and reranking.
- Experience building and deploying AI systems.
- Experience integrating with APIs and enterprise systems.
- Experience with at least one cloud platform (AWS, Azure, or GCP).
- Clear communication. You can explain a technical choice to an engineer and to a business leader.
- High ownership and comfort with ambiguity. You can take an unclear problem and turn it into a plan.
- Willingness to work onsite at client locations in India for extended periods, and to travel as the work needs.
Nice to have
- Experience with on-premises or private cloud (VPC) deployments.
- Experience with observability and tracing tools such as LangSmith or Braintrust.
- Experience with data engineering and pipelines.
- A history of side projects, open source contributions, or products you shipped end-to-end.
- Experience in embedded or forward-deployed roles before.
- Experience working at a consulting or professional services firm in a client-facing delivery role.
Stack and tools
- Languages: Python and TypeScript.
- Models: Claude and other frontier or open-source models, chosen to fit the customer.
- AI patterns: RAG, agents, prompt engineering, and evaluations.
- Vector and retrieval: vector databases and retrieval pipelines.
- Cloud: AWS, Azure, or GCP, on public or private cloud.
- Integration: REST APIs and enterprise system connectors.
Role: Full Stack GEN AI Engineer
Location: Remote - Bengaluru
Duration: Fulltime With VDart Digital
The role demands a developer who is not just familiar with Large Language Models (LLMs), but is an expert in building autonomous agentic workflows using the modern GenAI stack (LangChain, CrewAI, Vector DBs). Expertise in system design, cloud-native technologies, and CI/CD for AI-driven applications is essential for this high-impact delivery role.
Responsibilities
· Design, develop, and maintain full-stack applications that are scalable, robust, and meet the company's quality standards, with a specific focus on Generative AI integration.
· Agentic Orchestration: Build and deploy sophisticated multi-agent systems and autonomous workflows using frameworks like LangChain, CrewAI, or LangGraph.
· Collaborate effectively with cross-functional teams to define, design, and ship new features that bridge the gap between raw AI power and intuitive user workflows.
· Exhibit strong problem-solving skills with an emphasis on product development and driving architecture choices that enable a world-class user experience.
· Utilize a variety of modern web technologies and frameworks (React.js, Angular, or Vue.js) to build responsive and accessible user interfaces.
· Develop and maintain RESTful APIs and services with optimal performance and scalability, handling streaming AI responses and complex function-calling logic.
· Ensure code quality, organization, and automatization through best practices, including unit tests for prompts, model evaluation pipelines, and automated CI/CD for AI-driven features.
· Implement and optimize RAG (Retrieval-Augmented Generation) pipelines using Vector Databases and advanced retrieval techniques.
· Adapt to emerging technologies and frameworks, specifically new GenAI tools and frontier models, and apply them to operational and business needs.
· Manage individual project priorities, deadlines, and deliverables with minimal supervision.
Skill Requirements
· Excellent oral and written communication skills, with the ability to articulate complex AI and technical ideas to both technical and non-technical audiences.
· Profound knowledge of application development, data structures, networking, operating systems, and DBMS.
· Strong proficiency in backend programming languages for API development such as Python, Java, JavaScript (Node.js), or Go.
· Expertise in front-end technologies and frameworks such as React.js, Angular, or Vue.js.
· GenAI & Agentic Tools: Deep hands-on experience with LangChain, CrewAI, or AutoGen. Ability to manage agent memory, state, and tool-calling.
· In-depth understanding of SQL/NoSQL databases, data modeling, and experience with Vector Databases for RAG implementations.
· Solid grasp of system design, microservices architecture, and cloud-native technologies including Docker, Kubernetes, and GitHub Actions.
· Experience with distributed computing, machine learning frameworks, and tools, with a primary focus on Generative AI.
· Desirable (Good to Have): Experience with Generative or Adaptive UI development, where the interface dynamically adapts or renders components based on LLM outputs and real-time AI reasoning.
· A strong desire to learn and master new technologies and techniques in the rapidly evolving GenAI landscape.
Qualifications
· Bachelor’s degree in Computer Science, Engineering, or a related field.
· A minimum of 3-5 years of experience in full-stack development, with a proven track record of building and deploying Generative AI applications and agents.
· Portfolio of successfully deployed web applications and services, specifically showcasing AI agents, RAG implementations, or complex AI-driven features.
KEY RESPONSIBILITIES:
•Build agents with persistent context & memory
•Design self-learning feedback loops
•Implement RAG pipelines for domain knowledge
•Manage conversation state & orchestration
•Integrate with LLM APIs (OpenAI, Claude, open-source)
Iterate fast — ship daily, measure weekly
MUST-HAVE SKILLS
•Python / TypeScript proficiency
•LangChain, CrewAI, AutoGen or custom frameworks
•Experience with vector DBs (Pinecone, Weaviate, Qdrant)
•Prompt engineering & evaluation pipelines
•Understanding of agent architectures (ReAct, tool-use)
Git, CI/CD, containerization basics






