Cutshort logo
For Employers

Python, Gen AI at VY SYSTEMS PRIVATE LIMITED · Bengaluru (Bangalore) · 4 - 15 years · ₹5L - ₹25L / yr · Profitable · Posted 22 Sep 2026

VY SYSTEMS PRIVATE LIMITED's logo

Python, Gen AI

Banu S's profile picture
Posted by Banu S
4 - 15 yrs
₹5L - ₹25L / yr
Bengaluru (Bangalore)
Skills
skill iconPython
Agentic AI
LangChain

Role Overview

We are looking for a skilled Python Full Stack / Agentic AI Engineer to design, develop, and deploy AI-powered applications and intelligent agentic workflows. The ideal candidate should have strong expertise in Python, FastAPI, LLMs, RAG, LangChain/LangGraph, and modern full-stack development.

You will work on building scalable backend services, integrating Large Language Models, developing AI agents, implementing Retrieval-Augmented Generation (RAG) pipelines, and creating production-ready AI applications.

Key Responsibilities

  • Design and develop scalable backend applications using Python and FastAPI.
  • Build and deploy Agentic AI solutions using LLMs and agent frameworks.
  • Develop multi-step and multi-agent workflows using LangChain and LangGraph.
  • Design and implement RAG (Retrieval-Augmented Generation) pipelines.
  • Integrate LLMs such as OpenAI, Azure OpenAI, Anthropic, Gemini, or open-source models.
  • Develop prompt engineering strategies and structured LLM workflows.
  • Work with vector databases and embedding models for semantic search and knowledge retrieval.
  • Build APIs and microservices for AI-powered applications.
  • Integrate AI services with databases, third-party APIs, and enterprise systems.
  • Develop conversation memory, tool calling, function calling, and agent orchestration capabilities.
  • Implement evaluation, monitoring, logging, guardrails, and error handling for AI applications.
  • Optimize applications for performance, scalability, reliability, and cost.
  • Collaborate with product managers, frontend developers, data engineers, and other stakeholders.
  • Write clean, maintainable, well-tested, and production-ready code.
  • Participate in architecture discussions, code reviews, testing, and deployment activities.

Required Skills

Programming & Backend

  • Strong proficiency in Python.
  • Hands-on experience with FastAPI, REST APIs, and backend development.
  • Strong understanding of asynchronous programming, API design, authentication, and middleware.
  • Experience with SQL/NoSQL databases.

Generative AI / Agentic AI

  • Strong understanding of LLMs and Generative AI.
  • Hands-on experience building AI Agents / Agentic AI applications.
  • Experience with LangChain and/or LangGraph.
  • Knowledge of agent orchestration, tool calling, function calling, memory, and workflow management.
  • Strong understanding of prompt engineering.

RAG

  • Experience designing and implementing RAG architectures.
  • Knowledge of document ingestion, chunking, embeddings, vector search, retrieval, reranking, and response generation.
  • Experience with vector databases such as FAISS, Chroma, Pinecone, Weaviate, Qdrant, or similar.

LLM & AI Integration

  • Experience integrating commercial or open-source LLMs.
  • Understanding of embeddings, context windows, temperature, token usage, and model selection.
  • Experience with structured outputs and LLM-based workflows.
  • Familiarity with LLM evaluation and observability is a plus.

Full Stack

  • Working knowledge of HTML, CSS, JavaScript/TypeScript.
  • Experience with React.js or similar frontend frameworks is preferred.
  • Ability to integrate frontend applications with Python/FastAPI services.
Read more
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos

About VY SYSTEMS PRIVATE LIMITED

Founded :
2000
Type :
Services
Size :
100-1000
Stage :
Profitable

About

Vy Systems is a Global Technology consulting, Solutions, and Managed Technology Services company. We service our customers with ‘RESPONSIVENESS’ as a key factor and we believe that timely response to any transaction increases the operational efficiency and accelerates the revenue and profitability to our customers.

The Company is founded and managed by a team of professionals having more than two+ decades of global experience in the business of Technology Consulting and Services.


Read more

Tech stack

IT consulting

Company social profiles

instagramlinkedintwitterfacebook

Similar jobs (10)

company logo
Shruti mujbaile
Posted by Shruti mujbaile
Gurugram, Pune
6 - 12 yrs
₹8L - ₹25L / yr
Generative AI
Agentic AI
skill iconPython
skill iconMachine Learning (ML)

Location: Pune / Gurgaon

Position: AI Engineer

work mode: WFO


  Job Description.

 Job responsibilities:

  • Responsibility for design, implementation and deployment of Generative AI, Agentic frameworks at scale
  • Strong in programming - Python a
  • Previous experience of working on Computer Vision projects and VLM /VLAM models.
  • In depth awareness of Transformer architectures and End to End Deep neural networks
  • Full stack AI / ML development experience
  • Design, build & maintain efficient and reliable Agentic / Generative AI code leveraging pipelines
  • Hosting and deployment knowledge in GCP or AWS or Azure along with advanced engineering concepts to build user friendly UI interface for easy adoption.


    Requirements:

 ·      4 to 8 years overall years of experience (Agentic AI, Generative AI, VLM, VLAM and LLM) with significant exposure in Development, Architecture design, scaling and hosting in cloud.


    Must Have –

 ·      Architecting and solutioning experience with Python and FAST API, Agentic Ai frameworks, VLMs, VLAMs, Open source LLM’s and Code based LLM models at scale with - Langchain /      Ollama, embeddings, Memory      Management etc.,

·      Practical experience in implementing Explainable and ethical AI models  Practical experience in implementing frameworks like RAG/ CAG/ Self-reflective RAG etc.,

·      Experience in cloud hosting either AWS or Azure or GCP.

·      Experience in ML-OPS - Implement a feedback mechanism to continually improve the model over time through feedback loop and monitoring KPI’s in production.

·      Experience with Quantization and Kubernetes or docker


    Good to have

·      gRPC implementation to expose the API’s on a server for easy usage and good user interface

·      Streamlit front end creation

·      Experience with SAFe framework deliveries.


Read more
company logo
Anish N
Posted by Anish N
Bengaluru (Bangalore)
3 - 5 yrs
₹10L - ₹20L / yr
skill iconPython
Generative AI
Agentic AI
LangChain
LlamaIndex
+3 more

Job Description:

We are looking for a hands-on AI Engineer with experience in Generative AI and Agentic AI to build and deploy production-ready AI solutions.

Key Responsibilities:

  • Develop and deploy GenAI and Agentic AI applications.
  • Build RAG pipelines, LLM workflows, and AI agents.
  • Develop solutions using Python, LangChain, LangGraph, LlamaIndex, or similar frameworks.
  • Implement tool calling, context retrieval, and LLM orchestration.
  • Integrate AI solutions with APIs and cloud platforms.
  • Work with AWS/Azure/GCP, Docker, and CI/CD.

Required Skills:

  • Strong Python programming skills.
  • 3+ years of GenAI/Agentic AI experience.
  • RAG and LLM orchestration.
  • LangChain / LangGraph / LlamaIndex / AutoGen / CrewAI / Semantic Kernel.
  • MCP and A2A knowledge.
  • Cloud, APIs, Docker, and CI/CD experience.

Preferred Experience:

Hands-on experience building and deploying production-ready AI solutions.

Read more
company logo
Faisal AshrafNomani
Posted by Faisal AshrafNomani
Remote only
4 - 8 yrs
Best in industry
Agentic AI

Senior Agentic AI Engineer - (Freelance)

Positions: 2

Experience: Ideally 4(J–(J8 years with strong software-engineering fundamentals and recent hands-on Agentic AI experience.

Mission

Build UC2's governed AI agents capable of reasoning across and interacting safely with enterprise IT systems.

Mandatory capabilities

  • Python
  • LangGraph
  • Agentic AI
  • Tool/function calling
  • Stateful workflows
  • Structured outputs
  • Human-in-the-loop
  • Guardrails
  • Agent state/checkpointing
  • Agent evaluation
  • FastAPI
  • REST APIs
  • Async Python

Retry/timeout/error handling

Highly desirable

MCP, LangChain, Semantic Kernel, agent observability, event-driven architecture and experience integrating AI agents with ServiceNow/Splunk/Confluence or similar enterprise platforms.

The candidate should understand how to engineer:

Read more
One of the largest Paper product Manufacturing Conglomorate
One of the largest Paper product Manufacturing Conglomorate
Agency job
via by praveen somasundaram
Bengaluru (Bangalore)
5 - 6 yrs
₹35L - ₹40L / yr
Generative AI (GenAI)
Agentic AI
Azure OpenAI
skill iconPython
Large Language Models (LLM) tuning
+5 more

Senior Gen AI Full Stack Engineer:

• Strong background in AI/ML and Gen AI with a deep understanding of LLMs, NLP pipelines, and AI model lifecycle.

• Experience in designing and building guardrail systems for Gen AI applications – including prompt filtering, semantic validation, toxicity detection, and hallucination mitigation.

• Fast API experience for API development.

• Proficiency in Python with frameworks like LangChain, Transformers, OpenAI, and LLM orchestration tools.

• Strong DevOps skills including CI/CD, Docker, Kubernetes, and Git.

Experience integrating Gen AI models into enterprise platforms securely and ethically.

Read more
company logo
Meenal Patil
Posted by Meenal Patil
Pune
3 - 4 yrs
₹5L - ₹15L / yr
Agent development
legacy migration
AI Copilot
Claude AI APP

Role: AI Developer

Experience: 3–4 Years

Employment Type: Full-Time

Location: Goregaon, Mumbai


About the Role

We are looking for an experienced AI Developer with 3–4 years of software development experience and strong hands-on exposure to Generative AI, AI Agents, Copilots, and AI-powered application development.

The candidate will be responsible for building production-ready AI solutions, developing agentic workflows, modernizing legacy applications, and integrating LLM capabilities into enterprise applications.


Key Responsibilities

  • Design, develop, and deploy AI Agents and agentic workflows for enterprise use cases.
  • Build AI Copilots and LLM-powered applications using modern AI frameworks and APIs.
  • Develop RAG-based applications using embeddings, vector databases, and enterprise data.
  • Work on legacy application migration and modernization, leveraging AI-assisted development and code transformation techniques.
  • Analyze legacy codebases and design strategies for AI-driven migration, refactoring, and modernization.
  • Integrate LLMs with enterprise applications, APIs, databases, and third-party systems.
  • Implement tool calling, function calling, multi-agent workflows, and workflow automation.
  • Perform prompt engineering, context optimization, model evaluation, and AI application testing.
  • Take ownership of AI solutions from POC and prototyping through production deployment.
  • Collaborate with product managers, architects, and engineering teams to convert business requirements into scalable AI solutions.
  • Stay updated with emerging technologies in Generative AI, Agentic AI, LLMs, and AI-assisted software development.


Required Skills

  • 3–4 years of professional software development experience.
  • Strong proficiency in Python and/or JavaScript/TypeScript.
  • Hands-on experience developing Generative AI / LLM-based applications.
  • Strong understanding of AI Agents, RAG, Prompt Engineering, LLM APIs, and embeddings.
  • Experience with frameworks such as LangChain, LangGraph, Semantic Kernel, AutoGen, or equivalent.
  • Experience working with REST APIs, databases, Git, and cloud environments.
  • Hands-on experience with vector databases such as Pinecone, Weaviate, Chroma, FAISS, or equivalent.
  • Good understanding of software architecture, debugging, testing, and deployment practices.


Good to Have

  • Experience with Microsoft Copilot / Copilot Studio.
  • Experience working with Claude, OpenAI, Gemini, Azure OpenAI, or open-source LLMs.
  • Experience in legacy application migration, modernization, or code conversion.
  • Knowledge of Azure AI / AWS / Google Cloud AI services.
  • Experience with MCP, multi-agent systems, tool calling, and AI orchestration.
  • Experience building enterprise-grade AI solutions with focus on security, scalability, and performance.


Read more
New York, Los Angeles California
3 - 5 yrs
$2.5K - $5.5K / yr
skill iconPython
Artificial Intelligence (AI)
skill iconMachine Learning (ML)
Multi-Agent System
Full Stack Development
+17 more

We are building an advanced, AI-driven multi-agent software system designed to revolutionize task automation and code generation. This is a futuristic AI platform capable of:


✅ Real-time self-coding based on tasks  

✅ Autonomous multi-agent collaboration  

✅ AI-powered decision-making  

✅ Cross-platform compatibility (Desktop, Web, Mobile)  


We are hiring a highly skilled **AI Engineer & Full-Stack Developer** based in India, with a strong background in AI/ML, multi-agent architecture, and scalable, production-grade software development.


### Responsibilities:


- Build and maintain a multi-agent AI system (AutoGPT, BabyAGI, MetaGPT concepts)  

- Integrate large language models (GPT-4o, Claude, open-source LLMs)  

- Develop full-stack components (Backend: Python, FastAPI/Flask, Frontend: React/Next.js)  

- Work on real-time task execution pipelines  

- Build cross-platform apps using Electron or Flutter  

- Implement Redis, Vector databases, scalable APIs  

- Guide the architecture of autonomous, self-coding AI systems  


### Must-Have Skills:


- Python (advanced, AI applications)  

- AI/ML experience, including multi-agent orchestration  

- LLM integration knowledge  

- Full-stack development: React or Next.js  

- Redis, Vector Databases (e.g., Pinecone, FAISS)  

- Real-time applications (websockets, event-driven)  

- Cloud deployment (AWS, GCP)  


### Good to Have:


- Experience with code-generation AI models (Codex, GPT-4o coding abilities)  

- Microservices and secure system design  

- Knowledge of AI for workflow automation and productivity tools  


Join us to work on cutting-edge AI technology that builds the future of autonomous software.

Read more
company logo
Umama Sayed
Posted by Umama Sayed
Remote, Mumbai
2 - 4 yrs
Best in industry
skill iconPython
Large Language Models (LLM)
Generative AI
LangGraph
FastAPI
+7 more

AI Engineer

LLMs, Agents & AI Services

📍 Mumbai (On-site) | Full-time | 2-4 years


About the Role:

Unico Connect is an AI-first technology partner that builds custom mobile, web, and AI products for clients across multiple geographies.

AI is core to how we design, deliver, and scale software for our customers.

We are hiring an AI Engineer for a dedicated client engagement building a complex production AI platform, working on the AI capabilities and agentic features at the core of the product.

The mandatory requirement for this role is at least one AI feature personally shipped to production for real users, with operational ownership.

The role suits someone who thinks quickly on solutioning, can take an ambiguous problem to a working prototype in days, and has the discipline to carry it through to production with predictable economics.

You will work alongside the Senior AI Engineer and the wider pod, with ownership of parts of the AI surface area of the product.


Responsibilities:

Solutioning and POCs

Translate ambiguous customer problems into working POCs at speed.

Pick the right model, framework, and architecture, and demonstrate value early before scaling investment.


LLM Application Development

Build AI features and services using LLM APIs from OpenAI, Anthropic, Google, and self-hosted open-weight models (Llama, Qwen, Mistral).

Choose the right model per use case based on cost, latency, capability, and context-window trade-offs.


Agentic System Design

Design and implement agentic workflows using LangGraph, CrewAI, AutoGen, LlamaIndex Agents, or custom orchestration.

Cover tool use, planning, memory, and multi-step reasoning appropriate to the problem.


API and Service Development

Build production AI services and APIs using Python and FastAPI.

Handle streaming responses, async processing, structured outputs, retries, and graceful degradation when models or tools fail.


Retrieval and Tool Integration

Implement RAG pipelines with vector databases (Pinecone, Weaviate, Qdrant, pgvector, Chroma), embeddings, chunking strategies, hybrid search, and reranking.

Integrate external tools, internal APIs, and document sources through tool-calling and MCP-style patterns.


Cost Analysis and Unit Economics

Model the per-request and per-user cost of every AI feature before it ships.

Track token usage, prompt caching, batching, and model-routing strategies.

Drive measurable improvements in unit economics.


Production Hardening

Add observability and tracing (LangSmith, Langfuse, OpenTelemetry), guardrails, content safety checks, prompt injection defences, and fallback behaviour.


Prompt Engineering and Evaluation

Design, test, and iterate prompts with measured outcomes.

Build evaluation harnesses for accuracy, hallucination, latency, and cost.

Run benchmarks across models and prompt variants before locking in a design.


Requirements:

AI Feature Shipped to Production (Mandatory)

Must have personally built and shipped at least one AI feature that runs in production for real users, with operational ownership.

POCs, internal demos, and one-off scripts do not qualify.


2 to 4 Years of Professional Software or AI Engineering Experience

With at least one production AI feature owned end to end.


Strong Python Proficiency and API Development with FastAPI

Comfort with type hints, async, packaging, testing, streaming responses, and authentication.

Production-grade Python, not notebook-only code.


Hands-on Depth Across the LLM and Agent Stack

Working experience with at least two of OpenAI, Anthropic Claude, Google Gemini, or self-hosted open-weight models (vLLM, Ollama, Together, Replicate).

Working familiarity with at least one agent framework (LangGraph, CrewAI, AutoGen, LlamaIndex Agents) or hand-rolled equivalent.

Working knowledge of RAG, embeddings, and vector databases (Pinecone, Weaviate, Qdrant, pgvector, Chroma).


Solutioning Speed and POC Velocity

Demonstrated ability to move from a fuzzy problem to a working prototype in days.

Strong instinct for what to build first, what to defer, and what to throw away.


Cost Discipline for Production AI

Ability to calculate, monitor, and optimise the cost of LLM APIs, tokens, embeddings, vector store usage, and infrastructure.

Treats unit economics as a first-class concern.


AWS Familiarity

Working knowledge of EC2, S3, IAM, and at least one of Bedrock, SageMaker, or equivalent.


Comfortable in a Fast-Moving Environment

Self-directed, comfortable with ambiguity, takes ownership without being asked, and ships under shifting priorities.


Strong Written and Spoken English Communication

Able to explain trade-offs to non-AI engineers, designers, product managers, and clients in plain language.


Nice to Have

  • fine-tuning or LoRA, QLoRA, PEFT exposure
  • MCP server authoring
  • eval framework experience (LangSmith, Promptfoo, Ragas, DeepEval)
  • open-source AI contributions
  • multi-modal models (vision, audio)
Read more
company logo
Mishika Garg
Posted by Mishika Garg
Bengaluru (Bangalore)
7 - 9 yrs
Best in industry
Artificial Intelligence (AI)
User Interface (UI) Design
skill iconPython

Job Title

Python Full Stack Developer – AI

Experience: 6-9 Years

Location: Bangalore (Hybrid)

Employment Type: Full-Time

Job Summary

We are seeking a highly skilled Python Full Stack Developer with AI expertise to design, develop, and deploy scalable AI-powered applications. The ideal candidate should have strong experience in Python, Full Stack Development, REST APIs, modern frontend frameworks, and Generative AI technologies, including LLMs, prompt engineering, and AI integrations.

The role involves building end-to-end web applications, integrating AI models, and collaborating with cross-functional teams to deliver innovative AI-driven solutions.

Key Responsibilities

· Design, develop, and maintain scalable full-stack applications using Python.

· Build responsive and interactive user interfaces using React.js, Angular, or Vue.js.

· Develop backend services and RESTful APIs using Django, Flask, or FastAPI.

· Integrate Generative AI models such as OpenAI GPT, Claude, Gemini, or Llama into business applications.

· Develop AI-powered chatbots, assistants, document processing, and workflow automation solutions.

· Implement prompt engineering techniques to optimize AI model performance.

· Build Retrieval-Augmented Generation (RAG) pipelines using vector databases.

· Work with LangChain, LangGraph, or LlamaIndex for LLM orchestration.

· Integrate AI APIs and third-party services into enterprise applications.

· Design and optimize SQL and NoSQL databases.

· Deploy applications on AWS, Azure, or GCP.

· Develop CI/CD pipelines and manage deployments using Docker and Kubernetes.

· Write clean, reusable, and well-documented code following best practices.

· Participate in Agile ceremonies, code reviews, and sprint planning.

Required Technical Skills

Backend

· Python

· Django

· Flask

· FastAPI

Frontend

· React.js / Angular / Vue.js

· HTML5

· CSS3

· JavaScript (ES6+)

· TypeScript

AI / Generative AI

· OpenAI API

· Azure OpenAI

· Gemini API

· Claude API

· Llama Models

· LangChain

Databases

· PostgreSQL

· MySQL

· MongoDB

· Redis

Cloud & DevOps

· AWS / Azure / Google Cloud Platform

· Docker

· Kubernetes

· Git

· GitHub

· Jenkins

· CI/CD

API Development

· REST APIs

· GraphQL (Preferred)

· API Integration

Preferred Skills

· Machine Learning fundamentals

· NLP (Natural Language Processing)

· Hugging Face Transformers

· TensorFlow or PyTorch

· Kafka or RabbitMQ

· Elasticsearch

· Microservices Architecture

· Authentication (OAuth2, JWT)

Qualifications

· Bachelor's or Master's degree in Computer Science, Information Technology, or a related field.

· 6–9 years of experience in Python Full Stack Development.

· Hands-on experience with Generative AI and LLM-based application development.

· Experience working in Agile/Scrum environments. 

Read more
Leadsquared
Leadsquared
Agency job
via by Vrishali Mishra
Bengaluru (Bangalore)
2 - 4 yrs
₹25L - ₹45L / yr
Large Language Models (LLM) tuning

About LeadSquared

LeadSquared is a leading sales execution and marketing automation platform trusted by 2,000+ businesses globally, including healthcare, education, financial services, and real estate. Headquartered in Bengaluru with offices across the US, UK, UAE, and Southeast Asia, we empower sales teams to close faster, smarter, and at scale.

Our AI team is at the forefront of integrating cutting-edge large language model capabilities into enterprise workflows — building intelligent agents, copilots, and automation systems that redefine how businesses operate.

Role Overview

We are looking for a Senior AI Engineer with hands-on experience building LLM-powered agents and agentic AI systems. You will design, develop, and deploy autonomous AI pipelines that solve complex, multi-step business problems — from lead qualification and follow-up automation to intelligent CRM workflows and beyond.

This role is ideal for someone who is deeply excited about the frontier of AI, can move fast, and wants their work to directly impact millions of sales professionals worldwide.

Key Responsibilities

Design and build LLM-powered agentic systems using frameworks such as LangChain, LlamaIndex, AutoGen, or CrewAI to automate complex, multi-step workflows.

Develop and maintain Retrieval-Augmented Generation (RAG) pipelines with vector databases (Pinecone, Weaviate, Chroma, pgvector) for domain-specific knowledge grounding.

Build and integrate tool-use and function-calling capabilities into AI agents, enabling dynamic interaction with internal APIs, databases, and third-party services.

Implement prompt engineering strategies including chain-of-thought, few-shot prompting, and structured output parsing to ensure reliable agent behavior.

Design evaluation frameworks and observability pipelines (LangSmith, Helicone, custom metrics) to monitor agent performance, accuracy, and cost.

Collaborate with product, sales, and domain teams to translate business requirements into AI-driven solutions and features.

Optimize LLM inference for latency and cost using techniques like caching, model distillation, quantization, and batching.

Stay current with the rapidly evolving LLM ecosystem and proactively propose improvements and new approaches.

Contribute to internal best practices, documentation, and knowledge-sharing across the engineering org.

Required Qualifications

Experience

2–4 years of professional software engineering experience, with at least 1–2 years focused on LLM/AI systems.

Proven experience shipping LLM-based products or agentic AI systems into production environments.

Technical Skills

Strong proficiency in Python and familiarity with async programming patterns for AI pipelines.

Hands-on experience with LLM APIs: OpenAI (GPT-4o), Anthropic (Claude), Google (Gemini), or open-source models (Llama, Mistral).

Experience with agentic frameworks: LangChain, LangGraph, LlamaIndex, AutoGen, CrewAI, or similar.

Solid understanding of RAG architectures, embedding models, and semantic search.

Experience with vector databases and similarity search infrastructure.

Knowledge of REST APIs, microservices architecture, and containerization (Docker/Kubernetes).

Problem-Solving & Mindset

Strong ability to decompose ambiguous, open-ended problems into structured AI system designs.

Experience with prompt debugging, LLM evaluation, and iterative refinement workflows.

Ability to balance research exploration with engineering pragmatism to ship reliable systems.

Preferred Qualifications

Experience with multi-agent orchestration and agent memory systems (short-term and long-term).

Familiarity with fine-tuning or RLHF workflows for domain adaptation.

Background in NLP, information retrieval, or conversational AI.

Prior experience in B2B SaaS or CRM domain is a plus.

Contributions to open-source AI/ML projects or published research/blogs.

Experience with cloud platforms: AWS, GCP, or Azure — particularly AI/ML services

Read more
company logo
Bhawna Khemani
Posted by Bhawna Khemani
Bengaluru (Bangalore), Delhi, Gurugram, Noida, Ghaziabad, Faridabad
4 - 13 yrs
₹11L - ₹35L / yr
Generative AI
Large Language Models (LLM)
Retrieval Augmented Generation (RAG)
skill iconPython
LlamaIndex
+4 more

Generative AI Engineer 

Role Overview:

You will be responsible for the hands-on development, coding, and deployment of AI-powered features. Your focus is on writing clean, efficient code to integrate LLMs into our existing tech stack, building robust data pipelines for RAG, and ensuring the reliability of model outputs through rigorous testing and optimization.

Key Responsibilities

  • Application Implementation: Code and integrate LLM APIs (OpenAI, Anthropic, etc.) or local models into backend services using Python, FastAPI, etc.,
  • MCP Server Development: Design and implement custom MCP servers using the official SDKs (Python/TypeScript) to expose internal databases, APIs, and file systems to AI agents.
  • RAG Implementation: Build and maintain the "plumbing" for Retrieval-Augmented Generation—specifically coding the data ingestion scripts, text chunking logic, and metadata filtering.
  • Vector DB Management: Perform day-to-day operations on vector databases (Pinecone, Milvus, etc.), including indexing, querying, and optimizing search retrieval.
  • Prompt Programming: Develop, version-control, and refine complex prompt templates (using Jinja2 or similar) to ensure consistent structured outputs (JSON/YAML).
  • Agent Development: Implement multi-step workflows using LangChain, LangGraph, CrewAI etc.,, focusing on tool-calling logic and error handling.
  • Evaluation & Testing: Build automated test suites to detect "hallucinations" and measure accuracy using frameworks.
  • Performance Tuning: Implement caching layers and streaming responses to reduce latency and improve the end-user experience; Token optimization.
  • Data Pre-processing: Clean and tokenize datasets for model fine-tuning or high-quality context retrieval.

Technical Skills (The "Execution" Stack)

  • Language: Advanced Python (Asyncio, Pydantic) and optional TypeScript/Node.js (for full-stack integration).
  • AI Frameworks: Hands-on experience with any of LangChain, LlamaIndex, and Hugging Face Transformers. RAG and Vector search concepts.
  • Data Handling: Proficiency in SQL and handling unstructured data formats (PDFs, Markdown, JSON).
  • Deployment: Practical experience with Docker, GitHub Actions (CI/CD), and experience with OpenTelemetry, LangSmith, Weights & Biases etc., Understanding of evaluation/guardrails.
  • MCP/API Proficiency: Deep understanding of RESTful APIs, Streaming HTTP, MCP server vs client, JSONRPC
Read more
Why apply to jobs via Cutshort
people_solving_puzzle
Personalized job matches
Stop wasting time. Get matched with jobs that meet your skills, aspirations and preferences.
people_verifying_people
Verified hiring teams
See actual hiring teams, find common social connections or connect with them directly.
ai_chip
Move faster with AI
We use AI to get you faster responses, recommendations and unmatched user experience.
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos