Cutshort logo
For Employers

Gen AI Engineer at NeoGenCode Technologies Pvt Ltd Ā· Bengaluru (Bangalore) Ā· 3 - 8 years Ā· ₹10L - ₹24L / yr Ā· Raised funding Ā· Posted 27 May 2026

NeoGenCode Technologies Pvt Ltd's logo

Gen AI Engineer

Akshay Patil's profile picture
Posted by Akshay Patil
3 - 8 yrs
₹10L - ₹24L / yr
Bengaluru (Bangalore)
Skills
skill iconPython
Large Language Models (LLM)
Generative AI
AI Agents
LangChain
LlamaIndex
Semantic Kernel
Prompt engineering
Systems design
Vector database
Embeddings
RESTful APIs

šŸš€ Job Title : Gen AI Engineer

Experience : 3 to 5 Years

Location : Bengaluru (MG Road – Prestige Building)

Work Mode : Hybrid (3 Days WFO)

Open Positions : 2

Notice Period : Immediate to 15–20 Days Preferred


šŸŽÆ Role Overview :

We are looking for a Gen AI Engineer with hands-on experience in building and deploying LLM powered applications.

You will work on cutting-edge AI solutions, including real-world enterprise use cases and next-generation internal products.


šŸ›  Mandatory Skills :

  • Strong proficiency in Python.
  • Hands-on experience with LLMs & GenAI frameworks (LangChain, LlamaIndex, Semantic Kernel).
  • Experience in prompt engineering and system design.
  • Knowledge of vector databases & embeddings.
  • Experience integrating GenAI solutions into production systems.
  • Understanding of REST APIs, async processing, and streaming responses.

⚔ AI / ML Knowledge :

  • Strong understanding of NLP & transformer-based models.
  • Familiarity with fine-tuning, embeddings, and inference patterns.
  • Knowledge of GenAI evaluation metrics (accuracy, relevance, grounding).

☁ Infrastructure & Tooling :

  • Experience with Cloud platforms (AWS / Azure / GCP).
  • Familiarity with Docker & Kubernetes (good to have).
  • Exposure to CI/CD pipelines and MLOps practices.


🌟 Nice to Have :

  • Experience with multimodal models (vision, OCR, speech).
  • Knowledge of AIOps, RCA, observability, enterprise workflows.
  • Experience building AI agents & orchestration layers.
  • Understanding of AI governance, safety, and responsible AI.


šŸ’¼ Key Responsibilities :

  • Design, develop, and deploy GenAI applications using LLMs (OpenAI, Azure OpenAI, open-source models).
  • Build RAG pipelines using vector databases (FAISS, Pinecone, Weaviate, Chroma).
  • Develop high-quality prompts, system instructions, and structured outputs (JSON, function calling).
  • Integrate GenAI capabilities into backend systems via APIs & microservices.
  • Optimize models for performance, latency, cost, and accuracy.
  • Implement evaluation frameworks (hallucination detection, confidence scoring).
  • Ensure data security, privacy, and compliance.
  • Collaborate with product, design, and domain teams.
  • Document architecture, prompts, and best practices.


šŸ’¼šŸ¤ Interview Process :

1. Geektrust Assessment (AI Agent-based evaluation)

2. Final Interview with Founder (45 mins)

Read more
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos

About NeoGenCode Technologies Pvt Ltd

Founded :
2023
Type :
Services
Size
Stage :
Raised funding

About

Welcome to Neogencode Technologies, an IT services and consulting firm that provides innovative solutions to help businesses achieve their goals. Our team of experienced professionals is committed to providing tailored services to meet the specific needs of each client. Our comprehensive range of services includes software development, web design and development, mobile app development, cloud computing, cybersecurity, digital marketing, and skilled resource acquisition. We specialize in helping our clients find the right skilled resources to meet their unique business needs. At Neogencode Technologies, we prioritize communication and collaboration with our clients, striving to understand their unique challenges and provide customized solutions that exceed their expectations. We value long-term partnerships with our clients and are committed to delivering exceptional service at every stage of the engagement. Whether you are a small business looking to improve your processes or a large enterprise seeking to stay ahead of the competition, Neogencode Technologies has the expertise and experience to help you succeed. Contact us today to learn more about how we can support your business growth and provide skilled resources to meet your business needs.

Read more

Candid answers by the company

What does the company do?
What is the location preference of jobs?

IT & Engineering Talent Staffing

  • Provides full-time and contract-based hiring, delivering handpicked, pre‑screened developers across tech stacks—ranging from web, mobile, AI/ML, Web3/blockchain.
  • Maintains a bench o vetted candidates, offering fast delivery of interview-ready profiles—often within 24 hours.
  • Offers payroll management, handling compliance, tax, attendance, and documentation for both contractors and full-time employees.

2. End-to-End Project Delivery

  • Delivers full-stack development solutions: web, mobile, cloud, AI/ML, Blockchain/Web3.
  • Manages entire project lifecycle—requirements gathering, design (UI/UX), development, deployment, and ongoing support .

3. Additional Offerings

  • Expands into cybersecurity consulting, digital marketing, and cloud platform services (like AWS, GCP, Azure) .
  • Provides strategic IT consulting to align technology solutions with business objectives

Company social profiles

bloginstagramlinkedinfacebook

Similar jobs (10)

company logo
Mayank Choudhary
Posted by Mayank Choudhary
Bengaluru (Bangalore)
3 - 5 yrs
₹20L - ₹25L / yr
Artificial Intelligence (AI)

Strong AI/ML Engineer Profile

Mandatory (Experience) : Must have 3+ years of experience in software engineering with atleast 1+ years in GenAI application development and production deployment

Mandatory (GenAI Application Development): Must have proven experience building GenAI applications covering RAG pipelines, multi-agent systems, Text2SQL, and fine-tuning

Mandatory (Production GenAI Deployment): Must have expertise deploying production-grade GenAI applications including model evaluation, optimisation, and ownership of full production rollouts

Mandatory (ML & Data Science Tooling): Must have strong hands-on experience with core ML and data science tools including pandas, scikit-learn, and PyTorch

Mandatory (Cloud ML Infrastructure): Must have experience building and deploying production-grade ML workloads on at least one of AWS, Azure, or GCP

Mandatory (Communication): Must have strong English communication skills with the ability to work across time zones and collaborate cross-functionally with product, engineering, and business stakeholders

Mandatory (Note 1) : Role is Hybrid, WFH flexibility as well upto 6 days a month

Mandatory (Note 2) : CTC is inclusive of 10% variable

Mandatory (Note 3): Candidates should be available to join within May 31st or June first week max

Read more
company logo
Prithisha Kathiresan
Posted by Prithisha Kathiresan
Bengaluru (Bangalore)
3 - 8 yrs
Best in industry
Generative AI (GenAI)
Large Language Models (LLM) tuning
Retrieval Augmented Generation (RAG)
Azure OpenAI
skill iconAmazon Web Services (AWS)
+2 more

Senior Generative AI Engineer

Employment Type: Permanent with VDart Digital

Work Location: Marathalli, Bengaluru

Job Description

We are seeking a highly skilled Senior Generative AI Engineer with strong expertise in designing, developing, and deploying enterprise-scale AI solutions using Large Language Models (LLMs) and modern Generative AI frameworks. The ideal candidate should have hands-on production experience building scalable GenAI applications, AI agents, autonomous workflows, and Retrieval-Augmented Generation (RAG) systems in cloud-native environments.

This role requires deep technical expertise in LLM orchestration, AI application architecture, prompt engineering, vector databases, MLOps, and production deployment of AI systems. Candidates should have proven experience delivering real-world AI solutions in enterprise environments with strong exposure to cloud platforms and DevOps practices.

Key Responsibilities

  • Design, build, and deploy enterprise-grade Generative AI applications using Large Language Models (LLMs).
  • Develop intelligent AI agents and autonomous workflows using frameworks such as LangChain, CrewAI, LangGraph, AutoGen, or similar agentic AI frameworks.
  • Implement and optimize Retrieval-Augmented Generation (RAG) pipelines using vector databases and semantic search technologies.
  • Work extensively on prompt engineering, tool calling, memory management, agent orchestration, and multi-agent systems.
  • Integrate and manage LLMs such as OpenAI, Azure OpenAI, Claude, Gemini, Llama, Mistral, or similar foundation models.
  • Develop scalable AI services and APIs using Python and FastAPI.
  • Build production-ready AI solutions with high availability, scalability, monitoring, and observability.
  • Deploy and manage AI applications in cloud-native environments using Docker and Kubernetes.
  • Collaborate with Data Science, ML Engineering, and DevOps teams to operationalize AI solutions.
  • Implement CI/CD pipelines and automated deployment processes for AI workloads.
  • Monitor model performance, latency, reliability, and operational efficiency in production environments.
  • Ensure AI solutions follow enterprise security, governance, and responsible AI standards.
  • Evaluate and adopt emerging Generative AI tools, frameworks, and models.

Required Skills

Generative AI & LLM Expertise

  • Strong hands-on experience with Generative AI and Large Language Models (LLMs).
  • Production-level experience building and deploying GenAI applications.
  • Expertise in LangChain, CrewAI, LangGraph, AutoGen, or similar frameworks.
  • Experience with AI agents, autonomous workflows, and multi-agent architectures.
  • Strong understanding of prompt engineering, embeddings, model evaluation, and LLM orchestration.
  • Experience integrating OpenAI, Azure OpenAI, Claude, Gemini, Llama, Mistral, or similar models.

RAG & Vector Databases

  • Strong experience implementing RAG pipelines and semantic retrieval systems.
  • Experience with vector databases such as Pinecone, Weaviate, ChromaDB, FAISS, or Milvus.
  • Understanding of chunking strategies, embeddings, indexing, reranking, and retrieval optimization.

Python & AI Development

  • Strong proficiency in Python.
  • Experience with FastAPI for AI service and API development.
  • Experience with AI/ML libraries and data processing tools such as Pandas and NumPy.

Cloud & Production Deployment

  • Mandatory production experience on at least one cloud platform:
  • Microsoft Azure
  • Experience deploying scalable AI applications in enterprise production environments.
  • Hands-on experience with Docker, Kubernetes, Jenkins, Terraform, and CI/CD pipelines.
  • Strong understanding of MLOps, AI deployment lifecycle, monitoring, and observability.

Engineering & Operational Excellence

  • Strong understanding of software engineering best practices.
  • Experience with Git, version control, automated testing, and release management.
  • Experience building secure, scalable, and high-performance AI solutions.
  • Ability to troubleshoot production AI systems and optimize performance.

Preferred Skills

  • Experience with AI observability and evaluation frameworks.
  • Exposure to fine-tuning, PEFT, LoRA, or model optimization techniques.
  • Experience with enterprise AI governance and responsible AI practices.
  • Knowledge of distributed AI systems and scalable inference architectures.
  • Familiarity with AI security and compliance standards.

Qualifications

  • Bachelor’s or Master’s degree in Computer Science, Artificial Intelligence, Data Science, or a related field.
  • 3–8 years of overall software engineering experience.
  • Minimum 3+ years of hands-on experience in Generative AI and LLM-based application development,
  • Proven track record of delivering enterprise-scale AI solutions in production environments.
  • Strong communication and stakeholder management skills.
Read more
One of the largest Paper product Manufacturing Conglomorate
One of the largest Paper product Manufacturing Conglomorate
Agency job
via by praveen somasundaram
Bengaluru (Bangalore)
5 - 6 yrs
₹35L - ₹40L / yr
Generative AI (GenAI)
Agentic AI
Azure OpenAI
skill iconPython
Large Language Models (LLM) tuning
+5 more

Senior Gen AI Full Stack Engineer:

• Strong background in AI/ML and Gen AI with a deep understanding of LLMs, NLP pipelines, and AI model lifecycle.

• Experience in designing and building guardrail systems for Gen AI applications – including prompt filtering, semantic validation, toxicity detection, and hallucination mitigation.

• Fast API experience for API development.

• Proficiency in Python with frameworks like LangChain, Transformers, OpenAI, and LLM orchestration tools.

• Strong DevOps skills including CI/CD, Docker, Kubernetes, and Git.

Experience integrating Gen AI models into enterprise platforms securely and ethically.

Read more
company logo
Shruti mujbaile
Posted by Shruti mujbaile
Gurugram, Pune
6 - 12 yrs
₹8L - ₹25L / yr
Generative AI
Agentic AI
skill iconPython
skill iconMachine Learning (ML)

Location: Pune / Gurgaon

Position: AI Engineer

work mode: WFO


Ā Ā Job Description.

​

Ā Job responsibilities:

  • Responsibility for design, implementation and deployment of Generative AI,Ā Agentic frameworks at scale
  • Strong in programming - Python a
  • Previous experience of working on Computer Vision projects and VLM /VLAM models.
  • In depth awareness of Transformer architectures and End to End Deep neural networks
  • Full stack AI / ML development experience
  • Design, build & maintain efficient and reliable Agentic / Generative AI code leveraging pipelines
  • Hosting and deployment knowledge in GCP or AWS or Azure along with advanced engineering concepts to build user friendly UI interface for easy adoption.


Ā Ā Ā Ā Requirements:

Ā Ā·Ā Ā Ā Ā Ā Ā 4 to 8 years overall years of experience (Agentic AI, Generative AI, VLM, VLAM and LLM) with significant exposure in Development, Architecture design, scaling and hosting in cloud.


Ā Ā Ā Ā Must Have –

Ā Ā·Ā Ā Ā Ā Ā Ā Architecting and solutioning experience with Python and FAST API, Agentic Ai frameworks, VLMs, VLAMs, Open source LLM’s and Code based LLM models at scale with - Langchain /Ā Ā Ā Ā Ā Ā Ollama, embeddings, MemoryĀ Ā Ā Ā Ā Ā Management etc.,

Ā·Ā Ā Ā Ā Ā Ā Practical experience in implementing Explainable and ethical AI modelsĀ Ā Practical experience in implementing frameworks like RAG/ CAG/ Self-reflective RAG etc.,

Ā·Ā Ā Ā Ā Ā Ā Experience in cloud hosting either AWS or Azure or GCP.

Ā·Ā Ā Ā Ā Ā Ā Experience in ML-OPS - Implement a feedback mechanism to continually improve the model over time through feedback loop and monitoring KPI’s in production.

Ā·Ā Ā Ā Ā Ā Ā Experience with Quantization and Kubernetes or docker


Ā Ā Ā Ā Good to have

Ā·Ā Ā Ā Ā Ā Ā gRPC implementation to expose the API’s on a server for easy usage and good user interface

Ā·Ā Ā Ā Ā Ā Ā Streamlit front end creation

Ā·Ā Ā Ā Ā Ā Ā Experience with SAFe framework deliveries.


Read more
company logo
Agency job
via by Naveen Balne
Hyderabad, Bengaluru (Bangalore), Pune, Chennai, Kolkata
5 - 13 yrs
₹15L - ₹40L / yr
Artificial Intelligence (AI)
Generative AI
Generative AI (GenAI)
skill iconPython
Large Language Models (LLM)
+7 more

We are seeking Generative AI Developers with strong Python programming and AI/ML expertise to build, deploy, and optimize LLM-powered applications. The role involves developing RAG solutions, AI agents, and enterprise GenAI applications while collaborating with cross-functional teams.


Key Responsibilities


  • Develop and enhance Generative AI applications using LLMs and AI frameworks.
  • Build and optimize RAG pipelines, vector search, and AI-powered workflows.
  • Design effective prompts and fine-tune models using techniques such as LoRA and QLoRA.
  • Develop REST APIs and integrate AI capabilities into enterprise applications.
  • Deploy, monitor, and maintain AI solutions in cloud and containerized environments.
  • Ensure code quality through testing, debugging, documentation, and code reviews.
  • Follow Responsible AI, security, and data governance practices.


Required Technical Skills


  • Strong proficiency in Python, OOP, APIs, debugging, and software development best practices.
  • Good understanding of Data Structures & Algorithms, complexity analysis, and problem-solving.
  • Hands-on experience with LLMs, Prompt Engineering, RAG, AI Agents, and embeddings.
  • Experience with LangChain, LangGraph, LlamaIndex, Hugging Face, or similar frameworks.
  • Knowledge of vector databases, semantic/hybrid search, and retrieval architectures.
  • Experience with PyTorch, TensorFlow, or Keras.
  • Familiarity with Docker, Git, CI/CD, and cloud platforms (Azure/AWS/GCP).
  • Understanding of AI governance, data privacy, and Responsible AI principles.


Preferred Skills


  • Experience with Agentic AI frameworks (CrewAI, AutoGen, Semantic Kernel).
  • Exposure to Azure AI Foundry, Databricks, or enterprise AI platforms.
  • Knowledge of multimodal AI applications.


Qualifications


  • Bachelor's or Master's degree in Computer Science, AI, Data Science, or a related field.
  • 5 years of software development experience, including AI/ML or Generative AI projects.
  • Experience building and deploying production-grade AI solutions.

Assessment Focus Areas


Candidates will be evaluated on:

  • Python coding and problem-solving
  • Data Structures & Algorithms
  • LLMs, RAG, and Agentic AI concepts
  • API development and system design
  • Cloud deployment and AI solution architecture
Read more
company logo
Anish N
Posted by Anish N
Bengaluru (Bangalore)
3 - 5 yrs
₹10L - ₹20L / yr
skill iconPython
Generative AI
Agentic AI
LangChain
LlamaIndex
+3 more

Job Description:

We are looking for a hands-onĀ AI EngineerĀ with experience inĀ Generative AI and Agentic AIĀ to build and deploy production-ready AI solutions.

Key Responsibilities:

  • Develop and deploy GenAI and Agentic AI applications.
  • BuildĀ RAG pipelines, LLM workflows, and AI agents.
  • Develop solutions usingĀ Python, LangChain, LangGraph, LlamaIndex, or similar frameworks.
  • Implement tool calling, context retrieval, and LLM orchestration.
  • Integrate AI solutions with APIs and cloud platforms.
  • Work withĀ AWS/Azure/GCP, Docker, and CI/CD.

Required Skills:

  • Strong Python programming skills.
  • 3+ years of GenAI/Agentic AI experience.
  • RAG and LLM orchestration.
  • LangChain / LangGraph / LlamaIndex / AutoGen / CrewAI / Semantic Kernel.
  • MCP and A2A knowledge.
  • Cloud, APIs, Docker, and CI/CD experience.

Preferred Experience:

Hands-on experience building and deployingĀ production-ready AI solutions.

Read more
company logo
Anishka Burde
Posted by Anishka Burde
Mumbai
6 - 13 yrs
Best in industry
skill iconJava
skill iconPython
LangGraph
LangChain
Retrieval Augmented Generation (RAG)
+2 more

Job Summary:

Wissen Technology is hiring an AI Implementation Engineer to build, deploy, and scale enterprise-grade Generative AI solutions across business-critical applications. The role involves developing production-ready AI systems using Azure AI services, Large Language Models (LLMs), RAG architectures, and agent-based frameworks while collaborating closely with engineering teams to drive AI adoption and innovation.


Experience

6-12 years


Location

Mumbai / Bangalore


Mode of Work

Hybrid


Mandatory Skills (Must Have)

• Python programming (6+ years) including asynchronous programming and backend application development

• Java and Spring Framework (3+ years) for enterprise-scale application integration

• Azure OpenAI Service, Azure AI Foundry, and Azure AI Search for production GenAI applications

• Retrieval Augmented Generation (RAG) architecture including embeddings, chunking, vector databases, reranking, and grounding techniques

• Agent Frameworks such as Microsoft Agent Framework, Semantic Kernel, AutoGen, LangChain, or LangGraph

• Snowflake and Cortex AI (Cortex Search, LLM Functions) with strong SQL expertise

• Prompt Engineering, LLM evaluation frameworks, testing, and model performance optimization

• DevOps and Cloud Deployment using Azure DevOps, GitHub Actions, Docker, AKS, Azure Functions, and observability tools


Optional Skills (Good to Have)

• React.js for AI-powered user interfaces and conversational applications

• Azure AI Content Safety and Responsible AI implementation experience

• Financial Services, Banking, or other regulated industry domain experience

• Real-time streaming applications and token-level LLM operations

• Performance optimization, caching strategies, and cost optimization for AI workloads

• Microsoft Azure AI Engineer Associate Certification

Read more
company logo
HR  GYTWorkz
Posted by HR GYTWorkz
Hyderabad
2 - 6 yrs
₹10L - ₹40L / yr
Retrieval Augmented Generation (RAG)
LLM Evaluation Frameworks
Model Context Protocol (MCP)
Large Language Models (LLM) tuning
Fine-tuning LLMs
+6 more

Design and develop Agentic AI systems using LLMs, tools, memory,

workflows, and MCP.

Build production-grade RAG pipelines, including ingestion, chunking,

embeddings, retrieval, reranking, and evaluation.

Implement context engineering strategies for improving LLM accuracy,

relevance, and reliability.

Develop and integrate MCP-based tools and services for AI agents.

Work with LLMs, SLMs, quantized models, and model optimization

techniques for efficient inference.

Develop scalable backend services and APIs for AI applications.

Design databases and data models supporting AI/agentic applications.

Implement AI observability covering latency, token usage, cost, failures,

quality, and agent/tool execution.

Apply AI governance and responsible AI practices, including security,

access control, data privacy, and auditability.

Optimize AI systems for latency, scalability, cost, and reliability.

Collaborate with engineering and product teams to take AI solutions from

POC to production.

Strong hands-on experience with GenAI, LLMs, and Agentic AI.

Experience building RAG applications.

Strong understanding of Context Engineering and prompt/context

optimization.

Role Overview

We are looking for a hands-on AI/ML Engineer to design, develop, and deploy

production-ready GenAI and Agentic AI applications. The role involves building

intelligent agents, RAG pipelines, AI APIs, backend services, and scalable AI

infrastructure with a strong focus on context engineering, observability,

governance, and model optimisation.

Key Responsibilities

Required Skills

Practical experience with MCP (Model Context Protocol).

Experience with frameworks such as LangChain, LangGraph,

LlamaIndex, or equivalent.

Knowledge of LLM/SLM deployment and quantization techniques.

Strong Python backend development experience.

Experience developing REST APIs using FastAPI/Flask or equivalent.

Strong understanding of SQL/NoSQL databases and database design.

Experience with vector databases such as Qdrant, Pinecone, Weaviate,

ChromaDB, or FAISS.

Understanding of AI observability, evaluation, monitoring, and

governance.

Experience with cloud platforms and production deployment is preferred.

Strong understanding of software engineering principles, Git, testing, and

CI/CD.

Read more
company logo
Meenal Patil
Posted by Meenal Patil
Pune
3 - 4 yrs
₹5L - ₹15L / yr
Agent development
legacy migration
AI Copilot
Claude AI APP

Role: AI Developer

Experience: 3–4 Years

Employment Type: Full-Time

Location: Goregaon, Mumbai


About the Role

We are looking for an experienced AI Developer with 3–4 years of software development experience and strong hands-on exposure to Generative AI, AI Agents, Copilots, and AI-powered application development.

The candidate will be responsible for building production-ready AI solutions, developing agentic workflows, modernizing legacy applications, and integrating LLM capabilities into enterprise applications.


Key Responsibilities

  • Design, develop, and deploy AI Agents and agentic workflows for enterprise use cases.
  • Build AI Copilots and LLM-powered applications using modern AI frameworks and APIs.
  • Develop RAG-based applications using embeddings, vector databases, and enterprise data.
  • Work on legacy application migration and modernization, leveraging AI-assisted development and code transformation techniques.
  • Analyze legacy codebases and design strategies for AI-driven migration, refactoring, and modernization.
  • Integrate LLMs with enterprise applications, APIs, databases, and third-party systems.
  • Implement tool calling, function calling, multi-agent workflows, and workflow automation.
  • Perform prompt engineering, context optimization, model evaluation, and AI application testing.
  • Take ownership of AI solutions from POC and prototyping through production deployment.
  • Collaborate with product managers, architects, and engineering teams to convert business requirements into scalable AI solutions.
  • Stay updated with emerging technologies in Generative AI, Agentic AI, LLMs, and AI-assisted software development.


Required Skills

  • 3–4 years of professional software development experience.
  • Strong proficiency in Python and/or JavaScript/TypeScript.
  • Hands-on experience developing Generative AI / LLM-based applications.
  • Strong understanding of AI Agents, RAG, Prompt Engineering, LLM APIs, and embeddings.
  • Experience with frameworks such as LangChain, LangGraph, Semantic Kernel, AutoGen, or equivalent.
  • Experience working with REST APIs, databases, Git, and cloud environments.
  • Hands-on experience with vector databases such as Pinecone, Weaviate, Chroma, FAISS, or equivalent.
  • Good understanding of software architecture, debugging, testing, and deployment practices.


Good to Have

  • Experience with Microsoft Copilot / Copilot Studio.
  • Experience working with Claude, OpenAI, Gemini, Azure OpenAI, or open-source LLMs.
  • Experience in legacy application migration, modernization, or code conversion.
  • Knowledge of Azure AI / AWS / Google Cloud AI services.
  • Experience with MCP, multi-agent systems, tool calling, and AI orchestration.
  • Experience building enterprise-grade AI solutions with focus on security, scalability, and performance.


Read more
company logo
Nirmala Lama
Posted by Nirmala Lama
Mumbai
1 - 5 yrs
₹18L - ₹23L / yr
Generative AI
Retrieval Augmented Generation (RAG)
LangGraph
LangChain
Large Language Models (LLM)
+2 more

About the role

We are seeking an AI Engineer to build and implement AI systems for content production at scale. You'll work at the intersection of engineering and contentĀ designing prompt pipelines, integrating generative models, and building the tooling that turns source material into finished creative output. The ideal candidate is technically strong but also has taste: someone who understands story and craft, and can tell the difference between output that's technically correct and output that's actually good.


Responsibilities

  • Build and iterate on prompt pipelines and multi-agent workflow components
  • Design and integrate agentic workflowsĀ orchestrate multi-step, tool-using agents that plan, call models, and hand off between stages in production
  • Deploy and serve open-source models set up inference endpoints, manage GPU compute, and optimize for latency and cost
  • Write evals compare outputs against references, quantify quality, and feed results back into the pipeline
  • Work on data pipelines: structured extraction from messy source text, localization, similarity/dedup
  • Debug and maintain pipeline stages in production


What you bring:Ā Ā 

  • (1+/3+) years of engineering experience, or a strong portfolio of shipped projects
  • Solid Python fundamentals clean, working, readable code
  • Hands-on experience with LLM APIs and prompt engineering (personal projects count)
  • Comfort with Git, REST APIs, and working in a Linux environment
  • A feel for content and narrative you can judge whether generated output is actually good, not just valid
  • Curiosity and clear communication you ask good questions and don't stay stuck silently


Ā Preferred

  • Exposure to agent/orchestration frameworks (LangGraph, LangChain, CrewAI)
  • Familiarity with vector databases, embeddings, or RAG (Qdrant, pgvector)
  • Hands-on work with open-source generative media models Flux, LTX, Wan, or similar
  • Experience deploying open-source models for inference (vLLM, ComfyUI, Replicate/Cog, Docker + GPU)
  • Experience writing evals or LLM-as-judge scoring
  • Node.js and Fastapi familiarity, or experience deploying on AWS


Read more
Why apply to jobs via Cutshort
people_solving_puzzle
Personalized job matches
Stop wasting time. Get matched with jobs that meet your skills, aspirations and preferences.
people_verifying_people
Verified hiring teams
See actual hiring teams, find common social connections or connect with them directly.
ai_chip
Move faster with AI
We use AI to get you faster responses, recommendations and unmatched user experience.
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos