Cutshort logo
For Employers
httpswwwicloudemscomvlog logo
Junior Software Engineer – AI/ML
Junior Software Engineer – AI/ML

Junior Software Engineer – AI/ML at httpswwwicloudemscomvlog · Remote only · 2 - 4 years · ₹3L - ₹6L / yr · Profitable · Remote only · Posted 24 Sep 2026

httpswwwicloudemscomvlog's logo

Junior Software Engineer – AI/ML

AMISHA SRIVASTAVA's profile picture
Posted by AMISHA SRIVASTAVA
2 - 4 yrs
₹3L - ₹6L / yr
Remote only
Skills
skill iconPython
Artificial Intelligence (AI)
skill iconMachine Learning (ML)
Generative AI
Large Language Models (LLM)
AI Agents

Software Engineer – AI/ML to build next-generation AI solutions for the Higher Education domain. You will work on LLM-powered applications, AI Agents, RAG systems, and intelligent automation for our enterprise SaaS platform.

Required Skills

  • Strong Python programming and software engineering skills
  • Hands-on experience with LLMs (GPT, Claude, Gemini, Llama, Mistral, DeepSeek)
  • Experience with LangChain, LlamaIndex, AI Agents, and RAG
  • Knowledge of PyTorch/TensorFlow and Hugging Face
  • Experience with FastAPI, REST APIs, and Microservices
  • Hands-on experience with Vector Databases (Pinecone, Qdrant, ChromaDB, pgvector)
  • Experience with AWS (Bedrock, SageMaker, EC2, Lambda, S3)
  • Docker & Kubernetes
  • Strong understanding of System Design and scalable AI applications

Preferred Skills

  • Agentic AI & MCP (Model Context Protocol)
  • Prompt Engineering & Fine-tuning
  • Enterprise AI application development
  • Experience in EdTech or SaaS products

What You'll Build

  • AI Agents and intelligent workflows
  • RAG-based enterprise search solutions
  • LLM-powered applications
  • Scalable AI APIs and microservices
  • AI features for enterprise SaaS products

If you're passionate about building production-grade AI systems and working with the latest Generative AI technologies, we'd love to hear from you.

Read more
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos

About httpswwwicloudemscomvlog

Founded :
2017
Type :
Product
Size :
100-500
Stage :
Profitable

About

Vlog Videos
Read more

Company social profiles

bloglinkedintwitterfacebook

Similar jobs (10)

company logo
Prithisha Kathiresan
Posted by Prithisha Kathiresan
Bengaluru (Bangalore)
3 - 8 yrs
Best in industry
Generative AI (GenAI)
Large Language Models (LLM) tuning
Retrieval Augmented Generation (RAG)
Azure OpenAI
skill iconAmazon Web Services (AWS)
+2 more

Senior Generative AI Engineer

Employment Type: Permanent with VDart Digital

Work Location: Marathalli, Bengaluru

Job Description

We are seeking a highly skilled Senior Generative AI Engineer with strong expertise in designing, developing, and deploying enterprise-scale AI solutions using Large Language Models (LLMs) and modern Generative AI frameworks. The ideal candidate should have hands-on production experience building scalable GenAI applications, AI agents, autonomous workflows, and Retrieval-Augmented Generation (RAG) systems in cloud-native environments.

This role requires deep technical expertise in LLM orchestration, AI application architecture, prompt engineering, vector databases, MLOps, and production deployment of AI systems. Candidates should have proven experience delivering real-world AI solutions in enterprise environments with strong exposure to cloud platforms and DevOps practices.

Key Responsibilities

  • Design, build, and deploy enterprise-grade Generative AI applications using Large Language Models (LLMs).
  • Develop intelligent AI agents and autonomous workflows using frameworks such as LangChain, CrewAI, LangGraph, AutoGen, or similar agentic AI frameworks.
  • Implement and optimize Retrieval-Augmented Generation (RAG) pipelines using vector databases and semantic search technologies.
  • Work extensively on prompt engineering, tool calling, memory management, agent orchestration, and multi-agent systems.
  • Integrate and manage LLMs such as OpenAI, Azure OpenAI, Claude, Gemini, Llama, Mistral, or similar foundation models.
  • Develop scalable AI services and APIs using Python and FastAPI.
  • Build production-ready AI solutions with high availability, scalability, monitoring, and observability.
  • Deploy and manage AI applications in cloud-native environments using Docker and Kubernetes.
  • Collaborate with Data Science, ML Engineering, and DevOps teams to operationalize AI solutions.
  • Implement CI/CD pipelines and automated deployment processes for AI workloads.
  • Monitor model performance, latency, reliability, and operational efficiency in production environments.
  • Ensure AI solutions follow enterprise security, governance, and responsible AI standards.
  • Evaluate and adopt emerging Generative AI tools, frameworks, and models.

Required Skills

Generative AI & LLM Expertise

  • Strong hands-on experience with Generative AI and Large Language Models (LLMs).
  • Production-level experience building and deploying GenAI applications.
  • Expertise in LangChain, CrewAI, LangGraph, AutoGen, or similar frameworks.
  • Experience with AI agents, autonomous workflows, and multi-agent architectures.
  • Strong understanding of prompt engineering, embeddings, model evaluation, and LLM orchestration.
  • Experience integrating OpenAI, Azure OpenAI, Claude, Gemini, Llama, Mistral, or similar models.

RAG & Vector Databases

  • Strong experience implementing RAG pipelines and semantic retrieval systems.
  • Experience with vector databases such as Pinecone, Weaviate, ChromaDB, FAISS, or Milvus.
  • Understanding of chunking strategies, embeddings, indexing, reranking, and retrieval optimization.

Python & AI Development

  • Strong proficiency in Python.
  • Experience with FastAPI for AI service and API development.
  • Experience with AI/ML libraries and data processing tools such as Pandas and NumPy.

Cloud & Production Deployment

  • Mandatory production experience on at least one cloud platform:
  • Microsoft Azure
  • Experience deploying scalable AI applications in enterprise production environments.
  • Hands-on experience with Docker, Kubernetes, Jenkins, Terraform, and CI/CD pipelines.
  • Strong understanding of MLOps, AI deployment lifecycle, monitoring, and observability.

Engineering & Operational Excellence

  • Strong understanding of software engineering best practices.
  • Experience with Git, version control, automated testing, and release management.
  • Experience building secure, scalable, and high-performance AI solutions.
  • Ability to troubleshoot production AI systems and optimize performance.

Preferred Skills

  • Experience with AI observability and evaluation frameworks.
  • Exposure to fine-tuning, PEFT, LoRA, or model optimization techniques.
  • Experience with enterprise AI governance and responsible AI practices.
  • Knowledge of distributed AI systems and scalable inference architectures.
  • Familiarity with AI security and compliance standards.

Qualifications

  • Bachelor’s or Master’s degree in Computer Science, Artificial Intelligence, Data Science, or a related field.
  • 3–8 years of overall software engineering experience.
  • Minimum 3+ years of hands-on experience in Generative AI and LLM-based application development,
  • Proven track record of delivering enterprise-scale AI solutions in production environments.
  • Strong communication and stakeholder management skills.
Read more
Service Co
Service Co
Agency job
via by Rishika Teja
Pune
6 - 8 yrs
₹14L - ₹18L / yr
skill iconPython
Large Language Models (LLM)
Retrieval Augmented Generation (RAG)
skill iconDocker
skill iconKubernetes
+1 more

Hiring for AI Engineer


Exp: 6 - 8 yrs

Edu : BE/B.Tech/MCA

Work Location : Pune


Skill Set:


- Total experience ranging from 6–8 years in software engineering/AI roles

- Min 5 years strong programming experience in Python is a MUST

- Min 3.5 years hands-on experience in AI with LLMs, RAG pipelines, and AI frameworks

- Experience with cloud platforms (AWS/Azure/GCP)






Read more
One of the largest Paper product Manufacturing Conglomorate
One of the largest Paper product Manufacturing Conglomorate
Agency job
via by praveen somasundaram
Bengaluru (Bangalore)
5 - 6 yrs
₹35L - ₹40L / yr
Generative AI (GenAI)
Agentic AI
Azure OpenAI
skill iconPython
Large Language Models (LLM) tuning
+5 more

Senior Gen AI Full Stack Engineer:

• Strong background in AI/ML and Gen AI with a deep understanding of LLMs, NLP pipelines, and AI model lifecycle.

• Experience in designing and building guardrail systems for Gen AI applications – including prompt filtering, semantic validation, toxicity detection, and hallucination mitigation.

• Fast API experience for API development.

• Proficiency in Python with frameworks like LangChain, Transformers, OpenAI, and LLM orchestration tools.

• Strong DevOps skills including CI/CD, Docker, Kubernetes, and Git.

Experience integrating Gen AI models into enterprise platforms securely and ethically.

Read more
icon

The recruiter has not been active on this job recently. You may apply but please expect a delayed response.

Jaipur
3 - 6 yrs
₹7L - ₹10L / yr
skill iconPython
Large Language Models (LLM)
Generative AI
Retrieval Augmented Generation (RAG)
LangChain
+23 more

Location: Jaipur (Work From Office)

Employment Type: Full-Time


We're looking for a GenAI Engineer (LLM Engineer) to build scalable AI-powered SaaS applications using Large Language Models (LLMs). You'll develop intelligent AI workflows, integrate LLMs into production systems, and build secure, high-performance AI solutions.


Key Responsibilities

  • Integrate LLM APIs (OpenAI, Claude, Hugging Face) into production applications.
  • Design and optimize RAG pipelines and prompt engineering workflows.
  • Build and manage Vector Databases (Pinecone, Weaviate, pgvector).
  • Optimize AI performance, latency, and operational cost.
  • Ensure secure, scalable AI architecture.
  • Collaborate with Product and Engineering teams to deliver AI-powered features.


Requirements

  • 3+ years of backend development using Python, Go, or Node.js.
  • Hands-on experience with LLMs, LangChain or LlamaIndex.
  • Strong understanding of RAG, Prompt Engineering, and Vector Databases.
  • Experience with AWS, GCP, or Azure.
  • Knowledge of APIs, Microservices, and AI application development.


Preferred: Experience in SaaS/FinTech, LLMOps, or Model Fine-tuning.

Education: B.Tech, BCA, or equivalent technical qualification.


Apply Now

Application Form: https://zfrmz.com/pAKb2ynfomIsuNwRfRbV?utm_source=cutshort

Read more
company logo
Anish N
Posted by Anish N
Bengaluru (Bangalore)
3 - 5 yrs
₹10L - ₹20L / yr
skill iconPython
Generative AI
Agentic AI
LangChain
LlamaIndex
+3 more

Job Description:

We are looking for a hands-on AI Engineer with experience in Generative AI and Agentic AI to build and deploy production-ready AI solutions.

Key Responsibilities:

  • Develop and deploy GenAI and Agentic AI applications.
  • Build RAG pipelines, LLM workflows, and AI agents.
  • Develop solutions using Python, LangChain, LangGraph, LlamaIndex, or similar frameworks.
  • Implement tool calling, context retrieval, and LLM orchestration.
  • Integrate AI solutions with APIs and cloud platforms.
  • Work with AWS/Azure/GCP, Docker, and CI/CD.

Required Skills:

  • Strong Python programming skills.
  • 3+ years of GenAI/Agentic AI experience.
  • RAG and LLM orchestration.
  • LangChain / LangGraph / LlamaIndex / AutoGen / CrewAI / Semantic Kernel.
  • MCP and A2A knowledge.
  • Cloud, APIs, Docker, and CI/CD experience.

Preferred Experience:

Hands-on experience building and deploying production-ready AI solutions.

Read more
company logo
Shakthi M
Posted by Shakthi M
Bengaluru (Bangalore), Mumbai
5 - 14 yrs
Best in industry
Anti money laundering
Fraud
skill iconPython
AML
skill iconDjango

Must of Skills/Experience 

• System Design

• Python

• TensorFlow

• Google ADK or Lang Graph

• Lang Chain , Lang Graph

• Spark

• Agentic AI Design

• ML Ops

• MCP (client and server)

• FastAPI

• Doc Factory

• RAG

• Golang

• LLMs – Gemini, Open AI

• NLP

• Dev Assistant - AI based code - generation

(Qwen or Claude or Copilot)

• CI/CD

• Good in oral and written communication,

collaboration and be a team player

Good to have skills 

• DevOps with K8

• Scripting

• Java

• REST API

• UV

• ReACT

• DocFactory

• Unix

Read more
company logo
Agency job
via by Naveen Balne
Hyderabad, Bengaluru (Bangalore), Pune, Chennai, Kolkata
5 - 13 yrs
₹15L - ₹40L / yr
Artificial Intelligence (AI)
Generative AI
Generative AI (GenAI)
skill iconPython
Large Language Models (LLM)
+7 more

We are seeking Generative AI Developers with strong Python programming and AI/ML expertise to build, deploy, and optimize LLM-powered applications. The role involves developing RAG solutions, AI agents, and enterprise GenAI applications while collaborating with cross-functional teams.


Key Responsibilities


  • Develop and enhance Generative AI applications using LLMs and AI frameworks.
  • Build and optimize RAG pipelines, vector search, and AI-powered workflows.
  • Design effective prompts and fine-tune models using techniques such as LoRA and QLoRA.
  • Develop REST APIs and integrate AI capabilities into enterprise applications.
  • Deploy, monitor, and maintain AI solutions in cloud and containerized environments.
  • Ensure code quality through testing, debugging, documentation, and code reviews.
  • Follow Responsible AI, security, and data governance practices.


Required Technical Skills


  • Strong proficiency in Python, OOP, APIs, debugging, and software development best practices.
  • Good understanding of Data Structures & Algorithms, complexity analysis, and problem-solving.
  • Hands-on experience with LLMs, Prompt Engineering, RAG, AI Agents, and embeddings.
  • Experience with LangChain, LangGraph, LlamaIndex, Hugging Face, or similar frameworks.
  • Knowledge of vector databases, semantic/hybrid search, and retrieval architectures.
  • Experience with PyTorch, TensorFlow, or Keras.
  • Familiarity with Docker, Git, CI/CD, and cloud platforms (Azure/AWS/GCP).
  • Understanding of AI governance, data privacy, and Responsible AI principles.


Preferred Skills


  • Experience with Agentic AI frameworks (CrewAI, AutoGen, Semantic Kernel).
  • Exposure to Azure AI Foundry, Databricks, or enterprise AI platforms.
  • Knowledge of multimodal AI applications.


Qualifications


  • Bachelor's or Master's degree in Computer Science, AI, Data Science, or a related field.
  • 5 years of software development experience, including AI/ML or Generative AI projects.
  • Experience building and deploying production-grade AI solutions.

Assessment Focus Areas


Candidates will be evaluated on:

  • Python coding and problem-solving
  • Data Structures & Algorithms
  • LLMs, RAG, and Agentic AI concepts
  • API development and system design
  • Cloud deployment and AI solution architecture
Read more
Leadsquared
Leadsquared
Agency job
via by Vrishali Mishra
Bengaluru (Bangalore)
2 - 4 yrs
₹25L - ₹45L / yr
Large Language Models (LLM) tuning

About LeadSquared

LeadSquared is a leading sales execution and marketing automation platform trusted by 2,000+ businesses globally, including healthcare, education, financial services, and real estate. Headquartered in Bengaluru with offices across the US, UK, UAE, and Southeast Asia, we empower sales teams to close faster, smarter, and at scale.

Our AI team is at the forefront of integrating cutting-edge large language model capabilities into enterprise workflows — building intelligent agents, copilots, and automation systems that redefine how businesses operate.

Role Overview

We are looking for a Senior AI Engineer with hands-on experience building LLM-powered agents and agentic AI systems. You will design, develop, and deploy autonomous AI pipelines that solve complex, multi-step business problems — from lead qualification and follow-up automation to intelligent CRM workflows and beyond.

This role is ideal for someone who is deeply excited about the frontier of AI, can move fast, and wants their work to directly impact millions of sales professionals worldwide.

Key Responsibilities

•

Design and build LLM-powered agentic systems using frameworks such as LangChain, LlamaIndex, AutoGen, or CrewAI to automate complex, multi-step workflows.

•

Develop and maintain Retrieval-Augmented Generation (RAG) pipelines with vector databases (Pinecone, Weaviate, Chroma, pgvector) for domain-specific knowledge grounding.

•

Build and integrate tool-use and function-calling capabilities into AI agents, enabling dynamic interaction with internal APIs, databases, and third-party services.

•

Implement prompt engineering strategies including chain-of-thought, few-shot prompting, and structured output parsing to ensure reliable agent behavior.

•

Design evaluation frameworks and observability pipelines (LangSmith, Helicone, custom metrics) to monitor agent performance, accuracy, and cost.

•

Collaborate with product, sales, and domain teams to translate business requirements into AI-driven solutions and features.

•

Optimize LLM inference for latency and cost using techniques like caching, model distillation, quantization, and batching.

•

Stay current with the rapidly evolving LLM ecosystem and proactively propose improvements and new approaches.

•

Contribute to internal best practices, documentation, and knowledge-sharing across the engineering org.

Required Qualifications

Experience

•

2–4 years of professional software engineering experience, with at least 1–2 years focused on LLM/AI systems.

•

Proven experience shipping LLM-based products or agentic AI systems into production environments.

Technical Skills

•

Strong proficiency in Python and familiarity with async programming patterns for AI pipelines.

•

Hands-on experience with LLM APIs: OpenAI (GPT-4o), Anthropic (Claude), Google (Gemini), or open-source models (Llama, Mistral).

•

Experience with agentic frameworks: LangChain, LangGraph, LlamaIndex, AutoGen, CrewAI, or similar.

•

Solid understanding of RAG architectures, embedding models, and semantic search.

•

Experience with vector databases and similarity search infrastructure.

•

Knowledge of REST APIs, microservices architecture, and containerization (Docker/Kubernetes).

Problem-Solving & Mindset

•

Strong ability to decompose ambiguous, open-ended problems into structured AI system designs.

•

Experience with prompt debugging, LLM evaluation, and iterative refinement workflows.

•

Ability to balance research exploration with engineering pragmatism to ship reliable systems.

Preferred Qualifications

•

Experience with multi-agent orchestration and agent memory systems (short-term and long-term).

•

Familiarity with fine-tuning or RLHF workflows for domain adaptation.

•

Background in NLP, information retrieval, or conversational AI.

•

Prior experience in B2B SaaS or CRM domain is a plus.

•

Contributions to open-source AI/ML projects or published research/blogs.

•

Experience with cloud platforms: AWS, GCP, or Azure — particularly AI/ML services

Read more
company logo
Bhawna Khemani
Posted by Bhawna Khemani
icon

The recruiter has not been active on this job recently. You may apply but please expect a delayed response.

Bengaluru (Bangalore), Delhi, Gurugram, Noida, Ghaziabad, Faridabad
4 - 13 yrs
₹11L - ₹35L / yr
Generative AI
Large Language Models (LLM)
Retrieval Augmented Generation (RAG)
skill iconPython
LlamaIndex
+4 more

Generative AI Engineer 

Role Overview:

You will be responsible for the hands-on development, coding, and deployment of AI-powered features. Your focus is on writing clean, efficient code to integrate LLMs into our existing tech stack, building robust data pipelines for RAG, and ensuring the reliability of model outputs through rigorous testing and optimization.

Key Responsibilities

  • Application Implementation: Code and integrate LLM APIs (OpenAI, Anthropic, etc.) or local models into backend services using Python, FastAPI, etc.,
  • MCP Server Development: Design and implement custom MCP servers using the official SDKs (Python/TypeScript) to expose internal databases, APIs, and file systems to AI agents.
  • RAG Implementation: Build and maintain the "plumbing" for Retrieval-Augmented Generation—specifically coding the data ingestion scripts, text chunking logic, and metadata filtering.
  • Vector DB Management: Perform day-to-day operations on vector databases (Pinecone, Milvus, etc.), including indexing, querying, and optimizing search retrieval.
  • Prompt Programming: Develop, version-control, and refine complex prompt templates (using Jinja2 or similar) to ensure consistent structured outputs (JSON/YAML).
  • Agent Development: Implement multi-step workflows using LangChain, LangGraph, CrewAI etc.,, focusing on tool-calling logic and error handling.
  • Evaluation & Testing: Build automated test suites to detect "hallucinations" and measure accuracy using frameworks.
  • Performance Tuning: Implement caching layers and streaming responses to reduce latency and improve the end-user experience; Token optimization.
  • Data Pre-processing: Clean and tokenize datasets for model fine-tuning or high-quality context retrieval.

Technical Skills (The "Execution" Stack)

  • Language: Advanced Python (Asyncio, Pydantic) and optional TypeScript/Node.js (for full-stack integration).
  • AI Frameworks: Hands-on experience with any of LangChain, LlamaIndex, and Hugging Face Transformers. RAG and Vector search concepts.
  • Data Handling: Proficiency in SQL and handling unstructured data formats (PDFs, Markdown, JSON).
  • Deployment: Practical experience with Docker, GitHub Actions (CI/CD), and experience with OpenTelemetry, LangSmith, Weights & Biases etc., Understanding of evaluation/guardrails.
  • MCP/API Proficiency: Deep understanding of RESTful APIs, Streaming HTTP, MCP server vs client, JSONRPC
Read more
company logo
Dharshini A
Posted by Dharshini A
Hyderabad
7 - 10 yrs
₹2L - ₹15L / yr
skill iconPython
Agentic AI
Large Language Models (LLM)
Fullstack Developer

 

Python (Gen AI or Agentic AI) - Hyderabad

7 + years of exp with more than 2 + years on Gen AI/Agentic AI.

Design and implement Generative AI and Agentic AI capabilities using LLM platforms and frameworks such as LangChain, LangGraph, Google ADK, Semantic Kernel, or equivalent.

Implement tool calling, RAG, memory, planning, reasoning, multi-agent orchestration, structured outputs, and human approval controls.

Integrate applications with REST APIs, relational and NoSQL databases, vector stores, message queues, and enterprise systems.

Read more
Why apply to jobs via Cutshort
people_solving_puzzle
Personalized job matches
Stop wasting time. Get matched with jobs that meet your skills, aspirations and preferences.
people_verifying_people
Verified hiring teams
See actual hiring teams, find common social connections or connect with them directly.
ai_chip
Move faster with AI
We use AI to get you faster responses, recommendations and unmatched user experience.
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos