Cutshort logo
For Employers

LLM Engineer at TIFIN FINTECH INDIA LLP · Mumbai, Bengaluru (Bangalore) · 3 - 5 years · ₹20L - ₹40L / yr (ESOP available) · Raised funding · Posted 2 Feb 2026

TIFIN FINTECH INDIA LLP's logo

LLM Engineer

Yashika Shah's profile picture
Posted by Yashika Shah
3 - 5 yrs
₹20L - ₹40L / yr (ESOP available)
Mumbai, Bengaluru (Bangalore)
Skills
Large Language Models (LLM) tuning
PEFT (Parameter-Efficient Fine-Tuning)
Retrieval Augmented Generation (RAG)
Generative AI
Large Language Models (LLM)

Key responsibilities:

● Work closely with the design and product teams to help craft conversational experiences for our users

● Ability to work independently to deliver features that the team design

● You need to own outcomes without someone watching over you

● Understand the data we have and identify how we could leverage it to create more personalised experiences for our users.

● Fine tune models with data specific to our use case

● Use different RAG approaches to augment data that the LLM already has

● Ability to be a leader and independent contributor (we’re a startup so be ready to do whatever it takes)

● Setup new workflow, systems and tool from the ground up (you will have support, but we’re building things from scratch)


Requirements:

● 3+ years of experience

● Experience working with LLMs and Generative AI

● Experience building conversational bots

● Experience fine tuning models

● Experience using RAG base approaches

● Understanding of financial concepts and investing would be a big plus but not required

Read more
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos

About TIFIN FINTECH INDIA LLP

Founded :
2017
Type :
Products & Services
Size :
100-1000
Stage :
Raised funding

About

WHO WE ARE:

TIFIN is a fintech platform backed by industry leaders including JP Morgan, Morningstar,

Broadridge, Hamilton Lane, Franklin Templeton, Motive Partners and a who’s who of the financial

service industry. We are creating engaging wealth experiences to better financial lives through AI

and investment intelligence-powered personalization. We are working to change the world of

wealth in ways that personalization has changed the world of movies, music and more but with the

added responsibility of delivering better wealth outcomes.

We use design and behavioral thinking to enable engaging experiences through software and

application programming interfaces (APIs). We use investment science and intelligence to build

algorithmic engines inside the software and APIs to enable better investor outcomes.

In a world where every individual is unique, we match them to financial advice and investments

with a recognition of their distinct needs and goals across our investment marketplace and our

advice and planning divisions.

OUR VALUES: Go with your GUT

● Grow at the Edge. We are driven by personal growth. We get out of our comfort zone and keep

egos aside to find our genius zones. With self-awareness and integrity we strive to be the best we

can possibly be. No excuses.

●Understanding through Listening and Speaking the Truth. We value transparency. We

communicate with radical candor, authenticity and precision to create a shared understanding. We

challenge, but once a decision is made, commit fully.

●I Win for Teamwin. We believe in staying within our genius zones to succeed and we take full

ownership of our work. We inspire each other with our energy and attitude. We fly in formation to

win together.

Read more

Tech stack

skill iconPython
skill iconNodeJS (Node.js)
skill iconGo Programming (Golang)

Candid answers by the company

What is the location preference of jobs?

Bangalore

Mumbai

Company social profiles

instagramlinkedintwitter

Similar jobs (10)

Service Co
Service Co
Agency job
via by Rishika Teja
Pune, Mumbai
5 - 10 yrs
₹15L - ₹40L / yr
Artificial Intelligence (AI)
Generative AI
Large Language Models (LLM)
Retrieval Augmented Generation (RAG)

Hiring for AI Engineer


Exp: 5 - 10 yrs

Edu : BE/B.Tech/MCA

Work Location : Pune / Mumbai


Skill Set:


Total experience ranging from 5–10 years in software engineering/AI roles

Min 5 years strong programming experience in Python is a MUST

Min 3.5 years hands-on experience in AI with LLMs, RAG pipelines, and AI frameworks

2+ years shipping LLM systems in production

Experience with cloud platforms (AWS/Azure/GCP)

Read more
company logo
Pramila Ranjane
Posted by Pramila Ranjane
Pune
1 - 4 yrs
₹12L - ₹25L / yr
GEO
AEO
SGE initiatives
Generative AI
Fine-tuning LLMs
+5 more

🔹 Key Responsibilities


• Design, develop, and deploy production-grade AI/ML and Generative AI solutions

• Work on GEO, AEO, and SGE initiatives to improve visibility across AI-driven search platforms

• Optimize content and digital experiences for conversational queries and LLM-based search

• Develop solutions using LLMs, NLP, embeddings, semantic search, RAG, and vector databases

• Analyze search intent, AI-generated responses, citations, retrieval patterns, and content discoverability

• Build frameworks to measure GEO/AEO strategies and AI-search performance

• Collaborate with Product, Engineering, Content, SEO, Marketing, and Business teams

• Improve solution accuracy, relevance, latency, and user experience


🔹 Mandatory Requirements


✅ 1–4 years of professional experience

✅ Minimum 1 year of hands-on experience in GEO, AEO, or SGE

✅ Experience with prompt engineering, embeddings, vector search, or RAG systems

✅ Understanding of semantic search and entity-based optimization

✅ Exposure to ChatGPT, Google Gemini, or similar LLM platforms

✅ Knowledge of schema, context building, content structuring, and knowledge representation


🎓 Preferred Education


B.Tech, M.Tech, Integrated M.Sc., or MS from a Tier-1 engineering institute such as IIT, NIT, BITS, VIT, DTU, or NSUT.

Read more
company logo
Banu S
Posted by Banu S
Bengaluru (Bangalore)
4 - 12 yrs
₹4L - ₹25L / yr
skill iconData Science
skill iconPython
Large Language Models (LLM) tuning
RAG
Langchain

Support with design and build to prove out agentic AI solution flow by working with other data 

scientists and engineers to build, train Large Language Model (LLM) architectures, RAG 

systems, and autonomous agentic workflows 

Key qualifications: 

  

>> AI solution design & Development: Design Agentic AI solutions using RAG (Retrieval-

Augmented Generation) and orchestration frameworks like LangGraph or LangChain. 

  

>> Model Fine-Tuning: Solid understanding and experience with Pre-train, fine-tune, and 

optimize open-source like BERT, LLama, and other proprietary foundation models for domain-

specific tasks 

  

>> Solid Stats and ML foundations and (vibe) coding skills with Python, PySpark 

  

>>  Implement validation frameworks and tracing practices (using tools like Arize) to monitor 

agent behavior, guard against model drift, and ensure compliance 

  

>> Collaborate with Engineering to deploy models securely on cloud and on-prem ecosystems 

 

Read more
company logo
Anish N
Posted by Anish N
Bengaluru (Bangalore)
3 - 5 yrs
₹10L - ₹20L / yr
skill iconPython
Generative AI
Agentic AI
LangChain
LlamaIndex
+3 more

Job Description:

We are looking for a hands-on AI Engineer with experience in Generative AI and Agentic AI to build and deploy production-ready AI solutions.

Key Responsibilities:

  • Develop and deploy GenAI and Agentic AI applications.
  • Build RAG pipelines, LLM workflows, and AI agents.
  • Develop solutions using Python, LangChain, LangGraph, LlamaIndex, or similar frameworks.
  • Implement tool calling, context retrieval, and LLM orchestration.
  • Integrate AI solutions with APIs and cloud platforms.
  • Work with AWS/Azure/GCP, Docker, and CI/CD.

Required Skills:

  • Strong Python programming skills.
  • 3+ years of GenAI/Agentic AI experience.
  • RAG and LLM orchestration.
  • LangChain / LangGraph / LlamaIndex / AutoGen / CrewAI / Semantic Kernel.
  • MCP and A2A knowledge.
  • Cloud, APIs, Docker, and CI/CD experience.

Preferred Experience:

Hands-on experience building and deploying production-ready AI solutions.

Read more
company logo
Agency job
via by Naveen Balne
Hyderabad, Bengaluru (Bangalore), Pune, Chennai, Kolkata
5 - 13 yrs
₹15L - ₹40L / yr
Artificial Intelligence (AI)
Generative AI
Generative AI (GenAI)
skill iconPython
Large Language Models (LLM)
+7 more

We are seeking Generative AI Developers with strong Python programming and AI/ML expertise to build, deploy, and optimize LLM-powered applications. The role involves developing RAG solutions, AI agents, and enterprise GenAI applications while collaborating with cross-functional teams.


Key Responsibilities


  • Develop and enhance Generative AI applications using LLMs and AI frameworks.
  • Build and optimize RAG pipelines, vector search, and AI-powered workflows.
  • Design effective prompts and fine-tune models using techniques such as LoRA and QLoRA.
  • Develop REST APIs and integrate AI capabilities into enterprise applications.
  • Deploy, monitor, and maintain AI solutions in cloud and containerized environments.
  • Ensure code quality through testing, debugging, documentation, and code reviews.
  • Follow Responsible AI, security, and data governance practices.


Required Technical Skills


  • Strong proficiency in Python, OOP, APIs, debugging, and software development best practices.
  • Good understanding of Data Structures & Algorithms, complexity analysis, and problem-solving.
  • Hands-on experience with LLMs, Prompt Engineering, RAG, AI Agents, and embeddings.
  • Experience with LangChain, LangGraph, LlamaIndex, Hugging Face, or similar frameworks.
  • Knowledge of vector databases, semantic/hybrid search, and retrieval architectures.
  • Experience with PyTorch, TensorFlow, or Keras.
  • Familiarity with Docker, Git, CI/CD, and cloud platforms (Azure/AWS/GCP).
  • Understanding of AI governance, data privacy, and Responsible AI principles.


Preferred Skills


  • Experience with Agentic AI frameworks (CrewAI, AutoGen, Semantic Kernel).
  • Exposure to Azure AI Foundry, Databricks, or enterprise AI platforms.
  • Knowledge of multimodal AI applications.


Qualifications


  • Bachelor's or Master's degree in Computer Science, AI, Data Science, or a related field.
  • 5 years of software development experience, including AI/ML or Generative AI projects.
  • Experience building and deploying production-grade AI solutions.

Assessment Focus Areas


Candidates will be evaluated on:

  • Python coding and problem-solving
  • Data Structures & Algorithms
  • LLMs, RAG, and Agentic AI concepts
  • API development and system design
  • Cloud deployment and AI solution architecture
Read more
company logo
Prithisha Kathiresan
Posted by Prithisha Kathiresan
Bengaluru (Bangalore)
3 - 8 yrs
Best in industry
Generative AI (GenAI)
Large Language Models (LLM) tuning
Retrieval Augmented Generation (RAG)
Azure OpenAI
skill iconAmazon Web Services (AWS)
+2 more

Senior Generative AI Engineer

Employment Type: Permanent with VDart Digital

Work Location: Marathalli, Bengaluru

Job Description

We are seeking a highly skilled Senior Generative AI Engineer with strong expertise in designing, developing, and deploying enterprise-scale AI solutions using Large Language Models (LLMs) and modern Generative AI frameworks. The ideal candidate should have hands-on production experience building scalable GenAI applications, AI agents, autonomous workflows, and Retrieval-Augmented Generation (RAG) systems in cloud-native environments.

This role requires deep technical expertise in LLM orchestration, AI application architecture, prompt engineering, vector databases, MLOps, and production deployment of AI systems. Candidates should have proven experience delivering real-world AI solutions in enterprise environments with strong exposure to cloud platforms and DevOps practices.

Key Responsibilities

  • Design, build, and deploy enterprise-grade Generative AI applications using Large Language Models (LLMs).
  • Develop intelligent AI agents and autonomous workflows using frameworks such as LangChain, CrewAI, LangGraph, AutoGen, or similar agentic AI frameworks.
  • Implement and optimize Retrieval-Augmented Generation (RAG) pipelines using vector databases and semantic search technologies.
  • Work extensively on prompt engineering, tool calling, memory management, agent orchestration, and multi-agent systems.
  • Integrate and manage LLMs such as OpenAI, Azure OpenAI, Claude, Gemini, Llama, Mistral, or similar foundation models.
  • Develop scalable AI services and APIs using Python and FastAPI.
  • Build production-ready AI solutions with high availability, scalability, monitoring, and observability.
  • Deploy and manage AI applications in cloud-native environments using Docker and Kubernetes.
  • Collaborate with Data Science, ML Engineering, and DevOps teams to operationalize AI solutions.
  • Implement CI/CD pipelines and automated deployment processes for AI workloads.
  • Monitor model performance, latency, reliability, and operational efficiency in production environments.
  • Ensure AI solutions follow enterprise security, governance, and responsible AI standards.
  • Evaluate and adopt emerging Generative AI tools, frameworks, and models.

Required Skills

Generative AI & LLM Expertise

  • Strong hands-on experience with Generative AI and Large Language Models (LLMs).
  • Production-level experience building and deploying GenAI applications.
  • Expertise in LangChain, CrewAI, LangGraph, AutoGen, or similar frameworks.
  • Experience with AI agents, autonomous workflows, and multi-agent architectures.
  • Strong understanding of prompt engineering, embeddings, model evaluation, and LLM orchestration.
  • Experience integrating OpenAI, Azure OpenAI, Claude, Gemini, Llama, Mistral, or similar models.

RAG & Vector Databases

  • Strong experience implementing RAG pipelines and semantic retrieval systems.
  • Experience with vector databases such as Pinecone, Weaviate, ChromaDB, FAISS, or Milvus.
  • Understanding of chunking strategies, embeddings, indexing, reranking, and retrieval optimization.

Python & AI Development

  • Strong proficiency in Python.
  • Experience with FastAPI for AI service and API development.
  • Experience with AI/ML libraries and data processing tools such as Pandas and NumPy.

Cloud & Production Deployment

  • Mandatory production experience on at least one cloud platform:
  • Microsoft Azure
  • Experience deploying scalable AI applications in enterprise production environments.
  • Hands-on experience with Docker, Kubernetes, Jenkins, Terraform, and CI/CD pipelines.
  • Strong understanding of MLOps, AI deployment lifecycle, monitoring, and observability.

Engineering & Operational Excellence

  • Strong understanding of software engineering best practices.
  • Experience with Git, version control, automated testing, and release management.
  • Experience building secure, scalable, and high-performance AI solutions.
  • Ability to troubleshoot production AI systems and optimize performance.

Preferred Skills

  • Experience with AI observability and evaluation frameworks.
  • Exposure to fine-tuning, PEFT, LoRA, or model optimization techniques.
  • Experience with enterprise AI governance and responsible AI practices.
  • Knowledge of distributed AI systems and scalable inference architectures.
  • Familiarity with AI security and compliance standards.

Qualifications

  • Bachelor’s or Master’s degree in Computer Science, Artificial Intelligence, Data Science, or a related field.
  • 3–8 years of overall software engineering experience.
  • Minimum 3+ years of hands-on experience in Generative AI and LLM-based application development,
  • Proven track record of delivering enterprise-scale AI solutions in production environments.
  • Strong communication and stakeholder management skills.
Read more
Neosapien
Neosapien
Agency job
via by Nehlata Pandey
Bengaluru (Bangalore)
3 - 7 yrs
₹15L - ₹40L / yr (ESOP available)
skill iconPython
Large Language Models (LLM) tuning
Agentic AI
Retrieval Augmented Generation (RAG)
skill iconMachine Learning (ML)
+2 more

The Role

You own AI systems end to end. From the speech-to-text models that turn audio into text, to the diarization that separates and identifies speakers, to the agentic layer that turns conversation into memory and action, to the observability and evaluation that keep all of it honest in production. This is a wide role by design. You will own model selection, serving, and production reliability. If you want to tune one model and ignore the system around it, this is not the role.

What You Will Own

•     Speech-to-text. Evaluate, integrate, and optimize STT models across cloud and self-hosted. Drive accuracy and cost trade-offs with ground-truth metrics.

•     Speaker diarization and identification. Push accuracy on hard, real-world, multi-speaker audio.

•     Agentic AI. Build the memory and retrieval pipeline, LLM orchestration, and the agent workflows that sit on top of captured conversation.

•     Model serving and infrastructure. Stand up and optimize self-hosted serving (vLLM, Triton class). Own latency, throughput, and cost per user.

Observability

An always-on wearable means models run in production every second, on messy real-world audio. You own the visibility into that.

•     Instrument the full audio-to-memory pipeline: STT, diarization, retrieval, and LLM calls.

•     Define and track model-quality SLOs in production: transcription drift, diarization error over time, retrieval relevance, latency, throughput, and cost per user.

•     Build dashboards and alerting so model degradation is caught before users feel it.

•     Trace failures across a distributed, always-on system using metrics, logs, and traces.

•     Close the loop. Production signals feed back into evaluation and model selection.

Evaluation

We do not ship what we cannot measure. You own the systems that prove a model is actually better, not just newer.

•     Build and own ground-truth evaluation harnesses for every model in the stack.

•     Measure with real metrics: WER for transcription, DER for diarization, Recall and F1 for retrieval and speaker identification.

•     Build and maintain labeled benchmark datasets that reflect real, messy, multi-speaker audio.

•     Run regression and A/B evaluations on every model swap, prompt change, or pipeline update. Nothing ships on a vibe.

•     Reject anecdotal proxies, single confidence scores, and cherry-picked examples as evidence of quality.

What We Are Looking For

•     3 to 5 years as an AI/ML engineer with production systems behind you. Engineering and production experience is non-negotiable.

•     Depth across the modern AI stack: LLMs, speech models, vector retrieval, model serving.

•     Strong software engineering. You write code that ships and survives contact with real users.

•     Fluency in Python and the production ML ecosystem.

•     Comfort with cloud infrastructure (GCP a plus) and containerized deployment on Kubernetes.

•     A working command of observability and evaluation. You measure first and trust metrics over intuition.

•     First-principles reasoning and metric discipline.

Nice to Have

•     Research background or publications. A strong signal, not a substitute for production work.

•     Audio and speech ML experience (STT, diarization, voice).

•     Experience self-hosting and optimizing open models.

•     Experience with LLM gateway and agent orchestration patterns.

•     Experience building eval harnesses or production model-monitoring systems.


Requirements

Agentic work is must. Audio is good to have

. Self hosting models is a must

 Experience with LLM gateway and agent orchestration is a must have

Read more
company logo
Akash Bhatt
Posted by Akash Bhatt
Bengaluru (Bangalore), Delhi, Gurugram, Noida, Ghaziabad, Faridabad, Hyderabad, Mumbai, Chennai, Pune
1 - 10 yrs
₹8L - ₹40L / yr
LangChain
Retrieval Augmented Generation (RAG)
Vector database
Prompt engineering
LlamaIndex

We are hiring a Generative AI Engineer to build production LLM applications.


Responsibilities

  • Build RAG pipelines with LangChain or LlamaIndex
  • Design prompts and evaluate model outputs
  • Manage embeddings in vector databases such as Pinecone, Weaviate or FAISS
  • Deploy and monitor LLM features in production


Requirements

  • 1+ years building LLM-powered applications
  • Hands-on with LangChain or LlamaIndex and vector databases
  • Experience with the OpenAI, Anthropic or open-source model APIs
Read more
company logo
Aishwarya SURENDRAN
Posted by Aishwarya SURENDRAN
Hyderabad
5 - 10 yrs
₹7L - ₹8L / yr
Generative AI
Agentic AI

Job Title: AI/ML Consultant

Company Name- WINIT

Location: Hyderabad

Duration- 3 months


Role Overview:

We are looking for a passionate and skilled AI/ML Consultant to join our dynamic team. In this role, you will play a key part in designing and implementing intelligent systems, including the development of a cutting-edge Sales Supervisor Agent. You will work on projects involving Generative AI, Voice AI, sales performance analysis, and recommendation systems that drive automation and strategic decision-making in sales operations. As a consultant, you will collaborate with cross-functional teams to understand business challenges, recommend AI-driven solutions, and deliver scalable, production-ready applications.


Project Knowledge: Generative AI & Voice AI


Experiment with Generative AI models (e.g., GPT, Claude) for tasks such as content creation, email generation, and chat-based assistance.

Build and integrate AI Voice solutions like speech-to-text, call summarization, and conversational agents using tools such as Whisper, ElevenLabs, or Dialogflow.

Integrate GenAI and Voice AI capabilities into the Sales Supervisor Agent for automation and decision support.

Read more
company logo
Aswathy Vimal
Posted by Aswathy Vimal
Remote only
6 - 10 yrs
₹24L - ₹36L / yr
skill iconPython
skill iconJavascript
TypeScript
Large Language Models (LLM) tuning

The Role

We’re looking for a Senior Applied AI & Data Engineer to become our first dedicated AI and data engineer.

You’ll build conversational AI experiences across web, mobile, and in-store channels while developing the data foundation behind them. You’ll make key technical decisions and own your work through to production.

What You’ll Do

• Build AI assistants using tool calling to work with real product, search, and order systems

• Design guardrails and evaluation sets to ensure AI responses are accurate and safe

• Build real-time and voice-enabled AI experiences

• Improve product data quality through AI-assisted enrichment and review workflows

• Build data pipelines, analytics, and personalisation systems

• Work closely with web and mobile developers and help guide technical implementation

What You’ll Need

• 6+ years of experience building and running production backend systems

• Strong Python skills, plus experience with JavaScript/TypeScript backends

• Experience shipping at least one LLM-powered feature to real users

• Experience with search and relevance

• Experience building data pipelines and analytics stores

• Comfortable deploying and monitoring services on a major cloud platform

• Strong communication skills and the ability to work independently

Nice to Have

• Experience with speech or voice AI

• E-commerce or retail technology experience

• Experience building multilingual products

Read more
Why apply to jobs via Cutshort
people_solving_puzzle
Personalized job matches
Stop wasting time. Get matched with jobs that meet your skills, aspirations and preferences.
people_verifying_people
Verified hiring teams
See actual hiring teams, find common social connections or connect with them directly.
ai_chip
Move faster with AI
We use AI to get you faster responses, recommendations and unmatched user experience.
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos