Cutshort logo
For Employers
 Global Digital Transformation Solutions Provider logo
Lead II - Software Engineering- AI, NLP, Python, Data science
Global Digital Transformation Solutions Provider
Lead II - Software Engineering- AI, NLP, Python, Data science

Lead II - Software Engineering- AI, NLP, Python, Data science at Global Digital Transformation Solutions Provider · Bengaluru (Bangalore) · 7 - 9 years · ₹10L - ₹28L / yr · Posted 1 Nov 2025

Peak Hire Solutions's logo

Lead II - Software Engineering- AI, NLP, Python, Data science

at Global Digital Transformation Solutions Provider

Agency job
7 - 9 yrs
₹10L - ₹28L / yr
Bengaluru (Bangalore)
Skills
Artificial Intelligence (AI)
Natural Language Processing (NLP)
skill iconPython
skill iconData Science
Generative AI
Retrieval Augmented Generation (RAG)
Large Language Models (LLM)
PyTorch
TensorFlow
Architecture
Vector database
skill iconJava
skill iconAmazon Web Services (AWS)
Google Cloud Storage
AWS CloudFormation

Job Details

Job Title: Lead II - Software Engineering- AI, NLP, Python, Data science

Industry: Technology

Domain - Information technology (IT)

Experience Required: 7-9 years

Employment Type: Full Time

Job Location: Bangalore

CTC Range: Best in Industry


Job Description:

Role Proficiency:

Act creatively to develop applications by selecting appropriate technical options optimizing application development maintenance and performance by employing design patterns and reusing proven solutions. Account for others' developmental activities; assisting Project Manager in day-to-day project execution.


Additional Comments:

Mandatory Skills Data Science Skill to Evaluate AI, Gen AI, RAG, Data Science

Experience 8 to 10 Years

Location Bengaluru

Job Description

Job Title AI Engineer Mandatory Skills Artificial Intelligence, Natural Language Processing, python, data science Position AI Engineer – LLM & RAG Specialization Company Name: Sony India Software Centre About the role: We are seeking a highly skilled AI Engineer with 8-10 years of experience to join our innovation-driven team. This role focuses on the design, development, and deployment of advanced enterprise-scale Large Language Models (eLLM) and Retrieval Augmented Generation (RAG) solutions. You will work on end-to-end AI pipelines, from data processing to cloud deployment, delivering impactful solutions that enhance Sony’s products and services. Key Responsibilities: Design, implement, and optimize LLM-powered applications, ensuring high performance and scalability for enterprise use cases. Develop and maintain RAG pipelines, including vector database integration (e.g., Pinecone, Weaviate, FAISS) and embedding model optimization. Deploy, monitor, and maintain AI/ML models in production, ensuring reliability, security, and compliance. Collaborate with product, research, and engineering teams to integrate AI solutions into existing applications and workflows. Research and evaluate the latest LLM and AI advancements, recommending tools and architectures for continuous improvement. Preprocess, clean, and engineer features from large datasets to improve model accuracy and efficiency. Conduct code reviews and enforce AI/ML engineering best practices. Document architecture, pipelines, and results; present findings to both technical and business stakeholders. Job Description: 8-10 years of professional experience in AI/ML engineering, with at least 4+ years in LLM development and deployment. Proven expertise in RAG architectures, vector databases, and embedding models. Strong proficiency in Python; familiarity with Java, R, or other relevant languages is a plus. Experience with AI/ML frameworks (PyTorch, TensorFlow, etc.) and relevant deployment tools. Hands-on experience with cloud-based AI platforms such as AWS SageMaker, AWS Q Business, AWS Bedrock or Azure Machine Learning. Experience in designing, developing, and deploying Agentic AI systems, with a focus on creating autonomous agents that can reason, plan, and execute tasks to achieve specific goals. Understanding of security concepts in AI systems, including vulnerabilities and mitigation strategies. Solid knowledge of data processing, feature engineering, and working with large-scale datasets. Experience in designing and implementing AI-native applications and agentic workflows using the Model Context Protocol (MCP) is nice to have. Strong problem-solving skills, analytical thinking, and attention to detail. Excellent communication skills with the ability to explain complex AI concepts to diverse audiences. Day-to-day responsibilities: Design and deploy AI-driven solutions to address specific security challenges, such as threat detection, vulnerability prioritization, and security automation. Optimize LLM-based models for various security use cases, including chatbot development for security awareness or automated incident response. Implement and manage RAG pipelines for enhanced LLM performance. Integrate AI models with existing security tools, including Endpoint Detection and Response (EDR), Threat and Vulnerability Management (TVM) platforms, and Data Science/Analytics platforms. This will involve working with APIs and understanding data flows. Develop and implement metrics to evaluate the performance of AI models. Monitor deployed models for accuracy and performance and retrain as needed. Adhere to security best practices and ensure that all AI solutions are developed and deployed securely. Consider data privacy and compliance requirements. Work closely with other team members to understand security requirements and translate them into AI-driven solutions. Communicate effectively with stakeholders, including senior management, to present project updates and findings. Stay up to date with the latest advancements in AI/ML and security and identify opportunities to leverage new technologies to improve our security posture. Maintain thorough documentation of AI models, code, and processes. What We Offer Opportunity to work on cutting-edge LLM and RAG projects with global impact. A collaborative environment fostering innovation, research, and skill growth. Competitive salary, comprehensive benefits, and flexible work arrangements. The chance to shape AI-powered features in Sony’s next-generation products. Be able to function in an environment where the team is virtual and geographically dispersed

Education Qualification: Graduate


Skills: AI, NLP, Python, Data science


Must-Haves

Skills

AI, NLP, Python, Data science

NP: Immediate – 30 Days

 

Read more
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos

Similar jobs (10)

company logo
Rishu Dutta
Posted by Rishu Dutta
Gurugram
7 - 12 yrs
₹20L - ₹50L / yr
Retrieval Augmented Generation (RAG)
Agentic AI
Multi-agent Systems

Role Overview 

We are looking for an AI Engineer to design, build, and ship production AI systems, including agentic AI applications, for enterprise clients. This is a hands-on engineering role: you will write production code, build and evaluate models and agents, and work closely with architects and product teams to take solutions from prototype to scale. 


Key Responsibilities 

Design and build agentic AI systems: agent workflows, tool/function-calling, memory, and human-in-the-loop patterns. Build and productionise RAG pipelines, prompt-based applications, and LLM integrations across providers. Develop and maintain data and ML pipelines: feature engineering, model training, evaluation, and monitoring. Integrate AI systems with enterprise applications (CRMs, ERPs, ITSM tools) via APIs, events, and MCP-based tool servers. Implement guardrails, prompt-injection defences, and evaluation frameworks to keep AI systems safe and reliable in production. 

Write clean, tested, production-grade code and participate actively in code and design reviews. 

Collaborate with architects, product managers, and delivery teams to translate requirements into working AI solutions. Troubleshoot and optimise AI systems for accuracy, latency, and cost in production. 


Required Qualifications 

8–12 years of hands-on software engineering experience, with a strong, unbroken technical track record. Hands-on experience building and shipping AI/ML systems in production, not just POCs. 

Practical experience with agentic AI systems and at least one major agent framework (LangGraph, CrewAI, AutoGen, OpenAI Agents SDK, Bedrock Agents/Strands, or Semantic Kernel). 

Experience with LLM/GenAI systems: RAG pipelines, prompt engineering, structured outputs, and tool calling across providers. 

Strong Python skills (TypeScript/Node.js a plus), with production-grade testing, CI/CD, and API design practices. Working knowledge of ML fundamentals: model evaluation, feature engineering, and experimentation. Cloud-native experience on AWS and/or Azure: containers, serverless, event backbones, and vector databases. Understanding of LLM safety and reliability practices: guardrails, prompt-injection defences, and observability. 



Read more
company logo
Priyanka Khandelwal
Posted by Priyanka Khandelwal
icon

The recruiter has not been active on this job recently. You may apply but please expect a delayed response.

Jaipur
3 - 8 yrs
₹10L - ₹12L / yr
Generative AI (GenAI)
Large Language Models (LLM)
Retrieval Augmented Generation (RAG)

Job Description – AI Engineer (End-to-End Development & Deployment)


Role Summary

We are looking for an AI Engineer with hands-on experience in designing, developing, deploying, and maintaining Generative/Agentic AI solutions in production. The ideal candidate should have end-to-end ownership of AI applications, from development to deployment, monitoring, and optimization.

Key Responsibilities

●        Design, build, and deploy Generative/Agentic AI solutions.

●        Develop applications using LLMs, RAG, AI agents, and vector databases.

●        Build scalable APIs and integrate AI solutions with enterprise applications.

●        Implement CI/CD pipelines, containerization, and MLOps best practices.

●        Monitor, optimize, and maintain production AI systems.

●        Collaborate with cross-functional teams to deliver business-driven AI solutions.

Required Skills

●       Strong programming skills in Python.

●       Experience with vector databases (e.g., Pinecone, FAISS, ChromaDB) and graph memory systems

●       Knowledge of atleast one agent development framework: Google ADK (preferred), LangChain/LangGraph/LlamaIndex, CrewAI

●       Experience with LLMs, RAG, GenAI, AgenticAI Agents

●       Hands-on experience with FastAPI, and REST APIs.

●       Knowledge of Docker, Kubernetes, Git, CI/CD.

●       Experience with AWS, Azure, or GCP

●       Experience with security compliance, monitoring and observability tools such as AWS CloudWatch, Azure Monitor, Google Cloud Monitoring.


Read more
company logo
Anish N
Posted by Anish N
Bengaluru (Bangalore)
3 - 5 yrs
₹10L - ₹20L / yr
skill iconPython
Generative AI
Agentic AI
LangChain
LlamaIndex
+3 more

Job Description:

We are looking for a hands-on AI Engineer with experience in Generative AI and Agentic AI to build and deploy production-ready AI solutions.

Key Responsibilities:

  • Develop and deploy GenAI and Agentic AI applications.
  • Build RAG pipelines, LLM workflows, and AI agents.
  • Develop solutions using Python, LangChain, LangGraph, LlamaIndex, or similar frameworks.
  • Implement tool calling, context retrieval, and LLM orchestration.
  • Integrate AI solutions with APIs and cloud platforms.
  • Work with AWS/Azure/GCP, Docker, and CI/CD.

Required Skills:

  • Strong Python programming skills.
  • 3+ years of GenAI/Agentic AI experience.
  • RAG and LLM orchestration.
  • LangChain / LangGraph / LlamaIndex / AutoGen / CrewAI / Semantic Kernel.
  • MCP and A2A knowledge.
  • Cloud, APIs, Docker, and CI/CD experience.

Preferred Experience:

Hands-on experience building and deploying production-ready AI solutions.

Read more
Jaipur
3 - 6 yrs
₹7L - ₹10L / yr
skill iconPython
Large Language Models (LLM)
Generative AI
Retrieval Augmented Generation (RAG)
LangChain
+23 more

Location: Jaipur (Work From Office)

Employment Type: Full-Time


We're looking for a GenAI Engineer (LLM Engineer) to build scalable AI-powered SaaS applications using Large Language Models (LLMs). You'll develop intelligent AI workflows, integrate LLMs into production systems, and build secure, high-performance AI solutions.


Key Responsibilities

  • Integrate LLM APIs (OpenAI, Claude, Hugging Face) into production applications.
  • Design and optimize RAG pipelines and prompt engineering workflows.
  • Build and manage Vector Databases (Pinecone, Weaviate, pgvector).
  • Optimize AI performance, latency, and operational cost.
  • Ensure secure, scalable AI architecture.
  • Collaborate with Product and Engineering teams to deliver AI-powered features.


Requirements

  • 3+ years of backend development using Python, Go, or Node.js.
  • Hands-on experience with LLMs, LangChain or LlamaIndex.
  • Strong understanding of RAG, Prompt Engineering, and Vector Databases.
  • Experience with AWS, GCP, or Azure.
  • Knowledge of APIs, Microservices, and AI application development.


Preferred: Experience in SaaS/FinTech, LLMOps, or Model Fine-tuning.

Education: B.Tech, BCA, or equivalent technical qualification.


Apply Now

Application Form: https://zfrmz.com/pAKb2ynfomIsuNwRfRbV?utm_source=cutshort

Read more
company logo
Orenda Finserv
Posted by Orenda Finserv
Ahmedabad
3 - 5 yrs
₹7L - ₹11L / yr
skill iconMachine Learning (ML)
Model Serving
Vision Models
skill iconPython
RESTful APIs
+2 more

About the role

We are building AI systems that read, understand and act on real business documents, bank statements, financial reports, policy documents and forms and putting them into production where accuracy and cost both matters.

This is not a research role and it is not a prompt-writing role. You will own features end to end: pick and deploy open-source models, build the pipelines around them, measure whether they actually work on our documents, drive the cost per document down, and keep the whole thing running in production.

You will work closely with the engineering and product teams, and your work will be directly used by business users from day one.


What you will do

Deploy and evaluate open-source models

  • Select, deploy and benchmark open-source LLMs and vision-language models for specific, narrow use cases not general chat.
  • Build evaluation sets from real documents and define what "good" means numerically (field-level accuracy, extraction recall, hallucination rate) before shipping.
  • Run structured comparisons between models and approaches, and write up the trade-offs so the team can make a decision.
  • Apply quantization, batching and other optimizations to fit models into a sensible GPU budget.

Build and optimize AI orchestration

  • Design multi-step pipelines that combine deterministic code, ML models and LLM calls and know when not to use an LLM.
  • Optimize for latency, cost and reliability: caching, batching, request routing, fallback tiers, retries and graceful degradation.
  • Instrument pipelines so failures are visible and traceable rather than silent.

Ship to production

  • Package models and services with Docker, expose them behind clean APIs, and deploy them to our GPU and CPU infrastructure.
  • Handle the unglamorous production concerns: cold starts, timeouts, concurrency limits, versioning, rollback and monitoring.
  • Own on-call-style responsibility for the AI features you build, including cost tracking.


Must-have skills


Programming & engineering

  • Strong Python: type hints, async/await, dataclasses/Pydantic, clean module design, testing.
  • REST API development with FastAPI (or Flask/Django with a willingness to move to FastAPI).
  • Git, code review discipline, and the ability to write code someone else can maintain.
  • Comfortable in Linux and on the command line.

Machine learning fundamentals

  • Working knowledge of PyTorch and the Hugging Face ecosystem (transformers, tokenizers, accelerate).
  • Understanding of inference-time concepts: tokenization, context windows, batching, precision (FP16/BF16/INT8), memory footprint.
  • Ability to read a model card and a paper well enough to judge whether a model fits a use case.

Document processing

  • Hands-on experience with at least two of: pypdfium2, PyMuPDF, pdfplumber, pdfminer.six, Docling, Unstructured, Surya, DocTR, LayoutLM family.
  • Practical OCR experience (Tesseract, PaddleOCR, or a cloud OCR) and an understanding of when OCR is the wrong tool.
  • Experience extracting tables from PDFs and dealing with merged cells, multi-line rows, and inconsistent column layouts.


Strongly preferred

You will be a much stronger candidate with any of these. We do not expect all of them.

Model serving & optimization

  • vLLM, TGI, Ollama, llama.cpp, or Triton Inference Server.
  • Quantization formats and tooling: GGUF, AWQ, GPTQ, bitsandbytes, ONNX Runtime, INT8 export.
  • Serverless GPU platforms: Modal, RunPod, Replicate, Baseten including cold-start and container-lifecycle management.
  • LoRA / QLoRA fine-tuning with PEFT for narrow, task-specific improvements.

Vision-language models

  • Practical use of open VLMs: Qwen2.5-VL, InternVL, Granite Vision, Molmo, Phi-Vision, or similar.
  • Awareness of where VLMs hallucinate especially on numeric and financial content and patterns for constraining them (using the model for layout only, sourcing values from the text layer, constrained decoding).

Orchestration & pipelines

  • Workflow orchestration: Dagster, Airflow, Prefect, or Temporal.
  • Async job patterns: Celery, RQ, or platform-native spawn/poll patterns.
  • LLM orchestration frameworks (LangGraph, LlamaIndex, Haystack) with the judgement to know when plain Python is a better answer.
  • Structured output enforcement: Instructor, Outlines, XGrammar, JSON schema / tool-use modes.

Evaluation & observability

  • Building golden datasets and regression suites for extraction tasks.
  • Eval tooling: promptfoo, DeepEval, Ragas, or in-house harnesses.
  • LLM tracing and monitoring: Langfuse, Arize Phoenix, LangSmith, OpenTelemetry.

Nice extras

  • Rule engines and policy evaluation (Open Policy Agent / Rego, Drools, rule-engine).
  • Experience in fintech, lending, insurance or accounting documents.
  • Handling of PII and data-security practices in document pipelines.
  • Contributions to open-source ML or document-processing projects.


Why join us

  • Real production ownership from month one your work goes to actual users, not a demo.
  • Genuinely hard technical problems in document AI, not wrappers over an API.
  • Small team, short decision cycles, direct access to leadership.
  • Budget and freedom to evaluate and adopt new open-source models as they land.


To apply: send your CV along with a short note on one AI system you have taken to production what it did, what the accuracy was, and what broke.


Read more
company logo
shwetha V
Posted by shwetha V
Remote only
6 - 12 yrs
Best in industry
skill iconPython
Large Language Models (LLM)
skill iconMachine Learning (ML)
MLOps
Large Language Models (LLM) tuning
+4 more

Principal Software Engineer

Company Summary :


As the recognized global standard for project-based businesses, Deltek delivers software and information solutions to help organizations achieve their purpose. Our market leadership stems from the work of our diverse employees who are united by a passion for learning, growing and making a difference. At Deltek, we take immense pride in creating a balanced, values-driven environment, where every employee feels included and empowered to do their best work. Our employees put our core values into action daily, creating a one-of-a-kind culture that has been recognized globally. Thanks to our incredible team, Deltek has been named one of America's Best Midsize Employers by Forbes, a Best Place to Work by Glassdoor, a Top Workplace by The Washington Post and a Best Place to Work in Asia by World HRD Congress. www.deltek.com


Position Responsibilities :


About the Role 

We are seeking a highly motivated AI Solutions Engineer to join Deltek’s growing AI Center of Excellence team to design, develop, deploy, and optimize internal Artificial Intelligence and Machine Learning solutions that solve complex business challenges. The ideal candidate combines deep expertise in AI, machine learning, Generative AI, Large Language Models (LLMs), SLMs, software engineering, cloud computing, and MLOps/LLMOps to build scalable, production-grade AI applications. 

The AI Solutions Engineer will collaborate with AI data scientists, architects, and engineering teams to deliver innovative AI-driven solutions while ensuring security, scalability, governance, and operational excellence. This role reports to the Senior AI Solutions Architect. 

Key Responsibilities 

AI & Machine Learning Development 

  • Design, build, train, evaluate, and deploy machine learning and deep learning models. 
  • Develop Generative AI solutions using Large Language Models (LLMs) such as GPT, Claude, Gemini, Llama, and Mistral. 
  • Implement Retrieval-Augmented Generation (RAG), prompt engineering, fine-tuning, and AI agent frameworks. 
  • Build NLP, recommendation systems, forecasting, predictive analytics, and intelligent automation solutions. 
  • Optimize model performance, scalability, latency, and cost. 

Software Engineering & Solution Development 

  • Develop production-grade AI applications using Python and modern software engineering practices. 
  • Build APIs, microservices, and AI-powered enterprise applications. 
  • Integrate AI services with enterprise systems, business applications, and data platforms. 
  • Apply coding standards, automated testing, CI/CD, and version control best practices. 

MLOps & AI Operations 

  • Design and implement MLOps pipelines for model development, deployment, monitoring, and lifecycle management. 
  • Automate model training, validation, testing, and deployment processes. 
  • Monitor model performance, data drift, hallucinations, and operational metrics. 
  • Support continuous improvement and reliability of AI platforms. 

Cloud & Platform Engineering 

  • Develop AI solutions on Azure, AWS, or Google Cloud platforms. 
  • Leverage cloud-native AI services, containerization, Kubernetes, and serverless technologies. 
  • Build scalable architectures supporting enterprise AI workloads and real-time inference. 

AI Governance & Security 

  • Ensure compliance with Responsible AI, security, privacy, and regulatory requirements. 
  • Implement model governance, explainability, bias mitigation, and risk management practices. 
  • Maintain standards for secure design, deployment, and operation of AI solutions. 




Required Qualifications 

Education 

  • Bachelor's or Master's degree in Computer Science, Artificial Intelligence, Data Science, Engineering, or a related technical field. 

Experience 

  • 5+ years of software engineering or machine learning development experience. 
  • 2+ years of hands-on experience developing and deploying Agentic AI, Generative AI or AI/ML solutions in production environments. 

Technical Skills 

Programming & Engineering 

  • Strong expertise in Python. 
  • Experience with Java, ReactJS, JavaScript, or similar programming languages. 
  • Solid understanding of algorithms, data structures, APIs, and software design principles. 

Artificial Intelligence & Machine Learning 

  • Machine Learning and Deep Learning concepts and frameworks. 
  • Model training, evaluation, optimization, and deployment. 

Generative AI 

  • Large Language Models (LLMs) & SLMs 
  • Prompt Engineering 
  • Retrieval-Augmented Generation (RAG) 
  • AI Agents and Agentic Workflows 
  • Fine-tuning and model customization 
  • Vector embeddings and semantic search 

Frameworks & Tools 

  • PyTorch, TensorFlow, Scikit-learn 
  • LangChain, LlamaIndex, Semantic Kernel, MCP, A2A and Transformers 
  • FastAPI, Flask 

Data & Analytics 

  • SQL and NoSQL databases 
  • Data pipelines, ETL, and data modeling 
  • Experience with AWS, Azure and Google 

MLOps & DevOps 

  • MLflow, Kubeflow, Azure ML, SageMaker 
  • Docker and Kubernetes 
  • Git, GitHub, Azure DevOps, Jenkins 
  • CI/CD automation and model monitoring 

Cloud Platforms 

  • AWS (preferred) 
  • AWS Bedrock or Azure OpenAI Service 
  • AWS SageMaker 
  • Google Vertex AI 

Preferred Qualifications 

  • Experience designing enterprise-scale AI platforms and products.  
  • Knowledge of multi-agent architectures and autonomous AI systems.  
  • Experience with vector databases such as Pinecone, Snowflake Cortex, Pgvector, Weaviate, Chroma, or Azure AI Search.  
  • Understanding of AI governance, compliance, and Responsible AI frameworks.  
  • Relevant certifications in Azure AI, AWS Machine Learning, or Google Cloud AI.
Read more
company logo
Prajakta Ranade
Posted by Prajakta Ranade
Surat, Mumbai, Navi Mumbai
8 - 15 yrs
₹10L - ₹18L / yr
skill iconPython
LangChain
Vector database
AI Agents
Enterprise architecture
+8 more

Position Overview

We are seeking a highly skilled and experienced Senior AI/ML Developer to lead the development and integration of advanced AI solutions within our product ecosystem. This role involves close collaboration with cross-functional teams including product managers, data scientists, and engineers to build AI models that solve real-world integration challenges. The ideal candidate will have a strong foundation in machine learning, deep learning, and software development, along with hands-on experience deploying AI models in production environments.


Bachelor’s degree in computer science, Data Science, Mathematics, Engineering, or a related field.

8+ years of experience in designing and implementing AI/ML solutions.

Demonstrated ability to integrate AI models into production software.

Excellent analytical thinking, communication, and problem-solving abilities.

Ability to work autonomously as well as in a collaborative team setup.


Skills Required

Dataset Development: Strong track record of building datasets for training and/or evaluating machine learning models.

LLM and NLP Experience: Hands-on experience working with LLMs, RAG architecture, Natural Language Processing (NLP), or applying Machine Learning to solve real-world problems.

Experience with LLM fine-tuning, prompt engineering, vector databases (e.g., Pinecone, FAISS) is highly desirable.

Test Harness Automation for LLM Agents

Familiarity with agent frameworks (e.g., Semantic Kernel, AutoGen, Lang Chain, etc.).

Proficiency in Python and libraries like Pandas, NumPy, Scikit-learn, etc. containerization (Docker), and API frameworks (Flask, Fast API).

Integration Knowledge: API development, data transformation, system integration

Soft Skills: Communication, teamwork, adaptability, critical thinking

Read more
Leadsquared
Leadsquared
Agency job
via by Vrishali Mishra
Bengaluru (Bangalore)
2 - 4 yrs
₹25L - ₹45L / yr
Large Language Models (LLM) tuning

About LeadSquared

LeadSquared is a leading sales execution and marketing automation platform trusted by 2,000+ businesses globally, including healthcare, education, financial services, and real estate. Headquartered in Bengaluru with offices across the US, UK, UAE, and Southeast Asia, we empower sales teams to close faster, smarter, and at scale.

Our AI team is at the forefront of integrating cutting-edge large language model capabilities into enterprise workflows — building intelligent agents, copilots, and automation systems that redefine how businesses operate.

Role Overview

We are looking for a Senior AI Engineer with hands-on experience building LLM-powered agents and agentic AI systems. You will design, develop, and deploy autonomous AI pipelines that solve complex, multi-step business problems — from lead qualification and follow-up automation to intelligent CRM workflows and beyond.

This role is ideal for someone who is deeply excited about the frontier of AI, can move fast, and wants their work to directly impact millions of sales professionals worldwide.

Key Responsibilities

Design and build LLM-powered agentic systems using frameworks such as LangChain, LlamaIndex, AutoGen, or CrewAI to automate complex, multi-step workflows.

Develop and maintain Retrieval-Augmented Generation (RAG) pipelines with vector databases (Pinecone, Weaviate, Chroma, pgvector) for domain-specific knowledge grounding.

Build and integrate tool-use and function-calling capabilities into AI agents, enabling dynamic interaction with internal APIs, databases, and third-party services.

Implement prompt engineering strategies including chain-of-thought, few-shot prompting, and structured output parsing to ensure reliable agent behavior.

Design evaluation frameworks and observability pipelines (LangSmith, Helicone, custom metrics) to monitor agent performance, accuracy, and cost.

Collaborate with product, sales, and domain teams to translate business requirements into AI-driven solutions and features.

Optimize LLM inference for latency and cost using techniques like caching, model distillation, quantization, and batching.

Stay current with the rapidly evolving LLM ecosystem and proactively propose improvements and new approaches.

Contribute to internal best practices, documentation, and knowledge-sharing across the engineering org.

Required Qualifications

Experience

2–4 years of professional software engineering experience, with at least 1–2 years focused on LLM/AI systems.

Proven experience shipping LLM-based products or agentic AI systems into production environments.

Technical Skills

Strong proficiency in Python and familiarity with async programming patterns for AI pipelines.

Hands-on experience with LLM APIs: OpenAI (GPT-4o), Anthropic (Claude), Google (Gemini), or open-source models (Llama, Mistral).

Experience with agentic frameworks: LangChain, LangGraph, LlamaIndex, AutoGen, CrewAI, or similar.

Solid understanding of RAG architectures, embedding models, and semantic search.

Experience with vector databases and similarity search infrastructure.

Knowledge of REST APIs, microservices architecture, and containerization (Docker/Kubernetes).

Problem-Solving & Mindset

Strong ability to decompose ambiguous, open-ended problems into structured AI system designs.

Experience with prompt debugging, LLM evaluation, and iterative refinement workflows.

Ability to balance research exploration with engineering pragmatism to ship reliable systems.

Preferred Qualifications

Experience with multi-agent orchestration and agent memory systems (short-term and long-term).

Familiarity with fine-tuning or RLHF workflows for domain adaptation.

Background in NLP, information retrieval, or conversational AI.

Prior experience in B2B SaaS or CRM domain is a plus.

Contributions to open-source AI/ML projects or published research/blogs.

Experience with cloud platforms: AWS, GCP, or Azure — particularly AI/ML services

Read more
company logo
Mayank Choudhary
Posted by Mayank Choudhary
Bengaluru (Bangalore)
3 - 5 yrs
₹20L - ₹25L / yr
Artificial Intelligence (AI)

Strong AI/ML Engineer Profile

Mandatory (Experience) : Must have 3+ years of experience in software engineering with atleast 1+ years in GenAI application development and production deployment

Mandatory (GenAI Application Development): Must have proven experience building GenAI applications covering RAG pipelines, multi-agent systems, Text2SQL, and fine-tuning

Mandatory (Production GenAI Deployment): Must have expertise deploying production-grade GenAI applications including model evaluation, optimisation, and ownership of full production rollouts

Mandatory (ML & Data Science Tooling): Must have strong hands-on experience with core ML and data science tools including pandas, scikit-learn, and PyTorch

Mandatory (Cloud ML Infrastructure): Must have experience building and deploying production-grade ML workloads on at least one of AWS, Azure, or GCP

Mandatory (Communication): Must have strong English communication skills with the ability to work across time zones and collaborate cross-functionally with product, engineering, and business stakeholders

Mandatory (Note 1) : Role is Hybrid, WFH flexibility as well upto 6 days a month

Mandatory (Note 2) : CTC is inclusive of 10% variable

Mandatory (Note 3): Candidates should be available to join within May 31st or June first week max

Read more
company logo
Faisal AshrafNomani
Posted by Faisal AshrafNomani
Remote only
4 - 15 yrs
Best in industry
Generative AI
Large Language Models (LLM) tuning
Agentic AI
AI Agents
Retrieval Augmented Generation (RAG)

About the Role:

We are looking for an ideal candidate with 5+ years of experience in Data Science / Machine Learning, with strong hands-on experience in Generative AI, Large Language Models (LLMs), NLP, and AI-powered applications. The candidate should be comfortable working across the complete AI lifecycle—from understanding business requirements and experimenting with models to building, evaluating, deploying, and monitoring production-grade GenAI solutions.

The role requires a combination of strong technical expertise, business understanding, problem-solving ability, and stakeholder management skills.



Key Responsibilities:

 

Generative AI & LLM

·      Design, develop, and deploy Generative AI and LLM-based solutions for enterprise use cases.

·      Work with models such as OpenAI, Azure OpenAI, Llama, Mistral, Gemini, or equivalent LLM platforms.

·      Develop applications using prompt engineering, structured outputs, function/tool calling, and LLM orchestration.

·      Design and implement Retrieval-Augmented Generation (RAG) solutions.

·      Work with vector databases and semantic search for enterprise knowledge retrieval.

·      Develop and evaluate AI agents and multi-step AI workflows.

·      Implement techniques such as prompt optimization, context management, grounding, and hallucination reduction.

·      Develop AI solutions for text classification, summarization, information extraction, question answering, document intelligence, and other enterprise use cases.


Machine Learning & Data Science

·      Develop and optimize traditional Machine Learning and statistical models where appropriate.

·      Perform data exploration, feature engineering, model selection, training, validation, and evaluation.

·      Apply appropriate ML and statistical techniques to solve business problems.

·      Work with structured, unstructured, and semi-structured data.

·      Develop scalable data pipelines to support AI/ML solutions.

·      Collaborate with Data Engineers to prepare and manage data for AI applications.


AI Evaluation & Productionization

·      Design evaluation frameworks to measure LLM accuracy, relevance, groundedness, toxicity, latency, and cost.

·      Implement guardrails and responsible AI practices.

·      Monitor model and application performance in production.

·      Identify model/data drift and implement appropriate improvement strategies.

·      Optimize AI solutions for performance, scalability, reliability, and cost.

·      Support deployment and productionization of AI/ML solutions.

·      Client & Delivery Responsibilities

·      Work closely with the CEO, Delivery team, Solution Architects, Engineering teams, and clients to understand business problems and identify AI opportunities.

·      Translate business requirements into practical AI/ML solutions.

·      Participate in client discussions, solution presentations, technical workshops, and POCs.

·      Develop rapid prototypes and demonstrate the feasibility of GenAI solutions.



·      Convert successful POCs into scalable, production-ready applications.

·      Provide technical guidance and contribute to AI solution architecture.

·      Prepare technical documentation, solution approaches, and project estimates where required.

·      Stay current with developments in Generative AI, LLMs, Agentic AI, and AI engineering.

Required Skills:

·       5+ years of hands-on experience in Data Science, Machine Learning, AI, or a related field.

·      Strong practical experience in Generative AI and LLM-based applications.

·      Strong proficiency in Python.

·      Strong understanding of Machine Learning and statistical concepts.

·      Hands-on experience with:

o       LLMs

o       Prompt Engineering

o       RAG

o       Vector Databases

o       Embeddings

o       Semantic Search

o       LLM Evaluation

o       AI Guardrails

·      Experience with frameworks/tools such as LangChain, LangGraph, LlamaIndex, or equivalent.

·      Experience with APIs and integrating LLMs into enterprise applications.

·      Strong SQL and data handling skills.

·      Experience working with large and complex datasets.

·      Strong understanding of NLP concepts.XX



Technical Skills:

·      Experience with OpenAI / Azure OpenAI / AWS Bedrock / Google Vertex AI.

·      Experience with vector databases such as Pinecone, Weaviate, Milvus, FAISS, or equivalent.

·      Experience with Databricks, Snowflake, or cloud data platforms.

·      Experience with Docker and CI/CD.

·      Exposure to AWS, Azure, or GCP.

·      Experience with ML/AI deployment and MLOps.

·       Knowledge of AI security, data privacy, governance, and responsible AI.

·      Experience building AI Agents / Agentic AI workflows.

·      Experience with multimodal AI is an added advantage

Key Competencies

·      Strong analytical and problem-solving ability.

·      Ability to translate business problems into practical AI solutions.

·      Strong communication and presentation skills.

·      Ability to interact confidently with senior stakeholders and clients.

·      Strong ownership and delivery mindset.

·      Ability to work independently in a fast-paced environment.

  • Strong experimentation and innovation mindset.
  • Ability to balance technical feasibility, business value, scalability, and cost.

Required Education & Experience:

·      Bachelor's or Master's degree in Computer Science, Data Science, Artificial Intelligence, Statistics, Mathematics, Engineering, or a related discipline


Read more
Why apply to jobs via Cutshort
people_solving_puzzle
Personalized job matches
Stop wasting time. Get matched with jobs that meet your skills, aspirations and preferences.
people_verifying_people
Verified hiring teams
See actual hiring teams, find common social connections or connect with them directly.
ai_chip
Move faster with AI
We use AI to get you faster responses, recommendations and unmatched user experience.
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos