AI/ML Engineer at Staffnixcom · Bengaluru (Bangalore) · 3 - 5 years · ₹17L - ₹25L / yr · Bootstrapped · Posted 26 May 2026

Strong AI/ML Engineer Profile
Mandatory (Experience) : Must have 3+ years of experience in software engineering with atleast 1+ years in GenAI application development and production deployment
Mandatory (GenAI Application Development): Must have proven experience building GenAI applications covering RAG pipelines, multi-agent systems, Text2SQL, and fine-tuning
Mandatory (Production GenAI Deployment): Must have expertise deploying production-grade GenAI applications including model evaluation, optimisation, and ownership of full production rollouts
Mandatory (ML & Data Science Tooling): Must have strong hands-on experience with core ML and data science tools including pandas, scikit-learn, and PyTorch
Mandatory (Cloud ML Infrastructure): Must have experience building and deploying production-grade ML workloads on at least one of AWS, Azure, or GCP
Mandatory (Communication): Must have strong English communication skills with the ability to work across time zones and collaborate cross-functionally with product, engineering, and business stakeholders

Similar jobs (10)
Strong AI/ML Engineer Profile
Mandatory (Experience) : Must have 3+ years of experience in software engineering with atleast 1+ years in GenAI application development and production deployment
Mandatory (GenAI Application Development): Must have proven experience building GenAI applications covering RAG pipelines, multi-agent systems, Text2SQL, and fine-tuning
Mandatory (Production GenAI Deployment): Must have expertise deploying production-grade GenAI applications including model evaluation, optimisation, and ownership of full production rollouts
Mandatory (ML & Data Science Tooling): Must have strong hands-on experience with core ML and data science tools including pandas, scikit-learn, and PyTorch
Mandatory (Cloud ML Infrastructure): Must have experience building and deploying production-grade ML workloads on at least one of AWS, Azure, or GCP
Mandatory (Communication): Must have strong English communication skills with the ability to work across time zones and collaborate cross-functionally with product, engineering, and business stakeholders
Mandatory (Note 1) : Role is Hybrid, WFH flexibility as well upto 6 days a month
Mandatory (Note 2) : CTC is inclusive of 10% variable
Mandatory (Note 3): Candidates should be available to join within May 31st or June first week max
Strong AI Engineer / Machine Learning Engineer profiles.
2
Mandatory (Experience 1) – Must have minimum 3+ years of hands-on experience in Data Science, Machine Learning, Applied AI, NLP, Deep Learning, or Generative AI solutions.
3
Mandatory (Experience 2) – Must have strong hands-on experience in Python programming, SQL, data analysis, feature engineering, model development, and production-grade ML applications.
4
Mandatory (Experience 3) – Must have experience working with Machine Learning and Deep Learning frameworks such as PyTorch, TensorFlow, Keras, Scikit-learn, or equivalent.
5
Mandatory (Experience 4) – Must have hands-on experience working on NLP, embeddings, semantic search, text classification, document understanding, recommendation systems, or similar AI/ML use cases.
6
Mandatory (Experience 5) – Must have experience working with Large Language Models (LLMs) such as GPT, Llama, Mistral, Claude, Gemini, Phi, or similar foundation models.
7
Mandatory (Experience 6) – Must have hands-on experience building or implementing RAG (Retrieval Augmented Generation) systems, vector search, knowledge retrieval, embeddings, chunking, indexing, or semantic retrieval solutions.
8
Mandatory (Experience 7) – Must have experience working with Git, CI/CD practices, production environments, and scalable AI/ML systems.
9
Mandatory (CTC) – The CTC breakup offered will be 75% fixed + 25% variable, as per company policy.
10
Mandatory (Age) - Candidate's Age should be below 28 Years
11
Preferred (Experience 1) – Experience with MLFlow, Kubeflow, Airflow, Prefect, Feature Stores, Model Registry, or MLOps/LLMOps frameworks.
12
Preferred (Experience 2) – Experience working with Vector Databases, Spark, PySpark, distributed ML pipelines, large-scale data processing, or real-time ML systems..
13
Preferred (Experience 3) – Familiarity with Docker, Kubernetes, Azure, AWS, GCP, cloud-native AI deployments, and scalable ML architecture.
14
Preferred (Company) – Candidates from AI-first startups, Fintech, Banking, Lending, Fraud Analytics, Risk Analytics, Product Companies, SaaS organizations, or data-driven technology companies
15
Mandatory ( Pedigree) - B.TECH / M.TECH from Tier 1 Colleges (IIT's, NIT's, BITS) are Considered.
Location: Jaipur (Work From Office)
Employment Type: Full-Time
We're looking for a GenAI Engineer (LLM Engineer) to build scalable AI-powered SaaS applications using Large Language Models (LLMs). You'll develop intelligent AI workflows, integrate LLMs into production systems, and build secure, high-performance AI solutions.
Key Responsibilities
- Integrate LLM APIs (OpenAI, Claude, Hugging Face) into production applications.
- Design and optimize RAG pipelines and prompt engineering workflows.
- Build and manage Vector Databases (Pinecone, Weaviate, pgvector).
- Optimize AI performance, latency, and operational cost.
- Ensure secure, scalable AI architecture.
- Collaborate with Product and Engineering teams to deliver AI-powered features.
Requirements
- 3+ years of backend development using Python, Go, or Node.js.
- Hands-on experience with LLMs, LangChain or LlamaIndex.
- Strong understanding of RAG, Prompt Engineering, and Vector Databases.
- Experience with AWS, GCP, or Azure.
- Knowledge of APIs, Microservices, and AI application development.
Preferred: Experience in SaaS/FinTech, LLMOps, or Model Fine-tuning.
Education: B.Tech, BCA, or equivalent technical qualification.
Apply Now
Application Form: https://zfrmz.com/pAKb2ynfomIsuNwRfRbV?utm_source=cutshort
The recruiter has not been active on this job recently. You may apply but please expect a delayed response.
Job Description – AI Engineer (End-to-End Development & Deployment)
Role Summary
We are looking for an AI Engineer with hands-on experience in designing, developing, deploying, and maintaining Generative/Agentic AI solutions in production. The ideal candidate should have end-to-end ownership of AI applications, from development to deployment, monitoring, and optimization.
Key Responsibilities
● Design, build, and deploy Generative/Agentic AI solutions.
● Develop applications using LLMs, RAG, AI agents, and vector databases.
● Build scalable APIs and integrate AI solutions with enterprise applications.
● Implement CI/CD pipelines, containerization, and MLOps best practices.
● Monitor, optimize, and maintain production AI systems.
● Collaborate with cross-functional teams to deliver business-driven AI solutions.
Required Skills
● Strong programming skills in Python.
● Experience with vector databases (e.g., Pinecone, FAISS, ChromaDB) and graph memory systems
● Knowledge of atleast one agent development framework: Google ADK (preferred), LangChain/LangGraph/LlamaIndex, CrewAI
● Experience with LLMs, RAG, GenAI, AgenticAI Agents
● Hands-on experience with FastAPI, and REST APIs.
● Knowledge of Docker, Kubernetes, Git, CI/CD.
● Experience with AWS, Azure, or GCP.
● Experience with security compliance, monitoring and observability tools such as AWS CloudWatch, Azure Monitor, Google Cloud Monitoring.
The recruiter has not been active on this job recently. You may apply but please expect a delayed response.
About NonStop io Technologies
NonStop io Technologies is a value-driven company with a strong focus on process-oriented software engineering. We specialize in Product Development and have a decade's worth of experience in building web and mobile applications across various domains. NonStop io Technologies follows core principles that guide its operations and believes in staying invested in a product's vision for the long term. We are a small but proud group of individuals who believe in the 'givers gain' philosophy and strive to provide value in order to seek value. We are committed to and specialize in building cutting-edge technology products and serving as trusted technology partners for startups and enterprises. We pride ourselves on fostering innovation, learning, and community engagement. Join us to work on impactful projects in a collaborative and vibrant environment.
Brief Description:
We're seeking an AI/ML Engineer to join our team. As AI/ML Engineer, you will be responsible for designing, developing, and implementing artificial intelligence (AI) and machine learning (ML) solutions to solve real-world business problems. You will work closely with engineering teams, including software engineers, domain experts, and product managers, to deploy and integrate Applied AI/ML solutions into the products that are being built at NonStop io. Your role will involve researching cutting-edge algorithms and data processing techniques, and implementing scalable solutions to drive innovation and improve the overall user experience.
Responsibilities
● Applied AI/ML engineering; Building engineering solutions on top of the AI/ML tooling available in the industry today. Eg: Engineering APIs around OpenAI
● AI/ML Model Development: Design, develop, and implement machine learning models and algorithms that address specific business challenges, such as natural language processing, computer vision, recommendation systems, anomaly detection, etc.
● Data Preprocessing and Feature Engineering: Cleanse, preprocess, and transform raw data into suitable formats for training and testing AI/ML models. Perform feature engineering to extract relevant features from the data
● Model Training and Evaluation: Train and validate AI/ML models using diverse datasets to achieve optimal performance. Employ appropriate evaluation metrics to assess model accuracy, precision, recall, and other relevant metrics
● Data Visualization: Create clear and insightful data visualizations to aid in understanding data patterns, model behaviour, and performance metrics
● Deployment and Integration: Collaborate with software engineers and DevOps teams to deploy AI/ML models into production environments and integrate them into various applications and systems
● Data Security and Privacy: Ensure compliance with data privacy regulations and implement security measures to protect sensitive information used in AI/ML processes
● Continuous Learning: Stay updated with the latest advancements in AI/ML research, tools, and technologies, and apply them to improve existing models and develop novel solutions
● Documentation: Maintain detailed documentation of the AI/ML development process, including code, models, algorithms, and methodologies for easy understanding and future reference.
Qualifications & Skills
● Bachelor's, Master's, or PhD in Computer Science, Data Science, Machine Learning, or a related field. Advanced degrees or certifications in AI/ML are a plus
● Proven experience as an AI/ML Engineer, Data Scientist, or related role, ideally with a strong portfolio of AI/ML projects
● Proficiency in programming languages commonly used for AI/ML. Preferably Python
● Familiarity with popular AI/ML libraries and frameworks, such as TensorFlow, PyTorch, scikit-learn, etc.
● Familiarity with popular AI/ML Models such as GPT3, GPT4, Llama2, BERT etc.
● Strong understanding of machine learning algorithms, statistics, and data structures
● Experience with data preprocessing, data wrangling, and feature engineering
● Knowledge of deep learning architectures, neural networks, and transfer learning
● Familiarity with cloud platforms and services (e.g., AWS, Azure, Google Cloud) for scalable AI/ML deployment
● Solid understanding of software engineering principles and best practices for writing maintainable and scalable code
● Excellent analytical and problem-solving skills, with the ability to think critically and propose innovative solutions
● Effective communication skills to collaborate with cross-functional teams and present complex technical concepts to non-technical stakeholders
Design and develop Agentic AI systems using LLMs, tools, memory,
workflows, and MCP.
Build production-grade RAG pipelines, including ingestion, chunking,
embeddings, retrieval, reranking, and evaluation.
Implement context engineering strategies for improving LLM accuracy,
relevance, and reliability.
Develop and integrate MCP-based tools and services for AI agents.
Work with LLMs, SLMs, quantized models, and model optimization
techniques for efficient inference.
Develop scalable backend services and APIs for AI applications.
Design databases and data models supporting AI/agentic applications.
Implement AI observability covering latency, token usage, cost, failures,
quality, and agent/tool execution.
Apply AI governance and responsible AI practices, including security,
access control, data privacy, and auditability.
Optimize AI systems for latency, scalability, cost, and reliability.
Collaborate with engineering and product teams to take AI solutions from
POC to production.
Strong hands-on experience with GenAI, LLMs, and Agentic AI.
Experience building RAG applications.
Strong understanding of Context Engineering and prompt/context
optimization.
Role Overview
We are looking for a hands-on AI/ML Engineer to design, develop, and deploy
production-ready GenAI and Agentic AI applications. The role involves building
intelligent agents, RAG pipelines, AI APIs, backend services, and scalable AI
infrastructure with a strong focus on context engineering, observability,
governance, and model optimisation.
Key Responsibilities
Required Skills
Practical experience with MCP (Model Context Protocol).
Experience with frameworks such as LangChain, LangGraph,
LlamaIndex, or equivalent.
Knowledge of LLM/SLM deployment and quantization techniques.
Strong Python backend development experience.
Experience developing REST APIs using FastAPI/Flask or equivalent.
Strong understanding of SQL/NoSQL databases and database design.
Experience with vector databases such as Qdrant, Pinecone, Weaviate,
ChromaDB, or FAISS.
Understanding of AI observability, evaluation, monitoring, and
governance.
Experience with cloud platforms and production deployment is preferred.
Strong understanding of software engineering principles, Git, testing, and
CI/CD.
The recruiter has not been active on this job recently. You may apply but please expect a delayed response.
Key Responsibilities
- Design, build, and optimize scalable data pipelines for AI/ML applications.
- Develop, train, evaluate, and deploy Machine Learning and Deep Learning models.
- Build production-ready LLM applications using Retrieval-Augmented Generation (RAG), prompt engineering, and vector databases.
- Fine-tune open-source and foundation models using domain-specific datasets.
- Develop and maintain end-to-end MLOps pipelines for model deployment, monitoring, and lifecycle management.
- Perform data preprocessing, feature engineering, exploratory data analysis (EDA), and model evaluation.
- Develop APIs and AI services for production deployment.
- Collaborate with cross-functional teams to deliver scalable AI-driven solutions.
- Monitor model performance, troubleshoot production issues, and maintain technical documentation.
Required Skills
Mandatory
- 1–3 years of experience in Data Science, Data Engineering, or AI/ML development.
- Strong programming skills in Python and SQL.
- Hands-on experience with Machine Learning frameworks such as PyTorch, TensorFlow, or Scikit-learn.
- Experience building LLM-powered applications using RAG, Prompt Engineering, and Embeddings.
- Hands-on experience with LangChain, LlamaIndex, CrewAI, or n8n for LLM orchestration and AI workflow automation.
- Experience in LLM fine-tuning and working with Hugging Face models.
- Knowledge of MLOps concepts including model deployment, monitoring, versioning, and CI/CD.
- Experience with Git, REST APIs, Linux environments, and data processing libraries.
Preferred
- Experience with vector databases such as Pinecone, Chroma, Milvus, or Weaviate.
- Familiarity with Docker, Kubernetes, and MLflow.
- Exposure to Apache Spark or Airflow for data engineering workflows.
- Experience with cloud platforms (AWS, Azure, or GCP).
Primary Technology Stack
- Languages & Data Processing: Python, SQL, Pandas, NumPy, Apache Spark
- AI & Machine Learning: PyTorch, TensorFlow, Scikit-learn
- Application Frameworks: LangChain, LlamaIndex, CrewAI, n8n
- Core Methodologies: Retrieval-Augmented Generation (RAG), Model Fine-Tuning, Prompt Engineering, Embeddings
- Models & Infrastructure: OpenAI APIs, Hugging Face Ecosystem, Embedding Models
- Vector Databases: Pinecone, Chroma, Milvus, Weaviate
- Databases: PostgreSQL, MongoDB
- MLOps & DevOps: Docker, Kubernetes, MLflow, CI/CD, Git
- Cloud Platforms: AWS, Azure, GCP
Experience: 1–3 Years
Domain: Data Science | Data Engineering | Machine Learning | Generative AI | MLOps
Job Description:
We are looking for a hands-on AI Engineer with experience in Generative AI and Agentic AI to build and deploy production-ready AI solutions.
Key Responsibilities:
- Develop and deploy GenAI and Agentic AI applications.
- Build RAG pipelines, LLM workflows, and AI agents.
- Develop solutions using Python, LangChain, LangGraph, LlamaIndex, or similar frameworks.
- Implement tool calling, context retrieval, and LLM orchestration.
- Integrate AI solutions with APIs and cloud platforms.
- Work with AWS/Azure/GCP, Docker, and CI/CD.
Required Skills:
- Strong Python programming skills.
- 3+ years of GenAI/Agentic AI experience.
- RAG and LLM orchestration.
- LangChain / LangGraph / LlamaIndex / AutoGen / CrewAI / Semantic Kernel.
- MCP and A2A knowledge.
- Cloud, APIs, Docker, and CI/CD experience.
Preferred Experience:
Hands-on experience building and deploying production-ready AI solutions.
Senior Generative AI Engineer
Employment Type: Permanent with VDart Digital
Work Location: Marathalli, Bengaluru
Job Description
We are seeking a highly skilled Senior Generative AI Engineer with strong expertise in designing, developing, and deploying enterprise-scale AI solutions using Large Language Models (LLMs) and modern Generative AI frameworks. The ideal candidate should have hands-on production experience building scalable GenAI applications, AI agents, autonomous workflows, and Retrieval-Augmented Generation (RAG) systems in cloud-native environments.
This role requires deep technical expertise in LLM orchestration, AI application architecture, prompt engineering, vector databases, MLOps, and production deployment of AI systems. Candidates should have proven experience delivering real-world AI solutions in enterprise environments with strong exposure to cloud platforms and DevOps practices.
Key Responsibilities
- Design, build, and deploy enterprise-grade Generative AI applications using Large Language Models (LLMs).
- Develop intelligent AI agents and autonomous workflows using frameworks such as LangChain, CrewAI, LangGraph, AutoGen, or similar agentic AI frameworks.
- Implement and optimize Retrieval-Augmented Generation (RAG) pipelines using vector databases and semantic search technologies.
- Work extensively on prompt engineering, tool calling, memory management, agent orchestration, and multi-agent systems.
- Integrate and manage LLMs such as OpenAI, Azure OpenAI, Claude, Gemini, Llama, Mistral, or similar foundation models.
- Develop scalable AI services and APIs using Python and FastAPI.
- Build production-ready AI solutions with high availability, scalability, monitoring, and observability.
- Deploy and manage AI applications in cloud-native environments using Docker and Kubernetes.
- Collaborate with Data Science, ML Engineering, and DevOps teams to operationalize AI solutions.
- Implement CI/CD pipelines and automated deployment processes for AI workloads.
- Monitor model performance, latency, reliability, and operational efficiency in production environments.
- Ensure AI solutions follow enterprise security, governance, and responsible AI standards.
- Evaluate and adopt emerging Generative AI tools, frameworks, and models.
Required Skills
Generative AI & LLM Expertise
- Strong hands-on experience with Generative AI and Large Language Models (LLMs).
- Production-level experience building and deploying GenAI applications.
- Expertise in LangChain, CrewAI, LangGraph, AutoGen, or similar frameworks.
- Experience with AI agents, autonomous workflows, and multi-agent architectures.
- Strong understanding of prompt engineering, embeddings, model evaluation, and LLM orchestration.
- Experience integrating OpenAI, Azure OpenAI, Claude, Gemini, Llama, Mistral, or similar models.
RAG & Vector Databases
- Strong experience implementing RAG pipelines and semantic retrieval systems.
- Experience with vector databases such as Pinecone, Weaviate, ChromaDB, FAISS, or Milvus.
- Understanding of chunking strategies, embeddings, indexing, reranking, and retrieval optimization.
Python & AI Development
- Strong proficiency in Python.
- Experience with FastAPI for AI service and API development.
- Experience with AI/ML libraries and data processing tools such as Pandas and NumPy.
Cloud & Production Deployment
- Mandatory production experience on at least one cloud platform:
- Microsoft Azure
- Experience deploying scalable AI applications in enterprise production environments.
- Hands-on experience with Docker, Kubernetes, Jenkins, Terraform, and CI/CD pipelines.
- Strong understanding of MLOps, AI deployment lifecycle, monitoring, and observability.
Engineering & Operational Excellence
- Strong understanding of software engineering best practices.
- Experience with Git, version control, automated testing, and release management.
- Experience building secure, scalable, and high-performance AI solutions.
- Ability to troubleshoot production AI systems and optimize performance.
Preferred Skills
- Experience with AI observability and evaluation frameworks.
- Exposure to fine-tuning, PEFT, LoRA, or model optimization techniques.
- Experience with enterprise AI governance and responsible AI practices.
- Knowledge of distributed AI systems and scalable inference architectures.
- Familiarity with AI security and compliance standards.
Qualifications
- Bachelor’s or Master’s degree in Computer Science, Artificial Intelligence, Data Science, or a related field.
- 3–8 years of overall software engineering experience.
- Minimum 3+ years of hands-on experience in Generative AI and LLM-based application development,
- Proven track record of delivering enterprise-scale AI solutions in production environments.
- Strong communication and stakeholder management skills.
Strong AI Engineer / Machine Learning Engineer profiles.
2
Mandatory (Experience 1) – Must have minimum 3+ years of hands-on experience in Data Science, Machine Learning, Applied AI, NLP, Deep Learning, or Generative AI solutions.
3
Mandatory (Experience 2) – Must have strong hands-on experience in Python programming, SQL, data analysis, feature engineering, model development, and production-grade ML applications.
4
Mandatory (Experience 3) – Must have experience working with Machine Learning and Deep Learning frameworks such as PyTorch, TensorFlow, Keras, Scikit-learn, or equivalent.
5
Mandatory (Experience 4) – Must have hands-on experience working on NLP, embeddings, semantic search, text classification, document understanding, recommendation systems, or similar AI/ML use cases.
6
Mandatory (Experience 5) – Must have experience working with Large Language Models (LLMs) such as GPT, Llama, Mistral, Claude, Gemini, Phi, or similar foundation models.
7
Mandatory (Experience 6) – Must have hands-on experience building or implementing RAG (Retrieval Augmented Generation) systems, vector search, knowledge retrieval, embeddings, chunking, indexing, or semantic retrieval solutions.
8
Mandatory (Experience 7) – Must have experience working with Git, CI/CD practices, production environments, and scalable AI/ML systems.
9
Mandatory (CTC) – The CTC breakup offered will be 75% fixed + 25% variable, as per company policy.
10
Mandatory (Age) - Candidate's Age should be below 30 Years
11
Preferred (Experience 1) – Experience with MLFlow, Kubeflow, Airflow, Prefect, Feature Stores, Model Registry, or MLOps/LLMOps frameworks.
12
Preferred (Experience 2) – Experience working with Vector Databases, Spark, PySpark, distributed ML pipelines, large-scale data processing, or real-time ML systems..
13
Preferred (Experience 3) – Familiarity with Docker, Kubernetes, Azure, AWS, GCP, cloud-native AI deployments, and scalable ML architecture.
14
Preferred (Company) – Candidates from AI-first startups, Fintech, Banking, Lending, Fraud Analytics, Risk Analytics, Product Companies, SaaS organizations, or data-driven technology companies
15
Mandatory ( Pedigree) - B.TECH / M.TECH from Tier 1 Colleges (IIT's, NIT's, BITS) are Considered.






