Cutshort logo
For Employers
Talent Pro  logo
Test Engineer - AI/ML,
Test Engineer - AI/ML,

Test Engineer - AI/ML, at Talent Pro · Delhi · 5 - 8 years · ₹11L - ₹15L / yr · Bootstrapped · Posted 17 Jan 2026

Talent Pro 's logo

Test Engineer - AI/ML,

Mayank choudhary's profile picture
Posted by Mayank choudhary
5 - 8 yrs
₹11L - ₹15L / yr
Delhi
Skills
AI/ML tools

Mandatory (Total Experience): Must have 5+ years of overall experience in Testing/QA

Mandatory (Experience 1): Must have 2+ years of experience in testing AI/ML models and data-driven applications, across NLP, recommendation engines, fraud detection, and advanced analytics models

Mandatory (Experience 2): Must have expertise in validating AI/ML models for accuracy, bias, explainability, and performance, ensuring decisions are fair, reliable, and transparent

Mandatory (Experience 3): Must have strong experience to design AI/ML test strategies, including boundary testing, adversarial input simulation, and anomaly monitoring to detect manipulation attempts by marketplace users (buyers/sellers)

Mandatory (Tools & Frameworks): Proficiency in AI/ML testing frameworks and tools (like PyTest, TensorFlow Model Analysis, MLflow, Python-based data validation libraries, Jupyter) with the ability to integrate into CI/CD pipelines

Mandatory (Business Awareness): Must understand marketplace misuse scenarios, such as manipulating recommendation algorithms, biasing fraud detection systems, or exploiting gaps in automated scoring

Mandatory (Communication Skills): Must have strong verbal and written communication skills, able to collaborate with data scientists, engineers, and business stakeholders to articulate testing outcomes and issues.

Mandatory (Education):Degree in Engineering, Computer Science, IT, Data Science, or a related discipline (B.E./B.Tech/M.Tech/MCA/MS or equivalent)

Mandatory Location: Candidate must be based within Delhi NCR (100 km radius)

Preferred

Preferred (Certifications): Certifications such as ISTQB AI Testing, TensorFlow, Cloud AI, or equivalent applied AI credentials are an added advantage.

Read more
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos

About Talent Pro

Founded :
2024
Type :
Services
Size
Stage :
Bootstrapped

About

N/A

Company social profiles

bloglinkedin

Similar jobs (10)

company logo
Leena Lahari
Posted by Leena Lahari
Mumbai
2 - 4 yrs
₹6L - ₹15L / yr
Generative AI
skill iconPython
Test Automation (QA)
Manual testing
Functional testing
+7 more

Title: AI/ML Test Engineer – GenAI

Location - Mumbai

Technical Skills


• Strong experience in Generative AI, LLMs, and Agentic AI systems

• Hands-on expertise with AI evaluation frameworks (RAGAS, DeepEval, TruLens, LangSmith, Promptfoo, etc.)

• Proficiency in Python and AI/ML development libraries

• Knowledge of Prompt Engineering, prompt testing, and optimization

Ability to define and track evaluation metrics such as accuracy, relevance, groundedness, hallucination rate, latency, and user satisfaction

• Experience in creating automated evaluation pipelines and benchmarking frameworks

• Strong understanding of AI safety, guardrails, bias testing, and responsible AI practices

• Familiarity with REST APIs, JSON, vector databases, and knowledge retrieval systems

• Experience in A/B testing, human-in-the-loop evaluation, and red teaming

• Strong experience in Manual Testing of AI/GenAI applications, including functional, exploratory, UAT, regression, and end-to-end testing

• Expertise in validating Agent Reasoning, Tool Calling, Workflow Execution, and Response Quality

• Hands-on experience in Automation Testing using Python frameworks


Key Responsibilities

  • Design, execute, and automate evaluation strategies for Agentic AI applications.
  • Develop evaluation datasets, test cases, and benchmark suites.
  • Measure and improve agent performance, reasoning quality, tool usage, and workflow effectiveness.
  • Analyze model outputs and identify hallucinations, biases, safety risks, and failure patterns.
  • Collaborate with AI Engineers, Product Teams, and Domain Experts to improve agent quality and reliability.
  • Generate evaluation reports, dashboards, and actionable recommendations.


Read more
company logo
Hyderabad
2 - 5 yrs
₹11L - ₹15L / yr
Manual testing
Automation testing

Strong QA Engineer Profile with manual + automation testing across web, API, and AI-driven features

2

Mandatory (Experience 1): Must have 2+ years in QA / software testing, covering both manual and automation testing

3

Mandatory (Experience 2): Must have experience designing, writing, and maintaining test plans, test cases, and test data for web, API, and AI-driven features

4

Mandatory (Experience 3): Must have experience performing functional, regression, integration, and exploratory testing across releases

5

Mandatory (Tech skill 1): Must have hands-on automation experience with a framework such as Selenium, Playwright, Cypress, or similar

6

Mandatory (Tech skill 2): Experience testing AI/ML or data products (validating accuracy, consistency, edge cases, bias, graceful failure handling)

7

Mandatory (Tech skill 3): Must have API testing experience with tools like Postman, REST Assured, or equivalent, and experience testing REST APIs and data pipelines

8

Mandatory (Tech skill 4): Must have working knowledge of Python, JavaScript, or Java for test automation

9

Mandatory (Tech skill 5): Must have SQL and data skills — able to write queries to validate data and back-end behaviour

10

Mandatory (Skill 1): Must have a strong grasp of QA methodologies, SDLC/STLC, and Agile/Scrum practices.

11

Mandatory (Skill 2): Must have an analytical mindset with sharp attention to detail and clear defect communication

12

Mandatory (Skill 3): Must have strong communication skills

13

Preferred (CI/CD): Experience integrating automated tests into CI/CD pipelines (Jenkins, GitHub Actions)

14

Preferred (Domain): Cloud/AWS and insurance domain exposure

Read more
company logo
Hyderabad
2 - 15 yrs
₹11L - ₹15L / yr
Manual testing
Automation testing

Strong QA Engineer Profile with manual + automation testing across web, API, and AI-driven features

2

Mandatory (Experience 1): Must have 2+ years in QA / software testing, covering both manual and automation testing

3

Mandatory (Experience 2): Must have experience designing, writing, and maintaining test plans, test cases, and test data for web, API, and AI-driven features

4

Mandatory (Experience 3): Must have experience performing functional, regression, integration, and exploratory testing across releases

5

Mandatory (Tech skill 1): Must have hands-on automation experience with a framework such as Selenium, Playwright, Cypress, or similar

6

Mandatory (Tech skill 2): Experience testing AI/ML or data products (validating accuracy, consistency, edge cases, bias, graceful failure handling)

7

Mandatory (Tech skill 3): Must have API testing experience with tools like Postman, REST Assured, or equivalent, and experience testing REST APIs and data pipelines

8

Mandatory (Tech skill 4): Must have working knowledge of Python, JavaScript, or Java for test automation

9

Mandatory (Tech skill 5): Must have SQL and data skills — able to write queries to validate data and back-end behaviour

10

Mandatory (Skill 1): Must have a strong grasp of QA methodologies, SDLC/STLC, and Agile/Scrum practices.

11

Mandatory (Skill 2): Must have an analytical mindset with sharp attention to detail and clear defect communication

12

Mandatory (Skill 3): Must have strong communication skills

13

Preferred (CI/CD): Experience integrating automated tests into CI/CD pipelines (Jenkins, GitHub Actions)

14

Preferred (Domain): Cloud/AWS and insurance domain exposure

Read more
Leadsquared
Leadsquared
Agency job
via by Vrishali Mishra
Bengaluru (Bangalore)
2 - 4 yrs
₹15L - ₹35L / yr
Test Automation (QA)
API
skill iconPython
skill iconJavascript
Regression Testing
+2 more

About Us

We’re building the next generation of AI-powered business software, and we’re looking for people who want to shape that future with us. With Lumen, we’re reimagining how users interact with CRM — moving beyond screens, menus and dashboards to an intelligent interface where users can simply ask AI to take actions, retrieve knowledge, generate insights and get work done. With Agent Studio, we’re enabling businesses to build, test and deploy their own AI agents for real-world workflows. And with Invorto, we’re bringing AI to voice, allowing businesses to create intelligent voice agents tailored to their customer and operational use cases.

What makes this especially exciting is the stage and scale of the opportunity. You’ll get to work on genuinely hard problems across LLMs, agents, reasoning, orchestration, voice AI, evaluation, reliability and enterprise security — not as isolated experiments, but as products used in real business workflows. You’ll have the opportunity to build zero-to-one, own meaningful parts of the product end-to-end, work closely with customers, experiment rapidly, and see your work reach production at scale.

Why join now? Because the playbook for enterprise AI is still being written. You won’t just be implementing someone else’s roadmap — you’ll help define the product, architecture and experiences that become that playbook. Expect high ownership, fast iteration, hard technical and product problems, direct customer impact, and the chance to build AI systems that have to work reliably in the real world — not just in a demo.

About the Role

We are looking for a QA Engineer who specializes in testing agentic AI platforms. You will design and automate quality processes for systems that involve LLMs, autonomous agents, tool use and orchestration across Lumen and Agent Studio — ensuring that AI-driven workflows behave reliably, safely and predictably in production, not just in a demo.

What You’ll Do

  • Design and build automated test suites and evaluation frameworks for agentic AI workflows, including multi-step and tool-calling behaviors.
  • Use AI/LLM-based QA tools and evaluation frameworks to test model outputs, agent decisions and end-to-end task completion at scale.
  • Define quality metrics and benchmarks for agent reliability, correctness, latency and safety, and track them over releases.
  • Identify edge cases, failure modes and regressions specific to non-deterministic AI systems, and build automated checks to catch them early.
  • Integrate automated agent/LLM testing into CI/CD pipelines to support fast, reliable iteration.
  • Partner closely with AI/ML and backend engineers to reproduce issues, root-cause failures and validate fixes.
  • Work with customers and customer-facing teams to understand real-world usage patterns and translate them into test scenarios.

What We’re Looking For

  • 2–4 years of QA/test automation experience, including hands-on work testing agentic AI or LLM-based platforms.
  • Practical experience using AI-focused QA/evaluation tools to test agent behavior, prompts and model outputs.
  • Strong scripting/automation skills (Python preferred) to build and maintain test frameworks.
  • Understanding of how LLM-based agents work — tool calling, orchestration, memory, reasoning chains — well enough to design meaningful test cases.
  • Comfort working with non-deterministic systems and designing evaluation approaches beyond traditional pass/fail testing.
  • Strong communication skills — this is a customer-facing role, and you will be expected to clearly articulate technical concepts, decisions and trade-offs to both technical and non-technical stakeholders, including customers.

Good to Have

  • Experience testing voice AI or real-time conversational systems.
  • Familiarity with CRM or enterprise SaaS platforms.
  • Exposure to enterprise security or compliance testing for AI systems.


Read more
company logo
Apoorva Jain
Posted by Apoorva Jain
Bengaluru (Bangalore)
4 - 9 yrs
₹12L - ₹22L / yr
Software Testing (QA)
Automation
skill iconPython
Large Language Models (LLM)
Retrieval Augmented Generation (RAG)
+1 more

Required Skills:-

  • 4+ years of experience in Software Testing (Manual & Automation).
  • Strong experience with Selenium or Playwright.
  • Hands-on experience in API Testing (Postman, REST Assured, Swagger, etc.).
  • Experience with CI/CD pipelines (Jenkins, GitLab CI, Azure DevOps, GitHub Actions).
  • Strong understanding of Functional, Regression, and Non-functional Testing.
  • Excellent analytical and debugging skills.

AI-Specific Skills

  • Hands-on experience in GenAI / AI Testing.
  • Experience testing LLM-based applications.
  • Strong understanding of prompt engineering and prompt validation.
  • Knowledge of LLM behavior, model variability, and non-deterministic outputs.
  • Experience validating AI outputs for accuracy, hallucinations, bias, and safety.
  • Test data management for AI applications.
  • Understanding of Responsible AI testing concepts.


Read more
Bengaluru (Bangalore)
8 - 10 yrs
₹26L - ₹40L / yr
IDP
User Interface (UI) Design

About the team


SecurITe’s mission is to build an Agentic‑AI driven security platform that protects critical infrastructure from modern cyber threats. Our focus is on delivering highly performant, resilient, and intelligent network security systems that help defenders stay ahead of adversaries.


About the Role

We’re looking for a seasoned Senior Quality Engineer to provide technical leadership and architectural oversight for our next‑generation cybersecurity AI platform. In this high-impact role, you will define the technical strategy for quality assurance, ensuring our agentic AI transforms cyber defense with unparalleled reliability.

You will be responsible for the end-to-end quality lifecycle, from architectural reviews to the deployment of scalable automation frameworks. Beyond technical execution, you will serve as a mentor to junior team members, fostering a culture of technical excellence and driving the strategy that ensures our solutions meet the rigorous demands of critical infrastructure protection.


What You’ll Do

●    Defining and driving comprehensive QA strategies and roadmaps for the cybersecurity platform.

●    Designing, developing, and executing test plans, test cases, and automated scripts to ensure software quality.

●    Performing functional, regression, performance, scalability and security testing to identify bugs or defects.

●    Collaborating with developers, product managers, and other stakeholders to understand product requirements and testing needs.

●    Identifying, documenting, and tracking software defects, ensuring clear communication of issues and their resolutions.

●    Leading deep-dive root-cause analysis for critical system defects and security vulnerabilities.

●    Conducting thorough reviews of product specifications and software design to identify potential areas of concern before testing.

●    Architecting and designing complex, scalable test automation frameworks to optimize CI/CD velocity.

●    Ensuring the software meets customer and business requirements by validating the functionality and performance.

●    Assisting in continuously improving QA processes, tools, and best practices to enhance software testing efficiency and effectiveness.

●    Supporting user acceptance testing (UAT) and assisting clients with product validation.

●    Mentoring junior and mid-level engineers, providing technical guidance and conducting architectural reviews.


Required Experience

●    A Bachelor’s degree in Computer Science, Information Technology, Computer Engineering, or a related field.

●    8-10 years of proven experience in quality engineering, specifically within network cybersecurity, Identity Providers, or AI-integrated platforms.

●    Expertise in manual and automated testing.

●    Deep domain expertise in complex system validation and advanced automation practices at scale.

●    Proficiency in programming languages like Python to build and run automated test scripts.

●    "Strong knowledge of software testing methodologies, performance testing tools (e.g., JMeter, k6), and security traffic generation/simulation tools (e.g., Ixia BreakingPoint, Scapy, or Snort/Suricata traffic generators)."

●    Understanding of continuous integration/continuous deployment (CI/CD) pipelines and version control systems like Git.

●    Strong communication skills for documenting test results and interacting with cross-functional teams.

●    Excellent analytical skills, attention to detail, and problem-solving ability.

●    Ability to work independently as well as collaboratively in a team environment.

●    A curious mindset with a willingness to quickly learn new technologies and testing tools.

Required Skills & Qualifications

●    Familiarity with cloud-based testing environments (GCP, AWS, Azure).

●    Experience with cybersecurity products or cloud services or IDP or Web UI


The Mindset

●    Problem Solver: You thrive on complex, ambiguous challenges and engineer elegant solutions.

●    Ownership‑Driven: You take initiative, move fast, and deliver outcomes without hand‑holding.

●    Continuous Learner: You stay ahead of the curve in AI, ML, and emerging technologies.

●    Startup DNA: You excel in fast‑moving environments where priorities evolve and impact is immediate.

 

Read more
NA
NA
Agency job
via by Ramya Munirathnam
Hyderabad, Pune
10 - 14 yrs
₹6L - ₹14L / yr
Generative AI
skill iconC#
skill icon.NET
AI Agents

AI Engineer


We are seeking an AI Engineering specialist focused on AI evaluation, and continuous quality improvement for Ezra MetLife's employee-facing AI platform. This role will establish and scale the testing strategy for enterprise AI agents, ensuring high response quality, reliability, and production readiness. The engineer will build automated regression testing framework (preferred Playwright ), define AI evaluation methodologies, analyze AI performance metrics, and partner with engineering teams to continuously improve answer quality, grounding accuracy, and customer experience. This position is critical to enabling confidence as Ezra expands its AI agent portfolio and employee-facing capabilities.


Required Skills & Experience


• C# and .NET development experience


• Experience with at least one AI evaluation framework (e.g., prompt evaluation, RAG evaluation, LLM quality assessment)


• Microsoft Agent Framework (preferred) or similar enterprise agent frameworks


• Experience with Azure OpenAI / Azure AI Foundry


• Microsoft 365 Agent SDK


• Azure AI Search, RAG pipelines, and retrieval quality testing


• Infrastructure as Code using Terraform


• Experience building automated testing and AI quality validation processes


• Familiarity with telemetry analysis, AI observability, and performance measurement


• Strong analytical skills with a passion for improving AI response quality and reliability


Skills: AI Agents~Core .NET Technologies~C# 5.0

Experience Required: 10 & Above

Location: Hyderabad :5+ relevant exp in AI + .NET

Read more
Remote only
5 - 7 yrs
₹7L - ₹12L / yr
Playwright
TypeScript
skill iconJavascript
  • Web Application Testing: Design, document, and execute comprehensive functional test plans and test cases for complex, highly interactive web applications, ensuring they meet specified requirements and provide an excellent user experience.
  • Backend API Testing: Possess deep expertise in validating backend RESTful and/or SOAP APIs. This includes testing request/response payloads, status codes, data integrity, security, and robust error handling mechanisms.
  • Data Validation with SQL: Write and execute complex SQL queries (joins, aggregations, conditional logic) to perform backend data checks, verify application states, and ensure data integrity across integration points.
  • I Automation (Playwright & TypeScript):
  • Design, develop, and maintain robust, scalable, and reusable UI automation scripts using Playwright and TypeScript.
  • Integrate automation suites into Continuous Integration/Continuous Deployment (CI/CD) pipelines.
  • Implement advanced automation patterns and frameworks (e.g., Page Object Model) to enhance maintainability.
  • Prompt-Based Automation: Demonstrate familiarity or hands-on experience with emerging AI-driven or prompt-based automation approaches and tools to accelerate test case generation and execution.
  • API Automation: Develop and maintain automated test suites for APIs to ensure reliability and performance.
  • JMeter Proficiency: Utilize Apache JMeter to design, script, and execute robust API load testing and stress testing scenarios.
  • Analyze performance metrics, identify bottlenecks (e.g., response time, throughput), and provide actionable reports to development teams.
  • Experience: 4+ years of professional experience in Quality Assurance and Software Testing, with a strong focus on automation.
  • Automation Stack: Expert-level proficiency in developing and maintaining automation scripts using Playwright and TypeScript.
  • Testing Tools: Proven experience with API testing tools (e.g., Postman, Swagger) and strong functional testing methodologies.
  • Database Skills: Highly proficient in writing and executing complex SQL queries for data validation and backend verification.
  • Performance: Hands-on experience with Apache JMeter for API performance and load testing.
  • Communication: Excellent communication and collaboration skills to work effectively with cross-functional teams (Developers, Product Managers).
  • Problem-Solving: Strong analytical and debugging skills to efficiently isolate and report defects.


Read more
company logo
manipriya Arumugam
Posted by manipriya Arumugam
Bengaluru (Bangalore), Mumbai, Hyderabad, Delhi, Gurugram, Noida, Ghaziabad, Faridabad
3 - 6 yrs
₹5L - ₹10L / yr
Software Testing (QA)
uipath

We are seeking a forward-thinking Mid-Level QA Automation Engineer specializing in RPA, financial systems validation, and AI engineering. You will build and scale end-to-end automation workflows using UiPath for our financial services platform. Additionally, you will pioneer our testing evolution by leveraging prompt engineering, creating custom AI agents, and implementing Model Context Protocol (MCP) servers to bridge AI models with our testing infrastructure.

 

Key Responsibilities and Duties

   RPA Automation:

·    Design, develop, and maintain robust automated testing workflows using UiPath Studio.

 AI Agent Engineering:

·   Architect, build, and orchestrate custom AI agents to autonomously generate,execute, and self-heal test scripts.

Prompt Engineering:

·   Design, optimize, and manage advanced prompt templates to drive deterministic,high-quality code and test case generation from LLMs.

 MCP Integration:

·    Implement and configure Model Context Protocol (MCP) ecosystems to securely connect AI agents with local data, development tools, and testing environments. 

 Financial System Testing:

·   Validate end-to-end trade life cycles, market data processing, and capital market clearing workflows. 

Regression Orchestration

·   Manage continuous regression schedules via UiPath Orchestrator to ensure financial system stability. 

Mandatory skills

Project Mandatory Skills

Project Desired Skills

UiPath Studio and Orchestrator (UI/API automation activities)

Selenium or Playwright for web-based application testing

Building, deploying, and scripting autonomous AI agents for complex engineering workflows

Calypso treasury and capital markets platform

Prompt engineering — context injection, few-shot prompting, and guiding LLMs to exact test requirements

FitNesse for acceptance testing and collaborative documentation

Model Context Protocol (MCP) — using or developing MCP servers/clients to connect AI tools to external data sources and developer environments

TypeScript, Python, or Java to support agent tool calling and framework extensions

 Qualifications

  • 3+ years of professional software QA automation or RPA development experience.
  •  Strong understanding of capital markets, trading lifecycles, treasury workflows, or core banking architectures.
  • Strong communication and collaboration skills to work across engineering, QA, and business teams.


Read more
company logo
Hyderabad, Chennai
12 - 17 yrs
Best in industry
Playwright
Test Automation (QA)
  • 12 - 16 years of experience in Enterprise Software Product Engineering (Development/Quality) 
  • Working experience with Enterprise SaaS product companies 
  • 6+ years of experience in leading/managing multi-discipline activities - manual, automation, performance and security testing. 
  • Experience with AI powered test automation frameworks like Playwright, TestRigor etc is preferable. 
  • Experience with security and performance testing tools like BurpSuite, Locust, K6 etc   
  • Strong experience in generating and presenting data/metric driven reports to exec/C-level stakeholders. 
  • Self drive to achieve team and organization goals with high ROI 
  • Result oriented collaborator and enabler of success. 
  • Preferred: 
  • Enterprise software experience withing the Life Sciences industry and good understanding of GxP requirements 
  • Knowledge of cloud native, distributed architecture 

Responsibilities:

  • Lead the overall strategy, execution, reporting of ValGenesis’s platform/integration testing activities 
  • Own the strategic vision and establish the Software QA processes, tools, frameworks. Ensure, Monitor and Measure of the same in terms of rollout, adoption, results. 
  • Change advocate of lean and efficient working – Processes, Tools, Resources. 
  • Work with the engineering and cross-functional leaders along with the developers/testers to achieve continuous and incremental quality improvement across all products. 
  • Own and deliver automation, security, performance testing frameworks/tools/strategies.  
  • Enable the adoption of such tools/frameworks during development phase. 
  • Build tools to monitor and generate metrics of all QA activities - test definition, execution, results, production incidents etc. 
  • Build a strategy and deliver capabilities to fully automate computer software validation for all product releases. 
  • Build and sustain high performance teams – Hiring, Performance and Talent Management 
Read more
Why apply to jobs via Cutshort
people_solving_puzzle
Personalized job matches
Stop wasting time. Get matched with jobs that meet your skills, aspirations and preferences.
people_verifying_people
Verified hiring teams
See actual hiring teams, find common social connections or connect with them directly.
ai_chip
Move faster with AI
We use AI to get you faster responses, recommendations and unmatched user experience.
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos