Cutshort logo
For Employers
KnackLabs logo
AI Transformation Engineer
AI Transformation Engineer

AI Transformation Engineer at KnackLabs · Hyderabad · 7 - 10 years · ₹30L - ₹45L / yr · Bootstrapped · Posted 2 Sep 2026

KnackLabs's logo

AI Transformation Engineer

Stuti Jain's profile picture
Posted by Stuti Jain
7 - 10 yrs
₹30L - ₹45L / yr
Hyderabad
Skills
skill iconPython
AI Agents
Large Language Models (LLM)
Prompt engineering
Retrieval Augmented Generation (RAG)
Vector database
LangChain
RESTful APIs
Fullstack Developer
API
skill iconAmazon Web Services (AWS)
Google Cloud Platform (GCP)

Location: Hyderabad, India (home base), deployed at client sites in India. Occasional Middle East exposure possible.

About the Role

You will work as a senior AI consultant who embeds inside a customer's business. Your job is to learn how the business makes money, find the highest value problem, and build a working system that solves it.


You will not hand over a document and walk away. You will show working software early, own the roadmap, own the client relationship, and stay after go-live to run and improve the system.


Four behaviors define this role:

  1. Go where the work happens. You work onsite with the customer, in the room where decisions are made.
  2. Show working software early. You build a prototype in days, not a document in weeks.
  3. One person owns the outcome. You are the single point of accountability for the result.
  4. Stay after go-live. You keep running and improving the system after launch.


You are the single point of accountability. You are not a solo builder. A full KnackLabs engineering team in Hyderabad builds and runs the production systems behind you.


This role involves extended onsite deployments at client locations in other cities, sometimes up to six months at a stretch. Please apply only if you are ready for this way of working.

What you'll own

  1. Discovery - Learn how the customer makes money. Find the highest value problem to solve first.
  2. The prototype - Build a working prototype fast, using real or sample data, to prove the idea.
  3. The roadmap - Decide what to build, in what order, and set clear success measures tied to business outcomes.
  4. The build - Design and ship the production system with the Hyderabad engineering team. This includes data integration, agents, retrieval, and evaluations.
  5. The client relationship - Be the trusted technical contact for the customer, from engineers to senior leaders.
  6. Go live and after - Deploy the system, watch how it performs, fix problems, and improve it over time.
  7. Feedback to the product - Share what you learn in the field so our platform and internal tools get better.

What we are looking for

  1. Around 7 or more years of software engineering experience, including customer-facing or client delivery work.
  2. Experience working at a consulting or professional services firm in a client-facing delivery role.
  3. A full-stack development experience with strength in backend technologies.
  4. Strong programming skills in Python. Working knowledge of TypeScript or JavaScript.
  5. Production experience with large language models, including prompt engineering and agent development.
  6. You build with AI coding tools like Claude Code as your default way of working. You have built real apps and agents this way, not just used it for document generation or review.
  7. Experience building retrieval-augmented generation (RAG) systems: chunking, embeddings, vector databases, retrieval, and reranking.
  8. Experience building and deploying AI systems.
  9. Experience integrating with APIs and enterprise systems.
  10. Experience with at least one cloud platform (AWS, Azure, or GCP).
  11. Experience building evaluations to measure accuracy, safety, latency, and cost.
  12. Clear communication. You can explain a technical choice to an engineer and to a business leader.
  13. High ownership and comfort with ambiguity. You can take an unclear problem and turn it into a plan.
  14. Willingness to work onsite at client locations in India for extended periods, and to travel as the work needs.

Nice to have

  1. Experience deploying AI systems in regulated industries such as insurance, banking, or the public sector.
  2. Experience with on-premises or private cloud (VPC) deployments.
  3. Experience with observability and tracing tools such as LangSmith or Braintrust.
  4. Experience with data engineering and pipelines.
  5. A history of side projects, open source contributions, or products you shipped end-to-end.
  6. Experience in embedded or forward-deployed roles before.

Stack and tools

  1. Languages: Python and TypeScript.
  2. Models: Claude and other frontier or open source models, chosen to fit the customer.
  3. AI patterns: RAG, agents, prompt engineering, and evaluations.
  4. Vector and retrieval: vector databases and retrieval pipelines.
  5. Cloud: AWS, Azure, or GCP, on public or private cloud.
  6. Integration: REST APIs and enterprise system connectors.


Read more
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos

About KnackLabs

Founded :
2017
Type :
Services
Size :
100-1000
Stage :
Bootstrapped

About

Knacklabs.ai is Product Development Studio. WE BUILD TRUE PRODUCT TEAMS for our clients. Each team is a small, well-balanced group of geeks and a product manager that together produce relevant and high-quality products. We use data to make decisions, bringing big data and analysis to software development. We believe the product development process is broken as most studios operate as IT Services. We operate like a software factory that applies manufacturing principles of product development to the software.

Read more

Connect with the team

Profile picture
Ranjana Singh
Profile picture
Kartik Bansal
Profile picture
RaviKiran V
Profile picture
Krishnaveni B
Profile picture
Stuti Jain

Company social profiles

linkedinfacebook

Similar jobs (10)

company logo
Archita Srivastava
Posted by Archita Srivastava
Hyderabad
4 - 8 yrs
₹15L - ₹25L / yr
skill iconPython
TypeScript
skill iconJavascript
Large Language Models (LLM)
Agentic AI
+1 more

Location: Hyderabad, India (home base), deployed at client sites in India. Occasional Middle East exposure possible.

About the Role

You will work as a senior AI engineer who embeds inside a customer's business. Your job is to learn how the business makes money, find the highest value problem, and build a working system that solves it.


Four behaviors define this role:

  1. Go where the work happens. You work onsite with the customer, in the room where decisions are made.
  2. Show working software early. You build a prototype in days, not a document in weeks.
  3. One person owns the outcome. You are the single point of accountability for the result.
  4. Stay after go-live. You keep running and improving the system after launch.


You are the single point of accountability. You are not a solo builder. A full KnackLabs engineering team in Hyderabad builds and runs the production systems behind you.


This role involves extended onsite deployments at client locations in other cities, sometimes up to six months at a stretch. Please apply only if you are ready for this way of working.

What you'll own

  1. Discovery - Learn how the customer makes money. Find the highest value problem to solve first.
  2. The prototype - Build a working prototype fast, using real or sample data, to prove the idea.
  3. The roadmap - Decide what to build, in what order, and set clear success measures tied to business outcomes.
  4. The build - Design and ship the production system with the Hyderabad engineering team. This includes data integration, agents, retrieval, and evaluations.
  5. The client relationship - Be the trusted technical contact for the customer, from engineers to senior leaders.
  6. Go live and after - Deploy the system, watch how it performs, fix problems, and improve it over time.
  7. Feedback to the product - Share what you learn in the field so the vendor's product and our internal tools get better.


What we are looking for

  1. Around 4 or more years of software engineering experience, including customer-facing or client delivery work.
  2. Strong programming skills in Python. Working knowledge of TypeScript or JavaScript.
  3. A full-stack development experience with strength in backend technologies.
  4. Production experience with large language models, including prompt engineering and agent development.
  5. You build with AI coding tools like Claude Code or Codex as your default way of working, and you have shipped real apps or agents this way.
  6. Experience building retrieval-augmented generation (RAG) systems: chunking, embeddings, vector databases, retrieval, and reranking.
  7. Experience building and deploying AI systems.
  8. Experience integrating with APIs and enterprise systems.
  9. Experience with at least one cloud platform (AWS, Azure, or GCP).
  10. Clear communication. You can explain a technical choice to an engineer and to a business leader.
  11. High ownership and comfort with ambiguity. You can take an unclear problem and turn it into a plan.
  12. Willingness to work onsite at client locations in India for extended periods, and to travel as the work needs.

Nice to have

  1. Experience with on-premises or private cloud (VPC) deployments.
  2. Experience with observability and tracing tools such as LangSmith or Braintrust.
  3. Experience with data engineering and pipelines.
  4. A history of side projects, open source contributions, or products you shipped end-to-end.
  5. Experience in embedded or forward-deployed roles before.
  6. Experience working at a consulting or professional services firm in a client-facing delivery role.

Stack and tools

  1. Languages: Python and TypeScript.
  2. Models: Claude and other frontier or open-source models, chosen to fit the customer.
  3. AI patterns: RAG, agents, prompt engineering, and evaluations.
  4. Vector and retrieval: vector databases and retrieval pipelines.
  5. Cloud: AWS, Azure, or GCP, on public or private cloud.
  6. Integration: REST APIs and enterprise system connectors.


Read more
company logo
Meenal Patil
Posted by Meenal Patil
Pune
3 - 4 yrs
₹5L - ₹15L / yr
Agent development
legacy migration
AI Copilot
Claude AI APP

Role: AI Developer

Experience: 3–4 Years

Employment Type: Full-Time

Location: Goregaon, Mumbai


About the Role

We are looking for an experienced AI Developer with 3–4 years of software development experience and strong hands-on exposure to Generative AI, AI Agents, Copilots, and AI-powered application development.

The candidate will be responsible for building production-ready AI solutions, developing agentic workflows, modernizing legacy applications, and integrating LLM capabilities into enterprise applications.


Key Responsibilities

  • Design, develop, and deploy AI Agents and agentic workflows for enterprise use cases.
  • Build AI Copilots and LLM-powered applications using modern AI frameworks and APIs.
  • Develop RAG-based applications using embeddings, vector databases, and enterprise data.
  • Work on legacy application migration and modernization, leveraging AI-assisted development and code transformation techniques.
  • Analyze legacy codebases and design strategies for AI-driven migration, refactoring, and modernization.
  • Integrate LLMs with enterprise applications, APIs, databases, and third-party systems.
  • Implement tool calling, function calling, multi-agent workflows, and workflow automation.
  • Perform prompt engineering, context optimization, model evaluation, and AI application testing.
  • Take ownership of AI solutions from POC and prototyping through production deployment.
  • Collaborate with product managers, architects, and engineering teams to convert business requirements into scalable AI solutions.
  • Stay updated with emerging technologies in Generative AI, Agentic AI, LLMs, and AI-assisted software development.


Required Skills

  • 3–4 years of professional software development experience.
  • Strong proficiency in Python and/or JavaScript/TypeScript.
  • Hands-on experience developing Generative AI / LLM-based applications.
  • Strong understanding of AI Agents, RAG, Prompt Engineering, LLM APIs, and embeddings.
  • Experience with frameworks such as LangChain, LangGraph, Semantic Kernel, AutoGen, or equivalent.
  • Experience working with REST APIs, databases, Git, and cloud environments.
  • Hands-on experience with vector databases such as Pinecone, Weaviate, Chroma, FAISS, or equivalent.
  • Good understanding of software architecture, debugging, testing, and deployment practices.


Good to Have

  • Experience with Microsoft Copilot / Copilot Studio.
  • Experience working with Claude, OpenAI, Gemini, Azure OpenAI, or open-source LLMs.
  • Experience in legacy application migration, modernization, or code conversion.
  • Knowledge of Azure AI / AWS / Google Cloud AI services.
  • Experience with MCP, multi-agent systems, tool calling, and AI orchestration.
  • Experience building enterprise-grade AI solutions with focus on security, scalability, and performance.


Read more
company logo
Stuti Jain
Posted by Stuti Jain
Hyderabad
7 - 10 yrs
₹25L - ₹35L / yr
Retrieval Augmented Generation (RAG)
skill iconAmazon Web Services (AWS)

Location: Hyderabad, India. Based at the KnackLabs headquarters, with occasional travel to client locations for workshops and reviews. This role does not involve extended onsite deployments.

About the Role

You will work as an AI Architect who designs the systems behind our client engagements: AI agents, RAG systems, automation platforms, and the conventional backend systems around them.

This is a hands-on design role, not a slideware role. You will scope architectures with clients, make the hard technical decisions, defend them in review, and stay accountable for how the systems perform in production.


You will work directly with clients. Everyone at KnackLabs does. You will sit in design discussions with client engineering teams, present architecture decisions to technical and business stakeholders, and answer for the choices you make.


A full KnackLabs engineering team in Hyderabad builds with you. You own the technical design and the quality of what ships.

What you'll own

  1. Architecture - Design AI agents, RAG systems, integrations, and the scalable backend systems around them, for multiple client engagements.
  2. Technical scoping - Work directly with clients to turn a business problem into a system design, with clear trade-offs and clear reasons.
  3. Scale and reliability - Make sure what we build handles real load: data stores, queues, caching, horizontal scaling, and fault tolerance.
  4. Design reviews - Review designs and builds across engagements. Set the technical bar and hold it.
  5. Evaluation strategy - Define how we measure accuracy, safety, latency, and cost for the AI systems we ship.
  6. Guiding engineers - Raise the level of the engineers building with you, through reviews and direct pairing.
  7. Feedback to the platform - Feed what you learn across engagements back into our platform and internal tools.

What we are looking for

  1. Around 7 or more years of software engineering experience, including direct work with customers on design or delivery.
  2. Full-stack development experience with strength in backend technologies.
  3. Experience designing and building scalable applications. You understand how large-scale distributed systems work: data partitioning, queues, caching, horizontal scaling, and fault tolerance.
  4. At least 2 years of strong, hands-on AI experience with large language models in production.
  5. You build with AI coding tools like Claude Code or Codex as your default way of working. You understand Claude Skills, have written skills yourself, use them actively, and have contributed to them.
  6. Hands-on experience building retrieval-augmented generation (RAG) systems: chunking, embeddings, vector databases, retrieval, and reranking.
  7. Hands-on experience building AI agents.
  8. Strong programming skills in Python. Working knowledge of TypeScript or JavaScript.
  9. Experience with at least one cloud platform (AWS, Azure, or GCP).
  10. Clear communication. You can explain an architecture decision to an engineer and to a business leader, and defend it under questioning.
  11. High ownership and comfort with ambiguity. You can take an unclear problem and turn it into a design.

Nice to have

  1. Experience building evaluations to measure accuracy, safety, latency, and cost.
  2. Experience with observability and tracing tools such as LangSmith or Braintrust.
  3. Experience with on-premises or private cloud (VPC) deployments.
  4. Experience deploying AI systems in regulated industries such as insurance, banking, or the public sector.
  5. Experience with data engineering and pipelines.
  6. A history of side projects, open source contributions, or products you shipped end-to-end.
  7. Experience working at a consulting or professional services firm in a client-facing delivery role.

Stack and tools

  1. Languages: Python and TypeScript.
  2. Models: Claude and other frontier or open-source models, chosen to fit the customer.
  3. AI patterns: RAG, agents, prompt engineering, skills, and evaluations.
  4. Vector and retrieval: vector databases and retrieval pipelines.
  5. Cloud: AWS, Azure, or GCP, on public or private cloud.
  6. Integration: REST APIs and enterprise system connectors.


Read more
Service Co
Service Co
Agency job
via by Rishika Teja
Pune
4 - 8 yrs
₹14L - ₹18L / yr
skill iconPython
TypeScript
Large Language Models (LLM)
Retrieval Augmented Generation (RAG)
AI Frameworks
+3 more

Skill Set

Large language,Artificial Intelligence,Machine Learning


- 4–7 years of experience in software engineering/AI roles

- Strong programming skills in Python or TypeScript (Java/Go is a plus)

- Hands-on experience with LLMs, RAG pipelines, and AI frameworks

- Experience building APIs and working with distributed systems

- Familiarity with Kubernetes, Docker, and CI/CD pipelines

- Experience with cloud platforms (AWS/Azure/GCP)

Excellent communication

Read more
New York, Los Angeles California
3 - 5 yrs
$2.5K - $5.5K / yr
skill iconPython
Artificial Intelligence (AI)
skill iconMachine Learning (ML)
Multi-Agent System
Full Stack Development
+17 more

We are building an advanced, AI-driven multi-agent software system designed to revolutionize task automation and code generation. This is a futuristic AI platform capable of:


✅ Real-time self-coding based on tasks  

✅ Autonomous multi-agent collaboration  

✅ AI-powered decision-making  

✅ Cross-platform compatibility (Desktop, Web, Mobile)  


We are hiring a highly skilled **AI Engineer & Full-Stack Developer** based in India, with a strong background in AI/ML, multi-agent architecture, and scalable, production-grade software development.


### Responsibilities:


- Build and maintain a multi-agent AI system (AutoGPT, BabyAGI, MetaGPT concepts)  

- Integrate large language models (GPT-4o, Claude, open-source LLMs)  

- Develop full-stack components (Backend: Python, FastAPI/Flask, Frontend: React/Next.js)  

- Work on real-time task execution pipelines  

- Build cross-platform apps using Electron or Flutter  

- Implement Redis, Vector databases, scalable APIs  

- Guide the architecture of autonomous, self-coding AI systems  


### Must-Have Skills:


- Python (advanced, AI applications)  

- AI/ML experience, including multi-agent orchestration  

- LLM integration knowledge  

- Full-stack development: React or Next.js  

- Redis, Vector Databases (e.g., Pinecone, FAISS)  

- Real-time applications (websockets, event-driven)  

- Cloud deployment (AWS, GCP)  


### Good to Have:


- Experience with code-generation AI models (Codex, GPT-4o coding abilities)  

- Microservices and secure system design  

- Knowledge of AI for workflow automation and productivity tools  


Join us to work on cutting-edge AI technology that builds the future of autonomous software.

Read more
company logo
Xclusive Interiors
Posted by Xclusive Interiors
Pune, pimple saudagar
0 - 3 yrs
₹2L - ₹3.6L / yr
skill iconHTML/CSS
SQL
API
skill iconPython
skill iconJavascript
+5 more

* Strong practical knowledge and interest in modern AI tools.

* Genuine curiosity and willingness to continuously learn.

* Ability to research, experiment, implement, troubleshoot and improve independently.

* Good understanding of prompting and AI workflows.

* Strong problem-solving mindset.

* Basic understanding of APIs, integrations and automation.

* Ability to explain technology clearly to non-technical people.

* Comfortable using AI to solve real-world business problems.

Read more
company logo
Hema V
Posted by Hema V
Remote, Gurugram, Noida, Bengaluru (Bangalore), Chennai
3 - 8 yrs
₹20L - ₹50L / yr
Multi-agent Systems
CrewAI
LangChain
Retrieval Augmented Generation (RAG)
LoRA / QLoRA
+2 more

KEY RESPONSIBILITIES:

•Build agents with persistent context & memory

•Design self-learning feedback loops

•Implement RAG pipelines for domain knowledge

•Manage conversation state & orchestration

•Integrate with LLM APIs (OpenAI, Claude, open-source)

Iterate fast — ship daily, measure weekly


MUST-HAVE SKILLS

•Python / TypeScript proficiency

•LangChain, CrewAI, AutoGen or custom frameworks

•Experience with vector DBs (Pinecone, Weaviate, Qdrant)

•Prompt engineering & evaluation pipelines

•Understanding of agent architectures (ReAct, tool-use)

Git, CI/CD, containerization basics

Read more
company logo
Aswathy Vimal
Posted by Aswathy Vimal
Remote only
6 - 10 yrs
₹24L - ₹36L / yr
skill iconPython
skill iconJavascript
TypeScript
Large Language Models (LLM) tuning

The Role

We’re looking for a Senior Applied AI & Data Engineer to become our first dedicated AI and data engineer.

You’ll build conversational AI experiences across web, mobile, and in-store channels while developing the data foundation behind them. You’ll make key technical decisions and own your work through to production.

What You’ll Do

• Build AI assistants using tool calling to work with real product, search, and order systems

• Design guardrails and evaluation sets to ensure AI responses are accurate and safe

• Build real-time and voice-enabled AI experiences

• Improve product data quality through AI-assisted enrichment and review workflows

• Build data pipelines, analytics, and personalisation systems

• Work closely with web and mobile developers and help guide technical implementation

What You’ll Need

• 6+ years of experience building and running production backend systems

• Strong Python skills, plus experience with JavaScript/TypeScript backends

• Experience shipping at least one LLM-powered feature to real users

• Experience with search and relevance

• Experience building data pipelines and analytics stores

• Comfortable deploying and monitoring services on a major cloud platform

• Strong communication skills and the ability to work independently

Nice to Have

• Experience with speech or voice AI

• E-commerce or retail technology experience

• Experience building multilingual products

Read more
company logo
Madhavan I
Posted by Madhavan I
icon

The recruiter has not been active on this job recently. You may apply but please expect a delayed response.

Bengaluru (Bangalore), Noida
5 - 10 yrs
₹20L - ₹40L / yr
Fullstack Developer
Adobe Creative Cloud
AI Coding Tools

The Opportunity

Join our “DevOps for Content” revolution as we partner with global brands and agencies

to transform their end-to-end creative workflows – from ideation to activation – to deliver

AI-powered content services with speed, scale, and governance.

Through an AI-first experimentation approach and deep expertise in both first-party and

third-party AI models, we develop new applications, platforms, and scalable patterns

that unlock the GenAI-powered Content Supply Chain. We function as a continuous

product innovation engine—feeding repeatable customer solutions straight into the

product roadmap to accelerate product-led growth.

As a Forward-Deployed AI Engineer, you will own the end-to-end development and

launch of GenAI applications through Adobe and third-party generative models. You are

an AI native who will embed directly into customer teams to spin up proof points in days

and co-develop features that enhance customer value. If you thrive in start-up

environments—scaling products from prototype to production, pushing innovative

research into real-world use cases, and owning product & business outcomes—you’ll fit

right in here!

What You’ll Do

 Collaborate & Innovate: Partner with Technical Architects, Engagement

Managers and Product teams to define customer requirements/use cases,

facilitate technical workshops, and co-create specific GenAI solutions.

 Prototype Rapidly: Apply an AI-first experimentation mindset—build proof

points in days, iterate on feedback, and drive quick wins. UX/UI and rapid

prototyping skills are a plus.

 Engineer End-to-End: Design, build, and deploy full-stack applications and

microservices—integrating Firefly APIs, Common Extensibility Platform,

GenStudio for Performance Marketing, headless CMS architectures and UX/UI

Designs.

 Bridge to Product: Capture field-proven use cases and funnel them back to

Product & Engineering—fueling the Firefly roadmap.

 Automate & Scale: Develop reusable components, CI/CD pipelines, and

governance checks to ensure consistent, repeatable delivery.

 Operate at Speed: Thrive in a constantly evolving, start-up-like

environment—own delivery sprints, tackle challenges on-site/remotely, and adapt

to evolving AI trends.

 Share & Elevate: Document your playbooks, prompts, and patterns in our

internal knowledge base—helping the whole organization move faster.

What You Bring


 Full-Stack Expertise: 5+ years as a full-stack developer or technical consultant

building and launching production software (React.js/Next.js/Angular;

Node.js/Java Spring Boot; REST/GraphQL, HTML, JavaScript, Java, XML).

 AI Tinkerer: AI-native tinkerer who embeds AI in your day-to-day workflows and

spends your free time inventing and tinkering with the possibilities of AI. You

keep ahead of market trends across third-party models, GPUs, and new AI

applications.

 Adobe Fluency: Deep understanding of Adobe Creative Cloud APIs/SDK/CEP,

Firefly Services APIs, and Adobe Experience Cloud integrations (AEM, Workfront

Fusion, AJO or equivalent experience).

 Cloud & DevOps Savvy: AWS/Azure/GCP, containerization

(Docker/Kubernetes), and CI/CD tooling (GitLab, Jenkins, etc.).

 Strong communicator: You translate complex AI concepts for both engineers

and executives and guide customers through technical decisions aligned to

business objectives.

 Customer-Centricity: Always prioritizing the customer and their needs and

expectations first in everything you do. You are outstanding at contributing to a

culture of customer-centricity among your team and your extended ecosystem.

 Startup DNA: Thrives in fastpaced, ambiguous environments—relentlessly

seeking better, faster, more innovative ways to solve customer objectives and

challenges

Read more
company logo
Umama Sayed
Posted by Umama Sayed
Remote, Mumbai
2 - 4 yrs
Best in industry
skill iconPython
Large Language Models (LLM)
Generative AI
LangGraph
FastAPI
+7 more

AI Engineer

LLMs, Agents & AI Services

📍 Mumbai (On-site) | Full-time | 2-4 years


About the Role:

Unico Connect is an AI-first technology partner that builds custom mobile, web, and AI products for clients across multiple geographies.

AI is core to how we design, deliver, and scale software for our customers.

We are hiring an AI Engineer for a dedicated client engagement building a complex production AI platform, working on the AI capabilities and agentic features at the core of the product.

The mandatory requirement for this role is at least one AI feature personally shipped to production for real users, with operational ownership.

The role suits someone who thinks quickly on solutioning, can take an ambiguous problem to a working prototype in days, and has the discipline to carry it through to production with predictable economics.

You will work alongside the Senior AI Engineer and the wider pod, with ownership of parts of the AI surface area of the product.


Responsibilities:

Solutioning and POCs

Translate ambiguous customer problems into working POCs at speed.

Pick the right model, framework, and architecture, and demonstrate value early before scaling investment.


LLM Application Development

Build AI features and services using LLM APIs from OpenAI, Anthropic, Google, and self-hosted open-weight models (Llama, Qwen, Mistral).

Choose the right model per use case based on cost, latency, capability, and context-window trade-offs.


Agentic System Design

Design and implement agentic workflows using LangGraph, CrewAI, AutoGen, LlamaIndex Agents, or custom orchestration.

Cover tool use, planning, memory, and multi-step reasoning appropriate to the problem.


API and Service Development

Build production AI services and APIs using Python and FastAPI.

Handle streaming responses, async processing, structured outputs, retries, and graceful degradation when models or tools fail.


Retrieval and Tool Integration

Implement RAG pipelines with vector databases (Pinecone, Weaviate, Qdrant, pgvector, Chroma), embeddings, chunking strategies, hybrid search, and reranking.

Integrate external tools, internal APIs, and document sources through tool-calling and MCP-style patterns.


Cost Analysis and Unit Economics

Model the per-request and per-user cost of every AI feature before it ships.

Track token usage, prompt caching, batching, and model-routing strategies.

Drive measurable improvements in unit economics.


Production Hardening

Add observability and tracing (LangSmith, Langfuse, OpenTelemetry), guardrails, content safety checks, prompt injection defences, and fallback behaviour.


Prompt Engineering and Evaluation

Design, test, and iterate prompts with measured outcomes.

Build evaluation harnesses for accuracy, hallucination, latency, and cost.

Run benchmarks across models and prompt variants before locking in a design.


Requirements:

AI Feature Shipped to Production (Mandatory)

Must have personally built and shipped at least one AI feature that runs in production for real users, with operational ownership.

POCs, internal demos, and one-off scripts do not qualify.


2 to 4 Years of Professional Software or AI Engineering Experience

With at least one production AI feature owned end to end.


Strong Python Proficiency and API Development with FastAPI

Comfort with type hints, async, packaging, testing, streaming responses, and authentication.

Production-grade Python, not notebook-only code.


Hands-on Depth Across the LLM and Agent Stack

Working experience with at least two of OpenAI, Anthropic Claude, Google Gemini, or self-hosted open-weight models (vLLM, Ollama, Together, Replicate).

Working familiarity with at least one agent framework (LangGraph, CrewAI, AutoGen, LlamaIndex Agents) or hand-rolled equivalent.

Working knowledge of RAG, embeddings, and vector databases (Pinecone, Weaviate, Qdrant, pgvector, Chroma).


Solutioning Speed and POC Velocity

Demonstrated ability to move from a fuzzy problem to a working prototype in days.

Strong instinct for what to build first, what to defer, and what to throw away.


Cost Discipline for Production AI

Ability to calculate, monitor, and optimise the cost of LLM APIs, tokens, embeddings, vector store usage, and infrastructure.

Treats unit economics as a first-class concern.


AWS Familiarity

Working knowledge of EC2, S3, IAM, and at least one of Bedrock, SageMaker, or equivalent.


Comfortable in a Fast-Moving Environment

Self-directed, comfortable with ambiguity, takes ownership without being asked, and ships under shifting priorities.


Strong Written and Spoken English Communication

Able to explain trade-offs to non-AI engineers, designers, product managers, and clients in plain language.


Nice to Have

  • fine-tuning or LoRA, QLoRA, PEFT exposure
  • MCP server authoring
  • eval framework experience (LangSmith, Promptfoo, Ragas, DeepEval)
  • open-source AI contributions
  • multi-modal models (vision, audio)
Read more
Why apply to jobs via Cutshort
people_solving_puzzle
Personalized job matches
Stop wasting time. Get matched with jobs that meet your skills, aspirations and preferences.
people_verifying_people
Verified hiring teams
See actual hiring teams, find common social connections or connect with them directly.
ai_chip
Move faster with AI
We use AI to get you faster responses, recommendations and unmatched user experience.
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos