Cutshort logo
For Employers

AI Tech Lead at Appiness Interactive Pvt. Ltd. · Pune · 7 - 12 years · ₹20L - ₹28L / yr · Bootstrapped · Posted 22 Dec 2025

7 - 12 yrs
₹20L - ₹28L / yr
Pune
Skills
Artificial Intelligence (AI)
Large Language Models (LLM)
Retrieval Augmented Generation (RAG)

Position Overview:

The AI Tech Lead will architect and guide the implementation of real-time AI systems across voice automation, LLM pipelines, and knowledge-enhanced applications. The role requires strong architectural judgment, hands-on expertise in AI/LLM systems, and the ability to define high-performance, scalable best practices.


Key Responsibilities:

• Architect and guide implementation of real-time AI systems across voice automation, LLM pipelines, and knowledge-enhanced applications

• Design and develop distributed, provider-agnostic AI architectures with performance guarantees including low latency, resilient failover, distributed scaling, and cost-efficiency

• Define architectural best practices for GenAI systems including prompt design, context shaping, fallback logic, caching, and real-time agent orchestration

• Lead AI model governance including evaluation and selection frameworks for multiple LLM providers, routing logic, benchmarking, and cost management

• Establish and monitor KPIs for LLM quality, latency, reliability, grounding accuracy, and system stability

• Ensure data security practices for handling of voice and transcript data, applying PII-safe methods, and supporting multi-tenant data isolation

• Own knowledge integration and RAG architecture including vector databases, retrieval strategies, chunking policies, and hybrid grounding methods

• Continuously evaluate new model capabilities and AI technology trends


Required Skills:

• Architectural judgment and hands-on expertise in AI and LLM systems • Experience designing scalable and low-latency AI architectures

• Knowledge of multi-provider LLM integration and orchestration

• Understanding of distributed systems, microservices, and load balancing

• Strong grounding in AI model governance and benchmarking

• Awareness of data security and privacy best practices

• Experience with retrieval-augmented generation (RAG) and vector databases


Preferred (Bonus) Skills:

• Experience with function-calling and knowledge graphs

• Familiarity with hybrid grounding and retrieval enhancement strategies • Experience with voice automation systems (STT/TTS pipelines). give me in a proper formate

Read more
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos

About Appiness Interactive Pvt. Ltd.

Founded :
2011
Type :
Services
Size :
100-1000
Stage :
Bootstrapped

About

We are a Bangalore based Product Development and UX firm specializing in Digital Services for the whole spectrum, from startups to fortune-500’s. We do not redefine anything or reinvent the wheel.We work closely with startups to create their product and help it get of the ground, and continue support in scaling up strategies and technologies. For our blue-chip clients we bring in to the table our vast experience in bleeding edge technology and pleasurable UX.We create a comprehensive soul for a brand in the online world. We create digital products that infiltrates the public realm and engages the targeted customers of our clients all over the globe through a multiple platforms of Digital Media, including web, mobile and wearables.We have a young, passionate and aggressive team who are not afraid to think out of the box or tread the un-trodden path in order to deliver the best result for our clients. We call it Practical Creativity, where the belief is that the idea is only as good as the returns it fetches for our clients. We are not here to create immortal art; we are here to generate revenue for our clients. We believe in a dynamic and meaningful conversation between our clients and their customers for which we facilitate a medium.
Read more

Connect with the team

Profile picture
HR Appiness

Company social profiles

blogpinterestlinkedintwitterfacebook

Similar jobs (10)

The industry’s only Manufacturing Operating System
The industry’s only Manufacturing Operating System
Agency job
via by Ariba Khan
Hyderabad
10 - 15 yrs
Best in industry
Artificial Intelligence (AI)
Generative AI (GenAI)
Retrieval Augmented Generation (RAG)
Large Language Models (LLM)

We’re on hunt for AI Architect


Responsibilities:

  • 10–15+ years overall experience, with recent hands-on AI/GenAI architecture ownership.
  • Must have architected enterprise AI platforms/solutions end-to-end, not just individual ML models or PoCs.
  • Strong GenAI/LLM production experience: RAG, embeddings, vector DBs, hybrid search, reranking, evaluation, guardrails.
  • Strong Agentic AI understanding: agents, tool calling, workflows, orchestration, human-in-the-loop.
  • Experience taking AI solutions from architecture → production → scale, ideally across multiple business teams/use cases.
  • Strong cloud architecture — Azure/AWS preferred; hybrid/on-prem experience is a plus.
  • Must understand enterprise security, governance, Responsible AI, observability and LLMOps/MLOps.
  • Should be able to articulate build-vs-buy, MVP-vs-target architecture, cost/performance/security tradeoffs.
  • Strong stakeholder-facing / consulting ability — can work with business leaders, engineering, security and data teams and influence without authority.


There is scope to move to the US for this role if you are aligned for the same, else this will be a WFO role from Hyderabad location

Read more
company logo
Sandeep C
Posted by Sandeep C
Bengaluru (Bangalore)
8 - 16 yrs
₹1L - ₹2L / yr (ESOP available)
Large Language Models (LLM)
Agentic AI
Applied mathematics

Key Responsibilities:

·      Architectural Leadership: Design and lead the development of robust, scalable AI architectures, ensuring high performance, reliability, and security.

·      Applied Mathematics & Statistics: Apply statistical analysis, numerical computation, and mathematical modeling to derive insights from large-scale data and optimize model performance.

·      Deep Learning Development: Design, train, and deploy advanced Deep Learning (DL) models.

·      Technical Mentorship: Mentor engineering teams on best practices for AI/ML, coding standards, and architectural design.

·      Model Optimization: Optimize models for speed, efficiency, and accuracy using techniques like pruning, quantization, or GPU acceleration.

·      Strategy & Innovation: Evaluate and select appropriate AI frameworks, tools, and platforms, staying abreast of cutting-edge research and industry trends.

Qualifications:

Required:

·      Education: Master's or PhD in Computer Science, Applied Mathematics, Statistics, Physics, or a related quantitative field.

·      Experience: 10+ years of experience in software development, with at least 3-5 years in a Applied Mathematics and Deep learning.

·      AI/ML Expertise: Proven experience designing and deploying deep learning models in production using frameworks.

·      Mathematics/Statistics: Strong proficiency in linear algebra, calculus, probability, and statistical methods.

·      Programming Skills: Expert-level coding skills in Python (NumPy, Pandas, Scikit-learn) and experience with languages like Java or C++.

Key Competencies:

  • Strategic mindset with deep operational awareness.
  • Excellent communication and stakeholder management skills.
  • Ability to simplify complex technical concepts for executive reporting.
  • Strong leadership, people development, and cross-functional influencing skills.

Bias for action and a relentless focus on continuous improvement.

Read more
company logo
Stuti Jain
Posted by Stuti Jain
Hyderabad
7 - 10 yrs
₹25L - ₹35L / yr
Retrieval Augmented Generation (RAG)
skill iconAmazon Web Services (AWS)

Location: Hyderabad, India. Based at the KnackLabs headquarters, with occasional travel to client locations for workshops and reviews. This role does not involve extended onsite deployments.

About the Role

You will work as an AI Architect who designs the systems behind our client engagements: AI agents, RAG systems, automation platforms, and the conventional backend systems around them.

This is a hands-on design role, not a slideware role. You will scope architectures with clients, make the hard technical decisions, defend them in review, and stay accountable for how the systems perform in production.


You will work directly with clients. Everyone at KnackLabs does. You will sit in design discussions with client engineering teams, present architecture decisions to technical and business stakeholders, and answer for the choices you make.


A full KnackLabs engineering team in Hyderabad builds with you. You own the technical design and the quality of what ships.

What you'll own

  1. Architecture - Design AI agents, RAG systems, integrations, and the scalable backend systems around them, for multiple client engagements.
  2. Technical scoping - Work directly with clients to turn a business problem into a system design, with clear trade-offs and clear reasons.
  3. Scale and reliability - Make sure what we build handles real load: data stores, queues, caching, horizontal scaling, and fault tolerance.
  4. Design reviews - Review designs and builds across engagements. Set the technical bar and hold it.
  5. Evaluation strategy - Define how we measure accuracy, safety, latency, and cost for the AI systems we ship.
  6. Guiding engineers - Raise the level of the engineers building with you, through reviews and direct pairing.
  7. Feedback to the platform - Feed what you learn across engagements back into our platform and internal tools.

What we are looking for

  1. Around 7 or more years of software engineering experience, including direct work with customers on design or delivery.
  2. Full-stack development experience with strength in backend technologies.
  3. Experience designing and building scalable applications. You understand how large-scale distributed systems work: data partitioning, queues, caching, horizontal scaling, and fault tolerance.
  4. At least 2 years of strong, hands-on AI experience with large language models in production.
  5. You build with AI coding tools like Claude Code or Codex as your default way of working. You understand Claude Skills, have written skills yourself, use them actively, and have contributed to them.
  6. Hands-on experience building retrieval-augmented generation (RAG) systems: chunking, embeddings, vector databases, retrieval, and reranking.
  7. Hands-on experience building AI agents.
  8. Strong programming skills in Python. Working knowledge of TypeScript or JavaScript.
  9. Experience with at least one cloud platform (AWS, Azure, or GCP).
  10. Clear communication. You can explain an architecture decision to an engineer and to a business leader, and defend it under questioning.
  11. High ownership and comfort with ambiguity. You can take an unclear problem and turn it into a design.

Nice to have

  1. Experience building evaluations to measure accuracy, safety, latency, and cost.
  2. Experience with observability and tracing tools such as LangSmith or Braintrust.
  3. Experience with on-premises or private cloud (VPC) deployments.
  4. Experience deploying AI systems in regulated industries such as insurance, banking, or the public sector.
  5. Experience with data engineering and pipelines.
  6. A history of side projects, open source contributions, or products you shipped end-to-end.
  7. Experience working at a consulting or professional services firm in a client-facing delivery role.

Stack and tools

  1. Languages: Python and TypeScript.
  2. Models: Claude and other frontier or open-source models, chosen to fit the customer.
  3. AI patterns: RAG, agents, prompt engineering, skills, and evaluations.
  4. Vector and retrieval: vector databases and retrieval pipelines.
  5. Cloud: AWS, Azure, or GCP, on public or private cloud.
  6. Integration: REST APIs and enterprise system connectors.


Read more
company logo
Faisal AshrafNomani
Posted by Faisal AshrafNomani
Chennai, Bengaluru (Bangalore), Gurugram, Hyderabad
8 - 10 yrs
Best in industry
Artificial Intelligence (AI)
skill iconMachine Learning (ML)
Generative AI

About the Role We are seeking a highly technical, hands-on Senior AI/ML Tech Lead to drive the design, development, and deployment of cutting-edge Generative AI applications. In this dual-impact role, you wi l act as a primary individual contributor architecting core AI engines while simultaneously leading a team of engineers through task alocation, code reviews, and technical mentorship. The ideal candidate bridges the gap between state-of-the-art AI research (LLMs, Agentic frameworks, Advanced RAG, OCR) and production-grade ful-stack engineering (Python, FastAPI, React).


Key Responsibilities

Technical Leadership & Team Management (40%)

● Technical Oversight: Lead a team of AI, backend, and ful-stack engineers; alocate tasks, establish sprint priorities, and ensure timely delivery.

● Code Quality & Reviews: Conduct rigorous code reviews to maintain high engineering standards, security, performance, and scalability across AI and fu l-stack codebases.

● Architecture & Governance: Design end-to-end system architectures for AI solutions, ensuring seamless integration between frontend interfaces, backend APIs, and AI models.

● Mentorship: Guide and upskil team members on modern software practices, LLM engineering, and agentic design patterns. Hands-On Engineering & Development (60%)

● Generative AI & Agentic Systems: Architect, build, and optimize LLM-powered applications, multi-agent workflows (e.g., CrewAI, AutoGen, LangGraph), and autonomous AI agents.

● RAG & OCR Pipelines: Design and deploy advanced RAG (Retrieval-Augmented Generation) architectures and document processing pipelines utilizing OCR techniques (e.g., LayoutLM, PaddleOCR, Tesseract, Vision LLMs) to extract structured data from unstructured sources.

● Backend Systems: Build robust, asynchronous, high-throughput microservices and RESTful APIs using Python and FastAPI.

● Frontend Integration: Colaborate on or build modern web interfaces using React (e.g., Control Towers, operations dashboards, interactive chat interfaces).

● MLOps & Vector DBs: Oversee model deployment, prompt engineering, fine-tuning, vector database integration (Pinecone, Qdrant, Chroma, PGVector), and cloud infrastructure setup (Azure/AWS).


Required Qualifications & Skills

● Overall Experience: 8 to 10 years of professional software engineering experience.

● AI/ML Domain Experience: 3 to 4+ years of dedicated, hands-on experience building and deploying AI/ML, OCR, and Generative AI solutions in production.

● Core Technical Stack: ○ Generative AI & LLMs: Extensive experience with commercial and open-source LLMs (OpenAI, Anthropic Claude, Llama), Agentic frameworks (LangChain, LlamaIndex, AutoGen, CrewAI), and LLM evaluation frameworks (LangSmith, TruLens, Ragas). ○ RAG & Unstructured Data: Strong knowledge of hybrid search, re-ranking, chunking strategies, vector databases, and document inte ligence workflows. ○ OCR & Vision Techniques: Hands-on experience with OCR engines (Tesseract, PaddleOCR, Azure Document Inteligence) and Multi-Modal/Vision LLMs for document extraction. ○ Backend: Deep expertise in Python and asynchronous frameworks (FastAPI, AsyncIO). ○ Frontend: Working proficiency in React (TypeScript/JavaScript) for building interactive web UI components. ○ Cloud & DevOps: Hands-on experience with cloud platforms (Azure / AWS), Docker, Kubernetes, and CI/CD pipelines.


Preferred / Good-to-Have Skills


● Experience with cloud-native data platforms (e.g., Microsoft Fabric, Snowflake, Azure SQL).

● Familiarity with cost optimization and latency reduction techniques for LLM inference (caching, semantic routing, model quantization).

● Prior experience in client-facing technical leadership or agile consulting environments.


What We Offer


● Opportunity to lead and build high-impact, state-of-the-art Generative AI systems.

● Colaborative engineering culture with room for technical ownership and direct business impact.

● Flexible work arrangements and competitive compensation package.

Read more
Remote only
10 - 16 yrs
Best in industry
Large Language Models (LLM) tuning
Generative AI (GenAI)

Principal Enterprise GenAI / Agentic AI Architect - (Freelance) 

Positions: 1


Experience: Ideally 10–16 years overall, with significant recent hands-on GenAI/LLM architecture experience.


Mission

Own the end-to-end architecture across RAG, Agentic AI, enterprise integrations, model serving, security, evaluation, observability and production deployment.

This should not be a PowerPoint-only architect. We need someone technically deep enough to review code, challenge engineering decisions, troubleshoot RAG/agent behaviour and interact credibly with customer architecture/security/platform teams.

Mandatory capabilities

  • Enterprise GenAI architecture
  • Production RAG
  • Agentic AI architecture
  • Python
  • LangGraph or comparable stateful orchestration
  • Tool/function calling
  • Human-in-the-loop workflows
  • Vector databases
  • Embeddings/reranking
  • LLM/RAG/agent evaluation
  • REST APIs/microservices
  • Enterprise IAM
  • RBAC/ABAC
  • AI security and prompt-injection mitigation
  • Kubernetes
  • CI/CD and LLMOps/MLOps
  • Enterprise observability


Highly desirable

OpenShift/OpenShift AI, NVIDIA NIM, KServe, NVIDIA GPU Operator, open-weight LLM deployment, ServiceNow, Splunk, Microsoft Graph and previous banking/financial-services experience.



Read more
company logo
Agency job
via by aarushi Mahajan
Hyderabad, Bengaluru (Bangalore)
10 - 18 yrs
₹35L - ₹60L / yr
Artificial Intelligence (AI)
Large Language Models (LLM) tuning
skill iconPython
Architecture
Technical Architecture
+4 more
  • We are looking for experienced AI/ML Architects to join our AI Engineering service line. In this role, you will anchor the technical delivery of enterprise AI/Agentic AI projects post deal-closure. You will take over from solution architects and lead the design, build, deployment, and optimization of AI/ML systems — ensuring production-grade quality, scalability, and compliance. 
  • You will interface with cross-functional teams, manage engineering complexity, and ensure value realization for customers across industries such as BFSI, HLS, Manufacturing, CMT, Retail, and Energy. 


Key Responsibilities

Architecture & Technical Leadership

Hands-on Engineering & Problem Solving

Required Qualifications

Education : B.Tech/M.Tech or equivalent in Computer Science, Data Science, or a related field.


Experience

● 10+ years in software architecture or engineering with 5+ years in applied AI/ML

system delivery.

● Experience in productionizing AI/ML models and building full-stack AI applications in

enterprise settings.

● Strong Python development skills; proficiency in ML/AI frameworks (PyTorch,

TensorFlow, Scikit-learn).

● Strong understanding of LLMs, RAG pipelines, vector databases (Weaviate, Qdrant,

Pinecone).


● Experience with MLOps/LLMOps tools: MLflow, Argo, KServe, Feast, Kubeflow.

● Proficiency in data pipeline engineering using Spark, Airflow, or DataFlow.

● Exposure to agent orchestration frameworks: LangChain, LangGraph, AutoGen,

CrewAI is a big plus.

● Cloud & Infrastructure

● Hands-on experience with GCP (Vertex AI, BigQuery, Document AI, AI Gateway)

and/or Azure (Azure ML, OpenAI, Synapse).

● Expertise in containerization (Docker) and orchestration (Kubernetes).

● Familiarity with Infrastructure as Code (Terraform, Pulumi, CDK).


Soft Skills

Strong architectural thinking and problem-solving in fast-paced delivery environments.

Excellent communication and collaboration skills to work across cross-functional teams and

clients.

Proactive, structured, and detail-oriented with a bias for execution.

Nice to Have

Experience in real-world deployments of Agentic AI systems or collaborative multi-agent setups.

Exposure to regulatory/ethical concerns in AI such as fairness, transparency, or bias mitigation.

Familiarity with AI observability, explainability, and governance tooling (e.g., Arize, Fiddler,

TruEra).

Read more
Remote only
6 - 12 yrs
₹45L - ₹50L / yr
skill iconPython
skill iconReact.js
skill iconJavascript
API management
RESTful APIs

About the Role

We are seeking a hands-on Tech Lead to design, build, and integrate AI-driven systems that automate and enhance real-world business workflows. This is a high-impact role for someone who enjoys full-stack ownership — from backend AI architecture to frontend user experiences — and can align engineering decisions with measurable product outcomes.

You will begin as a strong individual contributor, independently architecting and deploying AI-powered solutions. As the product portfolio scales, you will lead a distributed team across India and Australia, acting as a System Integrator to align engineering, data, and AI contributions into cohesive production systems.

Example Project

Design and deploy a multi-agent AI system to automate critical stages of a company’s sales cycle, including:

  • Generating client proposals using historical SharePoint data and CRM insights
  • Summarizing meeting transcripts
  • Drafting follow-up communications
  • Feeding structured insights into dashboards and workflow tools

The solution will combine RAG pipelines, LLM reasoning, and React-based interfaces to deliver measurable productivity gains.

Key Responsibilities

  • Architect and implement AI workflows using LLMs, vector databases, and automation frameworks
  • Act as a System Integrator, coordinating deliverables across distributed engineering and AI teams
  • Develop frontend interfaces using React/JavaScript to enable seamless human-AI collaboration
  • Design APIs and microservices integrating AI systems with enterprise platforms (SharePoint, Teams, Databricks, Azure)
  • Drive architecture decisions balancing scalability, performance, and security
  • Collaborate with product managers, clients, and data teams to translate business use cases into production-ready systems
  • Mentor junior engineers and evolve into a broader leadership role as the team grows

Ideal Candidate Profile

Experience Requirements

  • 5+ years in full-stack development (Python backend + React/JavaScript frontend)
  • Strong experience in API and microservice integration
  • 2+ years leading technical teams and coordinating distributed engineering efforts
  • 1+ year of hands-on AI project experience (LLMs, Transformers, LangChain, OpenAI/Azure AI frameworks)
  • Prior experience in B2B SaaS environments, particularly in AI, automation, or enterprise productivity solutions

Technical Expertise

  • Designing and implementing AI workflows including RAG pipelines, vector databases, and prompt orchestration
  • Ensuring backend and AI systems are scalable, reliable, observable, and secure
  • Familiarity with enterprise integrations (SharePoint, Teams, Databricks, Azure)
  • Experience building production-grade AI systems within enterprise SaaS ecosystems




Read more
Service Co
Service Co
Agency job
via by Rishika Teja
Pune, Mumbai
5 - 10 yrs
₹15L - ₹40L / yr
Artificial Intelligence (AI)
Generative AI
Large Language Models (LLM)
Retrieval Augmented Generation (RAG)

Hiring for AI Engineer


Exp: 5 - 10 yrs

Edu : BE/B.Tech/MCA

Work Location : Pune / Mumbai


Skill Set:


Total experience ranging from 5–10 years in software engineering/AI roles

Min 5 years strong programming experience in Python is a MUST

Min 3.5 years hands-on experience in AI with LLMs, RAG pipelines, and AI frameworks

2+ years shipping LLM systems in production

Experience with cloud platforms (AWS/Azure/GCP)

Read more
company logo
Sandra Saravanan
Posted by Sandra Saravanan
Bengaluru (Bangalore)
8 - 11 yrs
Best in industry
skill iconPython
Generative AI
Generative AI (GenAI)
LLM
Agentic AI Frameworks
+7 more

AI Developer  


Primary Skill-set (Must have) 

  • Generative AI Expertise: 2-3 years of experience in designing and implementing generative AI solutions, including knowledge of various generative and autoregressive models. Ability to apply generative AI techniques to diverse use cases such as image generation, text generation, and creative content synthesis. 

• 2 years of experience in prompt engineering, fine tuning, agentic framework, GenAI SDK’s 

• 1-2 years of experience in Agentic AI frameworks like Autogen, Lanngraph, MS Agent SDK, A2A, MCP, A2P, memory concepts, multi agent orchestration 

• 7+ years of experience in Python 

• 5+ years of experience in software development 

• Azure Proficiency: 3-5 years of experience with Azure cloud services relevant to AI, including Azure Machine Learning, Azure Cognitive Services, Azure Databricks, and Azure Kubernetes Service (AKS). 2+ years of experience in Azure's capabilities to architect end-to-end AI solutions and optimize performance. 

• Architecture Design: 3-5 years of skills with the ability to design scalable, reliable, and cost-effective architectures for AI solutions. Proficiency in designing distributed systems, microservices architectures, and containerized solutions using technologies such as Docker and Kubernetes. 



Secondary Skills (Good to have) 

• Security and Compliance: Understanding of security principles and best practices in AI development, with the ability to implement security controls, encryption mechanisms, and access management policies to protect AI models and sensitive data. 

• Integration and Deployment: Proficiency in implementing CI/CD pipelines, automation scripts, and infrastructure as code (IaC) using tools such as Azure DevOps, Terraform, or Ansible. Experience in containerization and orchestration of AI workloads using Docker and Kubernetes. 

• Software Development: Strong programming skills in languages such as Python, with experience in developing AI applications, RESTful APIs, and microservices architectures. Familiarity with software development methodologies such as Agile or Scrum. 

• Communication and Presentation: Excellent communication skills with the ability to convey complex technical concepts to non-technical stakeholders. Experience in preparing and delivering technical presentations, architecture diagrams, and documentation to communicate architectural decisions and design rationale effectively. 

Read more
company logo
anju kushwaha
Posted by anju kushwaha
Gurugram
4 - 6 yrs
₹20L - ₹50L / yr
Generative AI (GenAI)
MLOps
Large Language Models (LLM)
skill iconData Science
PyTorch
+2 more

Job Description:

We are seeking a versatile and highly skilled Lead AI/ML Engineer with deep expertise in Generative AI (GenAI) and Large Language Models (LLMs). This role requires a leader who can take full ownership of the

AI lifecycle—from initial architectural design to final production execution. You will lead the development of scalable AI-powered applications, demonstrating exceptional execution skills and the ability to deliver high-performance results under pressure in demanding production environments.


Machine Learning & LLM Capability:

 End-to-End ML Engineering: Build and manage comprehensive ML pipelines, including data ingestion, preprocessing, training, and evaluation using frameworks like PyTorch, TensorFlow, and Scikit-learn.  Advanced LLM Systems: Design and implement sophisticated LLM-based applications such as autonomous agents, chatbots, and complex automation tools.

 Generative AI Specialization: Architect and optimize Retrieval-Augmented Generation (RAG) pipelines using vector databases like FAISS, Pinecone, or Weaviate.

 Model Optimization: Fine-tune open-source and proprietary models (e.g., LLaMA, GPT) using advanced techniques like LoRA, QLoRA, or instruction tuning.

 Agentic Frameworks: Develop complex agentic workflows utilizing frameworks such as LangChain or LlamaIndex.

 Prompt Engineering: Implement expert-level prompt engineering, tool/function calling, and structured output generation.

 Project Ownership & Execution

 Full Lifecycle Ownership: Take complete accountability for the full ML and GenAI lifecycle, spanning data processing, model development, monitoring, and optimization.

 Architectural Leadership: Drive strategic architectural decisions for AI platforms, ensuring they are modular, scalable, and maintainable.

 Execution Excellence: Write clean, high-performance Python code following strict OOP principles and manage CI/CD pipelines for seamless project execution.

 Leadership & Mentoring: Act as a key technical leader, managing stakeholders and mentoring team members to ensure all project milestones are met with quality.

 System Integrity: Manage model and prompt versioning, experiment tracking, and comprehensive documentation for all pipelines and workflows.

 Performance Under Pressure

 Production Reliability: Ensure all AI systems maintain extreme scalability and performance under heavy production workloads, including both batch and real-time processing.

 High-Pressure Optimization: Rapidly optimize inference latency and system costs for ML and LLM systems to meet urgent business and technical requirements.

 Proactive Problem Solving: Apply strong analytical thinking to address complex challenges such as system drift, hallucinations, and latency in fast-paced environments.

 Robust Guardrails: Implement and manage strict evaluation frameworks and feedback loops to maintain system quality under stress.


Qualifications:

 Bachelor’s or Master’s degree in Computer Science, AI, ML, or a related field.

 Proven expertise in Python, system design, and scalable AI/ML architecture.

 Deep knowledge of NLP, Computer Vision, and Deep Learning models.

 Hands-on experience with Docker, Kubernetes, MLOps, and major cloud platforms (AWS, GCP, or Azure).

Read more
Why apply to jobs via Cutshort
people_solving_puzzle
Personalized job matches
Stop wasting time. Get matched with jobs that meet your skills, aspirations and preferences.
people_verifying_people
Verified hiring teams
See actual hiring teams, find common social connections or connect with them directly.
ai_chip
Move faster with AI
We use AI to get you faster responses, recommendations and unmatched user experience.
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos