Cutshort logo
For Employers
Big Four Organization logo
Generative AI Engineer
Big Four Organization
Generative AI Engineer

Generative AI Engineer at Big Four Organization · Hyderabad, Bengaluru (Bangalore), Gurugram · 6 - 9 years · ₹15L - ₹32L / yr · Posted 24 Jun 2026

ANLAGE INFOTECH's logo

Generative AI Engineer

at Big Four Organization

Agency job
6 - 9 yrs
₹15L - ₹32L / yr
Hyderabad, Bengaluru (Bangalore), Gurugram
Skills
Generative AI
Generative AI (GenAI)
Large Language Models (LLM)
Retrieval Augmented Generation (RAG)
Agentic AI
AI Agents
Artificial Intelligence (AI)
skill iconData Science
Prompt engineering
LangChain
OpenAI API
Azure OpenAI
LangGraph
GitHub Copilot
skill iconPython

Exp - 6-8.5 Yrs

Level - Assistant Manager

Skill - AI ML/Gen AI Engineer

Location - Hyderabad/ Bangalore / Gurgaon

Key Skills - GitHub Copilot and Gemini; AI/ML and Generative AI (GenAI); LLMs and GenAI agents/assistants (e.g., LangChain); cloud integration; and Python/R workflows


Key Qualifications:


§ A bachelor’s degree in computer science, software engineering, or a related discipline. An advanced degree (e.g., MS) is preferred but not required. Experience is the most relevant factor.

§ Strong software engineering foundation with deep understanding of OOPs, data-structure, algorithms, code instrumentations, beautiful coding practices etc.

§ 5+ years of experience with AI/ML, with last 2 years focused on GenAI as well as technologies like OpenAI, Claude, Gemini, LangChain, Agents, Vector databases, and approaches like Prompt Engineering, fine-tuning, etc.

§ Proven experience in: Python, R, TensorFlow, PyTorch, Keras, Julia, ML libraries, NLP, etc.

§ Proven experience with big data technologies, Angular, React, NodeJS, Python, C#, .NET Core, Java, Golang, SQL/NoSQL.

§ Proven experience with cloud-native engineering, using FaaS/PaaS/micro-services on cloud hyper-scalers like Azure, AWS, and GCP.

§ Strong understanding of methodologies & tools like, XP, Lean, SAFe, DevSecOps, SRE, ADO, GitHub, SonarQube, etc.

§ Excellent interpersonal and organizational skills, with the ability to handle diverse situations, complex projects, and changing priorities, behaving with passion, empathy, and care.

Read more
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos

Similar jobs (10)

company logo
Anish N
Posted by Anish N
Bengaluru (Bangalore)
3 - 5 yrs
₹10L - ₹20L / yr
skill iconPython
Generative AI
Agentic AI
LangChain
LlamaIndex
+3 more

Job Description:

We are looking for a hands-on AI Engineer with experience in Generative AI and Agentic AI to build and deploy production-ready AI solutions.

Key Responsibilities:

  • Develop and deploy GenAI and Agentic AI applications.
  • Build RAG pipelines, LLM workflows, and AI agents.
  • Develop solutions using Python, LangChain, LangGraph, LlamaIndex, or similar frameworks.
  • Implement tool calling, context retrieval, and LLM orchestration.
  • Integrate AI solutions with APIs and cloud platforms.
  • Work with AWS/Azure/GCP, Docker, and CI/CD.

Required Skills:

  • Strong Python programming skills.
  • 3+ years of GenAI/Agentic AI experience.
  • RAG and LLM orchestration.
  • LangChain / LangGraph / LlamaIndex / AutoGen / CrewAI / Semantic Kernel.
  • MCP and A2A knowledge.
  • Cloud, APIs, Docker, and CI/CD experience.

Preferred Experience:

Hands-on experience building and deploying production-ready AI solutions.

Read more
company logo
Prithisha Kathiresan
Posted by Prithisha Kathiresan
Bengaluru (Bangalore)
3 - 8 yrs
Best in industry
Generative AI (GenAI)
Large Language Models (LLM) tuning
Retrieval Augmented Generation (RAG)
Azure OpenAI
skill iconAmazon Web Services (AWS)
+2 more

Senior Generative AI Engineer

Employment Type: Permanent with VDart Digital

Work Location: Marathalli, Bengaluru

Job Description

We are seeking a highly skilled Senior Generative AI Engineer with strong expertise in designing, developing, and deploying enterprise-scale AI solutions using Large Language Models (LLMs) and modern Generative AI frameworks. The ideal candidate should have hands-on production experience building scalable GenAI applications, AI agents, autonomous workflows, and Retrieval-Augmented Generation (RAG) systems in cloud-native environments.

This role requires deep technical expertise in LLM orchestration, AI application architecture, prompt engineering, vector databases, MLOps, and production deployment of AI systems. Candidates should have proven experience delivering real-world AI solutions in enterprise environments with strong exposure to cloud platforms and DevOps practices.

Key Responsibilities

  • Design, build, and deploy enterprise-grade Generative AI applications using Large Language Models (LLMs).
  • Develop intelligent AI agents and autonomous workflows using frameworks such as LangChain, CrewAI, LangGraph, AutoGen, or similar agentic AI frameworks.
  • Implement and optimize Retrieval-Augmented Generation (RAG) pipelines using vector databases and semantic search technologies.
  • Work extensively on prompt engineering, tool calling, memory management, agent orchestration, and multi-agent systems.
  • Integrate and manage LLMs such as OpenAI, Azure OpenAI, Claude, Gemini, Llama, Mistral, or similar foundation models.
  • Develop scalable AI services and APIs using Python and FastAPI.
  • Build production-ready AI solutions with high availability, scalability, monitoring, and observability.
  • Deploy and manage AI applications in cloud-native environments using Docker and Kubernetes.
  • Collaborate with Data Science, ML Engineering, and DevOps teams to operationalize AI solutions.
  • Implement CI/CD pipelines and automated deployment processes for AI workloads.
  • Monitor model performance, latency, reliability, and operational efficiency in production environments.
  • Ensure AI solutions follow enterprise security, governance, and responsible AI standards.
  • Evaluate and adopt emerging Generative AI tools, frameworks, and models.

Required Skills

Generative AI & LLM Expertise

  • Strong hands-on experience with Generative AI and Large Language Models (LLMs).
  • Production-level experience building and deploying GenAI applications.
  • Expertise in LangChain, CrewAI, LangGraph, AutoGen, or similar frameworks.
  • Experience with AI agents, autonomous workflows, and multi-agent architectures.
  • Strong understanding of prompt engineering, embeddings, model evaluation, and LLM orchestration.
  • Experience integrating OpenAI, Azure OpenAI, Claude, Gemini, Llama, Mistral, or similar models.

RAG & Vector Databases

  • Strong experience implementing RAG pipelines and semantic retrieval systems.
  • Experience with vector databases such as Pinecone, Weaviate, ChromaDB, FAISS, or Milvus.
  • Understanding of chunking strategies, embeddings, indexing, reranking, and retrieval optimization.

Python & AI Development

  • Strong proficiency in Python.
  • Experience with FastAPI for AI service and API development.
  • Experience with AI/ML libraries and data processing tools such as Pandas and NumPy.

Cloud & Production Deployment

  • Mandatory production experience on at least one cloud platform:
  • Microsoft Azure
  • Experience deploying scalable AI applications in enterprise production environments.
  • Hands-on experience with Docker, Kubernetes, Jenkins, Terraform, and CI/CD pipelines.
  • Strong understanding of MLOps, AI deployment lifecycle, monitoring, and observability.

Engineering & Operational Excellence

  • Strong understanding of software engineering best practices.
  • Experience with Git, version control, automated testing, and release management.
  • Experience building secure, scalable, and high-performance AI solutions.
  • Ability to troubleshoot production AI systems and optimize performance.

Preferred Skills

  • Experience with AI observability and evaluation frameworks.
  • Exposure to fine-tuning, PEFT, LoRA, or model optimization techniques.
  • Experience with enterprise AI governance and responsible AI practices.
  • Knowledge of distributed AI systems and scalable inference architectures.
  • Familiarity with AI security and compliance standards.

Qualifications

  • Bachelor’s or Master’s degree in Computer Science, Artificial Intelligence, Data Science, or a related field.
  • 3–8 years of overall software engineering experience.
  • Minimum 3+ years of hands-on experience in Generative AI and LLM-based application development,
  • Proven track record of delivering enterprise-scale AI solutions in production environments.
  • Strong communication and stakeholder management skills.
Read more
company logo
Shruti mujbaile
Posted by Shruti mujbaile
Gurugram, Pune
6 - 12 yrs
₹8L - ₹25L / yr
Generative AI
Agentic AI
skill iconPython
skill iconMachine Learning (ML)

Location: Pune / Gurgaon

Position: AI Engineer

work mode: WFO


  Job Description.

​

 Job responsibilities:

  • Responsibility for design, implementation and deployment of Generative AI, Agentic frameworks at scale
  • Strong in programming - Python a
  • Previous experience of working on Computer Vision projects and VLM /VLAM models.
  • In depth awareness of Transformer architectures and End to End Deep neural networks
  • Full stack AI / ML development experience
  • Design, build & maintain efficient and reliable Agentic / Generative AI code leveraging pipelines
  • Hosting and deployment knowledge in GCP or AWS or Azure along with advanced engineering concepts to build user friendly UI interface for easy adoption.


    Requirements:

 ·      4 to 8 years overall years of experience (Agentic AI, Generative AI, VLM, VLAM and LLM) with significant exposure in Development, Architecture design, scaling and hosting in cloud.


    Must Have –

 ·      Architecting and solutioning experience with Python and FAST API, Agentic Ai frameworks, VLMs, VLAMs, Open source LLM’s and Code based LLM models at scale with - Langchain /      Ollama, embeddings, Memory      Management etc.,

·      Practical experience in implementing Explainable and ethical AI models  Practical experience in implementing frameworks like RAG/ CAG/ Self-reflective RAG etc.,

·      Experience in cloud hosting either AWS or Azure or GCP.

·      Experience in ML-OPS - Implement a feedback mechanism to continually improve the model over time through feedback loop and monitoring KPI’s in production.

·      Experience with Quantization and Kubernetes or docker


    Good to have

·      gRPC implementation to expose the API’s on a server for easy usage and good user interface

·      Streamlit front end creation

·      Experience with SAFe framework deliveries.


Read more
company logo
Bhawna Khemani
Posted by Bhawna Khemani
icon

The recruiter has not been active on this job recently. You may apply but please expect a delayed response.

Bengaluru (Bangalore), Delhi, Gurugram, Noida, Ghaziabad, Faridabad
4 - 13 yrs
₹11L - ₹35L / yr
Generative AI
Large Language Models (LLM)
Retrieval Augmented Generation (RAG)
skill iconPython
LlamaIndex
+4 more

Generative AI Engineer 

Role Overview:

You will be responsible for the hands-on development, coding, and deployment of AI-powered features. Your focus is on writing clean, efficient code to integrate LLMs into our existing tech stack, building robust data pipelines for RAG, and ensuring the reliability of model outputs through rigorous testing and optimization.

Key Responsibilities

  • Application Implementation: Code and integrate LLM APIs (OpenAI, Anthropic, etc.) or local models into backend services using Python, FastAPI, etc.,
  • MCP Server Development: Design and implement custom MCP servers using the official SDKs (Python/TypeScript) to expose internal databases, APIs, and file systems to AI agents.
  • RAG Implementation: Build and maintain the "plumbing" for Retrieval-Augmented Generation—specifically coding the data ingestion scripts, text chunking logic, and metadata filtering.
  • Vector DB Management: Perform day-to-day operations on vector databases (Pinecone, Milvus, etc.), including indexing, querying, and optimizing search retrieval.
  • Prompt Programming: Develop, version-control, and refine complex prompt templates (using Jinja2 or similar) to ensure consistent structured outputs (JSON/YAML).
  • Agent Development: Implement multi-step workflows using LangChain, LangGraph, CrewAI etc.,, focusing on tool-calling logic and error handling.
  • Evaluation & Testing: Build automated test suites to detect "hallucinations" and measure accuracy using frameworks.
  • Performance Tuning: Implement caching layers and streaming responses to reduce latency and improve the end-user experience; Token optimization.
  • Data Pre-processing: Clean and tokenize datasets for model fine-tuning or high-quality context retrieval.

Technical Skills (The "Execution" Stack)

  • Language: Advanced Python (Asyncio, Pydantic) and optional TypeScript/Node.js (for full-stack integration).
  • AI Frameworks: Hands-on experience with any of LangChain, LlamaIndex, and Hugging Face Transformers. RAG and Vector search concepts.
  • Data Handling: Proficiency in SQL and handling unstructured data formats (PDFs, Markdown, JSON).
  • Deployment: Practical experience with Docker, GitHub Actions (CI/CD), and experience with OpenTelemetry, LangSmith, Weights & Biases etc., Understanding of evaluation/guardrails.
  • MCP/API Proficiency: Deep understanding of RESTful APIs, Streaming HTTP, MCP server vs client, JSONRPC
Read more
Insurance expertise
Insurance expertise
Agency job
via by Priyanka Bisht
Gurugram, Noida
5 - 9 yrs
Best in industry
skill iconPython
"AIML
skill iconMachine Learning (ML)
Artificial Intelligence (AI)
MLOps
+2 more

Job Summary/ Job Opportunity:

This is an excellent opportunity for an ideal candidate with a high level of technical proficiency and meeting the below mentioned criteria -- • Strong experience in Machine Learning, Deep Learning, Generative AI, and Large Language Models (LLMs). • Hands-on experience building and deploying production-grade solutions using Azure OpenAI, OpenAI, LangChain, LangGraph, Semantic Kernel, LlamaIndex, and Agentic AI frameworks. • Strong expertise in Python, API development, microservices, and cloud-native architectures. • Experience designing and implementing RAG solutions, vector databases, embeddings, knowledge retrieval systems, and AI copilots. • Experience with Azure cloud services, MLOps, CI/CD pipelines, monitoring, and model lifecycle management. • Strong understanding of AI governance, responsible AI, security, compliance, and model evaluation frameworks. • Ability to lead technical discussions, provide architectural recommendations, mentor team members, and interact with business stakeholde


Key Objectives and Major Responsibilities:

• Design, develop, and implement scalable AI/ML and Generative AI solutions for enterprise applications. • Lead development of intelligent applications leveraging LLMs, RAG pipelines, AI agents, and document intelligence solutions. • Collaborate with business stakeholders, architects, and product teams to translate business requirements into technical solutions. • Design and optimize data pipelines, vector search solutions, embeddings, and retrieval mechanisms. • Build and maintain REST APIs, microservices, and cloud-native AI applications. • Ensure best practices in coding standards, performance optimization, security, scalability, and maintainability. • Drive AI solution deployment using MLOps practices, CI/CD pipelines, monitoring, and observability frameworks. • Perform code reviews, mentor junior developers, and contribute to capability building within the team


Key Capabilities and Competencies:

Knowledge, Skills, Qualification and Experience

• Degree in B.Tech/M.Tech (Computer Science/IT/Data Science) or related discipline preferred, with 3–4 years of relevant experience in AI/ML, GenAI and total 5-7 years of experience. • Proficiency in Python and hands-on experience with ML libraries (scikit-learn, TensorFlow, PyTorch) and GenAI frameworks/tools. • Strong understanding of machine learning, deep learning, LLMs, prompt engineering, and techniques like RAG and fine-tuning. • Experience with data processing, embeddings, vector databases, APIs, and building scalable AI driven applications. • Good communication skills, ability to work on multiple projects, and eagerness to learn and adapt to evolving AI technologies. 

Read more
Pune
3 - 6 yrs
₹21L - ₹32L / yr
skill iconPython
Artificial Intelligence (AI)

Strong AI Engineer / Machine Learning Engineer profiles.

2

Mandatory (Experience 1) – Must have minimum 3+ years of hands-on experience in Data Science, Machine Learning, Applied AI, NLP, Deep Learning, or Generative AI solutions.

3

Mandatory (Experience 2) – Must have strong hands-on experience in Python programming, SQL, data analysis, feature engineering, model development, and production-grade ML applications.

4

Mandatory (Experience 3) – Must have experience working with Machine Learning and Deep Learning frameworks such as PyTorch, TensorFlow, Keras, Scikit-learn, or equivalent.

5

Mandatory (Experience 4) – Must have hands-on experience working on NLP, embeddings, semantic search, text classification, document understanding, recommendation systems, or similar AI/ML use cases.

6

Mandatory (Experience 5) – Must have experience working with Large Language Models (LLMs) such as GPT, Llama, Mistral, Claude, Gemini, Phi, or similar foundation models.

7

Mandatory (Experience 6) – Must have hands-on experience building or implementing RAG (Retrieval Augmented Generation) systems, vector search, knowledge retrieval, embeddings, chunking, indexing, or semantic retrieval solutions.

8

Mandatory (Experience 7) – Must have experience working with Git, CI/CD practices, production environments, and scalable AI/ML systems.

9

Mandatory (CTC) – The CTC breakup offered will be 75% fixed + 25% variable, as per company policy.

10

Mandatory (Age) - Candidate's Age should be below 28 Years

11

Preferred (Experience 1) – Experience with MLFlow, Kubeflow, Airflow, Prefect, Feature Stores, Model Registry, or MLOps/LLMOps frameworks.

12

Preferred (Experience 2) – Experience working with Vector Databases, Spark, PySpark, distributed ML pipelines, large-scale data processing, or real-time ML systems..

13

Preferred (Experience 3) – Familiarity with Docker, Kubernetes, Azure, AWS, GCP, cloud-native AI deployments, and scalable ML architecture.

14

Preferred (Company) – Candidates from AI-first startups, Fintech, Banking, Lending, Fraud Analytics, Risk Analytics, Product Companies, SaaS organizations, or data-driven technology companies

15

Mandatory ( Pedigree) - B.TECH / M.TECH from Tier 1 Colleges (IIT's, NIT's, BITS) are Considered.

Read more
company logo
Banu S
Posted by Banu S
Bengaluru (Bangalore), Hyderabad
5 - 12 yrs
₹4L - ₹25L / yr
skill iconData Science
skill iconPython
Large Language Models (LLM) tuning
RAG
Langchain

Support with design and build to prove out agentic AI solution flow by working with other data 

scientists and engineers to build, train Large Language Model (LLM) architectures, RAG 

systems, and autonomous agentic workflows 

Key qualifications: 

  

>> AI solution design & Development: Design Agentic AI solutions using RAG (Retrieval-

Augmented Generation) and orchestration frameworks like LangGraph or LangChain. 

  

>> Model Fine-Tuning: Solid understanding and experience with Pre-train, fine-tune, and 

optimize open-source like BERT, LLama, and other proprietary foundation models for domain-

specific tasks 

  

>> Solid Stats and ML foundations and (vibe) coding skills with Python, PySpark 

  

>>  Implement validation frameworks and tracing practices (using tools like Arize) to monitor 

agent behavior, guard against model drift, and ensure compliance 

  

>> Collaborate with Engineering to deploy models securely on cloud and on-prem ecosystems 

 

Read more
icon

The recruiter has not been active on this job recently. You may apply but please expect a delayed response.

Jaipur
3 - 6 yrs
₹7L - ₹10L / yr
skill iconPython
Large Language Models (LLM)
Generative AI
Retrieval Augmented Generation (RAG)
LangChain
+23 more

Location: Jaipur (Work From Office)

Employment Type: Full-Time


We're looking for a GenAI Engineer (LLM Engineer) to build scalable AI-powered SaaS applications using Large Language Models (LLMs). You'll develop intelligent AI workflows, integrate LLMs into production systems, and build secure, high-performance AI solutions.


Key Responsibilities

  • Integrate LLM APIs (OpenAI, Claude, Hugging Face) into production applications.
  • Design and optimize RAG pipelines and prompt engineering workflows.
  • Build and manage Vector Databases (Pinecone, Weaviate, pgvector).
  • Optimize AI performance, latency, and operational cost.
  • Ensure secure, scalable AI architecture.
  • Collaborate with Product and Engineering teams to deliver AI-powered features.


Requirements

  • 3+ years of backend development using Python, Go, or Node.js.
  • Hands-on experience with LLMs, LangChain or LlamaIndex.
  • Strong understanding of RAG, Prompt Engineering, and Vector Databases.
  • Experience with AWS, GCP, or Azure.
  • Knowledge of APIs, Microservices, and AI application development.


Preferred: Experience in SaaS/FinTech, LLMOps, or Model Fine-tuning.

Education: B.Tech, BCA, or equivalent technical qualification.


Apply Now

Application Form: https://zfrmz.com/pAKb2ynfomIsuNwRfRbV?utm_source=cutshort

Read more
company logo
Nirmala Lama
Posted by Nirmala Lama
Mumbai
1 - 5 yrs
₹18L - ₹23L / yr
Generative AI
Retrieval Augmented Generation (RAG)
LangGraph
LangChain
Large Language Models (LLM)
+2 more

About the role

We are seeking an AI Engineer to build and implement AI systems for content production at scale. You'll work at the intersection of engineering and content designing prompt pipelines, integrating generative models, and building the tooling that turns source material into finished creative output. The ideal candidate is technically strong but also has taste: someone who understands story and craft, and can tell the difference between output that's technically correct and output that's actually good.


Responsibilities

  • Build and iterate on prompt pipelines and multi-agent workflow components
  • Design and integrate agentic workflows orchestrate multi-step, tool-using agents that plan, call models, and hand off between stages in production
  • Deploy and serve open-source models set up inference endpoints, manage GPU compute, and optimize for latency and cost
  • Write evals compare outputs against references, quantify quality, and feed results back into the pipeline
  • Work on data pipelines: structured extraction from messy source text, localization, similarity/dedup
  • Debug and maintain pipeline stages in production


What you bring:  

  • (1+/3+) years of engineering experience, or a strong portfolio of shipped projects
  • Solid Python fundamentals clean, working, readable code
  • Hands-on experience with LLM APIs and prompt engineering (personal projects count)
  • Comfort with Git, REST APIs, and working in a Linux environment
  • A feel for content and narrative you can judge whether generated output is actually good, not just valid
  • Curiosity and clear communication you ask good questions and don't stay stuck silently


 Preferred

  • Exposure to agent/orchestration frameworks (LangGraph, LangChain, CrewAI)
  • Familiarity with vector databases, embeddings, or RAG (Qdrant, pgvector)
  • Hands-on work with open-source generative media models Flux, LTX, Wan, or similar
  • Experience deploying open-source models for inference (vLLM, ComfyUI, Replicate/Cog, Docker + GPU)
  • Experience writing evals or LLM-as-judge scoring
  • Node.js and Fastapi familiarity, or experience deploying on AWS


Read more
company logo
Waseem Shariff
Posted by Waseem Shariff
Mumbai
1 - 3 yrs
₹15L - ₹24L / yr
Generative AI
Retrieval Augmented Generation (RAG)
LangGraph
LangChain
Large Language Models (LLM) tuning
+3 more

About the role

We are seeking an AI Engineer to build and implement AI systems for content production at scale. You'll work at the intersection of engineering and content designing prompt pipelines, integrating generative models, and building the tooling that turns source material into finished creative output. The ideal candidate is technically strong but also has taste: someone who understands story and craft, and can tell the difference between output that's technically correct and output that's actually good.


Responsibilities

  • Build and iterate on prompt pipelines and multi-agent workflow components
  • Design and integrate agentic workflows orchestrate multi-step, tool-using agents that plan, call models, and hand off between stages in production
  • Deploy and serve open-source models set up inference endpoints, manage GPU compute, and optimize for latency and cost
  • Write evals compare outputs against references, quantify quality, and feed results back into the pipeline
  • Work on data pipelines: structured extraction from messy source text, localization, similarity/dedup
  • Debug and maintain pipeline stages in production


What you bring:  

  • (1+/3+) years of engineering experience, or a strong portfolio of shipped projects
  • Solid Python fundamentals clean, working, readable code
  • Hands-on experience with LLM APIs and prompt engineering (personal projects count)
  • Comfort with Git, REST APIs, and working in a Linux environment
  • A feel for content and narrative you can judge whether generated output is actually good, not just valid
  • Curiosity and clear communication you ask good questions and don't stay stuck silently


 Preferred

  • Exposure to agent/orchestration frameworks (LangGraph, LangChain, CrewAI)
  • Familiarity with vector databases, embeddings, or RAG (Qdrant, pgvector)
  • Hands-on work with open-source generative media models Flux, LTX, Wan, or similar
  • Experience deploying open-source models for inference (vLLM, ComfyUI, Replicate/Cog, Docker + GPU)
  • Experience writing evals or LLM-as-judge scoring
  • Node.js and Fastapi familiarity, or experience deploying on AWS


Read more
Why apply to jobs via Cutshort
people_solving_puzzle
Personalized job matches
Stop wasting time. Get matched with jobs that meet your skills, aspirations and preferences.
people_verifying_people
Verified hiring teams
See actual hiring teams, find common social connections or connect with them directly.
ai_chip
Move faster with AI
We use AI to get you faster responses, recommendations and unmatched user experience.
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos