Cutshort logo
GetSetYo Technology Labs Private Limited's logo

Senior AI Engineer

Abhishek Agrawal's profile picture
Posted by Abhishek Agrawal
1 - 10 yrs
₹20L - ₹60L / yr (ESOP available)
Bengaluru (Bangalore)
Skills
Generative AI
Agentic AI

AI Engineer (1–8 years)

Location: Bellandur, Outer Ring Road, Bengaluru (Work from Office only)

Organization Size: 20 members across functions, and growing

Reports to: Head of Engineering

GetSetYo is a Bangalore-based early-stage travel tech startup, built by internet industry veterans (ex-Makemytrip, Flipkart, Ola, PhonePe, Zynga, MagicPin, etc.) and premier academic institutes (IIT Delhi, IIT BHU, ISB, DCE, NIT Surathkal, etc.), and funded by multiple unicorn founders (of companies such as MakeMyTrip, Zomato, Groww, Udaan, MaMaEarth, etc.) and we’re growing fast. Look us up here - https://www.getsetyo.com/about

We’re building something exciting in the travel tech space and looking for a senior AI Engineer to join our core engineering team in Bangalore.

Who Are You:

  • 1–8 years of total software engineering experience, including at least 1 year building and shipping AI/ML or LLM-powered products in production
  • Engineering degree from a top-ranked college
  • Strong engineering foundation in Python or Java, with the ability to build reliable backend services, APIs, evaluation pipelines, and developer tooling around AI systems
  • Hands-on experience with LLM application patterns such as RAG, tool/function calling, structured output generation, vector search, reranking, and agentic workflows
  • Familiarity with agent frameworks and orchestration patterns, including multi-step workflows, planner/executor patterns, tool routing, and guardrails
  • Working knowledge of MCP (Model Context Protocol) or similar patterns for connecting models to internal tools, data sources, and external systems
  • Strong understanding of context engineering, prompt design, and how to manage instructions, conversation state, tools, memory, and retrieved context for consistent model behavior
  • Experience with evaluation and observability for AI systems: offline evals, online metrics, regression testing, trace inspection, cost/latency monitoring, and failure analysis
  • Comfortable working in a fast-paced startup where you can own problem statements end to end — from prototype to production rollout
  • Must have experience using AI-native developer tools such as Claude Code / coding agents / AI-assisted workflows to accelerate delivery

What You’ll Do

  • Build and own production-grade AI features across the stack, from experimentation and prototyping to backend integration, deployment, monitoring, and iterative improvement
  • Design and implement agentic workflows for real user problems — combining LLM reasoning, retrieval, tool use, business rules, and backend APIs into reliable multi-step systems
  • Build and optimize RAG and search systems: document ingestion, chunking strategies, embedding pipelines, vector indexes, hybrid retrieval, reranking, and citation/grounding flows
  • Integrate models with internal and external systems through tool calling, APIs, and where relevant MCP-compatible interfaces, so models can safely access the right context and take useful actions
  • Drive context engineering for AI products: decide what memory, instructions, retrieved context, tool outputs, and interaction history should be passed to the model at each step for maximum quality and efficiency 
  • Build evaluation systems for prompts, agents, and retrieval quality — including benchmark datasets, golden test cases, automated regression checks, and human-in-the-loop review workflows
  • Establish observability and debugging for AI pipelines: traces, tool execution logs, latency/cost tracking, hallucination analysis, and failure-mode investigation 
  • Help define engineering standards for AI systems across security, guardrails, versioning, rollback, experimentation, and cost-performance tradeoffs

What We Offer:

  • AI Impact from Day 1: Lead the development of our core ML capabilities
  • Fast Iteration: Weekly releases and direct user feedback
  • Collaborative Culture: Flat structure and transparent communication
  • Vibrant Office: In-person energy in Bangalore HQ
  • Perks: Employee travel discounts and exclusive deals.


Read more
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos

About GetSetYo Technology Labs Private Limited

Founded :
2022
Type :
Products & Services
Size :
0-20
Stage :
Raised funding

About

We are a funded Travel Tech Startup based out of Bengaluru, with a highly experienced founding team hailing from Institutions such as IIT Delhi, ISB, MakeMyTrip, Flipkart, Ola, PhonePe.


GetSetYo is building a new distribution model for travel. We combine AI-powered technology, travel expertise, and creator-led distribution to help travellers discover and book trips through trusted influencers and communities.


Our platform enables travel creators (YouTubers, Instagrammers, and travel experts) to monetize their audience by generating travel leads and transactions.Travellers either book directly online on the GetSetYo platform or take assistance from our expert travel sales team to plan and book customised trips.


About our Investors

We are funded by respected funds and entrepreneurs including co-founders of Zomato, MakeMyTrip, Groww, MamaEarth, Udaan, MoneyView, Niyo, to name a few.


About our Leadership Team

Abhishek was Vice President at MakeMyTrip. He has rich and diverse experience across other consumer sectors also having served as Senior Director at Flipkart and Ola. He studied Computer Science from IIT Delhi.


Sahil brings in deep experience in engineering, with his previous role being Senior Director Engg at Magicpin, and with companies such as Policybazaar He received his education from Delhi College of Engineering.

Read more

Company social profiles

bloginstagramlinkedinfacebook

Similar jobs

Hashone Career
Madhavan I
Posted by Madhavan I
Bengaluru (Bangalore), Chennai, Coimbatore
5 - 10 yrs
₹20L - ₹40L / yr
skill iconPython
Agentic AI
skill iconMachine Learning (ML)

AI/ML Engineer AI Operating System for Capital Markets Location Bangalore/Chennai Experience 5+ years Function Artificial Intelligence / Machine Learning Employment Type About Transient.AI Full-time Transient.AI is building a next-generation AI Operating System for capital markets — a unified intelligence layer that connects research, trading, compliance, and sales functions at banks and hedge funds. Today, these teams largely operate on disconnected legacy systems, forcing manual, expensive workarounds. Transient.AI replaces that fragmentation with a single AI-native layer built for institutional-grade compliance, security, and auditability. The company already has live products in market, including Caddie.AI (a research automation tool that cuts hedge fund research time significantly), ClarityRIA (helping sales teams identify the right investors in seconds), and CapFlo.AI (automated parsing of complex derivatives contracts). Founded by former traders and technologists from Goldman Sachs, Credit Suisse, UBS, and McKinsey, Transient.AI is headquartered in New York, with teams in Miami, Singapore, and India. The company has raised Series A funding and is scaling its engineering and product organization globally. Role Overview Transient.AI is hiring an experienced AI/ML Engineer to join its India engineering team in Bangalore/Chennai. This is a hands-on, build-from-scratch role — you'll be designing and shipping the core machine learning systems that power the company's flagship products, working closely with founders and senior engineers rather than inheriting existing infrastructure. Key Responsibilities • Design, build, and deploy machine learning and AI models that power Transient.AI's core products (research automation, document intelligence, investor matching, and workflow orchestration). • Workonapplied NLP/LLMsystems, including retrieval-augmented generation, structured extraction from unstructured financial documents, and model evaluation pipelines. • Partner closely with product and founding engineers to translate capital markets workflows into scalable AI systems. • Ownmodelperformance, reliability, and cost — from experimentation through production deployment. • Build and maintain data pipelines, feature stores, and evaluation frameworks to support rapid iteration. • Ensuresystems meet the compliance, auditability, and security standards required in regulated financial environments. What We're Looking For • 5+years ofexperience building and deploying machine learning or AI systems in production.• Strong hands-on experience with Python and modern ML/AI frameworks (PyTorch, TensorFlow, Hugging Face, LangChain, or equivalent). • Experience with LLMs — fine-tuning, prompt engineering, RAG architectures, or agentic systems — is highly valued. • Solid grounding in data structures, distributed systems, and MLOps practices (model serving, monitoring, versioning). • Prior experience at a strong product company, high-growth startup, or a top-tier engineering background • Comfort operating in an early-stage, high-ownership environment with limited process and high ambiguity. • Exposure to fintech, capital markets, or other regulated industries is a plus, though not mandatory. WhyJoin Transient.AI • Build core AI systems from the ground up — not maintain legacy code. • Workdirectly with founders who have deep, first-hand Wall Street experience (Goldman Sachs, Credit Suisse, UBS, McKinsey). • JoinaSeries A-funded company solving a real, expensive problem for institutional finance. • Bepart ofasmall, global team with outsized ownership and impact. .

Read more
Impacto Digifin Technologies
at Impacto Digifin Technologies
4 candid answers
1 recruiter
Navitha Reddy
Posted by Navitha Reddy
Bengaluru (Bangalore)
2 - 3 yrs
₹6L - ₹8L / yr
skill iconMachine Learning (ML)
skill iconDeep Learning
Natural Language Processing (NLP)
Voice Over IP (VoIP)
Artificial Intelligence (AI)
+4 more

Job Title: AI/ML Engineer – Voice (2–3 Years)

Location: Bengaluru (On-site)

Employment Type: Full-time


About Impacto Digifin Technologies

Impacto Digifin Technologies enables enterprises to adopt digital transformation through intelligent, AI-powered solutions. Our platforms reduce manual work, improve accuracy, automate complex workflows, and ensure compliance—empowering organizations to operate with speed, clarity, and confidence.


We combine automation where it’s fastest with human oversight where it matters most. This hybrid approach ensures trust, reliability, and measurable efficiency across fintech and enterprise operations.


Role Overview

We are looking for an AI Engineer Voice with strong applied experience in machine learning, deep learning, NLP, GenAI, and full-stack voice AI systems.


This role requires someone who can design, build, deploy, and optimize end-to-end voice AI pipelines, including speech-to-text, text-to-speech, real-time streaming voice interactions, voice-enabled AI applications, and voice-to-LLM integrations.


You will work across core ML/DL systems, voice models, predictive analytics, banking-domain AI applications, and emerging AGI-aligned frameworks. The ideal candidate is an applied engineer with strong fundamentals, the ability to prototype quickly, and the maturity to contribute to R&D when needed.


This role is collaborative, cross-functional, and hands-on.


Key Responsibilities

Voice AI Engineering

  • Build end-to-end voice AI systems, including STT, TTS, VAD, audio processing, and conversational voice pipelines.
  • Implement real-time voice pipelines involving streaming interactions with LLMs and AI agents.
  • Design and integrate voice calling workflows, bi-directional audio streaming, and voice-based user interactions.
  • Develop voice-enabled applications, voice chat systems, and voice-to-AI integrations for enterprise workflows.
  • Build and optimize audio preprocessing layers (noise reduction, segmentation, normalization)
  • Implement voice understanding modules, speech intent extraction, and context tracking.

Machine Learning & Deep Learning

  • Build, deploy, and optimize ML and DL models for prediction, classification, and automation use cases.
  • Train and fine-tune neural networks for text, speech, and multimodal tasks.
  • Build traditional ML systems where needed (statistical, rule-based, hybrid systems).
  • Perform feature engineering, model evaluation, retraining, and continuous learning cycles.

NLP, LLMs & GenAI

  • Implement NLP pipelines including tokenization, NER, intent, embeddings, and semantic classification.
  • Work with LLM architectures for text + voice workflows
  • Build GenAI-based workflows and integrate models into production systems.
  • Implement RAG pipelines and agent-based systems for complex automation.

Fintech & Banking AI

  • Work on AI-driven features related to banking, financial risk, compliance automation, fraud patterns, and customer intelligence.
  • Understand fintech data structures and constraints while designing AI models.

Engineering, Deployment & Collaboration

  • Deploy models on cloud or on-prem (AWS / Azure / GCP / internal infra).
  • Build robust APIs and services for voice and ML-based functionalities.
  • Collaborate with data engineers, backend developers, and business teams to deliver end-to-end AI solutions.
  • Document systems and contribute to internal knowledge bases and R&D.

Security & Compliance

  • Follow fundamental best practices for AI security, access control, and safe data handling.
  • Awareness of financial compliance standards (plus, not mandatory).
  • Follow internal guidelines on PII, audio data, and model privacy.

Primary Skills (Must-Have)

Core AI

  • Machine Learning fundamentals
  • Deep Learning architectures
  • NLP pipelines and transformers
  • LLM usage and integration
  • GenAI development
  • Voice AI (STT, TTS, VAD, real-time pipelines)
  • Audio processing fundamentals
  • Model building, tuning, and retraining
  • RAG systems
  • AI Agents (orchestration, multi-step reasoning)

Voice Engineering

  • End-to-end voice application development
  • Voice calling & telephony integration (framework-agnostic)
  • Realtime STT ↔ LLM ↔ TTS interactive flows
  • Voice chat system development
  • Voice-to-AI model integration for automation

Fintech/Banking Awareness

  • High-level understanding of fintech and banking AI use cases
  • Data patterns in core banking analytics (advantageous)

Programming & Engineering

  • Python (strong competency)
  • Cloud deployment understanding (AWS/Azure/GCP)
  • API development
  • Data processing & pipeline creation

Secondary Skills (Good to Have)

  • MLOps & CI/CD for ML systems
  • Vector databases
  • Prompt engineering
  • Model monitoring & evaluation frameworks
  • Microservices experience
  • Basic UI integration understanding for voice/chat
  • Research reading & benchmarking ability

Qualifications

  • 2–3 years of practical experience in AI/ML/DL engineering.
  • Bachelor’s/Master’s degree in CS, AI, Data Science, or related fields.
  • Proven hands-on experience building ML/DL/voice pipelines.
  • Experience in fintech or data-intensive domains preferred.

Soft Skills

  • Clear communication and requirement understanding
  • Curiosity and research mindset
  • Self-driven problem solving
  • Ability to collaborate cross-functionally
  • Strong ownership and delivery discipline
  • Ability to explain complex AI concepts simply



Read more
Remote only
7 - 12 yrs
₹40L - ₹70L / yr
Agentic AI
skill iconPython
API management
Anthropic Claude

Experience: 8+ years, senior candidates only | Type: Full-time | Location: Remote (India)


 ---

 WHAT WE'RE BUILDING


 See http://www.juliet.space


 We're building Juliet, an AI that runs marketing end to end. Our users are marketers, founders, CEOs, growth leads, agencies, and SMBs — not developers. They

 tell Juliet the goal. She plans, writes production code, and ships real marketing: conversion-optimized websites, launch assets, campaigns, audits, autonomously.


 That's the engineering problem in one line: the humans in the loop can't read code, so the agent has to get it right on her own — plan, build, self-correct,

 recover, ship.


 Under the hood: a browser-based studio backed by cloud sandboxes, a real-time SSE streaming pipeline, and a LangGraph agent working across 83 tools and 63 skill

 modules. The agent isn't bolted onto the product. She is the product.


 Small team, big ambitions. You'll ship things users touch daily, not write tickets about them.


 ---

 THE ROLE


 We're hiring one architect-level backend engineer to own Juliet's agentic infrastructure end to end. That means the agent graph, the execution environment, the

 streaming pipeline, the state and memory systems — and setting technical direction for the engineers working alongside you.


 This is a player-coach seat. You'll still write code every day, and your architectural calls become the product. You'll work directly with the founder. No PMs in

 between.


 Frontend is part of the system. You won't be leading it, but you'll need to understand how the agent's output reaches the browser and be able to ship full-stack

 features when needed.


 ---

 THE STACK


 AI agent (primary): Python 3.11, LangGraph 1.x + LangChain, Anthropic / Google / OpenAI model providers


 API (primary): NestJS 11, Supabase, Redis, PostgreSQL, Server-Sent Events


 Infra (primary): Modal cloud sandboxes, Docker, Netlify deployments


 Frontend (secondary): Next.js 15, React 19, TypeScript, Zustand, CodeMirror 6, XTerm.js


 Monorepo: Turborepo, pnpm


 ---

 WHAT YOU'LL WORK ON


 The majority of your time is here:


 Agentic AI workflows — Design, extend, and harden the LangGraph agent graph: multi-step planning, code generation, tool dispatch, self-correction, and recovery

 across 83 tools and 63 skill modules. This is the core of the product.


 Real-time streaming architecture — The SSE pipeline that carries every agent action from the Python backend through NestJS to the browser: event framing,

 reconnection, health monitoring, interrupt handling for plan approvals and clarifying questions.


 Agent execution environments — Sandbox lifecycle on Modal: container spin-up, file sync, terminal I/O, command execution, and live preview with per-asset esbuild

 bundling. The agent lives here.


 State and memory systems — LangGraph Postgres checkpointers, middleware-injected context (goals, design docs, memory anchors), conversation summarization. How

 the agent knows what it knows.


 Backend API and data layer — NestJS services, Supabase schema, Redis caching, quota enforcement, webhook handling. The plumbing the agent depends on.


 Marketing intelligence pipelines — AEO, CRO, and brand-perception audit engines: multi-LLM probing, parallel inference, streamed structured reports, result

 caching. Audit-at-scale infrastructure.


 The remaining ~25% of your time:


 Full-stack product features — Collaboration (roles and permissions), the Netlify deployment pipeline, subscription and quota flows, onboarding. You'll ship these

 end to end — backend first, frontend to close the loop.


 ---

 WHAT WE'RE LOOKING FOR


 Must-have:


 - 8+ years of professional software engineering, including meaningful time as a tech lead or systems architect who owned something end to end. Closer to ten is

 the norm for people who thrive here.

 - Both worlds on your resume: engineering rigor inside a large company and 0-to-1 ownership at an early-stage startup.

 - Production agentic systems experience. You've built and operated LLM agent systems in production with LangGraph, LangChain, or equivalent — agent graphs, tool

 use, state management, prompt engineering, evals. This means well beyond calling a chat endpoint.

 - Strong Python. You design and ship production Python daily. The agent codebase is yours to own.

 - Architect-level system design. You can own how data flows across four services, make tradeoffs under uncertainty, and defend every call.

 - AI-native development workflow. You drive Claude Code, Codex, or similar agentic tools as everyday instruments — not occasionally. You have opinions about

 working with coding agents because you do it constantly.

 - Real-time backend systems. You've built SSE, WebSocket, or streaming API infrastructure in production — not just consumed it.

 - Strong TypeScript. The API layer and most product features are in TypeScript. You're productive in it.


 Strong plus:


 - Background in developer tools, IDEs, or coding/execution platforms

 - Container runtimes and sandboxed execution (Modal, E2B, Firecracker, or similar)

 - Depth in PostgreSQL, Redis, and Supabase

 - LLM observability and evals tooling (LangSmith or similar)

 - NestJS or equivalent Node.js API framework experience

 - React/Next.js — enough to ship a full-stack feature without handoff

 - Exposure to marketing, growth, or publisher-facing products


 ---

 WHY THIS ROLE IS DIFFERENT

  

 You own the architecture. Not a feature factory. Not someone else's design doc. The technical execution of an AI product is yours to lead.


 The agent is the product. You're not adding AI to an existing system. You're building and operating the system that is the AI. Every architectural decision

 touches what Juliet can and can't do.


 Hard problems, always. The system spans cloud sandboxes, streaming infrastructure, multi-step agent graphs, and a full-stack web product — for non-technical

 users who can't course-correct a broken output. The bar is high.


 Small team, real leverage. Your code ships to users the same week. No layers of approval.


 ---

 HOW TO APPLY


 Send us:


 1. A short note on the most complex agentic system you've shipped: what broke, and what you'd redo. A link to something you've built that involves agent graphs, tool use, or autonomous multi-step execution


 2. What is one thing you would improve about Juliet? It could be a feature or a bug.

Read more
Terrabase
Remote only
3 - 15 yrs
₹20L - ₹50L / yr
skill iconPython
Large Language Models (LLM)
AI Agents
Performance Evaluation
Generative AI
+2 more

Experience: 5+ years production software engineering, with 2+ years working directly on LLM or agent systems in production. 

Location: Remote 

To streamline and fast-track screening, please submit your details here (if you haven’t already): https://airtable.com/appbtkr4odapnb5I6/pagqo91lKv3VJg3GT/form 


We’ll review your responses as part of the initial screening process. Please make sure you complete and submit all details through the form to be considered for the next stage. Submissions outside the form may not be considered.


Why This Role Matters

Terrabase builds agent infrastructure that enterprise customers rely on daily for SQL generation, forecasting, data analysis, and artifact delivery. Our orchestration layer routes between specialized sub-agents, manages typed handoff contracts, runs structured eval suites, and enforces correctness across every turn.


This is not a research-prototype role. You will build and evolve agent architecture, but always in service of making the system observable, typed, evaluated, recoverable, and boringly reliable in production.



What You Will Do


Own the harness architecture and middleware stack. Our LangGraph orchestrator routes between sub-agents through a layered middleware stack: file upload handling, source resolution, local context, workspace sync, state hydration, aggregation barriers, and typed handoff contracts. You will extend this stack, enforce its contracts in code, and keep it operational as routing logic and agent surfaces evolve.


Maintain typed contracts and boundaries. Agent handoffs at Terrabase carry typed contracts with barrier conditions and retry predicates. You will design these contracts, enforce them with strict typing, manage backward compatibility when contracts change, and write the contract tests that prevent silent regressions.


Own the eval suites. We run structured eval suites across routing decisions, context-resolution accuracy, multi-turn coherence, visual reference alignment, and artifact correctness. You will extend coverage, write new evals where gaps exist, and build CI gates that block releases when regressions are detected. A routing change or prompt change with no eval coverage does not ship.


Triage production failures and close the loop. When an agent turn fails in production, you will trace it in LangSmith, identify the failure class, and convert it into a durable regression test. You will own the release gates, keep prompts and runtime contracts in sync, manage feature flag rollout risk, and remove dead paths as the system evolves.


Own SQL and artifact correctness. Our agents generate SQL over customer schemas and produce structured artifacts (reports, dashboards, data sheets) under a strict schema contract. You will own the correctness layer: source grounding, schema-aware validation, provenance surfaces, and the eval infrastructure that catches generated artifact failures before they reach customers.


Build and maintain HITL workflows. Human-in-the-loop checkpoints let users intervene, redirect, or approve mid-chain. You will design these workflows, enforce their resumable state contracts, and ensure they degrade gracefully when interrupted.


Instrument for traceability. You will extend LangSmith tracing coverage, add structured span annotations, and build the tooling that lets us diagnose a bad agent turn from production trace data alone, without requiring a local reproduction.


What We Are Looking For


  • 5+ years production software engineering, with strong Python fundamentals
  • 2+ years working hands-on with LLM-based systems: agent loops, tool use, context management, or inference pipelines
  • Experience with LangGraph, LangChain, OpenAI/Anthropic tool-use systems, or equivalent multi-step agent/runtime orchestration
  • Practical eval engineering: you have built or extended eval harnesses, written automated test cases for agent behavior, and treated evaluation as an ongoing engineering discipline
  • Strong engineering hygiene: strict typing, small interfaces, contract tests, clear schema migrations, and CI discipline
  • Ability to debug from production traces and artifacts, not only local reproductions
  • Comfort working across prompts, Python runtime code, TypeScript product surfaces, data systems, and eval infrastructure
  • Systems thinking: you design for observability, recovery, and state management, not just the happy path
  • Maintenance ownership mindset: you triage, close loops, and leave systems more debuggable than you found them
  • Pragmatic judgment: you can distinguish between reliability-critical infrastructure and speculative abstraction

Bonus Points

  • HITL workflow design: checkpoints, approvals, mid-chain interrupts, resumable state
  • Context engineering depth: chunking strategies, retrieval-augmented generation, semantic routing, re-ranking
  • Experience with LangSmith, Weights and Biases, or similar trace and evaluation platforms
  • Prior work shipping agent systems to enterprise customers where SQL or data correctness is a hard requirement
  • Experience with mypy, Pydantic contracts, or strict typing disciplines in a production Python codebase


Life at Terrabase

We are a sharp, focused, fully remote team building agent infrastructure that enterprise customers trust with their data. You will work directly alongside the engineer who designed this harness, with broad ownership, generous compute budgets, and a culture that treats reliability as a product requirement, not a research topic.


Terrabase is an equal-opportunity employer. We celebrate diversity and are committed to building an inclusive environment for every team member.

Read more
Vertexcover Labs
at Vertexcover Labs
1 recruiter
Ritesh Kadmawala
Posted by Ritesh Kadmawala
Remote only
1 - 6 yrs
₹12L - ₹32L / yr
Generative AI
skill iconPython
skill iconGo Programming (Golang)
skill iconRust
TypeScript
+3 more

AI Engineer — Vertexcover Labs


Who We Are

Vertexcover Labs is an employee-focused, engineer-run software studio. We partner with fast-growing, funded startups around the world—Rephrase.ai, Dhiwise, Dunzo, Dubdub, Xapo, Arintra etc.—to crack their toughest engineering problems. Everyone is an individual contributor; no management layers. Engineers choose the projects they work on, see each project's P&L, and share directly in the profits.

Whether it's building multi-agent systems for production, scaling ML pipelines across GPU clusters, or shipping RAG-powered products that end-users rely on—we solve hard AI problems for real companies.


A Few Problems We Are Currently Working On

  • AI Agents Test-authoring agents that convert natural language into e2e tests.
  • Ad-performance agents that learn what works for your brand and generate winning creatives automatically.
  • Scale + MLOps Optimise ML pipelines and fix autoscaling by applying queuing theory while juggling CPU / GPU memory contention.
  • RAG & Knowledge Systems Design and deploy retrieval-augmented generation pipelines—chunking strategies, embedding models, reranking, and evaluation loops.
  • AI Video & Image Processing Build rendering pipelines, diffusion-model integrations, and real-time video processing at scale. You'll own at least one project like these—design, build, iterate.


Signals You're Probably the Right Fit

  • Strong in at least one other language (Python, Go, Rust, TypeScript, Java …); happy to learn more.
  • First principle understanding of how LLMs, embeddings, vector databases and AI Agents work
  • Evidence of shipping real AI-powered software—OSS, side projects, or production features.
  • Comfortable navigating the fast-moving AI landscape—papers, new model releases, evolving APIs.
  • Clear written & spoken communication; async collaboration is our default.
  • Self-directed—you ask for context, not permission.


How We Operate

  • Project choice. Engineers vote on which engagements we take.
  • Stack agnostic. We pick tools that fit the job, not the résumé.
  • Pragmatic craftsmanship. Durable design, no gold-plating.
  • Transparent economics. Know what your work is worth, share in the profits.
  • Remote-native. Async by default; sync when it helps.


Hiring Process (Lean & Human)

  1. 2–3 technical deep-dives with future teammates.
  2. 30-minute culture chat.
  3. Offer. No LeetCode marathons, no trick puzzles.


We read every application and reply to all candidates.

Read more
Techjays
at Techjays
4 candid answers
1 product
SREEHARIVASU S
Posted by SREEHARIVASU S
Coimbatore
6 - 10 yrs
Best in industry
Retrieval Augmented Generation (RAG)
skill iconPython
Generative AI
Agentic AI
Data Structures
+10 more

About Techjays

At Techjays, we build production-grade AI platforms for global clients. We operate at the intersection of backend engineering, distributed systems, and applied AI — delivering secure, scalable, and enterprise-ready intelligent systems. Our team has built and scaled products at Google, Akamai, NetApp, ADP, Cognizant, and Capgemini.

About the Role

This is not a feature-delivery role. We are looking for an AI Lead who can architect, own, and scale intelligent backend systems end-to-end. You will drive both technical direction and execution — working across LLM integrations, RAG pipelines, agentic AI workflows, and cloud-native backend systems for global clients.

What You'll Do

  • Architect and scale backend systems powering AI-driven applications
  • Design and implement RAG pipelines, AI agents, and LLM integrations
  • Own systems end-to-end — from architecture to deployment and scaling
  • Integrate and optimize LLMs (Claude, GPT, Gemini) for real-world production use cases
  • Build high-performance distributed systems with observability and cost efficiency
  • Lead backend and AI initiatives with strong technical ownership
  • Mentor engineers and raise the technical bar across teams
  • Collaborate with product and AI teams to deliver AI-native solutions

What We're Looking For

  • 6–10 years of strong backend engineering experience
  • Hands-on expertise in Python (FastAPI / Django / Flask)
  • Deep understanding of Generative AI and LLM-based systems
  • Strong experience with RAG pipelines and Vector Databases (Pinecone, FAISS, ChromaDB, Weaviate)
  • Solid knowledge of Agentic AI — building autonomous agents and multi-agent workflows
  • Proficiency in AWS or GCP in production environments
  • Experience with distributed systems, microservices, and system design
  • Strong grasp of Data Structures, Algorithms, and Design Patterns
  • Familiarity with WebSockets, Git, Linux/Unix, and CI/CD

Nice to Have

  • Experience with Anthropic Claude API and Claude Code
  • Familiarity with real-time data systems or streaming (Kafka, etc.)
  • MLOps and AI system lifecycle experience
  • Optimizing AI systems for latency, cost, and scalability

Who You Are

  • You think in systems, not just features
  • You take full ownership of what you build
  • You are comfortable navigating fast-moving, ambiguous environments
  • You stay updated with the latest in Generative AI and backend technologies
  • Strong communicator who can collaborate across teams and global clients

What We Offer

  • Competitive compensation (Best in Industry)
  • Work on production-grade AI systems used by global clients
  • Exposure to cutting-edge AI tools and frameworks
  • A culture that values clarity, integrity, and continuous growth
Read more
Wissen Technology
at Wissen Technology
4 recruiters
Janane Mohanasankaran
Posted by Janane Mohanasankaran
Pune
7 - 13 yrs
Best in industry
skill iconPython
skill iconDjango
RESTful APIs
Microservices
Generative AI
+2 more

7+ years of experience in Python Development

Good experience in Microservices and APIs development.

Must have exposure to large scale data

Good to have Gen AI experience

Code versioning and collaboration. (Git)

Knowledge for Libraries for extracting data from websites.

Knowledge of SQL and NoSQL databases

Familiarity with RESTful APIs

Familiarity with Cloud (Azure /AWS) technologies


About Wissen Technology:


• The Wissen Group was founded in the year 2000. Wissen Technology, a part of Wissen Group, was established in the year 2015.

• Wissen Technology is a specialized technology company that delivers high-end consulting for organizations in the Banking & Finance, Telecom, and Healthcare domains. We help clients build world class products.

• Our workforce consists of 550+ highly skilled professionals, with leadership and senior management executives who have graduated from Ivy League Universities like Wharton, MIT, IITs, IIMs, and NITs and with rich work experience in some of the biggest companies in the world.

• Wissen Technology has grown its revenues by 400% in these five years without any external funding or investments.

• Globally present with offices US, India, UK, Australia, Mexico, and Canada.

• We offer an array of services including Application Development, Artificial Intelligence & Machine Learning, Big Data & Analytics, Visualization & Business Intelligence, Robotic Process Automation, Cloud, Mobility, Agile & DevOps, Quality Assurance & Test Automation.

• Wissen Technology has been certified as a Great Place to Work®.

• Wissen Technology has been voted as the Top 20 AI/ML vendor by CIO Insider in 2020.

• Over the years, Wissen Group has successfully delivered $650 million worth of projects for more than 20 of the Fortune 500 companies.

We have served client across sectors like Banking, Telecom, Healthcare, Manufacturing, and Energy. They include likes of Morgan Stanley, MSCI, StateStreet, Flipkart, Swiggy, Trafigura, GE to name a few.

Website : www.wissen.com


Read more
JK Technosoft Ltd
Akanksh Gupta
Posted by Akanksh Gupta
Bengaluru (Bangalore), Delhi, Gurugram, Noida, Ghaziabad, Faridabad
6 - 10 yrs
₹30L - ₹42L / yr
Generative AI
GenAI
skill iconPython
skill iconFlask
FastAPI
+3 more

We are looking for a Technical Lead - GenAI with a strong foundation in Python, Data Analytics, Data Science or Data Engineering, system design, and practical experience in building and deploying Agentic Generative AI systems. The ideal candidate is passionate about solving complex problems using LLMs, understands the architecture of modern AI agent frameworks like LangChain/LangGraph, and can deliver scalable, cloud-native back-end services with a GenAI focus.


Key Responsibilities :


- Design and implement robust, scalable back-end systems for GenAI agent-based platforms.


- Work closely with AI researchers and front-end teams to integrate LLMs and agentic workflows into production services.


- Develop and maintain services using Python (FastAPI/Django/Flask), with best practices in modularity and performance.


- Leverage and extend frameworks like LangChain, LangGraph, and similar to orchestrate tool-augmented AI agents.


- Design and deploy systems in Azure Cloud, including usage of serverless functions, Kubernetes, and scalable data services.


- Build and maintain event-driven / streaming architectures using Kafka, Event Hubs, or other messaging frameworks.


- Implement inter-service communication using gRPC and REST.


- Contribute to architectural discussions, especially around distributed systems, data flow, and fault tolerance.


Required Skills & Qualifications :


- Strong hands-on back-end development experience in Python along with Data Analytics or Data Science.


- Strong track record on platforms like LeetCode or in real-world algorithmic/system problem-solving.


- Deep knowledge of at least one Python web framework (e.g., FastAPI, Flask, Django).


- Solid understanding of LangChain, LangGraph, or equivalent LLM agent orchestration tools.


- 2+ years of hands-on experience in Generative AI systems and LLM-based platforms.


- Proven experience with system architecture, distributed systems, and microservices.


- Strong familiarity with Any Cloud infrastructure and deployment practices.


- Should know about any Data Engineering or Analytics expertise (Preferred) e.g. Azure Data Factory, Snowflake, Databricks, ETL tools Talend, Informatica or Power BI, Tableau, Data modelling, Datawarehouse development.


Read more
Remote only
5 - 9 yrs
₹20L - ₹27L / yr
skill iconPython
FastAPI
Agentic AI
skill iconNextJs (Next.js)
skill iconReact.js

Key Responsibilities


Platform Build & Architecture

• Refactor an existing Python-based prototype into a modular, production-grade platform

• Define clear service boundaries (API layer, orchestration, agent runtime, data access)

• Build reusable components that allow extension without exposing core engine logic Agent Framework & Orchestration

• Design and implement frameworks for AI agents and reporting workflows

• Build orchestration for multi-step execution (deterministic + AI-driven)

• Ensure outputs are traceable, auditable, and suitable for financial reporting Developer Enablement

• Enable internal/client developers to: o build and deploy reporting agents o reuse approved components o access platform capabilities via APIs (without direct code access)

• Implement access controls and abstraction layers Full Stack Development

• Lead development across: o Backend: Python, FastAPI o Frontend: Next.js o Real-time: WebSockets

• Build simple internal interfaces for: o job execution o monitoring o output review Code Governance & DevOps

Own development workflows in Azure DevOps • branching strategy, PRs, code reviews, merges • Set up CI/CD pipelines, environments (dev/test/prod), and release processes • Ensure code quality, testing, and maintainability standards Team Leadership (Near-term) • Act as the technical anchor o shore • Mentor and guide future hires as the team scales • Establish best practices across code, documentation, and delivery


Required Skills

• Strong experience in Python backend development (FastAPI or similar)

• Experience with React / Next.js

• Familiarity with WebSockets or real-time systems

• Experience building APIs and scalable backend systems

• Hands-on experience with Azure DevOps (repos, pipelines, PR workflows)

• Understanding of modular architecture, access control, and system design • Ability to operate in an early-stage, fast-evolving environment

Read more
Noodle.ai
at Noodle.ai
2 recruiters
Ankita Ghosh
Posted by Ankita Ghosh
Remote only
8 - 15 yrs
₹20L - ₹70L / yr
TensorFlow
pandas
skill iconPython
Artificial Intelligence (AI)
skill iconMachine Learning (ML)
+2 more

Must have:

  • 8+ years of experience with a significant focus on developing, deploying & supporting AI solutions in production environments.
  • Proven experience in building enterprise software products for B2B businesses, particularly in the supply chain domain.
  • Good understanding of Generics, OOPs concepts & Design Patterns
  • Solid engineering and coding skills. Ability to write high-performance production quality code in Python
  • Proficiency with ML libraries and frameworks (e.g., Pandas, TensorFlow, PyTorch, scikit-learn).
  • Strong expertise in time series forecasting using stat, ML, DL and foundation models
  • Experience of working on processing time series data employing techniques such as decomposition, clustering, outlier detection & treatment
  • Exposure to generative AI models and agent architectures on platforms such as AWS Bedrock, Crew AI, Mosaic/Databricks, Azure
  • Experience of working with modern data architectures, including data lakes and data warehouses, having leveraged one or more of the frameworks such as Airbyte, Airflow, Dagster, AWS Glue, Snowflake,, DBT
  • Hands-on experience with cloud platforms (e.g., AWS, Azure, GCP) and deploying ML models in cloud environments.
  • Excellent problem-solving skills and the ability to work independently as well as in a collaborative team environment.
  • Effective communication skills, with the ability to convey complex technical concepts to non-technical stakeholders


Good To Have:

  • Experience with MLOps tools and practices for continuous integration and deployment of ML models.
  • Has familiarity with deploying applications on Kubernetes
  • Knowledge of supply chain management principles and challenges.
  • A Master's or Ph.D. in Computer Science, Machine Learning, Data Science, or a related field is preferred
Read more
Why apply to jobs via Cutshort
people_solving_puzzle
Personalized job matches
Stop wasting time. Get matched with jobs that meet your skills, aspirations and preferences.
people_verifying_people
Verified hiring teams
See actual hiring teams, find common social connections or connect with them directly.
ai_chip
Move faster with AI
We use AI to get you faster responses, recommendations and unmatched user experience.
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos