Cutshort logo
For Employers
Talent Pro  logo
AI Runtime Lead (LLM DevOps pytorch
AI Runtime Lead (LLM DevOps pytorch

AI Runtime Lead (LLM DevOps pytorch at Talent Pro · Bengaluru (Bangalore) · 4 - 8 years · ₹11L - ₹17L / yr · Bootstrapped · Posted 6 Jan 2026

Talent Pro 's logo

AI Runtime Lead (LLM DevOps pytorch

Mayank choudhary's profile picture
Posted by Mayank choudhary
4 - 8 yrs
₹11L - ₹17L / yr
Bengaluru (Bangalore)
Skills
Large Language Models (LLM)
skill iconPython

Mandatory (Experience 1) – Must have 4+ years of software engineering experience

Mandatory (Experience 2) – Must have proven1+ years of experience designing, building, and owning AI runtime infrastructure supporting distributed training and/or inference at scale

Mandatory (Experience 3) – Must have hands-on experience optimizing deep learning runtimes such as PyTorch, TensorFlow, etc

Mandatory (Experience 4) – Must have strong low-level performance engineering experience, including profiling, debugging, and optimizing system throughput, latency, and reliability

Mandatory (Team Lead) – Must have experience leading or mentoring a team, including technical guidance, code reviews, and delivery ownership

Mandatory (Programming Skills) – Must have strong programming skills in Python, Java, C++ , etc

Read more
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos

About Talent Pro

Founded :
2024
Type :
Services
Size
Stage :
Bootstrapped

About

N/A

Company social profiles

bloglinkedin

Similar jobs (10)

company logo
Sandeep C
Posted by Sandeep C
Bengaluru (Bangalore)
8 - 16 yrs
₹1L - ₹2L / yr (ESOP available)
Large Language Models (LLM)
Agentic AI
Applied mathematics

Key Responsibilities:

·      Architectural Leadership: Design and lead the development of robust, scalable AI architectures, ensuring high performance, reliability, and security.

·      Applied Mathematics & Statistics: Apply statistical analysis, numerical computation, and mathematical modeling to derive insights from large-scale data and optimize model performance.

·      Deep Learning Development: Design, train, and deploy advanced Deep Learning (DL) models.

·      Technical Mentorship: Mentor engineering teams on best practices for AI/ML, coding standards, and architectural design.

·      Model Optimization: Optimize models for speed, efficiency, and accuracy using techniques like pruning, quantization, or GPU acceleration.

·      Strategy & Innovation: Evaluate and select appropriate AI frameworks, tools, and platforms, staying abreast of cutting-edge research and industry trends.

Qualifications:

Required:

·      Education: Master's or PhD in Computer Science, Applied Mathematics, Statistics, Physics, or a related quantitative field.

·      Experience: 10+ years of experience in software development, with at least 3-5 years in a Applied Mathematics and Deep learning.

·      AI/ML Expertise: Proven experience designing and deploying deep learning models in production using frameworks.

·      Mathematics/Statistics: Strong proficiency in linear algebra, calculus, probability, and statistical methods.

·      Programming Skills: Expert-level coding skills in Python (NumPy, Pandas, Scikit-learn) and experience with languages like Java or C++.

Key Competencies:

  • Strategic mindset with deep operational awareness.
  • Excellent communication and stakeholder management skills.
  • Ability to simplify complex technical concepts for executive reporting.
  • Strong leadership, people development, and cross-functional influencing skills.

Bias for action and a relentless focus on continuous improvement.

Read more
Remote only
6 - 12 yrs
₹45L - ₹50L / yr
skill iconPython
skill iconReact.js
skill iconJavascript
API management
RESTful APIs

About the Role

We are seeking a hands-on Tech Lead to design, build, and integrate AI-driven systems that automate and enhance real-world business workflows. This is a high-impact role for someone who enjoys full-stack ownership — from backend AI architecture to frontend user experiences — and can align engineering decisions with measurable product outcomes.

You will begin as a strong individual contributor, independently architecting and deploying AI-powered solutions. As the product portfolio scales, you will lead a distributed team across India and Australia, acting as a System Integrator to align engineering, data, and AI contributions into cohesive production systems.

Example Project

Design and deploy a multi-agent AI system to automate critical stages of a company’s sales cycle, including:

  • Generating client proposals using historical SharePoint data and CRM insights
  • Summarizing meeting transcripts
  • Drafting follow-up communications
  • Feeding structured insights into dashboards and workflow tools

The solution will combine RAG pipelines, LLM reasoning, and React-based interfaces to deliver measurable productivity gains.

Key Responsibilities

  • Architect and implement AI workflows using LLMs, vector databases, and automation frameworks
  • Act as a System Integrator, coordinating deliverables across distributed engineering and AI teams
  • Develop frontend interfaces using React/JavaScript to enable seamless human-AI collaboration
  • Design APIs and microservices integrating AI systems with enterprise platforms (SharePoint, Teams, Databricks, Azure)
  • Drive architecture decisions balancing scalability, performance, and security
  • Collaborate with product managers, clients, and data teams to translate business use cases into production-ready systems
  • Mentor junior engineers and evolve into a broader leadership role as the team grows

Ideal Candidate Profile

Experience Requirements

  • 5+ years in full-stack development (Python backend + React/JavaScript frontend)
  • Strong experience in API and microservice integration
  • 2+ years leading technical teams and coordinating distributed engineering efforts
  • 1+ year of hands-on AI project experience (LLMs, Transformers, LangChain, OpenAI/Azure AI frameworks)
  • Prior experience in B2B SaaS environments, particularly in AI, automation, or enterprise productivity solutions

Technical Expertise

  • Designing and implementing AI workflows including RAG pipelines, vector databases, and prompt orchestration
  • Ensuring backend and AI systems are scalable, reliable, observable, and secure
  • Familiarity with enterprise integrations (SharePoint, Teams, Databricks, Azure)
  • Experience building production-grade AI systems within enterprise SaaS ecosystems




Read more
company logo
Sandli Srivastava
Posted by Sandli Srivastava
Remote only
3 - 6 yrs
Best in industry
skill iconPython
Artificial Intelligence (AI)
skill iconReact.js
TypeScript
skill iconJavascript

About Us

We believe the future of software development is AI-native — where engineers operate at a higher level of abstraction and quality remains non-negotiable. 

Incubyte is a software craft consultancy where the “how” of building software matters as much as the “what”.  

We partner with companies of all sizes, from helping enterprises build, scale, and modernize to early-stage founders bring their ideas to life. 

Our engineers operate in an AI-native development model, using AI as a collaborator across the SDLC to accelerate development while upholding the discipline of software craftsmanship. Guided by Software Craftsmanship and Extreme Programming practices, we build reliable, maintainable, and scalable systems with speed, without compromising quality. If this way of building software resonates with you, we’d like to talk. 


Our Guiding Principles 

These principles define how we work at Incubyte. They are non-negotiable. 


Relentless Pursuit of Quality with Pragmatism 

  We build high-quality systems without losing sight of delivery. 

Extreme Ownership 

  We take responsibility end-to-end for decisions, execution, and outcomes. 

Proactive Collaboration 

  We collaborate closely, challenge each other, and solve problems together. 

Active Pursuit of Mastery 

  We continuously improve our craft and raise our bar. 

Invite, Give, and Act on Feedback 

We seek, give, and act on feedback to get better every day. 

Ensuring Client Success 

We act as trusted partners and focus on real outcomes, not just output. 


Experience Level


This role is ideal for engineers with total 3+ years of experience with a proven track record of shipping complex projects successfully.

An experienced individual contributor and leader who thrives in large, complex projects with widespread impact.


What You’ll Do as a Software Craftsperson 


  • Design and build high-quality, maintainable systems using disciplined engineering practices such as TDD, continuous refactoring, and pair programming 
  • Operate in an AI-native development model, using AI as a collaborator to explore architecture and design, accelerate development, and continuously improve systems while applying strong judgment to ensure that speed never compromises quality 
  • Take end-to-end ownership of outcomes from problem understanding and system design to implementation, deployment, and operation in production 
  • Make thoughtful design decisions that balance simplicity, scalability, and long-term maintainability in real-world systems 
  • Maintain a high bar for engineering quality through rigorous testing, code reviews, and continuous feedback 
  • Investigate and resolve production issues, and implement systemic improvements to prevent recurrence 
  • Work directly with clients, navigate ambiguity, and translate business problems into well-designed technical solutions 
  • Contribute to improving team practices, tooling, and systems to raise the overall quality and effectiveness of engineering 


Requirements


What You’ll Bring 


  • 3+ years of experience building high-quality, production systems (flexible based on demonstrated capability) 
  • Strong fundamentals in software engineering, including object-oriented design, system design, and testing practices such as TDD 
  • Demonstrated ability to build simple, maintainable, and scalable systems with a focus on long-term reliability 
  • Proficiency in one or more modern technologies, Python, PHP, JavaScript, or TypeScript, with the ability to learn new technologies quickly 
  • Deep experience working with Git in collaborative environments, including managing shared codebases, conducting code reviews, and maintaining a high bar for quality 
  • Ability to operate effectively in an AI-native workflow using AI as a collaborator to explore solutions and accelerate development, while applying strong judgment to ensure correctness, quality, and maintainability 
  • Clear thinking and strong problem-solving ability, with the capacity to break down complex problems into simple, well-structured solutions 
  • A strong sense of ownership — you take responsibility for outcomes, care deeply about quality, and are not comfortable shipping work that does not meet your standards.



Benefits


Life at Incubyte 


We are a remote-first company with structured flexibility. Teams commit to shared rhythms during core hours, ensuring smooth collaboration while maintaining autonomy. Twice a year, we come together in person for a co-working sprint and once a year for a retreat - with all travel expenses covered. 

Our environment is built for crafters: pairing, refactoring, experimenting with AI, and pushing the boundaries of software excellence. We are all lifelong learners, and our work is our passion. 


Perks

  • Dedicated learning & development budget. 
  • Sponsorship for conference talks. 
  • Comprehensive medical & term insurance. 
  • Employee-friendly leave policies. 
  • Home Office fund 
  • Medical Insurance
Read more
company logo
Stuti Jain
Posted by Stuti Jain
Hyderabad
7 - 10 yrs
₹25L - ₹35L / yr
Retrieval Augmented Generation (RAG)
skill iconAmazon Web Services (AWS)

Location: Hyderabad, India. Based at the KnackLabs headquarters, with occasional travel to client locations for workshops and reviews. This role does not involve extended onsite deployments.

About the Role

You will work as an AI Architect who designs the systems behind our client engagements: AI agents, RAG systems, automation platforms, and the conventional backend systems around them.

This is a hands-on design role, not a slideware role. You will scope architectures with clients, make the hard technical decisions, defend them in review, and stay accountable for how the systems perform in production.


You will work directly with clients. Everyone at KnackLabs does. You will sit in design discussions with client engineering teams, present architecture decisions to technical and business stakeholders, and answer for the choices you make.


A full KnackLabs engineering team in Hyderabad builds with you. You own the technical design and the quality of what ships.

What you'll own

  1. Architecture - Design AI agents, RAG systems, integrations, and the scalable backend systems around them, for multiple client engagements.
  2. Technical scoping - Work directly with clients to turn a business problem into a system design, with clear trade-offs and clear reasons.
  3. Scale and reliability - Make sure what we build handles real load: data stores, queues, caching, horizontal scaling, and fault tolerance.
  4. Design reviews - Review designs and builds across engagements. Set the technical bar and hold it.
  5. Evaluation strategy - Define how we measure accuracy, safety, latency, and cost for the AI systems we ship.
  6. Guiding engineers - Raise the level of the engineers building with you, through reviews and direct pairing.
  7. Feedback to the platform - Feed what you learn across engagements back into our platform and internal tools.

What we are looking for

  1. Around 7 or more years of software engineering experience, including direct work with customers on design or delivery.
  2. Full-stack development experience with strength in backend technologies.
  3. Experience designing and building scalable applications. You understand how large-scale distributed systems work: data partitioning, queues, caching, horizontal scaling, and fault tolerance.
  4. At least 2 years of strong, hands-on AI experience with large language models in production.
  5. You build with AI coding tools like Claude Code or Codex as your default way of working. You understand Claude Skills, have written skills yourself, use them actively, and have contributed to them.
  6. Hands-on experience building retrieval-augmented generation (RAG) systems: chunking, embeddings, vector databases, retrieval, and reranking.
  7. Hands-on experience building AI agents.
  8. Strong programming skills in Python. Working knowledge of TypeScript or JavaScript.
  9. Experience with at least one cloud platform (AWS, Azure, or GCP).
  10. Clear communication. You can explain an architecture decision to an engineer and to a business leader, and defend it under questioning.
  11. High ownership and comfort with ambiguity. You can take an unclear problem and turn it into a design.

Nice to have

  1. Experience building evaluations to measure accuracy, safety, latency, and cost.
  2. Experience with observability and tracing tools such as LangSmith or Braintrust.
  3. Experience with on-premises or private cloud (VPC) deployments.
  4. Experience deploying AI systems in regulated industries such as insurance, banking, or the public sector.
  5. Experience with data engineering and pipelines.
  6. A history of side projects, open source contributions, or products you shipped end-to-end.
  7. Experience working at a consulting or professional services firm in a client-facing delivery role.

Stack and tools

  1. Languages: Python and TypeScript.
  2. Models: Claude and other frontier or open-source models, chosen to fit the customer.
  3. AI patterns: RAG, agents, prompt engineering, skills, and evaluations.
  4. Vector and retrieval: vector databases and retrieval pipelines.
  5. Cloud: AWS, Azure, or GCP, on public or private cloud.
  6. Integration: REST APIs and enterprise system connectors.


Read more
company logo
Taher Ujjainwala
Posted by Taher Ujjainwala
Pune
12 - 25 yrs
₹70L - ₹120L / yr (ESOP available)
skill iconPython
skill iconKubernetes
Google Cloud Platform (GCP)
skill iconAmazon Web Services (AWS)
Windows Azure
+15 more

About the Role

We are hiring Staff / Principal Engineers to take full, hands-on ownership of Blitzy's most critical production-grade systems and to deliver high-leverage features that materially improve customer outcomes and engineering velocity. This is the most senior individual contributor role at the company today.

This is not a Senior-plus role, an architecture-only role, or a promotion-track role. We are looking for someone who has already operated at Principal / Staff+ scope in a highly technical environment and expects to spend their time writing, reviewing, and shipping production code.

This role is 100% hands-on. Leverage comes from system ownership, execution quality, and durable technical decisions — not people management or process.


Responsibilities

  • Own mission-critical production systems end-to-end, ensuring correctness, scalability, performance, reliability, and operational excellence.
  • Design, build, and ship high-impact backend systems and features that improve product reliability, performance, and customer value.
  • Architect scalable services and cloud infrastructure using technologies such as Python, REST, gRPC, Kubernetes, and Terraform.
  • Identify and resolve complex technical bottlenecks that limit engineering quality, system performance, or organizational velocity.
  • Build and operate LLM-powered systems and validation loops that evaluate correctness, consistency, durability, and production performance.
  • Design and evolve data architectures incorporating relational, NoSQL, graph, and vector databases to support complex enterprise applications and semantic retrieval.
  • Modernize and improve complex enterprise systems while balancing reliability, maintainability, scalability, and delivery speed.
  • Set and uphold engineering quality standards through hands-on technical leadership, sound technical judgment, and ownership of long-term technical decisions.


Qualifications

  • Direct experience with Python as a primary programming language, backend frameworks, and microservices architectures.
  • Expertise in REST and gRPC, with proficiency in Node.js and JavaScript.
  • Proficiency in GCP, along with experience using at least one additional cloud platform such as AWS or Azure.
  • Advanced knowledge of Kubernetes and Terraform in production environments.
  • Experience operating highly available production systems, including monitoring, scalability, reliability, performance optimization, and operational tooling.
  • Strong knowledge of SQL and NoSQL databases, including PostgreSQL, MySQL, MongoDB, Cassandra, or DynamoDB.
  • Familiarity with graph databases such as Neo4j and vector databases or embedding infrastructure for semantic search and retrieval.
  • Hands-on experience building and operating LLM-powered systems in production, including evaluation, validation, regression testing, tracing, and failure analysis.
  • Working knowledge of LangSmith or comparable LLM observability and evaluation tools; familiarity with OpenAI, Anthropic, or similar model providers is a plus.
  • Ability to contribute across the full stack, with a strong understanding of frontend architecture and the ability to debug, design, and ship across frontend, backend, infrastructure, and AI systems.
  • Understanding of large-scale enterprise software systems, including architecture, integration, deployment, modernization, and long-term maintainability.
  • Proven track record of operating at Staff+, Principal Engineer, or equivalent level, independently driving complex technical initiatives and delivering high-impact outcomes with minimal supervision.


Blitzy is a Cambridge, MA based AI software development platform on a mission to revolutionize the software development life cycle by autonomously building custom software to unlock the next industrial revolution. We're transforming how enterprises build software, turning enterprise requirements into enterprise grade code with an agentic software development platform that can autonomously execute 80% of the quantum of software development work. We're backed by multiple tier 1 investors, and have proven success as founders of previous start-ups.


Our Culture

Who we are:

Led by two pioneering co-founders we are one of the fastest growing companies in the U.S., creating our own category of enterprise autonomous software development. We automate thousands of hours of software development for our customers, which includes strong representation within the Fortune 500.


How we work:

  • We move Blitzy Fast: Time is both our company’s and our clients’ most precious asset. We move quickly and decisively to innovate internally and deliver exceptional software externally.
  • Championship Mindset: We operate like a professional sports team. We win as a team by holding ourselves and each other to high standards, collaborating in-person, and remaining focused on the mission.
  • Passion for Invention: We’re pushing the frontier of what’s possible, requiring constant innovation and iteration.
  • We Work for the Customer: We focus on delivering outsized value to the customers we work with and expanding those relationships into deep, meaningful partnerships.
  • We believe in being ‘everyday athletes’: taking care of ourselves so we can bring our best minds to work. We promote great sleep, movement, and restorative activities for 


Blitzy is an equal opportunity employer committed to building a diverse and inclusive team. We believe different perspectives make us stronger.

Read more
icon

The recruiter has not been active on this job recently. You may apply but please expect a delayed response.

Pune
3 - 6 yrs
₹27L - ₹32L / yr
Artificial Intelligence (AI)
skill iconMachine Learning (ML)
skill iconPython

Strong AI Engineer / Machine Learning Engineer profiles.

2

Mandatory (Experience 1) – Must have minimum 5+ years of hands-on experience in Data Science, Machine Learning, Applied AI, NLP, Deep Learning, or Generative AI solutions.

3

Mandatory (Experience 2) – Must have strong hands-on experience in Python programming, SQL, data analysis, feature engineering, model development, and production-grade ML applications.

4

Mandatory (Experience 3) – Must have experience working with Machine Learning and Deep Learning frameworks such as PyTorch, TensorFlow, Keras, Scikit-learn, or equivalent.

5

Mandatory (Experience 4) – Must have hands-on experience working on NLP, embeddings, semantic search, text classification, document understanding, recommendation systems, or similar AI/ML use cases.

6

Mandatory (Experience 5) – Must have experience working with Large Language Models (LLMs) such as GPT, Llama, Mistral, Claude, Gemini, Phi, or similar foundation models.

7

Mandatory (Experience 6) – Must have hands-on experience building or implementing RAG (Retrieval Augmented Generation) systems, vector search, knowledge retrieval, embeddings, chunking, indexing, or semantic retrieval solutions.

8

Mandatory (Experience 7) – Must have experience working with Git, CI/CD practices, production environments, and scalable AI/ML systems.

9

Mandatory (CTC) – The CTC breakup offered will be 75% fixed + 25% variable, as per company policy.

10

Mandatory (Age) - Candidate's Age should be below 30 Years

11

Preferred (Experience 1) – Experience with MLFlow, Kubeflow, Airflow, Prefect, Feature Stores, Model Registry, or MLOps/LLMOps frameworks.

12

Preferred (Experience 2) – Experience working with Vector Databases, Spark, PySpark, distributed ML pipelines, large-scale data processing, or real-time ML systems..

13

Preferred (Experience 3) – Familiarity with Docker, Kubernetes, Azure, AWS, GCP, cloud-native AI deployments, and scalable ML architecture.

14

Preferred (Company) – Candidates from AI-first startups, Fintech, Banking, Lending, Fraud Analytics, Risk Analytics, Product Companies, SaaS organizations, or data-driven technology companies

15

Mandatory ( Pedigree) - B.TECH / M.TECH from Tier 1 Colleges (IIT's, NIT's, BITS) are Considered.

Read more
Pune
3 - 6 yrs
₹27L - ₹32L / yr
Artificial Intelligence (AI)

Strong AI Engineer / Machine Learning Engineer profiles.

2

Mandatory (Experience 1) – Must have minimum 3+ years of hands-on experience in Data Science, Machine Learning, Applied AI, NLP, Deep Learning, or Generative AI solutions.

3

Mandatory (Experience 2) – Must have strong hands-on experience in Python programming, SQL, data analysis, feature engineering, model development, and production-grade ML applications.

4

Mandatory (Experience 3) – Must have experience working with Machine Learning and Deep Learning frameworks such as PyTorch, TensorFlow, Keras, Scikit-learn, or equivalent.

5

Mandatory (Experience 4) – Must have hands-on experience working on NLP, embeddings, semantic search, text classification, document understanding, recommendation systems, or similar AI/ML use cases.

6

Mandatory (Experience 5) – Must have experience working with Large Language Models (LLMs) such as GPT, Llama, Mistral, Claude, Gemini, Phi, or similar foundation models.

7

Mandatory (Experience 6) – Must have hands-on experience building or implementing RAG (Retrieval Augmented Generation) systems, vector search, knowledge retrieval, embeddings, chunking, indexing, or semantic retrieval solutions.

8

Mandatory (Experience 7) – Must have experience working with Git, CI/CD practices, production environments, and scalable AI/ML systems.

9

Mandatory (CTC) – The CTC breakup offered will be 75% fixed + 25% variable, as per company policy.

10

Mandatory (Age) - Candidate's Age should be below 28 Years

Read more
company logo
Umama Sayed
Posted by Umama Sayed
Mumbai
5 - 8 yrs
Best in industry
skill iconPython
Large Language Models (LLM)
Artificial Intelligence (AI)
Prompt engineering
LangGraph
+6 more

Senior AI Engineer

Code Generation, Agent Architecture & LLM Systems

📍 Mumbai (On-site) | Full-time | 5+ years


About the Role:

Unico Connect is an AI-first technology partner that builds custom mobile, web, and AI products for clients across multiple geographies.

We are hiring a Senior AI Engineer for a dedicated client engagement focused on building an AI-powered application builder platform - a product where users describe software in plain English and the system generates, previews, and iteratively refines working code.

The mandatory requirement for this role is hands-on production experience shipping LLM-powered systems with agent architectures, with experience in code generation or developer tooling contexts a strong advantage.


The role is product-focused and deeply hands-on. You will own everything between the user's prompt and correct code landing in the project: the agentic loop, code generation pipeline, context management, evaluation suite, and model cost strategy.

You will work alongside the Senior MLOps Engineer who operationalises the infrastructure around your system, and collaborate closely with backend, frontend, and DevOps engineers.


Responsibilities:


Agent Architecture

Design and own the agentic loop for the platform - request interpretation, planning, tool-calling sequence (read file, edit file, run build, search code, install package), and stop conditions.

Make and revisit architectural decisions on single-agent vs. multi-agent designs, including planner/executor splits and dedicated build-repair sub-agents.


Code Generation Pipeline

Own the end-to-end generation flow: task classification, context gathering, planning, targeted edits, verification, and commit.

Implement diff/search-replace-based file editing with fuzzy matching and fallback strategies.

Enforce scope discipline so the agent makes minimal diffs and does not modify code it was not asked to touch.


Self-Repair Loop

Build and tune the automated repair loop that pipes compiler, lint, build, and runtime errors back to the model with retry budgets and model escalation.

This loop is the primary quality lever - the difference between 60-70% and 90%+ build success rates.


Context Management

Build file-relevance retrieval so the agent sees the right files, not the whole codebase: dependency graphs, AST/tree-sitter-based chunking, embeddings, recency signals, and hybrid retrieval.

Implement conversation summarisation and memory for long sessions, and address long-project degradation through codebase summaries and periodic consistency passes.

Own token budgeting and prompt caching strategy.


Prompt Engineering as a Discipline

Own the system prompt and per-task prompt variants (new feature, bug fix, styling change).

Maintain few-shot examples and enforce coding conventions, stack rules, and prohibited behaviours such as no hardcoded secrets and no whole-file rewrites.

Version prompts like code with changelogs and rollback capability.


Evaluation and Quality Measurement

Design and own the evaluation suite: representative test prompts run on every prompt and model change, scored on build success rate, instruction adherence, and output quality including LLM-as-judge and visual/screenshot checks where relevant.

Define regression gates that block quality-degrading changes from shipping.

Treat evals the way engineers treat automated testing: versioned, automated, and tracked over time.

This responsibility is non-negotiable at this level.


Model Strategy and Cost

Design model routing - cheap and fast models for classification and small edits, frontier models for complex generation.

Drive cost optimisation through prompt caching, diff-based edits over full-file rewrites, and tighter context selection.

Track cost per agent run and tokens per task; evaluate new model releases against the eval suite and lead migrations when results justify it.


Safety and Reliability of Agent Behaviour

Defend against prompt injection from user content and fetched web content.

Ensure secrets never appear in generated client code.

Define what the agent's tools may and may not do in collaboration with the platform team.

Contribute to output moderation and abuse-pattern awareness.


Mentorship and Engineering Standards

Run code reviews, define engineering conventions for AI work, and raise the engineering bar across the AI team.

Work closely with the Senior MLOps Engineer on handoff of eval design, prompt configurations, and model routing logic.


Requirements:


Hands-on Production Ownership of LLM-Powered Systems with Agent Architectures (Mandatory)

Must have personally shipped and operated at least one complex production AI system - agentic, multi-step, or code generation - with end-to-end ownership of architecture, evaluation, and cost.

POCs, internal demos, and tutorial-grade work do not qualify.


5+ Years of Professional Software or AI Engineering Experience

With at least 3 years focused on LLM applications, AI engineering, or production AI systems.

Candidates with strong backend backgrounds and a clear, substantive pivot into LLM systems qualify.


Strong Python Proficiency and Service Development

Production-grade Python with FastAPI or equivalent: type hints, async patterns, streaming responses, testing, and packaging.

Not notebook-only.


Depth Across LLM APIs and Agent Systems

Production experience with at least two of OpenAI, Anthropic Claude, Google Gemini, or open-weight models (vLLM, Ollama, Together).

Production experience with at least one agent framework (LangGraph, CrewAI, AutoGen, LlamaIndex Agents) or hand-rolled equivalent.

Hands-on with tool calling, structured outputs, and multi-step reasoning.


Demonstrated, Systematic Evaluation Practice - Non-Negotiable

Must have built evaluation harnesses that gate production releases, not ad-hoc testing.

Hands-on with at least one of LangSmith, Langfuse, Promptfoo, Ragas, or DeepEval.

Candidates with no systematic answer to evaluation should not be considered at senior level regardless of other strengths.


Cost Discipline for Production AI

Track record of measurable cost optimisation on production AI features.

Able to speak in specifics: cost per request, savings achieved through caching or model routing, context reduction decisions.


AWS Working Knowledge

Hands-on with EC2, S3, IAM, and Docker.

Comfort with CI/CD workflows and deploying AI services.


Awareness of LLM Security Failure Modes

Familiar with prompt injection patterns, understands that system prompt rules alone are insufficient, and has experience with output validation and content safety in production.


Nice to Have

  • Experience with AST/tree-sitter tooling, diff-based editing systems, or compiler-adjacent work
  • MCP server authoring
  • Open-source AI contributions
  • Published technical writing on LLM systems
  • Multi-modal model experience
  • Fine-tuning exposure (LoRA, QLoRA, PEFT)
Read more
a leading provider of electronic trading solutions in India. With over 1,000 clients and a presence in more than 400 cities, we have established ourselves as a trusted partner for brokerages across the nation. Our commitment to excellence is reflected in millions of active end users and our reputation for delivering the best customer service in the industry.
a leading provider of electronic trading solutions in India. With over 1,000 clients and a presence in more than 400 cities, we have established ourselves as a trusted partner for brokerages across the nation. Our commitment to excellence is reflected in millions of active end users and our reputation for delivering the best customer service in the industry.
Agency job
via by Shwetha Naik
Bengaluru (Bangalore)
10 - 16 yrs
₹35L - ₹80L / yr
skill iconPython
skill iconGo Programming (Golang)
Microservices

Required Experience: 10–15 years (with at least 3–5 years in leadership roles)


● 10–15 years of overall experience in backend engineering, with strong exposure to

Python and/or Golang.

● 3–5 years of experience managing engineering teams.

● Proven experience delivering large-scale, distributed systems in production

environments.

● Strong understanding of microservices, cloud-native architecture, and DevOps

practices.

● Hands-on background in backend engineering (able to guide teams technically, even

if not coding daily).

● Familiarity with CI/CD pipelines, observability, and performance optimization.

● Experience in financial services or high-transaction domains is a plus.

● Experience leading teams that have utilized AI-driven development practices (e.g.,

agentic coding, LLM integration) to improve productivity and innovation is a

significant advantage.


Skills

● Excellent leadership and people management abilities.

● Strong communication and stakeholder management skills.

● Ability to balance technical depth with business priorities.

● Problem-solving mindset with a focus on delivery and impact.

● Passion for building engineering culture and improving developer experience.

Read more
company logo
Remote only
5 - 8 yrs
Best in industry
skill iconPython
skill iconReact.js
Artificial Intelligence (AI)

About Us


We believe the future of software development is AI-native — where engineers operate at a higher level of abstraction and quality remains non-negotiable. 

Incubyte is a software craft consultancy where the “how” of building software matters as much as the “what”.  


We partner with companies of all sizes, from helping enterprises build, scale, and modernize to early-stage founders bring their ideas to life. 

Our engineers operate in an AI-native development model, using AI as a collaborator across the SDLC to accelerate development while upholding the discipline of software craftsmanship. Guided by Software Craftsmanship and Extreme Programming practices, we build reliable, maintainable, and scalable systems with speed, without compromising quality. If this way of building software resonates with you, we’d like to talk. 


Our Guiding Principles 


These principles define how we work at Incubyte. They are non-negotiable. 


Relentless Pursuit of Quality with Pragmatism 


  We build high-quality systems without losing sight of delivery. 


Extreme Ownership 

  We take responsibility end-to-end for decisions, execution, and outcomes. 


Proactive Collaboration 


  We collaborate closely, challenge each other, and solve problems together. 


Active Pursuit of Mastery 


  We continuously improve our craft and raise our bar. 


Invite, Give, and Act on Feedback 


We seek, give, and act on feedback to get better every day. 


Ensuring Client Success 


We act as trusted partners and focus on real outcomes, not just output. 


Job Description


This is a remote position.


Experience Level


This role is ideal for engineers with total 5+ years of experience with a proven track record of shipping complex projects successfully.

An experienced individual contributor and leader who thrives in large, complex projects with widespread impact.


What You’ll Do as a Software Craftsperson 


  • Design and build high-quality, maintainable systems using disciplined engineering practices such as TDD, continuous refactoring, and pair programming 
  • Operate in an AI-native development model, using AI as a collaborator to explore architecture and design, accelerate development, and continuously improve systems while applying strong judgment to ensure that speed never compromises quality. 
  • Take end-to-end ownership of outcomes from problem understanding and system design to implementation, deployment, and operation in production 
  • Make thoughtful design decisions that balance simplicity, scalability, and long-term maintainability in real-world systems 
  • Maintain a high bar for engineering quality through rigorous testing, code reviews, and continuous feedback 
  • Investigate and resolve production issues, and implement systemic improvements to prevent recurrence 
  • Work directly with clients, navigate ambiguity, and translate business problems into well-designed technical solutions 
  • Contribute to improving team practices, tooling, and systems to raise the overall quality and effectiveness of engineering 



Requirements


What You’ll Bring 


  • 5+ years of experience building high-quality, production systems (flexible based on demonstrated capability) 
  • Strong fundamentals in software engineering, including object-oriented design, system design, and testing practices such as TDD 
  • Demonstrated ability to build simple, maintainable, and scalable systems with a focus on long-term reliability 
  • Proficiency in one or more modern technologies Python, React, AI, JavaScript, or TypeScript, with the ability to learn new technologies quickly 
  • Deep experience working with Git in collaborative environments, including managing shared codebases, conducting code reviews, and maintaining a high bar for quality 
  • Ability to operate effectively in an AI-native workflow using AI as a collaborator to explore solutions and accelerate development, while applying strong judgment to ensure correctness, quality, and maintainability 
  • Clear thinking and strong problem-solving ability, with the capacity to break down complex problems into simple, well-structured solutions 
  • A strong sense of ownership — you take responsibility for outcomes, care deeply about quality, and are not comfortable shipping work that does not meet your standards.



Benefits


Life at Incubyte ​


We are a remote-first company with structured flexibility. Teams commit to shared rhythms during core hours, ensuring smooth collaboration while maintaining autonomy. Twice a year, we come together in person for a co-working sprint and once a year for a retreat - with all travel expenses covered. 

 

Our environment is built for crafters: pairing, refactoring, experimenting with AI, and pushing the boundaries of software excellence. We are all lifelong learners, and our work is our passion. 


Benefits 



  • Dedicated learning & development budget. 
  • Sponsorship for conference talks. 
  • Comprehensive medical & term insurance. 
  • Employee-friendly leave policies. 
  • Home Office fund 
  • Medical Insurance 
Read more
Why apply to jobs via Cutshort
people_solving_puzzle
Personalized job matches
Stop wasting time. Get matched with jobs that meet your skills, aspirations and preferences.
people_verifying_people
Verified hiring teams
See actual hiring teams, find common social connections or connect with them directly.
ai_chip
Move faster with AI
We use AI to get you faster responses, recommendations and unmatched user experience.
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos