Cutshort logo
For Employers
Fortune 200 MNC  logo
Software Engineer
Fortune 200 MNC

Software Engineer at Fortune 200 MNC · Bengaluru (Bangalore) · 11 - 14 years · Posted 23 Sep 2026

Bean HR Consulting's logo

Software Engineer

at Fortune 200 MNC

Agency job
11 - 14 yrs
Best in industry
Bengaluru (Bangalore)
Skills
Computer Vision
skill iconMachine Learning (ML)
YOLO
Object Oriented Programming (OOPs)
Image Processing

Hi,

Greetings !!

We/re are looking for someone who has Hands-on experience with CV/ML

The location for the same is Bangalore.


Requirements

  • 11–14 years total experience
  • Computer Vision – strong hands-on experience
  • Object Detection – YOLO(Preferred), Faster R-CNN, SSD, etc.
  • Image Processing – OpenCV, image enhancement, segmentation, feature extraction
  • Machine Learning / Deep Learning – CNNs, model training, evaluation, optimization
  • AI/ML – production-level AI solution development
  • LLM / GenAI – practical exposure to LLMs, multimodal AI, RAG, VLMs, or GenAI
  • Python – strong programming skills
  • Model deployment – preferably TensorRT, ONNX, Docker, Kubernetes, cloud, or edge deployment
  • Bangalore – candidate should be based in / willing to work from Bangalore


Preferred

  • Vision Transformers / ViT
  • YOLOv8/YOLOv9/YOLOv10/YOLO11
  • PyTorch / TensorFlow
  • NLP / LLM / VLM
  • Generative AI
  • CUDA / GPU optimization
  • Edge AI / NVIDIA
  • Experience leading CV/AI projects or teams


If interested, Share CV at: snigdhaattheratebeanhr.com

Read more
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos

Similar jobs (10)

company logo
Bengaluru (Bangalore), Mumbai, Hyderabad, Delhi, Gurugram
5 - 12 yrs
₹20L - ₹40L / yr
Computer Vision
OpenCV

AuxoAI is hiring a Senior Applied AI Engineer to design and deploy production-grade computer vision systems that operate reliably in real-world environments.

This role focuses on building end-to-end visual intelligence systems, combining deep learning, classical computer vision techniques, and multimodal models. It is not limited to model training and requires strong ownership of system design, deployment, and real-world performance.

You will work on systems that perform perception, understanding, and reasoning over visual data, and integrate these capabilities into larger AI platforms and agent-based workflows.

You will also work on problems where existing approaches may not be sufficient, and will be expected to combine deep learning, geometric methods, and multimodal reasoning to build robust, production-grade systems.

Location – Mumbai / Bangalore / Hyderabad / Gurgaon (Hybrid – 3 days per week in office)


Responsibilities:

  • Design and deploy computer vision systems for tasks such as:
  • Object detection, segmentation, and tracking
  • Scene understanding and structured perception
  • Video understanding and temporal reasoning
  • Build and optimize models using architectures such as:
  • CNNs (ResNet, EfficientNet)
  • Vision Transformers (ViT, Swin, DeiT)
  • Detection/segmentation models (YOLO, DETR, Mask R-CNN)
  • Develop multimodal systems combining vision and language:
  • CLIP-style models
  • Vision-language models (VLMs)
  • Visual grounding and captioning systems
  • Implement algorithms for:
  • Multi-object tracking (SORT, DeepSORT, ByteTrack)
  • Feature matching and representation learning
  • Temporal modeling (RNNs, Transformers for video)
  • Apply geometric and classical computer vision methods where relevant:
  • Camera calibration
  • Epipolar geometry
  • Pose estimation
  • 3D reconstruction or depth estimation
  • Optimize systems for:
  • Low-latency, real-time inference
  • Throughput and scalability
  • Edge and distributed deployment
  • Design and build data pipelines for:
  • Annotation workflows
  • Dataset curation
  • Synthetic data generation
  • Integrate vision systems into:
  • Multimodal AI pipelines
  • Agent-based systems
  • Decision-making workflows



Requirements:

  • 5+ years of experience building computer vision systems in production environments
  • Strong experience with deep learning frameworks (PyTorch / TensorFlow)
  • Hands-on experience with:
  • Detection, segmentation, or tracking systems
  • Model training, fine-tuning, and evaluation
  • Strong understanding of:
  • Representation learning
  • Loss functions (contrastive loss, focal loss, etc.)
  • Evaluation metrics (mAP, IoU, precision/recall)
  • Experience building and deploying end-to-end vision systems, not just training models


Candidates whose primary experience is limited to academic projects or model experimentation without real-world deployment may not be a fit for this role.


Nice to Have:

  • Experience with multimodal systems (vision + language)
  • Familiarity with models such as:
  • CLIP, BLIP, Flamingo, or similar
  • Experience with 3D vision:
  • NeRFs
  • SLAM
  • Point clouds
  • Experience with video understanding:
  • Action recognition
  • Event detection
  • Experience building data engines:
  • Active learning
  • Hard negative mining
  • Experience working with large-scale datasets and distributed training pipelines



Read more
company logo
Neeta Trivedi
Posted by Neeta Trivedi
Bengaluru (Bangalore)
2 - 3 yrs
₹8L - ₹15L / yr
Image Processing
Digital Signal Processing
Computer Vision
OpenCV
skill iconC++
+8 more

Position: Computer Vision Engineer

Experience: 2–3 Years

Location: Bengaluru, Karnataka

Employment Type: Full-time


About the Role

We are seeking a highly motivated Computer Vision Engineer to join our autonomy and avionics team. The role involves developing, implementing, and validating computer vision models and algorithms and pipelines for UAVs operating in both GNSS-available and GNSS-denied environments.

The ideal candidate should have a strong foundation in theory of deep learning and machine learning, strong understanding of electromagnetic spectrum, imaging fundamentals, camera principles, and mathematical concepts with hands-on experience in implementing these algorithms on embedded or real-time systems.


Key Responsibilities

  • Design, develop, and optimise AI Models
  • Make custom CNNs/ modify existing CNNs to suit specific problems at hand
  • Handle end-to-end training flow
  • Implement end to end inference pipelines on standard PCs as well as on embedded systems
  • Understand performance benchmarks and assess the accuracy and inference times
  • Implement traditional image processing algorithms
  • Factor the code to leverage underlying hardware architecture
  • Prune the networks for efficiency
  • Integrate the system within the application framework using C++
  • Work closely with perception, controls, embedded software, and systems engineering teams.


Required Qualifications

  • B.E./B.Tech/M.E./M.Tech in Computer Science and Engineering, Electronics, ECE, Mechatronics, or a related discipline.
  • 2–3 years of experience in relevant area
  • Strong understanding of: Linear Algebra, Probability and Statistics, AI-ML-DL fundamentals, Image processing, Camera Functioning
  • Strong programming skills in C++ and Python.
  • Experience with MATLAB for algorithm development and validation.
  • Familiarity with Linux development environments.
  • Experience with Git version control.


Preferred Skills

  • Experience with Camera, IMU Calibration and Synchronisation
  • Experience with multi-sensor fusion.
  • Experience working with NVIDIA devices
  • Experience on FPGA will be an added advantage
  • Full understanding of Git functionality
  • Exposure to airborne software development processes and coding standards (e.g., MISRA C++).


Personal Attributes

  • Strong analytical and problem-solving skills.
  • Ability to work independently on challenging technical problems.
  • Good communication and documentation skills.
  • Passion for solving challenging problems
  • Willingness to participate in field trials and flight testing.
  • Team playwe
Read more
company logo
Neeta Trivedi
Posted by Neeta Trivedi
Bengaluru (Bangalore)
2 - 3 yrs
₹6L - ₹12L / yr
skill iconPython
skill iconDeep Learning
skill iconMachine Learning (ML)
skill iconC++
CUDA
+11 more

AI based systems design and development, entire pipeline from image/ video ingest, metadata ingest, processing, encoding, transmitting.


Implementation and testing of advanced computer vision algorithms.

Dataset search, preparation, annotation, training, testing, fine tuning of vision CNN models. Multimodal AI, LLMs, hardware deployment, explainability.


Detailed analysis of results. Documentation, version control, client support, upgrades.

Read more
Remote only
2 - 7 yrs
₹12L - ₹24L / yr
skill iconC++
CUDA
TensorRT
Computer Vision
3D modeling
+1 more

Position Title: Real-Time Computer Vision & Edge AI Engineer (Founding Engineering Team / Core LLD) 

Reporting Structure: High-Level AI Architect (Principal ML Scientist, Google) 

Domain: Sub-16ms Edge AI, 3D Pose & Shape Estimation (SMPL-X), TensorRT C++ Inference, Zero-Copy Systems 

Performance Benchmark: Hard locked 60 FPS (<16.6 ms total frame budget) on dedicated RTX hardware


1. Position Overview & Architecture 

We are building a proprietary, ultra-low-latency spatial computing platform centered on high-fidelity 100% 3D Digital Twin architecture and real-time human digitization. 

In this role, you will serve as the Low-Level Design (LLD) Core AI Engineer, working directly alongside a Lead AI Scientist from Google. Your primary mandate is to solve complex surface occlusion and volumetric estimation challenges by building an ultra-fast C++ inference pipeline. This system must accurately regress a subject's true underlying 3D body shape and skeletal pose directly from a live camera feed. You will deploy models that extract parametric data (SMPL-X shape/pose parameters) and bridge these joint rotations seamlessly into our Vulkan graphics engine via shared GPU memory.

System Architecture: 

● Hardware Camera Ingestion: (V4L2 / GStreamer / CUDA) 

↓ Raw RGB Frames (Zero CPU Copy) 

● Edge AI Inference: (TensorRT / ONNX C++ API for 3D Pose Tracking, Kinematic Anchoring, SMPL-X Shape) 

↓ 3D Skeletal Transforms & Shape Parameters 

● Zero-Copy Shared Memory: (CUDA-Vulkan Bridge feeding directly into OpenRigLogic / MetaHuman Engine) 


2. Key Responsibilities & Deliverables 

A. Real-Time 3D Pose & Shape Estimation 

● Deploy and optimize state-of-the-art 3D human body reconstruction models (e.g., Shapy, SMPLify-X, CLIFF) to accurately regress the user's underlying skeletal structure and body volume, effectively bypassing unpredictable surface topologies and complex environmental occlusions. 

● Extract mathematically stable shape parameters (β) and pose parameters (θ) to drive the skeletal hierarchy of a high-fidelity digital avatar. 

B. Edge Inference Pipeline (TensorRT) 

● Translate Python-based research models into production-grade C++ inference engines using NVIDIA TensorRT and ONNX Runtime. 

● Implement INT8/FP16 quantization, layer fusion, and custom CUDA plugins to ensure the entire AI inference pass executes within a strict <10 ms budget per frame. 

C. Temporal Smoothing & Anti-Jitter Kinematics

● Implement highly optimized temporal filters (Kalman filters, One-Euro filters, optical flow tracking) in native C++ to eliminate all high-frequency jitter from the output joint rotations before they reach the graphics engine. 

● Ensure kinematic constraints (e.g., fixed bone lengths) are strictly maintained to prevent the digital asset from stretching or warping dynamically. 

D. Zero-Copy Ingestion & Engine Synchronization 

● Build hardware-accelerated video capture pipelines using V4L2 or GStreamer to ingest raw camera frames directly into GPU memory. 

● Bridge the output coordinate data and transformation matrices to the graphics team using POSIX shared memory and CUDA-Vulkan interop (VK_KHR_external_memory_fd), eliminating CPU staging overhead. 


3. Technical Qualifications & Tech Stack 

● Core Programming: Production-level Modern C++ (C++17/20), Python (strictly for model training/validation), and CUDA C/C++. 

● AI & Acceleration Frameworks: NVIDIA TensorRT, ONNX Runtime (C++ API), PyTorch. 

● Computer Vision Libraries: OpenCV (CUDA backend), MediaPipe C++ bindings. 

● Mathematical Foundations: 3D Kinematics, Matrix Transformations, Quaternions/Euler angles, statistical body modeling (SMPL/SMPL-X architecture). 

● Systems Architecture: Low-latency memory management, multi-threading (std::jthread, lock-free queues), SIMD vectorization. 


4. Relevant Projects & Demonstrable Experience (Preferred) 

Candidates will be preferred if they present functional codebases, GitHub repositories, or thesis work covering:

● Real-Time Body Fitting / Pose Estimation: Practical experience deploying 3D human pose or shape reconstruction models on live video feeds. 

● TensorRT / C++ Deployment: Demonstrable experience stripping a PyTorch model out of Python and running it natively in C++ using TensorRT or ONNX, ideally with custom CUDA layers or INT8 calibration. 

● High-Throughput Vision Pipelines: Built a C++ video processing pipeline that aggressively minimizes latency and avoids memory garbage collection pauses. 

● Kinematics & Smoothing: Applied mathematical filters to raw sensor or AI data to produce smooth, mechanically accurate 3D rotations. 


5. Compensation & Engagement Structure 

Compensation: ₹1,50,000 to ₹2,00,000/month 

Mentorship: Direct architectural guidance, algorithm review, and technical leadership from a Principal ML Scientist at Google. 

Hardware: Dedicated high-end workstation equipped with discrete NVIDIA RTX hardware.



Read more
company logo
Orenda Finserv
Posted by Orenda Finserv
Ahmedabad
3 - 5 yrs
₹7L - ₹11L / yr
skill iconMachine Learning (ML)
Model Serving
Vision Models
skill iconPython
RESTful APIs
+2 more

About the role

We are building AI systems that read, understand and act on real business documents, bank statements, financial reports, policy documents and forms and putting them into production where accuracy and cost both matters.

This is not a research role and it is not a prompt-writing role. You will own features end to end: pick and deploy open-source models, build the pipelines around them, measure whether they actually work on our documents, drive the cost per document down, and keep the whole thing running in production.

You will work closely with the engineering and product teams, and your work will be directly used by business users from day one.


What you will do

Deploy and evaluate open-source models

  • Select, deploy and benchmark open-source LLMs and vision-language models for specific, narrow use cases not general chat.
  • Build evaluation sets from real documents and define what "good" means numerically (field-level accuracy, extraction recall, hallucination rate) before shipping.
  • Run structured comparisons between models and approaches, and write up the trade-offs so the team can make a decision.
  • Apply quantization, batching and other optimizations to fit models into a sensible GPU budget.

Build and optimize AI orchestration

  • Design multi-step pipelines that combine deterministic code, ML models and LLM calls and know when not to use an LLM.
  • Optimize for latency, cost and reliability: caching, batching, request routing, fallback tiers, retries and graceful degradation.
  • Instrument pipelines so failures are visible and traceable rather than silent.

Ship to production

  • Package models and services with Docker, expose them behind clean APIs, and deploy them to our GPU and CPU infrastructure.
  • Handle the unglamorous production concerns: cold starts, timeouts, concurrency limits, versioning, rollback and monitoring.
  • Own on-call-style responsibility for the AI features you build, including cost tracking.


Must-have skills


Programming & engineering

  • Strong Python: type hints, async/await, dataclasses/Pydantic, clean module design, testing.
  • REST API development with FastAPI (or Flask/Django with a willingness to move to FastAPI).
  • Git, code review discipline, and the ability to write code someone else can maintain.
  • Comfortable in Linux and on the command line.

Machine learning fundamentals

  • Working knowledge of PyTorch and the Hugging Face ecosystem (transformers, tokenizers, accelerate).
  • Understanding of inference-time concepts: tokenization, context windows, batching, precision (FP16/BF16/INT8), memory footprint.
  • Ability to read a model card and a paper well enough to judge whether a model fits a use case.

Document processing

  • Hands-on experience with at least two of: pypdfium2, PyMuPDF, pdfplumber, pdfminer.six, Docling, Unstructured, Surya, DocTR, LayoutLM family.
  • Practical OCR experience (Tesseract, PaddleOCR, or a cloud OCR) and an understanding of when OCR is the wrong tool.
  • Experience extracting tables from PDFs and dealing with merged cells, multi-line rows, and inconsistent column layouts.


Strongly preferred

You will be a much stronger candidate with any of these. We do not expect all of them.

Model serving & optimization

  • vLLM, TGI, Ollama, llama.cpp, or Triton Inference Server.
  • Quantization formats and tooling: GGUF, AWQ, GPTQ, bitsandbytes, ONNX Runtime, INT8 export.
  • Serverless GPU platforms: Modal, RunPod, Replicate, Baseten including cold-start and container-lifecycle management.
  • LoRA / QLoRA fine-tuning with PEFT for narrow, task-specific improvements.

Vision-language models

  • Practical use of open VLMs: Qwen2.5-VL, InternVL, Granite Vision, Molmo, Phi-Vision, or similar.
  • Awareness of where VLMs hallucinate especially on numeric and financial content and patterns for constraining them (using the model for layout only, sourcing values from the text layer, constrained decoding).

Orchestration & pipelines

  • Workflow orchestration: Dagster, Airflow, Prefect, or Temporal.
  • Async job patterns: Celery, RQ, or platform-native spawn/poll patterns.
  • LLM orchestration frameworks (LangGraph, LlamaIndex, Haystack) with the judgement to know when plain Python is a better answer.
  • Structured output enforcement: Instructor, Outlines, XGrammar, JSON schema / tool-use modes.

Evaluation & observability

  • Building golden datasets and regression suites for extraction tasks.
  • Eval tooling: promptfoo, DeepEval, Ragas, or in-house harnesses.
  • LLM tracing and monitoring: Langfuse, Arize Phoenix, LangSmith, OpenTelemetry.

Nice extras

  • Rule engines and policy evaluation (Open Policy Agent / Rego, Drools, rule-engine).
  • Experience in fintech, lending, insurance or accounting documents.
  • Handling of PII and data-security practices in document pipelines.
  • Contributions to open-source ML or document-processing projects.


Why join us

  • Real production ownership from month one your work goes to actual users, not a demo.
  • Genuinely hard technical problems in document AI, not wrappers over an API.
  • Small team, short decision cycles, direct access to leadership.
  • Budget and freedom to evaluate and adopt new open-source models as they land.


To apply: send your CV along with a short note on one AI system you have taken to production what it did, what the accuracy was, and what broke.


Read more
company logo
Vijay Vijay V
Posted by Vijay Vijay V
Bengaluru (Bangalore)
2 - 3 yrs
₹20L - ₹25L / yr
skill iconMachine Learning (ML)
skill iconPython
PyTorch
Convolutional Neural Network (CNN)
Object Detection
+10 more

GEMBA CONCEPTS

Experience: ~3–5 years Type: Full-time

AI/ML Engineer

Location: Bengaluru, India (Hybrid)

About Gemba Concepts

Gemba Concepts is a lean manufacturing and technology consulting firm helping clients across pharma, manufacturing, and logistics

modernize how they operate. We build production systems that sit close to the shop floor — warehouse management, manufacturing

traceability, and an applied AI/ML platform whose flagship use cases are visual quality inspection and predictive maintenance. We’re a

tight engineering team that ships real systems for demanding, often regulated, environments.

The Role

We’re looking for an AI/ML Engineer to take ML capabilities from prototype to production. You’ll own models end-to-end — framing the

problem with stakeholders, building and validating the model, and deploying it as a reliable service that holds up against real-world, messy

industrial data. This is a hands-on building role, not a pure research seat: your work goes into client-facing systems.

What You’ll Do

Build and ship computer vision models for visual quality inspection (defect detection, classification, segmentation) that perform under

real factory lighting, throughput, and edge-case conditions.

Develop predictive maintenance models using sensor/time-series data — anomaly detection, remaining-useful-life estimation, failure

prediction.

Own the full ML lifecycle: data pipelines, feature engineering, training, evaluation, and deployment, with proper versioning and monitoring.

Deploy and serve models in production on Azure (AKS), and keep them healthy — track drift, retraining triggers, and latency.

Integrate LLM-based capabilities (we use the Claude API and self-hosted open models) into delivery and product workflows where they

add leverage.

Collaborate with product, engineering, and domain experts to translate fuzzy operational problems into well-scoped ML solutions — and to

know when ML is not the right answer.

Communicate results and limitations clearly to non-ML stakeholders, including clients.

What We’re Looking For

3–5 years of hands-on experience building and deploying ML models in production (not just notebooks or coursework).

Strong Python and the modern ML stack — PyTorch or TensorFlow, scikit-learn, NumPy/Pandas.

Solid grounding in at least one of: computer vision (CNNs, object detection/segmentation, image preprocessing) or time-series /

anomaly detection.

Practical MLOps experience: containerization (Docker), model serving, experiment tracking, and deploying on a cloud platform — Azure /

Kubernetes (AKS) is a strong plus.

Comfort working with imperfect, real-world data — labeling strategy, class imbalance, data drift, and validation that reflects production

reality.

Good engineering hygiene (Git, testing, code review) and the ability to write code others can build on.

Nice to Have

Experience with industrial / manufacturing data or regulated environments (pharma, 21 CFR Part 11 awareness).

Hands-on LLM integration experience — RAG, prompt engineering, working with APIs or self-hosted models (vLLM, Qwen, etc.).

Edge deployment experience (running CV models on-device / near the line).

Exposure to data pipeline tooling and orchestration.

What You’ll Get

Real ownership of ML systems that go into production for serious clients.

A lean, senior-heavy team where you ship fast and learn across the stack.

Direct exposure to applied AI in manufacturing — a domain where the work has tangible, physical impact

Read more
Pune
3 - 6 yrs
₹27L - ₹32L / yr
Artificial Intelligence (AI)
skill iconMachine Learning (ML)
skill iconPython

Strong AI Engineer / Machine Learning Engineer profiles.

2

Mandatory (Experience 1) – Must have minimum 5+ years of hands-on experience in Data Science, Machine Learning, Applied AI, NLP, Deep Learning, or Generative AI solutions.

3

Mandatory (Experience 2) – Must have strong hands-on experience in Python programming, SQL, data analysis, feature engineering, model development, and production-grade ML applications.

4

Mandatory (Experience 3) – Must have experience working with Machine Learning and Deep Learning frameworks such as PyTorch, TensorFlow, Keras, Scikit-learn, or equivalent.

5

Mandatory (Experience 4) – Must have hands-on experience working on NLP, embeddings, semantic search, text classification, document understanding, recommendation systems, or similar AI/ML use cases.

6

Mandatory (Experience 5) – Must have experience working with Large Language Models (LLMs) such as GPT, Llama, Mistral, Claude, Gemini, Phi, or similar foundation models.

7

Mandatory (Experience 6) – Must have hands-on experience building or implementing RAG (Retrieval Augmented Generation) systems, vector search, knowledge retrieval, embeddings, chunking, indexing, or semantic retrieval solutions.

8

Mandatory (Experience 7) – Must have experience working with Git, CI/CD practices, production environments, and scalable AI/ML systems.

9

Mandatory (CTC) – The CTC breakup offered will be 75% fixed + 25% variable, as per company policy.

10

Mandatory (Age) - Candidate's Age should be below 30 Years

11

Preferred (Experience 1) – Experience with MLFlow, Kubeflow, Airflow, Prefect, Feature Stores, Model Registry, or MLOps/LLMOps frameworks.

12

Preferred (Experience 2) – Experience working with Vector Databases, Spark, PySpark, distributed ML pipelines, large-scale data processing, or real-time ML systems..

13

Preferred (Experience 3) – Familiarity with Docker, Kubernetes, Azure, AWS, GCP, cloud-native AI deployments, and scalable ML architecture.

14

Preferred (Company) – Candidates from AI-first startups, Fintech, Banking, Lending, Fraud Analytics, Risk Analytics, Product Companies, SaaS organizations, or data-driven technology companies

15

Mandatory ( Pedigree) - B.TECH / M.TECH from Tier 1 Colleges (IIT's, NIT's, BITS) are Considered.

Read more
company logo
Shefali Gupta
Posted by Shefali Gupta
Remote, Delhi, Gurugram, Noida, Ghaziabad, Faridabad, Bengaluru (Bangalore)
2 - 10 yrs
₹5L - ₹15L / yr
skill iconAmazon Web Services (AWS)
Google Cloud Platform (GCP)
skill iconDocker
API
skill iconFlask
+4 more

Job Title: Senior AI/ML Engineer

Company: Timble Technologies Pvt. Ltd

Location: Gurugram (Hybrid)

Experience: 2 TO 5 Years


About Us

Timble Glance is a high-growth AI RegTech and B2B SaaS company catering to top-tier BFSI and enterprise clients. We build cutting-edge systems powering 30+ high-scale APIs for digital identity verification, fraud detection, document intelligence, and compliance automation.

Role Overview

We are looking for a hands-on Senior AI/ML Engineer to design, develop, and productionize high-throughput AI/ML and Generative AI systems. You will own the full lifecycle—from problem formulation and data pipelines to deep learning architectures, RAG systems, LLMOps, and model governance—delivering sub-second latency and high reliability across our enterprise products.


Key Responsibilities


·       Model Architecture & Deployment: Design, train, and deploy production-scale ML/Deep Learning and GenAI systems (computer vision, document intelligence, OCR, NLP, fraud risk classification, and LLM applications).

·       GenAI & LLM Solutions: Develop robust LLM workflows including prompt engineering, fine-tuning, RAG pipelines, semantic search, vector indexing (Pinecone/Milvus/Chroma), and safety guardrails.

·       Pipelines & Engineering: Build performant feature extraction and data pipelines; write modular, vectorized, production-grade Python (NumPy, Pandas) and advanced SQL.

·       MLOps & Monitoring: Establish end-to-end MLOps/LLMOps standards—model registries, CI/CD, experiment tracking, drift detection, A/B testing, latency optimization, and cost governance.

·       Responsible AI & Security: Ensure model decisions comply with enterprise data security, privacy standards, and auditability required by the BFSI sector.

·       Collaboration & Ownership: Translate complex business requirements into technical roadmaps, conduct rigorous code reviews, and mentor junior engineers.


Required Qualifications & Skills


·       Education: B.Tech / M.Tech in Computer Science, AI/ML, Mathematics, or a related field—Tier-1 institutes (IIT, IIIT, NIT) strongly preferred.

·       Experience: 2+ years of hands-on experience developing, deploying, and maintaining ML/Deep Learning or GenAI models in production environments.

·       GenAI & NLP Stack: Hands-on experience with LLMs, embeddings, RAG architectures, and frameworks such as LangChain, LlamaIndex, or Hugging Face.

·       Deep Learning Frameworks: Strong proficiency in PyTorch or TensorFlow, with deep knowledge of transformer architectures and modern NLP/CV models.

·       Software & Data Engineering: Expert-level Python skills (pytest, Git, OOP, asynchronous programming), solid SQL proficiency, and familiarity with data workflows.

·       Deployment & Cloud: Practical exposure to cloud platforms (AWS/GCP), containerization (Docker), API frameworks (FastAPI/Flask), and basic orchestration (Kubernetes).


Preferred Qualifications

·       Prior domain experience in Fintech, RegTech, Identity Verification (KYC/AML), Fraud Intelligence, or B2B SaaS.

·       Experience optimizing models for low latency and inference cost (e.g., ONNX, TensorRT, model quantization).

·       Familiarity with workflow orchestrators such as Airflow, Prefect, or Kubeflow.

Read more
Pune
3 - 6 yrs
₹27L - ₹32L / yr
Artificial Intelligence (AI)

Strong AI Engineer / Machine Learning Engineer profiles.

2

Mandatory (Experience 1) – Must have minimum 3+ years of hands-on experience in Data Science, Machine Learning, Applied AI, NLP, Deep Learning, or Generative AI solutions.

3

Mandatory (Experience 2) – Must have strong hands-on experience in Python programming, SQL, data analysis, feature engineering, model development, and production-grade ML applications.

4

Mandatory (Experience 3) – Must have experience working with Machine Learning and Deep Learning frameworks such as PyTorch, TensorFlow, Keras, Scikit-learn, or equivalent.

5

Mandatory (Experience 4) – Must have hands-on experience working on NLP, embeddings, semantic search, text classification, document understanding, recommendation systems, or similar AI/ML use cases.

6

Mandatory (Experience 5) – Must have experience working with Large Language Models (LLMs) such as GPT, Llama, Mistral, Claude, Gemini, Phi, or similar foundation models.

7

Mandatory (Experience 6) – Must have hands-on experience building or implementing RAG (Retrieval Augmented Generation) systems, vector search, knowledge retrieval, embeddings, chunking, indexing, or semantic retrieval solutions.

8

Mandatory (Experience 7) – Must have experience working with Git, CI/CD practices, production environments, and scalable AI/ML systems.

9

Mandatory (CTC) – The CTC breakup offered will be 75% fixed + 25% variable, as per company policy.

10

Mandatory (Age) - Candidate's Age should be below 28 Years

Read more
Service Co
Service Co
Agency job
via by Rishika Teja
Pune
4 - 8 yrs
₹14L - ₹18L / yr
skill iconPython
Large Language Models (LLM)
Retrieval Augmented Generation (RAG)
skill iconDocker
skill iconKubernetes
+1 more

Hiring for AI Engineer


Exp: 4 - 8 yrs

Edu : BE/B.Tech/MCA

Work Location : Pune


Skill Set

Large language,Artificial Intelligence,Machine Learning


- 4–7 years of experience in software engineering/AI roles

- Strong programming skills in Python or TypeScript (Java/Go is a plus)

- Hands-on experience with LLMs, RAG pipelines, and AI frameworks

- Experience building APIs and working with distributed systems

- Familiarity with Kubernetes, Docker, and CI/CD pipelines

- Experience with cloud platforms (AWS/Azure/GCP)

Excellent communication

Read more
Why apply to jobs via Cutshort
people_solving_puzzle
Personalized job matches
Stop wasting time. Get matched with jobs that meet your skills, aspirations and preferences.
people_verifying_people
Verified hiring teams
See actual hiring teams, find common social connections or connect with them directly.
ai_chip
Move faster with AI
We use AI to get you faster responses, recommendations and unmatched user experience.
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos