Cutshort logo
For Employers
APEAF logo
Founding Computer Vision & Edge AI Engineer- C++ / CUDA / TensorRT
Founding Computer Vision & Edge AI Engineer- C++ / CUDA / TensorRT

Founding Computer Vision & Edge AI Engineer- C++ / CUDA / TensorRT at APEAF · Remote only · 2 - 7 years · ₹12L - ₹24L / yr · Remote only · Posted 17 Sep 2026

APEAF's logo

Founding Computer Vision & Edge AI Engineer- C++ / CUDA / TensorRT

Vipra Goyal's profile picture
Posted by Vipra Goyal
2 - 7 yrs
₹12L - ₹24L / yr
Remote only
Skills
skill iconC++
CUDA
TensorRT
Computer Vision
3D modeling
ONNX

Position Title: Real-Time Computer Vision & Edge AI Engineer (Founding Engineering Team / Core LLD) 

Reporting Structure: High-Level AI Architect (Principal ML Scientist, Google) 

Domain: Sub-16ms Edge AI, 3D Pose & Shape Estimation (SMPL-X), TensorRT C++ Inference, Zero-Copy Systems 

Performance Benchmark: Hard locked 60 FPS (<16.6 ms total frame budget) on dedicated RTX hardware


1. Position Overview & Architecture 

We are building a proprietary, ultra-low-latency spatial computing platform centered on high-fidelity 100% 3D Digital Twin architecture and real-time human digitization. 

In this role, you will serve as the Low-Level Design (LLD) Core AI Engineer, working directly alongside a Lead AI Scientist from Google. Your primary mandate is to solve complex surface occlusion and volumetric estimation challenges by building an ultra-fast C++ inference pipeline. This system must accurately regress a subject's true underlying 3D body shape and skeletal pose directly from a live camera feed. You will deploy models that extract parametric data (SMPL-X shape/pose parameters) and bridge these joint rotations seamlessly into our Vulkan graphics engine via shared GPU memory.

System Architecture: 

● Hardware Camera Ingestion: (V4L2 / GStreamer / CUDA) 

↓ Raw RGB Frames (Zero CPU Copy) 

● Edge AI Inference: (TensorRT / ONNX C++ API for 3D Pose Tracking, Kinematic Anchoring, SMPL-X Shape) 

↓ 3D Skeletal Transforms & Shape Parameters 

● Zero-Copy Shared Memory: (CUDA-Vulkan Bridge feeding directly into OpenRigLogic / MetaHuman Engine) 


2. Key Responsibilities & Deliverables 

A. Real-Time 3D Pose & Shape Estimation 

● Deploy and optimize state-of-the-art 3D human body reconstruction models (e.g., Shapy, SMPLify-X, CLIFF) to accurately regress the user's underlying skeletal structure and body volume, effectively bypassing unpredictable surface topologies and complex environmental occlusions. 

● Extract mathematically stable shape parameters (β) and pose parameters (θ) to drive the skeletal hierarchy of a high-fidelity digital avatar. 

B. Edge Inference Pipeline (TensorRT) 

● Translate Python-based research models into production-grade C++ inference engines using NVIDIA TensorRT and ONNX Runtime. 

● Implement INT8/FP16 quantization, layer fusion, and custom CUDA plugins to ensure the entire AI inference pass executes within a strict <10 ms budget per frame. 

C. Temporal Smoothing & Anti-Jitter Kinematics

● Implement highly optimized temporal filters (Kalman filters, One-Euro filters, optical flow tracking) in native C++ to eliminate all high-frequency jitter from the output joint rotations before they reach the graphics engine. 

● Ensure kinematic constraints (e.g., fixed bone lengths) are strictly maintained to prevent the digital asset from stretching or warping dynamically. 

D. Zero-Copy Ingestion & Engine Synchronization 

● Build hardware-accelerated video capture pipelines using V4L2 or GStreamer to ingest raw camera frames directly into GPU memory. 

● Bridge the output coordinate data and transformation matrices to the graphics team using POSIX shared memory and CUDA-Vulkan interop (VK_KHR_external_memory_fd), eliminating CPU staging overhead. 


3. Technical Qualifications & Tech Stack 

● Core Programming: Production-level Modern C++ (C++17/20), Python (strictly for model training/validation), and CUDA C/C++. 

● AI & Acceleration Frameworks: NVIDIA TensorRT, ONNX Runtime (C++ API), PyTorch. 

● Computer Vision Libraries: OpenCV (CUDA backend), MediaPipe C++ bindings. 

● Mathematical Foundations: 3D Kinematics, Matrix Transformations, Quaternions/Euler angles, statistical body modeling (SMPL/SMPL-X architecture). 

● Systems Architecture: Low-latency memory management, multi-threading (std::jthread, lock-free queues), SIMD vectorization. 


4. Relevant Projects & Demonstrable Experience (Preferred) 

Candidates will be preferred if they present functional codebases, GitHub repositories, or thesis work covering:

● Real-Time Body Fitting / Pose Estimation: Practical experience deploying 3D human pose or shape reconstruction models on live video feeds. 

● TensorRT / C++ Deployment: Demonstrable experience stripping a PyTorch model out of Python and running it natively in C++ using TensorRT or ONNX, ideally with custom CUDA layers or INT8 calibration. 

● High-Throughput Vision Pipelines: Built a C++ video processing pipeline that aggressively minimizes latency and avoids memory garbage collection pauses. 

● Kinematics & Smoothing: Applied mathematical filters to raw sensor or AI data to produce smooth, mechanically accurate 3D rotations. 


5. Compensation & Engagement Structure 

● Compensation: ₹1,50,000 to ₹2,00,000/month 

● Mentorship: Direct architectural guidance, algorithm review, and technical leadership from a Principal ML Scientist at Google. 

● Hardware: Dedicated high-end workstation equipped with discrete NVIDIA RTX hardware.



Read more
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos

About APEAF

Founded
Type
Size
Stage

About

Facilitating 100 crore+ Indians achieve the Aatmanirbhar Bharat Dream by evolving individual actions into revolutionary national outcomes. Building Water, Energy and Employment Self-Sufficient and Self-Sustainable Model Gram Panchayat for rural India comprising of 2.53 lakh+ Gram Panchayats.
Read more

Company social profiles

blogtwitterfacebook

Similar jobs (10)

Inferigence Quotient
at Inferigence Quotient
1 recruiter
Neeta Trivedi
Posted by Neeta Trivedi
Bengaluru (Bangalore)
2 - 3 yrs
₹8L - ₹15L / yr
Image Processing
Digital Signal Processing
Computer Vision
OpenCV
skill iconC++
+8 more

Position: Computer Vision Engineer

Experience: 2–3 Years

Location: Bengaluru, Karnataka

Employment Type: Full-time


About the Role

We are seeking a highly motivated Computer Vision Engineer to join our autonomy and avionics team. The role involves developing, implementing, and validating computer vision models and algorithms and pipelines for UAVs operating in both GNSS-available and GNSS-denied environments.

The ideal candidate should have a strong foundation in theory of deep learning and machine learning, strong understanding of electromagnetic spectrum, imaging fundamentals, camera principles, and mathematical concepts with hands-on experience in implementing these algorithms on embedded or real-time systems.


Key Responsibilities

  • Design, develop, and optimise AI Models
  • Make custom CNNs/ modify existing CNNs to suit specific problems at hand
  • Handle end-to-end training flow
  • Implement end to end inference pipelines on standard PCs as well as on embedded systems
  • Understand performance benchmarks and assess the accuracy and inference times
  • Implement traditional image processing algorithms
  • Factor the code to leverage underlying hardware architecture
  • Prune the networks for efficiency
  • Integrate the system within the application framework using C++
  • Work closely with perception, controls, embedded software, and systems engineering teams.


Required Qualifications

  • B.E./B.Tech/M.E./M.Tech in Computer Science and Engineering, Electronics, ECE, Mechatronics, or a related discipline.
  • 2–3 years of experience in relevant area
  • Strong understanding of: Linear Algebra, Probability and Statistics, AI-ML-DL fundamentals, Image processing, Camera Functioning
  • Strong programming skills in C++ and Python.
  • Experience with MATLAB for algorithm development and validation.
  • Familiarity with Linux development environments.
  • Experience with Git version control.


Preferred Skills

  • Experience with Camera, IMU Calibration and Synchronisation
  • Experience with multi-sensor fusion.
  • Experience working with NVIDIA devices
  • Experience on FPGA will be an added advantage
  • Full understanding of Git functionality
  • Exposure to airborne software development processes and coding standards (e.g., MISRA C++).


Personal Attributes

  • Strong analytical and problem-solving skills.
  • Ability to work independently on challenging technical problems.
  • Good communication and documentation skills.
  • Passion for solving challenging problems
  • Willingness to participate in field trials and flight testing.
  • Team playwe
Read more
Auxo AI
Bengaluru (Bangalore), Mumbai, Hyderabad, Delhi, Gurugram
5 - 12 yrs
₹20L - ₹40L / yr
Computer Vision
OpenCV

AuxoAI is hiring a Senior Applied AI Engineer to design and deploy production-grade computer vision systems that operate reliably in real-world environments.

This role focuses on building end-to-end visual intelligence systems, combining deep learning, classical computer vision techniques, and multimodal models. It is not limited to model training and requires strong ownership of system design, deployment, and real-world performance.

You will work on systems that perform perception, understanding, and reasoning over visual data, and integrate these capabilities into larger AI platforms and agent-based workflows.

You will also work on problems where existing approaches may not be sufficient, and will be expected to combine deep learning, geometric methods, and multimodal reasoning to build robust, production-grade systems.

Location – Mumbai / Bangalore / Hyderabad / Gurgaon (Hybrid – 3 days per week in office)


Responsibilities:

  • Design and deploy computer vision systems for tasks such as:
  • Object detection, segmentation, and tracking
  • Scene understanding and structured perception
  • Video understanding and temporal reasoning
  • Build and optimize models using architectures such as:
  • CNNs (ResNet, EfficientNet)
  • Vision Transformers (ViT, Swin, DeiT)
  • Detection/segmentation models (YOLO, DETR, Mask R-CNN)
  • Develop multimodal systems combining vision and language:
  • CLIP-style models
  • Vision-language models (VLMs)
  • Visual grounding and captioning systems
  • Implement algorithms for:
  • Multi-object tracking (SORT, DeepSORT, ByteTrack)
  • Feature matching and representation learning
  • Temporal modeling (RNNs, Transformers for video)
  • Apply geometric and classical computer vision methods where relevant:
  • Camera calibration
  • Epipolar geometry
  • Pose estimation
  • 3D reconstruction or depth estimation
  • Optimize systems for:
  • Low-latency, real-time inference
  • Throughput and scalability
  • Edge and distributed deployment
  • Design and build data pipelines for:
  • Annotation workflows
  • Dataset curation
  • Synthetic data generation
  • Integrate vision systems into:
  • Multimodal AI pipelines
  • Agent-based systems
  • Decision-making workflows



Requirements:

  • 5+ years of experience building computer vision systems in production environments
  • Strong experience with deep learning frameworks (PyTorch / TensorFlow)
  • Hands-on experience with:
  • Detection, segmentation, or tracking systems
  • Model training, fine-tuning, and evaluation
  • Strong understanding of:
  • Representation learning
  • Loss functions (contrastive loss, focal loss, etc.)
  • Evaluation metrics (mAP, IoU, precision/recall)
  • Experience building and deploying end-to-end vision systems, not just training models


Candidates whose primary experience is limited to academic projects or model experimentation without real-world deployment may not be a fit for this role.


Nice to Have:

  • Experience with multimodal systems (vision + language)
  • Familiarity with models such as:
  • CLIP, BLIP, Flamingo, or similar
  • Experience with 3D vision:
  • NeRFs
  • SLAM
  • Point clouds
  • Experience with video understanding:
  • Action recognition
  • Event detection
  • Experience building data engines:
  • Active learning
  • Hard negative mining
  • Experience working with large-scale datasets and distributed training pipelines



Read more
Fortune 200 MNC
Fortune 200 MNC
Agency job
via Bean HR Consulting by snigdha Khurana
Bengaluru (Bangalore)
11 - 14 yrs
Best in industry
Computer Vision
skill iconMachine Learning (ML)
YOLO
Object Oriented Programming (OOPs)
Image Processing

Hi,

Greetings !!

We/re are looking for someone who has Hands-on experience with CV/ML

The location for the same is Bangalore.


Requirements

  • 11–14 years total experience
  • Computer Vision – strong hands-on experience
  • Object Detection – YOLO(Preferred), Faster R-CNN, SSD, etc.
  • Image Processing – OpenCV, image enhancement, segmentation, feature extraction
  • Machine Learning / Deep Learning – CNNs, model training, evaluation, optimization
  • AI/ML – production-level AI solution development
  • LLM / GenAI – practical exposure to LLMs, multimodal AI, RAG, VLMs, or GenAI
  • Python – strong programming skills
  • Model deployment – preferably TensorRT, ONNX, Docker, Kubernetes, cloud, or edge deployment
  • Bangalore – candidate should be based in / willing to work from Bangalore


Preferred

  • Vision Transformers / ViT
  • YOLOv8/YOLOv9/YOLOv10/YOLO11
  • PyTorch / TensorFlow
  • NLP / LLM / VLM
  • Generative AI
  • CUDA / GPU optimization
  • Edge AI / NVIDIA
  • Experience leading CV/AI projects or teams


If interested, Share CV at: snigdhaattheratebeanhr.com

Read more
Noida
3 - 4 yrs
₹20L - ₹25L / yr
skill iconDeep Learning
skill iconPython
PyTorch
TensorFlow
OpenCV
+2 more

About Naicos

Naicos, a fast-paced startup, builds AI-native products for algorithmic commerce: the future of how e-commerce runs. Our first products are already live with paying customers, and we are shipping new ones continuously.

Your Role

You will drive the research behind our imaging products, finding approaches to hard, unsolved problems in product and apparel imagery that work at production scale. You will run the experiments, prove what is viable, and hand a working approach to the engineering team.

Who We Are Looking For

•     Total experience: 3 years or more, with a strong research orientation

•     Deep learning frameworks in Python: PyTorch or TensorFlow

•     Image processing in Python: OpenCV, Pillow, scikit-image

•     Working knowledge of diffusion and other image generation models

We are looking for a strong research or research-student profile: someone who investigates, experiments and proves an approach, working closely with the AI Architect. Someone who is driven to build solutions, not just desk research.

AI Skills and Experience

•     Computer vision: classical CV alongside deep learning.

•     Segmentation, image-to-image translation, geometry and lighting; 

•     Generative imaging: diffusion models, conditioning and control, fine-tuning and LoRA,

•     Reads academic papers, judges what is reproducible, and turns one into a working prototype in days

Good to have

•     3D and rendering; published research or open-source contributions; model optimisation for inference cost

Research and innovative problem solving

•     Comfortable where there is no known answer, and defines the approach yourself

•     Solves problems inventively rather than reaching for the biggest model; many results come from classical image processing, fitment and geometric transformation

Other Relevant Skills and Experience

•     Designs experiments: baselines, measurable success criteria, honest reporting of negative results

•     Explains findings to a non-research audience and guides engineers to production

•     Git and reproducible experiment tracking (Weights & Biases, MLflow or similar)

Educational Qualification

•     BE / B.Tech / ME / M.Tech in Computer Science

•     BE / B.Tech / ME / M.Tech in any discipline with proven Computer Vision coursework or work

•     MSc / MS in Computer Science, Maths, Statistics or Computer Vision

•     PhD in Computer Vision or Machine Learning: an advantage, not a requirement

•     Reputed Tier 1 university preferred

Read more
New York, Los Angeles California
3 - 5 yrs
$2.5K - $5.5K / yr
skill iconPython
Artificial Intelligence (AI)
skill iconMachine Learning (ML)
Multi-Agent System
Full Stack Development
+17 more

We are building an advanced, AI-driven multi-agent software system designed to revolutionize task automation and code generation. This is a futuristic AI platform capable of:


✅ Real-time self-coding based on tasks  

✅ Autonomous multi-agent collaboration  

✅ AI-powered decision-making  

✅ Cross-platform compatibility (Desktop, Web, Mobile)  


We are hiring a highly skilled **AI Engineer & Full-Stack Developer** based in India, with a strong background in AI/ML, multi-agent architecture, and scalable, production-grade software development.


### Responsibilities:


- Build and maintain a multi-agent AI system (AutoGPT, BabyAGI, MetaGPT concepts)  

- Integrate large language models (GPT-4o, Claude, open-source LLMs)  

- Develop full-stack components (Backend: Python, FastAPI/Flask, Frontend: React/Next.js)  

- Work on real-time task execution pipelines  

- Build cross-platform apps using Electron or Flutter  

- Implement Redis, Vector databases, scalable APIs  

- Guide the architecture of autonomous, self-coding AI systems  


### Must-Have Skills:


- Python (advanced, AI applications)  

- AI/ML experience, including multi-agent orchestration  

- LLM integration knowledge  

- Full-stack development: React or Next.js  

- Redis, Vector Databases (e.g., Pinecone, FAISS)  

- Real-time applications (websockets, event-driven)  

- Cloud deployment (AWS, GCP)  


### Good to Have:


- Experience with code-generation AI models (Codex, GPT-4o coding abilities)  

- Microservices and secure system design  

- Knowledge of AI for workflow automation and productivity tools  


Join us to work on cutting-edge AI technology that builds the future of autonomous software.

Read more
Metadome.ai
at Metadome.ai
2 candid answers
Ananya  Arenavaru
Posted by Ananya Arenavaru
Bengaluru (Bangalore)
10 - 15 yrs
₹45L - ₹55L / yr
Fine-tuning LLMs
post-training SFT, RLHF, DPO
PyTorch
skill iconPython
Model deployment
+8 more

PRINCIPAL AI ENGINEER @ METADOME.AI

Company Description

Metadome.ai builds frontier AI models that transform text, drawings, and CAD into production-ready, interactive 3D experiences. The company advances a full generative pipeline—text-to-CAD, 2D-to-3D

reconstruction, CAD completion and harmonization, and real-time interactive rendering—engineered for the precision required in the physical world. Its technology currently powers the modernization of

OEM aftersales for more than 30 automotive and heavy-equipment manufacturers worldwide, delivering accurate, scalable, and fast 3D solutions. Metadome.ai’s platform enables shoppable 3D parts, step-by-step repair animations, and a headless API that feeds consistent 3D assets into commerce, dealer, training, and service systems. The broader mission is to allow anyone to move from an idea, drawing, or specification to a production-grade 3D model and beyond in seconds.


Role Description

As a Principal AI Engineer — Generative CAD & 3D, you will lead the design, development, and deployment of advanced AI models that convert text, 2D drawings, and CAD files into engineering-grade 3D content. You will architect end-to-end generative pipelines, including

text-to-CAD, 2D-to-3D reconstruction, CAD completion, and real-time rendering, collaborating closely with product, design, and engineering teams to ship robust production systems. Day-to-day, you will experiment with novel neural network architectures, optimize model performance on large-scale CAD datasets, write high-quality production code, and guide the integration of AI services into customer-facing platforms. You will mentor other engineers, establish best practices for AI development, and contribute to technical strategy and roadmap. This is a full-time, hybrid role based in Bengaluru, with a mix of on-site collaboration and work-from-home flexibility.


Qualifications

  • Strong foundation in Computer Science and Software Development, including data structures, algorithms, system design, and production-grade coding in languages such as Python, C++, or similar.
  • Deep expertise in Neural Networks and Pattern Recognition, with hands-on experience designing, training, and deploying modern deep learning architectures for complex, high-dimensional data.
  • Experience with Natural Language Processing (NLP), including working with text encoders, multimodal models, and integrating language understanding into generative workflows.
  • Advanced degree (Master’s or PhD) in Computer Science, Electrical Engineering, Applied Mathematics, or a related field, or equivalent practical experience in AI/ML research and engineering.
  • Background in 3D geometry, CAD, computer graphics, or related domains, with familiarity in 3D representations, mesh processing, and rendering pipelines.
Read more
Pune
3 - 6 yrs
₹27L - ₹32L / yr
Artificial Intelligence (AI)

Strong AI Engineer / Machine Learning Engineer profiles.

2

Mandatory (Experience 1) – Must have minimum 3+ years of hands-on experience in Data Science, Machine Learning, Applied AI, NLP, Deep Learning, or Generative AI solutions.

3

Mandatory (Experience 2) – Must have strong hands-on experience in Python programming, SQL, data analysis, feature engineering, model development, and production-grade ML applications.

4

Mandatory (Experience 3) – Must have experience working with Machine Learning and Deep Learning frameworks such as PyTorch, TensorFlow, Keras, Scikit-learn, or equivalent.

5

Mandatory (Experience 4) – Must have hands-on experience working on NLP, embeddings, semantic search, text classification, document understanding, recommendation systems, or similar AI/ML use cases.

6

Mandatory (Experience 5) – Must have experience working with Large Language Models (LLMs) such as GPT, Llama, Mistral, Claude, Gemini, Phi, or similar foundation models.

7

Mandatory (Experience 6) – Must have hands-on experience building or implementing RAG (Retrieval Augmented Generation) systems, vector search, knowledge retrieval, embeddings, chunking, indexing, or semantic retrieval solutions.

8

Mandatory (Experience 7) – Must have experience working with Git, CI/CD practices, production environments, and scalable AI/ML systems.

9

Mandatory (CTC) – The CTC breakup offered will be 75% fixed + 25% variable, as per company policy.

10

Mandatory (Age) - Candidate's Age should be below 28 Years

Read more
Bengaluru (Bangalore), Hyderabad, Pune, Chennai, Mumbai, Delhi, Gurugram, Noida, Ghaziabad, Faridabad
1 - 10 yrs
₹6L - ₹35L / yr
OpenCV
yolov7
Image Processing
Image segmentation

We are hiring a Computer Vision Engineer to build vision models that run in real products.


Responsibilities

  • Build object detection and segmentation models
  • Develop image-processing pipelines with OpenCV
  • Train and fine-tune YOLO-family models
  • Optimise models for real-time inference


Requirements

  • 1+ years in computer vision
  • Hands-on with OpenCV and YOLO or similar detectors
  • Experience deploying vision models
Read more
o9 Solutions, Inc.
icon

The recruiter has not been active on this job recently. You may apply but please expect a delayed response.

Bengaluru (Bangalore)
8 - 16 yrs
₹1L - ₹2L / yr (ESOP available)
Large Language Models (LLM)
Agentic AI
Applied mathematics

Key Responsibilities:

·      Architectural Leadership: Design and lead the development of robust, scalable AI architectures, ensuring high performance, reliability, and security.

·      Applied Mathematics &amp; Statistics: Apply statistical analysis, numerical computation, and mathematical modeling to derive insights from large-scale data and optimize model performance.

·      Deep Learning Development: Design, train, and deploy advanced Deep Learning (DL) models.

·      Technical Mentorship: Mentor engineering teams on best practices for AI/ML, coding standards, and architectural design.

·      Model Optimization: Optimize models for speed, efficiency, and accuracy using techniques like pruning, quantization, or GPU acceleration.

·      Strategy &amp; Innovation: Evaluate and select appropriate AI frameworks, tools, and platforms, staying abreast of cutting-edge research and industry trends.

Qualifications:

Required:

·      Education: Master&#39;s or PhD in Computer Science, Applied Mathematics, Statistics, Physics, or a related quantitative field.

·      Experience: 10+ years of experience in software development, with at least 3-5 years in a Applied Mathematics and Deep learning.

·      AI/ML Expertise: Proven experience designing and deploying deep learning models in production using frameworks.

·      Mathematics/Statistics: Strong proficiency in linear algebra, calculus, probability, and statistical methods.

·      Programming Skills: Expert-level coding skills in Python (NumPy, Pandas, Scikit-learn) and experience with languages like Java or C++.

Key Competencies:

  • Strategic mindset with deep operational awareness.
  • Excellent communication and stakeholder management skills.
  • Ability to simplify complex technical concepts for executive reporting.
  • Strong leadership, people development, and cross-functional influencing skills.

Bias for action and a relentless focus on continuous improvement.

Read more
Ampera Technologies
Chennai, Bengaluru (Bangalore), Delhi, Gurugram, Noida, Ghaziabad, Faridabad, Mumbai, Pune, Hyderabad, Kolkata
5 - 15 yrs
Best in industry
Retrieval Augmented Generation (RAG)
Fullstack Developer

Title                                 : Senior GenAI Engineer — RAG (Full-Stack)

Experience                    : 5+ years

Location                         : Remote

Work type                      : Chennai - Work from Office/ Remote – Other locations

Employment Type      : Full Time

Notice Period               : Immediate

Work Day                       :Mon to Fri

 

Key Responsibilities:

  • RAG pipeline end to end: ingestion integration, hybrid retrieval with reranking, prompt/context strategy, citation resolution, refusal behavior
  • Permission-aware retrieval: source ACL mapping (SharePoint/Entra, Confluence) to fail-closed retrieval filters; zero-leakage test suite partnership with QA
  • Vector database design and operations (Milvus or pgvector): schema, metadata filters, sync, performance
  • Full-stack product build: React/TypeScript chat and citation experience, Python/FastAPI services, REST APIs, SSO/OIDC integration, admin configuration UI
  • Evaluation-driven development: retrieval precision, faithfulness, and citation-accuracy metrics as the daily working loop; A/B testing of retrieval and prompt variants
  • Latency engineering to the 3–5s first-token / ~15s complete-answer targets at concurrency


Technical Skills:

  • 5+ years software engineering with 2+ years building RAG/LLM applications in production — with real users and real quality metrics, not notebooks
  • Deep retrieval craft: chunking strategy, embeddings, hybrid search, rerankers; you can explain why retrieval fails and how you measured the fix
  • Genuine full-stack evidence: shipped React/TypeScript front ends AND Python back-end services in production; API design; OIDC/SAML integration
  • Vector database production experience (Milvus, pgvector, Weaviate, or equivalent) including permission/metadata filtering
  • Evaluation fluency: has built or operated a retrieval/answer quality harness with numeric thresholds



Strongly Preferred:

  • Permission-aware/multi-tenant retrieval specifically; Microsoft Graph API; NIM/OpenAI-compatible serving endpoints; streaming UX; enterprise design systems; banking content domains

 


About Ampera: 

Ampera Technologies, a purpose driven Digital IT Services with primary focus on supporting our client with their Data, AI / ML, Accessibility and other Digital IT needs. We also ensure that equal opportunities are provided to Persons with Disabilities Talent. Ampera Technologies has its Global Headquarters in Chicago, USA and its Global Delivery Center is based out of Chennai, India. We are actively expanding our Tech Delivery team in Chennai and across India. We offer exciting benefits for our teams, such as 1) Hybrid and Remote work options available, 2) Opportunity to work directly with our Global Enterprise Clients, 3) Opportunity to learn and implement evolving Technologies, 4) Comprehensive healthcare, and 5) Conducive environment for Persons with Disability Talent meeting Physical and Digital Accessibility standards 

Read more
Why apply to jobs via Cutshort
people_solving_puzzle
Personalized job matches
Stop wasting time. Get matched with jobs that meet your skills, aspirations and preferences.
people_verifying_people
Verified hiring teams
See actual hiring teams, find common social connections or connect with them directly.
ai_chip
Move faster with AI
We use AI to get you faster responses, recommendations and unmatched user experience.
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos