Cutshort logo
For Employers
The world’s leading image  editing AI software, to capture logo
Computer Vision Lead
The world’s leading image editing AI software, to capture
Computer Vision Lead

Computer Vision Lead at The world’s leading image editing AI software, to capture · Gurugram · 5 - 10 years · ₹25L - ₹40L / yr · Posted 17 Aug 2022

Qrata's logo

Computer Vision Lead

at The world’s leading image editing AI software, to capture

Agency job
via Qrata
5 - 10 yrs
₹25L - ₹40L / yr
Gurugram
Skills
skill iconMachine Learning (ML)
skill iconData Science
Computer Vision
Machine vision
Artificial Intelligence (AI)
skill iconDeep Learning
GAN
About the company

 Changing the way cataloging is done across the Globe. Our vision is to empower the smallest of
sellers, situated in the farthest of corners, to create superior product images and videos, without the need
for any external professional help. Imagine 30M+ merchants shooting Product Images or Videos using
their Smartphones, and then choosing Filters for Amazon, Asos, Airbnb, Doordash, etc to instantly
compose High-Quality "tuned-in" product visuals, instantly.  
editing AI software, to capture and process beautiful product images for online selling. We are also
fortunate and proud to be backed by the biggest names in the investment community including the likes of
Accel Partners, Angellist and prominent Founders and Internet company operators, who believe that
there is an intelligent and efficient way of doing Digital Production than how the world operates currently.

Job Description :

- We are looking for a seasoned Computer Vision Engineer with AI/ML/CV and Deep Learning skills to
play a senior leadership role in our Product & Technology Research Team.
- You will be leading a team of CV researchers to build models that automatically transform millions of ecommerce, automobiles, food, real-estate ram images into processed final images.
- You will be responsible for researching the latest art of the possible in the field of computer vision,
designing the solution architecture for our offerings and lead the Computer Vision teams to build the core
algorithmic models & deploy them on Cloud Infrastructure.
- Working with the Data team to ensure your data pipelines are well set up and
models are being constantly trained and updated
- Working alongside product team to ensure that AI capabilities are built as democratized tools that
provides internal as well external stakeholders to innovate on top of it and make our customers
successful
- You will work closely with the Product & Engineering teams to convert the models into beautiful products
that will be used by thousands of Businesses everyday to transform their images and videos.

Job Requirements:

- 4-5 years of experience overall.
- Looking for strong Computer Vision experience and knowledge.
- BS/MS/ Phd degree in Computer Science, Engineering or a related subject from a ivy league institute
- Exposure on Deep Learning Techniques, TensorFlow/Pytorch
- Prior expertise on building Image processing applications using GANs, CNNs, Diffusion models
- Expertise with Image Processing Python libraries like OpenCV, etc.
- Good hands-on experience on Python, Flask or Django framework
- Authored publications at peer-reviewed AI conferences (e.g. NeurIPS, CVPR, ICML, ICLR,ICCV, ACL)
- Prior experience of managing teams and building large scale AI / CV projects is a big plus
- Great interpersonal and communication skills
- Critical thinker and problem-solving skills
Read more
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos

Similar jobs (6)

company logo
Noida
3 - 4 yrs
₹20L - ₹25L / yr
skill iconDeep Learning
skill iconPython
PyTorch
TensorFlow
OpenCV
+2 more

About Naicos

Naicos, a fast-paced startup, builds AI-native products for algorithmic commerce: the future of how e-commerce runs. Our first products are already live with paying customers, and we are shipping new ones continuously.

Your Role

You will drive the research behind our imaging products, finding approaches to hard, unsolved problems in product and apparel imagery that work at production scale. You will run the experiments, prove what is viable, and hand a working approach to the engineering team.

Who We Are Looking For

•     Total experience: 3 years or more, with a strong research orientation

•     Deep learning frameworks in Python: PyTorch or TensorFlow

•     Image processing in Python: OpenCV, Pillow, scikit-image

•     Working knowledge of diffusion and other image generation models

We are looking for a strong research or research-student profile: someone who investigates, experiments and proves an approach, working closely with the AI Architect. Someone who is driven to build solutions, not just desk research.

AI Skills and Experience

•     Computer vision: classical CV alongside deep learning.

•     Segmentation, image-to-image translation, geometry and lighting; 

•     Generative imaging: diffusion models, conditioning and control, fine-tuning and LoRA,

•     Reads academic papers, judges what is reproducible, and turns one into a working prototype in days

Good to have

•     3D and rendering; published research or open-source contributions; model optimisation for inference cost

Research and innovative problem solving

•     Comfortable where there is no known answer, and defines the approach yourself

•     Solves problems inventively rather than reaching for the biggest model; many results come from classical image processing, fitment and geometric transformation

Other Relevant Skills and Experience

•     Designs experiments: baselines, measurable success criteria, honest reporting of negative results

•     Explains findings to a non-research audience and guides engineers to production

•     Git and reproducible experiment tracking (Weights & Biases, MLflow or similar)

Educational Qualification

•     BE / B.Tech / ME / M.Tech in Computer Science

•     BE / B.Tech / ME / M.Tech in any discipline with proven Computer Vision coursework or work

•     MSc / MS in Computer Science, Maths, Statistics or Computer Vision

•     PhD in Computer Vision or Machine Learning: an advantage, not a requirement

•     Reputed Tier 1 university preferred

Read more
company logo
Neeta Trivedi
Posted by Neeta Trivedi
Bengaluru (Bangalore)
2 - 3 yrs
₹8L - ₹15L / yr
Image Processing
Digital Signal Processing
Computer Vision
OpenCV
skill iconC++
+8 more

Position: Computer Vision Engineer

Experience: 2–3 Years

Location: Bengaluru, Karnataka

Employment Type: Full-time


About the Role

We are seeking a highly motivated Computer Vision Engineer to join our autonomy and avionics team. The role involves developing, implementing, and validating computer vision models and algorithms and pipelines for UAVs operating in both GNSS-available and GNSS-denied environments.

The ideal candidate should have a strong foundation in theory of deep learning and machine learning, strong understanding of electromagnetic spectrum, imaging fundamentals, camera principles, and mathematical concepts with hands-on experience in implementing these algorithms on embedded or real-time systems.


Key Responsibilities

  • Design, develop, and optimise AI Models
  • Make custom CNNs/ modify existing CNNs to suit specific problems at hand
  • Handle end-to-end training flow
  • Implement end to end inference pipelines on standard PCs as well as on embedded systems
  • Understand performance benchmarks and assess the accuracy and inference times
  • Implement traditional image processing algorithms
  • Factor the code to leverage underlying hardware architecture
  • Prune the networks for efficiency
  • Integrate the system within the application framework using C++
  • Work closely with perception, controls, embedded software, and systems engineering teams.


Required Qualifications

  • B.E./B.Tech/M.E./M.Tech in Computer Science and Engineering, Electronics, ECE, Mechatronics, or a related discipline.
  • 2–3 years of experience in relevant area
  • Strong understanding of: Linear Algebra, Probability and Statistics, AI-ML-DL fundamentals, Image processing, Camera Functioning
  • Strong programming skills in C++ and Python.
  • Experience with MATLAB for algorithm development and validation.
  • Familiarity with Linux development environments.
  • Experience with Git version control.


Preferred Skills

  • Experience with Camera, IMU Calibration and Synchronisation
  • Experience with multi-sensor fusion.
  • Experience working with NVIDIA devices
  • Experience on FPGA will be an added advantage
  • Full understanding of Git functionality
  • Exposure to airborne software development processes and coding standards (e.g., MISRA C++).


Personal Attributes

  • Strong analytical and problem-solving skills.
  • Ability to work independently on challenging technical problems.
  • Good communication and documentation skills.
  • Passion for solving challenging problems
  • Willingness to participate in field trials and flight testing.
  • Team playwe
Read more
Fortune 200 MNC
Fortune 200 MNC
Agency job
via by snigdha Khurana
Bengaluru (Bangalore)
11 - 14 yrs
Best in industry
Computer Vision
skill iconMachine Learning (ML)
YOLO
Object Oriented Programming (OOPs)
Image Processing

Hi,

Greetings !!

We/re are looking for someone who has Hands-on experience with CV/ML

The location for the same is Bangalore.


Requirements

  • 11–14 years total experience
  • Computer Vision – strong hands-on experience
  • Object Detection – YOLO(Preferred), Faster R-CNN, SSD, etc.
  • Image Processing – OpenCV, image enhancement, segmentation, feature extraction
  • Machine Learning / Deep Learning – CNNs, model training, evaluation, optimization
  • AI/ML – production-level AI solution development
  • LLM / GenAI – practical exposure to LLMs, multimodal AI, RAG, VLMs, or GenAI
  • Python – strong programming skills
  • Model deployment – preferably TensorRT, ONNX, Docker, Kubernetes, cloud, or edge deployment
  • Bangalore – candidate should be based in / willing to work from Bangalore


Preferred

  • Vision Transformers / ViT
  • YOLOv8/YOLOv9/YOLOv10/YOLO11
  • PyTorch / TensorFlow
  • NLP / LLM / VLM
  • Generative AI
  • CUDA / GPU optimization
  • Edge AI / NVIDIA
  • Experience leading CV/AI projects or teams


If interested, Share CV at: snigdhaattheratebeanhr.com

Read more
company logo
Bengaluru (Bangalore), Mumbai, Hyderabad, Delhi, Gurugram
5 - 12 yrs
₹20L - ₹40L / yr
Computer Vision
OpenCV

AuxoAI is hiring a Senior Applied AI Engineer to design and deploy production-grade computer vision systems that operate reliably in real-world environments.

This role focuses on building end-to-end visual intelligence systems, combining deep learning, classical computer vision techniques, and multimodal models. It is not limited to model training and requires strong ownership of system design, deployment, and real-world performance.

You will work on systems that perform perception, understanding, and reasoning over visual data, and integrate these capabilities into larger AI platforms and agent-based workflows.

You will also work on problems where existing approaches may not be sufficient, and will be expected to combine deep learning, geometric methods, and multimodal reasoning to build robust, production-grade systems.

Location – Mumbai / Bangalore / Hyderabad / Gurgaon (Hybrid – 3 days per week in office)


Responsibilities:

  • Design and deploy computer vision systems for tasks such as:
  • Object detection, segmentation, and tracking
  • Scene understanding and structured perception
  • Video understanding and temporal reasoning
  • Build and optimize models using architectures such as:
  • CNNs (ResNet, EfficientNet)
  • Vision Transformers (ViT, Swin, DeiT)
  • Detection/segmentation models (YOLO, DETR, Mask R-CNN)
  • Develop multimodal systems combining vision and language:
  • CLIP-style models
  • Vision-language models (VLMs)
  • Visual grounding and captioning systems
  • Implement algorithms for:
  • Multi-object tracking (SORT, DeepSORT, ByteTrack)
  • Feature matching and representation learning
  • Temporal modeling (RNNs, Transformers for video)
  • Apply geometric and classical computer vision methods where relevant:
  • Camera calibration
  • Epipolar geometry
  • Pose estimation
  • 3D reconstruction or depth estimation
  • Optimize systems for:
  • Low-latency, real-time inference
  • Throughput and scalability
  • Edge and distributed deployment
  • Design and build data pipelines for:
  • Annotation workflows
  • Dataset curation
  • Synthetic data generation
  • Integrate vision systems into:
  • Multimodal AI pipelines
  • Agent-based systems
  • Decision-making workflows



Requirements:

  • 5+ years of experience building computer vision systems in production environments
  • Strong experience with deep learning frameworks (PyTorch / TensorFlow)
  • Hands-on experience with:
  • Detection, segmentation, or tracking systems
  • Model training, fine-tuning, and evaluation
  • Strong understanding of:
  • Representation learning
  • Loss functions (contrastive loss, focal loss, etc.)
  • Evaluation metrics (mAP, IoU, precision/recall)
  • Experience building and deploying end-to-end vision systems, not just training models


Candidates whose primary experience is limited to academic projects or model experimentation without real-world deployment may not be a fit for this role.


Nice to Have:

  • Experience with multimodal systems (vision + language)
  • Familiarity with models such as:
  • CLIP, BLIP, Flamingo, or similar
  • Experience with 3D vision:
  • NeRFs
  • SLAM
  • Point clouds
  • Experience with video understanding:
  • Action recognition
  • Event detection
  • Experience building data engines:
  • Active learning
  • Hard negative mining
  • Experience working with large-scale datasets and distributed training pipelines



Read more
company logo
Neeta Trivedi
Posted by Neeta Trivedi
Bengaluru (Bangalore)
2 - 3 yrs
₹6L - ₹12L / yr
skill iconPython
skill iconDeep Learning
skill iconMachine Learning (ML)
skill iconC++
CUDA
+11 more

AI based systems design and development, entire pipeline from image/ video ingest, metadata ingest, processing, encoding, transmitting.


Implementation and testing of advanced computer vision algorithms.

Dataset search, preparation, annotation, training, testing, fine tuning of vision CNN models. Multimodal AI, LLMs, hardware deployment, explainability.


Detailed analysis of results. Documentation, version control, client support, upgrades.

Read more
Bengaluru (Bangalore)
3 - 9 yrs
₹20L - ₹40L / yr (ESOP available)
Artificial Intelligence (AI)
Large Language Models (LLM) tuning
MULTIMODAL AI
skill iconPython
skill iconProgramming

[Please refrain from applying if you have over 10 years of experience. This is a hands-on role that requires building from the ground up.]


Location: Bengaluru (In-Office)


Employment Type: Full-Time


About Logikality


Logikality is building an AI-native mortgage intelligence platform for the U.S. mortgage industry. We are reimagining how mortgage operations are executed by combining AI, workflow automation, and domain expertise to solve one of the most document-intensive and decision-heavy industries in the world.

Our platform goes beyond document extraction. We are building AI systems that understand mortgage files, reason across multiple sources of information, identify risks and exceptions, support underwriting and quality control decisions, and continuously improve through expert feedback and rigorous evaluation.

As we expand our AI capabilities, we are looking for a Director, AI Engineering to define and drive the research direction behind our next generation of intelligent systems.

About the Role

This is a hands-on technical leadership role for someone who enjoys solving difficult AI problems and turning research into production impact.

You will lead the research agenda across large language models, reasoning systems, agentic AI, multimodal learning, and intelligent decision support while working closely with engineering, product, and mortgage domain experts. You will prototype new ideas, validate them through rigorous experimentation, and help productionize solutions that directly improve customer outcomes.

This role is ideal for someone with deep research expertise who enjoys building real-world AI systems rather than research for its own sake.


What You'll Do


  • Define and execute the Applied AI research roadmap aligned with company and product goals.
  • Design novel approaches for document understanding, reasoning, planning, retrieval, and decision support.
  • Build agentic AI systems capable of orchestrating tools, workflows, and domain knowledge to solve complex mortgage use cases.
  • Develop multimodal AI models that combine documents, structured data, images, and operational context.
  • Lead research on long-context reasoning, knowledge integration, memory, retrieval-augmented generation (RAG), and workflow automation.
  • Design robust evaluation frameworks, benchmarks, and automated testing pipelines to measure model quality, reliability, explainability, and business impact.
  • Rapidly prototype, experiment, and iterate on new AI techniques, evaluating state-of-the-art research for production adoption.
  • Work closely with software engineers to translate research prototypes into scalable, production-ready systems.
  • Mentor AI engineers and contribute to building a strong research culture within the organisation.
  • Collaborate with mortgage domain experts to deeply understand operational workflows, compliance requirements, and decision-making processes.
  • Stay current with advances in AI research and identify opportunities to leverage emerging techniques within our platform.
  • Represent Logikality in customer interactions, strategic discussions, industry conferences, and business forums, communicating our AI vision, gathering market insights, and helping shape research priorities through direct engagement with customers and ecosystem partners.


What We're Looking For


  • PhD in Computer Science, Artificial Intelligence, Machine Learning, or a related discipline; or an engineering degree in Computer Science or related disciplines from a premier engineering institution (e.g., IITs, IISc, NITs, BITS Pilani, or top-tier global universities).
  • 3–8 years of professional experience in Applied AI, Machine Learning, or AI Research, with experience building production-grade AI systems
  • Strong expertise in modern AI, including Large Language Models, transformers, agentic AI, reasoning systems, retrieval, multimodal learning, or adjacent areas.
  • Strong software engineering skills with Python and modern machine learning frameworks.
  • Experience designing and implementing production-grade AI systems that solve complex real-world problems.
  • Strong understanding of model evaluation, benchmarking, experimentation, and AI system reliability.
  • Experience balancing research innovation with engineering pragmatism and product delivery.
  • Excellent problem-solving and communication skills with the ability to collaborate across engineering, product, and business teams.


Why Join Logikality?


At Logikality, you'll work on problems that require genuine reasoning, not just text generation. You'll help build AI systems that understand complex documents, synthesise information across workflows, explain decisions, identify exceptions, and improve through continuous learning and expert feedback.


This is an opportunity to work at the intersection of cutting-edge AI research and real-world impact, where your ideas won't remain as papers or prototypes; they'll power intelligent systems used every day by mortgage professionals. We are looking for someone who can connect AI, platform engineering, product thinking and customer outcomes.

For the right person, this could develop into a CTO and co-founder track over the next 6–9 months, based on contribution, technical leadership and mutual fit.


Interested candidates are requested to apply via the Google Form given: https://forms.gle/jFqKzfLhNCcCFU5t9


This will be a full-time in-office role based in Bangalore. Immediate joiners are preferred.

Read more
Why apply to jobs via Cutshort
people_solving_puzzle
Personalized job matches
Stop wasting time. Get matched with jobs that meet your skills, aspirations and preferences.
people_verifying_people
Verified hiring teams
See actual hiring teams, find common social connections or connect with them directly.
ai_chip
Move faster with AI
We use AI to get you faster responses, recommendations and unmatched user experience.
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos