Computer Vision Lead at The world’s leading image editing AI software, to capture · Gurugram · 5 - 10 years · ₹25L - ₹40L / yr · Posted 17 Aug 2022

Changing the way cataloging is done across the Globe. Our vision is to empower the smallest of
sellers, situated in the farthest of corners, to create superior product images and videos, without the need
for any external professional help. Imagine 30M+ merchants shooting Product Images or Videos using
their Smartphones, and then choosing Filters for Amazon, Asos, Airbnb, Doordash, etc to instantly
compose High-Quality "tuned-in" product visuals, instantly.
editing AI software, to capture and process beautiful product images for online selling. We are also
fortunate and proud to be backed by the biggest names in the investment community including the likes of
Accel Partners, Angellist and prominent Founders and Internet company operators, who believe that
there is an intelligent and efficient way of doing Digital Production than how the world operates currently.
Job Description :
- We are looking for a seasoned Computer Vision Engineer with AI/ML/CV and Deep Learning skills to
play a senior leadership role in our Product & Technology Research Team.
- You will be leading a team of CV researchers to build models that automatically transform millions of ecommerce, automobiles, food, real-estate ram images into processed final images.
- You will be responsible for researching the latest art of the possible in the field of computer vision,
designing the solution architecture for our offerings and lead the Computer Vision teams to build the core
algorithmic models & deploy them on Cloud Infrastructure.
- Working with the Data team to ensure your data pipelines are well set up and
models are being constantly trained and updated
- Working alongside product team to ensure that AI capabilities are built as democratized tools that
provides internal as well external stakeholders to innovate on top of it and make our customers
successful
- You will work closely with the Product & Engineering teams to convert the models into beautiful products
that will be used by thousands of Businesses everyday to transform their images and videos.
Job Requirements:
- 4-5 years of experience overall.
- Looking for strong Computer Vision experience and knowledge.
- BS/MS/ Phd degree in Computer Science, Engineering or a related subject from a ivy league institute
- Exposure on Deep Learning Techniques, TensorFlow/Pytorch
- Prior expertise on building Image processing applications using GANs, CNNs, Diffusion models
- Expertise with Image Processing Python libraries like OpenCV, etc.
- Good hands-on experience on Python, Flask or Django framework
- Authored publications at peer-reviewed AI conferences (e.g. NeurIPS, CVPR, ICML, ICLR,ICCV, ACL)
- Prior experience of managing teams and building large scale AI / CV projects is a big plus
- Great interpersonal and communication skills
- Critical thinker and problem-solving skills

Similar jobs (6)
About Naicos
Naicos, a fast-paced startup, builds AI-native products for algorithmic commerce: the future of how e-commerce runs. Our first products are already live with paying customers, and we are shipping new ones continuously.
Your Role
You will drive the research behind our imaging products, finding approaches to hard, unsolved problems in product and apparel imagery that work at production scale. You will run the experiments, prove what is viable, and hand a working approach to the engineering team.
Who We Are Looking For
• Total experience: 3 years or more, with a strong research orientation
• Deep learning frameworks in Python: PyTorch or TensorFlow
• Image processing in Python: OpenCV, Pillow, scikit-image
• Working knowledge of diffusion and other image generation models
We are looking for a strong research or research-student profile: someone who investigates, experiments and proves an approach, working closely with the AI Architect. Someone who is driven to build solutions, not just desk research.
AI Skills and Experience
• Computer vision: classical CV alongside deep learning.
• Segmentation, image-to-image translation, geometry and lighting;
• Generative imaging: diffusion models, conditioning and control, fine-tuning and LoRA,
• Reads academic papers, judges what is reproducible, and turns one into a working prototype in days
Good to have
• 3D and rendering; published research or open-source contributions; model optimisation for inference cost
Research and innovative problem solving
• Comfortable where there is no known answer, and defines the approach yourself
• Solves problems inventively rather than reaching for the biggest model; many results come from classical image processing, fitment and geometric transformation
Other Relevant Skills and Experience
• Designs experiments: baselines, measurable success criteria, honest reporting of negative results
• Explains findings to a non-research audience and guides engineers to production
• Git and reproducible experiment tracking (Weights & Biases, MLflow or similar)
Educational Qualification
• BE / B.Tech / ME / M.Tech in Computer Science
• BE / B.Tech / ME / M.Tech in any discipline with proven Computer Vision coursework or work
• MSc / MS in Computer Science, Maths, Statistics or Computer Vision
• PhD in Computer Vision or Machine Learning: an advantage, not a requirement
• Reputed Tier 1 university preferred
Position: Computer Vision Engineer
Experience: 2–3 Years
Location: Bengaluru, Karnataka
Employment Type: Full-time
About the Role
We are seeking a highly motivated Computer Vision Engineer to join our autonomy and avionics team. The role involves developing, implementing, and validating computer vision models and algorithms and pipelines for UAVs operating in both GNSS-available and GNSS-denied environments.
The ideal candidate should have a strong foundation in theory of deep learning and machine learning, strong understanding of electromagnetic spectrum, imaging fundamentals, camera principles, and mathematical concepts with hands-on experience in implementing these algorithms on embedded or real-time systems.
Key Responsibilities
- Design, develop, and optimise AI Models
- Make custom CNNs/ modify existing CNNs to suit specific problems at hand
- Handle end-to-end training flow
- Implement end to end inference pipelines on standard PCs as well as on embedded systems
- Understand performance benchmarks and assess the accuracy and inference times
- Implement traditional image processing algorithms
- Factor the code to leverage underlying hardware architecture
- Prune the networks for efficiency
- Integrate the system within the application framework using C++
- Work closely with perception, controls, embedded software, and systems engineering teams.
Required Qualifications
- B.E./B.Tech/M.E./M.Tech in Computer Science and Engineering, Electronics, ECE, Mechatronics, or a related discipline.
- 2–3 years of experience in relevant area
- Strong understanding of: Linear Algebra, Probability and Statistics, AI-ML-DL fundamentals, Image processing, Camera Functioning
- Strong programming skills in C++ and Python.
- Experience with MATLAB for algorithm development and validation.
- Familiarity with Linux development environments.
- Experience with Git version control.
Preferred Skills
- Experience with Camera, IMU Calibration and Synchronisation
- Experience with multi-sensor fusion.
- Experience working with NVIDIA devices
- Experience on FPGA will be an added advantage
- Full understanding of Git functionality
- Exposure to airborne software development processes and coding standards (e.g., MISRA C++).
Personal Attributes
- Strong analytical and problem-solving skills.
- Ability to work independently on challenging technical problems.
- Good communication and documentation skills.
- Passion for solving challenging problems
- Willingness to participate in field trials and flight testing.
- Team playwe
Hi,
Greetings !!
We/re are looking for someone who has Hands-on experience with CV/ML
The location for the same is Bangalore.
Requirements
- 11–14 years total experience
- Computer Vision – strong hands-on experience
- Object Detection – YOLO(Preferred), Faster R-CNN, SSD, etc.
- Image Processing – OpenCV, image enhancement, segmentation, feature extraction
- Machine Learning / Deep Learning – CNNs, model training, evaluation, optimization
- AI/ML – production-level AI solution development
- LLM / GenAI – practical exposure to LLMs, multimodal AI, RAG, VLMs, or GenAI
- Python – strong programming skills
- Model deployment – preferably TensorRT, ONNX, Docker, Kubernetes, cloud, or edge deployment
- Bangalore – candidate should be based in / willing to work from Bangalore
Preferred
- Vision Transformers / ViT
- YOLOv8/YOLOv9/YOLOv10/YOLO11
- PyTorch / TensorFlow
- NLP / LLM / VLM
- Generative AI
- CUDA / GPU optimization
- Edge AI / NVIDIA
- Experience leading CV/AI projects or teams
If interested, Share CV at: snigdhaattheratebeanhr.com
AuxoAI is hiring a Senior Applied AI Engineer to design and deploy production-grade computer vision systems that operate reliably in real-world environments.
This role focuses on building end-to-end visual intelligence systems, combining deep learning, classical computer vision techniques, and multimodal models. It is not limited to model training and requires strong ownership of system design, deployment, and real-world performance.
You will work on systems that perform perception, understanding, and reasoning over visual data, and integrate these capabilities into larger AI platforms and agent-based workflows.
You will also work on problems where existing approaches may not be sufficient, and will be expected to combine deep learning, geometric methods, and multimodal reasoning to build robust, production-grade systems.
Location – Mumbai / Bangalore / Hyderabad / Gurgaon (Hybrid – 3 days per week in office)
Responsibilities:
- Design and deploy computer vision systems for tasks such as:
- Object detection, segmentation, and tracking
- Scene understanding and structured perception
- Video understanding and temporal reasoning
- Build and optimize models using architectures such as:
- CNNs (ResNet, EfficientNet)
- Vision Transformers (ViT, Swin, DeiT)
- Detection/segmentation models (YOLO, DETR, Mask R-CNN)
- Develop multimodal systems combining vision and language:
- CLIP-style models
- Vision-language models (VLMs)
- Visual grounding and captioning systems
- Implement algorithms for:
- Multi-object tracking (SORT, DeepSORT, ByteTrack)
- Feature matching and representation learning
- Temporal modeling (RNNs, Transformers for video)
- Apply geometric and classical computer vision methods where relevant:
- Camera calibration
- Epipolar geometry
- Pose estimation
- 3D reconstruction or depth estimation
- Optimize systems for:
- Low-latency, real-time inference
- Throughput and scalability
- Edge and distributed deployment
- Design and build data pipelines for:
- Annotation workflows
- Dataset curation
- Synthetic data generation
- Integrate vision systems into:
- Multimodal AI pipelines
- Agent-based systems
- Decision-making workflows
Requirements:
- 5+ years of experience building computer vision systems in production environments
- Strong experience with deep learning frameworks (PyTorch / TensorFlow)
- Hands-on experience with:
- Detection, segmentation, or tracking systems
- Model training, fine-tuning, and evaluation
- Strong understanding of:
- Representation learning
- Loss functions (contrastive loss, focal loss, etc.)
- Evaluation metrics (mAP, IoU, precision/recall)
- Experience building and deploying end-to-end vision systems, not just training models
Candidates whose primary experience is limited to academic projects or model experimentation without real-world deployment may not be a fit for this role.
Nice to Have:
- Experience with multimodal systems (vision + language)
- Familiarity with models such as:
- CLIP, BLIP, Flamingo, or similar
- Experience with 3D vision:
- NeRFs
- SLAM
- Point clouds
- Experience with video understanding:
- Action recognition
- Event detection
- Experience building data engines:
- Active learning
- Hard negative mining
- Experience working with large-scale datasets and distributed training pipelines
AI based systems design and development, entire pipeline from image/ video ingest, metadata ingest, processing, encoding, transmitting.
Implementation and testing of advanced computer vision algorithms.
Dataset search, preparation, annotation, training, testing, fine tuning of vision CNN models. Multimodal AI, LLMs, hardware deployment, explainability.
Detailed analysis of results. Documentation, version control, client support, upgrades.
[Please refrain from applying if you have over 10 years of experience. This is a hands-on role that requires building from the ground up.]
Location: Bengaluru (In-Office)
Employment Type: Full-Time
About Logikality
Logikality is building an AI-native mortgage intelligence platform for the U.S. mortgage industry. We are reimagining how mortgage operations are executed by combining AI, workflow automation, and domain expertise to solve one of the most document-intensive and decision-heavy industries in the world.
Our platform goes beyond document extraction. We are building AI systems that understand mortgage files, reason across multiple sources of information, identify risks and exceptions, support underwriting and quality control decisions, and continuously improve through expert feedback and rigorous evaluation.
As we expand our AI capabilities, we are looking for a Director, AI Engineering to define and drive the research direction behind our next generation of intelligent systems.
About the Role
This is a hands-on technical leadership role for someone who enjoys solving difficult AI problems and turning research into production impact.
You will lead the research agenda across large language models, reasoning systems, agentic AI, multimodal learning, and intelligent decision support while working closely with engineering, product, and mortgage domain experts. You will prototype new ideas, validate them through rigorous experimentation, and help productionize solutions that directly improve customer outcomes.
This role is ideal for someone with deep research expertise who enjoys building real-world AI systems rather than research for its own sake.
What You'll Do
- Define and execute the Applied AI research roadmap aligned with company and product goals.
- Design novel approaches for document understanding, reasoning, planning, retrieval, and decision support.
- Build agentic AI systems capable of orchestrating tools, workflows, and domain knowledge to solve complex mortgage use cases.
- Develop multimodal AI models that combine documents, structured data, images, and operational context.
- Lead research on long-context reasoning, knowledge integration, memory, retrieval-augmented generation (RAG), and workflow automation.
- Design robust evaluation frameworks, benchmarks, and automated testing pipelines to measure model quality, reliability, explainability, and business impact.
- Rapidly prototype, experiment, and iterate on new AI techniques, evaluating state-of-the-art research for production adoption.
- Work closely with software engineers to translate research prototypes into scalable, production-ready systems.
- Mentor AI engineers and contribute to building a strong research culture within the organisation.
- Collaborate with mortgage domain experts to deeply understand operational workflows, compliance requirements, and decision-making processes.
- Stay current with advances in AI research and identify opportunities to leverage emerging techniques within our platform.
- Represent Logikality in customer interactions, strategic discussions, industry conferences, and business forums, communicating our AI vision, gathering market insights, and helping shape research priorities through direct engagement with customers and ecosystem partners.
What We're Looking For
- PhD in Computer Science, Artificial Intelligence, Machine Learning, or a related discipline; or an engineering degree in Computer Science or related disciplines from a premier engineering institution (e.g., IITs, IISc, NITs, BITS Pilani, or top-tier global universities).
- 3–8 years of professional experience in Applied AI, Machine Learning, or AI Research, with experience building production-grade AI systems
- Strong expertise in modern AI, including Large Language Models, transformers, agentic AI, reasoning systems, retrieval, multimodal learning, or adjacent areas.
- Strong software engineering skills with Python and modern machine learning frameworks.
- Experience designing and implementing production-grade AI systems that solve complex real-world problems.
- Strong understanding of model evaluation, benchmarking, experimentation, and AI system reliability.
- Experience balancing research innovation with engineering pragmatism and product delivery.
- Excellent problem-solving and communication skills with the ability to collaborate across engineering, product, and business teams.
Why Join Logikality?
At Logikality, you'll work on problems that require genuine reasoning, not just text generation. You'll help build AI systems that understand complex documents, synthesise information across workflows, explain decisions, identify exceptions, and improve through continuous learning and expert feedback.
This is an opportunity to work at the intersection of cutting-edge AI research and real-world impact, where your ideas won't remain as papers or prototypes; they'll power intelligent systems used every day by mortgage professionals. We are looking for someone who can connect AI, platform engineering, product thinking and customer outcomes.
For the right person, this could develop into a CTO and co-founder track over the next 6–9 months, based on contribution, technical leadership and mutual fit.
Interested candidates are requested to apply via the Google Form given: https://forms.gle/jFqKzfLhNCcCFU5t9
This will be a full-time in-office role based in Bangalore. Immediate joiners are preferred.







