computer vision engineer at create high quality product images/videos at scale using AI · Gurugram · 3 - 9 years · ₹25L - ₹50L / yr · Posted 19 Apr 2022

computer vision engineer
at create high quality product images/videos at scale using AI
Job Description : - We are looking for a seasoned Computer Vision Engineer with AI/ML/CV and Deep Learning skills to play a senior leadership role in our Product & Technology Research Team. -
You will be leading a team of CV researchers to build models that automatically transform millions of e-commerce, automobiles, food, real-estate ram images into processed final images. -
You will be responsible for researching the latest art of the possible in the field of computer vision, designing the solution architecture for our offerings and lead the Computer Vision teams to build the core algorithmic models & deploy them on Cloud Infrastructure. -
Working with the Data team to ensure your data pipelines are well set up and models are being constantly trained and updated - Working alongside product team to ensure that AI capabilities are built as democratized tools that provides internal as well external stakeholders to innovate on top of it and make our customers successful - You will work closely with the Product & Engineering teams to convert the models into beautiful products that will be used by thousands of Businesses everyday to transform their images and videos.
Job Requirements: - Min 3+ years of work experience in Computer Vision with 5-8 years work experience overall - BS/MS/ Phd degree in Computer Science, Engineering or a related subject from a ivy league institute - Exposure on Deep Learning Techniques, TensorFlow/Pytorch - Prior expertise on building Image processing applications using GANs, CNNs, Diffusion models - Expertise with Image Processing Python libraries like OpenCV, etc. - Good hands-on experience on Python, Flask or Django framework - Authored publications at peer-reviewed AI conferences (e.g. NeurIPS, CVPR, ICML, ICLR,ICCV, ACL)
- Prior experience of managing teams and building large scale AI / CV projects is a big plus - Great interpersonal and communication skills - Critical thinker and problem-solving skills In Media.

Similar jobs (8)
About Naicos
Naicos, a fast-paced startup, builds AI-native products for algorithmic commerce: the future of how e-commerce runs. Our first products are already live with paying customers, and we are shipping new ones continuously.
Your Role
You will drive the research behind our imaging products, finding approaches to hard, unsolved problems in product and apparel imagery that work at production scale. You will run the experiments, prove what is viable, and hand a working approach to the engineering team.
Who We Are Looking For
• Total experience: 3 years or more, with a strong research orientation
• Deep learning frameworks in Python: PyTorch or TensorFlow
• Image processing in Python: OpenCV, Pillow, scikit-image
• Working knowledge of diffusion and other image generation models
We are looking for a strong research or research-student profile: someone who investigates, experiments and proves an approach, working closely with the AI Architect. Someone who is driven to build solutions, not just desk research.
AI Skills and Experience
• Computer vision: classical CV alongside deep learning.
• Segmentation, image-to-image translation, geometry and lighting;
• Generative imaging: diffusion models, conditioning and control, fine-tuning and LoRA,
• Reads academic papers, judges what is reproducible, and turns one into a working prototype in days
Good to have
• 3D and rendering; published research or open-source contributions; model optimisation for inference cost
Research and innovative problem solving
• Comfortable where there is no known answer, and defines the approach yourself
• Solves problems inventively rather than reaching for the biggest model; many results come from classical image processing, fitment and geometric transformation
Other Relevant Skills and Experience
• Designs experiments: baselines, measurable success criteria, honest reporting of negative results
• Explains findings to a non-research audience and guides engineers to production
• Git and reproducible experiment tracking (Weights & Biases, MLflow or similar)
Educational Qualification
• BE / B.Tech / ME / M.Tech in Computer Science
• BE / B.Tech / ME / M.Tech in any discipline with proven Computer Vision coursework or work
• MSc / MS in Computer Science, Maths, Statistics or Computer Vision
• PhD in Computer Vision or Machine Learning: an advantage, not a requirement
• Reputed Tier 1 university preferred
Hi,
Greetings !!
We/re are looking for someone who has Hands-on experience with CV/ML
The location for the same is Bangalore.
Requirements
- 11–14 years total experience
- Computer Vision – strong hands-on experience
- Object Detection – YOLO(Preferred), Faster R-CNN, SSD, etc.
- Image Processing – OpenCV, image enhancement, segmentation, feature extraction
- Machine Learning / Deep Learning – CNNs, model training, evaluation, optimization
- AI/ML – production-level AI solution development
- LLM / GenAI – practical exposure to LLMs, multimodal AI, RAG, VLMs, or GenAI
- Python – strong programming skills
- Model deployment – preferably TensorRT, ONNX, Docker, Kubernetes, cloud, or edge deployment
- Bangalore – candidate should be based in / willing to work from Bangalore
Preferred
- Vision Transformers / ViT
- YOLOv8/YOLOv9/YOLOv10/YOLO11
- PyTorch / TensorFlow
- NLP / LLM / VLM
- Generative AI
- CUDA / GPU optimization
- Edge AI / NVIDIA
- Experience leading CV/AI projects or teams
If interested, Share CV at: snigdhaattheratebeanhr.com
Position: Computer Vision Engineer
Experience: 2–3 Years
Location: Bengaluru, Karnataka
Employment Type: Full-time
About the Role
We are seeking a highly motivated Computer Vision Engineer to join our autonomy and avionics team. The role involves developing, implementing, and validating computer vision models and algorithms and pipelines for UAVs operating in both GNSS-available and GNSS-denied environments.
The ideal candidate should have a strong foundation in theory of deep learning and machine learning, strong understanding of electromagnetic spectrum, imaging fundamentals, camera principles, and mathematical concepts with hands-on experience in implementing these algorithms on embedded or real-time systems.
Key Responsibilities
- Design, develop, and optimise AI Models
- Make custom CNNs/ modify existing CNNs to suit specific problems at hand
- Handle end-to-end training flow
- Implement end to end inference pipelines on standard PCs as well as on embedded systems
- Understand performance benchmarks and assess the accuracy and inference times
- Implement traditional image processing algorithms
- Factor the code to leverage underlying hardware architecture
- Prune the networks for efficiency
- Integrate the system within the application framework using C++
- Work closely with perception, controls, embedded software, and systems engineering teams.
Required Qualifications
- B.E./B.Tech/M.E./M.Tech in Computer Science and Engineering, Electronics, ECE, Mechatronics, or a related discipline.
- 2–3 years of experience in relevant area
- Strong understanding of: Linear Algebra, Probability and Statistics, AI-ML-DL fundamentals, Image processing, Camera Functioning
- Strong programming skills in C++ and Python.
- Experience with MATLAB for algorithm development and validation.
- Familiarity with Linux development environments.
- Experience with Git version control.
Preferred Skills
- Experience with Camera, IMU Calibration and Synchronisation
- Experience with multi-sensor fusion.
- Experience working with NVIDIA devices
- Experience on FPGA will be an added advantage
- Full understanding of Git functionality
- Exposure to airborne software development processes and coding standards (e.g., MISRA C++).
Personal Attributes
- Strong analytical and problem-solving skills.
- Ability to work independently on challenging technical problems.
- Good communication and documentation skills.
- Passion for solving challenging problems
- Willingness to participate in field trials and flight testing.
- Team player
AuxoAI is hiring a Senior Applied AI Engineer to design and deploy production-grade computer vision systems that operate reliably in real-world environments.
This role focuses on building end-to-end visual intelligence systems, combining deep learning, classical computer vision techniques, and multimodal models. It is not limited to model training and requires strong ownership of system design, deployment, and real-world performance.
You will work on systems that perform perception, understanding, and reasoning over visual data, and integrate these capabilities into larger AI platforms and agent-based workflows.
You will also work on problems where existing approaches may not be sufficient, and will be expected to combine deep learning, geometric methods, and multimodal reasoning to build robust, production-grade systems.
Location – Mumbai / Bangalore / Hyderabad / Gurgaon (Hybrid – 3 days per week in office)
Responsibilities:
- Design and deploy computer vision systems for tasks such as:
- Object detection, segmentation, and tracking
- Scene understanding and structured perception
- Video understanding and temporal reasoning
- Build and optimize models using architectures such as:
- CNNs (ResNet, EfficientNet)
- Vision Transformers (ViT, Swin, DeiT)
- Detection/segmentation models (YOLO, DETR, Mask R-CNN)
- Develop multimodal systems combining vision and language:
- CLIP-style models
- Vision-language models (VLMs)
- Visual grounding and captioning systems
- Implement algorithms for:
- Multi-object tracking (SORT, DeepSORT, ByteTrack)
- Feature matching and representation learning
- Temporal modeling (RNNs, Transformers for video)
- Apply geometric and classical computer vision methods where relevant:
- Camera calibration
- Epipolar geometry
- Pose estimation
- 3D reconstruction or depth estimation
- Optimize systems for:
- Low-latency, real-time inference
- Throughput and scalability
- Edge and distributed deployment
- Design and build data pipelines for:
- Annotation workflows
- Dataset curation
- Synthetic data generation
- Integrate vision systems into:
- Multimodal AI pipelines
- Agent-based systems
- Decision-making workflows
Requirements:
- 5+ years of experience building computer vision systems in production environments
- Strong experience with deep learning frameworks (PyTorch / TensorFlow)
- Hands-on experience with:
- Detection, segmentation, or tracking systems
- Model training, fine-tuning, and evaluation
- Strong understanding of:
- Representation learning
- Loss functions (contrastive loss, focal loss, etc.)
- Evaluation metrics (mAP, IoU, precision/recall)
- Experience building and deploying end-to-end vision systems, not just training models
Candidates whose primary experience is limited to academic projects or model experimentation without real-world deployment may not be a fit for this role.
Nice to Have:
- Experience with multimodal systems (vision + language)
- Familiarity with models such as:
- CLIP, BLIP, Flamingo, or similar
- Experience with 3D vision:
- NeRFs
- SLAM
- Point clouds
- Experience with video understanding:
- Action recognition
- Event detection
- Experience building data engines:
- Active learning
- Hard negative mining
- Experience working with large-scale datasets and distributed training pipelines
[Please refrain from applying if you have over 10 years of experience. This is a hands-on role that requires building from the ground up.]
Location: Bengaluru (In-Office)
Employment Type: Full-Time
About Logikality
Logikality is building an AI-native mortgage intelligence platform for the U.S. mortgage industry. We are reimagining how mortgage operations are executed by combining AI, workflow automation, and domain expertise to solve one of the most document-intensive and decision-heavy industries in the world.
Our platform goes beyond document extraction. We are building AI systems that understand mortgage files, reason across multiple sources of information, identify risks and exceptions, support underwriting and quality control decisions, and continuously improve through expert feedback and rigorous evaluation.
As we expand our AI capabilities, we are looking for a Director, AI Engineering to define and drive the research direction behind our next generation of intelligent systems.
About the Role
This is a hands-on technical leadership role for someone who enjoys solving difficult AI problems and turning research into production impact.
You will lead the research agenda across large language models, reasoning systems, agentic AI, multimodal learning, and intelligent decision support while working closely with engineering, product, and mortgage domain experts. You will prototype new ideas, validate them through rigorous experimentation, and help productionize solutions that directly improve customer outcomes.
This role is ideal for someone with deep research expertise who enjoys building real-world AI systems rather than research for its own sake.
What You'll Do
- Define and execute the Applied AI research roadmap aligned with company and product goals.
- Design novel approaches for document understanding, reasoning, planning, retrieval, and decision support.
- Build agentic AI systems capable of orchestrating tools, workflows, and domain knowledge to solve complex mortgage use cases.
- Develop multimodal AI models that combine documents, structured data, images, and operational context.
- Lead research on long-context reasoning, knowledge integration, memory, retrieval-augmented generation (RAG), and workflow automation.
- Design robust evaluation frameworks, benchmarks, and automated testing pipelines to measure model quality, reliability, explainability, and business impact.
- Rapidly prototype, experiment, and iterate on new AI techniques, evaluating state-of-the-art research for production adoption.
- Work closely with software engineers to translate research prototypes into scalable, production-ready systems.
- Mentor AI engineers and contribute to building a strong research culture within the organisation.
- Collaborate with mortgage domain experts to deeply understand operational workflows, compliance requirements, and decision-making processes.
- Stay current with advances in AI research and identify opportunities to leverage emerging techniques within our platform.
- Represent Logikality in customer interactions, strategic discussions, industry conferences, and business forums, communicating our AI vision, gathering market insights, and helping shape research priorities through direct engagement with customers and ecosystem partners.
What We're Looking For
- PhD in Computer Science, Artificial Intelligence, Machine Learning, or a related discipline; or an engineering degree in Computer Science or related disciplines from a premier engineering institution (e.g., IITs, IISc, NITs, BITS Pilani, or top-tier global universities).
- 3–8 years of professional experience in Applied AI, Machine Learning, or AI Research, with experience building production-grade AI systems
- Strong expertise in modern AI, including Large Language Models, transformers, agentic AI, reasoning systems, retrieval, multimodal learning, or adjacent areas.
- Strong software engineering skills with Python and modern machine learning frameworks.
- Experience designing and implementing production-grade AI systems that solve complex real-world problems.
- Strong understanding of model evaluation, benchmarking, experimentation, and AI system reliability.
- Experience balancing research innovation with engineering pragmatism and product delivery.
- Excellent problem-solving and communication skills with the ability to collaborate across engineering, product, and business teams.
Why Join Logikality?
At Logikality, you'll work on problems that require genuine reasoning, not just text generation. You'll help build AI systems that understand complex documents, synthesise information across workflows, explain decisions, identify exceptions, and improve through continuous learning and expert feedback.
This is an opportunity to work at the intersection of cutting-edge AI research and real-world impact, where your ideas won't remain as papers or prototypes; they'll power intelligent systems used every day by mortgage professionals. We are looking for someone who can connect AI, platform engineering, product thinking and customer outcomes.
For the right person, this could develop into a CTO and co-founder track over the next 6–9 months, based on contribution, technical leadership and mutual fit.
Interested candidates are requested to apply via the Google Form given: https://forms.gle/jFqKzfLhNCcCFU5t9
This will be a full-time in-office role based in Bangalore. Immediate joiners are preferred.
About the Role We are seeking a highly technical, hands-on Senior AI/ML Tech Lead to drive the design, development, and deployment of cutting-edge Generative AI applications. In this dual-impact role, you wi l act as a primary individual contributor architecting core AI engines while simultaneously leading a team of engineers through task alocation, code reviews, and technical mentorship. The ideal candidate bridges the gap between state-of-the-art AI research (LLMs, Agentic frameworks, Advanced RAG, OCR) and production-grade ful-stack engineering (Python, FastAPI, React).
Key Responsibilities
Technical Leadership & Team Management (40%)
● Technical Oversight: Lead a team of AI, backend, and ful-stack engineers; alocate tasks, establish sprint priorities, and ensure timely delivery.
● Code Quality & Reviews: Conduct rigorous code reviews to maintain high engineering standards, security, performance, and scalability across AI and fu l-stack codebases.
● Architecture & Governance: Design end-to-end system architectures for AI solutions, ensuring seamless integration between frontend interfaces, backend APIs, and AI models.
● Mentorship: Guide and upskil team members on modern software practices, LLM engineering, and agentic design patterns. Hands-On Engineering & Development (60%)
● Generative AI & Agentic Systems: Architect, build, and optimize LLM-powered applications, multi-agent workflows (e.g., CrewAI, AutoGen, LangGraph), and autonomous AI agents.
● RAG & OCR Pipelines: Design and deploy advanced RAG (Retrieval-Augmented Generation) architectures and document processing pipelines utilizing OCR techniques (e.g., LayoutLM, PaddleOCR, Tesseract, Vision LLMs) to extract structured data from unstructured sources.
● Backend Systems: Build robust, asynchronous, high-throughput microservices and RESTful APIs using Python and FastAPI.
● Frontend Integration: Colaborate on or build modern web interfaces using React (e.g., Control Towers, operations dashboards, interactive chat interfaces).
● MLOps & Vector DBs: Oversee model deployment, prompt engineering, fine-tuning, vector database integration (Pinecone, Qdrant, Chroma, PGVector), and cloud infrastructure setup (Azure/AWS).
Required Qualifications & Skills
● Overall Experience: 8 to 10 years of professional software engineering experience.
● AI/ML Domain Experience: 3 to 4+ years of dedicated, hands-on experience building and deploying AI/ML, OCR, and Generative AI solutions in production.
● Core Technical Stack: ○ Generative AI & LLMs: Extensive experience with commercial and open-source LLMs (OpenAI, Anthropic Claude, Llama), Agentic frameworks (LangChain, LlamaIndex, AutoGen, CrewAI), and LLM evaluation frameworks (LangSmith, TruLens, Ragas). ○ RAG & Unstructured Data: Strong knowledge of hybrid search, re-ranking, chunking strategies, vector databases, and document inte ligence workflows. ○ OCR & Vision Techniques: Hands-on experience with OCR engines (Tesseract, PaddleOCR, Azure Document Inteligence) and Multi-Modal/Vision LLMs for document extraction. ○ Backend: Deep expertise in Python and asynchronous frameworks (FastAPI, AsyncIO). ○ Frontend: Working proficiency in React (TypeScript/JavaScript) for building interactive web UI components. ○ Cloud & DevOps: Hands-on experience with cloud platforms (Azure / AWS), Docker, Kubernetes, and CI/CD pipelines.
Preferred / Good-to-Have Skills
● Experience with cloud-native data platforms (e.g., Microsoft Fabric, Snowflake, Azure SQL).
● Familiarity with cost optimization and latency reduction techniques for LLM inference (caching, semantic routing, model quantization).
● Prior experience in client-facing technical leadership or agile consulting environments.
What We Offer
● Opportunity to lead and build high-impact, state-of-the-art Generative AI systems.
● Colaborative engineering culture with room for technical ownership and direct business impact.
● Flexible work arrangements and competitive compensation package.
Kody Technolab Limited
AI/ML Lead (7+ Years Experience)
Location: Ahmedabad / Gandhinagar
Experience: 7+ Years
Employment Type: Full-Time
Kody Technolab Limited is seeking an experienced AI/ML Engineer to design, develop, and deploy
cutting-edge Artificial Intelligence and Machine Learning solutions. The ideal candidate will have
strong expertise in Machine Learning, Deep Learning, Generative AI, LLMs, MLOps, and
cloud-based AI deployments.
Key Responsibilities
• Design, develop, and deploy Machine Learning and Deep Learning models for classification,
regression, recommendation systems, NLP, Computer Vision, and Generative AI applications.
• Build and maintain end-to-end ML pipelines including data preprocessing, feature
engineering, model training, validation, evaluation, and deployment.
• Develop AI solutions using PyTorch, TensorFlow, Scikit-learn, Hugging Face, and related
frameworks.
• Work with Large Language Models (LLMs) and foundation models such as GPT, BERT, Llama,
Claude, and Stable Diffusion.
• Collaborate with product, engineering, and business teams to translate requirements into
scalable AI solutions.
• Optimize model performance, scalability, and reliability for production environments.
• Implement MLOps best practices using tools such as MLflow, Docker, Kubernetes, and Kubeflow.
• Stay updated with emerging trends and research in AI, ML, Deep Learning, and Generative AI.
Required Qualifications
• Bachelor’s or Master’s degree in Computer Science, Data Science, Artificial Intelligence,
Mathematics, or a related field.
• 7+ years of hands-on experience in AI/ML product development.
• Strong proficiency in Python and ML frameworks including Scikit-learn, TensorFlow, PyTorch, and
Hugging Face.
• Experience with Generative AI, LLMs, GANs, VAEs, diffusion models, and prompt engineering.
• Strong understanding of the ML lifecycle including model training, tuning, deployment,
monitoring, and optimization.
• Experience with MLOps tools such as MLflow, Docker, Kubeflow, and CI/CD pipelines.
• Experience with AWS, Azure, or GCP cloud platforms.
• Strong problem-solving and analytical skills.Preferred Skills
• Fine-tuning and deployment of Large Language Models.
• Experience with RAG (Retrieval Augmented Generation) architectures.
• Contributions to open-source AI projects or research publications.
• Knowledge of model interpretability, data annotation, and feature engineering.
• C++ experience for high-performance AI applications.
Why Join Kody Technolab Limited?
Opportunity to work on innovative AI products, Generative AI solutions, robotics integrations,
and enterprise-scale applications while collaborating with a highly skilled technology team.
Visit the Website to know more about us.
Company Website - Kody Technolab | Deep Tech Company in Robotics & AI Solution
Kody Robots | Robotics Company in India for Autonomous Robots
Kody Technolab Limited is seeking an experienced AI/ML Engineer to design, develop, and deploy cutting-edge Artificial Intelligence and Machine Learning solutions. The ideal candidate will have
strong expertise in Machine Learning, Deep Learning, Generative AI, LLMs, MLOps, and cloud-based AI deployments.
Key Responsibilities
• Design, develop, and deploy Machine Learning and Deep Learning models for classification, regression, recommendation systems, NLP, Computer Vision, and Generative AI applications.
• Build and maintain end-to-end ML pipelines including data preprocessing, feature engineering, model training, validation, evaluation, and deployment.
• Develop AI solutions using PyTorch, TensorFlow, Scikit-learn, Hugging Face, and related frameworks.
• Work with Large Language Models (LLMs) and foundation models such as GPT, BERT, Llama, Claude, and Stable Diffusion.
• Collaborate with product, engineering, and business teams to translate requirements into scalable AI solutions.
• Optimize model performance, scalability, and reliability for production environments.
• Implement MLOps best practices using tools such as MLflow, Docker, Kubernetes, and Kubeflow.
• Stay updated with emerging trends and research in AI, ML, Deep Learning, and Generative AI.
Required Qualifications
• Bachelor’s or Master’s degree in Computer Science, Data Science, Artificial Intelligence, Mathematics, or a related field.
• 7+ years of hands-on experience in AI/ML product development.
• Strong proficiency in Python and ML frameworks including Scikit-learn, TensorFlow, PyTorch, and Hugging Face.
• Experience with Generative AI, LLMs, GANs, VAEs, diffusion models, and prompt engineering.
• Strong understanding of the ML lifecycle including model training, tuning, deployment, monitoring, and optimization.
• Experience with MLOps tools such as MLflow, Docker, Kubeflow, and CI/CD pipelines.
• Experience with AWS, Azure, or GCP cloud platforms.
• Strong problem-solving and analytical skills.
Preferred Skills
• Fine-tuning and deployment of Large Language Models.
• Experience with RAG (Retrieval Augmented Generation) architectures.
• Contributions to open-source AI projects or research publications.
• Knowledge of model interpretability, data annotation, and feature engineering.
• C++ experience for high-performance AI applications.
Why Join Kody Technolab Limited?
Opportunity to work on innovative AI products, Generative AI solutions, robotics integrations,
and enterprise-scale applications while collaborating with a highly skilled technology team.
Visit the Website to know more about us.
Company Website - Kody Technolab | Deep Tech Company in Robotics & AI Solution
Kody Robots | Robotics Company in India for Autonomous Robots










