Computer Vision Engineer at Leena AI · Remote, Gurugram · 2 - 6 years · ₹5L - ₹30L / yr (ESOP available) · Raised funding · Remote friendly · Posted 2 Dec 2021

Responsibilities:
• Develop computer vision systems for enterprises to be used by hundreds of our
customers
• Enhance existing Computer vision systems to achieve high performance
• Prototype new algorithms rapidly, iterating to achieve high levels of performance
• Package these prototypes as robust models written in production level code to be
integrated into the product
• Work closely with the ML engineers to explore and enhance new product features
leading to new areas of business
Requirements:
Strong understanding of linear algebra, optimisation, probability, statistics
• Experience in the data science methodology from exploratory data analysis, feature
engineering, model selection, deployment of the model at scale and model evaluation
• Background in machine learning with experience in large scale training and
convolutional neural networks
• Deep understanding of evaluation metrics for different computer vision tasks
• Knowledge of common architectures for various computer vision tasks like object
detection, recognition, and semantic segmentation
• Experience with model quantization is a plus
• Experience with Python Web Framework (Django/Flask/FastAPI), Machine Learning
frameworks like Tensorflow/Keras/Pytorch

About Leena AI
About
Connect with the team
Similar jobs (8)
Position: Computer Vision Engineer
Experience: 2–3 Years
Location: Bengaluru, Karnataka
Employment Type: Full-time
About the Role
We are seeking a highly motivated Computer Vision Engineer to join our autonomy and avionics team. The role involves developing, implementing, and validating computer vision models and algorithms and pipelines for UAVs operating in both GNSS-available and GNSS-denied environments.
The ideal candidate should have a strong foundation in theory of deep learning and machine learning, strong understanding of electromagnetic spectrum, imaging fundamentals, camera principles, and mathematical concepts with hands-on experience in implementing these algorithms on embedded or real-time systems.
Key Responsibilities
- Design, develop, and optimise AI Models
- Make custom CNNs/ modify existing CNNs to suit specific problems at hand
- Handle end-to-end training flow
- Implement end to end inference pipelines on standard PCs as well as on embedded systems
- Understand performance benchmarks and assess the accuracy and inference times
- Implement traditional image processing algorithms
- Factor the code to leverage underlying hardware architecture
- Prune the networks for efficiency
- Integrate the system within the application framework using C++
- Work closely with perception, controls, embedded software, and systems engineering teams.
Required Qualifications
- B.E./B.Tech/M.E./M.Tech in Computer Science and Engineering, Electronics, ECE, Mechatronics, or a related discipline.
- 2–3 years of experience in relevant area
- Strong understanding of: Linear Algebra, Probability and Statistics, AI-ML-DL fundamentals, Image processing, Camera Functioning
- Strong programming skills in C++ and Python.
- Experience with MATLAB for algorithm development and validation.
- Familiarity with Linux development environments.
- Experience with Git version control.
Preferred Skills
- Experience with Camera, IMU Calibration and Synchronisation
- Experience with multi-sensor fusion.
- Experience working with NVIDIA devices
- Experience on FPGA will be an added advantage
- Full understanding of Git functionality
- Exposure to airborne software development processes and coding standards (e.g., MISRA C++).
Personal Attributes
- Strong analytical and problem-solving skills.
- Ability to work independently on challenging technical problems.
- Good communication and documentation skills.
- Passion for solving challenging problems
- Willingness to participate in field trials and flight testing.
- Team playwe
Hi,
Greetings !!
We/re are looking for someone who has Hands-on experience with CV/ML
The location for the same is Bangalore.
Requirements
- 11–14 years total experience
- Computer Vision – strong hands-on experience
- Object Detection – YOLO(Preferred), Faster R-CNN, SSD, etc.
- Image Processing – OpenCV, image enhancement, segmentation, feature extraction
- Machine Learning / Deep Learning – CNNs, model training, evaluation, optimization
- AI/ML – production-level AI solution development
- LLM / GenAI – practical exposure to LLMs, multimodal AI, RAG, VLMs, or GenAI
- Python – strong programming skills
- Model deployment – preferably TensorRT, ONNX, Docker, Kubernetes, cloud, or edge deployment
- Bangalore – candidate should be based in / willing to work from Bangalore
Preferred
- Vision Transformers / ViT
- YOLOv8/YOLOv9/YOLOv10/YOLO11
- PyTorch / TensorFlow
- NLP / LLM / VLM
- Generative AI
- CUDA / GPU optimization
- Edge AI / NVIDIA
- Experience leading CV/AI projects or teams
If interested, Share CV at: snigdhaattheratebeanhr.com
About Naicos
Naicos, a fast-paced startup, builds AI-native products for algorithmic commerce: the future of how e-commerce runs. Our first products are already live with paying customers, and we are shipping new ones continuously.
Your Role
You will drive the research behind our imaging products, finding approaches to hard, unsolved problems in product and apparel imagery that work at production scale. You will run the experiments, prove what is viable, and hand a working approach to the engineering team.
Who We Are Looking For
• Total experience: 3 years or more, with a strong research orientation
• Deep learning frameworks in Python: PyTorch or TensorFlow
• Image processing in Python: OpenCV, Pillow, scikit-image
• Working knowledge of diffusion and other image generation models
We are looking for a strong research or research-student profile: someone who investigates, experiments and proves an approach, working closely with the AI Architect. Someone who is driven to build solutions, not just desk research.
AI Skills and Experience
• Computer vision: classical CV alongside deep learning.
• Segmentation, image-to-image translation, geometry and lighting;
• Generative imaging: diffusion models, conditioning and control, fine-tuning and LoRA,
• Reads academic papers, judges what is reproducible, and turns one into a working prototype in days
Good to have
• 3D and rendering; published research or open-source contributions; model optimisation for inference cost
Research and innovative problem solving
• Comfortable where there is no known answer, and defines the approach yourself
• Solves problems inventively rather than reaching for the biggest model; many results come from classical image processing, fitment and geometric transformation
Other Relevant Skills and Experience
• Designs experiments: baselines, measurable success criteria, honest reporting of negative results
• Explains findings to a non-research audience and guides engineers to production
• Git and reproducible experiment tracking (Weights & Biases, MLflow or similar)
Educational Qualification
• BE / B.Tech / ME / M.Tech in Computer Science
• BE / B.Tech / ME / M.Tech in any discipline with proven Computer Vision coursework or work
• MSc / MS in Computer Science, Maths, Statistics or Computer Vision
• PhD in Computer Vision or Machine Learning: an advantage, not a requirement
• Reputed Tier 1 university preferred
AuxoAI is hiring a Senior Applied AI Engineer to design and deploy production-grade computer vision systems that operate reliably in real-world environments.
This role focuses on building end-to-end visual intelligence systems, combining deep learning, classical computer vision techniques, and multimodal models. It is not limited to model training and requires strong ownership of system design, deployment, and real-world performance.
You will work on systems that perform perception, understanding, and reasoning over visual data, and integrate these capabilities into larger AI platforms and agent-based workflows.
You will also work on problems where existing approaches may not be sufficient, and will be expected to combine deep learning, geometric methods, and multimodal reasoning to build robust, production-grade systems.
Location – Mumbai / Bangalore / Hyderabad / Gurgaon (Hybrid – 3 days per week in office)
Responsibilities:
- Design and deploy computer vision systems for tasks such as:
- Object detection, segmentation, and tracking
- Scene understanding and structured perception
- Video understanding and temporal reasoning
- Build and optimize models using architectures such as:
- CNNs (ResNet, EfficientNet)
- Vision Transformers (ViT, Swin, DeiT)
- Detection/segmentation models (YOLO, DETR, Mask R-CNN)
- Develop multimodal systems combining vision and language:
- CLIP-style models
- Vision-language models (VLMs)
- Visual grounding and captioning systems
- Implement algorithms for:
- Multi-object tracking (SORT, DeepSORT, ByteTrack)
- Feature matching and representation learning
- Temporal modeling (RNNs, Transformers for video)
- Apply geometric and classical computer vision methods where relevant:
- Camera calibration
- Epipolar geometry
- Pose estimation
- 3D reconstruction or depth estimation
- Optimize systems for:
- Low-latency, real-time inference
- Throughput and scalability
- Edge and distributed deployment
- Design and build data pipelines for:
- Annotation workflows
- Dataset curation
- Synthetic data generation
- Integrate vision systems into:
- Multimodal AI pipelines
- Agent-based systems
- Decision-making workflows
Requirements:
- 5+ years of experience building computer vision systems in production environments
- Strong experience with deep learning frameworks (PyTorch / TensorFlow)
- Hands-on experience with:
- Detection, segmentation, or tracking systems
- Model training, fine-tuning, and evaluation
- Strong understanding of:
- Representation learning
- Loss functions (contrastive loss, focal loss, etc.)
- Evaluation metrics (mAP, IoU, precision/recall)
- Experience building and deploying end-to-end vision systems, not just training models
Candidates whose primary experience is limited to academic projects or model experimentation without real-world deployment may not be a fit for this role.
Nice to Have:
- Experience with multimodal systems (vision + language)
- Familiarity with models such as:
- CLIP, BLIP, Flamingo, or similar
- Experience with 3D vision:
- NeRFs
- SLAM
- Point clouds
- Experience with video understanding:
- Action recognition
- Event detection
- Experience building data engines:
- Active learning
- Hard negative mining
- Experience working with large-scale datasets and distributed training pipelines
AI based systems design and development, entire pipeline from image/ video ingest, metadata ingest, processing, encoding, transmitting.
Implementation and testing of advanced computer vision algorithms.
Dataset search, preparation, annotation, training, testing, fine tuning of vision CNN models. Multimodal AI, LLMs, hardware deployment, explainability.
Detailed analysis of results. Documentation, version control, client support, upgrades.
This is a remote position.
About Leegality:
Leegality works with large Indian businesses to digitally transform critical compliance processes in a fast, easy and secure way.
We have multiple products across 2 categories:
Document Infrastructure:
Products that help businesses build paperless processes at scale:
- Document Execution Workflow: A unified platform for businesses to digitally execute (eSign, eStamp, Template Pre-fill, Document Fraud Prevention etc.) agreements, forms and other documents in a compliant way. Currently in use by 2000+ Indian businesses from giants like HDFC and SBI Cards to high-growth disruptors like goDigit and Cars24.
- Contract Management: An AI-powered platform for businesses to quickly review, negotiate and take action on contract
- Signstation: A simple platform for businesses to digitally sign simple documents like invoices, policies and letters in a cost effective manner
Consent Infrastructure:
- Consentin: An end-to-end DPDP and Privacy compliance platform for Indian businesses
- Consentin Lens: A data discovery platform for businesses to identify the personal data they collect and store.
If you’re interested in building mission critical software that operates at population scale (75 million + Indians have signed at least one document through Leegality) then join Leegality.
Curious about our impact? Explore our customer success stories: leegality.com/case-studies
Our Culture
At Leegality, trust, ownership, transparency, and having fun while doing meaningful work are core to how we operate — not just values on paper. Our team rated us an incredible 97 eNPS for FY 2023–24 — the highest among 175+ startups surveyed.
We focus deeply on helping our people grow and stay motivated. Some of the perks you’ll enjoy:
- Flexible working hours
- Hybrid work setup
- Bi-annual performance appraisals
- A culture that rewards initiative, curiosity, and impact
If you're looking for a place where you can make a real difference while working with smart, driven, and genuinely nice people, welcome to Leegality.
Location: Hybrid
Job Brief:
- As a Machine Learning Engineer specializing in Computer Vision (CV) and Natural Language Processing (NLP), you will develop solutions to interesting technical problems, exploring exciting growth opportunities and having a real impact on our product, particularly focusing on document and content intelligence.
- To ensure success, you should demonstrate solid data science knowledge and experience in a related ML, CV, or NLP role. A first-class engineer will be someone whose expertise enhances our systems for document intelligence and content processing
Responsibilities:
- Designing machine learning systems, self-running artificial intelligence (AI) software, and specialized models for Computer Vision and Natural Language Processing applications.
- Transforming data science prototypes and applying appropriate deep learning algorithms and tools to text and image/document data.
- Solving complex CV and NLP problems with multi-layered data types, such as image/document classification, information extraction, semantic search, and object detection.
- Optimizing existing machine learning models, with a focus on high-performance model deployment for CV and NLP tasks.
- Developing ML algorithms (including large language models/LLMs and computer vision models) to analyze huge volumes of historical text, image, and document data to make predictions and automate workflows.
- Running tests, performing statistical analysis, and interpreting test results for CV/NLP model performance.
- Documenting machine learning processes, model architectures, and data pipelines.
- Keeping abreast of developments in machine learning, Computer Vision, and Natural Language Processing.
Requirements:
- 3+ years of relevant experience in Machine Learning Engineering, with a strong focus on Computer Vision and/or Natural Language Processing.
- Advanced proficiency with Python.
- Extensive knowledge of ML frameworks, libraries (e.g., PyTorch, Transformers), data structures, data modeling, and software architecture.
- Experience with building and maintaining scalable RESTful APIs (e.g., FastAPI).
- In-depth knowledge of mathematics, statistics, deep learning (CNNs, RNNs, Transformers), and algorithms.
- Superb analytical and problem-solving abilities, especially for unstructured data challenges.
- Great communication and collaboration skills.
- Excellent time management and organizational abilities.
- Experience with cloud platforms (e.g., AWS) for model deployment and MLOps.
Recruitment Process:
- Our hiring process combines AI-powered evaluations with structured interviews to ensure a fair and seamless experience.
- You will be contacted via email with the next steps upon being shortlisted.
- The process may include Assessments, AI-enabled interviews, and In-Person Interviews with our team.
- Final selection and CTC will be based on your overall performance and experience.
Apply directly through our career page: https://careers.leegality.com/jobs/Careers
For more information about us please visit our:
Our Company and Culture: https://bit.ly/3Iqm5SB
Our Website: www.leegality.com/
Our LinkedIn Page: www.linkedin.com/company/leegality/
Leegality's Privacy Notice: https://www.leegality.com/employee-privacy-notice
We are hiring a Computer Vision Engineer to build vision models that run in real products.
Responsibilities
- Build object detection and segmentation models
- Develop image-processing pipelines with OpenCV
- Train and fine-tune YOLO-family models
- Optimise models for real-time inference
Requirements
- 1+ years in computer vision
- Hands-on with OpenCV and YOLO or similar detectors
- Experience deploying vision models
ML DEVELOPER
Hyperworks Imaging is a cutting-edge technology company based out of Bengaluru, India since 2016. Our team uses the latest advances in deep learning and multi-modal machine learning techniques to solve diverse real world problems. We are rapidly growing, working with multiple companies around the world.
JOB OVERVIEW
We are seeking a talented and results-oriented ML Developer to join our growing team in India. In this role, you will be responsible for developing and implementing new advanced ML algorithms and AI agents for creating AI assistants of the future.
The ideal candidate will work on a complete ML pipeline starting from extraction, transformation and analysis of data to developing novel ML algorithms. The candidate will implement latest research papers and closely work with various stakeholders to ensure data-driven decisions and integrate the solutions into a robust ML pipeline.
RESPONSIBILITIES:
- Create AI agents using Model Context Protocols (MCPs), Claude Code, DsPy etc.
- Develop custom evals for AI agents.
- Build and maintain ML pipelines
- Optimize and evaluate ML models to ensure accuracy and performance.
- Define system requirements and integrate ML algorithms into cloud based workflows.
- Write clean, well-documented, and maintainable code following best practices
REQUIREMENTS:
- 2-3+ years of experience in data science, machine learning, or a similar role.
- Demonstrated expertise with python, PyTorch, and TensorFlow.
- Graduated/Graduating with B.Tech/M.Tech/PhD degrees in Electrical Engg./Electronics Engg./Computer Science/Maths and Computing/Physics
- Has done coursework in Linear Algebra, Probability, Image Processing, Deep Learning and Machine Learning.
- Has demonstrated experience with Model Context Protocols (MCPs), DSPy, AI Agents, MLOps etc
WHO CAN APPLY:
Only those candidates will be considered who,
- have relevant skills and interests
- can commit full time
- Can show prior work and deployed projects
- can start immediately
Please note that we will reach out to ONLY those applicants who satisfy the criteria listed above.
SALARY DETAILS: Commensurate with experience.
JOINING DATE: Immediate
JOB TYPE: Full-time






