Generative AI/GEN AI at Einfochips · Indore, Pune, Ahmedabad · 4 - 6 years · ₹7L - ₹10L / yr · Posted 7 Sep 2026
Experience - 4 to 6 year
Location – Ahmedabad/Pune/Indore
- Additional Job Description
Additional Job Description
Required Skills and Experience:
- Strong proficiency in Python and experience with ML/AI libraries (scikit-learn, TensorFlow, PyTorch, Hugging Face ecosystem).
- Hands-on experience with LLMs, RAG, vector databases, and retrieval pipelines.
- Practical experience deploying agentic workflows and building multi-step, tool-enabled agents.
- Experience using Garak (or similar LLM red-teaming/vulnerability scanners) to identify model weaknesses and harden deployments.
- Demonstrated experience implementing content filtering / moderation systems.
- Solid skills working with structured and unstructured data and advanced feature engineering.
- Familiarity with cloud GenAI platforms and services (Azure AI Services preferred; AWS/GCP acceptable).
- Experience building APIs/microservices; containerization (Docker), orchestration (Kubernetes).
- Strong understanding of model evaluation, performance profiling, inference cost optimization, and observability.
- Good knowledge of security, data governance, and privacy best practices for AI systems.

About Einfochips
About
Similar jobs (10)
Hiring for AI Engineer
Exp: 6 - 8 yrs
Edu : BE/B.Tech/MCA
Work Location : Pune
Skill Set:
- Total experience ranging from 6–8 years in software engineering/AI roles
- Min 5 years strong programming experience in Python is a MUST
- Min 3.5 years hands-on experience in AI with LLMs, RAG pipelines, and AI frameworks
- Experience with cloud platforms (AWS/Azure/GCP)
Job Summary/ Job Opportunity:
This is an excellent opportunity for an ideal candidate with a high level of technical proficiency and meeting the below mentioned criteria -- • Strong experience in Machine Learning, Deep Learning, Generative AI, and Large Language Models (LLMs). • Hands-on experience building and deploying production-grade solutions using Azure OpenAI, OpenAI, LangChain, LangGraph, Semantic Kernel, LlamaIndex, and Agentic AI frameworks. • Strong expertise in Python, API development, microservices, and cloud-native architectures. • Experience designing and implementing RAG solutions, vector databases, embeddings, knowledge retrieval systems, and AI copilots. • Experience with Azure cloud services, MLOps, CI/CD pipelines, monitoring, and model lifecycle management. • Strong understanding of AI governance, responsible AI, security, compliance, and model evaluation frameworks. • Ability to lead technical discussions, provide architectural recommendations, mentor team members, and interact with business stakeholde
Key Objectives and Major Responsibilities:
• Design, develop, and implement scalable AI/ML and Generative AI solutions for enterprise applications. • Lead development of intelligent applications leveraging LLMs, RAG pipelines, AI agents, and document intelligence solutions. • Collaborate with business stakeholders, architects, and product teams to translate business requirements into technical solutions. • Design and optimize data pipelines, vector search solutions, embeddings, and retrieval mechanisms. • Build and maintain REST APIs, microservices, and cloud-native AI applications. • Ensure best practices in coding standards, performance optimization, security, scalability, and maintainability. • Drive AI solution deployment using MLOps practices, CI/CD pipelines, monitoring, and observability frameworks. • Perform code reviews, mentor junior developers, and contribute to capability building within the team
Key Capabilities and Competencies:
Knowledge, Skills, Qualification and Experience
• Degree in B.Tech/M.Tech (Computer Science/IT/Data Science) or related discipline preferred, with 3–4 years of relevant experience in AI/ML, GenAI and total 5-7 years of experience. • Proficiency in Python and hands-on experience with ML libraries (scikit-learn, TensorFlow, PyTorch) and GenAI frameworks/tools. • Strong understanding of machine learning, deep learning, LLMs, prompt engineering, and techniques like RAG and fine-tuning. • Experience with data processing, embeddings, vector databases, APIs, and building scalable AI driven applications. • Good communication skills, ability to work on multiple projects, and eagerness to learn and adapt to evolving AI technologies.
The recruiter has not been active on this job recently. You may apply but please expect a delayed response.
Strong AI Engineer / Machine Learning Engineer profiles.
2
Mandatory (Experience 1) – Must have minimum 3+ years of hands-on experience in Data Science, Machine Learning, Applied AI, NLP, Deep Learning, or Generative AI solutions.
3
Mandatory (Experience 2) – Must have strong hands-on experience in Python programming, SQL, data analysis, feature engineering, model development, and production-grade ML applications.
4
Mandatory (Experience 3) – Must have experience working with Machine Learning and Deep Learning frameworks such as PyTorch, TensorFlow, Keras, Scikit-learn, or equivalent.
5
Mandatory (Experience 4) – Must have hands-on experience working on NLP, embeddings, semantic search, text classification, document understanding, recommendation systems, or similar AI/ML use cases.
6
Mandatory (Experience 5) – Must have experience working with Large Language Models (LLMs) such as GPT, Llama, Mistral, Claude, Gemini, Phi, or similar foundation models.
7
Mandatory (Experience 6) – Must have hands-on experience building or implementing RAG (Retrieval Augmented Generation) systems, vector search, knowledge retrieval, embeddings, chunking, indexing, or semantic retrieval solutions.
8
Mandatory (Experience 7) – Must have experience working with Git, CI/CD practices, production environments, and scalable AI/ML systems.
9
Mandatory (CTC) – The CTC breakup offered will be 75% fixed + 25% variable, as per company policy.
10
Mandatory (Age) - Candidate's Age should be below 28 Years
11
Preferred (Experience 1) – Experience with MLFlow, Kubeflow, Airflow, Prefect, Feature Stores, Model Registry, or MLOps/LLMOps frameworks.
12
Preferred (Experience 2) – Experience working with Vector Databases, Spark, PySpark, distributed ML pipelines, large-scale data processing, or real-time ML systems..
13
Preferred (Experience 3) – Familiarity with Docker, Kubernetes, Azure, AWS, GCP, cloud-native AI deployments, and scalable ML architecture.
14
Preferred (Company) – Candidates from AI-first startups, Fintech, Banking, Lending, Fraud Analytics, Risk Analytics, Product Companies, SaaS organizations, or data-driven technology companies
15
Mandatory ( Pedigree) - B.TECH / M.TECH from Tier 1 Colleges (IIT's, NIT's, BITS) are Considered.
The recruiter has not been active on this job recently. You may apply but please expect a delayed response.
Strong AI Engineer / Machine Learning Engineer profiles.
2
Mandatory (Experience 1) – Must have minimum 3+ years of hands-on experience in Data Science, Machine Learning, Applied AI, NLP, Deep Learning, or Generative AI solutions.
3
Mandatory (Experience 2) – Must have strong hands-on experience in Python programming, SQL, data analysis, feature engineering, model development, and production-grade ML applications.
4
Mandatory (Experience 3) – Must have experience working with Machine Learning and Deep Learning frameworks such as PyTorch, TensorFlow, Keras, Scikit-learn, or equivalent.
5
Mandatory (Experience 4) – Must have hands-on experience working on NLP, embeddings, semantic search, text classification, document understanding, recommendation systems, or similar AI/ML use cases.
6
Mandatory (Experience 5) – Must have experience working with Large Language Models (LLMs) such as GPT, Llama, Mistral, Claude, Gemini, Phi, or similar foundation models.
7
Mandatory (Experience 6) – Must have hands-on experience building or implementing RAG (Retrieval Augmented Generation) systems, vector search, knowledge retrieval, embeddings, chunking, indexing, or semantic retrieval solutions.
8
Mandatory (Experience 7) – Must have experience working with Git, CI/CD practices, production environments, and scalable AI/ML systems.
9
Mandatory (CTC) – The CTC breakup offered will be 75% fixed + 25% variable, as per company policy.
10
Mandatory (Age) - Candidate's Age should be below 30 Years
11
Preferred (Experience 1) – Experience with MLFlow, Kubeflow, Airflow, Prefect, Feature Stores, Model Registry, or MLOps/LLMOps frameworks.
12
Preferred (Experience 2) – Experience working with Vector Databases, Spark, PySpark, distributed ML pipelines, large-scale data processing, or real-time ML systems..
13
Preferred (Experience 3) – Familiarity with Docker, Kubernetes, Azure, AWS, GCP, cloud-native AI deployments, and scalable ML architecture.
14
Preferred (Company) – Candidates from AI-first startups, Fintech, Banking, Lending, Fraud Analytics, Risk Analytics, Product Companies, SaaS organizations, or data-driven technology companies
15
Mandatory ( Pedigree) - B.TECH / M.TECH from Tier 1 Colleges (IIT's, NIT's, BITS) are Considered.
About NonStop io Technologies
NonStop io Technologies is a value-driven company with a strong focus on process-oriented software engineering. We specialize in Product Development and have a decade's worth of experience in building web and mobile applications across various domains. NonStop io Technologies follows core principles that guide its operations and believes in staying invested in a product's vision for the long term. We are a small but proud group of individuals who believe in the 'givers gain' philosophy and strive to provide value in order to seek value. We are committed to and specialize in building cutting-edge technology products and serving as trusted technology partners for startups and enterprises. We pride ourselves on fostering innovation, learning, and community engagement. Join us to work on impactful projects in a collaborative and vibrant environment.
Brief Description:
We're seeking an AI/ML Engineer to join our team. As AI/ML Engineer, you will be responsible for designing, developing, and implementing artificial intelligence (AI) and machine learning (ML) solutions to solve real-world business problems. You will work closely with engineering teams, including software engineers, domain experts, and product managers, to deploy and integrate Applied AI/ML solutions into the products that are being built at NonStop io. Your role will involve researching cutting-edge algorithms and data processing techniques, and implementing scalable solutions to drive innovation and improve the overall user experience.
Responsibilities
● Applied AI/ML engineering; Building engineering solutions on top of the AI/ML tooling available in the industry today. Eg: Engineering APIs around OpenAI
● AI/ML Model Development: Design, develop, and implement machine learning models and algorithms that address specific business challenges, such as natural language processing, computer vision, recommendation systems, anomaly detection, etc.
● Data Preprocessing and Feature Engineering: Cleanse, preprocess, and transform raw data into suitable formats for training and testing AI/ML models. Perform feature engineering to extract relevant features from the data
● Model Training and Evaluation: Train and validate AI/ML models using diverse datasets to achieve optimal performance. Employ appropriate evaluation metrics to assess model accuracy, precision, recall, and other relevant metrics
● Data Visualization: Create clear and insightful data visualizations to aid in understanding data patterns, model behaviour, and performance metrics
● Deployment and Integration: Collaborate with software engineers and DevOps teams to deploy AI/ML models into production environments and integrate them into various applications and systems
● Data Security and Privacy: Ensure compliance with data privacy regulations and implement security measures to protect sensitive information used in AI/ML processes
● Continuous Learning: Stay updated with the latest advancements in AI/ML research, tools, and technologies, and apply them to improve existing models and develop novel solutions
● Documentation: Maintain detailed documentation of the AI/ML development process, including code, models, algorithms, and methodologies for easy understanding and future reference.
Qualifications & Skills
● Bachelor's, Master's, or PhD in Computer Science, Data Science, Machine Learning, or a related field. Advanced degrees or certifications in AI/ML are a plus
● Proven experience as an AI/ML Engineer, Data Scientist, or related role, ideally with a strong portfolio of AI/ML projects
● Proficiency in programming languages commonly used for AI/ML. Preferably Python
● Familiarity with popular AI/ML libraries and frameworks, such as TensorFlow, PyTorch, scikit-learn, etc.
● Familiarity with popular AI/ML Models such as GPT3, GPT4, Llama2, BERT etc.
● Strong understanding of machine learning algorithms, statistics, and data structures
● Experience with data preprocessing, data wrangling, and feature engineering
● Knowledge of deep learning architectures, neural networks, and transfer learning
● Familiarity with cloud platforms and services (e.g., AWS, Azure, Google Cloud) for scalable AI/ML deployment
● Solid understanding of software engineering principles and best practices for writing maintainable and scalable code
● Excellent analytical and problem-solving skills, with the ability to think critically and propose innovative solutions
● Effective communication skills to collaborate with cross-functional teams and present complex technical concepts to non-technical stakeholders
The recruiter has not been active on this job recently. You may apply but please expect a delayed response.
Strong AI Engineer / Machine Learning Engineer profiles.
2
Mandatory (Experience 1) – Must have minimum 5+ years of hands-on experience in Data Science, Machine Learning, Applied AI, NLP, Deep Learning, or Generative AI solutions.
3
Mandatory (Experience 2) – Must have strong hands-on experience in Python programming, SQL, data analysis, feature engineering, model development, and production-grade ML applications.
4
Mandatory (Experience 3) – Must have experience working with Machine Learning and Deep Learning frameworks such as PyTorch, TensorFlow, Keras, Scikit-learn, or equivalent.
5
Mandatory (Experience 4) – Must have hands-on experience working on NLP, embeddings, semantic search, text classification, document understanding, recommendation systems, or similar AI/ML use cases.
6
Mandatory (Experience 5) – Must have experience working with Large Language Models (LLMs) such as GPT, Llama, Mistral, Claude, Gemini, Phi, or similar foundation models.
7
Mandatory (Experience 6) – Must have hands-on experience building or implementing RAG (Retrieval Augmented Generation) systems, vector search, knowledge retrieval, embeddings, chunking, indexing, or semantic retrieval solutions.
8
Mandatory (Experience 7) – Must have experience working with Git, CI/CD practices, production environments, and scalable AI/ML systems.
9
Mandatory (CTC) – The CTC breakup offered will be 75% fixed + 25% variable, as per company policy.
10
Mandatory (Age) - Candidate's Age should be below 30 Years
11
Preferred (Experience 1) – Experience with MLFlow, Kubeflow, Airflow, Prefect, Feature Stores, Model Registry, or MLOps/LLMOps frameworks.
12
Preferred (Experience 2) – Experience working with Vector Databases, Spark, PySpark, distributed ML pipelines, large-scale data processing, or real-time ML systems..
13
Preferred (Experience 3) – Familiarity with Docker, Kubernetes, Azure, AWS, GCP, cloud-native AI deployments, and scalable ML architecture.
14
Preferred (Company) – Candidates from AI-first startups, Fintech, Banking, Lending, Fraud Analytics, Risk Analytics, Product Companies, SaaS organizations, or data-driven technology companies
15
Mandatory ( Pedigree) - B.TECH / M.TECH from Tier 1 Colleges (IIT's, NIT's, BITS) are Considered.
Strong AI Engineer / Machine Learning Engineer profiles.
2
Mandatory (Experience 1) – Must have minimum 3+ years of hands-on experience in Data Science, Machine Learning, Applied AI, NLP, Deep Learning, or Generative AI solutions.
3
Mandatory (Experience 2) – Must have strong hands-on experience in Python programming, SQL, data analysis, feature engineering, model development, and production-grade ML applications.
4
Mandatory (Experience 3) – Must have experience working with Machine Learning and Deep Learning frameworks such as PyTorch, TensorFlow, Keras, Scikit-learn, or equivalent.
5
Mandatory (Experience 4) – Must have hands-on experience working on NLP, embeddings, semantic search, text classification, document understanding, recommendation systems, or similar AI/ML use cases.
6
Mandatory (Experience 5) – Must have experience working with Large Language Models (LLMs) such as GPT, Llama, Mistral, Claude, Gemini, Phi, or similar foundation models.
7
Mandatory (Experience 6) – Must have hands-on experience building or implementing RAG (Retrieval Augmented Generation) systems, vector search, knowledge retrieval, embeddings, chunking, indexing, or semantic retrieval solutions.
8
Mandatory (Experience 7) – Must have experience working with Git, CI/CD practices, production environments, and scalable AI/ML systems.
9
Mandatory (CTC) – The CTC breakup offered will be 75% fixed + 25% variable, as per company policy.
10
Mandatory (Age) - Candidate's Age should be below 28 Years
The Role
You own AI systems end to end. From the speech-to-text models that turn audio into text, to the diarization that separates and identifies speakers, to the agentic layer that turns conversation into memory and action, to the observability and evaluation that keep all of it honest in production. This is a wide role by design. You will own model selection, serving, and production reliability. If you want to tune one model and ignore the system around it, this is not the role.
What You Will Own
• Speech-to-text. Evaluate, integrate, and optimize STT models across cloud and self-hosted. Drive accuracy and cost trade-offs with ground-truth metrics.
• Speaker diarization and identification. Push accuracy on hard, real-world, multi-speaker audio.
• Agentic AI. Build the memory and retrieval pipeline, LLM orchestration, and the agent workflows that sit on top of captured conversation.
• Model serving and infrastructure. Stand up and optimize self-hosted serving (vLLM, Triton class). Own latency, throughput, and cost per user.
Observability
An always-on wearable means models run in production every second, on messy real-world audio. You own the visibility into that.
• Instrument the full audio-to-memory pipeline: STT, diarization, retrieval, and LLM calls.
• Define and track model-quality SLOs in production: transcription drift, diarization error over time, retrieval relevance, latency, throughput, and cost per user.
• Build dashboards and alerting so model degradation is caught before users feel it.
• Trace failures across a distributed, always-on system using metrics, logs, and traces.
• Close the loop. Production signals feed back into evaluation and model selection.
Evaluation
We do not ship what we cannot measure. You own the systems that prove a model is actually better, not just newer.
• Build and own ground-truth evaluation harnesses for every model in the stack.
• Measure with real metrics: WER for transcription, DER for diarization, Recall and F1 for retrieval and speaker identification.
• Build and maintain labeled benchmark datasets that reflect real, messy, multi-speaker audio.
• Run regression and A/B evaluations on every model swap, prompt change, or pipeline update. Nothing ships on a vibe.
• Reject anecdotal proxies, single confidence scores, and cherry-picked examples as evidence of quality.
What We Are Looking For
• 3 to 5 years as an AI/ML engineer with production systems behind you. Engineering and production experience is non-negotiable.
• Depth across the modern AI stack: LLMs, speech models, vector retrieval, model serving.
• Strong software engineering. You write code that ships and survives contact with real users.
• Fluency in Python and the production ML ecosystem.
• Comfort with cloud infrastructure (GCP a plus) and containerized deployment on Kubernetes.
• A working command of observability and evaluation. You measure first and trust metrics over intuition.
• First-principles reasoning and metric discipline.
Nice to Have
• Research background or publications. A strong signal, not a substitute for production work.
• Audio and speech ML experience (STT, diarization, voice).
• Experience self-hosting and optimizing open models.
• Experience with LLM gateway and agent orchestration patterns.
• Experience building eval harnesses or production model-monitoring systems.
Requirements
Agentic work is must. Audio is good to have
. Self hosting models is a must
Experience with LLM gateway and agent orchestration is a must have
This is a remote position.
About Leegality:
Leegality works with large Indian businesses to digitally transform critical compliance processes in a fast, easy and secure way.
We have multiple products across 2 categories:
Document Infrastructure:
Products that help businesses build paperless processes at scale:
- Document Execution Workflow: A unified platform for businesses to digitally execute (eSign, eStamp, Template Pre-fill, Document Fraud Prevention etc.) agreements, forms and other documents in a compliant way. Currently in use by 2000+ Indian businesses from giants like HDFC and SBI Cards to high-growth disruptors like goDigit and Cars24.
- Contract Management: An AI-powered platform for businesses to quickly review, negotiate and take action on contract
- Signstation: A simple platform for businesses to digitally sign simple documents like invoices, policies and letters in a cost effective manner
Consent Infrastructure:
- Consentin: An end-to-end DPDP and Privacy compliance platform for Indian businesses
- Consentin Lens: A data discovery platform for businesses to identify the personal data they collect and store.
If you’re interested in building mission critical software that operates at population scale (75 million + Indians have signed at least one document through Leegality) then join Leegality.
Curious about our impact? Explore our customer success stories: leegality.com/case-studies
Our Culture
At Leegality, trust, ownership, transparency, and having fun while doing meaningful work are core to how we operate — not just values on paper. Our team rated us an incredible 97 eNPS for FY 2023–24 — the highest among 175+ startups surveyed.
We focus deeply on helping our people grow and stay motivated. Some of the perks you’ll enjoy:
- Flexible working hours
- Hybrid work setup
- Bi-annual performance appraisals
- A culture that rewards initiative, curiosity, and impact
If you're looking for a place where you can make a real difference while working with smart, driven, and genuinely nice people, welcome to Leegality.
Location: Hybrid
Job Brief:
- As a Machine Learning Engineer specializing in Computer Vision (CV) and Natural Language Processing (NLP), you will develop solutions to interesting technical problems, exploring exciting growth opportunities and having a real impact on our product, particularly focusing on document and content intelligence.
- To ensure success, you should demonstrate solid data science knowledge and experience in a related ML, CV, or NLP role. A first-class engineer will be someone whose expertise enhances our systems for document intelligence and content processing
Responsibilities:
- Designing machine learning systems, self-running artificial intelligence (AI) software, and specialized models for Computer Vision and Natural Language Processing applications.
- Transforming data science prototypes and applying appropriate deep learning algorithms and tools to text and image/document data.
- Solving complex CV and NLP problems with multi-layered data types, such as image/document classification, information extraction, semantic search, and object detection.
- Optimizing existing machine learning models, with a focus on high-performance model deployment for CV and NLP tasks.
- Developing ML algorithms (including large language models/LLMs and computer vision models) to analyze huge volumes of historical text, image, and document data to make predictions and automate workflows.
- Running tests, performing statistical analysis, and interpreting test results for CV/NLP model performance.
- Documenting machine learning processes, model architectures, and data pipelines.
- Keeping abreast of developments in machine learning, Computer Vision, and Natural Language Processing.
Requirements:
- 3+ years of relevant experience in Machine Learning Engineering, with a strong focus on Computer Vision and/or Natural Language Processing.
- Advanced proficiency with Python.
- Extensive knowledge of ML frameworks, libraries (e.g., PyTorch, Transformers), data structures, data modeling, and software architecture.
- Experience with building and maintaining scalable RESTful APIs (e.g., FastAPI).
- In-depth knowledge of mathematics, statistics, deep learning (CNNs, RNNs, Transformers), and algorithms.
- Superb analytical and problem-solving abilities, especially for unstructured data challenges.
- Great communication and collaboration skills.
- Excellent time management and organizational abilities.
- Experience with cloud platforms (e.g., AWS) for model deployment and MLOps.
Recruitment Process:
- Our hiring process combines AI-powered evaluations with structured interviews to ensure a fair and seamless experience.
- You will be contacted via email with the next steps upon being shortlisted.
- The process may include Assessments, AI-enabled interviews, and In-Person Interviews with our team.
- Final selection and CTC will be based on your overall performance and experience.
Apply directly through our career page: https://careers.leegality.com/jobs/Careers
For more information about us please visit our:
Our Company and Culture: https://bit.ly/3Iqm5SB
Our Website: www.leegality.com/
Our LinkedIn Page: www.linkedin.com/company/leegality/
Leegality's Privacy Notice: https://www.leegality.com/employee-privacy-notice
Must of Skills/Experience
• System Design
• Python
• TensorFlow
• Google ADK or Lang Graph
• Lang Chain , Lang Graph
• Spark
• Agentic AI Design
• ML Ops
• MCP (client and server)
• FastAPI
• Doc Factory
• RAG
• Golang
• LLMs – Gemini, Open AI
• NLP
• Dev Assistant - AI based code - generation
(Qwen or Claude or Copilot)
• CI/CD
• Good in oral and written communication,
collaboration and be a team player
Good to have skills
• DevOps with K8
• Scripting
• Java
• REST API
• UV
• ReACT
• DocFactory
• Unix






