Cutshort logo
For Employers
 Global Digital Transformation Solutions Provider logo
ML Engineer II - Aws, Aws Cloud
Global Digital Transformation Solutions Provider
ML Engineer II - Aws, Aws Cloud

ML Engineer II - Aws, Aws Cloud at Global Digital Transformation Solutions Provider · Pune · 6 - 12 years · ₹10L - ₹30L / yr · Posted 6 Nov 2025

Peak Hire Solutions's logo

ML Engineer II - Aws, Aws Cloud

at Global Digital Transformation Solutions Provider

Agency job
6 - 12 yrs
₹10L - ₹30L / yr
Pune
Skills
skill iconAmazon Web Services (AWS)
AWS CloudFormation
Amazon Redshift
skill iconElastic Search
ECS
skill iconDocker
skill iconKubernetes
skill iconMachine Learning (ML)
MLOps
Neo4J
SQL
Algorithms
Architecture
Statistical Modeling
Large Language Models (LLM)
skill iconDeep Learning

Job Details

- Job Title: ML Engineer II - Aws, Aws Cloud

- Industry: Technology

- Domain - Information technology (IT)

- Experience Required: 6-12 years

- Employment Type: Full Time

- Job Location: Pune

- CTC Range: Best in Industry


Job Description:

Core Responsibilities:

? The MLE will design, build, test, and deploy scalable machine learning systems, optimizing model accuracy and efficiency

? Model Development: Algorithms and architectures span traditional statistical methods to deep learning along with employing LLMs in modern frameworks.

? Data Preparation: Prepare, cleanse, and transform data for model training and evaluation.

? Algorithm Implementation: Implement and optimize machine learning algorithms and statistical models.

? System Integration: Integrate models into existing systems and workflows.

? Model Deployment: Deploy models to production environments and monitor performance.

? Collaboration: Work closely with data scientists, software engineers, and other stakeholders.

? Continuous Improvement: Identify areas for improvement in model performance and systems.


Skills:

? Programming and Software Engineering: Knowledge of software engineering best practices (version control, testing, CI/CD).

? Data Engineering: Ability to handle data pipelines, data cleaning, and feature engineering. Proficiency in SQL for data manipulation + Kafka, Chaossearch logs, etc for troubleshooting; Other tech touch points are ScyllaDB (like BigTable), OpenSearch, Neo4J graph

? Model Deployment and Monitoring: MLOps Experience in deploying ML models to production environments.

? Knowledge of model monitoring and performance evaluation.


Required experience:

? Amazon SageMaker: Deep understanding of SageMaker's capabilities for building, training, and deploying ML models; understanding of the Sagemaker pipeline with ability to analyze gaps and recommend/implement improvements

? AWS Cloud Infrastructure: Familiarity with S3, EC2, Lambda and using these services in


ML workflows

? AWS data: Redshift, Glue

? Containerization and Orchestration: Understanding of Docker and Kubernetes, and their implementation within AWS (EKS, ECS)


Skills: Aws, Aws Cloud, Amazon Redshift, Eks


Must-Haves

Aws, Aws Cloud, Amazon Redshift, Eks

NP: Immediate – 30 Days

 

Read more
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos

Similar jobs (10)

The Persona Labs
Bhilai, Raipur
0 - 4 yrs
₹7L - ₹13L / yr (ESOP available)
skill iconPython
PyTorch
TensorFlow
skill iconPostgreSQL
Retrieval Augmented Generation (RAG)
+11 more

ABOUT

The Persona Labs is building a new kind of social platform focused on something most social products do not explicitly optimize for: helping people become real friends.


We want to help people discover interesting people around them, find meaningful common ground, start low-pressure interactions, continue promising conversations, create shared experiences, and ultimately build real-life friendships.


DISCOVER → CURIOSITY → COMPATIBILITY → INTERACTION → UNDERSTAND → IRL EXPERIENCE → FRIENDSHIP


THE AI LAYER - COMPANION INTELLIGENCE

Alongside the platform, we are building a proactive personal AI companion that learns about the user and helps them navigate this journey through personalized recommendations, suggestions, reminders, conversations, and experiences. 


THE OPPORTUNITY

We are looking for a Founding ML Engineer to build the intelligence layer of the platform from the ground up. This is a 0→1 Applied AI / ML role where you will work directly with the founder and Product Engineer to turn ambiguous problems around users, relationships, recommendations and personal intelligence into working systems.


You will be expected to:

Understand the problem → identify the signals → design the intelligence system → prototype → evaluate → deploy → learn → improve.


WHAT YOU WILL BUILD & OWN


USER INTELLIGENCE

User representations, behavioural models, interests, preferences, contextual signals, and evolving understanding of the user. MEMORY Short- and long-term memory, episodic/preference/relationship memory, retrieval, relevance and updating.


RECOMMENDATION & MATCHING

People discovery, compatibility, activity/experience recommendations, and personalized ranking.


INTENT & INTEREST

Infer what the user is trying to do and learn what they care about from behaviour, not only declared interests.


RANKING

Decide what should appear first across potentially thousands of relevant people, activities or experiences.


CONTENT INTELLIGENCE

Classification, toxicity, spam, policy signals, quality, relevance, and semantic understanding.


RELATIONSHIP INTELLIGENCE

Reciprocity, interaction health, shared interests, progression, declining engagement and shared activity.


NEXT-BEST-ACTION

Determine the most useful action now: show a person, suggest a question, recommend an activity, reconnect, or do nothing.


TRUST / SAFETY INTELLIGENCE

Fake-account signals, spam, abuse, behavioural anomalies, risky interactions and moderation assistance.

COMPANION INTELLIGENCE

Use signals and outputs to help the companion decide what to say, suggest, recommend or not do. 


WHAT YOUR DAY-TO-DAY LOOKS LIKE

• Translate ambiguous product problems into ML/AI system designs.

• Build models and intelligence pipelines using behavioural, relational and contextual signals.

• Develop recommendation, matching and personalization systems.

• Design memory and retrieval systems that help the companion understand the user over time.

• Build and evaluate LLM-powered and agentic workflows.

• Decide when to use traditional ML, rules, retrieval, ranking or LLMs.

• Prototype quickly, test assumptions and iterate based on real user behaviour.

• Work closely with the founder and Product Engineer to turn intelligence into product experiences.

• Design APIs and production systems that bring ML/AI capabilities into the application.

• Build evaluation, monitoring and feedback loops so the intelligence improves over time.


WHO SHOULD APPLY

• Experience: 0–4 years’ experience, including exceptional fresh graduates. Strong 1–3 year engineers and experienced 3–4 year product builders are welcome.

• Strong foundations in ML, Python, statistics and software engineering.

• Evidence of Building: Experience with AI/ML projects, recommendation systems, LLM applications or personalization is highly valued.

• Strong evidence of building: Shipped projects, research, hackathons, internships, open source or startup work. 


WHAT WE LOOK FOR

MACHINE LEARNING DEPTH

Can you understand the modelling problem underneath the application?


RECOMMENDATION & PERSONALIZATION

Can you reason about relevance, ranking, cold start and behavioural signals?


AI ENGINEERING

Can you turn LLMs and agents into reliable product capabilities rather than simple API wrappers?


USER INTELLIGENCE

Can you design systems that gradually understand a person from sparse and changing signals?


SYSTEMS THINKING

Can you move from a model to a production system with APIs, data, latency, cost and monitoring?


EVALUATION MINDSET

Can you determine whether the intelligence actually helped the user?


PRODUCT JUDGMENT

Can you decide what the system should do when there is no predefined answer?


SPEED OF EXECUTION

Can you move from idea → prototype → evaluation → production quickly and responsibly?


BUILD WITH US

You will join at a stage where many of the answers do not exist yet. You will not simply implement a model someone else selected; you will help decide how the product learns to understand people.


CAREERS:

Apply with your resume, GitHub, portfolio or shipped work.

https://forms.gle/12YpUSBY2Sqs5xjp8

www.thepersonalabs.com

Read more
Quanteon Solutions
at Quanteon Solutions
1 recruiter
DurgaPrasad Sannamuri
Posted by DurgaPrasad Sannamuri
Hyderabad
2 - 5 yrs
₹5L - ₹25L / yr
Prompt engineering
skill iconMachine Learning (ML)
skill iconPython
Generative AI
Open-source LLMs
+15 more

Key Responsibilities:

  • Develop and deploy machine learning, deep learning, and NLP models for various business use cases.
  • Build end-to-end ML pipelines including data preprocessing, feature engineering, training, evaluation, and production deployment.
  • Optimize model performance and ensure scalability in production environments.
  • Work closely with data scientists, product teams, and engineers to translate business requirements into AI solutions.
  • Conduct data analysis to identify trends and insights.
  • Implement MLOps practices for versioning, monitoring, and automating ML workflows.
  • Research and evaluate new AI/ML techniques, tools, and frameworks.
  • Document system architecture, model design, and development processes.


Required Skills:

  • Strong programming skills in Python (NumPy, Pandas, Scikit-learn, TensorFlow, PyTorch, Keras).
  • Hands-on experience in building and deploying, finetuning ML/DL models in production.
  • Good understanding of machine learning algorithms, neural networks, NLP, and computer vision.
  • Experience with REST APIs, Docker, Kubernetes, and cloud platforms (AWS/GCP/Azure).
  • Working knowledge of MLOps tools such as MLflow, Airflow, DVC, or Kubeflow.
  • Familiarity with data pipelines and big data technologies (Spark, Hadoop) is a plus.
  • Strong analytical skills and ability to work with large datasets.
  • Excellent communication and problem-solving abilities.
  • Experience in deploying models using cloud services (AWS Sagemaker, GCP Vertex AI, etc.).
  • Experience in LLM fine-tuning or Generative AI, Voice AI, is an added advantage.


Educational Qualification:

  • Bachelor’s or Master’s degree in Computer Science, Data Science, AI, Machine Learning, IT, from IIT/NIT colleges strongly preferred
Read more
TalentOne HR Consulting LLP
Shivani Waghulkar
Posted by Shivani Waghulkar
Pune, Nagpur
8 - 12 yrs
₹15L - ₹28L / yr
skill iconMachine Learning (ML)
skill iconPython
Workflow
Prompt engineering

Greetings!

Hiring For Large Product Based Company!

Role- Mlops Engineer

Experience- 8-12 years

Location- Pune, Nagpur


JD-

  • 8-10 years of experience in DevOps, MLOps, Data Engineering, Software Engineering or Site Reliability Engineering  
  • Strong understanding of cloud infrastructure and experience working with at least one major cloud provider, preferably Azure  

Proficiency in at least one objected-oriented programming language, preferably python with hands-on experience in ml frameworks like TensorFlow, PyTorch or Scikit-learn  



Read more
Timble Technologies
at Timble Technologies
1 recruiter
Shefali Gupta
Posted by Shefali Gupta
Remote, Delhi, Gurugram, Noida, Ghaziabad, Faridabad, Bengaluru (Bangalore)
2 - 10 yrs
₹5L - ₹15L / yr
skill iconAmazon Web Services (AWS)
Google Cloud Platform (GCP)
skill iconDocker
API
skill iconFlask
+4 more

Job Title: Senior AI/ML Engineer

Company: Timble Technologies Pvt. Ltd

Location: Gurugram (Hybrid)

Experience: 2 TO 5 Years


About Us

Timble Glance is a high-growth AI RegTech and B2B SaaS company catering to top-tier BFSI and enterprise clients. We build cutting-edge systems powering 30+ high-scale APIs for digital identity verification, fraud detection, document intelligence, and compliance automation.

Role Overview

We are looking for a hands-on Senior AI/ML Engineer to design, develop, and productionize high-throughput AI/ML and Generative AI systems. You will own the full lifecycle—from problem formulation and data pipelines to deep learning architectures, RAG systems, LLMOps, and model governance—delivering sub-second latency and high reliability across our enterprise products.


Key Responsibilities


·       Model Architecture & Deployment: Design, train, and deploy production-scale ML/Deep Learning and GenAI systems (computer vision, document intelligence, OCR, NLP, fraud risk classification, and LLM applications).

·       GenAI & LLM Solutions: Develop robust LLM workflows including prompt engineering, fine-tuning, RAG pipelines, semantic search, vector indexing (Pinecone/Milvus/Chroma), and safety guardrails.

·       Pipelines & Engineering: Build performant feature extraction and data pipelines; write modular, vectorized, production-grade Python (NumPy, Pandas) and advanced SQL.

·       MLOps & Monitoring: Establish end-to-end MLOps/LLMOps standards—model registries, CI/CD, experiment tracking, drift detection, A/B testing, latency optimization, and cost governance.

·       Responsible AI & Security: Ensure model decisions comply with enterprise data security, privacy standards, and auditability required by the BFSI sector.

·       Collaboration & Ownership: Translate complex business requirements into technical roadmaps, conduct rigorous code reviews, and mentor junior engineers.


Required Qualifications & Skills


·       Education: B.Tech / M.Tech in Computer Science, AI/ML, Mathematics, or a related field—Tier-1 institutes (IIT, IIIT, NIT) strongly preferred.

·       Experience: 2+ years of hands-on experience developing, deploying, and maintaining ML/Deep Learning or GenAI models in production environments.

·       GenAI & NLP Stack: Hands-on experience with LLMs, embeddings, RAG architectures, and frameworks such as LangChain, LlamaIndex, or Hugging Face.

·       Deep Learning Frameworks: Strong proficiency in PyTorch or TensorFlow, with deep knowledge of transformer architectures and modern NLP/CV models.

·       Software & Data Engineering: Expert-level Python skills (pytest, Git, OOP, asynchronous programming), solid SQL proficiency, and familiarity with data workflows.

·       Deployment & Cloud: Practical exposure to cloud platforms (AWS/GCP), containerization (Docker), API frameworks (FastAPI/Flask), and basic orchestration (Kubernetes).


Preferred Qualifications

·       Prior domain experience in Fintech, RegTech, Identity Verification (KYC/AML), Fraud Intelligence, or B2B SaaS.

·       Experience optimizing models for low latency and inference cost (e.g., ONNX, TensorRT, model quantization).

·       Familiarity with workflow orchestrators such as Airflow, Prefect, or Kubeflow.

Read more
Smartsheet
Sandeep Selvan
Posted by Sandeep Selvan
Bengaluru (Bangalore)
4 - 12 yrs
Best in industry
MLOps
databricks
skill iconMachine Learning (ML)
MLFlow
LangGraph
+4 more

For over 20 years, Smartsheet has empowered teams to manage work seamlessly and scale solutions smarter. Now, in our most ambitious chapter yet, we are uniting human teams with AI agents. By orchestrating the work agents do best, automating manual tasks and uncovering insights at scale, we create the space for people to focus on what truly matters: judgment, creativity, and big thinking. That is magic at work, and it’s what we show up for every day.


Our India Global Capability Center isn't just supporting global operations—we’re leading global innovation. After scaling rapidly into a best-in-class hub, we deliver the product innovation and enterprise capabilities that accelerate our global growth, profitability, and scale. As we expand Smartsheet India, we’re searching for Senior AI/ML Ops Engineers who crave variety and ownership. You’ll have the opportunity to work across multiple teams and disciplines, building a versatile skillset while solving the complex challenges of a global platform.


You Will:

  • Designing, Developing and overseeing the strategy and architecture of scalable and reliable AI/ML Ops platforms / pipelines
  • Model Deployment: Package and deploy AI/ML services to production, ensuring they are reproducible and interpretable
  • CI/CD Pipeline Development: Design and implement automated CI/CD (Continuous Integration/Continuous Deployment) pipelines to accelerate model deployment using tools
  • Infrastructure Management: Provision and optimize infrastructure for training and serving, utilizing Docker, Kubernetes, or serverless platforms
  • Monitoring & Observability : Implement post-deployment monitoring for model performance, data drift, and latency using tools. Experience in Monte Carlo is preferable
  • Automation: Automate retraining and data pipeline workflows to ensure models stay accurate over time.
  • Manage the deployment of foundation models, fine-tuning workflows, and Retrieval-Augmented Generation (RAG) stacks (Vector DBs, Knowledge Graph. Experience with AWS Bedrock is preferable
  • Resource Optimization: Manage GPU/CPU utilization to minimize cloud costs while maintaining low-latency inference for users
  • Collaboration: Work closely with data scientists, data engineers, and software engineers to bridge the gap between model development and production.
  • Version Control & Governance: Manage versioning for data, code, and models using tools like MLflow.
  • Security & Compliance: Implementing data security measures, ensuring compliance with data governance policies, and protecting sensitive data
  • Technology Evaluation and Innovation: Staying abreast of emerging data technologies and exploring opportunities for innovation to improve the organisation’s data infrastructure
  • Troubleshooting and Problem Solving: Diagnosing and resolving complex data-related issues, ensuring the stability and reliability of the data platform
  • Perform other duties as assigned


You Have:

  • Enterprise SaaS software solutions with high availability and scalability
  • Solution handling large scale structured and unstructured data from varied data sources
  • Experience in building and maintaining AI/ML Ops platform systems ensuring scalability, reliability, efficiency and security
  • Working with Product engineering team to influence designs with data, AI and analytics use cases in mind
  • In depth experience in System design, AI/ML Frameworks and tools involving large Petabytes of data with Databricks Lakehouse ecosystem
  • AI/MLOps workflows on Databricks , MLFlow, Mosaic AI Agent Framework, Unity Catalog, Vector Search, Knowledge Graph
  • Knowledge of AI/ML frameworks like LangChain, LangGraph for AI/ML Ops pipeline integration
  • Cloud Platforms: Hands-on experience with at least one major cloud provider (AWS, Azure, or GCP). Experience in AWS hosted data platform is preferable
  • Programming languages like Python and SQL
  • Modern software engineering practices like Kubernetes, CI/CD, IAC tools (Preferably Terraform), Observability, monitoring and alerting
  • Solution Cost Optimisations and design to cost
  • Legally eligible to work in India on an ongoing basis

 

Get to Know Us:

At Smartsheet, your ideas are heard, your potential is supported, and your contributions have real impact. You’ll have the freedom to explore, push boundaries, and grow beyond your role. We welcome diverse perspectives and nontraditional paths—because we know that impact comes from individuals who care deeply and challenge thoughtfully. When you’re doing work that stretches you, excites you, and connects you to something bigger, that’s magic at work. Let’s build what’s next, together.


Equal Opportunity Employer:

Smartsheet is an Equal Opportunity (EEO) employer committed to fostering an inclusive environment with the best employees. It is our policy to provide equal employment opportunities to all qualified applicants in accordance with applicable laws in the US, UK, Australia, Germany, Costa Rica, Japan, Bulgaria, India, and Singapore. All qualified applicants will receive consideration without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, age, protected veteran or disabled status, or genetic information. 

If there are preparations we can make to help ensure you have a comfortable and positive interview experience, please let us know.



Job application link : https://grnh.se/z7qx2ehx1us

Read more
Appiness Interactive Pvt. Ltd.
S Suriya Kumar
Posted by S Suriya Kumar
Navi Mumbai
2 - 6 yrs
₹2L - ₹35L / yr
Artificial Intelligence (AI)
skill iconMachine Learning (ML)

Role Overview

We are looking for an AI/ML Engineer with 2 - 6 years of industry experience to develop and deploy machine learning solutions for industrial and smart manufacturing use cases. The role involves working with real-world industrial data, developing predictive and analytical models, and translating business and operational requirements into scalable AI/ML solutions.

The ideal candidate should have strong fundamentals in machine learning, Python, data processing, model development, and deployment, along with an interest in applying AI to manufacturing, industrial automation, and operational optimization.


Key Responsibilities

  • Design, develop, train, and evaluate machine learning models for industrial and manufacturing use cases.
  • Analyze large and complex datasets to identify patterns, trends, anomalies, and opportunities for optimization.
  • Perform data preprocessing, feature engineering, model selection, and performance evaluation.
  • Develop solutions for use cases such as predictive maintenance, anomaly detection, quality prediction, process optimization, forecasting, and equipment monitoring.
  • Work with structured and time-series data generated from industrial equipment, machines, sensors, and operational systems.
  • Develop and optimize ML pipelines for data preparation, model training, validation, and deployment.
  • Collaborate with domain experts, data engineers, software engineers, and business stakeholders to understand requirements and translate them into technical solutions.
  • Deploy machine learning models into production environments and monitor model performance.
  • Troubleshoot model and data-related issues and continuously improve model accuracy, reliability, and scalability.
  • Develop reusable code, APIs, and components to integrate ML models with enterprise and industrial applications.
  • Document models, methodologies, experiments, results, and technical implementations.
  • Stay updated with emerging AI/ML techniques, industrial AI trends, and smart manufacturing technologies.


Required Technical Skills

  • Strong programming experience in Python.
  • Good understanding of Machine Learning algorithms and concepts.
  • Experience with libraries/frameworks such as Scikit-learn, Pandas, NumPy, and preferably TensorFlow or PyTorch.
  • Strong understanding of data preprocessing, feature engineering, model training, validation, and evaluation.
  • Experience working with time-series data is preferred.
  • Good understanding of statistical concepts and data analysis.
  • Experience with SQL and relational databases.
  • Understanding of model deployment and productionization of ML solutions.
  • Familiarity with Git and software development best practices.
  • Good problem-solving and analytical skills.


Preferred Skills

  • Experience in Industrial AI, Smart Manufacturing, Industry 4.0, or Industrial IoT (IIoT).
  • Experience with predictive maintenance, anomaly detection, forecasting, or quality inspection.
  • Exposure to sensor data, machine/equipment data, telemetry, or real-time industrial data.
  • Knowledge of computer vision for manufacturing or quality inspection use cases.
  • Exposure to cloud platforms such as AWS, Azure, or GCP.
  • Familiarity with Docker, Kubernetes, or CI/CD for ML deployment.
  • Knowledge of MLOps concepts and tools.
  • Exposure to Generative AI/LLMs is an added advantage. 
Read more
Hiring for Top Product based company
Hiring for Top Product based company
Agency job
via TalentOne HR Consulting LLP by Manasi Chavan
Pune
7.5 - 12 yrs
₹25L - ₹29L / yr
skill iconMachine Learning (ML)
MLOps
skill iconPython
skill iconAmazon Web Services (AWS)
SQL Azure
+1 more
  • Bachelor’s degree in computer science, Data Science, Information Systems, or a related field  
  • 8-10 years of experience in DevOps, MLOps, Data Engineering, Software Engineering or Site Reliability Engineering  
  • Strong understanding of cloud infrastructure and experience working with at least one major cloud provider, preferably Azure  
  • Proficiency in at least one objected-oriented programming language, preferably python with hands-on experience in ml frameworks like TensorFlow, PyTorch or Scikit-learn  


Read more
Jumio
at Jumio
1 recruiter
Agency job
via Talentojcom by Bhavesh Kiroula
Remote only
3 - 15 yrs
₹15L - ₹42L / yr
MLOps
Aws sagemaker
skill iconAmazon Web Services (AWS)
MLFlow
Systems Development Life Cycle (SDLC)
+3 more

Example Responsibilities:

  • Build and optimize model serving infrastructure with a focus on inference latency and cost optimization
  • Architect efficient inference pipelines that balance latency, throughput, and cost across various acceleration options
  • Develop monitoring and observability solutions for ML systems
  • Collaborate with ML Engineers to establish best practices for optimized model deployment
  • Implement cost-efficient, enterprise-scale solutions
  • Collaborate in a cross-functional, distributed team for continuous system improvement
  • Work with MLEs, QA Engineers, and DevOps Engineers
  • Evaluate and implement new technologies and tools
  • Contribute to architectural decisions for distributed ML systems


Experience and Qualifications:

  • 5+ years of experience in software engineering with Python
  • Experience with ML frameworks, particularly PyTorch
  • Experience optimizing ML models with hardware acceleration (AWS Neuron , ONNX, TensorRT)
  • Experience with AWS ML services and hardware-accelerated instances (Sagemaker, Inferentia,Trainium)
  • Proven experience building and operating AWS serverless architectures
  • Deep understanding of event-driven processing patterns, SQS/SNS and serverless caching solutions
  • Experience with containerization using Docker and orchestration tools
  • Strong knowledge of RESTful API design and implementation
  • Proficiency in writing good quality & secure code and be familiar with static code analysis tools
  • Excellent analytical, conceptual and communication skills in spoken and written English
  • Experience applying Computer Science fundamentals in algorithm design, problem solving, and complexity analysis


Great to have Experience and Qualifications:

  • Experience with any of the following: model compilation and quantization, performance profiling and benchmarking ML inference systems
  • Experience working in regulated industries with strict compliance requirements for cloud-native solutions
Read more
Egnyte
at Egnyte
4 recruiters
Bhavana Kapalganti
Posted by Bhavana Kapalganti
Remote only
5 - 10 yrs
Best in industry
LoRA / QLoRA
skill iconPython
SLM
Large Language Models (LLM)
PyTorch

EGNYTE YOUR CAREER. SPARK YOUR PASSION.


Egnyte is a place where we spark opportunities for amazing people. We believe that every role has meaning, and every Egnyter should be respected. With 23,000 customers worldwide and growing, you can make an impact by protecting their valuable data. When joining Egnyte, you’re not just landing a new career; you become part of a team of Egnyters who are doers, thinkers, and collaborators who embrace and live by our values:


Invested Relationships


Fiscal Prudence


Candid Conversations

 

ABOUT EGNYTE


Egnyte is the secure multi-cloud platform for content security and governance that enables organizations to better protect and collaborate on their most valuable content. Established in 2008, Egnyte has democratized cloud content security for more than 23,000 organizations, helping customers improve data security, maintain compliance, prevent and detect ransomware threats, and boost employee productivity on any app, any cloud, anywhere.

 

WHAT YOU’LL DO: 


  • Fine-tune and train SLMs using Hugging Face, TRL, and adapter methods (LoRA, QLoRA, PEFT)
  • Optimize models for inference via quantization, pruning, and knowledge distillation
  • Deploy models to edge devices, mobile, and local servers with strict latency targets
  • Build end-to-end MLOps pipelines from data ingestion to deployment
  • Monitor model accuracy, latency, and hardware utilization in production
  • Evaluate model quality using benchmarking frameworks and custom evaluation suites


YOUR QUALIFICATIONS:


  • SLM Development & Fine-tuning: Train and fine-tune SLMs using Hugging Face and Knowledge on Adaptors.
  • Model Optimization: Apply quantization, pruning, knowledge distillation, and optimization for lightweight, efficient models.
  • Edge Deployment: Deploy models to edge devices, mobile, and local servers, etc.
  • Pipeline Engineering: Build end-to-end MLOps pipelines — from data ingestion to deployment.
  • Performance Monitoring: Track model accuracy, latency, and CPU/GPU usage in production.


Good to have


  • Deployment experience on edge or mobile environments
  • Knowledge of ONNX export and cross-platform inference
  • MLOps tooling — experiment tracking, model registries, CI/CD for ML


EQUAL EMPLOYMENT OPPORTUNITY


At Egnyte, we celebrate our unique differences and thrive on our diversity for our employees, our products, our customers, our investors, and our communities. Our global Egnyte Employee Communities (EECs) support representation and inclusion across our diverse workplace. Egnyters are encouraged to bring their whole selves to work and to appreciate the many differences that collectively make Egnyte a higher-performing company and a great place to be.


Egnyte will not allow any form of retaliation against employees who raise issues of equal employment opportunity. To ensure the workplace is free of artificial barriers, violation of this policy including any improper retaliatory conduct will lead to discipline, up to and including discharge. All employees must cooperate with all investigations conducted pursuant to this policy.

Read more
Einfochips
Arpita Pathak
Posted by Arpita Pathak
Indore, Pune, Ahmedabad
4 - 6 yrs
₹7L - ₹10L / yr
skill iconPython
skill iconMachine Learning (ML)
Artificial Intelligence (AI)
Generative AI
Large Language Models (LLM) tuning
+5 more

Experience - 4 to 6 year

Location – Ahmedabad/Pune/Indore

  • Additional Job Description

Additional Job Description

Required Skills and Experience: 

  • Strong proficiency in Python and experience with ML/AI libraries (scikit-learn, TensorFlow, PyTorch, Hugging Face ecosystem).
  • Hands-on experience with LLMs, RAG, vector databases, and retrieval pipelines.
  • Practical experience deploying agentic workflows and building multi-step, tool-enabled agents.
  • Experience using Garak (or similar LLM red-teaming/vulnerability scanners) to identify model weaknesses and harden deployments.
  • Demonstrated experience implementing content filtering / moderation systems.
  • Solid skills working with structured and unstructured data and advanced feature engineering.
  • Familiarity with cloud GenAI platforms and services (Azure AI Services preferred; AWS/GCP acceptable).
  • Experience building APIs/microservices; containerization (Docker), orchestration (Kubernetes).
  • Strong understanding of model evaluation, performance profiling, inference cost optimization, and observability.
  • Good knowledge of security, data governance, and privacy best practices for AI systems.


Read more
Why apply to jobs via Cutshort
people_solving_puzzle
Personalized job matches
Stop wasting time. Get matched with jobs that meet your skills, aspirations and preferences.
people_verifying_people
Verified hiring teams
See actual hiring teams, find common social connections or connect with them directly.
ai_chip
Move faster with AI
We use AI to get you faster responses, recommendations and unmatched user experience.
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos