Cutshort logo
For Employers
Deqode logo
Azure Machine Learning
Azure Machine Learning

Azure Machine Learning at Deqode · Remote only · 5 - 8 years · ₹20L - ₹24L / yr · Bootstrapped · Remote only · Posted 20 Feb 2026

Deqode's logo

Azure Machine Learning

Shubham Das's profile picture
Posted by Shubham Das
5 - 8 yrs
₹20L - ₹24L / yr
Remote only
Skills
skill iconMachine Learning (ML)
Windows Azure
Microsoft Visual Studio
  1. Strong experience in Azure – mainly Azure ML Studio, AKS, Blob Storage, ADF, ADO Pipelines.
  2. Ability and experience to register and deploy ML/AI/GenAI models via Azure ML Studio.
  3. Working knowledge of deploying models in AKS clusters.
  4. Design and implement data processing, training, inference, and monitoring pipelines using Azure ML.
  5. Excellent Python skills – environment setup and dependency management, coding as per best practices, and knowledge of automatic code review tools like linting and Black.
  6. Experience with MLflow for model experiments, logging artifacts and models, and monitoring.
  7. Experience in orchestrating machine learning pipelines using MLOps best practices.
  8. Experience in DevOps with CI/CD knowledge (Git in Azure DevOps).
  9. Experience in model monitoring (drift detection and performance monitoring).
  10. Fundamentals of data engineering.
  11. Docker-based deployment is good to have.


Read more
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos

About Deqode

Founded :
2016
Type :
Products & Services
Size :
100-1000
Stage :
Bootstrapped

About

At Deqode, our purpose is to help businesses solve complex problems using new-age technologies. We provide enterprise blockchain solutions to businesses.

Read more

Connect with the team

Profile picture
Mohini Bansal

Company social profiles

bloglinkedintwitterfacebook

Similar jobs (10)

company logo
Sandeep Selvan
Posted by Sandeep Selvan
Bengaluru (Bangalore)
4 - 12 yrs
Best in industry
MLOps
databricks
skill iconMachine Learning (ML)
MLFlow
LangGraph
+4 more

For over 20 years, Smartsheet has empowered teams to manage work seamlessly and scale solutions smarter. Now, in our most ambitious chapter yet, we are uniting human teams with AI agents. By orchestrating the work agents do best, automating manual tasks and uncovering insights at scale, we create the space for people to focus on what truly matters: judgment, creativity, and big thinking. That is magic at work, and it’s what we show up for every day.


Our India Global Capability Center isn't just supporting global operations—we’re leading global innovation. After scaling rapidly into a best-in-class hub, we deliver the product innovation and enterprise capabilities that accelerate our global growth, profitability, and scale. As we expand Smartsheet India, we’re searching for Senior AI/ML Ops Engineers who crave variety and ownership. You’ll have the opportunity to work across multiple teams and disciplines, building a versatile skillset while solving the complex challenges of a global platform.


You Will:

  • Designing, Developing and overseeing the strategy and architecture of scalable and reliable AI/ML Ops platforms / pipelines
  • Model Deployment: Package and deploy AI/ML services to production, ensuring they are reproducible and interpretable
  • CI/CD Pipeline Development: Design and implement automated CI/CD (Continuous Integration/Continuous Deployment) pipelines to accelerate model deployment using tools
  • Infrastructure Management: Provision and optimize infrastructure for training and serving, utilizing Docker, Kubernetes, or serverless platforms
  • Monitoring & Observability : Implement post-deployment monitoring for model performance, data drift, and latency using tools. Experience in Monte Carlo is preferable
  • Automation: Automate retraining and data pipeline workflows to ensure models stay accurate over time.
  • Manage the deployment of foundation models, fine-tuning workflows, and Retrieval-Augmented Generation (RAG) stacks (Vector DBs, Knowledge Graph. Experience with AWS Bedrock is preferable
  • Resource Optimization: Manage GPU/CPU utilization to minimize cloud costs while maintaining low-latency inference for users
  • Collaboration: Work closely with data scientists, data engineers, and software engineers to bridge the gap between model development and production.
  • Version Control & Governance: Manage versioning for data, code, and models using tools like MLflow.
  • Security & Compliance: Implementing data security measures, ensuring compliance with data governance policies, and protecting sensitive data
  • Technology Evaluation and Innovation: Staying abreast of emerging data technologies and exploring opportunities for innovation to improve the organisation’s data infrastructure
  • Troubleshooting and Problem Solving: Diagnosing and resolving complex data-related issues, ensuring the stability and reliability of the data platform
  • Perform other duties as assigned


You Have:

  • Enterprise SaaS software solutions with high availability and scalability
  • Solution handling large scale structured and unstructured data from varied data sources
  • Experience in building and maintaining AI/ML Ops platform systems ensuring scalability, reliability, efficiency and security
  • Working with Product engineering team to influence designs with data, AI and analytics use cases in mind
  • In depth experience in System design, AI/ML Frameworks and tools involving large Petabytes of data with Databricks Lakehouse ecosystem
  • AI/MLOps workflows on Databricks , MLFlow, Mosaic AI Agent Framework, Unity Catalog, Vector Search, Knowledge Graph
  • Knowledge of AI/ML frameworks like LangChain, LangGraph for AI/ML Ops pipeline integration
  • Cloud Platforms: Hands-on experience with at least one major cloud provider (AWS, Azure, or GCP). Experience in AWS hosted data platform is preferable
  • Programming languages like Python and SQL
  • Modern software engineering practices like Kubernetes, CI/CD, IAC tools (Preferably Terraform), Observability, monitoring and alerting
  • Solution Cost Optimisations and design to cost
  • Legally eligible to work in India on an ongoing basis

 

Get to Know Us:

At Smartsheet, your ideas are heard, your potential is supported, and your contributions have real impact. You’ll have the freedom to explore, push boundaries, and grow beyond your role. We welcome diverse perspectives and nontraditional paths—because we know that impact comes from individuals who care deeply and challenge thoughtfully. When you’re doing work that stretches you, excites you, and connects you to something bigger, that’s magic at work. Let’s build what’s next, together.


Equal Opportunity Employer:

Smartsheet is an Equal Opportunity (EEO) employer committed to fostering an inclusive environment with the best employees. It is our policy to provide equal employment opportunities to all qualified applicants in accordance with applicable laws in the US, UK, Australia, Germany, Costa Rica, Japan, Bulgaria, India, and Singapore. All qualified applicants will receive consideration without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, age, protected veteran or disabled status, or genetic information. 

If there are preparations we can make to help ensure you have a comfortable and positive interview experience, please let us know.



Job application link : https://grnh.se/z7qx2ehx1us

Read more
company logo
Bengaluru (Bangalore), Delhi, Gurugram, Noida, Ghaziabad, Faridabad, Chennai
4 - 15 yrs
₹30L - ₹40L / yr
skill iconMachine Learning (ML)
Natural Language Processing (NLP)
Generative AI
skill iconPython
Scikit-Learn
+4 more

About the Role

 

We are looking for a highly skilled Data Scientist with strong expertise in Machine Learning, MLOps, and Generative AI. The ideal candidate will have hands-on experience in building scalable ML models, deploying them in production, and working with modern AI frameworks, including GenAI technologies.

 

 

 

Key Responsibilities

 

·      Design, develop, and deploy machine learning models for real-world business problems

·      Work on end-to-end ML lifecycle: data preprocessing, model building, evaluation, deployment, and monitoring

·      Implement and manage MLOps pipelines for scalable and reproducible workflows

·      Utilize tools like MLflow for experiment tracking, model versioning, and lifecycle management

·      Develop and integrate Generative AI (GenAI) solutions such as LLM-based applications

·      Collaborate with cross-functional teams (engineering, product, business) to translate requirements into AI solutions

·      Optimize model performance and ensure production stability

·      Stay updated with the latest advancements in AI/ML and GenAI ecosystems

 

 

 

Required Skills & Qualifications

 

·      4+ years of experience in Data Science / Machine Learning

·      Strong programming skills in Python

·      Hands-on experience with ML modeling techniques (supervised, unsupervised, NLP, etc.)

·      Solid understanding of MLOps practices and tools

·      Experience with MLflow or similar model lifecycle tools 

·      Practical experience in Generative AI (GenAI), including working with LLMs

·      Experience with libraries/frameworks like Scikit-learn, TensorFlow, PyTorch

·      Strong understanding of data structures, algorithms, and statistics

·      Experience with cloud platforms (AWS/GCP/Azure) is a plus


Good to Have

 

·      Experience with LLM fine-tuning, prompt engineering, or RAG pipelines

·      Exposure to Docker, Kubernetes, and CI/CD pipelines

·      Knowledge of data engineering workflows 



Read more
company logo
Akash Bhatt
Posted by Akash Bhatt
Delhi, Gurugram, Noida, Ghaziabad, Faridabad, Hyderabad, Bengaluru (Bangalore), Mumbai, Pune, Chennai
2 - 12 yrs
₹8L - ₹38L / yr
MLFlow
kubeflow
Deployment management
skill iconMachine Learning (ML)

We are looking for an MLOps Engineer to take ML models from notebook to production reliably.


Responsibilities

  • Build ML training and deployment pipelines
  • Track experiments and models with MLflow
  • Run pipelines on Kubeflow or SageMaker
  • Monitor model drift and performance


Requirements

  • 2+ years in MLOps or ML engineering
  • Hands-on with MLflow and Kubeflow or SageMaker
  • Experience serving models at scale


Read more
company logo
Shefali Gupta
Posted by Shefali Gupta
Remote, Delhi, Gurugram, Noida, Ghaziabad, Faridabad, Bengaluru (Bangalore)
2 - 10 yrs
₹5L - ₹15L / yr
skill iconAmazon Web Services (AWS)
Google Cloud Platform (GCP)
skill iconDocker
API
skill iconFlask
+4 more

Job Title: Senior AI/ML Engineer

Company: Timble Technologies Pvt. Ltd

Location: Gurugram (Hybrid)

Experience: 2 TO 5 Years


About Us

Timble Glance is a high-growth AI RegTech and B2B SaaS company catering to top-tier BFSI and enterprise clients. We build cutting-edge systems powering 30+ high-scale APIs for digital identity verification, fraud detection, document intelligence, and compliance automation.

Role Overview

We are looking for a hands-on Senior AI/ML Engineer to design, develop, and productionize high-throughput AI/ML and Generative AI systems. You will own the full lifecycle—from problem formulation and data pipelines to deep learning architectures, RAG systems, LLMOps, and model governance—delivering sub-second latency and high reliability across our enterprise products.


Key Responsibilities


·       Model Architecture & Deployment: Design, train, and deploy production-scale ML/Deep Learning and GenAI systems (computer vision, document intelligence, OCR, NLP, fraud risk classification, and LLM applications).

·       GenAI & LLM Solutions: Develop robust LLM workflows including prompt engineering, fine-tuning, RAG pipelines, semantic search, vector indexing (Pinecone/Milvus/Chroma), and safety guardrails.

·       Pipelines & Engineering: Build performant feature extraction and data pipelines; write modular, vectorized, production-grade Python (NumPy, Pandas) and advanced SQL.

·       MLOps & Monitoring: Establish end-to-end MLOps/LLMOps standards—model registries, CI/CD, experiment tracking, drift detection, A/B testing, latency optimization, and cost governance.

·       Responsible AI & Security: Ensure model decisions comply with enterprise data security, privacy standards, and auditability required by the BFSI sector.

·       Collaboration & Ownership: Translate complex business requirements into technical roadmaps, conduct rigorous code reviews, and mentor junior engineers.


Required Qualifications & Skills


·       Education: B.Tech / M.Tech in Computer Science, AI/ML, Mathematics, or a related field—Tier-1 institutes (IIT, IIIT, NIT) strongly preferred.

·       Experience: 2+ years of hands-on experience developing, deploying, and maintaining ML/Deep Learning or GenAI models in production environments.

·       GenAI & NLP Stack: Hands-on experience with LLMs, embeddings, RAG architectures, and frameworks such as LangChain, LlamaIndex, or Hugging Face.

·       Deep Learning Frameworks: Strong proficiency in PyTorch or TensorFlow, with deep knowledge of transformer architectures and modern NLP/CV models.

·       Software & Data Engineering: Expert-level Python skills (pytest, Git, OOP, asynchronous programming), solid SQL proficiency, and familiarity with data workflows.

·       Deployment & Cloud: Practical exposure to cloud platforms (AWS/GCP), containerization (Docker), API frameworks (FastAPI/Flask), and basic orchestration (Kubernetes).


Preferred Qualifications

·       Prior domain experience in Fintech, RegTech, Identity Verification (KYC/AML), Fraud Intelligence, or B2B SaaS.

·       Experience optimizing models for low latency and inference cost (e.g., ONNX, TensorRT, model quantization).

·       Familiarity with workflow orchestrators such as Airflow, Prefect, or Kubeflow.

Read more
Hiring for IT Consulting Firm (MNC)
Hiring for IT Consulting Firm (MNC)
Agency job
via by Sneha k
Pune, Nagpur
5 - 10 yrs
₹20L - ₹30L / yr
MLOps
DevOps
Artificial Intelligence (AI)
skill iconData Science
Data engineering
+5 more

Position Overview 

The AI Observability Engineer will be instrumental in implementation of scalable, cloud-native solutions to meet the growing needs of our Data & Development team. The successful candidate will demonstrate the ability to abstract complexity and create reusable, scalable patterns that accelerate development. The AI Observability Engineer will build and maintain a robust framework to ensure the reliability and maintainability of DPR Construction's complex AI systems. 

 

Responsibilities 

  • Standardize observability practices across AI/ML and other development teams including logging, metrics, tracing, and model performance monitoring, ingesting data from multiple platforms 
  • Lead hands-on implementation of automation-first DevOps and MLOps practices, enabling infrastructure-as-code and consistent, repeatable environment provisioning 
  • Design and manage intelligent DataOps pipelines with automated data quality monitoring and anomaly detection 
  • Deploy, maintain and monitor containerized ML workloads 
  • Extend existing CI/CD pipelines to support automated infrastructure changes and ML workflows 
  • Implement AI-driven data validation, schema and concept drift detection and metadata management. 
  • Establish governance frameworks for AI systems, including bias detection, explainability, and auditability 
  • Extend existing Azure RBAC strategy by automating role and permission management to reduce manual intervention 
  • Develop automated test suites for model performance, regression, edge cases and bias validation 
  • Monitor model KPIs (accuracy, precision, recall, latency, calibration) 
  • Ensure reproducability of experiments and production models 
  • Act as a technical point of contact for DevOps and MLOps practices, developing reusable patterns, documentation, and proof-of-concepts to drive adoption 

Qualifications 

  • Bachelor’s degree in computer science, Data Science, Information Systems, or a related field 
  • 5+ years of experience in DevOps, MLOps, Data Engineering, Software Engineering or Site Reliability Engineering 
  • Strong understanding of cloud infrastructure and experience working with at least one major cloud provider, preferably Azure 
  • Proficiency in at least one objected-oriented programming language, preferably python with hands-on experience in ml frameworks like TensorFlow, PyTorch or Scikit-learn 
Read more
company logo
Arpita Pathak
Posted by Arpita Pathak
Indore, Pune, Ahmedabad
4 - 6 yrs
₹7L - ₹10L / yr
skill iconPython
skill iconMachine Learning (ML)
Artificial Intelligence (AI)
Generative AI
Large Language Models (LLM) tuning
+5 more

Experience - 4 to 6 year

Location – Ahmedabad/Pune/Indore

  • Additional Job Description

Additional Job Description

Required Skills and Experience: 

  • Strong proficiency in Python and experience with ML/AI libraries (scikit-learn, TensorFlow, PyTorch, Hugging Face ecosystem).
  • Hands-on experience with LLMs, RAG, vector databases, and retrieval pipelines.
  • Practical experience deploying agentic workflows and building multi-step, tool-enabled agents.
  • Experience using Garak (or similar LLM red-teaming/vulnerability scanners) to identify model weaknesses and harden deployments.
  • Demonstrated experience implementing content filtering / moderation systems.
  • Solid skills working with structured and unstructured data and advanced feature engineering.
  • Familiarity with cloud GenAI platforms and services (Azure AI Services preferred; AWS/GCP acceptable).
  • Experience building APIs/microservices; containerization (Docker), orchestration (Kubernetes).
  • Strong understanding of model evaluation, performance profiling, inference cost optimization, and observability.
  • Good knowledge of security, data governance, and privacy best practices for AI systems.


Read more
Insurance expertise
Insurance expertise
Agency job
via by Priyanka Bisht
Gurugram, Noida
5 - 9 yrs
Best in industry
skill iconPython
"AIML
skill iconMachine Learning (ML)
Artificial Intelligence (AI)
MLOps
+2 more

Job Summary/ Job Opportunity:

This is an excellent opportunity for an ideal candidate with a high level of technical proficiency and meeting the below mentioned criteria -- • Strong experience in Machine Learning, Deep Learning, Generative AI, and Large Language Models (LLMs). • Hands-on experience building and deploying production-grade solutions using Azure OpenAI, OpenAI, LangChain, LangGraph, Semantic Kernel, LlamaIndex, and Agentic AI frameworks. • Strong expertise in Python, API development, microservices, and cloud-native architectures. • Experience designing and implementing RAG solutions, vector databases, embeddings, knowledge retrieval systems, and AI copilots. • Experience with Azure cloud services, MLOps, CI/CD pipelines, monitoring, and model lifecycle management. • Strong understanding of AI governance, responsible AI, security, compliance, and model evaluation frameworks. • Ability to lead technical discussions, provide architectural recommendations, mentor team members, and interact with business stakeholde


Key Objectives and Major Responsibilities:

• Design, develop, and implement scalable AI/ML and Generative AI solutions for enterprise applications. • Lead development of intelligent applications leveraging LLMs, RAG pipelines, AI agents, and document intelligence solutions. • Collaborate with business stakeholders, architects, and product teams to translate business requirements into technical solutions. • Design and optimize data pipelines, vector search solutions, embeddings, and retrieval mechanisms. • Build and maintain REST APIs, microservices, and cloud-native AI applications. • Ensure best practices in coding standards, performance optimization, security, scalability, and maintainability. • Drive AI solution deployment using MLOps practices, CI/CD pipelines, monitoring, and observability frameworks. • Perform code reviews, mentor junior developers, and contribute to capability building within the team


Key Capabilities and Competencies:

Knowledge, Skills, Qualification and Experience

• Degree in B.Tech/M.Tech (Computer Science/IT/Data Science) or related discipline preferred, with 3–4 years of relevant experience in AI/ML, GenAI and total 5-7 years of experience. • Proficiency in Python and hands-on experience with ML libraries (scikit-learn, TensorFlow, PyTorch) and GenAI frameworks/tools. • Strong understanding of machine learning, deep learning, LLMs, prompt engineering, and techniques like RAG and fine-tuning. • Experience with data processing, embeddings, vector databases, APIs, and building scalable AI driven applications. • Good communication skills, ability to work on multiple projects, and eagerness to learn and adapt to evolving AI technologies. 

Read more
company logo
Remote only
5 - 10 yrs
Best in industry
skill iconPython
SQL
skill iconMachine Learning (ML)
databricks
Apache Airflow
+1 more

Description

We’re seeking a highly skilled, execution-focused Senior Data Scientist with a minimum of 5 years of experience. This role demands hands-on expertise in building, deploying, and optimizing machine learning models at scale, while working with big data technologies and modern cloud platforms. You will be responsible for driving data-driven solutions from experimentation to production, leveraging advanced tools and frameworks across Python, SQL, Spark, and AWS. The role requires strong technical depth, problem-solving ability, and ownership in delivering business impact through data science.


Responsibilities

  • Design, build, and deploy scalable machine learning models into production systems.
  • Develop advanced analytics and predictive models using Python, SQL, and popular ML/DL frameworks (Pandas, Scikit-learn, TensorFlow, PyTorch).
  • Leverage Databricks, Apache Spark, and Hadoop for large-scale data processing and model training.
  • Implement workflows and pipelines using Airflow and AWS EMR for automation and orchestration.
  • Collaborate with engineering teams to integrate models into cloud-based applications on AWS.
  • Optimize query performance, storage usage, and data pipelines for efficiency.
  • Conduct end-to-end experiments, including data preprocessing, feature engineering, model training, validation, and deployment.
  • Drive initiatives independently with high ownership and accountability.
  • Stay up to date with industry best practices in machine learning, big data, and cloud-native deployments.


Requirements

  • Minimum 5 years of experience in Data Science or Applied Machine Learning.
  • Strong proficiency in Python, SQL, and ML libraries (Pandas, Scikit-learn, TensorFlow, PyTorch).
  • Proven expertise in deploying ML models into production systems.
  • Experience with big data platforms (Hadoop, Spark) and distributed data processing.
  • Hands-on experience with Databricks, Airflow, and AWS EMR.
  • Strong knowledge of AWS cloud services (S3, Lambda, SageMaker, EC2, etc.).
  • Solid understanding of query optimization, storage systems, and data pipelines.
  • Excellent problem-solving skills, with the ability to design scalable solutions.
  • Strong communication and collaboration skills to work in cross-functional teams.


Benefits

  • Best-in-class salary: We hire strong talent and compensate accordingly.
  • Proximity Talks: Meet and learn from designers, engineers, product leaders, and AI practitioners.
  • Continuous learning: Work with a world-class team and stay close to the latest in AI, engineering, and product development.
  • High-impact work: Build AI-first systems and products used at scale by global clients.



About Us

Proximity is the trusted technology, design, and consulting partner for some of the biggest Sports, Media, and Entertainment companies in the world. We’re headquartered in San Francisco and have offices in Palo Alto, Dubai, Mumbai, and Bangalore.

Since 2019, Proximity has built high-impact, scalable products used by millions of users every day. Today, we are a global team of engineers, designers, product managers, and experts solving complex problems and building cutting-edge technology at scale.


Read more
company logo
Anupam Arya
Posted by Anupam Arya
Bengaluru (Bangalore), Mumbai, Delhi, Gurugram, Hyderabad
5 - 12 yrs
₹30L - ₹40L / yr
skill iconMachine Learning (ML)
skill iconDeep Learning
Generative AI (GenAI)

AuxoAI is hiring a Senior Data Scientist with strong expertise in AI, machine learning engineering (MLE), and generative AI. You will play a leading role in designing, deploying, and scaling production-grade ML systems — including large language model (LLM)-based pipelines, AI copilots, and agentic workflows. This role is ideal for someone who thrives on balancing cutting-edge research with production rigor and loves mentoring while building impact-first AI applications. 


Location - Mumbai/Bangalore/Hyderabad/Gurgaon (Hybrid - 3 Days a week in Office)​​


Responsibilities: 

  • Own the full ML lifecycle: model design, training, evaluation, deployment 
  • Design production-ready ML pipelines with CI/CD, testing, monitoring, and drift detection 
  • Fine-tune LLMs and implement retrieval-augmented generation (RAG) pipelines 
  • Build agentic workflows for reasoning, planning, and decision-making 
  • Develop both real-time and batch inference systems using Docker, Kubernetes, and Spark 
  • Leverage state-of-the-art architectures: transformers, diffusion models, RLHF, and multimodal pipelines 
  • Collaborate with product and engineering teams to integrate AI models into business applications 
  • Mentor junior team members and promote MLOps, scalable architecture, and responsible AI best practices 



Requirements

  • 5+ years of experience in designing, deploying, and scaling ML/DL systems in production 
  • Proficient in Python and deep learning frameworks such as PyTorch, TensorFlow, or JAX 
  • Experience with LLM fine-tuning, LoRA/QLoRA, vector search (Weaviate/PGVector), and RAG pipelines 
  • Familiarity with agent-based development (e.g., ReAct agents, function-calling, orchestration) 
  • Solid understanding of MLOps: Docker, Kubernetes, Spark, model registries, and deployment workflows 
  • Strong software engineering background with experience in testing, version control, and APIs 
  • Proven ability to balance innovation with scalable deployment 
  • B.S./M.S./Ph.D. in Computer Science, Data Science, or a related field 
  • Bonus: Open-source contributions, GenAI research, or applied systems at scale 


Read more
company logo
Agency job
via by Bhavesh Kiroula
Remote only
3 - 15 yrs
₹15L - ₹42L / yr
MLOps
Aws sagemaker
skill iconAmazon Web Services (AWS)
MLFlow
Systems Development Life Cycle (SDLC)
+3 more

Example Responsibilities:

  • Build and optimize model serving infrastructure with a focus on inference latency and cost optimization
  • Architect efficient inference pipelines that balance latency, throughput, and cost across various acceleration options
  • Develop monitoring and observability solutions for ML systems
  • Collaborate with ML Engineers to establish best practices for optimized model deployment
  • Implement cost-efficient, enterprise-scale solutions
  • Collaborate in a cross-functional, distributed team for continuous system improvement
  • Work with MLEs, QA Engineers, and DevOps Engineers
  • Evaluate and implement new technologies and tools
  • Contribute to architectural decisions for distributed ML systems


Experience and Qualifications:

  • 5+ years of experience in software engineering with Python
  • Experience with ML frameworks, particularly PyTorch
  • Experience optimizing ML models with hardware acceleration (AWS Neuron , ONNX, TensorRT)
  • Experience with AWS ML services and hardware-accelerated instances (Sagemaker, Inferentia,Trainium)
  • Proven experience building and operating AWS serverless architectures
  • Deep understanding of event-driven processing patterns, SQS/SNS and serverless caching solutions
  • Experience with containerization using Docker and orchestration tools
  • Strong knowledge of RESTful API design and implementation
  • Proficiency in writing good quality & secure code and be familiar with static code analysis tools
  • Excellent analytical, conceptual and communication skills in spoken and written English
  • Experience applying Computer Science fundamentals in algorithm design, problem solving, and complexity analysis


Great to have Experience and Qualifications:

  • Experience with any of the following: model compilation and quantization, performance profiling and benchmarking ML inference systems
  • Experience working in regulated industries with strict compliance requirements for cloud-native solutions
Read more
Why apply to jobs via Cutshort
people_solving_puzzle
Personalized job matches
Stop wasting time. Get matched with jobs that meet your skills, aspirations and preferences.
people_verifying_people
Verified hiring teams
See actual hiring teams, find common social connections or connect with them directly.
ai_chip
Move faster with AI
We use AI to get you faster responses, recommendations and unmatched user experience.
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos