Cutshort logo
For Employers
Its for a IT Service MNC logo
DevOps AI Engineer
Its for a IT Service MNC
DevOps AI Engineer
Freelancer's logo

DevOps AI Engineer

at Its for a IT Service MNC

Agency job
6 - 10 yrs
₹10L - ₹30L / yr
Bengaluru (Bangalore), Mumbai, Chennai
Skills
DevOps
skill icongrafana
Generative AI

Primary Skills: 

Observability - ELK (Elactic/Kibana), Prometheus, Grafana, PromQL

Software and automation - Java, Python/Shell/Bash, Rest-SOAP API, docker containerization, Kubernetes, Kafka

Reliability and DR engineering - Distributed architecture and distributed system fundamentals, micro services, and event-driven architecture.

Cross-team coordination, incident triage and resolution, leadership and stakeholder management.

 

Secondary Skills:

Lang-chain, Langraph, RAG, MCP

Experience with working on LLM's and integrating with the existing applications

Python - FastAPI

Cache - Redis

 

Program Details:

Design, build, and ship LLM-powered and agentic product features that enhance the team efforts and outcomes.

Build agentic AI systems that reason over context, invoke tools, take real actions, and recover gracefully from failure.

Work on integrating the existing AI tools and should know major AI frameworks and libraries. 

Own service reliability and operational governance by defining SLA's, managing error budgets, and reporting reliability (MTTD, MTTR) to leadership for prioritization, risk decisions and planning.

Architect and continuously optimize the observability of platform using Kibana/Elastic (ELF) along with other observability tools like Prometheus, Grafana (dashboards, metrics, alert lifecycle), improving detection quality, reducing noise/toil, and enabling faster triage and measurable uptime improvements.

Engineer advance alerting and automation capabilities with Kibana alerting and anomaly detections and integrating response workflows (routing, runbooks, remediation scripts) to standardize on-call execution and accelerate restoration of services.

Lead incident response for customer-impacting issues across teams-coordination, communications, service restoration, and blameless RCA-then corrective actions that prevent recurrence and reduce operational risk.

Design, automate and validate Disaster Recovery and failover for critical services/journeys (RTO/RPO alignment, DR Drills), ensuring resiliency under failure scenarios and improving recovery.

Consult and partner with application teams by providing production readiness inputs (Resiliency patterns, availability, performance/capacity considerations) and driving platform enhancements that improve stability while optimizing infrastructure and observability spend.


If you are interested for this role, kindly acknowledge this email with your interest.

Read more
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos

Similar jobs

NeoGenCode Technologies Pvt Ltd
Divya Sharma
Posted by Divya Sharma
Gurugram
5 - 10 yrs
₹15L - ₹20L / yr
DevOps
Bare metal
Physical server
Onpremises
AWS
+2 more

Job Title: Senior DevOps Engineer

Location: Gurgaon – Sector 39

Work Mode: 5 Days Onsite

Experience: 5+ Years

About the Role

We are looking for an experienced Senior DevOps Engineer to build, manage, and maintain highly reliable, scalable, and secure infrastructure. The role involves deploying product updates, handling production issues, implementing customer integrations, and leading DevOps best practices across teams.

Key Responsibilities

  • Manage and maintain production-grade infrastructure ensuring high availability and performance.
  • Deploy application updates, patches, and bug fixes across environments.
  • Handle Level-2 support and resolve escalated production issues.
  • Perform root cause analysis and implement preventive solutions.
  • Build automation tools and scripts to improve system reliability and efficiency.
  • Develop monitoring, logging, alerting, and reporting systems.
  • Ensure secure deployments following data encryption and cybersecurity best practices.
  • Collaborate with development, product, and QA teams for smooth releases.
  • Lead and mentor a small DevOps team (3–4 engineers).

Core Focus Areas

Server Setup & Management (60%)

  • Hands-on management of bare-metal servers.
  • Server provisioning, configuration, and lifecycle management.
  • Network configuration including redundancy, bonding, and performance tuning.

Queue Systems – Kafka / RabbitMQ (15%)

  • Implementation and management of message queues for distributed systems.

Storage Systems – SAN / NAS (15%)

  • Setup and management of enterprise storage systems.
  • Ensure backup, recovery, and data availability.

Database Knowledge (5%)

  • Working experience with Redis, MySQL/PostgreSQL, MongoDB, Elasticsearch.
  • Basic database administration and performance tuning.

Telecom Exposure (Good to Have – 5%)

  • Experience with SMS, voice systems, or real-time data processing environments.

Technical Skills Required

  • Linux administration & Shell scripting
  • CI/CD tools – Jenkins
  • Git (GitHub / SVN) and branching strategies
  • Docker & Kubernetes
  • AWS cloud services
  • Ansible for configuration management
  • Databases: MySQL, MariaDB, MongoDB
  • Web servers: Apache, Tomcat
  • Load balancing & HA: HAProxy, Keepalived
  • Monitoring tools: Nagios and related observability stacks


Read more
Recro
at Recro
1 video
32 recruiters
Shaista Rafik
Posted by Shaista Rafik
Bengaluru (Bangalore)
3 - 7 yrs
Best in industry
Infrastructure
DevOps
skill iconAmazon Web Services (AWS)
skill iconKubernetes
CI/CD
+4 more

On the Job

● Build and maintain scalable cloud infrastructure on AWS

● Improve CI/CD pipelines, deployment systems, and developer workflows

● Automate infrastructure provisioning using Infrastructure-as-Code (Terraform/Pulumi)

● Manage and optimize Kubernetes clusters and containerized workloads

● Improve observability across systems through monitoring, logging, and alerting

● Drive reliability initiatives including incident response, root cause analysis, and operational

improvements

● Collaborate with engineering teams to improve service scalability, performance, and security

● Implement IAM, secrets management, and infrastructure security best practices

● Optimize infrastructure costs and improve resource efficiency

● Actively leverage AI tools and workflows to improve engineering productivity and automation


Must Haves

● 3–6 years of experience in DevOps, Platform Engineering, or SRE roles

● Strong hands-on experience with AWS infrastructure and services

● Good understanding of Kubernetes, Docker, and container orchestration

● Experience with Infrastructure-as-Code tools like Terraform or Pulumi

● Strong scripting/coding skills in Python, Go, or Bash

● Experience building and maintaining CI/CD pipelines (GitHub Actions, Jenkins, GitLab CI,

etc.)

● Understanding of networking fundamentals, Linux systems, and cloud security practices

● Familiarity with monitoring and observability tools like Prometheus, Grafana, ELK, Datadog,

etc.

● Strong debugging and problem-solving skills

● Ability to work independently in a fast-moving environment


Good To Haves

● Experience working in fintech or high-scale startup environments

● Exposure to service mesh, zero-trust security, or secrets management systems

● Experience with multi-cluster Kubernetes environments

● Familiarity with incident management, SLOs, and reliability engineering practices

● Experience building internal developer platforms or automation tooling

Read more
Agentic AI Platform
Agentic AI Platform
Agency job
via Peak Hire Solutions by Dharati Thakkar
Gurugram
3 - 6 yrs
₹10L - ₹25L / yr
DevOps
skill iconPython
Google Cloud Platform (GCP)
Linux/Unix
CI/CD
+21 more

Review Criteria

  • Strong DevOps /Cloud Engineer Profiles
  • Must have 3+ years of experience as a DevOps / Cloud Engineer
  • Must have strong expertise in cloud platforms – AWS / Azure / GCP (any one or more)
  • Must have strong hands-on experience in Linux administration and system management
  • Must have hands-on experience with containerization and orchestration tools such as Docker and Kubernetes
  • Must have experience in building and optimizing CI/CD pipelines using tools like GitHub Actions, GitLab CI, or Jenkins
  • Must have hands-on experience with Infrastructure-as-Code tools such as Terraform, Ansible, or CloudFormation
  • Must be proficient in scripting languages such as Python or Bash for automation
  • Must have experience with monitoring and alerting tools like Prometheus, Grafana, ELK, or CloudWatch
  • Top tier Product-based company (B2B Enterprise SaaS preferred)


Preferred

  • Experience in multi-tenant SaaS infrastructure scaling.
  • Exposure to AI/ML pipeline deployments or iPaaS / reverse ETL connectors.


Role & Responsibilities

We are seeking a DevOps Engineer to design, build, and maintain scalable, secure, and resilient infrastructure for our SaaS platform and AI-driven products. The role will focus on cloud infrastructure, CI/CD pipelines, container orchestration, monitoring, and security automation, enabling rapid and reliable software delivery.


Key Responsibilities:

  • Design, implement, and manage cloud-native infrastructure (AWS/Azure/GCP).
  • Build and optimize CI/CD pipelines to support rapid release cycles.
  • Manage containerization & orchestration (Docker, Kubernetes).
  • Own infrastructure-as-code (Terraform, Ansible, CloudFormation).
  • Set up and maintain monitoring & alerting frameworks (Prometheus, Grafana, ELK, etc.).
  • Drive cloud security automation (IAM, SSL, secrets management).
  • Partner with engineering teams to embed DevOps into SDLC.
  • Troubleshoot production issues and drive incident response.
  • Support multi-tenant SaaS scaling strategies.


Ideal Candidate

  • 3–6 years' experience as DevOps/Cloud Engineer in SaaS or enterprise environments.
  • Strong expertise in AWS, Azure, or GCP.
  • Strong expertise in LINUX Administration.
  • Hands-on with Kubernetes, Docker, CI/CD tools (GitHub Actions, GitLab, Jenkins).
  • Proficient in Terraform/Ansible/CloudFormation.
  • Strong scripting skills (Python, Bash).
  • Experience with monitoring stacks (Prometheus, Grafana, ELK, CloudWatch).
  • Strong grasp of cloud security best practices.



Read more
Bengaluru (Bangalore)
4 - 6 yrs
₹30L - ₹37L / yr
DevOps

Candidate must be from a product-based company with experience handling large-scale production traffic.

2. Candidate must have strong Linux expertise with hands-on production troubleshooting and working knowledge of databases and middleware (Mongo, Redis, Cassandra, Elasticsearch, Kafka).

3. Candidate must have solid experience with Kubernetes.

4. Candidate should have strong knowledge of configuration management tools like Ansible, Terraform, and Chef / Puppet. Add on- Prometheus & Grafana etc.

5. Candidate must be an individual contributor with strong ownership.

6. Candidate must have hands-on experience with DATABASE MIGRATIONS and observability tools such as Prometheus and Grafana.

7. Candidate must have working knowledge of Go/Python and Java.

8. Candidate should have working experience on Cloud platform - AWS

9. Candidate should have Minimum 1.5 years stability per organization, and a clear reason for relocation

Read more
Pluginlive
at Pluginlive
1 recruiter
Harsha Saggi
Posted by Harsha Saggi
Remote only
4 - 6 yrs
₹2L - ₹6L / yr
skill iconDocker
skill iconJenkins
skill iconKubernetes
DevOps
skill iconPython
+4 more

NOTE- This is a contractual role for a period of 3-6 months.


Responsibilities:

● Set up and maintain CI/CD pipelines across services and environments

● Monitor system health and set up alerts/logs for performance & errors ● Work closely with backend/frontend teams to improve deployment velocity

● Manage cloud environments (staging, production) with cost and reliability in mind

● Ensure secure access, role policies, and audit logging

● Contribute to internal tooling, CLI automation, and dev workflow improvements


Must-Haves:

● 2–3 years of hands-on experience in DevOps, SRE, or Platform Engineering

● Experience with Docker, CI/CD (especially GitHub Actions), and cloud providers (AWS/GCP)

● Proficiency in writing scripts (Bash, Python) for automation

● Good understanding of system monitoring, logs, and alerting

● Strong debugging skills, ownership mindset, and clear documentation habits

● Infra monitoring tools like Grafana dashboards

Read more
Indventur Partner
at Indventur Partner
2 recruiters
Vanshika kaur
Posted by Vanshika kaur
Remote only
5 - 6 yrs
₹24L - ₹25L / yr
DevOps
CI/CD
GitOps
Client: The client is one of the largest API platforms in SE Asia whose mission is to shape the digital transformation and be the leading player in building the local API economy.

Position: DevOps Lead 

Job Description

● Research, evangelize and implement best practices and tools for GitOps, DevOps, continuous integration, build automation, deployment automation, configuration management, infrastructure as code.

● Develop software solutions to support DevOps tooling; including investigation of bug fixes, feature enhancements, and software/tools updates

● Participate in the full systems life cycle with solution design, development, implementation, and product support using Scrum and/or other Agile practices

● Evaluating, implementing, and streamlining DevOps practices.

● Design and drive the implementation of fully automated CI/CD pipelines.

● Designing and creating Cloud services and architecture for highly available and scalable environments. Lead the monitoring, debugging, and enhancing pipelines for optimal operation and performance. Supervising, examining, and handling technical operations.

Qualifications

● 5 years of experience in managing application development, software delivery lifecycle, and/or infrastructure development and/or administration

● Experience with source code repository management tools, code merge and quality checks, continuous integration, and automated deployment & management using tools like Bitbucket, Git, Ansible, Terraform, Artifactory, Service Now, Sonarqube, Selenium.

● Minimum of 4 years of experience with approaches and tooling for automated build, delivery, and release of the software

● Experience and/or knowledge of CI/CD tools: Jenkins, Bitbucket Pipelines, Gitlab CI, GoCD.

● Experience with Linux systems: CentOS, RHEL, Ubuntu, Secure Linux... and Linux Administration.

● Minimum of 4 years experience with managing medium/large teams including progress monitoring and reporting

● Experience and/or knowledge of Docker, Cloud, and Orchestration: GCP, AWS, Kubernetes.

● Experience and/or knowledge of system monitoring, logging, high availability, redundancy, autoscaling, and failover.

● Experience automating manual and/or repetitive processes.

● Experience and/or knowledge with networking and load balancing: Nginx, Firewall, IP network
Read more
Euromonitor International
Bengaluru (Bangalore)
8 - 16 yrs
₹40L - ₹45L / yr
DevOps
skill iconKubernetes
skill iconDocker
skill iconAmazon Web Services (AWS)
Windows Azure
+1 more

Requirements:

  • Experience of managing Engineering teams in an Agile environment.
  • Expert knowledge of delivering solutions in Azure cloud within a large-scale enterprise environment.
  • Great understanding of DevOps principles and how they assist in taking products to market in an effective manner.
  • Experience of Automation/Configuration management tools as well as working in a continuous delivery environment, monitoring and tooling.
  • Knowledge and experience in Azure, Kubernetes, Containerisation, Azure DevOps pipelines.
  • Experience in managing permissions in Azure DevOps.
  • Working experience in Application Gateways, App Services, Front-Door, Azure Service Bus, etc.
  • Troubleshooting experience in virtual/cloud infrastructures.
  • Experience in delivery of projects using IAC (Infrastructure as Code).
Read more
provides mobile application development & support services.
provides mobile application development & support services.
Agency job
via Jobdost by Ankitha Vyas
Hyderabad
8 - 10 yrs
₹20L - ₹35L / yr
DevOps
skill iconDocker
skill iconKubernetes
Terraform
skill iconAmazon Web Services (AWS)
+18 more
Sr Cloud & DevOps Engineer

We are looking for a self motivated and goal oriented candidate to lead in architecting, developing, deploying, and maintaining first class, highly scalable, highly available SaaS platforms.

This is a very hands-on role.  You will have a significant impact on Wenable's success.

Technical Requirements:

    8+ years SaaS and Cloud Architecture and Development with frameworks such as:
        - AWS, GoogleCloud, Azure, and/or other
        - Kafka, RabbitMQ, Redis, MongoDB, Cassandra, ElasticSearch
        - Docker, Kubernetes, Helm, Terraform, Mesos, VMs, and/or similar orchestration, scaling, and deployment frameworks
        - ProtoBufs, JSON modeling
        - CI/CD utilities like Jenkins, CircleCi, etc.. 
        - Log aggregation systems like Graylog or ELK
        - Additional development tools typically used in orchestration and automation like Python, Shell, etc...
        - Strong security best practices background
        - Strong software development a plus

Leadership Requirements:
    
    - Strong written and verbal skills.  This role will entail significant coordination both internally and externally.
    - Ability to lead projects of blended teams, on/offshore, of various sizes.
    - Ability to report to executive and leadership teams.
    - Must be data driven, and objective/goal oriented.
Read more
A Series-B funded, Fintech Company based out of Bangalore.
A Series-B funded, Fintech Company based out of Bangalore.
Agency job
via Nexusrize Solutions by Indira Cowkur
Remote, Bengaluru (Bangalore)
3 - 12 yrs
₹15L - ₹35L / yr
DevOps
skill iconDocker
skill iconKubernetes
Terraform
CI/CD
+4 more

Requirements and Qualifications

  • Bachelor’s degree in Computer Science Engineering or in a related field
  • 4+ years of experience
  • Excellent analytical and problem-solving skills
  • Strong knowledge of Linux systems and internals
  • Programming experience in Python/Shell scripting
  • Strong AWS skills with knowledge of EC2, VPC, S3, RDS, Cloudfront, Route53, etc
  • Experience in containerization (Docker) and container orchestration (Kubernetes)
  • Experience in DevOps & CI/CD tools such as Git, Jenkins, Terraform, Helm
  • Experience with SQL & NoSQL databases such as MySql, MongoDB, and ElasticSearch
  • Debugging and troubleshooting skills using tools such as strace, tcpdump, etc
  • Good understanding of networking protocol and security concerns (VPN, VPC, IG, NAT, AZ, Subnet)
  • Experience with monitoring and data analysis tools such as Prometheus, EFK, etc
  • Good communication & collaboration skills and attention to details
  • Participation in rotating on-call duties
Read more
Opt IT technologies
at Opt IT technologies
4 recruiters
Niranjini R
Posted by Niranjini R
Bengaluru (Bangalore)
3 - 7 yrs
₹4L - ₹6L / yr
DevOps
Linux administration
Shell Scripting
skill iconAmazon Web Services (AWS)
Continuous Integration
+1 more
Requirements
    • Strong Understanding of Linux administration
    • Good understanding of using Python or Shell scripting (Automation mindset is key in this role)
    • Hands on experience with Implementation of CI/CD Processes
      Experience working with one of these cloud platforms (AWS, Azure or Google Cloud)
    • Experience working with configuration management tools such as Ansible, Chef
      Experience in Source Control Management including SVN, Bitbucket and GitHub
      Experience with setup & management of monitoring tools like Nagios, Sensu & Prometheus
      Troubleshoot and triage development and Production issues
    • Understanding of micro-services is a plus

Roles & Responsibilities
  • Implementation and troubleshooting on Linux technologies related to OS, Virtualization, server and storage, backup, scripting / automation, Performance fine tuning
  • LAMP stack skills
  • Monitoring tools deployment / management (Nagios, New Relic, Zabbix, etc)
  • Infra provisioning using Infra as code mindset
  • CI/CD automation
 
 
Read more
Why apply to jobs via Cutshort
people_solving_puzzle
Personalized job matches
Stop wasting time. Get matched with jobs that meet your skills, aspirations and preferences.
people_verifying_people
Verified hiring teams
See actual hiring teams, find common social connections or connect with them directly.
ai_chip
Move faster with AI
We use AI to get you faster responses, recommendations and unmatched user experience.
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos