Cutshort logo
For Employers
Agentic Universe logo
Senior DevOps (Reliability, Cost & Release Systems)
Senior DevOps (Reliability, Cost & Release Systems)

Senior DevOps (Reliability, Cost & Release Systems) at Agentic Universe · Bengaluru (Bangalore) · 2 - 5 years · ₹4.2L - ₹6L / yr · Raised funding · Posted 24 Apr 2026

Agentic Universe's logo

Senior DevOps (Reliability, Cost & Release Systems)

Anubhav Kumar Rai's profile picture
Posted by Anubhav Kumar Rai
2 - 5 yrs
₹4.2L - ₹6L / yr
Bengaluru (Bangalore)
Skills
DevOps
Agile/Scrum
CI/CD

Location: Bangalore

Experience: 2–5 years

Type: Full-time | On-site

Open Roles: 1

Start: Immediate

Why this role exists

Most engineering teams choose between speed and stability.

We need both.

Today:

  • Deployments carry risk
  • Cloud costs are higher than they should be
  • Compliance is reactive, not built-in

This role exists to build a platform where:

  • We can deploy fast without breaking production
  • We can scale without runaway cost
  • We can pass enterprise InfoSec reviews without firefighting

What you’ll do

You will not just manage infrastructure.

You will build the platform that engineering runs on.

1. Drive cloud cost efficiency

  • Reduce Azure compute spend by 40%
  • Implement:
  • Reserved Instances / savings plans
  • Right-sizing of workloads
  • Scheduling for non-critical workloads
  • Continuously monitor and optimize cost vs performance

2. Build zero-downtime deployment systems

  • Ship a deployment pipeline that supports:
  • 5+ production deployments per week
  • Zero customer-visible downtime
  • Implement:
  • Blue-green / canary deployments
  • Automated health checks
  • Safe rollout strategies

3. Enable fast and safe releases

  • Reduce time-to-launch significantly
  • Ensure:
  • High reliability in every release
  • Ability to rollback instantly if something breaks
  • Create systems where:
  • Scaling up is seamless when things go right
  • Failures are contained when they don’t

4. Build disaster recovery and compliance readiness

  • Create DR/BCP systems that pass enterprise audits from:
  • HDFC Life, SBI Life
  • Ensure:
  • Backup and recovery processes are defined and tested
  • Failover strategies are documented and executable
  • Build compliance as part of the system, not an afterthought

5. Embed security into the pipeline

  • Integrate:
  • SAST (Static Application Security Testing)
  • DAST (Dynamic Application Security Testing)
  • SCA (Software Composition Analysis)
  • Secret scanning
  • Container scanning
  • IaC scanning
  • Ensure vulnerabilities are caught before deployment

6. Enforce policy-as-code

  • Implement:
  • OPA / Gatekeeper
  • Azure Policy
  • Prevent non-compliant infrastructure from being deployed
  • Ensure consistency across environments

7. Build a scalable platform layer

  • Create systems that:
  • Support increasing deployment frequency
  • Maintain reliability under scale
  • Work closely with backend and SRE teams to:
  • Improve system stability
  • Reduce operational overhead

What success looks like

  • Cloud costs reduce by ≥ 40%
  • Deployments are:
  • Frequent
  • Safe
  • Invisible to customers
  • Rollbacks are instant and reliable
  • DR/BCP passes enterprise audits in the first attempt
  • Security is embedded in the pipeline, not patched later
  • Engineering teams ship faster with confidence

Who you are

  • You have 2-5 years of experience in DevOps / Platform Engineering
  • You have worked with:
  • Cloud platforms (Azure preferred)
  • CI/CD systems
  • Infrastructure as Code
  • You think in:
  • Systems
  • Trade-offs (speed vs reliability vs cost)
  • You are comfortable owning:
  • Production infrastructure
  • Deployment systems

What will make you stand out

  • Experience with:
  • High-frequency deployment systems
  • Cost optimization at scale
  • Security-first pipelines
  • Strong understanding of:
  • Kubernetes / container orchestration
  • Monitoring and observability
  • Distributed system reliability
  • Experience passing enterprise security/compliance audits

Why join

  • You will define how engineering ships and scales
  • Your work directly impacts:
  • Reliability
  • Cost
  • Deployment velocity
  • You will build a platform that moves from:
  • Fragile → predictable and scalable

What this role is not

  • Not manual infra management
  • Not reactive firefighting
  • Not limited to CI/CD maintenance

What this role is

  • A builder of deployment systems
  • A driver of cost efficiency
  • A guardian of reliability and compliance

One question to self-evaluate

Can you build a platform where we deploy faster, spend less, and never break production?


Read more
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos

About Agentic Universe

Founded :
2022
Type :
Product
Size :
20-100
Stage :
Raised funding

About

Agentic Universe - AI Agents that run outcomes for your teams
Read more

Company social profiles

N/A

Similar jobs (10)

company logo
Santhanalakshmi A
Posted by Santhanalakshmi A
Bengaluru (Bangalore), Hyderabad
4 - 8 yrs
₹8L - ₹12L / yr
skill iconKubernetes
skill iconPython
DevOps
skill iconAmazon Web Services (AWS)
Windows Azure

Devops + Python 

Mandatory: Python+Kubernetes Containerisation+ DevOps Exposure+ Cloud Exposure(any)

4-8 years /16 LPA

10+ years /26 LPA

Bangalore/Hyderabad

immediate to 15 days.

Read more
company logo
Akshay Patil
Posted by Akshay Patil
Remote only
7 - 10 yrs
₹8L - ₹12L / yr
DevOps
CI/CD
Azure DevOps
GitHub Actions
skill iconKubernetes
+10 more

Job Title : DevOps Engineer

Experience : 7+ Years

Working Hours : 3:30 AM – 11:30 AM ET

Duration : 12–18 Months

Engagement : Contract

Freelance : Not Applicable

BGV : Required


Role Overview :

We are looking for an experienced DevOps Engineer to design, implement, and optimize DevOps practices, CI/CD pipelines, deployment automation, and cloud-native delivery platforms. The role will focus on enabling secure, reliable, and efficient software delivery while improving infrastructure automation, release processes, platform reliability, and operational resilience.


Mandatory Skills :

DevOps, CI/CD, Azure DevOps / GitHub Actions, Kubernetes, Docker, PowerShell / Python / Bash, Automation, Monitoring & Observability, DevSecOps, IaC, Troubleshooting, and Production Support.


Key Responsibilities :

  • Design, build, and maintain CI/CD pipelines for enterprise application delivery and release management.
  • Develop deployment automation frameworks to improve release quality, consistency, scalability, and efficiency.
  • Administer Kubernetes and containerized environments and supporting orchestration platforms.
  • Improve build, test, deployment, and release workflows for development teams.
  • Integrate automated testing, security scanning, validation, and compliance controls into CI/CD pipelines.
  • Implement monitoring, logging, alerting, and observability solutions for platform and application environments.
  • Troubleshoot deployment failures, production issues, and platform incidents; perform root cause analysis.
  • Develop automation scripts and tools using PowerShell, Python, Bash, or similar technologies.
  • Implement Infrastructure as Code and automation practices to reduce manual effort and operational risk.
  • Collaborate with Cloud, Infrastructure, Security, and Application teams to deliver scalable and resilient solutions.
  • Maintain deployment standards, release procedures, technical documentation, and operational runbooks.
  • Apply DevOps, DevSecOps, and SRE best practices to improve delivery speed and platform reliability.
  • Evaluate emerging DevOps technologies and recommend improvements.
  • Mentor team members on DevOps methodologies, automation, and platform engineering practices.
  • Act as an escalation point for critical deployment, infrastructure, and platform issues.


Required Technical Skills :

1. DevOps & CI/CD :

  • 7+ years of experience in DevOps, Platform Engineering, Infrastructure Engineering, or Software Engineering.
  • Strong experience designing and supporting CI/CD pipelines and release management.
  • Hands-on experience with Azure DevOps, GitHub, GitHub Actions, or equivalent platforms.
  • Strong understanding of Git, branching strategies, release management, and deployment methodologies.


2. Containers & Cloud :

  • Strong hands-on experience with Kubernetes and Docker/containerization.
  • Understanding of cloud-native application architectures and distributed systems.
  • Experience supporting production-grade containerized environments.


3. Automation & IaC :

  • Strong proficiency in PowerShell, Python, Bash, or similar scripting languages.
  • Experience developing automation frameworks and deployment tooling.
  • Knowledge of Infrastructure as Code (IaC) concepts.
  • Experience with Terraform, Bicep, or similar tools is preferred.


4. Monitoring & Reliability :

  • Experience with monitoring, logging, observability, and alerting.
  • Strong troubleshooting and root cause analysis skills.
  • Experience supporting highly available and resilient production systems.
  • Knowledge of SRE practices is preferred.


5. Security & DevSecOps :

  • Understanding of DevSecOps and secure software delivery practices.
  • Experience integrating vulnerability scanning, security controls, and compliance checks into CI/CD pipelines.
  • Knowledge of IAM, secrets management, and access control within DevOps environments.


Preferred Qualifications :

  • Hands-on experience with Microsoft Azure.
  • Experience with Terraform/Bicep or other IaC tools.
  • Knowledge of SRE and platform engineering practices.
  • Relevant Azure, Kubernetes, Cloud, or DevOps certifications.
Read more
company logo
Arshiya Shaikh
Posted by Arshiya Shaikh
Mumbai
4 - 7 yrs
₹5L - ₹9L / yr
AWS CloudFormation
Bitbucket
skill iconDocker
skill iconKubernetes

Key Responsibilities

  • Automate application deployments from Bitbucket to servers using CI/CD pipelines.
  • Design and manage scalable, highly available AWS infrastructure.
  • Implement Auto Scaling, ELB, and Route 53 for traffic management and high availability.
  • Work with AWS services including IAM, RDS, DynamoDB, EC2, and other cloud services.
  • Build and manage Docker containers and server images.
  • Deploy and manage applications using Kubernetes.
  • Implement Infrastructure as Code using Terraform, CloudFormation, or Ansible.
  • Develop automation scripts using Python and Bash.
  • Implement monitoring and logging using tools such as Prometheus, Grafana, and ELK.
  • Integrate security and compliance practices into CI/CD pipelines.
  • Optimize infrastructure for security, scalability, performance, and cost.

Required Skills

  • 3+ years of experience in DevOps or a similar role.
  • Strong knowledge of AWS beyond EC2.
  • Hands-on experience with Jenkins or similar CI/CD tools.
  • Experience with Docker and Kubernetes.
  • Good understanding of Terraform/IaC and automation.
  • Proficiency in Python and/or Bash scripting.
  • Knowledge of DevSecOps, security, and compliance best practices.
  • Strong troubleshooting and problem-solving skills.


Read more
company logo
Gurugram
3 - 6 yrs
₹4L - ₹9L / yr
DevOps
Linux/Unix
Networking
Bare Metal
Server administration
+13 more

Job Title : DevOps Engineer – Linux, AWS & Infrastructure

Experience : 3+ Years

Location : Sector 48, Gurgaon

Work Mode : 6 Days WFO – Monday to Saturday

Week Off : 01st & 03rd Saturday Off

Employment Type : Full-Time


Job Summary :

We are looking for a DevOps Engineer with strong hands-on experience in Linux, Networking, Server Administration, AWS, CI/CD, Containers, Kubernetes, and Infrastructure Automation. The ideal candidate should have strong troubleshooting skills, production ownership, and the ability to manage and automate infrastructure reliably.


Key Responsibilities :

  • Manage and troubleshoot Linux servers, bare-metal infrastructure, and server environments.
  • Perform system administration, networking, performance monitoring, and production troubleshooting.
  • Design, maintain, and optimize Jenkins-based CI/CD pipelines and deployment workflows.
  • Manage Docker containers and Kubernetes environments.
  • Work with AWS Cloud services, infrastructure, security, and deployment environments.
  • Monitor system health, application performance, logs, and infrastructure using appropriate monitoring and logging tools.
  • Implement and maintain Infrastructure as Code (IaC) using tools such as Terraform or CloudFormation.
  • Automate repetitive operational tasks using Python, Bash, Shell scripting, or similar technologies.
  • Implement infrastructure and application security, access controls, patching, and hardening.
  • Investigate production incidents, perform root-cause analysis (RCA), and drive issues to resolution.
  • Take end-to-end ownership of infrastructure reliability, availability, and operational issues.
  • Collaborate with development and other engineering teams to improve deployment, scalability, and system reliability.


Mandatory Skills :

Linux & Networking | Bare Metal / Server Administration | Jenkins / CI-CD | Docker / Containers | AWS Cloud | Kubernetes | Git | Monitoring & Logging | Security | Infrastructure as Code (IaC) | Automation / Scripting | Production Troubleshooting & Ownership


Preferred Skills :

  • Strong understanding of TCP/IP, DNS, HTTP/HTTPS, SSH, load balancing, and networking fundamentals.
  • Hands-on experience with Terraform / CloudFormation.
  • Experience with Prometheus, Grafana, ELK / EFK, CloudWatch, or similar monitoring / logging tools.
  • Good understanding of Linux performance troubleshooting, processes, memory, disk, networking, and file systems.
  • Experience handling production incidents, RCA, deployments, and system reliability.
  • Exposure to cloud security, IAM, secrets management, and server hardening.


What We’re Looking For :

  • 3+ years of hands-on experience in DevOps / SRE / Infrastructure Engineering.
  • Strong practical knowledge rather than certification-based/theoretical understanding.
  • Good troubleshooting and analytical skills.
  • Strong sense of ownership and accountability for production systems.
  • Comfortable working in a 6-day work-from-office environment.
Read more
Gurugram
4 - 10 yrs
₹4L - ₹10L / yr
DevOps
Site Reliability Engineer (SRE)
skill iconAmazon Web Services (AWS)
skill iconDocker
skill iconKubernetes
+14 more

🚀 Job Title : DevOps Engineer / Site Reliability Engineer (SRE)

Experience Level : 4+ Years

Location : Gurugram Sector 48, Haryana (On-site)

Employment Type : Full Time Opportunity


About the Role :

We are looking for a proactive DevOps / Site Reliability Engineer (SRE) with around 4 years of hands-on experience designing, automating, and scaling cloud infrastructure and CI/CD delivery pipelines.

In this role, you will bridge the gap between development and operations. You will be responsible for orchestrating containerized applications, automating infrastructure via Code (IaC), establishing SRE best practices (SLIs, SLOs, SLAs), and ensuring maximum uptime, resiliency, and operational efficiency across multi-cloud environments (AWS/Azure/GCP).


Mandatory Skills :

AWS, Kubernetes, Docker, Terraform, Ansible, Jenkins, GitLab CI/CD, GitHub Actions, Python, Bash, CI/CD, Infrastructure as Code (IaC), Grafana, Prometheus, ELK, New Relic, CloudWatch, SRE, SLI/SLO/SLA, Linux


Key Responsibilities :

1. Cloud Infrastructure & Infrastructure as Code (IaC) :

  • Provision, configure, and maintain scalable, high-availability infrastructure on multi-cloud platforms, primarily AWS (EC2, VPC, IAM, S3, RDS, Route53, ALB/ASG, Lambda, EBS).
  • Build, deploy, and manage Infrastructure as Code (IaC) using Terraform, Ansible, and CloudFormation to enforce consistency and eliminate configuration drift.
  • Execute disaster recovery (DR) planning, automated failover / failback mechanisms, and chaos engineering exercises to validate system resiliency.

2. CI/CD, Automation & Development :

  • Design, end-to-end maintain, and optimize robust CI/CD pipelines using Jenkins, GitLab CI, and GitHub Actions.
  • Automate release pipelines, versioning, branching strategies, and approval gates using Groovy, Python, and Bash scripting. Integrate automated code quality and security scanning tools (SonarQube, Black Duck, or Fortify) directly into delivery pipelines.
  • Develop custom tools, scripts, or microservices (e.g., Python / Node.js) to automate manual operational tasks and operational toil.

3. Containerization & Orchestration :

  • Onboard and orchestrate containerized microservices utilizing Docker and Kubernetes (including Helm charts).
  • Ensure high availability, auto-scaling, resource management, and fault tolerance for Kubernetes pod deployments.

4. Observability, SRE & Incident Management :

  • Drive Site Reliability Engineering (SRE) maturity by establishing, tracking, and reporting SLIs, SLOs, and SLAs with cross-functional engineering teams.
  • Build, configure, and manage full-stack observability tools : Grafana, Prometheus, New Relic, Elasticsearch / Logstash / Kibana (ELK), Sentry, and AWS CloudWatch.
  • Set up real-time alerting, custom metric dashboards, and automated log rotation / pruning scripts.
  • Handle production incidents, lead Root Cause Analysis (RCA) investigations, and implement preventive measures to reduce Mean Time to Resolution (MTTR).


Required Qualifications & Skills :

  • Education : Bachelor’s Degree in Electronics and Communication Engineering, Computer Science, or a related technical field.
  • Experience : ~4 years of experience in DevOps, SRE, or Cloud System Administration roles.
  • Cloud & Infrastructure : Hands-on experience with AWS (Core services like EC2, S3, VPC, RDS, IAM, Lambda, Auto Scaling) and exposure to Azure / GCP.
  • CI/CD & Version Control : Proficiency with Jenkins, GitLab CI, GitHub Actions, and Git workflows.
  • Containerization : Core proficiency in Docker and Kubernetes cluster management / onboarding.
  • Infrastructure as Code : Expertise in Ansible, Terraform, or AWS CloudFormation.
  • Scripting & Languages : Strong hands-on automation skills with Python, Bash, and foundational knowledge of Node.js, Java or C++.
  • Observability & Logging : Strong experience with Grafana, Prometheus, New Relic, ELK stack, or Splunk.
  • Database & SQL : Familiarity with relational databases (MySQL, RDS) for monitoring setup and operational analytics.
Read more
MNC
MNC
Agency job
via by Jaya Mishra
Remote only
5.5 - 16 yrs
Best in industry
skill iconKubernetes
Platform engineering
Azure networking
Terraform
DevOps

Azure DevOps Engineer

Experience

5-10 years of hands-on experience in Platform Engineering, Cloud Engineering, SRE, DevOps or Infrastructure Engineering roles.

Priority 1 – Must Have (Hands-On)

Azure Cloud Platform

·        Strong hands-on experience supporting Azure workloads in production environments.

·        Experience designing, building and supporting Azure infrastructure using Terraform.

·        Good understanding of Azure networking and connectivity patterns.

·        Experience supporting:

Ø AKS

Ø Application Gateway

Ø Azure Traffic Manager

Ø Key Vault

Ø Azure Monitor / Log Analytics

Ø Managed Identities

Ø Service Principals

Ø Private Endpoints

Ø VNets, NSGs and Route Tables

 

Kubernetes / AKS

  • Strong practical experience operating and supporting AKS.
  • Ability to troubleshoot:

Ø Pod failures

Ø Ingress issues

Ø DNS issues

Ø SSL/TLS certificate issues

Ø Network routing issues

Ø Performance and availability incidents

  • Experience with:

Ø Helm

Ø Ingress Controllers

Ø Cluster upgrades

Ø Scaling

Ø Monitoring

 

Terraform

·        Strong hands-on experience writing and maintaining Terraform.

·        Experience creating reusable modules.

·        Experience managing:

·        State files

·        Remote backends

·        Environment promotion

·        Infrastructure lifecycle

Linux & Scripting

·        Strong Linux administration fundamentals.

·        Practical experience troubleshooting production issues.

·        Bash scripting mandatory.

·        Python desirable.

Application Support / Troubleshooting

Must be comfortable supporting business applications end-to-end.

 

 Priority 2 – Highly Desirable

GitHub & DevOps Platform

Hands-on experience with:

·        GitHub Enterprise

·        GitHub Actions

·        Shared workflows

·        Reusable pipelines

·        Repository onboarding

·        Branch protections

·        GitHub security features

Experience supporting:

·        Runner issues

·        Disk space issues

·        Network connectivity issues

·        Dependency failures

·        Self-hosted runners lifecycle management

API Management

Pipeline failures

GitOps

Experience with:

·        ArgoCD

·        GitOps deployment models

·        Kubernetes deployment automation

 

Monitoring & Observability

Experience working with:

·        Prometheus

·        Grafana

·        Azure Monitor

·        Log Analytics

·        Application Insights

Read more
company logo
Sakshi Mittal
Posted by Sakshi Mittal
Bengaluru (Bangalore)
3 - 5 yrs
₹6L - ₹12L / yr
skill iconAmazon Web Services (AWS)
DevOps
skill iconKubernetes
Terraform
CI/CD
+1 more

Job Summary :

We are looking for a proactive and skilled DevOps Engineer to join our team and play a key role in building, managing, and scaling infrastructure for high-performance systems. The ideal candidate will have hands-on experience with Kubernetes, Docker, Python scripting, cloud platforms, and DevOps practices around CI/CD, monitoring, and incident response.

Key Responsibilities :

- Design, build, and maintain scalable, reliable, and secure infrastructure on cloud platforms such as AWS.

- Implement Infrastructure as Code (IaC) using tools like Terraform, Cloud Formation, or similar.

- Manage Kubernetes clusters, configure namespaces, services, deployments, and auto scaling. CI/CD & Release Management

- Build and optimize CI/CD pipelines for automated testing, building, and deployment of services.

- Collaborate with developers to ensure smooth and frequent deployments to production.

- Manage versioning and rollback strategies for critical deployments.

- Containerization & Orchestration using Kubernetes.

- Containerize applications using Docker, and manage them using Kubernetes.

- Write automation scripts using Python or Shell for infrastructure tasks, monitoring, and deployment flows.

- Develop utilities and tools to enhance operational efficiency and reliability.

- Monitoring & Incident Management

- Analyze system performance and implement infrastructure scaling strategies based on load and usage trends.

- Optimize application and system performance through proactive monitoring and configuration tuning.

Desired Skills and Experience :

- Experience Required - 6+ yrs.

- Hands-on experience on cloud services like AWS, EKS etc.

- Ability to design a good cloud solution.

- Strong Linux troubleshooting, Shell Scripting, Kubernetes, Docker, Ansible, Jenkins Skills.

- Design and implement the CI/CD pipeline following the best industry practices using open-source tools.

- Use knowledge and research to constantly modernize our applications and infrastructure stacks.

- Be a team player and strong problem-solver to work with a diverse team.

- Having good communication skills.

Read more
company logo
Agency job
via by aarushi Mahajan
UAE
7 - 13 yrs
₹20L - ₹32L / yr
skill iconAmazon Web Services (AWS)
Azure OpenAI
Google Cloud Platform (GCP)
CI/CD
DevOps
+9 more

Job Description:

  • Infrastructure Management: Design, implement, and manage scalable, reliable, and secure cloud infrastructure using AWS, GCP, and/or Azure.
  • CI/CD Pipelines: Develop and maintain continuous integration and continuous deployment (CI/CD) pipelines to streamline the development lifecycle.
  • Automation: Automate infrastructure provisioning, configuration management, and application deployment processes.
  • Monitoring and Performance: Implement monitoring, logging, and alerting solutions to ensure system health, performance, and reliability.
  • Security: Ensure the security of cloud infrastructure and applications, including identity management and compliance with industry standards.
  • Collaboration: Work closely with client and development teams to integrate DevOps practices and deliver high-quality software.
  • Documentation: Maintain comprehensive documentation of infrastructure, configurations, and processes.
  • Innovation: Stay current with emerging technologies and industry trends, integrating them into the DevOps strategy as appropriate.


Qualifications: 

  • Education: Bachelor's degree in Computer Science, Information Technology, or a related field.
  • Experience: 7 - 10 years of overall experience with relevant experience of at least 7 years in DevOps and served as a lead or senior engineer.
Read more
Gurugram
5 - 10 yrs
₹12L - ₹18L / yr
DevOps
Reliability engineering
skill iconAmazon Web Services (AWS)
Terraform
Ansible
+18 more

Job Title : DevOps Engineer / Site Reliability Engineer (SRE)

Experience : 5+ Years

Location : Gurugram, Haryana

Work Mode : On-site (Full-time)


About the Role :

We are looking for a skilled DevOps Engineer with 5+ years of experience in cloud infrastructure, CI/CD, automation, Kubernetes, and Site Reliability Engineering (SRE). The ideal candidate will be responsible for building scalable cloud infrastructure, automating deployments, improving system reliability, and ensuring high availability across production environments.


Mandatory Skills :

AWS, Terraform, Ansible, CloudFormation, Jenkins, GitLab CI, GitHub Actions, Docker, Kubernetes, Helm, Python, Bash, Grafana, Prometheus, ELK Stack, CloudWatch, New Relic, SRE, CI/CD, Infrastructure as Code (IaC), Linux


Key Responsibilities :

  • Design, deploy, and manage cloud infrastructure primarily on AWS (EC2, VPC, IAM, S3, RDS, Route53, ALB, Auto Scaling, Lambda).
  • Build and maintain Infrastructure as Code (IaC) using Terraform, Ansible, and CloudFormation.
  • Develop and optimize CI/CD pipelines using Jenkins, GitLab CI, and GitHub Actions.
  • Deploy and manage containerized applications using Docker, Kubernetes, and Helm.
  • Implement monitoring and observability using Grafana, Prometheus, ELK Stack, CloudWatch, and New Relic.
  • Drive SRE practices by defining SLIs, SLOs, SLAs, handling production incidents, conducting RCA, and improving system reliability.
  • Automate operational tasks using Python, Bash, and Groovy scripting.
  • Collaborate with Development, QA, Security, and Operations teams to ensure reliable and secure software delivery.


Required Skills & Qualifications :

  • Bachelor's degree in Computer Science, IT, Electronics, or a related field.
  • 5+ years of experience in DevOps, SRE, or Cloud Infrastructure.
  • Strong expertise in AWS, with exposure to Azure/GCP.
  • Hands-on experience with Terraform, Ansible, CloudFormation, Docker, Kubernetes, Helm, Jenkins, GitLab CI, GitHub Actions, and Git.
  • Strong scripting skills in Python and Bash.
  • Experience with monitoring tools such as Grafana, Prometheus, ELK Stack, CloudWatch, and New Relic.
  • Good understanding of Linux, networking, SQL, and cloud security best practices.


Preferred Skills :

  • Experience with multi-cloud environments and DevSecOps practices.
  • Knowledge of disaster recovery, automation, and microservices architecture.
  • Strong troubleshooting, communication, and problem-solving skills.
Read more
Gurugram
4 - 8 yrs
₹6L - ₹15L / yr
DevOps
Linux administration
skill iconAmazon Web Services (AWS)
skill iconDocker
skill iconKubernetes
+20 more

Job Title : DevOps Engineer / Site Reliability Engineer (SRE)

Experience : 4+ Years

Location : Gurugram, Sector 48, Haryana (On-site)

Employment Type : Full-Time

Working Days : Monday to Saturday (1st & 3rd Saturday Off)


About the Role :

We are looking for a hands-on DevOps Engineer / Site Reliability Engineer (SRE) with strong experience in Linux, AWS, Kubernetes, Docker, CI/CD, Infrastructure as Code, and production application deployments.

The ideal candidate should have real-world production experience, excellent troubleshooting skills, and the ability to manage both infrastructure and application-level issues.


Mandatory Skills :

Linux, AWS, Docker, Kubernetes, Terraform, Ansible, Jenkins, GitHub Actions, GitLab CI/CD, CI/CD, Infrastructure as Code (IaC), Python, Bash, Git, Grafana, Prometheus, ELK, CloudWatch, New Relic, SRE (SLI/SLO/SLA), Networking (DNS, HTTP/HTTPS, TCP/IP, Load Balancer), Production Application Deployment & Troubleshooting


Key Responsibilities :

  • Manage and maintain AWS cloud infrastructure.
  • Build and optimize CI/CD pipelines using Jenkins, GitHub Actions, or GitLab CI.
  • Deploy, monitor, and troubleshoot applications across production environments.
  • Automate infrastructure using Terraform and Ansible.
  • Manage Docker containers and Kubernetes clusters.
  • Monitor systems using Grafana, Prometheus, ELK, CloudWatch, and New Relic.
  • Perform Linux server administration and troubleshooting.
  • Handle production incidents, Root Cause Analysis (RCA), and improve system reliability.
  • Collaborate with development teams to support application releases and automation.


Required Qualifications :

  • Bachelor's degree in Computer Science or related field.
  • 4+ years of hands-on experience in DevOps / SRE.
  • Strong Linux administration and production troubleshooting skills.
  • Experience with AWS and modern DevOps toolchains.
  • Hands-on experience with application deployment and production support.


What We're Looking For :

  • Strong practical Linux and cloud knowledge.
  • Real production experience with application deployments.
  • Ability to troubleshoot both infrastructure and application issues.
  • Experience handling live production incidents.
  • Excellent communication and problem-solving skills.
  • Candidates should be comfortable with scenario-based technical discussions and demonstrate genuine hands-on expertise.


Interview Process :

  1. HR Screening
  2. Technical Round
  3. Client Technical Round
  4. Final Discussion


Note : The interview will focus on practical hands-on experience in Linux, AWS, Kubernetes, Docker, CI/CD, Infrastructure as Code, application deployment, production troubleshooting, and real-world DevOps scenarios.

Read more
Why apply to jobs via Cutshort
people_solving_puzzle
Personalized job matches
Stop wasting time. Get matched with jobs that meet your skills, aspirations and preferences.
people_verifying_people
Verified hiring teams
See actual hiring teams, find common social connections or connect with them directly.
ai_chip
Move faster with AI
We use AI to get you faster responses, recommendations and unmatched user experience.
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos