Cutshort logo
For Employers
Wissen Technology logo
IAC SRE Engineers

IAC SRE Engineers at Wissen Technology · Bengaluru (Bangalore), Mumbai, Pune · 4 - 8 years · Profitable · Posted 9 Jul 2025

Wissen Technology's logo

IAC SRE Engineers

Moulina Dey's profile picture
Posted by Moulina Dey
4 - 8 yrs
Best in industry
Bengaluru (Bangalore), Mumbai, Pune
Skills
skill iconJava
Data Structures
DSA
Akamai
WAF

Job Title: IAC SRE Engineer

Location: Pune, Mumbai, Bangalore

Experience Required: 4 Years

Role Overview:

We are looking for experienced IAC Engineers with a strong background in Akamai, Data Structures & Algorithms (DSA), Java, and DevSecOps. The ideal candidate should have hands-on development experience, be proficient in writing Infrastructure as Code using Terraform, and demonstrate strong problem-solving skills.

Core Skills:

  • Akamai – Strong experience in CDN, caching, and performance optimization.
  • Data Structures & Algorithms (DSA) – Strong problem-solving and coding abilities.
  • Java – Solid programming background and experience in development.
  • DevSecOps – Understanding of integrating security in CI/CD pipelines and infrastructure.

Good to Have:

  • WAF (Web Application Firewall) – Knowledge of WAF is a plus, though not mandatory.

Additional Skills:

  • Experience with SRE (Site Reliability Engineering) practices is beneficial.
  • Strong hands-on with Terraform for managing cloud infrastructure.


Read more
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos

About Wissen Technology

Founded :
2015
Type :
Products & Services
Size :
1000-5000
Stage :
Profitable

About

The Wissen Group was founded in the year 2000. Wissen Technology, a part of Wissen Group, was established in the year 2015. Wissen Technology is a specialized technology company that delivers high-end consulting for organizations in the Banking & Finance, Telecom, and Healthcare domains.

With offices in US, India, UK, Australia, Mexico, and Canada, we offer an array of services including Application Development, Artificial Intelligence & Machine Learning, Big Data & Analytics, Visualization & Business Intelligence, Robotic Process Automation, Cloud, Mobility, Agile & DevOps, Quality Assurance & Test Automation.


Leveraging our multi-site operations in the USA and India and availability of world-class infrastructure, we offer a combination of on-site, off-site and offshore service models. Our technical competencies, proactive management approach, proven methodologies, committed support and the ability to quickly react to urgent needs make us a valued partner for any kind of Digital Enablement Services, Managed Services, or Business Services.


We believe that the technology and thought leadership that we command in the industry is the direct result of the kind of people we have been able to attract, to form this organization (you are one of them!).


Our workforce consists of 1000+ highly skilled professionals, with leadership and senior management executives who have graduated from Ivy League Universities like MIT, Wharton, IITs, IIMs, and BITS and with rich work experience in some of the biggest companies in the world.


Wissen Technology has been certified as a Great Place to Work®. The technology and thought leadership that the company commands in the industry is the direct result of the kind of people Wissen has been able to attract. Wissen is committed to providing them the best possible opportunities and careers, which extends to providing the best possible experience and value to our clients.

Read more

Connect with the team

Profile picture
Lokesh Manikappa
Profile picture
Vijayalakshmi Selvaraj
Profile picture
Adishi Sood
Profile picture
Shiva Kumar J Goud

Company social profiles

bloglinkedinfacebook

Similar jobs (10)

Gurugram
5 - 10 yrs
₹12L - ₹18L / yr
DevOps
Reliability engineering
skill iconAmazon Web Services (AWS)
Terraform
Ansible
+18 more

Job Title : DevOps Engineer / Site Reliability Engineer (SRE)

Experience : 5+ Years

Location : Gurugram, Haryana

Work Mode : On-site (Full-time)


About the Role :

We are looking for a skilled DevOps Engineer with 5+ years of experience in cloud infrastructure, CI/CD, automation, Kubernetes, and Site Reliability Engineering (SRE). The ideal candidate will be responsible for building scalable cloud infrastructure, automating deployments, improving system reliability, and ensuring high availability across production environments.


Mandatory Skills :

AWS, Terraform, Ansible, CloudFormation, Jenkins, GitLab CI, GitHub Actions, Docker, Kubernetes, Helm, Python, Bash, Grafana, Prometheus, ELK Stack, CloudWatch, New Relic, SRE, CI/CD, Infrastructure as Code (IaC), Linux


Key Responsibilities :

  • Design, deploy, and manage cloud infrastructure primarily on AWS (EC2, VPC, IAM, S3, RDS, Route53, ALB, Auto Scaling, Lambda).
  • Build and maintain Infrastructure as Code (IaC) using Terraform, Ansible, and CloudFormation.
  • Develop and optimize CI/CD pipelines using Jenkins, GitLab CI, and GitHub Actions.
  • Deploy and manage containerized applications using Docker, Kubernetes, and Helm.
  • Implement monitoring and observability using Grafana, Prometheus, ELK Stack, CloudWatch, and New Relic.
  • Drive SRE practices by defining SLIs, SLOs, SLAs, handling production incidents, conducting RCA, and improving system reliability.
  • Automate operational tasks using Python, Bash, and Groovy scripting.
  • Collaborate with Development, QA, Security, and Operations teams to ensure reliable and secure software delivery.


Required Skills & Qualifications :

  • Bachelor's degree in Computer Science, IT, Electronics, or a related field.
  • 5+ years of experience in DevOps, SRE, or Cloud Infrastructure.
  • Strong expertise in AWS, with exposure to Azure/GCP.
  • Hands-on experience with Terraform, Ansible, CloudFormation, Docker, Kubernetes, Helm, Jenkins, GitLab CI, GitHub Actions, and Git.
  • Strong scripting skills in Python and Bash.
  • Experience with monitoring tools such as Grafana, Prometheus, ELK Stack, CloudWatch, and New Relic.
  • Good understanding of Linux, networking, SQL, and cloud security best practices.


Preferred Skills :

  • Experience with multi-cloud environments and DevSecOps practices.
  • Knowledge of disaster recovery, automation, and microservices architecture.
  • Strong troubleshooting, communication, and problem-solving skills.
Read more
company logo
Arshiya Shaikh
Posted by Arshiya Shaikh
Mumbai
4 - 7 yrs
₹5L - ₹9L / yr
AWS CloudFormation
Bitbucket
skill iconDocker
skill iconKubernetes

Key Responsibilities

  • Automate application deployments from Bitbucket to servers using CI/CD pipelines.
  • Design and manage scalable, highly available AWS infrastructure.
  • Implement Auto Scaling, ELB, and Route 53 for traffic management and high availability.
  • Work with AWS services including IAM, RDS, DynamoDB, EC2, and other cloud services.
  • Build and manage Docker containers and server images.
  • Deploy and manage applications using Kubernetes.
  • Implement Infrastructure as Code using Terraform, CloudFormation, or Ansible.
  • Develop automation scripts using Python and Bash.
  • Implement monitoring and logging using tools such as Prometheus, Grafana, and ELK.
  • Integrate security and compliance practices into CI/CD pipelines.
  • Optimize infrastructure for security, scalability, performance, and cost.

Required Skills

  • 3+ years of experience in DevOps or a similar role.
  • Strong knowledge of AWS beyond EC2.
  • Hands-on experience with Jenkins or similar CI/CD tools.
  • Experience with Docker and Kubernetes.
  • Good understanding of Terraform/IaC and automation.
  • Proficiency in Python and/or Bash scripting.
  • Knowledge of DevSecOps, security, and compliance best practices.
  • Strong troubleshooting and problem-solving skills.


Read more
Gurugram
4 - 10 yrs
₹4L - ₹10L / yr
DevOps
Site Reliability Engineer (SRE)
skill iconAmazon Web Services (AWS)
skill iconDocker
skill iconKubernetes
+14 more

🚀 Job Title : DevOps Engineer / Site Reliability Engineer (SRE)

Experience Level : 4+ Years

Location : Gurugram Sector 48, Haryana (On-site)

Employment Type : Full Time Opportunity


About the Role :

We are looking for a proactive DevOps / Site Reliability Engineer (SRE) with around 4 years of hands-on experience designing, automating, and scaling cloud infrastructure and CI/CD delivery pipelines.

In this role, you will bridge the gap between development and operations. You will be responsible for orchestrating containerized applications, automating infrastructure via Code (IaC), establishing SRE best practices (SLIs, SLOs, SLAs), and ensuring maximum uptime, resiliency, and operational efficiency across multi-cloud environments (AWS/Azure/GCP).


Mandatory Skills :

AWS, Kubernetes, Docker, Terraform, Ansible, Jenkins, GitLab CI/CD, GitHub Actions, Python, Bash, CI/CD, Infrastructure as Code (IaC), Grafana, Prometheus, ELK, New Relic, CloudWatch, SRE, SLI/SLO/SLA, Linux


Key Responsibilities :

1. Cloud Infrastructure & Infrastructure as Code (IaC) :

  • Provision, configure, and maintain scalable, high-availability infrastructure on multi-cloud platforms, primarily AWS (EC2, VPC, IAM, S3, RDS, Route53, ALB/ASG, Lambda, EBS).
  • Build, deploy, and manage Infrastructure as Code (IaC) using Terraform, Ansible, and CloudFormation to enforce consistency and eliminate configuration drift.
  • Execute disaster recovery (DR) planning, automated failover / failback mechanisms, and chaos engineering exercises to validate system resiliency.

2. CI/CD, Automation & Development :

  • Design, end-to-end maintain, and optimize robust CI/CD pipelines using Jenkins, GitLab CI, and GitHub Actions.
  • Automate release pipelines, versioning, branching strategies, and approval gates using Groovy, Python, and Bash scripting. Integrate automated code quality and security scanning tools (SonarQube, Black Duck, or Fortify) directly into delivery pipelines.
  • Develop custom tools, scripts, or microservices (e.g., Python / Node.js) to automate manual operational tasks and operational toil.

3. Containerization & Orchestration :

  • Onboard and orchestrate containerized microservices utilizing Docker and Kubernetes (including Helm charts).
  • Ensure high availability, auto-scaling, resource management, and fault tolerance for Kubernetes pod deployments.

4. Observability, SRE & Incident Management :

  • Drive Site Reliability Engineering (SRE) maturity by establishing, tracking, and reporting SLIs, SLOs, and SLAs with cross-functional engineering teams.
  • Build, configure, and manage full-stack observability tools : Grafana, Prometheus, New Relic, Elasticsearch / Logstash / Kibana (ELK), Sentry, and AWS CloudWatch.
  • Set up real-time alerting, custom metric dashboards, and automated log rotation / pruning scripts.
  • Handle production incidents, lead Root Cause Analysis (RCA) investigations, and implement preventive measures to reduce Mean Time to Resolution (MTTR).


Required Qualifications & Skills :

  • Education : Bachelor’s Degree in Electronics and Communication Engineering, Computer Science, or a related technical field.
  • Experience : ~4 years of experience in DevOps, SRE, or Cloud System Administration roles.
  • Cloud & Infrastructure : Hands-on experience with AWS (Core services like EC2, S3, VPC, RDS, IAM, Lambda, Auto Scaling) and exposure to Azure / GCP.
  • CI/CD & Version Control : Proficiency with Jenkins, GitLab CI, GitHub Actions, and Git workflows.
  • Containerization : Core proficiency in Docker and Kubernetes cluster management / onboarding.
  • Infrastructure as Code : Expertise in Ansible, Terraform, or AWS CloudFormation.
  • Scripting & Languages : Strong hands-on automation skills with Python, Bash, and foundational knowledge of Node.js, Java or C++.
  • Observability & Logging : Strong experience with Grafana, Prometheus, New Relic, ELK stack, or Splunk.
  • Database & SQL : Familiarity with relational databases (MySQL, RDS) for monitoring setup and operational analytics.
Read more
company logo
Remote, Pune
5 - 10 yrs
₹20L - ₹32L / yr
Infrastructure Platform Engineer
skill iconAmazon Web Services (AWS)
Terraform
AWS CloudFormation
skill iconKubernetes
+13 more

Job Title : SDE 3 – Infrastructure Platform Engineer

Experience : 5.5 to 8.5 Years

Number of Positions : 2

Employment Type : C2H (Contract to Hire)

Work Mode : Remote during contractual period → 5 Days WFO after conversion

Contract Duration : 3 Months

Post-Conversion Location : Pune

Notice Period : Immediate Joiners / Serving Notice Period / Up to 15 Days preferred

(Candidates officially serving a 30-day notice period may also be considered if they are on the bench and have a negotiable joining date)


Role Overview :

We are looking for an experienced SDE 3 – Infrastructure Platform Engineer to design, build, and operate scalable, secure, and highly reliable cloud infrastructure and internal platform capabilities.


The ideal candidate will have strong hands-on experience in Cloud Infrastructure, Infrastructure as Code (IaC), CI/CD, Docker, Kubernetes, automation, observability, networking, and distributed systems.


Mandatory Skills : AWS / Azure / GCP, Terraform / CloudFormation, Kubernetes, Docker, CI/CD, Platform / Infrastructure Engineering, Python / Go / Java / Ruby, Networking, Cloud Security, Distributed Systems, Scalability & Reliability, Strong Coding & Automation.


Key Responsibilities :

  • Design and maintain scalable, highly available infrastructure on AWS / GCP / Azure.
  • Build and manage Infrastructure as Code (IaC) using Terraform, CloudFormation, or similar tools.
  • Develop automation for infrastructure provisioning, deployments, monitoring, and operations.
  • Manage and optimize Docker and Kubernetes workloads.
  • Build internal platform tools to improve developer productivity and engineering efficiency.
  • Implement monitoring, logging, alerting, and observability solutions.
  • Participate in incident response, RCA, postmortems, and reliability improvements.
  • Design and improve CI/CD pipelines and deployment automation.
  • Contribute to system design, architecture discussions, scalability, security, and cost optimization.
  • Collaborate with application, data, and product engineering teams.


Required Skills :

  • 5.5 to 8.5 years of experience in Infrastructure / Platform Engineering or similar roles.
  • Strong hands-on experience with AWS, GCP, or Azure.
  • Strong expertise in Terraform / CloudFormation.
  • Experience with CI/CD, Docker, and Kubernetes.
  • Strong programming skills in at least one of:
  • Python, Go, Java, or Ruby.
  • Good understanding of networking, cloud security, distributed systems, scalability, and reliability.
  • Experience working with production infrastructure and highly available systems.
  • Strong troubleshooting and problem-solving skills.


Nice to Have :

  • Experience with SRE practices and production on-call ownership.
  • Experience in fintech, payments, banking, or transaction-heavy systems.
  • Knowledge of cloud security, compliance, or FinOps/cost optimization.
  • Experience building internal developer platforms or productivity tools.
  • Previous product company experience.


Interview Process :

Round 1 : Take-Home Coding Assignment – Submit within 48 hours

Round 2 : Coding Assignment Discussion – 1 Hour

Round 3 : Technical Managerial Round – 30 Minutes


Note : The take-home coding assignment is mandatory. Candidates should be comfortable completing and submitting the assignment within 48 hours before proceeding.


Ideal Candidate :

Strong Platform / Infrastructure Engineer with hands-on experience in :

Cloud + Terraform / CloudFormation + Kubernetes + CI/CD + Programming + SRE / Production Operations


Pure DevOps profiles without strong coding and platform engineering experience are not preferred.

Read more
It is an Product Based Company(Domain- EV Charging)
It is an Product Based Company(Domain- EV Charging)
Agency job
via by Mantasha Naaz
Bengaluru (Bangalore)
6 - 8 yrs
₹18L - ₹20L / yr
SRE
Reliability engineering
on call Support
Incident management
skill iconAmazon Web Services (AWS)

Job Title: Senior Site Reliability Engineer 

Location: Bengaluru, India (Hybrid)

Employment Type: Full-time

Experience: 6+ years

About Compnay

It is driving the electric mobility revolution through cutting-edge software, infrastructure, and professional services. Our technology empowers utilities, cities, fleets, transit agencies, and automakers to deploy EV charging infrastructure at scale safely, efficiently, and sustainably. With a global footprint spanning three continents and operations in 13 countries, we are passionate about shaping the future of sustainable transport.

Operating over 70,000 charge points globally, It is driving the transition toward cleaner, smarter, and more efficient mobility. The India team serves as a critical operational hub, supporting global platforms focused on decarbonization, digitalization, and scalable infrastructure growth.

We value purpose-driven individuals who want to make a meaningful impact and help create a cleaner, smarter, and more connected world.

Role Overview

We are seeking a skilled and proactive Site Reliability Engineer (SRE) to join our growing team. In this role, you will be responsible for maintaining system reliability, scalability, and performance across our EV charging platforms. You will collaborate closely with development and operations teams to build resilient, automated, and observable systems.

Key Responsibilities

  • Ensure high availability, performance, and reliability of production systems
  • Design, implement, and manage scalable infrastructure solutions
  • Build and maintain CI/CD pipelines for efficient software delivery
  • Monitor system health using observability tools and respond to incidents proactively
  • Automate operational processes using scripting and Infrastructure as Code (IaC)
  • Manage containerized environments using Docker and Kubernetes
  • Collaborate with cross-functional teams to improve system architecture and resilience
  • Participate in on-call rotations and incident management processes
  • Continuously optimize cloud infrastructure for cost, performance, and scalability

Required Qualifications & Skills

  • Bachelor’s degree in Computer Science, IT, or related field
  • 4+ years of experience in Site Reliability Engineering, DevOps, or Infrastructure roles
  • Strong experience with containerization (Docker) and orchestration (Kubernetes)
  • Proficiency in Linux administration, networking, and system security
  • Hands-on experience with cloud platforms, especially AWS (EKS, EC2, S3, RDS, Lambda)
  • Experience with CI/CD tools such as Jenkins, GitLab CI/CD, or similar
  • Knowledge of Infrastructure as Code tools (Terraform, AWS CloudFormation, Ansible)
  • Proficiency in scripting languages (Python, Bash, or PowerShell)
  • Experience with monitoring tools like Dynatrace, Prometheus, Grafana, or Zabbix
  • Solid understanding of system architecture, microservices, and SaaS/PaaS models
  • Strong analytical and problem-solving skills   

What We Offer

  • Work with some of the brightest minds in the emerging EV industry.
  • Make a tangible impact in reducing carbon emissions and enabling sustainable energy.
  • Freedom to suggest, implement, and innovate on systems, processes, and technologies.
  • Daily ownership in a high-growth, challenging environment.
  • Flexible work environment with hybrid schedules and virtualization options.
  • Competitive pay and benefits including health coverage, innovative PTO program, and performance bonuses.


Read more
Remote only
8 - 12 yrs
Best in industry
Terraform
Artificial Intelligence (AI)
IAC
skill iconAmazon Web Services (AWS)
ECS
+6 more


Senior Platform & Site Reliability Engineer

Location: Remote Employment Type: Contract

The Role

This role carries full architectural and operational ownership of the platform layer across a growing SaaS portfolio. The Cloud Architect owns AWS infrastructure standards — VPCs, account structures, networking, and compute design. Everything outside that lane is yours: the CI/CD platform, the observability and reliability stack, the event streaming infrastructure, the deployment pipelines, and the incident engineering model.

Architectural decisions are yours to make and defend, standards are yours to define and enforce, and the reliability of 20+ enterprise SaaS products depends on what you and your team build.

This is an AI-native engineering organisation. Where it is practical and safe to do so, you are expected to use automation and AI-assisted tooling to reduce toil — in CI/CD triage, infrastructure provisioning, observability workflows, and acquisition onboarding. The expectation is not to replace engineering judgement with automation, but to free it up for the problems that genuinely require it.

The Scale You Will Operate At

The portfolio consists of 20+ live, enterprise-grade SaaS solutions running concurrently. Each product serves enterprise customers and processes millions to billions of real-time requests. The architecture is serious: event streaming for real-time data pipelines, batch processing workloads running alongside live transaction flows, and multi-tenant enterprise-grade reliability expectations across every product.

You will design and operate the platform infrastructure that underpins all of it — scaling horizontally as each new acquisition joins the portfolio, without proportionally scaling cost, complexity, or headcount.

What You Will Own

Platform Architecture

  • Full architectural ownership of the non-AWS toolchain: CI/CD, observability, event streaming, automation, secrets, and deployment infrastructure
  • Define, build, and enforce platform standards across portfolio products
  • Terraform IaC for all infrastructure — nothing provisioned manually, everything versioned and reviewed
  • Self-service developer platform so product teams ship without waiting on platform

Event Streaming & Pipeline Infrastructure

  • Own the event streaming architecture, operational standards, and health monitoring across all products using real-time pipelines
  • Design and maintain batch processing infrastructure alongside live event flows
  • Ensure pipeline reliability, throughput, and cost are actively managed at scale

CI/CD & Deployment

  • Build and maintain CI/CD pipelines (GitHub Actions) across all portfolio products
  • Automate triage and retry logic for known failure classes — flaky tests, dependency timeouts, OOM kills — so engineers are only paged for genuinely novel failures
  • Deployment standards: release management, rollback mechanisms, canary and blue-green patterns where justified

Observability & Reliability

  • Own the full observability stack: Grafana, Prometheus, and Loki across all products
  • SLOs and error budgets defined per product; reliability tracked consistently
  • Build alerting that correlates signals and surfaces diagnostic context alongside notifications — so on-call engineers arrive at an incident with hypotheses, not a blank screen
  • Incident response: on-call design, escalation playbooks, post-mortem facilitation
  • Automated remediation scoped to safe, idempotent actions — container restarts, ECS task scaling, known rollback patterns; novel or ambiguous failures escalate to a human with full context attached

Acquisition Onboarding

  • Platform audit and gap analysis for every new acquisition — assessing CI/CD maturity, IaC coverage, observability gaps, and security posture
  • Migration plan and execution for each portfolio company joining the platform — sequenced to avoid disrupting live operations
  • Target: full platform integration within a defined window per acquisition

A Note on Automation

Where automation is safe and failure modes are well understood — routine provisioning, known CI/CD failure classes, secrets rotation, cost anomaly flagging — aggressive automation is expected. Where automation would act on ambiguous signals or carry significant blast radius, human judgement stays in the loop. The goal is to reduce toil on solved problems, not to automate decisions that require engineering expertise.

Platform Stack

Area Stack / Standard IaC Terraform OSS / OpenTofu CI/CD GitHub Actions Event Streaming Architecture and tooling chosen for the workload Observability Grafana, Prometheus, Loki Log Management AWS CloudWatch, Grafana Loki Incident Management OpsGenie (startup tier) or Better Uptime Secrets AWS Secrets Manager / HashiCorp Vault OSS Containers ECS (default), EKS only where justified Cost Monitoring AWS Cost Explorer with custom dashboards What We’re Looking For

  • 8–12 years in platform engineering, DevOps, or SRE — with clear evidence of increasing ownership over time
  • Strong Terraform depth across multi-environment, multi-account setups
  • CI/CD ownership across a multi-product environment with GitHub Actions
  • Experience with event streaming infrastructure at production scale — design, operations, reliability, and cost management
  • Hands-on Grafana, Prometheus, and Loki in production
  • AWS operational depth: ECS, EKS, RDS, IAM, VPC, CloudWatch, Cost Explorer
  • SRE fundamentals: SLOs, error budgets, on-call design, post-mortem culture
  • Acquisition or greenfield platform integration experience strongly preferred

How You Work

  • Comfortable operating across multiple products simultaneously — context-switching without dropping standards
  • Cost-efficiency instinct — you optimise spend as a habit, not as a project
  • You treat automation as a tool for eliminating toil, not a substitute for engineering judgement
  • You document decisions, enforce standards through code, and build platforms that other engineers find intuitive to use

Why This Role

The platform function is being built from the ground up. You will have architectural ownership of the entire non-AWS platform layer across a growing portfolio of enterprise SaaS products, with the freedom — and responsibility — to build the reliability and delivery culture of the organisation.

This is not a role that inherits someone else’s decisions and maintains them. Every major architectural choice is still to be made. If you want to build something that lasts and that other engineers depend on, this is the role.

Read more
company logo
Ganesh Ram
Posted by Ganesh Ram
Bengaluru (Bangalore), Mumbai, Delhi, Gurugram, Noida, Ghaziabad, Faridabad, Hyderabad, Pune
7 - 10 yrs
₹15L - ₹20L / yr
CI/CD
skill iconKubernetes
helm
Terraform
yaml

Cloud Expertise(Azure):

• Strong understanding of cloud services and resources like AI services, webapp, database, including monitoring tools like Azure Monitor and Log Analytics.

• Experience with Infrastructure as Code (IaC) tools such as Arm template / Bicep/Terraform.

• Deep understanding of Networking concepts(DNS, DHCP , Hub and Spoke).

• Understanding on policies and security aspects of cloud.


Kubernetes & Helm:

• In-depth knowledge of Kubernetes concepts such as pods, services, ingress, config maps, and secrets.

• Understand of Kubernetes templates and its deployment.

• Proficiency with Helm/ Kustomize or equivalent for Kubernetes package management and deployment automation.

• Implement Kubernetes best practices, including security, networking, and scaling.

• Concepts of Docker and Containers


CI/CD & Programming:

• Hands-on experience with YAML-based CI/CD pipelines (e.g., Azure DevOps, GitHub Actions).

• Familiarity with scripting and automation tools such as PowerShell, Azure CLI, or Bash.

• Proven skill in python programming and concepts.


Monitoring and Observability:

Expertise in creating and managing Grafana dashboards for visualizing metrics and logs.

• Knowledge of Log Analytics & Azure Application Insights for performance monitoring and tracing.

 

Read more
company logo
Bhattacharjee Akash
Posted by Bhattacharjee Akash
Bengaluru (Bangalore), Chennai, Mumbai, Hyderabad, Pune, Gurugram
3 - 10 yrs
₹12L - ₹35L / yr
Linux/Unix
skill iconKubernetes
Monitoring
skill iconDocker
skill iconAmazon Web Services (AWS)
+4 more



We're looking for a Site Reliability Engineer to keep our production systems fast, reliable, and scalable. Sitting at the intersection of software engineering and operations, you'll treat infrastructure as code, automate away toil, and build the observability that lets us catch problems before customers do. You'll own uptime and on-call for critical services, lead incident response and blameless postmortems, and continuously harden the platform against failure. This role suits an engineer who is as comfortable debugging a production incident at 2 a.m. as they are writing the automation that prevents the next one.



Key Responsibilities

  • Own reliability, availability, and performance of production services, including on-call rotation
  • Build and maintain monitoring, alerting, and observability (metrics, logs, traces)
  • Automate deployments, scaling, and operational tasks to reduce manual toil
  • Manage containerized workloads on Kubernetes and cloud infrastructure
  • Design and maintain CI/CD pipelines for safe, frequent releases
  • Lead incident response and drive blameless postmortems with clear follow-ups
  • Perform capacity planning, performance tuning, and cost optimization
  • Define and track SLIs/SLOs and error budgets with product teams


Requirements

  • 3+ years in SRE, DevOps, or production-focused engineering
  • Strong Linux administration and hands-on Kubernetes experience
  • Solid experience with monitoring/observability tools (Prometheus, Grafana, ELK, or similar)
  • Cloud experience with AWS, GCP, or Azure
  • CI/CD pipelines and infrastructure-as-code (Terraform, CloudFormation)
  • Proficient scripting in Python and/or Bash


Nice to have

  • Experience with service meshes, Helm, or GitOps (ArgoCD/Flux)
  • Background in high-traffic or distributed systems
Read more
company logo
Srikanth Bajgur
Posted by Srikanth Bajgur
Bengaluru (Bangalore)
8 - 10 yrs
Best in industry
skill iconAmazon Web Services (AWS)
skill iconPython
Terraform
Microsoft Windows Server administration
CI/CD
+1 more


We are seeking a highly skilled Senior DevOps Engineer with 8+ years of professional experience to join our team. In this role, you will design, implement, and optimize cloud infrastructure and CI/CD processes.

You will collaborate closely with development, QA, and operations teams to deliver scalable, secure, automated, and reliable solutions on AWS.


The ideal candidate will have strong hands-on experience with AWS, Terraform, Git/GitHub, Jenkins, PowerShell, Python, AWS Systems Manager (SSM) Documents, and Active Directory (AD) administration, along with a passion for automation, efficiency, and operational excellence.


Key Responsibilities


  • Design, build, and maintain scalable cloud infrastructure on AWS.
  • Develop and manage Infrastructure as Code (IaC) using Terraform.
  • Build, maintain, and optimize CI/CD pipelines using Jenkins and GitHub.
  • Use AWS Systems Manager (SSM) to support operational automation and system administration.
  • Automate system tasks and administrative workflows using PowerShell and other scripting languages.
  • Create, maintain, and execute custom SSM Documents for configuration management, patching, automation, and troubleshooting.
  • Manage and administer Active Directory (AD), including users, groups, permissions, policies, authentication, and integration with AWS services.
  • Implement and manage version-control workflows in Git and GitHub.
  • Ensure infrastructure and deployment processes follow best practices for security, reliability, scalability, and cost optimization.
  • Monitor and troubleshoot production systems to ensure high availability, performance, and reliability.
  • Collaborate with development teams to improve software delivery processes and release management.
  • Mentor junior engineers and contribute to the development of DevOps standards and best practices.


Required Skills and Experience


  • 8+ years of professional experience, including at least 4 years in DevOps or Site Reliability Engineering (SRE) roles.
  • Strong hands-on experience with AWS services, including EC2, VPC, IAM, S3, EKS, Lambda, and related services.
  • Proven expertise in Terraform for Infrastructure as Code.
  • Experience administering both Linux and Windows operating systems.
  • Experience designing and managing CI/CD pipelines using Jenkins and GitHub.
  • Proficiency with Git workflows and source-code management best practices.
  • Strong PowerShell scripting skills; familiarity with Python and/or Bash is a plus.
  • Experience with AWS Systems Manager (SSM), including creating and managing SSM Documents.
  • Hands-on experience managing Active Directory, including users, groups, policies, authentication, permissions, and AWS integration.
  • Solid understanding of cloud networking, security, monitoring, and troubleshooting.
  • Excellent problem-solving, communication, collaboration, and decision-making skills.


Nice-to-Have Skills


  • Experience with containerization and orchestration technologies, such as Docker, Kubernetes, and Amazon EKS.
  • Knowledge of monitoring and observability tools, such as Amazon CloudWatch, New Relic, and Sumo Logic.
  • An AWS certification, such as AWS Certified DevOps Engineer – Professional or AWS Certified Solutions Architect.


Who You Are


  • You are eager to learn new technologies and continuously improve your skills.
  • You make sound decisions and take ownership of your work.
  • You are proactive, self-motivated, and comfortable taking initiative.
  • You are a strong communicator who enjoys collaborating with cross-functional teams.
  • You are committed to improving processes, automation, reliability, and operational efficiency.
Read more
company logo
Soni Sharma
Posted by Soni Sharma
Delhi
3 - 8 yrs
₹5L - ₹12L / yr
DevOps
skill iconPython

About the Role

We are looking for an experienced AWS DevOps Engineer with around 4 years of hands-on experience to join our infrastructure/DevOps team. The ideal candidate will independently design, deploy, and maintain cloud infrastructure primarily on AWS, drive automation initiatives, mentor junior engineers, and work closely with cross-functional teams to build scalable, secure, and highly available systems. Exposure to Azure is a strong plus.

Key Responsibilities

·      Design, deploy, and maintain robust, scalable, and secure infrastructure on AWS

·      Architect and manage core AWS services such as EC2, S3, VPC, IAM, RDS, Lambda, ECS/EKS, Route 53, and CloudFront

·      Build, own, and optimize CI/CD pipelines (e.g., CodePipeline, Jenkins, GitLab CI, GitHub Actions) to enable fast and reliable deployments

·      Design and implement Infrastructure as Code (IaC) using Terraform / AWS CloudFormation

·      Set up and manage monitoring, logging, and alerting solutions (CloudWatch, ELK, Prometheus, Grafana, Datadog, etc.)

·      Implement and enforce security best practices including IAM policies, security groups, NACLs, KMS, Secrets Manager, and compliance standards

·      Lead troubleshooting and root cause analysis for infrastructure, deployment, and production incidents

·      Drive backup, disaster recovery, high-availability, and cost-optimization strategies (Reserved Instances, Savings Plans, right-sizing)

·      Containerize applications and manage orchestration using Docker, ECS, and/or Kubernetes (EKS)

·      Automate repetitive operational tasks through scripting and tooling

·      Support any hybrid or multi-cloud initiatives involving Azure services, where applicable

·      Mentor junior engineers and review their work, providing technical guidance

·      Collaborate with development, QA, security, and product teams to support and streamline application deployments

·      Maintain comprehensive documentation of architecture, configurations, processes, and runbooks

·      Participate in on-call rotations and incident response as needed

Required Skills & Qualifications

·      Bachelor's degree in Computer Science, IT, or a related field (or equivalent practical experience)

·      4+ years of hands-on experience working with AWS cloud services in a production environment

·      Strong expertise in core AWS services: EC2, S3, VPC, IAM, RDS, Lambda, CloudWatch, ECS/EKS, Route 53, ELB/ALB, Auto Scaling

·      Solid understanding of networking concepts (subnets, routing, security groups, load balancers, VPNs, VPC peering, Direct Connect)

·      Strong scripting/programming skills in Python, Bash, or PowerShell for automation

·      Hands-on experience with Infrastructure as Code tools such as Terraform or AWS CloudFormation

·      Proven experience with Linux and/or Windows server administration

·      Strong understanding of CI/CD pipelines, Git-based version control, and branching strategies

·      Solid experience with containerization and orchestration (Docker, ECS, or EKS/Kubernetes)

·      Experience with configuration management tools (Ansible, Chef, or Puppet) is a plus

·      Application Server Management — strong knowledge of networking, firewalls, load balancers, Nginx, Apache, etc.

·      Ability to independently read, interpret AWS documentation, and troubleshoot complex issues

·      Experience with cost optimization, security audits, and compliance frameworks (e.g., ISO, SOC2) is a plus

·      AWS certification (Solutions Architect Associate/Professional, DevOps Engineer Professional) preferred

Good to Have

·      Working knowledge of Microsoft Azure services (Virtual Machines, VNets, Azure DevOps, Azure Storage, Azure AD/Entra ID, AKS)

·      Experience with multi-cloud or hybrid-cloud environments

·      Familiarity with Azure Resource Manager (ARM) templates or Bicep

·      Any Azure certification (AZ-104, AZ-400, etc.)

Soft Skills

·      Strong analytical, problem-solving, and decision-making ability

·      Excellent verbal and written communication skills

·      Ability to mentor and guide junior team members

·      Proactive, ownership-driven approach to infrastructure and incident management

·      Strong collaboration skills in a cross-functional, team-oriented environment

·      High attention to detail with a focus on reliability and scalability

Read more
Why apply to jobs via Cutshort
people_solving_puzzle
Personalized job matches
Stop wasting time. Get matched with jobs that meet your skills, aspirations and preferences.
people_verifying_people
Verified hiring teams
See actual hiring teams, find common social connections or connect with them directly.
ai_chip
Move faster with AI
We use AI to get you faster responses, recommendations and unmatched user experience.
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos