Cutshort logo
For Employers
A modern configuration management platform based on advanced logo
DevOps Engineer
A modern configuration management platform based on advanced

DevOps Engineer at A modern configuration management platform based on advanced · Bengaluru (Bangalore) · 3 - 5 years · ₹25L - ₹100L / yr (ESOP available) · Posted 8 Sep 2025

Scaling Theory's logo

DevOps Engineer

at A modern configuration management platform based on advanced

Agency job
3 - 5 yrs
₹25L - ₹100L / yr (ESOP available)
Bengaluru (Bangalore)
Skills
DevOps
skill iconKubernetes
Terraform
Ansible
skill iconDocker
skill iconAmazon Web Services (AWS)

Key Responsibilities:

Kubernetes Management:

Deploy, configure, and maintain Kubernetes clusters on AKS, EKS, GKE, and OKE.

Troubleshoot and resolve issues related to cluster performance and availability.

Database Migration:

 Plan and execute database migration strategies across multicloud environments, ensuring data integrity and minimal downtime.

Collaborate with database teams to optimize data flow and management.

Coding and Development:

 Develop, test, and optimize code with a focus on enhancing algorithms and data structures for system performance.

Implement best coding practices and contribute to code reviews.

Cross-Platform Integration:

  Facilitate seamless integration of services across different cloud providers to enhance interoperability.

Collaborate with development teams to ensure consistent application performance across environments.

Performance Optimization:

  Monitor system performance metrics, identify bottlenecks, and implement effective solutions to optimize resource utilization.

Conduct regular performance assessments and provide recommendations for improvements.

Experience:

  Minimum of 2+ years of experience in cloud computing, with a strong focus on Kubernetes management across multiple platforms.

Technical Skills:

  Proficient in cloud services and infrastructure, including networking and security considerations.

Strong programming skills in languages such as Python, Go, or Java, with a solid understanding of algorithms and data structures.

Problem-Solving:

 Excellent analytical and troubleshooting skills with a proactive approach to identifying and resolving issues.

Communication:

 Strong verbal and written communication skills, with the ability to collaborate effectively with cross-functional teams.

Preferred Skills:

- Familiarity with CI/CD tools and practices.

- Experience with container orchestration and management tools.

- Knowledge of microservices architecture and design patterns.


Read more
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos

Similar jobs (10)

company logo
Ashish Singh
Posted by Ashish Singh
Remote only
0 - 1 yrs
₹12000 - ₹18000 / mo
CI/CD
DevOps

Build production-grade cloud infrastructure that powers enterprise applications with cutting-edge DevOps practices.


What you'll do:

  • Design CI/CD pipelines (GitHub Actions, Jenkins, GitLab CI)
  • Containerize apps with Docker, deploy on Kubernetes clusters
  • Manage infrastructure as code (Terraform, CloudFormation)
  • Set up monitoring (Prometheus, Grafana, ELK Stack)
  • Cloud migrations (AWS EC2, EKS, RDS → GCP equivalent)
  • Optimize costs and performance for live production systems

What we need:

  • Basic Python/Bash scripting
  • Docker basics, Git workflows
  • Cloud exposure (AWS/GCP/Azure free tier projects)
  • Problem-solving mindset, eagerness to learn

Real impact:

  • Deploy apps used by 1000+ daily users
  • Work with senior DevOps engineers on client deliverables
  • Build portfolio for FAANG-level interviews


Read more
company logo
Ganesh Ram
Posted by Ganesh Ram
Bengaluru (Bangalore), Mumbai, Delhi, Gurugram, Noida, Ghaziabad, Faridabad, Hyderabad, Pune
7 - 10 yrs
₹15L - ₹20L / yr
CI/CD
skill iconKubernetes
helm
Terraform
yaml

Cloud Expertise(Azure):

• Strong understanding of cloud services and resources like AI services, webapp, database, including monitoring tools like Azure Monitor and Log Analytics.

• Experience with Infrastructure as Code (IaC) tools such as Arm template / Bicep/Terraform.

• Deep understanding of Networking concepts(DNS, DHCP , Hub and Spoke).

• Understanding on policies and security aspects of cloud.


Kubernetes & Helm:

• In-depth knowledge of Kubernetes concepts such as pods, services, ingress, config maps, and secrets.

• Understand of Kubernetes templates and its deployment.

• Proficiency with Helm/ Kustomize or equivalent for Kubernetes package management and deployment automation.

• Implement Kubernetes best practices, including security, networking, and scaling.

• Concepts of Docker and Containers


CI/CD & Programming:

• Hands-on experience with YAML-based CI/CD pipelines (e.g., Azure DevOps, GitHub Actions).

• Familiarity with scripting and automation tools such as PowerShell, Azure CLI, or Bash.

• Proven skill in python programming and concepts.


Monitoring and Observability:

Expertise in creating and managing Grafana dashboards for visualizing metrics and logs.

• Knowledge of Log Analytics & Azure Application Insights for performance monitoring and tracing.

 

Read more
MNC
MNC
Agency job
via by aafia parveen
Bengaluru (Bangalore)
7 - 11 yrs
₹15L - ₹18L / yr
skill iconAmazon Web Services (AWS)
skill iconKubernetes
Linux/Unix
openshift

AWS / Kubernetes / OpenShift / Linux –

Location: Bangalore

Experience: 7–10 Years

Mandatory Skills:

  • Strong hands-on experience with AWS
  • Experience in Kubernetes & OpenShift
  • Strong knowledge of Linux administration
  • Experience with Docker & containerization
  • Knowledge of CI/CD pipelines and DevOps practices
  • Troubleshooting, monitoring, and deployment experience

Role: Cloud/DevOps Engineer – AWS, Kubernetes & OpenShift

Read more
company logo
Mohammed Rabidheen
Posted by Mohammed Rabidheen
Coimbatore
3 - 8 yrs
Best in industry
Windows Azure
AKS
DevOps
Microsoft Windows Azure

Senior Cloud Site Reliability Engineer (CSRE) – Azure


About Searce:

Searce is an AI-native, engineering-led modern technology consultancy that empowers

clients to futurify their businesses by delivering real, intelligent business outcomes. As a

trusted partner for over 3,000 clients globally, Searce specializes in cloud modernization,

data engineering, applied AI, and robust cloud platform security. Driven by a "HAPPIER"

cultural mindset and our proprietary evlos problem-solving framework, we eliminate

bureaucratic fluff to build working prototypes fast and scale enterprise production

environments intelligently. We don't just fix systems; we leverage multi-cloud technologies

to transform client operations into distinct competitive advantages.

Position Overview:

We are looking for a high-caliber Senior or Lead Cloud Site Reliability Engineer (CSRE) to

architect, secure, and stabilize next-generation hybrid and multi-cloud environments.

Operating at the intersection of infrastructure design, security compliance, and production

operations, you will serve as the technical Subject Matter Expert (SME) across GCP, Azure,

and AWS.

Whether optimizing a microservice mesh on GKE, tuning autoscaling on AKS, or driving a

massive disaster recovery drill across AWS regions, your focus will be absolute reliability. For

the Lead path, you will couple this deep engineering toolkit with stakeholder management

and mentorship to drive an elite operational culture.


Experience & Level Expectation:

Years of Experience: 3 to 10 years of intensive, hands-on production operations

experience in a dedicated DevOps, Cloud Platform Engineering, or SRE role.

Associate level (3-5 Years): Expected to show flawless execution of IaC, advanced

triaging of infrastructure failures, and ownership of the CI/CD and deployment

lifecycles.

Intermediate level (5-10 Years): Expected to take architectural ownership, serve as

primary Incident Commander for complex outages, design cross-cloud governance

frameworks, and act as a reliable bridge between technical teams and client

leadership.


Key Responsibilities & Role Expectations:

Multi-Cloud Platforms & Orchestration: Design, configure, and maintain

production-grade Kubernetes clusters across major platforms (AKS).

Manage advanced network routing, service meshes (e.g., Istio), and multi-tenant

isolation.

Infrastructure as Code (IaC) & GitOps: Build declarative, enterprise-grade, reusable

infrastructure components using Terraform or Crossplane. Standardize automated

environment provisioning to eliminate configuration drift across multi-branch

environments.

Incident Management & Reliability (SRE): Own and optimize the production on-call

rotation. Lead rapid mitigation strategies for Sev-1/Sev-2 system outages, reducing

Mean Time to Recovery (MTTR) through centralized log and metric correlation.

Root Cause Analysis (RCA): Facilitate rigorous, blameless post-incident reviews to

identify core architectural vulnerabilities and establish long-term fixes preventing

recurrence.

Lifecycle, Patching & Upgrades: Plan and execute zero-downtime cluster upgrades,

operating system patching strategies (Linux/Windows), database lifecycle updates,

and multi-region Disaster Recovery (DR) failover drills.

Core Core Operations & Legacy Integration: Manage enterprise-level hybrid

networking architecture (VPCs, Firewalls, Load Balancers, DNS routing, and DHCP

configurations) while effectively connecting cloud native services to legacy

infrastructures like Active Directory.

Security & Governance: Embed Zero Trust policies, secure secrets management

(Secrets Manager/Key Vault), and continuous vulnerability patching into the

automated SDLC pipeline.


Required Technical Skills:

- Microsoft Azure: Azure Virtual Machines, Virtual Networks, Azure Active Directory, Azure Update Management.

- Containers & Orchestration

  • Production-level management of GKE, AKS, and EKS.
  • Advanced mastery of Docker, Helm, Kubernetes StatefulSets, Pod Disruption
Read more
company logo
Megha Shetty
Posted by Megha Shetty
Bengaluru (Bangalore)
6 - 8 yrs
₹10L - ₹20L / yr
DevOps
skill iconAmazon Web Services (AWS)
skill iconPython
Terraform
skill iconDocker

Job Description:

Pre-requisite skills required for a DevOps Engineer role include:


  • 6+yrs exp in DevOps
  • Experience working on Linux based infrastructure
  • knowledge in AWS, docker, CI/CD tools
  • Hands on exp in Python/shell scripting language
  • hands on exp in AWS and Azure
  • Work exp in Docker, Terraform, Ansible, Kubernetes, LINUX
  • Excellent understanding of Ruby, Python, Perl, and Java
  • Configuration and managing databases such as MySQL, Mongo
  • Excellent troubleshooting
  • Working knowledge of various tools, open-source technologies, and cloud services
Read more
company logo
Sakshi Mittal
Posted by Sakshi Mittal
Bengaluru (Bangalore)
3 - 5 yrs
₹6L - ₹12L / yr
skill iconAmazon Web Services (AWS)
DevOps
skill iconKubernetes
Terraform
CI/CD
+1 more

Job Summary :

We are looking for a proactive and skilled DevOps Engineer to join our team and play a key role in building, managing, and scaling infrastructure for high-performance systems. The ideal candidate will have hands-on experience with Kubernetes, Docker, Python scripting, cloud platforms, and DevOps practices around CI/CD, monitoring, and incident response.

Key Responsibilities :

- Design, build, and maintain scalable, reliable, and secure infrastructure on cloud platforms such as AWS.

- Implement Infrastructure as Code (IaC) using tools like Terraform, Cloud Formation, or similar.

- Manage Kubernetes clusters, configure namespaces, services, deployments, and auto scaling. CI/CD & Release Management

- Build and optimize CI/CD pipelines for automated testing, building, and deployment of services.

- Collaborate with developers to ensure smooth and frequent deployments to production.

- Manage versioning and rollback strategies for critical deployments.

- Containerization & Orchestration using Kubernetes.

- Containerize applications using Docker, and manage them using Kubernetes.

- Write automation scripts using Python or Shell for infrastructure tasks, monitoring, and deployment flows.

- Develop utilities and tools to enhance operational efficiency and reliability.

- Monitoring & Incident Management

- Analyze system performance and implement infrastructure scaling strategies based on load and usage trends.

- Optimize application and system performance through proactive monitoring and configuration tuning.

Desired Skills and Experience :

- Experience Required - 6+ yrs.

- Hands-on experience on cloud services like AWS, EKS etc.

- Ability to design a good cloud solution.

- Strong Linux troubleshooting, Shell Scripting, Kubernetes, Docker, Ansible, Jenkins Skills.

- Design and implement the CI/CD pipeline following the best industry practices using open-source tools.

- Use knowledge and research to constantly modernize our applications and infrastructure stacks.

- Be a team player and strong problem-solver to work with a diverse team.

- Having good communication skills.

Read more
Service Co
Pune
6 - 11 yrs
₹15L - ₹26L / yr
AWS Iaas
Platform as a Service (PaaS)
AWS VPS
Amazon EKS

Key Skills:


• Bachelor's or Master's degree in Computer Science or related field.


• Minimum 5 years of experience in Platform Engineering, DevOps, or Cloud Infrastructure Engineering.


• Experience migrating data and systems between AWS IaaS and PaaS.


• Experience operating and supporting applications using AWS VPC, EKS, and related services for multi-account operations.


• Experience developing fast and reliable Continuous Integration/Continuous Deployment (CI/CD) workflows used by hundreds of application teams.


• Experience administering and troubleshooting Operating Systems such as Linux, Windows, and MacOS.


• Professional Certifications in AWS Networks, CNCF Technologies, or Kubernetes.


• Experience using and configuring observability tools such as ELK, Prometheus/Grafana, AWS CloudWatch, and Jaeger.


• Experience of applied GitOps principles using ArgoCD or Flux.


• Public examples of code you've worked on with other people using any of these technologies:


o Configuration management/Infrastructure as Code (IAC) tools, such as AWS CDK, AWS CloudFormation, Terraform, Ansible, or Puppet.


o Systems solutions in one or more programming languages, such as Golang, Python, Java.


o Build, Release, Deploy or Ops Workflows using Bamboo, Argo Project, or GitHub Actions.

Read more
company logo
Taher Ujjainwala
Posted by Taher Ujjainwala
Pune
12 - 25 yrs
₹70L - ₹120L / yr (ESOP available)
skill iconPython
skill iconKubernetes
Google Cloud Platform (GCP)
skill iconAmazon Web Services (AWS)
Windows Azure
+15 more

About the Role

We are hiring Staff / Principal Engineers to take full, hands-on ownership of Blitzy's most critical production-grade systems and to deliver high-leverage features that materially improve customer outcomes and engineering velocity. This is the most senior individual contributor role at the company today.

This is not a Senior-plus role, an architecture-only role, or a promotion-track role. We are looking for someone who has already operated at Principal / Staff+ scope in a highly technical environment and expects to spend their time writing, reviewing, and shipping production code.

This role is 100% hands-on. Leverage comes from system ownership, execution quality, and durable technical decisions — not people management or process.


Responsibilities

  • Own mission-critical production systems end-to-end, ensuring correctness, scalability, performance, reliability, and operational excellence.
  • Design, build, and ship high-impact backend systems and features that improve product reliability, performance, and customer value.
  • Architect scalable services and cloud infrastructure using technologies such as Python, REST, gRPC, Kubernetes, and Terraform.
  • Identify and resolve complex technical bottlenecks that limit engineering quality, system performance, or organizational velocity.
  • Build and operate LLM-powered systems and validation loops that evaluate correctness, consistency, durability, and production performance.
  • Design and evolve data architectures incorporating relational, NoSQL, graph, and vector databases to support complex enterprise applications and semantic retrieval.
  • Modernize and improve complex enterprise systems while balancing reliability, maintainability, scalability, and delivery speed.
  • Set and uphold engineering quality standards through hands-on technical leadership, sound technical judgment, and ownership of long-term technical decisions.


Qualifications

  • Direct experience with Python as a primary programming language, backend frameworks, and microservices architectures.
  • Expertise in REST and gRPC, with proficiency in Node.js and JavaScript.
  • Proficiency in GCP, along with experience using at least one additional cloud platform such as AWS or Azure.
  • Advanced knowledge of Kubernetes and Terraform in production environments.
  • Experience operating highly available production systems, including monitoring, scalability, reliability, performance optimization, and operational tooling.
  • Strong knowledge of SQL and NoSQL databases, including PostgreSQL, MySQL, MongoDB, Cassandra, or DynamoDB.
  • Familiarity with graph databases such as Neo4j and vector databases or embedding infrastructure for semantic search and retrieval.
  • Hands-on experience building and operating LLM-powered systems in production, including evaluation, validation, regression testing, tracing, and failure analysis.
  • Working knowledge of LangSmith or comparable LLM observability and evaluation tools; familiarity with OpenAI, Anthropic, or similar model providers is a plus.
  • Ability to contribute across the full stack, with a strong understanding of frontend architecture and the ability to debug, design, and ship across frontend, backend, infrastructure, and AI systems.
  • Understanding of large-scale enterprise software systems, including architecture, integration, deployment, modernization, and long-term maintainability.
  • Proven track record of operating at Staff+, Principal Engineer, or equivalent level, independently driving complex technical initiatives and delivering high-impact outcomes with minimal supervision.


Blitzy is a Cambridge, MA based AI software development platform on a mission to revolutionize the software development life cycle by autonomously building custom software to unlock the next industrial revolution. We're transforming how enterprises build software, turning enterprise requirements into enterprise grade code with an agentic software development platform that can autonomously execute 80% of the quantum of software development work. We're backed by multiple tier 1 investors, and have proven success as founders of previous start-ups.


Our Culture

Who we are:

Led by two pioneering co-founders we are one of the fastest growing companies in the U.S., creating our own category of enterprise autonomous software development. We automate thousands of hours of software development for our customers, which includes strong representation within the Fortune 500.


How we work:

  • We move Blitzy Fast: Time is both our company’s and our clients’ most precious asset. We move quickly and decisively to innovate internally and deliver exceptional software externally.
  • Championship Mindset: We operate like a professional sports team. We win as a team by holding ourselves and each other to high standards, collaborating in-person, and remaining focused on the mission.
  • Passion for Invention: We’re pushing the frontier of what’s possible, requiring constant innovation and iteration.
  • We Work for the Customer: We focus on delivering outsized value to the customers we work with and expanding those relationships into deep, meaningful partnerships.
  • We believe in being ‘everyday athletes’: taking care of ourselves so we can bring our best minds to work. We promote great sleep, movement, and restorative activities for 


Blitzy is an equal opportunity employer committed to building a diverse and inclusive team. We believe different perspectives make us stronger.

Read more
company logo
Swathi S
Posted by Swathi S
Chennai
7 - 12 yrs
₹30L - ₹55L / yr
skill iconAmazon Web Services (AWS)
skill iconPython
CI/CD
DevOps
Platform as a Service (PaaS)
+7 more

Amura’s Vision 


We believe that the most under-appreciated route to releasing untapped human potential is to build a healthier body, and through which a better brain. This allows us to do more of everything that is important to each one of us.


Billions of healthier brains, sitting in healthier bodies, can take up more complex problems that defy solutions today, including many existential threats, and solve them in just a few decades.


Billions of healthier brains will make the world richer beyond what we can imagine today. The surplus wealth, combined with better human capabilities, will lead us to a new renaissance, giving us a richer and more beautiful culture.


These healthier brains will be equipped with deeper intellect, be less acrimonious, more magnanimous, and have a kinder outlook on the world, resulting in a world that is better than any previous time.

We find this vision of the future exhilarating. Our hopes and dreams are to create this future as quickly as possible and ensure that it is widely distributed and optimized to maximize all forms of human excellence. 


Role Overview 


We are looking for a highly skilled Senior DevOps Engineer (AI-Native Infrastructure & Platform Engineering) with deep expertise in AWS cloud infrastructure, automation, AI infrastructure operations, and modern DevOps/SRE practices.


This role goes beyond traditional DevOps and requires a seasoned specialist capable of building and operating AI-ready infrastructure platforms that support high-throughput APIs, LLM/AI workloads, GPU-based compute, data-intensive systems, real-time inference pipelines, and scalable ML platforms.


You will be responsible for architecting, automating, securing, and optimizing highly scalable and cost-efficient cloud environments that enable high-velocity engineering and AI teams. This is an ideal position for someone who combines technical ownership, an automation-first mindset, and a passion for developer productivity and platform reliability. 


Key Responsibilities 


Cloud Infrastructure & Platform Engineering (AWS) 

  • Architect, deploy, and manage highly scalable and secure infrastructure on AWS. Design cloud platforms supporting AI/ML workloads, data pipelines, real-time APIs, and high-concurrency backend systems.
  • Hands-on expertise with key AWS services including EC2, ECS/EKS, Lambda, RDS, DynamoDB, S3, VPC, CloudFront, IAM, CloudWatch, and GPU-enabled instances.
  • Build and maintain Infrastructure-as-Code (IaC) using Terraform, CloudFormation, or AWS CDK.
  • Design multi-AZ and multi-region architectures for high availability and disaster recovery (HA/DR).
  • Build reusable platform templates and shared infrastructure modules. 


AI/ML Infrastructure & MLOps 

  • Build and maintain infrastructure for LLM applications, AI inference workloads, model serving platforms, vector databases, and feature stores.
  • Support GPU-based workloads and optimize compute/storage usage.
  • Enable scalable deployment patterns for AI applications using Kubernetes/EKS. Collaborate with Data Science and ML Engineering teams on model deployment, training/tuning of models, CI/CD for ML systems, experiment environments, and reproducibility.
  • Support orchestration and deployment of AI workflows and inference services while implementing observability and reliability for AI pipelines. 


CI/CD, Automation & Developer Productivity 

  • Build and maintain CI/CD pipelines using GitHub Actions, GitLab CI, Jenkins, or AWS CodePipeline.
  • Automate deployments, environment provisioning, and release workflows.
  • Build self-service developer platforms, preview environments, and reusable deployment workflows to improve developer productivity.
  • Implement automated patching, scaling, backups, cleanup workflows, and drift detection. 


Containers, Kubernetes & Platform Reliability

  • Manage Docker-based environments, containerized applications, and optimize workloads using Kubernetes (EKS) or ECS/Fargate.
  • Manage autoscaling, cluster health, node pools, ingress, service mesh, and workload isolation.
  • Optimize infrastructure for performance, resilience, and cost-efficiency.
  • Implement progressive deployment strategies including blue/green, canary, and rolling deployments. 


Observability, Incident Response & SRE Practices

  • Implement observability stacks using CloudWatch, Prometheus, Grafana, ELK, Datadog, OpenTelemetry, or New Relic.
  • Build actionable dashboards and intelligent alerting systems while defining and tracking SLIs, SLOs, and SLAs.
  • Lead incident response, root cause analysis, and blameless postmortems to reduce operational toil and improve MTTR.

FinOps, Cost Governance & Security

  • Continuously monitor and optimize cloud costs (compute utilization, storage lifecycle, GPU usage, and data transfer) using AWS Cost Explorer, Budgets, Trusted Advisor, CloudHealth, or Kubecost.
  • Implement AWS security best practices for IAM, VPCs, security groups, NACLs, encryption, and manage secrets using KMS, SSM Parameter Store, or Vault.
  • Build secure CI/CD pipelines with automated security checks, least-privilege access, audit logging, and ensure compliance readiness for ISO 27001, SOC2, and GDPR.

Collaboration, Leadership & Platform Culture

  • Work closely with engineering, AI/ML, QA, product, and operations teams to drive a DevOps, SRE, GitOps, and automation-first culture.
  • Mentor junior DevOps and Platform Engineers while creating and maintaining detailed runbooks, architecture diagrams, and platform documentation.

Skills & Qualifications


Must-Have:

  • 7+ years of experience in DevOps, SRE, Platform Engineering, or Cloud Infrastructure Engineering.
  • Strong expertise in AWS cloud architecture, services, and deep understanding of Kubernetes (EKS), containers, and cloud-native systems.
  • Strong Infrastructure-as-Code expertise using Terraform, CloudFormation, or CDK. Strong Linux administration, networking, DNS, routing, and load balancing knowledge. Strong scripting/programming experience in Python, Bash, or Go (preferred). Experience with CI/CD automation, GitOps workflows, and observability platforms supporting scalable production systems.


Preferred / Nice-to-Have:

  • Experience with AI/ML infrastructure, MLOps, model serving, vector databases, GPU orchestration, and inference optimization.
  • Familiarity with Kafka, Redis, SQS, and event-driven systems.
  • Exposure to platform engineering, internal developer platforms, and tools like ArgoCD, Flux, Helm, and OpenTelemetry.
  • AWS Certifications: Solutions Architect, DevOps Engineer, or SysOps Administrator. Knowledge of distributed systems and large-scale platform operations. 


Preferred / Nice-to-Have:

  • Experience with AI/ML infrastructure, MLOps, model serving, vector databases, GPU orchestration, and inference optimization.
  • Familiarity with Kafka, Redis, SQS, and event-driven systems.
  • Exposure to platform engineering, internal developer platforms, and tools like ArgoCD, Flux, Helm, and OpenTelemetry.
  • AWS Certifications: Solutions Architect, DevOps Engineer, or SysOps Administrator. Knowledge of distributed systems and large-scale platform operations. 


Here are answers to some questions you may have

Where is your office?

Chennai (Velachery)

Work Model

Work from Office – because great stories are built in person!

Do you have an online presence?

https://amura.ai (we are @AmuraHealth on all social media)


Read more
company logo
Daniel Castellanos
Posted by Daniel Castellanos
Remote only
1 - 10 yrs
₹10L - ₹50L / yr
skill iconKubernetes
ArgoCD
Exim
Postfix

We are looking for an experienced DevOps Engineer to take ownership of production infrastructure, cloud environments, Kubernetes platforms, and infrastructure automation. This is a hands-on role for someone who enjoys solving complex infrastructure challenges and is comfortable being responsible for systems in production.


Key Responsibilities

  • Own and operate production infrastructure, including participating in an on-call rotation and responding to production incidents.
  • Design, operate, and continuously improve Kubernetes clusters in production.
  • Manage and automate infrastructure using Infrastructure as Code, primarily with Terraform.
  • Build, maintain, and optimise cloud infrastructure across AWS, GCP, or Azure.
  • Work extensively with Linux, including system administration, networking, troubleshooting, and system-level configuration.
  • Manage production deployment and GitOps workflows using ArgoCD.
  • Improve infrastructure reliability, scalability, security, monitoring, and operational efficiency.
  • Troubleshoot complex production issues and drive problems through to resolution.
  • Develop automation and processes that reduce manual operational work.

Essential Requirements

  • 4+ years of hands-on experience operating production infrastructure, with personal ownership and responsibility for live systems, including on-call experience.
  • Deep, hands-on Kubernetes experience — you must have operated and managed Kubernetes clusters, rather than simply deploying applications onto clusters managed by another team.
  • Strong experience with Infrastructure as Code, with Terraform strongly preferred. Experience with Pulumi or CloudFormation is also considered.
  • Strong experience with at least one major cloud platform, ideally AWS. Strong GCP or Azure experience is also welcome, provided you are willing to work with AWS.
  • Strong Linux skills and confidence working from the command line, including networking, troubleshooting, system configuration, and performance issues.
  • Production experience with ArgoCD and GitOps-based deployment workflows.
  • Strong troubleshooting and problem-solving skills, with the ability to take ownership of production incidents and infrastructure issues.


Nice to Have

Experience with email infrastructure would be a strong advantage, particularly:

  • Exim
  • IMAP / SMTP
  • Postfix
  • Dovecot
  • General mail server administration and maintenance


Read more
Why apply to jobs via Cutshort
people_solving_puzzle
Personalized job matches
Stop wasting time. Get matched with jobs that meet your skills, aspirations and preferences.
people_verifying_people
Verified hiring teams
See actual hiring teams, find common social connections or connect with them directly.
ai_chip
Move faster with AI
We use AI to get you faster responses, recommendations and unmatched user experience.
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos