Senior AWS Devops Engineer at ketteq · Remote only · 5 - 15 years · ₹20L - ₹35L / yr · Bootstrapped · Remote only · Posted 1 Dec 2023

ketteQ is a supply chain planning and automation platform. We are looking for an experienced AWS Devops Engineer to help manage AWS infrastructure and automation. This job comes with a attractive compensation package, work-from-home and flex-time benefits. You will get to work on projects for large global brands with a highly experienced team based in US and India. If you are high-energy, motivated, and initiative-taking individual then this could be a fantastic opportunity for you. Candidates must meet the following requirements:
Duties & Responsibilities
- Deployment, automation, management, and maintenance of AWS cloud-based production system
- Build a deployment pipeline for AWS and Salesforce
- Design cloud infrastructure that is secure, scalable, and highly available on AWS
- Work collaboratively with software engineering to define infrastructure and deployment requirements
- Provision, configure and maintain AWS cloud infrastructure defined as cloud formation template
- Ensure configuration and compliance with configuration management tools
- Administer and troubleshoot Linux based systems
- Troubleshoot problems across a wide array of services and functional areas
- Build and maintain operational tools for deployment, monitoring, and analysis of AWS infrastructure and systems
- Perform infrastructure cost analysis and optimization
Requirements
- At least 5 years of experience building and maintaining AWS infrastructure (VPC, EC2, Security Groups, IAM, ECS, Fargate, S3, Cloud Formation)
- Strong understanding of how to secure AWS environments and meet compliance requirements
- Solid foundation of networking and Linux administration
- Experience with Docker, GitHub, Jenkins, Cloud Formation and deploying applications on AWS
- Ability to learn/use a wide variety of open source technologies and tools
- Database experience to help with monitoring and performance; PostgreSql experience preferred
- AWS certification preferred
Education
- Bachelors in Engineering or related field

About ketteq
About
Connect with the team
Similar jobs (10)
Job Title : DevOps Engineer / Site Reliability Engineer (SRE)
Experience : 4+ Years
Location : Gurugram, Sector 48, Haryana (On-site)
Employment Type : Full-Time
Working Days : Monday to Saturday (1st & 3rd Saturday Off)
About the Role :
We are looking for a hands-on DevOps Engineer / Site Reliability Engineer (SRE) with strong experience in Linux, AWS, Kubernetes, Docker, CI/CD, Infrastructure as Code, and production application deployments.
The ideal candidate should have real-world production experience, excellent troubleshooting skills, and the ability to manage both infrastructure and application-level issues.
Mandatory Skills :
Linux, AWS, Docker, Kubernetes, Terraform, Ansible, Jenkins, GitHub Actions, GitLab CI/CD, CI/CD, Infrastructure as Code (IaC), Python, Bash, Git, Grafana, Prometheus, ELK, CloudWatch, New Relic, SRE (SLI/SLO/SLA), Networking (DNS, HTTP/HTTPS, TCP/IP, Load Balancer), Production Application Deployment & Troubleshooting
Key Responsibilities :
- Manage and maintain AWS cloud infrastructure.
- Build and optimize CI/CD pipelines using Jenkins, GitHub Actions, or GitLab CI.
- Deploy, monitor, and troubleshoot applications across production environments.
- Automate infrastructure using Terraform and Ansible.
- Manage Docker containers and Kubernetes clusters.
- Monitor systems using Grafana, Prometheus, ELK, CloudWatch, and New Relic.
- Perform Linux server administration and troubleshooting.
- Handle production incidents, Root Cause Analysis (RCA), and improve system reliability.
- Collaborate with development teams to support application releases and automation.
Required Qualifications :
- Bachelor's degree in Computer Science or related field.
- 4+ years of hands-on experience in DevOps / SRE.
- Strong Linux administration and production troubleshooting skills.
- Experience with AWS and modern DevOps toolchains.
- Hands-on experience with application deployment and production support.
What We're Looking For :
- Strong practical Linux and cloud knowledge.
- Real production experience with application deployments.
- Ability to troubleshoot both infrastructure and application issues.
- Experience handling live production incidents.
- Excellent communication and problem-solving skills.
- Candidates should be comfortable with scenario-based technical discussions and demonstrate genuine hands-on expertise.
Interview Process :
- HR Screening
- Technical Round
- Client Technical Round
- Final Discussion
Note : The interview will focus on practical hands-on experience in Linux, AWS, Kubernetes, Docker, CI/CD, Infrastructure as Code, application deployment, production troubleshooting, and real-world DevOps scenarios.
Job Title : DevOps Engineer – Linux, AWS & Infrastructure
Experience : 3+ Years
Location : Sector 48, Gurgaon
Work Mode : 6 Days WFO – Monday to Saturday
Week Off : 01st & 03rd Saturday Off
Employment Type : Full-Time
Job Summary :
We are looking for a DevOps Engineer with strong hands-on experience in Linux, Networking, Server Administration, AWS, CI/CD, Containers, Kubernetes, and Infrastructure Automation. The ideal candidate should have strong troubleshooting skills, production ownership, and the ability to manage and automate infrastructure reliably.
Key Responsibilities :
- Manage and troubleshoot Linux servers, bare-metal infrastructure, and server environments.
- Perform system administration, networking, performance monitoring, and production troubleshooting.
- Design, maintain, and optimize Jenkins-based CI/CD pipelines and deployment workflows.
- Manage Docker containers and Kubernetes environments.
- Work with AWS Cloud services, infrastructure, security, and deployment environments.
- Monitor system health, application performance, logs, and infrastructure using appropriate monitoring and logging tools.
- Implement and maintain Infrastructure as Code (IaC) using tools such as Terraform or CloudFormation.
- Automate repetitive operational tasks using Python, Bash, Shell scripting, or similar technologies.
- Implement infrastructure and application security, access controls, patching, and hardening.
- Investigate production incidents, perform root-cause analysis (RCA), and drive issues to resolution.
- Take end-to-end ownership of infrastructure reliability, availability, and operational issues.
- Collaborate with development and other engineering teams to improve deployment, scalability, and system reliability.
Mandatory Skills :
Linux & Networking | Bare Metal / Server Administration | Jenkins / CI-CD | Docker / Containers | AWS Cloud | Kubernetes | Git | Monitoring & Logging | Security | Infrastructure as Code (IaC) | Automation / Scripting | Production Troubleshooting & Ownership
Preferred Skills :
- Strong understanding of TCP/IP, DNS, HTTP/HTTPS, SSH, load balancing, and networking fundamentals.
- Hands-on experience with Terraform / CloudFormation.
- Experience with Prometheus, Grafana, ELK / EFK, CloudWatch, or similar monitoring / logging tools.
- Good understanding of Linux performance troubleshooting, processes, memory, disk, networking, and file systems.
- Experience handling production incidents, RCA, deployments, and system reliability.
- Exposure to cloud security, IAM, secrets management, and server hardening.
What We’re Looking For :
- 3+ years of hands-on experience in DevOps / SRE / Infrastructure Engineering.
- Strong practical knowledge rather than certification-based/theoretical understanding.
- Good troubleshooting and analytical skills.
- Strong sense of ownership and accountability for production systems.
- Comfortable working in a 6-day work-from-office environment.
Job Title : DevOps / Infrastructure Engineer
Experience : 3+ Years
Location : Gurugram Sector 48
Work Mode : 6 Days WFO (Monday to Saturday) / 01st & 03rd Saturdays are off
Employment Type : Full-time
Role Overview :
We are looking for a DevOps / Infrastructure Engineer with strong hands-on experience in Linux administration, Jenkins, Docker, networking, bare-metal servers, AWS, Redis, and MongoDB. The candidate should be capable of independently troubleshooting infrastructure, deployment, networking, and application-related issues in production environments.
Mandatory / Non-Negotiable Skills :
- Strong hands-on experience with Linux Administration & Troubleshooting.Strong experience with Jenkins and CI/CD pipelines.
- Hands-on experience with Docker and containerized environments.
- Strong understanding of Networking concepts – TCP/IP, DNS, HTTP/HTTPS, ports, routing, firewalls, load balancing, etc.
- Hands-on experience with Bare Metal Servers / Server Administration.
- Strong hands-on experience with AWS (EC2, VPC, IAM, Security Groups, Load Balancers, S3 & CloudWatch)
- Working knowledge of Redis
- Working knowledge of MongoDB
- Strong production troubleshooting and incident-resolution skills
Key Responsibilities :
- Manage, configure, monitor, and troubleshoot Linux and bare-metal servers
- Build, maintain, and troubleshoot Jenkins CI/CD pipelines
- Deploy, manage, and troubleshoot applications using Docker
- Manage and troubleshoot AWS infrastructure and services
- Configure and maintain networking, security groups, firewalls, ports, and connectivity
- Support and maintain Redis and MongoDB environments
- Perform server health checks, log analysis, performance troubleshooting, and incident resolution
- Work closely with development teams to support application deployments
- Identify root causes of infrastructure and production issues and implement preventive solutions
- Maintain infrastructure security, availability, and reliability
- Automate repetitive operational tasks wherever possible
Required Candidate Profile :
- 3+ years of relevant experience in DevOps, Infrastructure, System Administration, or related roles.
- Strong hands-on / production experience with all mandatory technologies.
- Good understanding of Linux systems and infrastructure.
- Strong troubleshooting and problem-solving abilities.
- Ability to take ownership of production infrastructure and deployment issues.
- Good communication and collaboration skills.
Job Title : DevOps Engineer / Site Reliability Engineer (SRE)
Experience : 5+ Years
Location : Gurugram, Haryana
Work Mode : On-site (Full-time)
About the Role :
We are looking for a skilled DevOps Engineer with 5+ years of experience in cloud infrastructure, CI/CD, automation, Kubernetes, and Site Reliability Engineering (SRE). The ideal candidate will be responsible for building scalable cloud infrastructure, automating deployments, improving system reliability, and ensuring high availability across production environments.
Mandatory Skills :
AWS, Terraform, Ansible, CloudFormation, Jenkins, GitLab CI, GitHub Actions, Docker, Kubernetes, Helm, Python, Bash, Grafana, Prometheus, ELK Stack, CloudWatch, New Relic, SRE, CI/CD, Infrastructure as Code (IaC), Linux
Key Responsibilities :
- Design, deploy, and manage cloud infrastructure primarily on AWS (EC2, VPC, IAM, S3, RDS, Route53, ALB, Auto Scaling, Lambda).
- Build and maintain Infrastructure as Code (IaC) using Terraform, Ansible, and CloudFormation.
- Develop and optimize CI/CD pipelines using Jenkins, GitLab CI, and GitHub Actions.
- Deploy and manage containerized applications using Docker, Kubernetes, and Helm.
- Implement monitoring and observability using Grafana, Prometheus, ELK Stack, CloudWatch, and New Relic.
- Drive SRE practices by defining SLIs, SLOs, SLAs, handling production incidents, conducting RCA, and improving system reliability.
- Automate operational tasks using Python, Bash, and Groovy scripting.
- Collaborate with Development, QA, Security, and Operations teams to ensure reliable and secure software delivery.
Required Skills & Qualifications :
- Bachelor's degree in Computer Science, IT, Electronics, or a related field.
- 5+ years of experience in DevOps, SRE, or Cloud Infrastructure.
- Strong expertise in AWS, with exposure to Azure/GCP.
- Hands-on experience with Terraform, Ansible, CloudFormation, Docker, Kubernetes, Helm, Jenkins, GitLab CI, GitHub Actions, and Git.
- Strong scripting skills in Python and Bash.
- Experience with monitoring tools such as Grafana, Prometheus, ELK Stack, CloudWatch, and New Relic.
- Good understanding of Linux, networking, SQL, and cloud security best practices.
Preferred Skills :
- Experience with multi-cloud environments and DevSecOps practices.
- Knowledge of disaster recovery, automation, and microservices architecture.
- Strong troubleshooting, communication, and problem-solving skills.
Job Description:
- Infrastructure Management: Design, implement, and manage scalable, reliable, and secure cloud infrastructure using AWS, GCP, and/or Azure.
- CI/CD Pipelines: Develop and maintain continuous integration and continuous deployment (CI/CD) pipelines to streamline the development lifecycle.
- Automation: Automate infrastructure provisioning, configuration management, and application deployment processes.
- Monitoring and Performance: Implement monitoring, logging, and alerting solutions to ensure system health, performance, and reliability.
- Security: Ensure the security of cloud infrastructure and applications, including identity management and compliance with industry standards.
- Collaboration: Work closely with client and development teams to integrate DevOps practices and deliver high-quality software.
- Documentation: Maintain comprehensive documentation of infrastructure, configurations, and processes.
- Innovation: Stay current with emerging technologies and industry trends, integrating them into the DevOps strategy as appropriate.
Qualifications:
- Education: Bachelor's degree in Computer Science, Information Technology, or a related field.
- Experience: 7 - 10 years of overall experience with relevant experience of at least 7 years in DevOps and served as a lead or senior engineer.
Location: Bangalore preferred / Hybrid as applicable
Experience: 3+ years
Education: B.E/B.Tech in Computer Science, Engineering or a related technical discipline
Salary: Above market standards, flexible for the right candidate
Career growth: Long-term opportunity with potential to lead DevOps architecture and cloud platform operations
About FrontM
FrontM builds software platforms for frontline workforces operating in remote and low-connectivity environments, with a strong focus on the maritime industry. The platform supports communication, collaboration, healthcare, learning, welfare and operational workflows across mobile, web, kiosk and connected device environments.
The platform runs across cloud infrastructure, constrained networks and specialised customer environments, requiring reliable DevOps practices, strong observability, secure architecture and careful operational discipline.
Role Summary
As a Senior DevOps Engineer, you will take ownership of FrontM’s AWS cloud infrastructure, CI/CD pipelines, platform reliability and technical operations. You will work closely with the VP of Delivery, CTO and CEO to maintain secure, scalable and high-availability infrastructure for FrontM’s production systems.
This role requires strong hands-on DevOps experience, broad AWS knowledge, Kubernetes experience and the ability to troubleshoot complex networking and production issues across multi-domain SaaS environments.
Key Responsibilities
Cloud Infrastructure & DevOps Architecture (≈45%)
· Own, maintain and improve AWS cloud infrastructure for FrontM platforms
· Create and maintain Terraform scripts for infrastructure deployment and management
· Manage Kubernetes workloads deployed within AWS EKS
· Support multi-zone AWS infrastructure design for availability, resilience and scale
· Maintain AWS services including Route 53, EC2, API Gateway, VPC, VPN, AWS Cognito, ElastiCache, DynamoDB and Lambda
· Contribute to DevOps architecture planning in line with FrontM’s platform roadmap
CI/CD, Operations & Platform Reliability (≈35%)
· Build, maintain and improve CI/CD pipelines for backend and platform services
· Oversee technical operations with hands-on administration, monitoring and release support
· Ensure continuous server uptime, stability, performance and maintainability
· Debug, respond to and restore system outages in production and staging environments
· Improve observability across infrastructure and applications, including migration from Elastic stack to logz.io
· Support backend stability, scale and performance across Node.js, Java and related services
Security, Networking & Production Support (≈20%)
· Maintain AWS security configurations, access controls and monitoring practices
· Support complex networking requirements across multi-domain SaaS implementations
· Troubleshoot network, infrastructure and access issues with internal teams and customer-side users
· Work with backend teams to support API integrations and infrastructure abstractions for complex requirements
· Document operational procedures, incident findings and technical support steps clearly
Required Technical Skills
Cloud Infrastructure & AWS
· Strong hands-on experience with AWS infrastructure and cloud operations
· Experience with Route 53, EC2, API Gateway, VPC, VPN, AWS Cognito, ElastiCache, DynamoDB and Lambda
· Experience with AWS security setup, monitoring and multi-zone infrastructure
· Ability to manage infrastructure using Terraform
Kubernetes, CI/CD & Observability
· Strong experience with Kubernetes, preferably AWS EKS
· Extensive CI/CD and DevOps experience
· Experience with infrastructure observability and application monitoring tools
· Ability to diagnose production bottlenecks, server failures and performance issues
Backend, Networking & SaaS Operations
· Experience supporting Node.js, Java and backend system procedures for stability and scale
· Good understanding of APIs, integrations and backend service dependencies
· Experience with complex networking and multi-domain SaaS implementations
· Ability to troubleshoot technical issues with non-technical end users
Nice to Have
· Experience with MongoDB clusters in MongoDB Atlas
Personal Attributes
· Strong ownership mindset for uptime, reliability and production stability
· Practical problem-solving approach with the ability to act quickly during incidents
· Clear written and spoken communication in English
· Ability to work independently and coordinate with senior management when required
· Comfortable working in fast-moving engineering teams
· Attention to detail in security, monitoring, documentation and operational processes
Why join FrontM?
Long-Term Career Growth
Opportunity to work on cloud infrastructure used by global maritime and remote workforce customers, with scope to grow into DevOps architecture and platform leadership roles.
Engineering Challenges That Matter
Work on infrastructure that supports applications used in remote, low-bandwidth and operationally demanding environments.
Broad Technical Ownership
Take responsibility across cloud infrastructure, Kubernetes, CI/CD, observability, networking, security and production reliability.
Apply now
Join a team focused on building reliable software infrastructure for real-world use cases and contribute to systems used across the global maritime workforce.
Job Description:
Pre-requisite skills required for a DevOps Engineer role include:
- 6+yrs exp in DevOps
- Experience working on Linux based infrastructure
- knowledge in AWS, docker, CI/CD tools
- Hands on exp in Python/shell scripting language
- hands on exp in AWS and Azure
- Work exp in Docker, Terraform, Ansible, Kubernetes, LINUX
- Excellent understanding of Ruby, Python, Perl, and Java
- Configuration and managing databases such as MySQL, Mongo
- Excellent troubleshooting
- Working knowledge of various tools, open-source technologies, and cloud services
Job Summary:
We are looking for an experienced Cloud / DevOps Engineer with strong hands-on experience in AWS, Kubernetes, OpenShift, and Linux. The candidate will be responsible for managing cloud infrastructure, containerized applications, deployments, and production environments.
Roles & Responsibilities:
- Design, deploy, and manage cloud infrastructure on AWS.
- Manage and administer Kubernetes and OpenShift environments.
- Deploy, configure, and troubleshoot containerized applications.
- Perform Linux administration, troubleshooting, and system monitoring.
- Monitor application and infrastructure performance and resolve production issues.
- Automate deployment and operational activities wherever possible.
- Work with development and operations teams to support application deployments.
- Ensure system availability, reliability, and security.
- Troubleshoot issues related to Kubernetes clusters, containers, networking, and Linux systems.
- Follow DevOps best practices for CI/CD, infrastructure management, and automation.
Required Skills:
- 7–10 years of relevant experience in Cloud/DevOps.
- Strong hands-on experience with AWS.
- Good knowledge of Kubernetes and OpenShift.
- Strong Linux administration and troubleshooting skills.
- Experience in containerization and deployment of applications.
- Good understanding of cloud infrastructure and production support.
Key Responsibilities
- Automate application deployments from Bitbucket to servers using CI/CD pipelines.
- Design and manage scalable, highly available AWS infrastructure.
- Implement Auto Scaling, ELB, and Route 53 for traffic management and high availability.
- Work with AWS services including IAM, RDS, DynamoDB, EC2, and other cloud services.
- Build and manage Docker containers and server images.
- Deploy and manage applications using Kubernetes.
- Implement Infrastructure as Code using Terraform, CloudFormation, or Ansible.
- Develop automation scripts using Python and Bash.
- Implement monitoring and logging using tools such as Prometheus, Grafana, and ELK.
- Integrate security and compliance practices into CI/CD pipelines.
- Optimize infrastructure for security, scalability, performance, and cost.
Required Skills
- 3+ years of experience in DevOps or a similar role.
- Strong knowledge of AWS beyond EC2.
- Hands-on experience with Jenkins or similar CI/CD tools.
- Experience with Docker and Kubernetes.
- Good understanding of Terraform/IaC and automation.
- Proficiency in Python and/or Bash scripting.
- Knowledge of DevSecOps, security, and compliance best practices.
- Strong troubleshooting and problem-solving skills.
The Role
As a **DevOps Engineer** you'll own the infrastructure and delivery backbone that
keeps our platform running as we grow. You'll build the CI/CD, cloud infrastructure, and
observability that let a small, fast-moving team ship confidently — and you'll keep our AI and
data workloads reliable and affordable at scale.
This is a hands-on role with real ownership: you won't be maintaining someone else's setup,
you'll be shaping ours. You'll work closely with the backend, AI/ML, and data teams to make
deployment boring, incidents rare, and scaling a non-event. ---
What You'll Own
**CI/CD & developer experience**
- Build and maintain fast, reliable CI/CD pipelines so engineers ship multiple times a day with
confidence. - Make the path from commit to production simple, safe, and repeatable, with sensible
automated testing, rollbacks, and release controls.
**Cloud infrastructure & IaC** - Own our cloud infrastructure (AWS/GCP) end to end, managed as code (Terraform or
similar) — no click-ops. - Design for scale and cost-efficiency as store and conversation volumes grow.
**Containers & orchestration** - Run our services on containers/Kubernetes: deployments, autoscaling, networking, and
resource management. - Support the specific needs of AI/ML workloads, including GPU-backed inference and batch
processing for the speech pipeline.
**Reliability & observability (SRE)** - Own uptime, performance, and incident response — monitoring, logging, tracing, alerting,
on-call, and blameless postmortems. - Define and defend SLOs; keep the platform dependable as it scales across clients.
**Data & pipeline infrastructure** - Support the infrastructure behind large-scale, edge-to-cloud data movement and
processing (audio ingestion, ASR/AI pipelines, analytics). - Keep data workloads reliable, performant, and cost-aware.
**Security & compliance** - Bake security into the platform: secrets management, IAM/least-privilege, encryption in
transit and at rest, network hardening, and vulnerability management. - Support compliance readiness (including India's DPDP Act and enterprise-client security
requirements) for a product that handles sensitive customer conversations.
**Cost & scale** - Own cloud cost visibility and optimization; make scaling decisions that balance reliability
and spend. ---
What You'll Bring - 6+ years in DevOps, SRE, platform, or infrastructure engineering, running production
systems at meaningful scale. - Strong hands-on experience with a major cloud provider (**AWS or Azure or GCP**) and
Infrastructure-as-Code (**Terraform** or equivalent). - Solid experience with **containers and Kubernetes** in production. - Experience building and owning **CI/CD** pipelines (e.g. GitHub Actions, GitLab CI,
Jenkins, Argo, or similar).
- Comfort with a scripting/automation language (Python, Go, or Bash) and a strong
automation-first mindset. - Real experience with **observability** (Prometheus/Grafana, ELK, Datadog,
OpenTelemetry, or similar) and running incident response / on-call. - A security-conscious approach — secrets, IAM, encryption, and least-privilege as defaults. - Startup temperament: ownership, pragmatism, and a bias to automate and ship. - Based in or willing to relocate to Bangalore, and up for an onsite/hybrid, in-person team.
Bonus Points - Experience running **ML/AI or GPU workloads** in production (inference serving, batch
pipelines, model deployment). - Experience with data-intensive infrastructure — streaming/queues (Kafka, SQS), data
pipelines, or large object/audio storage. - Exposure to **edge devices / IoT fleets**, OTA updates, or high-volume device-to-cloud
ingestion. - Experience with compliance/security frameworks (SOC 2, ISO 27001, DPDP). - FinOps / cloud cost-optimization experience. - Early-stage startup experience. ---
Why Join - Own infrastructure that's already live with leading retail brands and growing fast — real
scale, real impact. - Work across genuinely interesting workloads: speech AI, GPU inference, large-scale data,
and edge-to-cloud ingestion. - Small team, high ownership, direct line to engineering leadership — your decisions ship. - Build the platform foundation of a category-defining product from an






