Senior DevOps Engineer at wwwthehiveai · Gurugram · 5 - 20 years · ₹10L - ₹40L / yr · Raised funding · Posted 24 Jan 2023
About Hive
Hive is the leading provider of cloud-based AI solutions for content understanding,
trusted by the world’s largest, fastest growing, and most innovative organizations. The
company empowers developers with a portfolio of best-in-class, pre-trained AI models, serving billions of customer API requests every month. Hive also offers turnkey software applications powered by proprietary AI models and datasets, enabling breakthrough use cases across industries. Together, Hive’s solutions are transforming content moderation, brand protection, sponsorship measurement, context-based ad targeting, and more.
Hive has raised over $120M in capital from leading investors, including General Catalyst, 8VC, Glynn Capital, Bain & Company, Visa Ventures, and others. We have over 250 employees globally in our San Francisco, Seattle, and Delhi offices. Please reach out if you are interested in joining the future of AI!
About Role
Our unique machine learning needs led us to open our own data centers, with an
emphasis on distributed high performance computing integrating GPUs. Even with these data centers, we maintain a hybrid infrastructure with public clouds when the right fit. As we continue to commercialize our machine learning models, we also need to grow our DevOps and Site Reliability team to maintain the reliability of our enterprise SaaS offering for our customers. Our ideal candidate is someone who is
able to thrive in an unstructured environment and takes automation seriously. You believe there is no task that can’t be automated and no server scale too large. You take pride in optimizing performance at scale in every part of the stack and never manually performing the same task twice.
Responsibilities
● Create tools and processes for deploying and managing hardware for Private Cloud Infrastructure.
● Improve workflows of developer, data, and machine learning teams
● Manage integration and deployment tooling
● Create and maintain monitoring and alerting tools and dashboards for various services, and audit infrastructure
● Manage a diverse array of technology platforms, following best practices and
procedures
● Participate in on-call rotation and root cause analysis
Requirements
● Minimum 5 - 10 years of previous experience working directly with Software
Engineering teams as a developer, DevOps Engineer, or Site Reliability
Engineer.
● Experience with infrastructure as a service, distributed systems, and software design at a high-level.
● Comfortable working on Linux infrastructures (Debian) via the CLIAble to learn quickly in a fast-paced environment.
● Able to debug, optimize, and automate routine tasks
● Able to multitask, prioritize, and manage time efficiently independently
● Can communicate effectively across teams and management levels
● Degree in computer science, or similar, is an added plus!
Technology Stack
● Operating Systems - Linux/Debian Family/Ubuntu
● Configuration Management - Chef
● Containerization - Docker
● Container Orchestrators - Mesosphere/Kubernetes
● Scripting Languages - Python/Ruby/Node/Bash
● CI/CD Tools - Jenkins
● Network hardware - Arista/Cisco/Fortinet
● Hardware - HP/SuperMicro
● Storage - Ceph, S3
● Database - Scylla, Postgres, Pivotal GreenPlum
● Message Brokers: RabbitMQ
● Logging/Search - ELK Stack
● AWS: VPC/EC2/IAM/S3
● Networking: TCP / IP, ICMP, SSH, DNS, HTTP, SSL / TLS, Storage systems,
RAID, distributed file systems, NFS / iSCSI / CIFS
Who we are
We are a group of ambitious individuals who are passionate about creating a revolutionary AI company. At Hive, you will have a steep learning curve and an opportunity to contribute to one of the fastest growing AI start-ups in San Francisco. The work you do here will have a noticeable and direct impact on the
development of the company.
Thank you for your interest in Hive and we hope to meet you soon

Similar jobs (10)
Job Title : DevOps Engineer – Linux, AWS & Infrastructure
Experience : 3+ Years
Location : Sector 48, Gurgaon
Work Mode : 6 Days WFO – Monday to Saturday
Week Off : 01st & 03rd Saturday Off
Employment Type : Full-Time
Job Summary :
We are looking for a DevOps Engineer with strong hands-on experience in Linux, Networking, Server Administration, AWS, CI/CD, Containers, Kubernetes, and Infrastructure Automation. The ideal candidate should have strong troubleshooting skills, production ownership, and the ability to manage and automate infrastructure reliably.
Key Responsibilities :
- Manage and troubleshoot Linux servers, bare-metal infrastructure, and server environments.
- Perform system administration, networking, performance monitoring, and production troubleshooting.
- Design, maintain, and optimize Jenkins-based CI/CD pipelines and deployment workflows.
- Manage Docker containers and Kubernetes environments.
- Work with AWS Cloud services, infrastructure, security, and deployment environments.
- Monitor system health, application performance, logs, and infrastructure using appropriate monitoring and logging tools.
- Implement and maintain Infrastructure as Code (IaC) using tools such as Terraform or CloudFormation.
- Automate repetitive operational tasks using Python, Bash, Shell scripting, or similar technologies.
- Implement infrastructure and application security, access controls, patching, and hardening.
- Investigate production incidents, perform root-cause analysis (RCA), and drive issues to resolution.
- Take end-to-end ownership of infrastructure reliability, availability, and operational issues.
- Collaborate with development and other engineering teams to improve deployment, scalability, and system reliability.
Mandatory Skills :
Linux & Networking | Bare Metal / Server Administration | Jenkins / CI-CD | Docker / Containers | AWS Cloud | Kubernetes | Git | Monitoring & Logging | Security | Infrastructure as Code (IaC) | Automation / Scripting | Production Troubleshooting & Ownership
Preferred Skills :
- Strong understanding of TCP/IP, DNS, HTTP/HTTPS, SSH, load balancing, and networking fundamentals.
- Hands-on experience with Terraform / CloudFormation.
- Experience with Prometheus, Grafana, ELK / EFK, CloudWatch, or similar monitoring / logging tools.
- Good understanding of Linux performance troubleshooting, processes, memory, disk, networking, and file systems.
- Experience handling production incidents, RCA, deployments, and system reliability.
- Exposure to cloud security, IAM, secrets management, and server hardening.
What We’re Looking For :
- 3+ years of hands-on experience in DevOps / SRE / Infrastructure Engineering.
- Strong practical knowledge rather than certification-based/theoretical understanding.
- Good troubleshooting and analytical skills.
- Strong sense of ownership and accountability for production systems.
- Comfortable working in a 6-day work-from-office environment.
Devops + Python
Mandatory: Python+Kubernetes Containerisation+ DevOps Exposure+ Cloud Exposure(any)
4-8 years /16 LPA
10+ years /26 LPA
Bangalore/Hyderabad
immediate to 15 days.
Job Summary:
We are looking for an experienced Cloud / DevOps Engineer with strong hands-on experience in AWS, Kubernetes, OpenShift, and Linux. The candidate will be responsible for managing cloud infrastructure, containerized applications, deployments, and production environments.
Roles & Responsibilities:
- Design, deploy, and manage cloud infrastructure on AWS.
- Manage and administer Kubernetes and OpenShift environments.
- Deploy, configure, and troubleshoot containerized applications.
- Perform Linux administration, troubleshooting, and system monitoring.
- Monitor application and infrastructure performance and resolve production issues.
- Automate deployment and operational activities wherever possible.
- Work with development and operations teams to support application deployments.
- Ensure system availability, reliability, and security.
- Troubleshoot issues related to Kubernetes clusters, containers, networking, and Linux systems.
- Follow DevOps best practices for CI/CD, infrastructure management, and automation.
Required Skills:
- 7–10 years of relevant experience in Cloud/DevOps.
- Strong hands-on experience with AWS.
- Good knowledge of Kubernetes and OpenShift.
- Strong Linux administration and troubleshooting skills.
- Experience in containerization and deployment of applications.
- Good understanding of cloud infrastructure and production support.
* Deploy and maintain company websites and HRM Cloud applications.
* Configure and manage cloud servers and Linux environments.
* Set up CI/CD pipelines using GitHub/GitHub Actions.
* Configure domains, DNS, SSL certificates, and HTTPS.
* Deploy frontend, backend/API, and database services.
* Configure production and staging environments.
* Manage Docker containers where required.
* Monitor server performance, uptime, logs, and application health.
* Implement regular database and server backups.
* Maintain security, access controls, firewalls, and server permissions.
* Troubleshoot deployment, server, network, and application issues.
* Work closely with Full-Stack Developers and QA teams.
* Maintain deployment documentation and technical procedures.
Required Skills
* AWS / Azure / DigitalOcean or equivalent cloud platform
* Linux server administration
* Git & GitHub
* CI/CD and GitHub Actions
* Docker
* Nginx / Apache
* DNS & domain configuration
* SSL/TLS and HTTPS
* Database deployment and backup
* Basic networking and security
* Monitoring and troubleshooting
* Experience deploying web applications to production
Preferred Skills
* Kubernetes
* Terraform / Infrastructure as Code
* Cloud security
* Load balancing
* Redis / caching
* PostgreSQL / MySQL / MongoDB
* Experience with Node.js, React, PHP/Laravel, or .NET applications
What We Offer
* Opportunity to work on HRM Cloud and SaaS products
* Exposure to real-world cloud infrastructure and production systems
* Remote/hybrid working opportunity
* Growth opportunities within the technology team
* Opportunity to work with an international product-focused company
How to Apply
Send your CV,GitHub profile, LinkedIn profile, and details of your previous cloud/DevOps projects.
Subject: Application – DevOps & Cloud Engineer
DevOps / Infrastructure Engineer
Location: Chennai
Experience: 5+ Years
Role: DevOps / Infrastructure Engineer
Job Description
We are looking for an experienced DevOps / Infrastructure Engineer with strong hands-on experience in Linux administration, containerization, Kubernetes, automation, monitoring, and troubleshooting.
Mandatory Skills
- 5+ years of experience in DevOps / Infrastructure Administration
- Strong hands-on experience with Linux Administration
- Experience with Docker and Kubernetes
- Monitoring tools: AppDynamics, Prometheus, Grafana
- Strong Shell Scripting / Python Scripting skills
- Hands-on experience with Ansible 4.1
- Strong troubleshooting and problem-solving skills
- Experience in infrastructure/application monitoring and production support
- Good understanding of DevOps practices and automation
Key Responsibilities
- Manage and support Linux-based infrastructure and production environments.
- Deploy, manage, and troubleshoot applications using Docker and Kubernetes.
- Develop and maintain automation scripts using Shell/Python.
- Automate infrastructure and configuration management using Ansible.
- Monitor applications and infrastructure using AppDynamics, Prometheus, and Grafana.
- Perform root-cause analysis and resolve infrastructure/application issues.
- Handle incidents, troubleshoot performance issues, and ensure system availability.
- Collaborate with development and operations teams to improve deployment and operational processes.
Job Title : DevOps Engineer / Site Reliability Engineer (SRE)
Experience : 4+ Years
Location : Gurugram, Sector 48, Haryana (On-site)
Employment Type : Full-Time
Working Days : Monday to Saturday (1st & 3rd Saturday Off)
About the Role :
We are looking for a hands-on DevOps Engineer / Site Reliability Engineer (SRE) with strong experience in Linux, AWS, Kubernetes, Docker, CI/CD, Infrastructure as Code, and production application deployments.
The ideal candidate should have real-world production experience, excellent troubleshooting skills, and the ability to manage both infrastructure and application-level issues.
Mandatory Skills :
Linux, AWS, Docker, Kubernetes, Terraform, Ansible, Jenkins, GitHub Actions, GitLab CI/CD, CI/CD, Infrastructure as Code (IaC), Python, Bash, Git, Grafana, Prometheus, ELK, CloudWatch, New Relic, SRE (SLI/SLO/SLA), Networking (DNS, HTTP/HTTPS, TCP/IP, Load Balancer), Production Application Deployment & Troubleshooting
Key Responsibilities :
- Manage and maintain AWS cloud infrastructure.
- Build and optimize CI/CD pipelines using Jenkins, GitHub Actions, or GitLab CI.
- Deploy, monitor, and troubleshoot applications across production environments.
- Automate infrastructure using Terraform and Ansible.
- Manage Docker containers and Kubernetes clusters.
- Monitor systems using Grafana, Prometheus, ELK, CloudWatch, and New Relic.
- Perform Linux server administration and troubleshooting.
- Handle production incidents, Root Cause Analysis (RCA), and improve system reliability.
- Collaborate with development teams to support application releases and automation.
Required Qualifications :
- Bachelor's degree in Computer Science or related field.
- 4+ years of hands-on experience in DevOps / SRE.
- Strong Linux administration and production troubleshooting skills.
- Experience with AWS and modern DevOps toolchains.
- Hands-on experience with application deployment and production support.
What We're Looking For :
- Strong practical Linux and cloud knowledge.
- Real production experience with application deployments.
- Ability to troubleshoot both infrastructure and application issues.
- Experience handling live production incidents.
- Excellent communication and problem-solving skills.
- Candidates should be comfortable with scenario-based technical discussions and demonstrate genuine hands-on expertise.
Interview Process :
- HR Screening
- Technical Round
- Client Technical Round
- Final Discussion
Note : The interview will focus on practical hands-on experience in Linux, AWS, Kubernetes, Docker, CI/CD, Infrastructure as Code, application deployment, production troubleshooting, and real-world DevOps scenarios.
About the Role
The non-negotiable is deep, hands-on infrastructure expertise spanning hybrid cloud and self-hosted systems. You will architect and maintain our unique infrastructure combining AWS CDN, bare metal servers, and Kubernetes clusters. This is not just maintenance work: you will establish company-wide DevOps policy, build security-hardened environments, and create the documentation and processes that scale with us. You will design our CI/CD pipelines, implement zero-trust networking, and guide the technical team on how to keep our production systems robustly online. This role is hands-on, autonomous, and sets the standard for how we approach infrastructure as the company grows.
What You'll Build
- Hybrid Infrastructure Management: Architect and maintain our unique infrastructure spanning AWS CDN, bare metal servers, and Kubernetes clusters for computationally intensive facial analysis workloads.
- CI/CD Pipeline Architecture: Design and implement CircleCI or Jenkins pipelines with comprehensive build testing, versioning, and change logging.
- Zero-Trust Networking: Build and maintain mesh topology networks using Tailscale or Wireguard to securely connect our hybrid infrastructure.
- Security-First Culture: Establish and enforce security policies including key rotation, access controls, compliance frameworks, and employee security management.
- Infrastructure as Code: Document and codify all infrastructure decisions, creating repeatable, auditable deployments.
- Containerization Strategy: Implement and optimize Docker/K8s deployments for our AI/ML workloads.
- Cost Optimization: Continue our approach of strategic compute placement using owned, rented or borrowed infrastructure where it makes financial sense without sacrificing security or reliability.
- Observability & Monitoring: Implement comprehensive logging, monitoring, and alerting across our distributed systems.
What We're Looking For
- 5+ years of DevOps or infrastructure engineering experience, with a track record of building from scratch
- Hybrid infrastructure expertise: experience managing both cloud (AWS) and self-hosted infrastructure, understanding the tradeoffs and security risks of each
- Kubernetes production experience: deep knowledge of cluster design, operations, and scaling
- Networking mastery: strong understanding of VPCs, mesh networks, VPNs, and zero-trust architectures
- Security-first mindset: experience with security compliance, key management, IAM policies, and hardening production systems
- CI/CD expertise: hands-on experience building robust pipelines for build testing before deployment
- Infrastructure as Code: proficiency with Terraform, Ansible, or similar tools
- Scripting and automation: strong Python, Bash, or Go skills for tooling and automation
- Policy and documentation: ability to establish best practices and document them clearly for team adoption
- Leadership mentality: comfortable setting standards and directing technical decisions, not just executing them
Nice to Have
- Experience architecting infrastructure for AI/ML workloads
- Background in a fast-moving startup or scale-up environment
- H ands-on experience with cost optimization across cloud and on-premises infrastructure
Why Join
- Opportunity to solve real healthcare problems with cutting-edge technology
- Well-funded startup with a strong market presence
- Work with advanced AI technology in a healthcare context
- Collaborate with a talented team in a fast-paced environment
- Competitive salary with equity options
- Performance and quarterly bonuses
- Professional development opportunities
Compensation and Logistics
- Remote, full-time
- Reports to: Head of Engineering
- Competitive based on experience
Job Title : DevOps / Infrastructure Engineer
Experience : 3+ Years
Location : Gurugram Sector 48
Work Mode : 6 Days WFO (Monday to Saturday) / 01st & 03rd Saturdays are off
Employment Type : Full-time
Role Overview :
We are looking for a DevOps / Infrastructure Engineer with strong hands-on experience in Linux administration, Jenkins, Docker, networking, bare-metal servers, AWS, Redis, and MongoDB. The candidate should be capable of independently troubleshooting infrastructure, deployment, networking, and application-related issues in production environments.
Mandatory / Non-Negotiable Skills :
- Strong hands-on experience with Linux Administration & Troubleshooting.Strong experience with Jenkins and CI/CD pipelines.
- Hands-on experience with Docker and containerized environments.
- Strong understanding of Networking concepts – TCP/IP, DNS, HTTP/HTTPS, ports, routing, firewalls, load balancing, etc.
- Hands-on experience with Bare Metal Servers / Server Administration.
- Strong hands-on experience with AWS (EC2, VPC, IAM, Security Groups, Load Balancers, S3 & CloudWatch)
- Working knowledge of Redis
- Working knowledge of MongoDB
- Strong production troubleshooting and incident-resolution skills
Key Responsibilities :
- Manage, configure, monitor, and troubleshoot Linux and bare-metal servers
- Build, maintain, and troubleshoot Jenkins CI/CD pipelines
- Deploy, manage, and troubleshoot applications using Docker
- Manage and troubleshoot AWS infrastructure and services
- Configure and maintain networking, security groups, firewalls, ports, and connectivity
- Support and maintain Redis and MongoDB environments
- Perform server health checks, log analysis, performance troubleshooting, and incident resolution
- Work closely with development teams to support application deployments
- Identify root causes of infrastructure and production issues and implement preventive solutions
- Maintain infrastructure security, availability, and reliability
- Automate repetitive operational tasks wherever possible
Required Candidate Profile :
- 3+ years of relevant experience in DevOps, Infrastructure, System Administration, or related roles.
- Strong hands-on / production experience with all mandatory technologies.
- Good understanding of Linux systems and infrastructure.
- Strong troubleshooting and problem-solving abilities.
- Ability to take ownership of production infrastructure and deployment issues.
- Good communication and collaboration skills.
Job Description:
Pre-requisite skills required for a DevOps Engineer role include:
- 6+yrs exp in DevOps
- Experience working on Linux based infrastructure
- knowledge in AWS, docker, CI/CD tools
- Hands on exp in Python/shell scripting language
- hands on exp in AWS and Azure
- Work exp in Docker, Terraform, Ansible, Kubernetes, LINUX
- Excellent understanding of Ruby, Python, Perl, and Java
- Configuration and managing databases such as MySQL, Mongo
- Excellent troubleshooting
- Working knowledge of various tools, open-source technologies, and cloud services
Job Title : DevOps Engineer / Site Reliability Engineer (SRE)
Experience : 5+ Years
Location : Gurugram, Haryana
Work Mode : On-site (Full-time)
About the Role :
We are looking for a skilled DevOps Engineer with 5+ years of experience in cloud infrastructure, CI/CD, automation, Kubernetes, and Site Reliability Engineering (SRE). The ideal candidate will be responsible for building scalable cloud infrastructure, automating deployments, improving system reliability, and ensuring high availability across production environments.
Mandatory Skills :
AWS, Terraform, Ansible, CloudFormation, Jenkins, GitLab CI, GitHub Actions, Docker, Kubernetes, Helm, Python, Bash, Grafana, Prometheus, ELK Stack, CloudWatch, New Relic, SRE, CI/CD, Infrastructure as Code (IaC), Linux
Key Responsibilities :
- Design, deploy, and manage cloud infrastructure primarily on AWS (EC2, VPC, IAM, S3, RDS, Route53, ALB, Auto Scaling, Lambda).
- Build and maintain Infrastructure as Code (IaC) using Terraform, Ansible, and CloudFormation.
- Develop and optimize CI/CD pipelines using Jenkins, GitLab CI, and GitHub Actions.
- Deploy and manage containerized applications using Docker, Kubernetes, and Helm.
- Implement monitoring and observability using Grafana, Prometheus, ELK Stack, CloudWatch, and New Relic.
- Drive SRE practices by defining SLIs, SLOs, SLAs, handling production incidents, conducting RCA, and improving system reliability.
- Automate operational tasks using Python, Bash, and Groovy scripting.
- Collaborate with Development, QA, Security, and Operations teams to ensure reliable and secure software delivery.
Required Skills & Qualifications :
- Bachelor's degree in Computer Science, IT, Electronics, or a related field.
- 5+ years of experience in DevOps, SRE, or Cloud Infrastructure.
- Strong expertise in AWS, with exposure to Azure/GCP.
- Hands-on experience with Terraform, Ansible, CloudFormation, Docker, Kubernetes, Helm, Jenkins, GitLab CI, GitHub Actions, and Git.
- Strong scripting skills in Python and Bash.
- Experience with monitoring tools such as Grafana, Prometheus, ELK Stack, CloudWatch, and New Relic.
- Good understanding of Linux, networking, SQL, and cloud security best practices.
Preferred Skills :
- Experience with multi-cloud environments and DevSecOps practices.
- Knowledge of disaster recovery, automation, and microservices architecture.
- Strong troubleshooting, communication, and problem-solving skills.











