Senior Devops Engineer at Questionpro · Remote, Pune · 7 - 14 years · ₹14L - ₹26L / yr · Profitable · Remote friendly · Posted 7 Jan 2022
About QuestionPro:
QuestionPro is one of the leading market research platforms. We have a wide range of products in Market Research, Customer Experience, Employee Experience, Vehicle Experience. All our products are multi-tenant SAAS platforms built on the latest technologies.
Our infrastructure is spread across 6 Data Centers across the globe. The platform collects over 10Million Surveys every month. Our Customer Experience platform was named top provider in the Gartner Voice of the Customer Rankings. Ever since we launched in 2016, we have grown by over 200% YoY. All up we are on plan to hit $ 31M in 2021. We are bootstrapped and proud to get where we are without any funding or investments.
Our operations are spread across the globe with offices in the US, Mexico, Germany, UK, UAE, and Canada.
https://www.questionpro.com/blog/cx-top-provider-gartner-voc-rankings/
We are a bootstrapped company and proud to have not taken any funding or investment.
QuestionPro has a particularly exciting journey ahead, requiring a passionate individual to join our growing team. If you are a true technology craftsman and want to build cutting-edge software solutions, hit us!
We operate 100% remote. You will be working from any place you desire for this position.
Responsibilities
- You will be responsible for key deliverables that would help improve the quality and reliability of our infrastructure spread across 8 data centres, both cloud and hybrid.
- Streamline Life Cycle Management activities for the Infrastructure.
- You will be interacting with DevOps & Support teams to quickly investigate and mitigate the problems impacting customers at various levels in the Infrastructure including MySQL Database and Engineered Systems technology stack.
Skills & Requirements
Must-Have:
- 8+ years of experience in Linux System Administration / Development / QA roles with a thorough understanding of Software Development Life Cycle
- Excellent knowledge of Linux System administration activities
- A very good understanding of Linux kernel internals, Server Virtualization, Networking & Security layer
- Any experience in IO subsystem, Operating Systems, Storage technologies would be a definite plus
- Hands-on experience with automation of system administration activities
- Hands on experience in building OS images and testing based on standard test framework.
- Proven ability to triaging and resolve issues during patch testing & certification
- Experience in OEM, Linux OS Patching, Yum, KSplice, etc.
- Excellent Scripting skills in Bash, Perl, Python, or similar scripting languages
Good To Have:
- Good experience on private cloud management.
- Proxmox virtualization
- Ability to build and grow the team.

About Questionpro
About
Connect with the team
Similar jobs (10)
Location : Mumbai
Big Picture (The Opportunity) :
Are you looking for an opportunity to advance your Career? & If you are able to maintain a positive attitude even when everything goes wrong, if you are detail oriented and self motivated with a passion to learn and improve your skills and knowledge, we have a perfect job for you !
What do we want from you ? (Our Expectations) :
- Zero to 2 years experience in Linux Operating System.
- Flexible working hours - able to support occasional nights, weekends, and call-ins, able to quickly adapt to a constantly changing faced paced environment. Open to travel to sites.
- An ideal person who is excited and motivated about running and supporting a production - grade critical infrastructure and looks for opportunities to improve processes with automation.
What are you required to do ? (Your Responsibilities) :
- Proactively maintain and develop all linux infrastructure technology to maintain a 24*7*365 uptime service.
- Engineering of systems administration-related various solutions for our various SAAS products and projects as well as operational needs.
- Proactively monitoring system performance and capacity planning.
- Providing technical support to customers for Applications, Operating systems, and networking.
- Will also be the first point of contact for our clients, where installations are placed on a permanent basis, for basic troubleshooting and problem solving.
- Fault finding, analysis and logging information for reporting of performance exceptions.
- Maintain best practices on managing systems and services across all environments.
Skills & Qualification Required (Add Value) :
- Graduate - Preferred to have Bachelor‘s Degree in Engineering, Computer Science or related field.
- He / She should be familiar with the installation and configuration of Linux operating systems and setup and operation of TCP/IP networking on Linux systems also familiar with Internet concepts including SMTP, IMAP, POP, HTTP, DNS, LDAP and related protocols.
- You should possess excellent communication skills .
- Knowledge of Email concepts, Helpdesk Concept , VoIP Concept , Cloud computing will be an added advantage.
We are looking for a Linux Administrator to keep our servers secure, patched and reliable.
Responsibilities
- Administer RHEL and Ubuntu servers
- Automate tasks with Shell scripts and Ansible
- Patch, monitor and troubleshoot systems
- Manage users, storage and networking
Requirements
- 2+ years of Linux administration
- Strong shell scripting skills
- RHCSA or RHCE is a plus
We are looking for a System Administrator to manage our servers and core IT infrastructure.
Responsibilities
- Administer Linux (RHEL) and Windows Server systems
- Manage VMware virtualisation and Active Directory
- Handle patching, user access and system hardening
- Monitor systems and resolve infrastructure issues
Requirements
- 1+ years of system administration experience
- Strong Linux and Windows Server skills
- RHCSA or MCSA certification is a plus
Job Title : DevOps / Infrastructure Engineer
Experience : 3+ Years
Location : Gurugram Sector 48
Work Mode : 6 Days WFO (Monday to Saturday) / 01st & 03rd Saturdays are off
Employment Type : Full-time
Role Overview :
We are looking for a DevOps / Infrastructure Engineer with strong hands-on experience in Linux administration, Jenkins, Docker, networking, bare-metal servers, AWS, Redis, and MongoDB. The candidate should be capable of independently troubleshooting infrastructure, deployment, networking, and application-related issues in production environments.
Mandatory / Non-Negotiable Skills :
- Strong hands-on experience with Linux Administration & Troubleshooting.Strong experience with Jenkins and CI/CD pipelines.
- Hands-on experience with Docker and containerized environments.
- Strong understanding of Networking concepts – TCP/IP, DNS, HTTP/HTTPS, ports, routing, firewalls, load balancing, etc.
- Hands-on experience with Bare Metal Servers / Server Administration.
- Strong hands-on experience with AWS (EC2, VPC, IAM, Security Groups, Load Balancers, S3 & CloudWatch)
- Working knowledge of Redis
- Working knowledge of MongoDB
- Strong production troubleshooting and incident-resolution skills
Key Responsibilities :
- Manage, configure, monitor, and troubleshoot Linux and bare-metal servers
- Build, maintain, and troubleshoot Jenkins CI/CD pipelines
- Deploy, manage, and troubleshoot applications using Docker
- Manage and troubleshoot AWS infrastructure and services
- Configure and maintain networking, security groups, firewalls, ports, and connectivity
- Support and maintain Redis and MongoDB environments
- Perform server health checks, log analysis, performance troubleshooting, and incident resolution
- Work closely with development teams to support application deployments
- Identify root causes of infrastructure and production issues and implement preventive solutions
- Maintain infrastructure security, availability, and reliability
- Automate repetitive operational tasks wherever possible
Required Candidate Profile :
- 3+ years of relevant experience in DevOps, Infrastructure, System Administration, or related roles.
- Strong hands-on / production experience with all mandatory technologies.
- Good understanding of Linux systems and infrastructure.
- Strong troubleshooting and problem-solving abilities.
- Ability to take ownership of production infrastructure and deployment issues.
- Good communication and collaboration skills.

The recruiter has not been active on this job recently. You may apply but please expect a delayed response.
Job Overview
We are seeking a highly experienced Senior Linux Infrastructure Engineer with deep expertise in Linux administration, bare metal infrastructure, enterprise storage, and next-generation AI Factory / GPU infrastructure platforms. This role is focused on designing, deploying, operating, and troubleshooting large-scale Linux-based infrastructure that powers both traditional enterprise workloads and modern AI/ML environments.
This is not a DevOps-focused role. We already have a dedicated DevOps team and are looking for an engineer with extensive hands-on experience in Bare Metal as a Service (BMaaS), GPU infrastructure, high-performance storage, data center operations, and enterprise Linux platforms.
The ideal candidate will have experience building and managing infrastructure from the hardware layer up, including servers, networking, storage, GPU clusters, and AI-ready platforms. They should be comfortable working with high-performance computing (HPC), AI Factory environments, and large-scale Linux deployments where performance, reliability, and operational excellence are critical.
Key Responsibilities & Required Skills
Linux & Bare Metal Infrastructure
- Expert-level Linux administration (Ubuntu required; Red Hat and SUSE preferred)
- Deep expertise in bare metal server deployment, architecture, provisioning, and lifecycle management
- Experience operating Bare Metal as a Service (BMaaS) platforms and large-scale infrastructure environments
- Strong understanding of server hardware, including:
- BIOS/UEFI
- RAID controllers
- Firmware management
- iLO/iDRAC/IPMI
- NICs and SmartNICs
- HBA cards
- Hardware diagnostics and troubleshooting
- Experience designing, implementing, and supporting enterprise Linux infrastructure at scale
AI Factory & GPU Infrastructure
- Experience deploying and managing GPU-accelerated infrastructure for AI/ML workloads
- Understanding of NVIDIA GPU technologies including:
- A100, H100, H200, B200, or equivalent GPU platforms
- NVIDIA DGX and OEM GPU servers
- GPU provisioning and lifecycle management
- GPU monitoring and performance optimization
- Knowledge of AI Factory architecture and infrastructure requirements
- Experience supporting GPU clusters, AI training environments, and high-performance computing (HPC) workloads
- Understanding of:
- GPU resource allocation and scheduling
- Multi-GPU systems
- GPU networking requirements
- High-bandwidth, low-latency infrastructure design
- Familiarity with NVIDIA ecosystem technologies such as:
- CUDA
- NCCL
- GPUDirect Storage
- NVIDIA Fabric Manager
- NVIDIA Base Command (preferred)
Enterprise Storage & Data Platforms
- Advanced Linux storage administration:
- LVM
- XFS, EXT4
- NFS
- iSCSI
- Fibre Channel SAN
- Multipath I/O
- Strong hands-on experience with Ceph, including:
- Cluster architecture
- MON, OSD, MDS
- RBD, CephFS, RGW
- Capacity planning
- Performance tuning
- Failure recovery
- Experience with high-performance AI storage platforms such as:
- WEKA
- VAST Data
- Dell PowerScale
- Pure Storage FlashBlade
- NetApp
- Understanding of:
- NVMe-over-Fabrics (NVMe-oF)
- RDMA
- GPUDirect Storage
- Parallel file systems
- AI data pipelines
Networking & Infrastructure
- Strong networking knowledge:
- Bonding
- VLANs
- Routing
- MTU optimization
- DNS
- DHCP
- Experience with high-performance data center networking:
- 100G/200G/400G Ethernet
- RoCE
- RDMA
- Spine-Leaf architectures
- Familiarity with NVIDIA Spectrum-X, Mellanox/NVIDIA ConnectX adapters, or equivalent technologies
- Strong understanding of Layer 2 and Layer 3 infrastructure design and troubleshooting
Operations & Reliability
- Experience with high availability, clustering, and disaster recovery
- Strong troubleshooting skills across:
- Linux operating systems
- Hardware platforms
- GPU infrastructure
- Networking
- Enterprise storage
- Experience supporting mission-critical production environments
- Bash and Python scripting for automation and operational efficiency
- Experience creating operational documentation, runbooks, and infrastructure standards
Nice to Have
- Kubernetes infrastructure (especially AI/ML and GPU integration)
- KVM, VMware, OpenShift Virtualization, or similar virtualization platforms
- Ansible automation
- NVIDIA Base Command Manager
- Slurm or HPC workload schedulers
- Observability and monitoring platforms (Prometheus, Grafana, OpenTelemetry)
- Data Center Infrastructure Management (DCIM) tools
- IPAM solutions
- AWS, Azure, or hybrid cloud exposure
DevOps / Infrastructure Engineer
Location: Chennai
Experience: 5+ Years
Role: DevOps / Infrastructure Engineer
Job Description
We are looking for an experienced DevOps / Infrastructure Engineer with strong hands-on experience in Linux administration, containerization, Kubernetes, automation, monitoring, and troubleshooting.
Mandatory Skills
- 5+ years of experience in DevOps / Infrastructure Administration
- Strong hands-on experience with Linux Administration
- Experience with Docker and Kubernetes
- Monitoring tools: AppDynamics, Prometheus, Grafana
- Strong Shell Scripting / Python Scripting skills
- Hands-on experience with Ansible 4.1
- Strong troubleshooting and problem-solving skills
- Experience in infrastructure/application monitoring and production support
- Good understanding of DevOps practices and automation
Key Responsibilities
- Manage and support Linux-based infrastructure and production environments.
- Deploy, manage, and troubleshoot applications using Docker and Kubernetes.
- Develop and maintain automation scripts using Shell/Python.
- Automate infrastructure and configuration management using Ansible.
- Monitor applications and infrastructure using AppDynamics, Prometheus, and Grafana.
- Perform root-cause analysis and resolve infrastructure/application issues.
- Handle incidents, troubleshoot performance issues, and ensure system availability.
- Collaborate with development and operations teams to improve deployment and operational processes.
Job Description
Position Title: System Administrator
Employment Type: Full-Time, Permanent
Department: Research and Development
Reports To: R&D Manager, India Development Centre
Work from Office; No Remote; No Hybrid
About the Company
We are a software product and solutions company that builds mission-critical platforms used in regulated industries such as life sciences, biotech research, public health, and government compliance. Our systems are used to manage workflows, audits, research approvals, data security, compliance, and more—helping institutions reduce risk, increase transparency, and advance human well-being.
Our product line includes dynamic form engines, workflow automation, rules engines, analytics, AI/ML integrations, and user-centric digital tools for a global customer base.
Job Summary
We are looking for a skilled Windows & Linux System Administrator / Desktop Engineer responsible for managing, maintaining, monitoring, and troubleshooting the organization's IT infrastructure, including Windows and Linux servers, desktops, laptops, networks, applications, and end-user systems.
The role requires a combination of system administration, desktop support, troubleshooting, security, user management, and infrastructure monitoring. The candidate should be comfortable working independently, resolving technical issues, and providing timely support to employees.
Key Responsibilities
- Install, configure, maintain, and troubleshoot Windows and Linux servers.
- Manage Active Directory, DNS, DHCP, Group Policies, users, and access permissions.
- Perform server monitoring, patching, updates, backup checks, and basic security administration.
- Provide desktop and laptop support, including OS, software, hardware, printer, email, VPN, and connectivity troubleshooting.
- Troubleshoot basic LAN, Wi-Fi, TCP/IP, DNS, and VPN issues.
- Manage employee IT setup, onboarding/offboarding, and IT asset allocation.
- Monitor antivirus/endpoint security and ensure systems are updated and secure.
- Maintain IT inventory, system documentation, and helpdesk/ticket records.
- Coordinate with vendors and escalate infrastructure issues when required.
- Support virtualization, cloud, backup, and disaster-recovery activities as needed.
Required Skills
- Windows 10/11, Windows Server, Active Directory, DNS, DHCP, Group Policy
- Linux (Ubuntu/RHEL/CentOS) and basic Shell scripting
- Desktop/laptop hardware and software troubleshooting
- Basic networking: TCP/IP, LAN, Wi-Fi, VPN
- Backup, patch management, antivirus, and system monitoring
- Good troubleshooting, communication, and documentation skills
- Qualification
- Bachelor's degree/Diploma in Computer Science, IT, Electronics, or a related field.
- Preferred
- Exposure to VMware/Hyper-V, AWS/Azure, PowerShell/Bash, and IT helpdesk/ticketing tools.
Job Title : DevOps Engineer – Linux, AWS & Infrastructure
Experience : 3+ Years
Location : Sector 48, Gurgaon
Work Mode : 6 Days WFO – Monday to Saturday
Week Off : 01st & 03rd Saturday Off
Employment Type : Full-Time
Job Summary :
We are looking for a DevOps Engineer with strong hands-on experience in Linux, Networking, Server Administration, AWS, CI/CD, Containers, Kubernetes, and Infrastructure Automation. The ideal candidate should have strong troubleshooting skills, production ownership, and the ability to manage and automate infrastructure reliably.
Key Responsibilities :
- Manage and troubleshoot Linux servers, bare-metal infrastructure, and server environments.
- Perform system administration, networking, performance monitoring, and production troubleshooting.
- Design, maintain, and optimize Jenkins-based CI/CD pipelines and deployment workflows.
- Manage Docker containers and Kubernetes environments.
- Work with AWS Cloud services, infrastructure, security, and deployment environments.
- Monitor system health, application performance, logs, and infrastructure using appropriate monitoring and logging tools.
- Implement and maintain Infrastructure as Code (IaC) using tools such as Terraform or CloudFormation.
- Automate repetitive operational tasks using Python, Bash, Shell scripting, or similar technologies.
- Implement infrastructure and application security, access controls, patching, and hardening.
- Investigate production incidents, perform root-cause analysis (RCA), and drive issues to resolution.
- Take end-to-end ownership of infrastructure reliability, availability, and operational issues.
- Collaborate with development and other engineering teams to improve deployment, scalability, and system reliability.
Mandatory Skills :
Linux & Networking | Bare Metal / Server Administration | Jenkins / CI-CD | Docker / Containers | AWS Cloud | Kubernetes | Git | Monitoring & Logging | Security | Infrastructure as Code (IaC) | Automation / Scripting | Production Troubleshooting & Ownership
Preferred Skills :
- Strong understanding of TCP/IP, DNS, HTTP/HTTPS, SSH, load balancing, and networking fundamentals.
- Hands-on experience with Terraform / CloudFormation.
- Experience with Prometheus, Grafana, ELK / EFK, CloudWatch, or similar monitoring / logging tools.
- Good understanding of Linux performance troubleshooting, processes, memory, disk, networking, and file systems.
- Experience handling production incidents, RCA, deployments, and system reliability.
- Exposure to cloud security, IAM, secrets management, and server hardening.
What We’re Looking For :
- 3+ years of hands-on experience in DevOps / SRE / Infrastructure Engineering.
- Strong practical knowledge rather than certification-based/theoretical understanding.
- Good troubleshooting and analytical skills.
- Strong sense of ownership and accountability for production systems.
- Comfortable working in a 6-day work-from-office environment.
Primary Skills: Linux Administration, Production Support
Secondary Skills: Oracle SQL, Splunk, Grafana, AppDynamics, Cloud (OCP/AWS/Azure/GCP), Incident Management, AI/Automation.
• Vista Plus Development & Engineering
• Shell Scripting
• Application installation and upgrade in Linux environments (added advantage)
• File System Management
• Incident, Change, and Problem Management (ITIL)
• Basic Oracle / Database concepts
• Production Support and Health Check Monitoring
• Documentation and Operational Governance





