Linux Engineer (L2) at InfoGrowth private Limited · Bengaluru (Bangalore) · 5 - 7 years · ₹16L - ₹20L / yr · Profitable · Posted 4 Oct 2024

Join our dynamic team to provide advanced technical support for global HPE customers. We’re looking for a Linux Engineer with 5-7 years of experience who is passionate about solving complex technical issues and providing superior customer solutions. If you have a strong Linux background and thrive in a high-risk, mission-critical environment, this could be your next exciting career move!
Key Responsibilities:
- Provide technical support through phone, email, or remote sessions
- Troubleshoot and resolve issues related to Linux systems
- Perform proactive problem management and root cause analysis
- Manage escalations and provide leadership during crisis situations
- Apply patches, conduct security updates, and manage configurations
- Lead or participate in customer and internal projects, including Change Management, Incident Management, and Disaster Recovery planning
- Create and maintain Standard Operating Procedures (SOP) for smooth operations
- Collaborate with global teams to solve complex issues and deliver high-quality service to customers
- Maintain case documentation, meet SLA timeframes, and contribute to operational metrics
Technical Skills:
- Proficient in all flavors of Linux (installation, configuration, troubleshooting, administration)
- Experience with Redhat Satellite/Oracle Linux Ksplice
- Strong understanding of networking, Cluster Services, Oracle RAC clusters, SAN technologies, and HP Blade/Rackmount servers
- Knowledge in vulnerability assessment, security patching, and disaster recovery
- Experience in HP Service Guard cluster on RHEL
Non-Technical Skills:
- Strong written and verbal communication
- Excellent customer relationship management and teamwork
- Ability to lead technical escalations and crisis management
- Demonstrated commitment to delivering high-quality solutions
Eligibility:
- Bachelor’s degree in Engineering (or equivalent)
- Relevant certifications such as RHCE, Oracle Enterprise Linux, or ITIL preferred
- Flexibility to work in a 24x7 support environment

Similar jobs (10)

The recruiter has not been active on this job recently. You may apply but please expect a delayed response.
Job Description – Linux Administrator
Role: Linux Administrator
Experience: 8+Years
Location: chennai
Required Skills
Experience with Linux servers in virtualized environmentsExperience installing| configuring| and maintaining services such as Bind| Apache| nginx| etc.Familiarity with the fundamentals of Linux scripting languagesAbility to build and monitor services on production serversExperience with virtualization technologies| such as Nutanix or VMWareProficient with network tools such as iptables| firewalld| netstat| Experience on SSSD service for domain joiningMaintain and monitor all patch releases and design various patch installation strategies and maintain all systems according to NIST standardization.Proficient with network tools such as iptables| Linux IPVS| HAProxy| etcMonitor everyday systems and evaluate availability of all server resources and perform all activities for Linux servers.Nutanix AHV Administration Basic Ansible Scripting
Desirable Skills:
Keyword:
- Skills: RedHat Linux
We are looking for an experienced Application Support Engineer with strong expertise in Linux/Unix, Networking, Routing, Load Balancing, and Production Support. The ideal candidate should have hands-on experience troubleshooting application and infrastructure issues using enterprise monitoring and observability tools such as Splunk, Grafana, and AppDynamics.
Key Responsibilities
- Provide L2/L3 production support for enterprise applications and infrastructure.
- Troubleshoot application, Linux/Unix, network, routing, and connectivity-related issues.
- Monitor application and infrastructure health using Splunk, Grafana, and AppDynamics.
- Analyze alerts, logs, performance metrics, and system behavior to identify issues.
- Troubleshoot routing, load balancing, connectivity, and network-related problems.
- Perform incident investigation, troubleshooting, and root cause analysis.
- Coordinate with Network, Infrastructure, Application, and other technical teams for issue resolution.
- Handle production incidents and ensure timely resolution within defined SLAs.
- Participate in problem management and identify recurring issues.
- Perform application health checks and proactively identify potential failures.
- Maintain troubleshooting guides, runbooks, and operational documentation.
- Support planned changes, deployments, and maintenance activities.
Must-Have Skills
Application Support
Linux / Unix
Networking
Routing
Load Balancing
Splunk
Grafana
AppDynamics
Production Support
Incident Management
Troubleshooting
Required Skills & Tools
- 4–6 years of experience in Application / Production Support.
- Strong hands-on experience with Linux/Unix administration and troubleshooting.
- Good understanding of TCP/IP, networking, routing, and connectivity concepts.
- Practical experience troubleshooting load balancing issues.
- Hands-on experience with monitoring and observability tools such as Splunk, Grafana, or AppDynamics.
- Strong log analysis and troubleshooting skills.
- Experience handling production incidents and working within SLA-driven environments.
- Strong communication and coordination skills.
Preferred Experience
- Exposure to ITIL processes including Incident, Problem, and Change Management.
- Experience with scripting/automation using Shell or Python.
- Exposure to cloud environments is an added advantage.
- Experience supporting enterprise-scale applications and infrastructure.
Key Competencies
- Strong analytical and troubleshooting skills.
- Ability to work under pressure during critical production incidents.
- Strong ownership and accountability.
- Excellent communication and stakeholder management.
- Ability to work collaboratively with global technical teams.
Qualifications
Bachelor’s degree in Computer Science, Information Technology, Engineering, or a related discipline.
Job Overview
We are seeking a highly experienced Senior Linux Infrastructure Engineer with deep expertise in Linux administration, bare metal infrastructure, enterprise storage, and next-generation AI Factory / GPU infrastructure platforms. This role is focused on designing, deploying, operating, and troubleshooting large-scale Linux-based infrastructure that powers both traditional enterprise workloads and modern AI/ML environments.
This is not a DevOps-focused role. We already have a dedicated DevOps team and are looking for an engineer with extensive hands-on experience in Bare Metal as a Service (BMaaS), GPU infrastructure, high-performance storage, data center operations, and enterprise Linux platforms.
The ideal candidate will have experience building and managing infrastructure from the hardware layer up, including servers, networking, storage, GPU clusters, and AI-ready platforms. They should be comfortable working with high-performance computing (HPC), AI Factory environments, and large-scale Linux deployments where performance, reliability, and operational excellence are critical.
Key Responsibilities & Required Skills
Linux & Bare Metal Infrastructure
- Expert-level Linux administration (Ubuntu required; Red Hat and SUSE preferred)
- Deep expertise in bare metal server deployment, architecture, provisioning, and lifecycle management
- Experience operating Bare Metal as a Service (BMaaS) platforms and large-scale infrastructure environments
- Strong understanding of server hardware, including:
- BIOS/UEFI
- RAID controllers
- Firmware management
- iLO/iDRAC/IPMI
- NICs and SmartNICs
- HBA cards
- Hardware diagnostics and troubleshooting
- Experience designing, implementing, and supporting enterprise Linux infrastructure at scale
AI Factory & GPU Infrastructure
- Experience deploying and managing GPU-accelerated infrastructure for AI/ML workloads
- Understanding of NVIDIA GPU technologies including:
- A100, H100, H200, B200, or equivalent GPU platforms
- NVIDIA DGX and OEM GPU servers
- GPU provisioning and lifecycle management
- GPU monitoring and performance optimization
- Knowledge of AI Factory architecture and infrastructure requirements
- Experience supporting GPU clusters, AI training environments, and high-performance computing (HPC) workloads
- Understanding of:
- GPU resource allocation and scheduling
- Multi-GPU systems
- GPU networking requirements
- High-bandwidth, low-latency infrastructure design
- Familiarity with NVIDIA ecosystem technologies such as:
- CUDA
- NCCL
- GPUDirect Storage
- NVIDIA Fabric Manager
- NVIDIA Base Command (preferred)
Enterprise Storage & Data Platforms
- Advanced Linux storage administration:
- LVM
- XFS, EXT4
- NFS
- iSCSI
- Fibre Channel SAN
- Multipath I/O
- Strong hands-on experience with Ceph, including:
- Cluster architecture
- MON, OSD, MDS
- RBD, CephFS, RGW
- Capacity planning
- Performance tuning
- Failure recovery
- Experience with high-performance AI storage platforms such as:
- WEKA
- VAST Data
- Dell PowerScale
- Pure Storage FlashBlade
- NetApp
- Understanding of:
- NVMe-over-Fabrics (NVMe-oF)
- RDMA
- GPUDirect Storage
- Parallel file systems
- AI data pipelines
Networking & Infrastructure
- Strong networking knowledge:
- Bonding
- VLANs
- Routing
- MTU optimization
- DNS
- DHCP
- Experience with high-performance data center networking:
- 100G/200G/400G Ethernet
- RoCE
- RDMA
- Spine-Leaf architectures
- Familiarity with NVIDIA Spectrum-X, Mellanox/NVIDIA ConnectX adapters, or equivalent technologies
- Strong understanding of Layer 2 and Layer 3 infrastructure design and troubleshooting
Operations & Reliability
- Experience with high availability, clustering, and disaster recovery
- Strong troubleshooting skills across:
- Linux operating systems
- Hardware platforms
- GPU infrastructure
- Networking
- Enterprise storage
- Experience supporting mission-critical production environments
- Bash and Python scripting for automation and operational efficiency
- Experience creating operational documentation, runbooks, and infrastructure standards
Nice to Have
- Kubernetes infrastructure (especially AI/ML and GPU integration)
- KVM, VMware, OpenShift Virtualization, or similar virtualization platforms
- Ansible automation
- NVIDIA Base Command Manager
- Slurm or HPC workload schedulers
- Observability and monitoring platforms (Prometheus, Grafana, OpenTelemetry)
- Data Center Infrastructure Management (DCIM) tools
- IPAM solutions
- AWS, Azure, or hybrid cloud exposure
Job Description
Position Title: System Administrator
Employment Type: Full-Time, Permanent
Department: Research and Development
Reports To: R&D Manager, India Development Centre
Work from Office; No Remote; No Hybrid
About the Company
We are a software product and solutions company that builds mission-critical platforms used in regulated industries such as life sciences, biotech research, public health, and government compliance. Our systems are used to manage workflows, audits, research approvals, data security, compliance, and more—helping institutions reduce risk, increase transparency, and advance human well-being.
Our product line includes dynamic form engines, workflow automation, rules engines, analytics, AI/ML integrations, and user-centric digital tools for a global customer base.
Job Summary
We are looking for a skilled Windows & Linux System Administrator / Desktop Engineer responsible for managing, maintaining, monitoring, and troubleshooting the organization's IT infrastructure, including Windows and Linux servers, desktops, laptops, networks, applications, and end-user systems.
The role requires a combination of system administration, desktop support, troubleshooting, security, user management, and infrastructure monitoring. The candidate should be comfortable working independently, resolving technical issues, and providing timely support to employees.
Key Responsibilities
- Install, configure, maintain, and troubleshoot Windows and Linux servers.
- Manage Active Directory, DNS, DHCP, Group Policies, users, and access permissions.
- Perform server monitoring, patching, updates, backup checks, and basic security administration.
- Provide desktop and laptop support, including OS, software, hardware, printer, email, VPN, and connectivity troubleshooting.
- Troubleshoot basic LAN, Wi-Fi, TCP/IP, DNS, and VPN issues.
- Manage employee IT setup, onboarding/offboarding, and IT asset allocation.
- Monitor antivirus/endpoint security and ensure systems are updated and secure.
- Maintain IT inventory, system documentation, and helpdesk/ticket records.
- Coordinate with vendors and escalate infrastructure issues when required.
- Support virtualization, cloud, backup, and disaster-recovery activities as needed.
Required Skills
- Windows 10/11, Windows Server, Active Directory, DNS, DHCP, Group Policy
- Linux (Ubuntu/RHEL/CentOS) and basic Shell scripting
- Desktop/laptop hardware and software troubleshooting
- Basic networking: TCP/IP, LAN, Wi-Fi, VPN
- Backup, patch management, antivirus, and system monitoring
- Good troubleshooting, communication, and documentation skills
- Qualification
- Bachelor's degree/Diploma in Computer Science, IT, Electronics, or a related field.
- Preferred
- Exposure to VMware/Hyper-V, AWS/Azure, PowerShell/Bash, and IT helpdesk/ticketing tools.
We are seeking a skilled System Administrator and DevOps Engineer to join AmpereHour Energy, where you will play a crucial role in making sure our software systems run without downtime and our infrastructure scales well. This position involves working on cloud and physical servers across multiple operating systems and fundamental understanding of computers and networking.
At AmpereHour Energy, we are dedicated to advancing sustainable energy solutions. Our team thrives on collaboration and innovation, driving impactful projects that contribute to a greener future. We value expertise and creativity, making it an exciting place for tech professionals to grow.
If you are interested in this opportunity and meet the qualifications outlined in the job description, we encourage you to apply and explore how you can contribute to our mission.
Hiring Platform Engineer
Exp: 6 -- 10 yrs
Edu : BE/B.tech/MCA
Work Location : Pune
Skills :
Platform monitoring ,Incident trouble shooting, Incident recovery, openshift ,kubernetes.
2 years of IT operations, infrastructure, cloud or application support experience.
Exp in Linux command-line knowledge.
Exp in networking knowledge including IP addressing, DNS, ports and connectivity troubleshooting.
- Linux troubleshooting
- Hands-on AWS
- Production/Application Support
- Bash/Shell/Python
- Monitoring/log analysis
- Incident resolution
- Application deployment/support
- Basic networking and database knowledge
- Production/batch support exposure
- Willingness for rotational weekend/critical production support
We're looking for a Site Reliability Engineer to keep our production systems fast, reliable, and scalable. Sitting at the intersection of software engineering and operations, you'll treat infrastructure as code, automate away toil, and build the observability that lets us catch problems before customers do. You'll own uptime and on-call for critical services, lead incident response and blameless postmortems, and continuously harden the platform against failure. This role suits an engineer who is as comfortable debugging a production incident at 2 a.m. as they are writing the automation that prevents the next one.
Key Responsibilities
- Own reliability, availability, and performance of production services, including on-call rotation
- Build and maintain monitoring, alerting, and observability (metrics, logs, traces)
- Automate deployments, scaling, and operational tasks to reduce manual toil
- Manage containerized workloads on Kubernetes and cloud infrastructure
- Design and maintain CI/CD pipelines for safe, frequent releases
- Lead incident response and drive blameless postmortems with clear follow-ups
- Perform capacity planning, performance tuning, and cost optimization
- Define and track SLIs/SLOs and error budgets with product teams
Requirements
- 3+ years in SRE, DevOps, or production-focused engineering
- Strong Linux administration and hands-on Kubernetes experience
- Solid experience with monitoring/observability tools (Prometheus, Grafana, ELK, or similar)
- Cloud experience with AWS, GCP, or Azure
- CI/CD pipelines and infrastructure-as-code (Terraform, CloudFormation)
- Proficient scripting in Python and/or Bash
Nice to have
- Experience with service meshes, Helm, or GitOps (ArgoCD/Flux)
- Background in high-traffic or distributed systems
Location : Mumbai
Big Picture (The Opportunity) :
Are you looking for an opportunity to advance your Career? & If you are able to maintain a positive attitude even when everything goes wrong, if you are detail oriented and self motivated with a passion to learn and improve your skills and knowledge, we have a perfect job for you !
What do we want from you ? (Our Expectations) :
- Zero to 2 years experience in Linux Operating System.
- Flexible working hours - able to support occasional nights, weekends, and call-ins, able to quickly adapt to a constantly changing faced paced environment. Open to travel to sites.
- An ideal person who is excited and motivated about running and supporting a production - grade critical infrastructure and looks for opportunities to improve processes with automation.
What are you required to do ? (Your Responsibilities) :
- Proactively maintain and develop all linux infrastructure technology to maintain a 24*7*365 uptime service.
- Engineering of systems administration-related various solutions for our various SAAS products and projects as well as operational needs.
- Proactively monitoring system performance and capacity planning.
- Providing technical support to customers for Applications, Operating systems, and networking.
- Will also be the first point of contact for our clients, where installations are placed on a permanent basis, for basic troubleshooting and problem solving.
- Fault finding, analysis and logging information for reporting of performance exceptions.
- Maintain best practices on managing systems and services across all environments.
Skills & Qualification Required (Add Value) :
- Graduate - Preferred to have Bachelor‘s Degree in Engineering, Computer Science or related field.
- He / She should be familiar with the installation and configuration of Linux operating systems and setup and operation of TCP/IP networking on Linux systems also familiar with Internet concepts including SMTP, IMAP, POP, HTTP, DNS, LDAP and related protocols.
- You should possess excellent communication skills .
- Knowledge of Email concepts, Helpdesk Concept , VoIP Concept , Cloud computing will be an added advantage.
This is a Red Hat OpenShift / Kubernetes Platform Engineer role, mainly focused on OpenShift cluster administration, installation, upgrades, troubleshooting, security certificates, and BAU operations.
Key responsibilities:
- Install and configure OpenShift clusters on bare metal, on-prem VMware/virtual environments, and cloud platforms.
- Strong hands-on experience with Red Hat OpenShift 3.x/4.x.
- Administer OpenShift Container Platform (OCP) and Kubernetes environments.
- Perform cluster installation, upgrades, health checks, certificate renewal, and troubleshooting.
- Manage OpenShift projects, users, roles, access, and day-to-day platform activities.
- Monitor cluster performance, identify issues, and maintain high availability.
- Perform cluster scaling and performance optimization.
- Provide production/BAU support for OpenShift environments.





