Infrastructure Stream Lead at DevOpspatial Pvt Ltd · Remote only · 10 - 20 years · ₹16L - ₹27L / yr · Profitable · Remote only · Posted 2 Aug 2024

- On-Premise Infrastructure Management: Utilize 10-20 years of experience to oversee and optimize on-premise infrastructure, ensuring stability, scalability, and security.
- Cloud Infrastructure Management: Lead the implementation and management of AWS and Azure cloud environments, leveraging best practices and ensuring adherence to security standards. Must be cloud certified and proficient in managing PaaS services.
- Dynatrace Monitoring: Utilize Dynatrace expertise to implement and maintain effective monitoring solutions, ensuring proactive identification and resolution of performance issues.
- Technical Expertise: Demonstrate proficiency in a variety of technologies including Oracle, SQL Server, Shareplex for Oracle replication, Commvault, Windows, Linux, Solaris, SNMP Polling and Traps, Unix Extension, and more.
- Analytical Skills: Conduct complex analysis to optimize bulk metrics and alert profiles, ensuring efficient resource utilization and timely response to critical incidents.
- Cloud Integration: Facilitate inter-AWS account integration using Terraform IAM roles and integrate with other products such as Splunk, APIM, and VMWare to enhance overall infrastructure capabilities.
- Experience working with On-Premise Infrastructure
- Experience with AWS & Azure and Cloud Certified.
- Dynatrace Experience & Professional Certification

About DevOpspatial Pvt Ltd
About
Similar jobs (10)
Cloud Infrastructure Engineer – BANG | 10+ Years
Location: Bangalore
Experience: 10+ Years
Job Description:
- Design, implement, and manage cloud infrastructure across AWS/Azure/GCP environments.
- Strong experience in cloud architecture, compute, storage, networking, and security.
- Manage VMs, VPC/VNet, load balancers, DNS, DHCP, firewalls, and IAM.
- Hands-on experience with Windows/Linux servers, VMware, virtualization, and infrastructure operations.
- Automate infrastructure provisioning and configuration using Terraform, Ansible, or similar tools.
- Monitor infrastructure performance, availability, and capacity using tools such as Grafana, Prometheus, or CloudWatch/Azure Monitor.
- Handle incident management, troubleshooting, disaster recovery, backup, and high-availability requirements.
- Work with cross-functional teams to support cloud migration, infrastructure upgrades, and production environments.
- Ensure infrastructure follows security, compliance, and operational best practices.
Must-Have Skills:
Cloud Infrastructure | AWS/Azure/GCP | Networking | Linux/Windows | VMware | Terraform | Ansible | DNS/DHCP | IAM | Monitoring | Backup & DR
Primary Skills: Linux Administration, Production Support
Secondary Skills: Oracle SQL, Splunk, Grafana, AppDynamics, Cloud (OCP/AWS/Azure/GCP), Incident Management, AI/Automation.
Work Mode: WORK FROM OFFICE
Interview Mode: Virtual
Role Descriptions:
Exp Range: 5 - 8 years
City Locations: Bengaluru
*Key Responsibilities*
Devops Developer (Azure)
Azure L2 Support | Specialist | Infrastructure Operations-260006F2- Certificate & DNS Management- Monitoring & Observability- Azure Operations & Compliance- Migration: Open Cloud TnC- Reliability & Operations Excellence (OPS)
Profile – 5 to 8+ years in Cloud/Infrastructure Operations- Strong hands-on Azure skills- Certificates/PKI experience- DNS management- Grafana| Health Checks| Azure Monitor- IaC automation (Terraform/Bicep)- Networking fundamentals Nice-to-Have- Azure security and compliance- AKS monitoring basics- FinOps practices- French reading proficiency- Certifications: AZ-104| AZ-305| etc.
Desire candidate
- Candidate should have valid PF.
Job Description
We are looking for an Azure Cloud & Observability Engineer with strong experience in Azure infrastructure and enterprise monitoring tools such as Splunk, Grafana, and AppDynamics.
Responsibilities
- Design, deploy, and manage Azure cloud infrastructure and services.
- Monitor application and infrastructure performance using Splunk, Grafana, and AppDynamics.
- Configure dashboards, alerts, health rules, and monitoring metrics.
- Perform log analysis, troubleshooting, and root-cause analysis for production issues.
- Implement observability solutions for applications, cloud infrastructure, and services.
- Automate monitoring and operational activities using scripting.
- Support incident, problem, and change management processes.
- Collaborate with development, DevOps, and SRE teams to improve system reliability.
- Maintain monitoring standards, documentation, and operational procedures.
Primary Skills
- Microsoft Azure
- Splunk
- Grafana
- AppDynamics
- Cloud Monitoring & Observability
- Application Performance Monitoring (APM)
- Log Analysis & Troubleshooting
Secondary Skills
- Azure Monitor / Log Analytics
- Azure VMs, Storage, Networking
- Linux
- Python / PowerShell / Shell Scripting
- CI/CD
- Git
- ITIL / ServiceNow
Job Description
We are looking for an Azure Cloud & Observability Engineer with strong experience in Azure infrastructure and enterprise monitoring tools such as Splunk, Grafana, and AppDynamics.
Responsibilities
- Design, deploy, and manage Azure cloud infrastructure and services.
- Monitor application and infrastructure performance using Splunk, Grafana, and AppDynamics.
- Configure dashboards, alerts, health rules, and monitoring metrics.
- Perform log analysis, troubleshooting, and root-cause analysis for production issues.
- Implement observability solutions for applications, cloud infrastructure, and services.
- Automate monitoring and operational activities using scripting.
- Support incident, problem, and change management processes.
- Collaborate with development, DevOps, and SRE teams to improve system reliability.
- Maintain monitoring standards, documentation, and operational procedures.
Primary Skills
- Microsoft Azure
- Splunk
- Grafana
- AppDynamics
- Cloud Monitoring & Observability
- Application Performance Monitoring (APM)
- Log Analysis & Troubleshooting
Secondary Skills
- Azure Monitor / Log Analytics
- Azure VMs, Storage, Networking
- Linux
- Python / PowerShell / Shell Scripting
- CI/CD
- Git
- ITIL / ServiceNow
Hiring: DevOps Lead
📍 Kochi / Trivandrum / Remote
💼 Full-time
🕐 General Shift | Australian Overlap
We are looking for an experienced DevOps Lead to join our team.
🔹 Key Responsibilities
Implement and continually improve the observability platform using Datadog, particularly from a user experience perspective.
Work across engineering squads as a virtual team member, supporting their DevOps, infrastructure, and observability requirements.
Configure and manage Datadog RUM, Session Replay, Distributed Tracing, and APM.
Support campaign readiness activities, including load and performance testing.
Participate in gamedays and incident response activities for production systems.
Liaise closely with the managed infrastructure provider on infrastructure requirements and activities.
🔹 Essential Skills & Requirements
✅ 7+ years of relevant DevOps / Cloud / Observability experience
✅ Strong hands-on experience with Datadog
✅ Strong experience with RUM, Session Replay, Distributed Tracing, and APM
✅ Solid experience with AWS
✅ Strong Infrastructure-as-Code experience using CloudFormation and AWS CDK
✅ Good working knowledge of GitHub Actions and AWS CodePipeline
✅ Real incident response experience on high-traffic systems
✅ Experience with Load & Performance Testing
✅ Strong problem-solving and communication skills
✅ Ability to work closely with infrastructure partners and internal platform teams
🔹 Skills - Good to Have
⭐ Experience with high-traffic platforms
⭐ Experience with campaign readiness and gamedays
⭐ Infrastructure partner management
⭐ Advanced AWS observability
⭐ Performance engineering
📌 Experience: 7+ Years
📌 Work Location: Kochi / Trivandrum / Remote
📌 Shift: General Shift with Australian Overlap
📌 Remote: Mandatory 1 week at office
📩 Interested candidates can share their resume

Job Title: TechOps Engineer
Location: Bengaluru, India (Hybrid)
Employment Type: Full-time
Experience: 6 Month-2 years (Excluding Internship)
Shift Timing: 2 PM to 11 PM IST
Role Overview
We are excited to find a highly engaged engineer who is obsessed with technology that wants to be a part of a “world class” platform SRE team. Engineers must possess an "automation first" mindset, with a relentless focus on documentation, quality, scalability, and reliability using Infrastructure as Code tools. This position will be part of a platform team that is developing exciting products and solutions and playing a key part in driving forward the electrification of transportation.
What you’ll do:
- Ensure system reliability, uptime, and performance of global platform.
- Conduct real-time surveillance of our EV charging systems to proactively identify and mitigate performance issues and anomalies near 24/7 basis. As such, you collaborate with IDT and FMC players to ensure incident detection also happens outside office hours (monitoring shifts among team members subject to duty schedule)
- Deliver on change & releases like firmware changes and drive insights & intelligence back into testing processes and tech discussions with the wider organization.
- Successfully deliver and project manage first time right commissioning activities alongside our Engineering Procurement Contract Management (EPCM) partners to successfully bring charge points onto our Charge Point Management System (CPMS).
- End-to-end EV charger lifecycle management, including deployment, commissioning, monitoring, maintenance, and decommissioning activities.
- Provide technical guidance and support to DC specialists during the commissioning of EV charging solutions.
- Work closely with Shell, Engineering, and IT colleagues to ensure projects are completed on time and to specification.
- Act as a liaison with the Engineering Procurement Contract Management (EPCM) partner to manage projects from start to finish, ensuring charge points are successfully onboarded on the Charge Point Management System (CPMS).
- Collaborate with development, operations and support teams to build scalable and resilient systems.
- Contribute to incident response, root-cause analysis, and post-mortem reviews, driving continuous improvement.
- Participate in capacity planning, performance tuning, and resource optimization.
- Integrate security and compliance best practices into all infrastructure operations.
- Stay current with emerging SRE tools, frameworks, and cloud technologies to continuously improve reliability practices.
- Participate in and lead on-call rotations and incident response, conducting detailed postmortems and RCA reports.
- Flexible to resolve blocking issues during off hours or weekends if required.
What We’re Looking For:
Basic Qualifications and skills
- Bachelor’s degree in Engineering , Electrical, ECE, Computer Science, Information Technology, or related field.
- Overall 1 years of experience as a Site Reliability Engineer, Technical project coordinator role.
- Proven experience of SRE or Technical Project Coordination with IoT or connected devices based platforms.
- Experience with incident management and on-call best practices. Provide support to on call engineers.
- Excellent analytical and problem-solving skills with a proactive mindset.
- Hands-on experience with AWS Cloud and IaC tools such as Terraform or Ansible.
- Expertise with monitoring and observability tools (Dynatrace,Prometheus, Grafana, Zabbix, etc.).
- Proactively monitor the network, triage performance outliers, and coordinate correction actions to ensure optimal system functionality.
- Fluency in English (spoken and written).
- Successfully recommission or decommission chargers following changes in our network.
- Responsible for the go-live of the chargers on Shell’s public network following commissioning attempts.
Note: This role involves managing infrastructure for a global platform operating in over ten countries, requiring effective communication and collaboration across regions. Strong verbal and written communication skills, along with availability and flexibility to resolve blocking issues, are essential to support On-call Engineers. This role may involve EU or US time‑zone shifts based on business requirements. The shift timing will be 2 PM IST to 11 PM IST.
What We Offer
- Work with some of the brightest minds in the emerging EV industry.
- Make a tangible impact in reducing carbon emissions and enabling sustainable energy.
- Freedom to suggest, implement, and innovate on systems, processes, and technologies.
- Daily ownership in a high-growth, challenging environment.
- Flexible work environment with hybrid schedules and virtualization options.
- Competitive pay and benefits including health coverage, innovative PTO program, and performance bonuses.
Job Summary
We are looking for an experienced AWS Cloud Engineer with strong expertise in AWS infrastructure, deployment, migration, and cloud operations. The candidate should have hands-on experience managing AWS services such as EC2, EBS, S3, EFS, and FSx, along with infrastructure provisioning, migration, troubleshooting, and optimization.
Key Responsibilities
- Design, deploy, configure, and manage AWS infrastructure environments.
- Perform application and infrastructure migration to AWS.
- Provision and manage EC2 instances, including configuration, scaling, patching, and troubleshooting.
- Manage EBS volumes, snapshots, backups, and storage performance.
- Configure and administer S3 buckets, storage policies, lifecycle management, and access controls.
- Manage EFS for scalable shared file storage.
- Implement and manage Amazon FSx file systems based on application requirements.
- Monitor AWS infrastructure performance, availability, and capacity.
- Troubleshoot infrastructure, networking, storage, and deployment-related issues.
- Implement security best practices including IAM, security groups, encryption, and access controls.
- Support backup, disaster recovery, high availability, and business continuity requirements.
- Optimize AWS resources for performance, scalability, reliability, and cost.
Mandatory Skills
- Strong hands-on experience in AWS Cloud Infrastructure.
- Expertise in EC2, EBS, S3, EFS, and FSx.
- Experience in AWS deployment and migration projects.
- Strong knowledge of AWS networking concepts such as VPC, Subnets, Route Tables, Security Groups, and Load Balancers.
- Experience with IAM and AWS security best practices.
- Good knowledge of AWS monitoring and troubleshooting.
- Experience with cloud infrastructure automation using Terraform or CloudFormation is preferred.
- Strong Linux administration and troubleshooting skills.
- Good understanding of backup, disaster recovery, and high-availability concepts
We are looking for a hands-on Senior AWS Cloud Engineer to lead the infrastructure build, optimization, automation, and production deployment of a Multi-Agent AI Chatbot Platform hosted on AWS. The development environment is already in place, and the successful candidate will drive the solution through testing, integrations, and production go-live.
Key Responsibilities
- Review, validate, and optimize existing Terraform code and AWS infrastructure.
- Establish and manage integrations with enterprise platforms such as ServiceNow, Workday, and other third-party systems.
- Design, build, and support secure, scalable, and highly available AWS environments.
- Implement and automate CI/CD pipelines and Infrastructure-as-Code practices.
- Lead infrastructure testing, performance tuning, and production readiness activities.
- Drive deployment and operationalization of the platform in the Production environment.
- Implement cloud governance, security, monitoring, and FinOps best practices.
- Troubleshoot and resolve complex cloud infrastructure issues.
Required Skills & Experience
- 10+ years of IT experience with strong expertise in AWS Cloud Engineering.
- Proven experience designing, deploying, and managing AWS production environments.
- Strong hands-on experience with Terraform and Infrastructure-as-Code.
- Experience with CI/CD pipeline automation and DevOps practices.
- Expertise in AWS services including VPC, IAM, EC2, S3, Lambda, CloudWatch, and networking.
- Experience in performance optimization, reliability, and cloud cost management (FinOps).
- Strong scripting and automation skills.
- Experience integrating enterprise applications through APIs and secure connectivity patterns.
Device42 Administrator
We are looking for an experienced Device42 Administrator with strong expertise in Enterprise CMDB, Infrastructure Discovery, Asset Management, Application Dependency Mapping, and Cloud/On-Premises environments.
Role: Device42 Administrator
Experience: 8–12 Years
Location: Kolkata / Chennai / Hyderabad / Pune / Delhi
Work Mode: Onsite
Contract: 6 Months
Budget: ₹1.5–2 LPM Max
Openings: 5
Key Skills Required:
• Strong hands-on experience in Device42 Administration
• Enterprise CMDB & Infrastructure Discovery
• Device42 Auto-Discovery
• CMDB & Asset Management
• Application Dependency Mapping
• On-Premises & Cloud Infrastructure Discovery
• Device42 Monitoring Rules & Alerts
• ServiceNow CMDB Integration
• APM / Network / Infrastructure Monitoring integrations
• Reporting, Dashboards & Data Governance
• Backup, Security & Compliance
Good to Have:
• ServiceNow Configuration Management & CMDB
• ITIL v4 Foundation
• Incident / Problem / Change Management
• Enterprise ITSM & Monitoring integrations
Qualification:
Bachelor's degree in Computer Science / IT / Engineering or related field.





