5+ Monitoring Jobs in Chennai | Monitoring Job openings in Chennai
Apply to 5+ Monitoring Jobs in Chennai on CutShort.io. Explore the latest Monitoring Job opportunities across top companies like Google, Amazon & Adobe.
Role Overview
We are seeking a technically strong Java Support Engineer who combines solid development knowledge with a passion for support and operational excellence. The ideal candidate should have hands-on experience in Java, Spring Boot, and Angular, along with a strong understanding of application engineering concepts, and must be comfortable working in a production support environment handling incidents, troubleshooting, monitoring, and system stability.
Work Location: Hyderabad or Chennai
Key Responsibilities
· Provide L2/L3 production support for enterprise applications.
· Troubleshoot, debug, and resolve application issues within defined SLAs.
· Analyze logs, identify root causes, and implement fixes or workarounds.
· Collaborate with development teams for permanent issue resolution.
· Monitor application health, performance, and availability.
· Support deployments, releases, and environment validations.
· Perform minor code fixes and enhancements when required.
· Document issues, solutions, and support procedures.
· Participate in on-call rotations and handle incident management.
Required Skills & Qualifications
· Strong hands-on experience in Java and Spring Boot.
· Working knowledge of Angular for frontend understanding.
· Good understanding of application architecture, APIs, microservices, and debugging techniques.
· Experience with log analysis tools, monitoring tools, and ticketing systems.
· Knowledge of SQL databases and query troubleshooting.
· Familiarity with Linux/Unix environments.
· Understanding of CI/CD, release processes, and version control (Git).
· Strong analytical, problem-solving, and communication skills
Role & Responsibilities:
AWS Cloud Management:
• Lead the design, deployment, and management of AWS cloud infrastructure to ensure
scalability, security, and reliability.
• Oversee the implementation of best practices for cloud resource utilization.
Automated Provisioning:
• Drive the development and maintenance of automated provisioning processes for
infrastructure deployment, leveraging tools such as Terraform and Packer.
• Continuously enhance deployment workflows to optimize efficiency.
Financial Operations (FinOps):
• Implement and champion FinOps practices to optimize cloud costs and resource
utilization.
• Conduct regular cost analysis and identify opportunities for cost savings without
compromising performance.
Infrastructure as Code (IaC):
• Collaborate with teams to implement and maintain IaC scripts for infrastructure
configuration and deployment.
• Ensure version control and consistency in infrastructure code across projects.
Team Leadership:
• Lead and mentor a team of DevOps engineers, providing technical guidance and
support.\
• Foster a collaborative and innovative team culture focused on continuous improvement.
Continuous Integration/Continuous Deployment (CI/CD):
• Drive the implementation and maintenance of CI/CD pipelines to automate software
delivery processes.
• Ensure seamless and reliable application deployments across environments.
Monitoring and Optimization:
• Implement monitoring solutions for cloud resources and applications.
• Proactively identify and address performance bottlenecks, ensuring optimal system
performance.
JOB Description
- Monitoring entire infrastructure of Olam using various monitoring tools like
- SCOM, SolarWinds, Telegraph, OEM.
- Monitoring various types of alerts like
- CPU Utilization
- Memory Utilization
- Database related alerts
- DR Replication issues
- Backup Failure Alerts
- Exchange Mail Queue Threshold Alerts
- Service Mailbox quota breach alert
- Adobe Experience Manager / Site 24/7 Alerts
- Application URL Alerting
- Scheduling Maintenance Mode for planned Activity.
- Daily repeat CI analysis of events/alerts/incident and raising proactive problem tickets which helps in reduction of major incident.
- Handling Major Incidents, Driving the major incident bridge, sending communication about major incident to stake holders.
- CMDB Inventory Management – Onboarding and Offboarding of Device's are commissioned/decommissioned.
- Coordinating with Service Provider for MPLS related outage
- Daily follow ups with Regional and internal teams to ensure all the node are up and running fine.
This role is for Work from the office.
Job Description
Roles & Responsibilities
- Work across the entire landscape that spans network, compute, storage, databases, applications, and business domain
- Use the Big Data and AI-driven features of vuSmartMaps to provide solutions that will enable customers to improve the end-user experience for their applications
- Create detailed designs, solutions and validate with internal engineering and customer teams, and establish a good network of relationships with customers and experts
- Understand the application architecture and transaction-level workflow to identify touchpoints and metrics to be monitored and analyzed
- Analytics and analysis of data and provide insights and recommendations
- Constantly stay ahead in communicating with customers. Manage planning and execution of platform implementation at customer sites.
- Work with the product team in developing new features, identifying solution gaps, etc.
- Interest and aptitude in learning new technologies - Big Data, no SQL databases, Elastic Search, Mongo DB, DevOps.
Skills & Experience
- At least 2+ years of experience in IT Infrastructure Management
- Experience in working with large-scale IT infra, including applications, databases, and networks.
- Experience in working with monitoring tools, automation tools
- Hands-on experience in Linux and scripting.
- Knowledge/Experience in the following technologies will be an added plus: ElasticSearch, Kafka, Docker Containers, MongoDB, Big Data, SQL databases, ELK stack, REST APIs, web services, and JMX.


