Application Support at MNC · Hyderabad · 7 - 11 years · ₹14L - ₹20L / yr · Posted 29 Sep 2026

Job Description
We are looking for an experienced Application Support Engineer with strong expertise in production support, networking, DNS, Linux/Unix, and application monitoring.
Key Responsibilities
- Provide application and production support for business-critical applications.
- Monitor application performance, availability, and system health.
- Troubleshoot issues related to Network, DNS, Linux, and Unix.
- Analyze application and system logs to identify and resolve production issues.
- Monitor applications using Splunk, Grafana, and AppDynamics.
- Identify and resolve incidents within defined timelines.
- Perform root cause analysis and support issue resolution.
- Coordinate with technical teams for incident investigation and escalation.
- Conduct application health checks and proactively identify potential issues.
- Maintain proper documentation of incidents, troubleshooting steps, and resolutions.
- Follow standard production support and incident management processes.
Required Skills
- 7–11 years of experience in Application/Production Support.
- Strong knowledge of Networking and DNS.
- Hands-on experience with Linux/Unix.
- Experience with monitoring tools such as Splunk, Grafana, and AppDynamics.
- Strong troubleshooting and analytical skills.
- Good understanding of incident management and production support.
- Excellent communication and coordination skills.
- Ability to work in a fast-paced production support environment.

Similar jobs (10)
We are looking for an experienced Application Support Engineer with strong knowledge of Network, DNS, Linux/Unix, and application monitoring tools. The candidate will be responsible for production support, incident troubleshooting, monitoring, and ensuring application availability and stability.
Responsibilities:
- Provide L2/L3 application and production support for critical applications.
- Troubleshoot issues related to Network, DNS, Linux/Unix, and application connectivity.
- Monitor application health, performance, and availability using Splunk, Grafana, AppDynamics, or similar tools.
- Analyze logs, alerts, and performance metrics to identify and resolve incidents.
- Handle incident, problem, and change management activities as per ITIL processes.
- Perform root cause analysis (RCA) and implement preventive measures.
- Coordinate with Network, Infrastructure, Development, and other support teams for issue resolution.
- Participate in production deployments, maintenance activities, and on-call support.
Required Skills:
- Strong experience in Application/Production Support.
- Good knowledge of Linux/Unix administration and troubleshooting.
- Strong understanding of Networking concepts and DNS.
- Hands-on experience with Splunk, Grafana, AppDynamics, or similar monitoring tools.
- Good understanding of TCP/IP, HTTP/HTTPS, load balancing, and connectivity troubleshooting.
- Experience with incident, problem, and change management.
- Strong troubleshooting, communication, and stakeholder management skills.
- Good knowledge of SQL and application log analysis is preferred.
We are looking for an experienced Application Support Engineer with strong expertise in Linux/Unix, Networking, Routing, Load Balancing, and Production Support. The ideal candidate should have hands-on experience troubleshooting application and infrastructure issues using enterprise monitoring and observability tools such as Splunk, Grafana, and AppDynamics.
Key Responsibilities
- Provide L2/L3 production support for enterprise applications and infrastructure.
- Troubleshoot application, Linux/Unix, network, routing, and connectivity-related issues.
- Monitor application and infrastructure health using Splunk, Grafana, and AppDynamics.
- Analyze alerts, logs, performance metrics, and system behavior to identify issues.
- Troubleshoot routing, load balancing, connectivity, and network-related problems.
- Perform incident investigation, troubleshooting, and root cause analysis.
- Coordinate with Network, Infrastructure, Application, and other technical teams for issue resolution.
- Handle production incidents and ensure timely resolution within defined SLAs.
- Participate in problem management and identify recurring issues.
- Perform application health checks and proactively identify potential failures.
- Maintain troubleshooting guides, runbooks, and operational documentation.
- Support planned changes, deployments, and maintenance activities.
Must-Have Skills
Application Support
Linux / Unix
Networking
Routing
Load Balancing
Splunk
Grafana
AppDynamics
Production Support
Incident Management
Troubleshooting
Required Skills & Tools
- 4–6 years of experience in Application / Production Support.
- Strong hands-on experience with Linux/Unix administration and troubleshooting.
- Good understanding of TCP/IP, networking, routing, and connectivity concepts.
- Practical experience troubleshooting load balancing issues.
- Hands-on experience with monitoring and observability tools such as Splunk, Grafana, or AppDynamics.
- Strong log analysis and troubleshooting skills.
- Experience handling production incidents and working within SLA-driven environments.
- Strong communication and coordination skills.
Preferred Experience
- Exposure to ITIL processes including Incident, Problem, and Change Management.
- Experience with scripting/automation using Shell or Python.
- Exposure to cloud environments is an added advantage.
- Experience supporting enterprise-scale applications and infrastructure.
Key Competencies
- Strong analytical and troubleshooting skills.
- Ability to work under pressure during critical production incidents.
- Strong ownership and accountability.
- Excellent communication and stakeholder management.
- Ability to work collaboratively with global technical teams.
Qualifications
Bachelor’s degree in Computer Science, Information Technology, Engineering, or a related discipline.
Application Production Support with SRE, Linux/Unix, Splunk/AppD/Grafana, Troubleshooting
WFO-Immediate
8 to 12 Yrs
Bangalore/Hyderabad
Role Summary:
We are looking for an experienced Application Production Support Engineer with strong expertise in application support, incident and change management, Linux/Unix, SQL, Oracle, and monitoring tools. The candidate will be responsible for maintaining application availability, troubleshooting production issues, monitoring system performance, and coordinating with technical and business stakeholders.
Key Responsibilities
- Provide L2/L3 production support for business-critical applications.
- Monitor applications and infrastructure using Splunk, Grafana, and AppDynamics.
- Analyze and resolve production incidents within defined SLAs.
- Perform incident, problem, change, and service request management.
- Troubleshoot application issues across Linux/Unix, SQL, and Oracle environments.
- Perform SQL queries and database-level troubleshooting to identify application issues.
- Analyze application logs, alerts, and performance metrics to identify root causes.
- Coordinate with development, database, infrastructure, and other technical teams for issue resolution.
- Participate in Root Cause Analysis (RCA) and implement corrective/preventive actions.
- Support application deployments, releases, and production changes.
- Ensure effective communication with business users and stakeholders during critical incidents.
- Identify recurring issues and drive problem management and service improvement initiatives.
- Maintain support documentation, knowledge articles, and operational procedures.
- Participate in on-call/shift support as required.
Mandatory Skills
- 6+ years of experience in Application Production Support.
- Strong experience in Incident & Change Management.
- Hands-on experience with Linux/Unix.
- Good knowledge of SQL and Oracle database support.
- Experience with monitoring and observability tools:
- Splunk
- Grafana
- AppDynamics
- Strong troubleshooting and problem-solving skills.
- Good understanding of application monitoring, logs, alerts, and performance analysis.
- Strong stakeholder management and communication skills
Role Summary
We are seeking a proactive and technically skilled Python Application Support Engineer to join our Technical Operations team. This role is crucial for ensuring the stability and reliability of our mission-critical, Python-based applications. You will be responsible for timely incident resolution, deep-dive troubleshooting, implementing permanent fixes, and driving operational efficiency through automation.
🔑 Key Responsibilities
Technical Troubleshooting & Incident Management
- Incident Resolution: Serve as the primary point of contact for complex Level 2 and Level 3 production incidents, diagnosing root causes and resolving issues across our Python application stack.
- Deep-Dive Analysis: Utilize log analysis tools (e.g., Splunk, ELK Stack) and monitoring platforms (e.g., Prometheus, Grafana) to quickly identify and address anomalies in application behavior.
- Code Debugging: Analyze, debug, and fix application issues directly within the Python codebase, including Flask/Django services, worker queues, and custom scripts.
- Database Health: Troubleshoot performance issues and conduct basic SQL/NoSQL query tuning and health checks (e.g., for PostgreSQL, MongoDB, or Redis).
Operational Excellence & Automation
- Monitoring & Alerting: Continuously refine and optimize application monitoring, alerting, and logging configurations to improve mean time to detect (MTTD) and mean time to resolve (MTTR).
- Python Automation: Develop, maintain, and enhance automated scripts (primarily in Python) to streamline routine operational tasks, reporting, health checks, and system recovery processes.
- Documentation: Create and maintain comprehensive documentation, runbooks, and knowledge base articles for application support procedures and recurring issues.
Collaboration & Prevention
- Cross-Functional Fixes: Collaborate closely with the Development and DevOps teams to provide clear technical feedback on recurring issues and implement permanent, scalable solutions.
- Proactive Maintenance: Identify potential system bottlenecks, performance degradation points, and areas prone to failure, recommending and implementing preventative measures.
⚙️ Required Qualifications
- Experience: 3 to 5 years of professional experience in Application Support, Production Support, Site Reliability Engineering (SRE), or a similar technical role.
- Python Expertise (Mandatory): Strong hands-on experience with Python scripting and programming, including the ability to read, debug, and modify application code.
- Operating Systems: Proficient working knowledge of Linux/Unix environments and shell scripting.
- Databases: Solid experience with relational (e.g., PostgreSQL, MySQL) and/or NoSQL (e.g., MongoDB, Redis) databases, focusing on query analysis and performance.
Job Summary
We are looking for an experienced Observability Engineer with strong expertise in ThousandEyes, Splunk, database technologies, and Python automation. The candidate will be responsible for monitoring application, network, and infrastructure performance, identifying issues, and developing automation solutions to improve system visibility, reliability, and operational efficiency.
Required Skills
- Observability
- ThousandEyes / Cisco ThousandEyes
- Splunk
- OracleDB
- MongoDB
- SQL
- Python Automation
Roles & Responsibilities
- Implement and support observability and monitoring solutions across applications, networks, and infrastructure.
- Work with ThousandEyes for network and digital experience monitoring.
- Configure and maintain monitoring dashboards, alerts, and performance metrics.
- Use Splunk for log analysis, monitoring, troubleshooting, and reporting.
- Work with OracleDB, MongoDB, and SQL for data analysis and troubleshooting.
- Develop Python scripts for monitoring and operational automation.
- Analyze performance issues and identify root causes across applications and infrastructure.
- Collaborate with application, infrastructure, and network teams to resolve incidents.
- Improve monitoring, alerting, and automation processes.
- Prepare performance and monitoring reports and provide actionable insights.
Mandatory Skills
- 7+ years of relevant experience
- Strong experience in Observability / Monitoring
- Hands-on experience with ThousandEyes
- Splunk
- Python Automation
- SQL
- OracleDB / MongoDB
- Linux troubleshooting
- Hands-on AWS
- Production/Application Support
- Bash/Shell/Python
- Monitoring/log analysis
- Incident resolution
- Application deployment/support
- Basic networking and database knowledge
- Production/batch support exposure
- Willingness for rotational weekend/critical production support
Observability Engineer (AppDynamics)
Hyderabad
Exp: 8+years of exp
Mandate Skills: AppDynamics, Splunk, Python (Scripting knowledge)
Primary Skill Set:
- AppDynamics administration
- SPLOC administration
- Enterprise monitoring and observability
- Application performance monitoring
- Alerting and event management
- Monitoring strategy and design
- Platform configuration and governance
- Python scripting
- Monitoring automation
Secondary Skills:
- Glassbox monitoring
- Customer journey observability
- GenAI concepts for operations
- Log, metric, and trace telemetry
- Dashboarding and visualization
SRE – Network / DNS / Load Balancer
Experience: 7–11 Years
Location: Hyderabad
Work Mode: WFO
Availability: Immediate Joiner
Job Description:
- Strong experience in Site Reliability Engineering (SRE) with focus on infrastructure and application reliability.
- Hands-on experience with Network, DNS and Load Balancer troubleshooting and administration.
- Monitor system performance, availability, latency and infrastructure health.
- Troubleshoot network connectivity, DNS resolution, routing and load-balancing issues.
- Experience with load balancers such as F5, BIG-IP, HAProxy or similar technologies.
- Good understanding of TCP/IP, HTTP/HTTPS, LAN/WAN, SSL/TLS and networking concepts.
- Experience with DNS technologies such as BIND, Infoblox or equivalent.
- Work on incident management, root cause analysis and problem resolution.
- Collaborate with application, network, cloud and infrastructure teams to resolve production issues.
- Experience with monitoring and alerting tools such as Splunk, Grafana, Prometheus, AppDynamics or similar tools.
- Strong troubleshooting, production support and communication skills.
- Willingness to work from Hyderabad office (WFO) and join immediately.
The recruiter has not been active on this job recently. You may apply but please expect a delayed response.
About the Role
We are looking for a proactive and detail-oriented Senior Site Reliability Engineer (SRE) to ensure the reliability, performance, and availability of our applications. The role involves monitoring production systems, troubleshooting issues, and collaborating with cross-functional teams to drive faster resolution and continuous improvement. You will play a key role in maintaining system stability and enhancing observability across our microservices-based platform.
Key Responsibilities
- Handle MFS application issues by investigating, troubleshooting, and escalating to engineering teams when needed
- Perform initial root cause analysis (RCA) and support resolution of recurring or moderately complex issues
- Ensure timely incident resolution in line with SLAs, including proper documentation of fixes and workarounds
- Identify and analyze system bottlenecks, and assist in deploying fixes via change management processes
- Collaborate with cross-functional teams (Development, SRE/DevOps, QA, Business) to resolve incidents and improve systems
- Use observability tools (Grafana, Loki, ELK) to monitor system health, availability, performance, and resiliency
- Participate in incident/severity calls, ensuring clear communication and coordination
- Develop and maintain knowledge bases, SOPs, and runbooks for standardized operations and troubleshooting
Required Skills & Experience
- Strong understanding of Linux/Unix systems for application support
- Hands-on experience troubleshooting applications in staging and production environments
- Ability to monitor system performance and identify root causes using logs and metrics
- Experience working with Kubernetes and microservices-based architectures
- Proficiency in observability and monitoring tools such as Grafana, Loki, and ELK (Elasticsearch, Logstash, Kibana)
- Familiarity with CI/CD practices and tools (e.g., Jenkins, GitOps)
- Experience in API testing and validation using tools like Postman and Swagger/OpenAPI
- Hands-on experience with PostgreSQL and MongoDB for troubleshooting and ad-hoc reporting
- Experience with ticketing and documentation tools such as Jira and Confluence
- Minimum 4+ years of experience in application support or reliability engineering
Education & Certifications
- Bachelor's degree in Computer Science, Information Technology, or a related field
- Relevant certifications (Cloud, Kubernetes, Microservices) are a plus
Work Schedule
- Willingness to work in a 24x7 environment, including weekends and on-call rotations







