Cutshort logo
For Employers
Porter.in logo
Cloud Production Support Engineer - Level 2
Cloud Production Support Engineer - Level 2

Cloud Production Support Engineer - Level 2 at Porter.in · Bengaluru (Bangalore) · 3 - 5 years · ₹8L - ₹12L / yr · Profitable · Posted 12 Oct 2023

UPhill HR's logo

Cloud Production Support Engineer - Level 2

Agency job
3 - 5 yrs
₹8L - ₹12L / yr
Bengaluru (Bangalore)
Skills
skill iconDocker
skill iconAmazon Web Services (AWS)
Linux administration
Monitoring
AWS CloudFormation

Job Summary


Cloud Production Support Engineer(PSE) is responsible for fulfilling the day-to-day infrastructure and service requests from the application teams across AWS, CI/CD solutions and observability tools. You will be expected to handle production issues in collaboration with the cloud Infrastructure and application teams.


Responsibilities and Duties


  • Troubleshoot production Issues: When technical issues with the cloud infrastructure components arise, PSE must act quickly to analyse the available data and find the root cause of the problem. They may then develop a solution or escalate the problem to other engineering team members while providing stakeholders with progress updates.
  • Infrastructure provisioning and modification: Application teams may request to create new infrastructure or modify the existing ones in AWS based on their requirements via the ticketing tool. PSE should ensure that the required data/info is available on the ticket and provide a resolution based on the given SLA.
  • Alert Management: Alerts from the observability tools will be received on multiple channels according to the notification settings. PSEs are expected to acknowledge the alerts, troubleshoot the issue, close the alert based on the given SLA, or escalate to the cloud infra/DevOps team for further diagnosis.
  • Onboarding, Off-boarding and access management: Whenever an employee joins or leaves the organization, you will receive an onboarding or offboarding request.
  • Prepare Technical Documentation: PSEs must prepare documentation when logging product issues, as they must note all details, including their observations, diagnoses, and action steps. Other everyday tasks include weekly reports summarising production performance, upgrade release notes, and troubleshooting guides.
  • Product Improvements: Since PSEs have good exposure to the product issues, they should work closely with the PMs+EMs, pass the feedback on the product, and get the improvements/fixes included in the product roadmap.
  • Adherence to SLA and timelines: PSEs should always adhere to the timelines shared with other teams for closure of fixes and deliver outcomes as per the SLA guidance agreed with business teams
  • Reporting: Report & track weekly regarding SLA metrics, tickets being worked and closed by PSEs/transferred tickets. Identify and devise how productivity can be captured at the individual level and report the same monthly.


Qualifications and Skills


  • Degree in Computer Science/Information Technology.
  • Two years or more experience in Cloud and system administration.
  • Experience troubleshooting in complex environments using monitoring tools.
  • Demonstrated experience with containerisation technologies (Docker, Kubernetes, etc.)
  • Hands-on experience with the most common AWS services.
Read more
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos

About Porter.in

Founded :
2014
Type :
Product
Size :
1000-5000
Stage :
Profitable

About

Company Overview:

At Porter, we are passionate about improving productivity. We want to help businesses, large and small, optimize their last-mile operations and empower them to unleash the growth of their core functions. Last-mile delivery logistics is one of the biggest and fastest-growing sectors of the economy with a market cap upwards of 50 billion USD and a growth rate exceeding 15% CAGR.


Porter is the fastest-growing leader in this sector with operations in 14 major cities, a fleet size exceeding 1L registered and 50k active driver-partners and a customer base with 3.5M being monthly active. Our industry-best technology platform has raised over 50 million USD from investors including Sequoia Capital, Kae Capital, Mahindra Group and LGT Aspada. We are addressing a massive problem and going after a huge market. We’re trying to create a household name in transportation and our ambition is to disrupt all facets of last-mile logistics including warehousing and LTL transportation. At Porter, we’re here to do the best work of our lives. If you want to do the same and love the challenges and opportunities of a fast-paced work environment, then we believe Porter is the right place for you.


Company URL: https://porter.in



Read more

Connect with the team

Profile picture
Satyajit Mittra

Company social profiles

bloglinkedintwitterfacebook

Similar jobs (5)

MNC
MNC
Agency job
via by Sandhiya b
Bengaluru (Bangalore), Hyderabad
8 - 12 yrs
₹2L - ₹22L / yr
Production support
Reliability engineering
Linux/Unix
Monitoring
skill icongrafana

Application Production Support with SRE, Linux/Unix, Splunk/AppD/Grafana, Troubleshooting

WFO-Immediate

8 to 12 Yrs

Bangalore/Hyderabad

Read more
company logo
Pune, Gurugram, Bengaluru (Bangalore), Hyderabad
5 - 12 yrs
₹15L - ₹28L / yr
DevOps
skill iconKubernetes
Incident management
Observability
Reliability engineering
+4 more

Lead Cloud Reliability Engineer


Job Responsibilities

● Lead and manage the Cloud Reliability teams to provide strong Managed Services support to end-customers.

● Isolate, troubleshoot and resolve issues reported by CMS clients in their cloud environment

● Drive the communication with the customer providing details about the issue, current steps, next plan of action, ETA

● Gather client's requirements related to use of specic cloud services and provide assistance in seing them up and resolving issues

● Create SOPs and knowledge articles for use by the L1 teams to resolve common issues

● Identify recurring issues, perform root cause analysis and propose/implement preventive actions

● Follow change management procedure to identify, record and implement changes

● Plan and deploy OS, security patches in Windows/Linux environment and upgrade k8s clusters

● Identify the recurring manual activities and contribute to automation

● Provide technical guidance and educate team members on development and operations. Monitor metrics and develop ways to improve.

● System troubleshooting and problem-solving across plaorm and application domains. Ability to use a wide variety of open-source technologies and cloud services.

● Build, maintain, and monitor conguration standards.

● Ensuring critical system security through using best-in-class cloud security solutions.


Qualifications

● 4-7 years experience in Cloud Infrastructure and Operations domains and IT operational experience preferably in a global enterprise environment.

● Specialize in one or two cloud deployment platforms: AWS, GCP

● Hands on experience with AWS/GCP services (EKS, ECS, EC2, VPC, RDS, Lambda, GKE, Compute Engine)

● Understanding of one or more programming languages (Python, JavaScript, Ruby, Java, .Net)

● Logging and Monitoring tools (ELK, Stackdriver, CloudWatch)

● Knowledge on Conguration Management tools such as Ansible, Terraform, Puppet, Chef

● Experience working with deployment and orchestration technologies (such as Docker, Kubernetes, Mesos)

● Good analytical, communication, problem solving, and learning skills.

● Knowledge on programming against cloud plaorms such as Google Cloud Platform and lean development methodologies.

● Strong service aitude and a commitment to quality.

● Willingness to work in shifts.

Read more
company logo
Agency job
via by Soundarya Valli Chintapalli
Hyderabad
3 - 8 yrs
₹8L - ₹18L / yr
Linux/Unix
  • Linux troubleshooting
  • Hands-on AWS
  • Production/Application Support
  • Bash/Shell/Python
  • Monitoring/log analysis
  • Incident resolution
  • Application deployment/support
  • Basic networking and database knowledge
  • Production/batch support exposure
  • Willingness for rotational weekend/critical production support


Read more
company logo
Pune, Mumbai
4 - 8 yrs
₹6L - ₹22L / yr
Reliability engineering
DevOps
Google Cloud Platform (GCP)
Alerting and Monitoring
skill iconKubernetes
+6 more

Lead Cloud Reliability Engineer


Job Responsibilities

● Lead and manage the Cloud Reliability teams to provide strong Managed Services support to end-customers.

● Isolate, troubleshoot and resolve issues reported by CMS clients in their cloud environment

● Drive the communication with the customer providing details about the issue, current steps, next plan of action, ETA

● Gather client's requirements related to use of specic cloud services and provide assistance in seing them up and resolving issues

● Create SOPs and knowledge articles for use by the L1 teams to resolve common issues

● Identify recurring issues, perform root cause analysis and propose/implement preventive actions

● Follow change management procedure to identify, record and implement changes

● Plan and deploy OS, security patches in Windows/Linux environment and upgrade k8s clusters

● Identify the recurring manual activities and contribute to automation

● Provide technical guidance and educate team members on development and operations. Monitor metrics and develop ways to improve.

● System troubleshooting and problem-solving across plaorm and application domains. Ability to use a wide variety of open-source technologies and cloud services.

● Build, maintain, and monitor conguration standards.

● Ensuring critical system security through using best-in-class cloud security solutions.


Qualifications

● 4-7 years experience in Cloud Infrastructure and Operations domains and IT operational experience preferably in a global enterprise environment.

● Specialize in one or two cloud deployment platforms: AWS, GCP

● Hands on experience with AWS/GCP services (EKS, ECS, EC2, VPC, RDS, Lambda, GKE, Compute Engine)

● Understanding of one or more programming languages (Python, JavaScript, Ruby, Java, .Net)

● Logging and Monitoring tools (ELK, Stackdriver, CloudWatch)

● Knowledge on Conguration Management tools such as Ansible, Terraform, Puppet, Chef

● Experience working with deployment and orchestration technologies (such as Docker, Kubernetes, Mesos)

● Good analytical, communication, problem solving, and learning skills.

● Knowledge on programming against cloud plaorms such as Google Cloud Platform and lean development methodologies.

● Strong service aitude and a commitment to quality.

● Willingness to work in shifts.

Read less


Read more
MNC
MNC
Agency job
via by aafia parveen
Bengaluru (Bangalore), Hyderabad
6 - 12 yrs
₹2L - ₹14L / yr
Production support
skill icongrafana
AppDynamics
Splunk
Linux/Unix
+2 more

Role Summary:

We are looking for an experienced Application Production Support Engineer with strong expertise in application support, incident and change management, Linux/Unix, SQL, Oracle, and monitoring tools. The candidate will be responsible for maintaining application availability, troubleshooting production issues, monitoring system performance, and coordinating with technical and business stakeholders.

Key Responsibilities

  • Provide L2/L3 production support for business-critical applications.
  • Monitor applications and infrastructure using Splunk, Grafana, and AppDynamics.
  • Analyze and resolve production incidents within defined SLAs.
  • Perform incident, problem, change, and service request management.
  • Troubleshoot application issues across Linux/Unix, SQL, and Oracle environments.
  • Perform SQL queries and database-level troubleshooting to identify application issues.
  • Analyze application logs, alerts, and performance metrics to identify root causes.
  • Coordinate with development, database, infrastructure, and other technical teams for issue resolution.
  • Participate in Root Cause Analysis (RCA) and implement corrective/preventive actions.
  • Support application deployments, releases, and production changes.
  • Ensure effective communication with business users and stakeholders during critical incidents.
  • Identify recurring issues and drive problem management and service improvement initiatives.
  • Maintain support documentation, knowledge articles, and operational procedures.
  • Participate in on-call/shift support as required.

Mandatory Skills

  • 6+ years of experience in Application Production Support.
  • Strong experience in Incident & Change Management.
  • Hands-on experience with Linux/Unix.
  • Good knowledge of SQL and Oracle database support.
  • Experience with monitoring and observability tools:
  • Splunk
  • Grafana
  • AppDynamics
  • Strong troubleshooting and problem-solving skills.
  • Good understanding of application monitoring, logs, alerts, and performance analysis.
  • Strong stakeholder management and communication skills


Read more
Why apply to jobs via Cutshort
people_solving_puzzle
Personalized job matches
Stop wasting time. Get matched with jobs that meet your skills, aspirations and preferences.
people_verifying_people
Verified hiring teams
See actual hiring teams, find common social connections or connect with them directly.
ai_chip
Move faster with AI
We use AI to get you faster responses, recommendations and unmatched user experience.
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos