Platform Engineer at Service Co · Pune · 6 - 10 years · ₹20L - ₹45L / yr · Posted 1 Oct 2026
Hiring Platform Engineer
Exp: 6 -- 10 yrs
Edu : BE/B.tech/MCA
Work Location : Pune
Skills :
Platform monitoring ,Incident trouble shooting, Incident recovery, openshift ,kubernetes.
2 years of IT operations, infrastructure, cloud or application support experience.
Exp in Linux command-line knowledge.
Exp in networking knowledge including IP addressing, DNS, ports and connectivity troubleshooting.

Similar jobs (10)
This is a Red Hat OpenShift / Kubernetes Platform Engineer role, mainly focused on OpenShift cluster administration, installation, upgrades, troubleshooting, security certificates, and BAU operations.
Key responsibilities:
- Install and configure OpenShift clusters on bare metal, on-prem VMware/virtual environments, and cloud platforms.
- Strong hands-on experience with Red Hat OpenShift 3.x/4.x.
- Administer OpenShift Container Platform (OCP) and Kubernetes environments.
- Perform cluster installation, upgrades, health checks, certificate renewal, and troubleshooting.
- Manage OpenShift projects, users, roles, access, and day-to-day platform activities.
- Monitor cluster performance, identify issues, and maintain high availability.
- Perform cluster scaling and performance optimization.
- Provide production/BAU support for OpenShift environments.
🚀 Hiring – AWS / Kubernetes / OpenShift Engineer
📍 Location: Bangalore
💼 Experience: 7–10 Years
⚡ Joining: Immediate Joiners Only
🔑 Required Skills
- Strong hands-on experience in AWS
- Expertise in Kubernetes & OpenShift
- Strong Linux Administration & Troubleshooting
- Experience in containerized environments and platform operations
- Production support, monitoring and incident troubleshooting
- Good understanding of cloud and infrastructure technologies
📌 Interview Process
2nd Round – Face-to-Face Interview in Bangalore
👉 Please share profiles of candidates who are available for a F2F interview in Bangalore.
#Hiring #AWS #Kubernetes #OpenShift #Linux #CloudEngineer #PlatformEngineer #DevOps #BangaloreJobs #ImmediateJoiners #WFO #ITJobs
Site Reliability Engineer (SRE) / Production Support Engineer
Experience: 5–10 Years
Location: Hyderabad
Work Mode: Face-to-Face Drive
Shift: Rotational Shifts
Job Description
Looking for an experienced SRE / Production Support Engineer with strong experience in application and production support, incident management, monitoring, troubleshooting, and cloud operations.
Key Skills
Production Support, Incident Management, Splunk, APM, SLI/SLO, Cloud, Kubernetes, Docker, Terraform, Linux/Windows Administration, Shell Scripting and Python.
Good understanding of production deployments, batch monitoring, network/load balancing, and troubleshooting is required.
Candidates from SRE, Production Support, Application Support, Cloud Operations, or DevOps backgrounds are preferred.

Platform Engineer
Location: Bengaluru, India (Hybrid)
Employment Type: Full-time
Experience: 2-4 years
About Compnay
This is driving the electric mobility revolution through cutting-edge software, infrastructure, and professional services. Our technology empowers utilities, cities, fleets, transit agencies, and automakers to deploy EV charging infrastructure at scale safely, efficiently, and sustainably. With a global footprint spanning three continents and operations in 13 countries, we are passionate about shaping the future of sustainable transport.
Operating over 70,000 charge points globally, this is driving the transition toward cleaner, smarter, and more efficient mobility. The India team serves as a critical operational hub, supporting global platforms focused on decarbonization, digitalization, and scalable infrastructure growth.
Role Overview
What you’ll do:
- Ensure system reliability, uptime, and performance of global platform.
- Conduct real-time surveillance of our EV charging systems to proactively identify and mitigate performance issues and anomalies near 24/7 basis. As such, you collaborate with IDT and FMC players to ensure incident detection also happens outside office hours (monitoring shifts among team members subject to duty schedule).
- Deliver on change & releases like firmware changes and drive insights & intelligence back into testing processes and tech discussions with the wider organization.
- Successfully deliver and project manage first time right commissioning activities alongside our Engineering Procurement Contract Management (EPCM) partners to successfully bring charge points onto our Charge Point Management System (CPMS).
- Provide technical guidance and support to DC specialists during the commissioning of EV charging solutions.
- Work closely with Shell, Engineering, and IT colleagues to ensure projects are completed on time and to specification.
- Act as a liaison with the Engineering Procurement Contract Management (EPCM) partner to manage projects from start to finish, ensuring charge points are successfully onboarded on the Charge Point Management System (CPMS).
- Collaborate with development, operations and support teams to build scalable and resilient systems.
- Contribute to incident response, root-cause analysis, and post-mortem reviews, driving continuous improvement.
- Participate in capacity planning, performance tuning, and resource optimization.
- Integrate security and compliance best practices into all infrastructure operations.
- Stay current with emerging SRE tools, frameworks, and cloud technologies to continuously improve reliability practices.
- Participate in and lead on-call rotations and incident response, conducting detailed postmortems and RCA reports.
- Flexible to resolve blocking issues during off hours or weekends if required.
What We’re Looking For:
Basic Qualifications and Skills
- Bachelor’s degree in Engineering, Electrical, ECE, Computer Science, Information Technology, or related field.
- 2–4 years of overall experience with at least 1+ years of experience as a Site Reliability Engineer, DevOps Engineer, or Technical Project Coordinator.
- Proven experience of DevOps, SRE or Technical Project Coordination with IoT or connected devices-based platforms.
- Experience with incident management and on-call best practices. Provide support to on-call engineers.
- Excellent analytical and problem-solving skills with a proactive mindset.
- Expertise with monitoring and observability tools (Dynatrace, Prometheus, Grafana, Zabbix, etc.).
- Solid understanding of cloud platforms (AWS) and AWS native services (EKS, EC2, S3, RDS, Lambda).
- Proactively monitor the network, triage performance outliers, and coordinate correction actions to ensure optimal system functionality.
- Fluency in English (spoken and written).
- Successfully recommission or decommission chargers following changes in our network.
- Responsible for the go-live of the chargers on Shell’s public network following commissioning attempts.
Additional Information
- This role involves managing infrastructure for a global platform operating in over ten countries, requiring effective communication and collaboration across regions.
- Strong verbal and written communication skills, along with availability and flexibility to resolve blocking issues, are essential to support on-call engineers.
- This role may involve EU or US time-zone shifts based on business requirements.
- Shift timing: 2 PM IST to 11 PM IST.
What is required to be successful in this role:
- Global platform experience (B2C or B2B).
- AWS native service experience.
- Firmware deployment and cloud cost optimization experience.
- Strong exposure to monitoring and alerts.
- Experience with firmware rollout, IoT devices onboarding and offboarding will be an added advantage.
- Experience as an SRE or DevOps Engineer with some exposure to Project Management or Technical Project Management in IoT-based projects will be helpful.
What We Offer
- Work with some of the brightest minds in the emerging EV industry.
- Make a tangible impact in reducing carbon emissions and enabling sustainable energy.
- Freedom to suggest, implement, and innovate on systems, processes, and technologies.
- Daily ownership in a high-growth, challenging environment.
- Flexible work environment with hybrid schedules and virtualization options.
- Competitive pay and benefits including health coverage, innovative PTO program, and performance bonuses.

Jr Platform Engineer
Location: Bengaluru, India (Hybrid)
Employment Type: Full-time
Experience: 0.6-2 years
Shift Timing: 2 PM to 11 PM IST
About Company
It is driving the electric mobility revolution through cutting-edge software, infrastructure, and professional services. Our technology empowers utilities, cities, fleets, transit agencies, and automakers to deploy EV charging infrastructure at scale safely, efficiently, and sustainably. With a global footprint spanning three continents and operations in 13 countries, we are passionate about shaping the future of sustainable transport.
Operating over 70,000 charge points globally,It is driving the transition toward cleaner, smarter, and more efficient mobility. The India team serves as a critical operational hub, supporting global platforms focused on decarbonization, digitalization, and scalable infrastructure growth.
At this company, we value purpose-driven individuals who want to make a meaningful impact and help create a cleaner, smarter, and more connected world.
Role Overview
It is seeking a TechOps Engineer! We are excited to find a highly engaged engineer who is obsessed with technology that wants to be a part of a “world class” platform SRE team. It engineers must possess an "automation first" mindset, with a relentless focus on documentation, quality, scalability, and reliability using Infrastructure as Code tools. This position will be part of a platform team that is developing exciting products and solutions and playing a key part in driving forward the electrification of transportation.
What you’ll do:
- Ensure system reliability, uptime, and performance of global platform.
- Conduct real-time surveillance of our EV charging systems to proactively identify and mitigate performance issues and anomalies near 24/7 basis. As such, you collaborate with IDT and FMC players to ensure incident detection also happens outside office hours (monitoring shifts among team members subject to duty schedule)
- Deliver on change & releases like firmware changes and drive insights & intelligence back into testing processes and tech discussions with the wider organization.
- Successfully deliver and project manage first time right commissioning activities alongside our Engineering Procurement Contract Management (EPCM) partners to successfully bring charge points onto our Charge Point Management System (CPMS).
- End-to-end EV charger lifecycle management, including deployment, commissioning, monitoring, maintenance, and decommissioning activities.
- Provide technical guidance and support to DC specialists during the commissioning of EV charging solutions.
- Work closely with Shell, Engineering, and IT colleagues to ensure projects are completed on time and to specification.
- Act as a liaison with the Engineering Procurement Contract Management (EPCM) partner to manage projects from start to finish, ensuring charge points are successfully onboarded on the Charge Point Management System (CPMS).
- Collaborate with development, operations and support teams to build scalable and resilient systems.
- Contribute to incident response, root-cause analysis, and post-mortem reviews, driving continuous improvement.
- Participate in capacity planning, performance tuning, and resource optimization.
- Integrate security and compliance best practices into all infrastructure operations.
- Stay current with emerging SRE tools, frameworks, and cloud technologies to continuously improve reliability practices.
- Participate in and lead on-call rotations and incident response, conducting detailed postmortems and RCA reports.
- Flexible to resolve blocking issues during off hours or weekends if required.
What We’re Looking For:
Basic Qualifications and skills
- Bachelor’s degree in Engineering , Electrical, ECE, Computer Science, Information Technology, or related field.
- Overall 1+ years of experience as a Site Reliability Engineer, DevOps/ Technical project coordinator role.
- Proven experience of DevOps, SRE, or Technical Project Coordination with IoT or connected devices based platforms.
- Hands-on experience with cloud platforms such as AWS and Infrastructure as Code (IaC) tools such as Terraform.
- Experience with incident management and on-call best practices. Provide support to on call engineers.
- Excellent analytical and problem-solving skills with a proactive mindset.
- Expertise with monitoring and observability tools (Dynatrace,Prometheus, Grafana, Zabbix, etc.).
- Proactively monitor the network, triage performance outliers, and coordinate correction actions to ensure optimal system functionality.
- Fluency in English (spoken and written).
- Successfully recommission or decommission chargers following changes in our network.
- Responsible for the go-live of the chargers on Shell’s public network following commissioning attempts.
Note: This role involves managing infrastructure for a global platform operating in over ten countries, requiring effective communication and collaboration across regions. Strong verbal and written communication skills, along with availability and flexibility to resolve blocking issues, are essential to support On-call Engineers. This role may involve EU or US time‑zone shifts based on business requirements. The shift timing will be 2 PM IST to 11 PM IST.
What We Offer
- Work with some of the brightest minds in the emerging EV industry.
- Make a tangible impact in reducing carbon emissions and enabling sustainable energy.
- Freedom to suggest, implement, and innovate on systems, processes, and technologies.
- Daily ownership in a high-growth, challenging environment.
- Flexible work environment with hybrid schedules and virtualization options.
- Competitive pay and benefits including health coverage, innovative PTO program, and performance bonuses.
We're looking for a Site Reliability Engineer to keep our production systems fast, reliable, and scalable. Sitting at the intersection of software engineering and operations, you'll treat infrastructure as code, automate away toil, and build the observability that lets us catch problems before customers do. You'll own uptime and on-call for critical services, lead incident response and blameless postmortems, and continuously harden the platform against failure. This role suits an engineer who is as comfortable debugging a production incident at 2 a.m. as they are writing the automation that prevents the next one.
Key Responsibilities
- Own reliability, availability, and performance of production services, including on-call rotation
- Build and maintain monitoring, alerting, and observability (metrics, logs, traces)
- Automate deployments, scaling, and operational tasks to reduce manual toil
- Manage containerized workloads on Kubernetes and cloud infrastructure
- Design and maintain CI/CD pipelines for safe, frequent releases
- Lead incident response and drive blameless postmortems with clear follow-ups
- Perform capacity planning, performance tuning, and cost optimization
- Define and track SLIs/SLOs and error budgets with product teams
Requirements
- 3+ years in SRE, DevOps, or production-focused engineering
- Strong Linux administration and hands-on Kubernetes experience
- Solid experience with monitoring/observability tools (Prometheus, Grafana, ELK, or similar)
- Cloud experience with AWS, GCP, or Azure
- CI/CD pipelines and infrastructure-as-code (Terraform, CloudFormation)
- Proficient scripting in Python and/or Bash
Nice to have
- Experience with service meshes, Helm, or GitOps (ArgoCD/Flux)
- Background in high-traffic or distributed systems
AWS / Kubernetes / OpenShift / Linux –
Location: Bangalore
Experience: 7–10 Years
Mandatory Skills:
- Strong hands-on experience with AWS
- Experience in Kubernetes & OpenShift
- Strong knowledge of Linux administration
- Experience with Docker & containerization
- Knowledge of CI/CD pipelines and DevOps practices
- Troubleshooting, monitoring, and deployment experience
Role: Cloud/DevOps Engineer – AWS, Kubernetes & OpenShift
We are hiring a Kubernetes Engineer to run and scale our container platform.
Responsibilities
- Manage production Kubernetes clusters
- Write Helm charts and Terraform modules
- Set up monitoring, alerting and autoscaling
- Improve CI/CD pipelines and deployments
Requirements
- 2+ years running Kubernetes in production
- Strong Docker and CI/CD experience
- CKA certification is a plus

Job Title: TechOps Engineer
Location: Bengaluru, India (Hybrid)
Employment Type: Full-time
Experience: 6 Month-2 years (Excluding Internship)
Shift Timing: 2 PM to 11 PM IST
Role Overview
We are excited to find a highly engaged engineer who is obsessed with technology that wants to be a part of a “world class” platform SRE team. Engineers must possess an "automation first" mindset, with a relentless focus on documentation, quality, scalability, and reliability using Infrastructure as Code tools. This position will be part of a platform team that is developing exciting products and solutions and playing a key part in driving forward the electrification of transportation.
What you’ll do:
- Ensure system reliability, uptime, and performance of global platform.
- Conduct real-time surveillance of our EV charging systems to proactively identify and mitigate performance issues and anomalies near 24/7 basis. As such, you collaborate with IDT and FMC players to ensure incident detection also happens outside office hours (monitoring shifts among team members subject to duty schedule)
- Deliver on change & releases like firmware changes and drive insights & intelligence back into testing processes and tech discussions with the wider organization.
- Successfully deliver and project manage first time right commissioning activities alongside our Engineering Procurement Contract Management (EPCM) partners to successfully bring charge points onto our Charge Point Management System (CPMS).
- End-to-end EV charger lifecycle management, including deployment, commissioning, monitoring, maintenance, and decommissioning activities.
- Provide technical guidance and support to DC specialists during the commissioning of EV charging solutions.
- Work closely with Shell, Engineering, and IT colleagues to ensure projects are completed on time and to specification.
- Act as a liaison with the Engineering Procurement Contract Management (EPCM) partner to manage projects from start to finish, ensuring charge points are successfully onboarded on the Charge Point Management System (CPMS).
- Collaborate with development, operations and support teams to build scalable and resilient systems.
- Contribute to incident response, root-cause analysis, and post-mortem reviews, driving continuous improvement.
- Participate in capacity planning, performance tuning, and resource optimization.
- Integrate security and compliance best practices into all infrastructure operations.
- Stay current with emerging SRE tools, frameworks, and cloud technologies to continuously improve reliability practices.
- Participate in and lead on-call rotations and incident response, conducting detailed postmortems and RCA reports.
- Flexible to resolve blocking issues during off hours or weekends if required.
What We’re Looking For:
Basic Qualifications and skills
- Bachelor’s degree in Engineering , Electrical, ECE, Computer Science, Information Technology, or related field.
- Overall 1 years of experience as a Site Reliability Engineer, Technical project coordinator role.
- Proven experience of SRE or Technical Project Coordination with IoT or connected devices based platforms.
- Experience with incident management and on-call best practices. Provide support to on call engineers.
- Excellent analytical and problem-solving skills with a proactive mindset.
- Hands-on experience with AWS Cloud and IaC tools such as Terraform or Ansible.
- Expertise with monitoring and observability tools (Dynatrace,Prometheus, Grafana, Zabbix, etc.).
- Proactively monitor the network, triage performance outliers, and coordinate correction actions to ensure optimal system functionality.
- Fluency in English (spoken and written).
- Successfully recommission or decommission chargers following changes in our network.
- Responsible for the go-live of the chargers on Shell’s public network following commissioning attempts.
Note: This role involves managing infrastructure for a global platform operating in over ten countries, requiring effective communication and collaboration across regions. Strong verbal and written communication skills, along with availability and flexibility to resolve blocking issues, are essential to support On-call Engineers. This role may involve EU or US time‑zone shifts based on business requirements. The shift timing will be 2 PM IST to 11 PM IST.
What We Offer
- Work with some of the brightest minds in the emerging EV industry.
- Make a tangible impact in reducing carbon emissions and enabling sustainable energy.
- Freedom to suggest, implement, and innovate on systems, processes, and technologies.
- Daily ownership in a high-growth, challenging environment.
- Flexible work environment with hybrid schedules and virtualization options.
- Competitive pay and benefits including health coverage, innovative PTO program, and performance bonuses.
Key Responsibilities
Platform Monitoring and Reliability
- Monitor the health and performance of EDM’s advertiser, publisher, broker, and internal platforms.
- Maintain monitoring, alerting, logging, and system-health dashboards.
- Investigate platform outages, degraded performance, failed transactions, delayed data, and integration errors.
- Respond to production incidents and coordinate resolutions with the appropriate engineers and vendors.
- Perform root-cause analysis and document corrective and preventive actions.
- Help maintain defined uptime, response-time, recovery-time, and system-reliability targets.
- Identify recurring problems and recommend permanent solutions.
Cloud Infrastructure and Systems Operations
- Maintain and support cloud infrastructure, servers, databases, networks, storage, and production environments.
- Support development, staging, and production environments.
- Assist with infrastructure scaling, system upgrades, patching, backups, and disaster recovery.
- Monitor cloud usage and help control infrastructure and technology costs.
- Maintain access controls, service accounts, certificates, domain configurations, and environment variables.
- Ensure production systems are properly documented and recoverable.
Deployment and Release Support
- Support safe and consistent application deployments.
- Maintain or improve continuous integration and deployment workflows.
- Coordinate release schedules, deployment validation, rollback procedures, and post-release monitoring.
- Help engineering teams identify configuration or infrastructure problems before releases reach production.
- Maintain deployment documentation, technical checklists, and change logs.
- Reduce manual deployment work through automation.
API and Integration Support
- Monitor and troubleshoot third-party APIs, webhooks, postbacks, dialer connections, CRM integrations, payment systems, tracking platforms, and compliance services.
- Investigate failed lead deliveries, missing postbacks, duplicate records, delayed reporting, and authentication problems.
- Support ping-post, real-time bidding, call-routing, SIP, and data-transfer workflows.
- Work with advertisers, publishers, and vendors to diagnose technical integration problems.
- Create clear documentation for common integration methods and troubleshooting procedures.
- Develop alerts that identify integration failures before clients report them.
Call and Lead Operations
- Monitor call-routing, tracking, recording, attribution, and disposition systems.
- Investigate calls that fail to route, connect, record, track, or report correctly.
- Support number provisioning, routing rules, caps, schedules, geographic restrictions, buyer availability, and failover logic.
- Validate that leads, calls, and transactions are properly attributed to the correct advertiser, publisher, campaign, and payout.
- Assist with discrepancies involving call duration, billable events, conversions, payouts, and reporting.
- Help protect revenue by identifying technical leakage and delivery failures.
Data and Reporting Support
- Monitor data pipelines, scheduled jobs, reporting processes, and database performance.
- Investigate discrepancies between platform reporting, billing records, payment records, and third-party systems.
- Write and maintain database queries for troubleshooting, validation, and operational reporting.
- Assist with data corrections using controlled and documented procedures.
- Support dashboards and operational alerts for revenue, margin, consumption, conversion, and platform activity.
- Maintain appropriate controls around production data access and modification.
Security and Access Management
- Support role-based access controls, multifactor authentication, audit logging, encryption, and secure system configuration.
- Provision and remove employee, contractor, client, and vendor access.
- Monitor suspicious activity and report potential security incidents.
- Assist with vulnerability remediation, security reviews, access audits, and incident-response procedures.
- Protect consumer, advertiser, publisher, employee, and company information.
- Follow company policies for credentials, production access, sensitive data, and change management.
Automation and Process Improvement
- Automate repetitive operational tasks using scripts, workflows, APIs, and infrastructure tools.
- Reduce manual work associated with monitoring, deployments, reporting, reconciliation, onboarding, and support.
- Build internal tools that improve visibility and response times.
- Identify operational bottlenecks that affect revenue, margin, client satisfaction, or employee productivity.
- Maintain clear runbooks and standard operating procedures for recurring technical tasks.
Technical Support and Documentation
- Serve as an escalation point for complex platform and integration issues.
- Translate technical problems into clear explanations for nontechnical teams.
- Create and maintain architecture diagrams, system inventories, runbooks, troubleshooting guides, and incident reports.
- Track incidents and technical requests through completion.
- Document known issues, temporary workarounds, permanent resolutions, and system dependencies.
- Participate in an on-call rotation for urgent production incidents.
First 90-Day Priorities
The successful candidate will be expected to:
- Learn EDM’s platforms, infrastructure, integrations, call-routing systems, reporting processes, and revenue workflows.
- Document critical systems, dependencies, credentials ownership, vendor contacts, and escalation procedures.
- Review existing monitoring, alerting, backups, access controls, and deployment procedures.
- Establish baseline metrics for uptime, incident volume, response time, recovery time, deployment success, and integration failures.
- Resolve high-priority recurring production and integration issues.
- Improve alerting for call-routing failures, API errors, delayed data, failed jobs, and reporting discrepancies.
- Create runbooks for the company’s most common and highest-risk technical incidents.
- Identify at least three meaningful automation or cost-saving opportunities.
- Participate in production support and demonstrate ownership of incidents through resolution.
Performance Expectations
Success will be measured by:
- Platform uptime and reliability
- Mean time to acknowledge and resolve incidents
- Reduction in recurring production problems
- Deployment success and rollback rates
- API, postback, webhook, and call-routing reliability
- Reporting and data accuracy
- Backup and recovery readiness
- Quality of technical documentation
- Security and access-control compliance
- Reduction in manual operational work
- Infrastructure costs relative to platform volume
- Responsiveness to internal teams, clients, and technical partners
Required Qualifications
- Three or more years of experience in operations engineering, DevOps, site reliability engineering, cloud infrastructure, systems administration, or production support.
- Hands-on experience with cloud platforms such as AWS, Azure, or Google Cloud.
- Experience supporting Linux-based production environments.
- Working knowledge of networking, DNS, SSL certificates, firewalls, load balancing, and application security.
- Experience with relational databases and SQL.
- Experience troubleshooting REST APIs, webhooks, authentication, and third-party integrations.
- Familiarity with monitoring, logging, alerting, and incident-management tools.
- Experience with scripting languages such as Python, Bash, JavaScript, or PowerShell.
- Understanding of source control, deployment pipelines, and release management.
- Strong troubleshooting, documentation, and communication skills.
- Ability to prioritize incidents based on business and revenue impact.
- Availability to participate in an on-call rotation.
Preferred Qualifications
- Experience in ad-tech, mar-tech, affiliate marketing, lead generation, pay-per-call, telecommunications, or SaaS.
- Familiarity with SIP, VoIP, dialers, call-tracking platforms, routing systems, and phone-number provisioning.
- Experience with containers, infrastructure as code, and automated deployment tools.
- Experience with Docker, Kubernetes, Terraform, GitHub Actions, or similar technologies.
- Familiarity with payment processing, usage-based billing, reconciliation, and commission systems.
- Experience working with real-time bidding, ping-post, lead distribution, or high-volume transactional systems.
- Understanding of TCPA-related controls, consent records, DNC suppression, data privacy, or regulated marketing environments.
- Experience with security audits, disaster-recovery testing, and compliance documentation.
Ideal Candidate
The ideal candidate:
- Takes ownership instead of waiting for someone else to fix the problem.
- Remains calm and methodical during high-impact incidents.
- Understands the difference between applying a temporary fix and eliminating a root cause.
- Communicates technical problems clearly and directly.
- Recognizes that production reliability is a business and revenue responsibility.
- Automates repetitive work whenever practical.
- Documents systems so the company is not dependent on one person’s memory.
- Protects security and stability without creating unnecessary bureaucracy.
- Is comfortable working in a fast-moving entrepreneurial environment.
- Can manage competing priorities while maintaining attention to detail.






