Cutshort logo
For Employers
SSA  Company logo
Oracle Cloud Infrastructure (OCI) L3 / Technical Lead
Oracle Cloud Infrastructure (OCI) L3 / Technical Lead

Oracle Cloud Infrastructure (OCI) L3 / Technical Lead at SSA Company · Mumbai · 8 - 15 years · ₹10L - ₹25L / yr · Posted 12 Apr 2026

1HResource Solutions's logo

Oracle Cloud Infrastructure (OCI) L3 / Technical Lead

Agency job
8 - 15 yrs
₹10L - ₹25L / yr
Mumbai
Skills
Oracle
OCI

About the Role

We are seeking an experienced OCI L3 / Technical Lead to own the reliability, performance, security, and cost efficiency of our OCI workloads. You will serve as the highest technical escalation point for OCI operations, lead architecture and automation initiatives, mentor engineers, and collaborate cross-functionally to ensure resilient, compliant, and scalable solutions. This role combines hands-on engineering with leadership responsibilities across incident management, platform engineering, and cloud governance.

________________________________________

Key Responsibilities

1) L3 Operations & Escalation Management

• Act as the final technical escalation point for critical incidents, complex problems, and performance issues across OCI services (Compute, Networking, Storage, IAM, Load Balancer, WAF, OKE/Kubernetes, DBaaS/Autonomous DB, Exadata Cloud Service).

• Lead root cause analysis (RCA), produce corrective/preventive action plans, and drive problem management per ITIL.

• Own on-call rotations for priority incidents; coordinate across L2/L1 teams and vendors for swift resolution.

2) Architecture, Design & Governance

• Design and review high-availability, disaster recovery (HA/DR) architectures leveraging OCI regions, ADs, Fault Domains, Backup/Archive Storage, Data Guard (for Oracle DB), and multi-cloud patterns as needed.

• Define landing zone architectures, tenancy/subscription structure, compartment strategy, IAM policies, tagging, and cost governance.

• Establish standards for network segmentation (VCNs, subnets), routing, VPN/ FastConnect, NSGs/Security Lists, and WAF

3) Observability, Performance & Reliability

• Implement and optimize Monitoring, Logging, Alarms, APM, Tracing, and Log Analytics in OCI.

• Capacity plans, and performance baselines; drive performance tuning of compute, networking, databases, and storage.

4) Security, Compliance & Risk

• Enforce OCI security best practices: IAM least privilege, vaults/keys, secrets management, Cloud Guard, vulnerability scanning, CIS benchmarks, and Security Zones.

• Partner with GRC teams on audit readiness, regulatory compliance (e.g., ISO 27001, SOC 2, PCI DSS), data residency, and incident response tabletop exercises.

• Drive patching baselines, image hardening, and secure configuration drift detection.

5) Cost Management & FinOps

• Implement tagging, budgets, usage reports, and cost policies; recommend rightsizing, storage tiers, autoscaling, and reservations/committed use discounts.

• Run monthly cost reviews and produce optimization recommendations; integrate with FinOps dashboards/tools.

6) Migration & Modernization

• Lead migrations into OCI (re-host, re-platform, re-architect) for workloads including Oracle Databases, app servers, microservices, and data pipelines.

• Guide adoption of managed services (Autonomous DB, OKE, Streaming, Functions, API Gateway, Data Integration) and container strategies.

7) Stakeholder Leadership & Mentoring

• Serve as technical lead for cross-functional projects; translate business needs into robust cloud designs.

• Mentor L1/L2 engineers; deliver runbooks, playbooks, and capability uplift training.

• Collaborate with DBAs, App Owners, Security, Network, DevOps, and Product teams for end-to-end outcomes.

________________________________________

Required Qualifications

• 8–12+ years in cloud/infra engineering; 4+ years hands-on with OCI at scale.

• Deep expertise across OCI core services: Compute, VCN/Networking, Block/Object/Archive Storage, Load Balancer, WAF, IAM, Cloud Guard, Logging/Monitoring, OKE/Kubernetes, Autonomous DB/Exadata Cloud.

• Strong in scripting (Python/Bash/PowerShell)

• Solid understanding of ITIL and incident/problem/change processes.

• Proven experience with HA/DR architectures, performance tuning, and cost optimization.

• Hands-on with security hardening, compliance frameworks, and audit support.

________________________________________

Preferred Certifications (Nice to Have)

• OCI Architect Professional

• OCI Cloud Operations Associate / OCI Security Professional

• Oracle Autonomous Database / Exadata Cloud certifications

• CKA/CKAD (Kubernetes), Terraform Associate

• ITIL v4 Foundation/Managing Professional

________________________________________

Technical Stack (Representative)

• Cloud: OCI (Tenancy, Compartments, IAM, Policies, Tags, Budgets)

• Compute/Containers: Compute instances, OKE, OCI Registry, Functions

• Networking: VCN, Subnets, DRG, NAT, Service Gateway, VPN, FastConnect, NSG, WAF, Load Balancer

• Storage/DB: Block/Object/Archive, File Storage, Autonomous DB, Exadata Cloud Service, Data Guard

• Observability: OCI Monitoring, Alarms, Logging, Log Analytics, APM

• Security: Cloud Guard, Security Zones, Vault, KMS, IAM, Policies

• Automation: Terraform, Ansible, Python/Bash/PowerShell, OCI CLI/SDK

• ITSM: Remedy/ServiceNow/Jira (incidents, changes, CMDB), Confluence/Wiki

________________________________________

Soft Skills & Attributes

• Systems thinking, strong analytical and troubleshooting skills.

• Clear communication (can articulate trade-offs and risk).

• Ownership mindset; calm under pressure during Major Incidents.

• Collaborative leadership and mentorship; ability to influence without authority.

________________________________________

Typical Day / Week

• Morning: Review alarms, dashboards, capacity & cost trends; act on exceptions.

• Daytime: Lead solution designs, review plan of actions, mentor engineers, handle L3 escalations.

• Weekly: Architecture council, change advisory board (CAB), cost/security review, RCA review.

• Monthly/Quarterly: DR drills, failure mode analysis, audit support, roadmap updates.

________________________________________

Education

• Bachelor’s/Master’s in Computer Science, Information Technology, or equivalent experience.


Read more
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos

About SSA Company

Founded
Type
Size
Stage

About

N/A

Company social profiles

N/A

Similar jobs (10)

company logo
Ganesh Ram
Posted by Ganesh Ram
Bengaluru (Bangalore), Mumbai, Delhi, Gurugram, Noida, Ghaziabad, Faridabad, Hyderabad, Pune
7 - 10 yrs
₹15L - ₹20L / yr
CI/CD
skill iconKubernetes
helm
Terraform
yaml

Cloud Expertise(Azure):

• Strong understanding of cloud services and resources like AI services, webapp, database, including monitoring tools like Azure Monitor and Log Analytics.

• Experience with Infrastructure as Code (IaC) tools such as Arm template / Bicep/Terraform.

• Deep understanding of Networking concepts(DNS, DHCP , Hub and Spoke).

• Understanding on policies and security aspects of cloud.


Kubernetes & Helm:

• In-depth knowledge of Kubernetes concepts such as pods, services, ingress, config maps, and secrets.

• Understand of Kubernetes templates and its deployment.

• Proficiency with Helm/ Kustomize or equivalent for Kubernetes package management and deployment automation.

• Implement Kubernetes best practices, including security, networking, and scaling.

• Concepts of Docker and Containers


CI/CD & Programming:

• Hands-on experience with YAML-based CI/CD pipelines (e.g., Azure DevOps, GitHub Actions).

• Familiarity with scripting and automation tools such as PowerShell, Azure CLI, or Bash.

• Proven skill in python programming and concepts.


Monitoring and Observability:

Expertise in creating and managing Grafana dashboards for visualizing metrics and logs.

• Knowledge of Log Analytics & Azure Application Insights for performance monitoring and tracing.

 

Read more
company logo
Atharva K
Posted by Atharva K
Pune
7 - 10 yrs
₹30L - ₹45L / yr
skill iconAmazon Web Services (AWS)
IDP
Terraform

Platform Engineering Lead (For client company)

Location: Pune, India

Experience: 7+ years 


What Success Looks Like

  • Engineering teams ship faster with confidence and built-in guardrails.
  • Cloud cost, security, and reliability are predictable, measurable, and well-managed.
  • CI/CD pipelines are trusted, standardized, and production-ready.
  • Platform decisions reduce cognitive load instead of introducing unnecessary process. 

Scope & Expectations

This is a hands-on leadership role combining architecture and implementation.

You will:

  • Build, not just review.
  • Own the platform roadmap—not just infrastructure tickets.
  • Act as a force multiplier for product engineering teams rather than becoming a bottleneck.
  • Drive platform strategy while remaining deeply involved in execution.

Key Responsibilities

Platform & Cloud Architecture

  • Own Zoop's platform and cloud architecture across GCP and AWS.
  • Design reusable, opinionated platform patterns instead of one-off infrastructure.
  • Build and evolve Zoop's Internal Developer Platform (IDP), including:
  • Self-service environments
  • Golden paths (paved roads)
  • Standardized templates
  • Built-in engineering guardrails
  • Lead Kubernetes and cloud-native adoption at scale.
  • Drive infrastructure automation using Terraform, Pulumi, or similar Infrastructure-as-Code (IaC) tools.

CI/CD, Reliability & Developer Experience

  • Establish robust CI/CD practices with quality gates and production readiness.
  • Improve deployment safety through automation and testing.
  • Define and monitor:
  • Golden Signals
  • SLIs
  • SLOs
  • Incident response processes
  • Reduce operational toil and improve developer productivity.
  • Make observability a first-class capability using cost-efficient monitoring systems.
  • Build an observability platform that multiple engineering teams can easily integrate into their applications.

Security, Privacy & Compliance

  • Build security-by-default into infrastructure and deployment pipelines.
  • Lead implementation and continuous compliance for:
  • DPDP Act (India)
  • ISO 27001:2022
  • SOC 2 Type II
  • Implement:
  • Zero Trust architecture
  • Least-privilege access
  • Secure data isolation

FinOps & Cloud Optimization

  • Make cloud costs transparent and accountable across engineering teams.
  • Establish FinOps practices including:
  • Budgets
  • Cost alerts
  • Optimization routines
  • Drive build-vs-buy decisions using clear ROI analysis.

AI, Data & MLOps Foundations

  • Build secure and scalable foundations for AI and MLOps workloads.
  • Define guardrails for AI systems and sensitive data handling.

Leadership & Collaboration

  • Partner closely with engineering teams to align infrastructure strategy with product goals.
  • Mentor engineers and guide teams through technical change.
  • Balance long-term platform initiatives with practical execution.

What We're Looking For

Experience

  • 7+ years of experience building and operating production infrastructure.
  • Experience scaling engineering platforms in high-growth or regulated companies.
  • Strong hands-on expertise in:
  • Kubernetes and the cloud-native ecosystem
  • Service Mesh technologies
  • Policy Engines
  • GCP, AWS (Azure exposure is a plus)
  • Terraform and Infrastructure as Code

Engineering & Operations

  • Strong understanding of SDLC and modern CI/CD systems (Jenkins, GitOps, etc.).
  • Experience with observability tools such as:
  • Grafana
  • Prometheus
  • New Relic
  • Comfortable reading and contributing to production systems written in:
  • Go
  • Python
  • Node.js

Security & Compliance

  • Practical experience implementing ISO 27001 and SOC 2 controls.
  • Strong understanding of:
  • Data protection
  • Privacy
  • Identity and access management
  • Security best practices

Mindset

We're looking for someone who is:

  • Action-oriented with sound engineering judgment.
  • Analytical, cost-conscious, and reliability-focused.
  • Collaborative, calm under pressure, and open to feedback.
  • Comfortable challenging decisions and explaining trade-offs when necessary.

Nice to Have

  • Experience in fintech, identity, or other regulated industries.
  • Built Internal Developer Platforms (IDPs) or shared infrastructure tooling.
  • Contributions to open-source projects.







Read more
company logo
Pune
7 - 12 yrs
Best in industry
Google Cloud Platform (GCP)
Terraform
skill iconKubernetes
GKE
Reliability engineering
+1 more

About Searce

Searce is a global, AI-native, engineering-led technology consultancy and a Premier Google

Cloud Partner — recognized as the Google Cloud Workplace AI Transformation Partner of the

Year, APAC (2026). With 20+ years of experience and 3,000+ clients across 10+ countries, we

help businesses stay ahead of the cloud curve.


The Role

We're looking for a Lead Cloud Security & Reliability Engineer with deep GCP expertise to own

end-to-end cloud reliability and security forAPAC enterprise clients. As Lead, you'll set the architectural direction, mentor your squad, and drive measurable client outcomes across multi-

cloud environments.


What You'll Do

Own Client Delivery — Lead 24x7 GCP cloud operations forAPAC clients. Define SLO frameworks and ensure adherence.


Architect Solutions — Design scalable, secure GCP-primary architectures with multi-cloud awareness.


Drive Reliability — Lead incident response, RCA, and long-term remediation across production systems.


Mentor & Elevate — Coach and grow a squad of Senior CSREs.


Drive FinOps — Own cloud cost governance and optimization with quantified impact.


Be the Expert — Represent Searce's technical depth in global client conversations.


What We're Looking For

Experience

7–12 years total with 5+ years on GCP cloud infrastructure

Strong background in Cloud Managed Services / MSP environments

Proven experience leading a team in client-facing delivery

Multi-cloud exposure (AWS/Azure secondary) preferred


Technical Skills (Must-Have)

  • GCP: GKE, IAM, VPC, Cloud Monitoring, Stackdriver, KMS — demonstrated in work
  • experience
  • Kubernetes: GKE — production cluster management, Helm
  • IaC: Terraform — module-level, reusable frameworks
  • Observability: Prometheus, Grafana, Thanos or equivalent
  • Security: IAM, Zero-trust, DevSecOps, CSPM tools
  • Scripting: Python or Go
  • FinOps: GCP cost governance demonstrated


Nice to Have

  • GCP Professional Cloud Architect / Pro DevOps Engineer certification
  • AWS / Azure secondary experience
  • CKA (Certified Kubernetes Administrator)
  • ITIL / change management awareness
  • APAC client delivery experience


Why Searce?

🏆 Google Cloud Partner of the Year — APAC 2026

🌍 Work with APAC enterprise clients across multiple industries

🤖 AI-first, engineering-led culture

📈 Lead-level ownership with real career growth

🤝 HAPPIER values — Humble, Adaptable, Positive, Passionate, Innovative, Excellence,

Responsible

Read more
company logo
Remote only
10 - 15 yrs
₹30L - ₹35L / yr
skill iconAmazon Web Services (AWS)
Microsoft Windows Azure
Generative AI
Implementation
System deployment
+9 more

Job Description: Lead - Cloud Engineering (AWS / Azure)

Role Title: Lead - Cloud Engineering

Experience Level: 10+ Years

Domain Focus: Healthcare AI & Cloud Infrastructure

Location: Remote

Job Overview

We are seeking an experienced Lead - Cloud Engineering with over 10 years of IT experience to lead our cloud strategy, architecture, and infrastructure teams. In this role, you will oversee end-to-end cloud deployment, multi-cloud migration, and scalable architecture designed to support cutting-edge Generative AI applications in the healthcare technology domain.

The ideal candidate brings deep technical expertise in both AWS and Azure, strong hands-on capability in cloud infrastructure, and proven leadership experience driving security, compliance, and team growth.

Key Responsibilities

Cloud Architecture & Migration

  • Lead the architecture, design, and execution of cloud migrations, deployments, and modernizations across AWS and Azure environments.
  • Drive Infrastructure as Code (IaC) standards using Terraform, CloudFormation, or Bicep to ensure scalable, automated infrastructure provisioning.
  • Build high-availability, low-latency architectures optimized for data-intensive Generative AI and Machine Learning workloads.

Security & Healthcare Compliance

  • Enforce healthcare security standards including HIPAA, HITRUST, SOC 2, and data governance best practices across all cloud assets.
  • Implement Zero-Trust security, Identity Access Management (IAM), data encryption key management, and continuous vulnerability monitoring.

Leadership & Team Management

  • Manage, mentor, and scale a high-performing team of DevOps, Cloud, and SRE Engineers.
  • Drive Agile workflows, sprint planning, incident response frameworks, and SLA compliance.
  • Collaborate closely with Data Engineering, AI/ML, and Software Product teams to align infrastructure with business roadmaps.

Operations & FinOps

  • Establish cloud cost optimization strategies (FinOps) to manage computing costs associated with AI models and large-scale data processing.
  • Manage monitoring, alerting, and telemetry frameworks (e.g., Prometheus, Datadog, CloudWatch) to ensure 99.99% uptime.

Key Requirements

  • Experience: 10+ years of overall IT experience with at least 5+ years in a cloud leadership or lead architect role.
  • Cloud Platforms: Advanced hands-on expertise with both AWS (e.g., EC2, S3, EKS, Bedrock, SageMaker) and Azure (e.g., AKS, Azure OpenAI, Blob, Virtual Machines).
  • DevOps & IaC: Strong background in Terraform, Docker, Kubernetes, CI/CD pipelines (GitHub Actions, GitLab CI, or Jenkins).
  • Domain Knowledge: Prior experience building or managing cloud environments within Healthcare, Life Sciences, or HealthTech is strongly preferred.
  • AI/ML Familiarity: Experience supporting cloud infrastructure for machine learning pipelines, LLM deployments, or GPU compute management.
  • Certifications (Preferred): AWS Certified Solutions Architect – Professional, Azure Solutions Architect Expert, or Certified Kubernetes Administrator (CKA).


Read more
MNC
MNC
Agency job
via by Gauri Naik
Remote only
10 - 18 yrs
Best in industry
Landing page optimization
skill iconKubernetes
Terraform
CI/CD

🚀 Hiring: Senior Azure Platform Engineer | Azure | Terraform | Kubernetes


📍 Location: India

💼 Employment: Full-Time | Long-Term

🎯 Experience: 10+ Years | 8+ Years Hands-on Azure


🔑 What We’re Looking For

▪️ Strong expertise in Azure Cloud Platform & Landing Zone Architecture

▪️ Advanced hands-on experience with Terraform & reusable IaC modules

▪️ Expertise in Kubernetes, Helm & container platforms

▪️ Strong understanding of Azure Networking, Entra ID, RBAC & Azure Policy

▪️ Experience with GitHub, GitHub Actions / Azure Pipelines & CI/CD

▪️ Hands-on GitOps & ArgoCD experience

▪️ Strong observability skills with Prometheus, Grafana & Azure Monitor

▪️ Experience designing and supporting microservices architectures

▪️ Knowledge of security, policy-as-code and IaC security scanning

▪️ Exposure to Azure Arc / Azure Local / Azure Stack HCI is a plus

Read more
company logo
Lakshit Bagga
Posted by Lakshit Bagga
Remote only
8 - 30 yrs
₹1L - ₹60L / yr
DevOps
skill iconAmazon Web Services (AWS)
prometheus
skill icongrafana
Terraform

We're hiring a Cloud Architect (Contract) to work with our Equity Partners who builds profitable growth by acquiring and operating enterprise software companies. Refining a proprietary operating model across 40+ acquisitions and two decades of hands-on experience, now supercharged by our patented agentic AI platform . In this role, you'll take full architectural control of our CI/CD, observability, and event streaming infrastructure, build the standards every new acquisition plugs into, and use AI-assisted automation to keep 20+ products reliable without proportionally scaling headcount.


Job title: Cloud/Platform Architect (SRE)

Type: Global Remote | Contract


What You Bring

  • 8–12 years in platform engineering, DevOps, or SRE, with growing ownership over time
  • Deep Terraform experience across multi-account, multi-env setups
  • Real production experience with event streaming at scale
  • Hands-on Grafana, Prometheus, Loki, and strong AWS depth (ECS, EKS, IAM, VPC, RDS)
  • SRE fundamentals: SLOs, error budgets, on-call design, post-mortems
  • Bonus: acquisition or greenfield platform-building experience


Roles and Responsibilities

  • Own everything outside core AWS infra: CI/CD, observability, event streaming, deployment, incidents
  • Define the standards every future acquisition will plug into
  • Keep 20+ enterprise products running at serious scale (millions–billions of requests)
  • Build self-service tooling so product teams never wait on you
  • Use AI/automation to kill toil — not to replace engineering judgement


Ready to build the platform that scales an entire portfolio? — let's connect.

Read more
company logo
Shakthi M
Posted by Shakthi M
Bengaluru (Bangalore)
5 - 14 yrs
Best in industry
skill iconPython
Azure
Terraform
DevOps
  • Strong hands-on experience in Microsoft Azure Cloud.
  • Good understanding of Azure services such as Compute, Storage, Event Hub, Event Subscription, Storage Queue, and PaaS services.
  • Basic understanding of Azure AI Foundry and AI-related Azure service setup.
  • Good Azure networking basics: VNet, subnet, routing, and basic troubleshooting.
  • Strong knowledge of Terraform, especially:
  • Terraform state
  • plan / apply
  • troubleshooting failures
  • migration risks
  • Terraform Enterprise concepts
  • Strong Python coding capability, not just basic scripting.
  • Experience using Python for API integration, automation, JSON/YAML handling, and internal tooling.
  • Good understanding of CI/CD pipelines.
  • Ability to troubleshoot pipeline failures.
  • Comfortable with YAML and JSON.
  • Ability to troubleshoot Azure infrastructure/platform issues.
  • Ability to collect logs/evidence and coordinate with network/app/Microsoft support teams.
  • Basic awareness of agentic AI / LLM concepts.
  • Awareness of security and cost best practices.

Good to Have Skills

  • Hands-on experience with Harness.
  • Hands-on experience with Terraform Enterprise.
  • Exposure to LangGraph / LangChain.
  • Exposure to agentic AI workflows or skill creation.
  • Exposure to Claude or enterprise LLM integrations.
  • Knowledge of Azure ML Workspace, model registry, and managed endpoints.
  • MLOps / LLMOps knowledge.
  • FinOps / Azure cost optimization experience.
  • Azure certifications: AZ-104, AZ-305, AZ-400, AZ-500.

 

Screening Priority:

Azure Cloud + Terraform + Python Coding + CI/CD Troubleshooting + YAML/JSON + Basic Agentic AI Awareness

 

Read more
It is an Product Based Company(Domain- EV Charging)
It is an Product Based Company(Domain- EV Charging)
Agency job
via by Mantasha Naaz
Bengaluru (Bangalore)
0.6 - 2 yrs
₹8L - ₹10L / yr
Site Reliability Engineer
Reliability engineering
skill iconAmazon Web Services (AWS)
Terraform
Ansible
+3 more

Job Title: TechOps Engineer

Location: Bengaluru, India (Hybrid)

Employment Type: Full-time

Experience: 6 Month-2 years (Excluding Internship)

Shift Timing: 2 PM to 11 PM IST


Role Overview

We are excited to find a highly engaged engineer who is obsessed with technology that wants to be a part of a “world class” platform SRE team. Engineers must possess an "automation first" mindset, with a relentless focus on documentation, quality, scalability, and reliability using Infrastructure as Code tools. This position will be part of a platform team that is developing exciting products and solutions and playing a key part in driving forward the electrification of transportation.

What you’ll do:  

  • Ensure system reliability, uptime, and performance of global platform.
  • Conduct real-time surveillance of our EV charging systems to proactively identify and mitigate performance issues and anomalies near 24/7 basis. As such, you collaborate with IDT and FMC players to ensure incident detection also happens outside office hours (monitoring shifts among team members subject to duty schedule) 
  • Deliver on change & releases like firmware changes and drive insights & intelligence back into testing processes and tech discussions with the wider organization. 
  •  Successfully deliver and project manage first time right commissioning activities alongside our Engineering Procurement Contract Management (EPCM) partners to successfully bring charge points onto our Charge Point Management System (CPMS).
  • End-to-end EV charger lifecycle management, including deployment, commissioning, monitoring, maintenance, and decommissioning activities.
  • Provide technical guidance and support to DC specialists during the commissioning of EV charging solutions.
  •  Work closely with Shell, Engineering, and IT colleagues to ensure projects are completed on time and to specification.
  • Act as a liaison with the Engineering Procurement Contract Management (EPCM) partner to manage projects from start to finish, ensuring charge points are successfully onboarded on the Charge Point Management System (CPMS).
  • Collaborate with development, operations and support  teams to build scalable and resilient systems.
  • Contribute to incident response, root-cause analysis, and post-mortem reviews, driving continuous improvement.
  • Participate in capacity planning, performance tuning, and resource optimization.
  • Integrate security and compliance best practices into all infrastructure operations.
  • Stay current with emerging SRE tools, frameworks, and cloud technologies to continuously improve reliability practices.
  • Participate in and lead on-call rotations and incident response, conducting detailed postmortems and RCA reports.
  • Flexible to resolve blocking issues during off hours or weekends if required.  

 

What We’re Looking For: 

Basic Qualifications and skills

  • Bachelor’s degree in Engineering , Electrical, ECE, Computer Science, Information Technology, or related field.
  • Overall 1 years of experience as a Site Reliability Engineer, Technical project coordinator role.
  • Proven experience of SRE or Technical Project Coordination with IoT or connected devices based platforms.
  • Experience with incident management and on-call best practices. Provide support to on call engineers.
  • Excellent analytical and problem-solving skills with a proactive mindset. 
  • Hands-on experience with AWS Cloud and IaC tools such as Terraform or Ansible.
  • Expertise with monitoring and observability tools (Dynatrace,Prometheus, Grafana, Zabbix, etc.).
  • Proactively monitor the network, triage performance outliers, and coordinate correction actions to ensure optimal system functionality.
  • Fluency in English (spoken and written). 
  • Successfully recommission or decommission chargers following changes in our network.
  • Responsible for the go-live of the chargers on Shell’s public network following commissioning attempts.

 Note: This role involves managing infrastructure for a global platform operating in over ten countries, requiring effective communication and collaboration across regions. Strong verbal and written communication skills, along with availability and flexibility to resolve blocking issues, are essential to support On-call Engineers. This role may involve EU or US time‑zone shifts based on business requirements. The shift timing will be 2 PM IST to 11 PM IST.  


What We Offer

  • Work with some of the brightest minds in the emerging EV industry.
  • Make a tangible impact in reducing carbon emissions and enabling sustainable energy.
  • Freedom to suggest, implement, and innovate on systems, processes, and technologies.
  • Daily ownership in a high-growth, challenging environment.
  • Flexible work environment with hybrid schedules and virtualization options.
  • Competitive pay and benefits including health coverage, innovative PTO program, and performance bonuses.


Read more
company logo
Swathi S
Posted by Swathi S
Chennai
7 - 12 yrs
₹30L - ₹55L / yr
skill iconAmazon Web Services (AWS)
skill iconPython
CI/CD
DevOps
Platform as a Service (PaaS)
+7 more

Amura’s Vision 


We believe that the most under-appreciated route to releasing untapped human potential is to build a healthier body, and through which a better brain. This allows us to do more of everything that is important to each one of us.


Billions of healthier brains, sitting in healthier bodies, can take up more complex problems that defy solutions today, including many existential threats, and solve them in just a few decades.


Billions of healthier brains will make the world richer beyond what we can imagine today. The surplus wealth, combined with better human capabilities, will lead us to a new renaissance, giving us a richer and more beautiful culture.


These healthier brains will be equipped with deeper intellect, be less acrimonious, more magnanimous, and have a kinder outlook on the world, resulting in a world that is better than any previous time.

We find this vision of the future exhilarating. Our hopes and dreams are to create this future as quickly as possible and ensure that it is widely distributed and optimized to maximize all forms of human excellence. 


Role Overview 


We are looking for a highly skilled Senior DevOps Engineer (AI-Native Infrastructure & Platform Engineering) with deep expertise in AWS cloud infrastructure, automation, AI infrastructure operations, and modern DevOps/SRE practices.


This role goes beyond traditional DevOps and requires a seasoned specialist capable of building and operating AI-ready infrastructure platforms that support high-throughput APIs, LLM/AI workloads, GPU-based compute, data-intensive systems, real-time inference pipelines, and scalable ML platforms.


You will be responsible for architecting, automating, securing, and optimizing highly scalable and cost-efficient cloud environments that enable high-velocity engineering and AI teams. This is an ideal position for someone who combines technical ownership, an automation-first mindset, and a passion for developer productivity and platform reliability. 


Key Responsibilities 


Cloud Infrastructure & Platform Engineering (AWS) 

  • Architect, deploy, and manage highly scalable and secure infrastructure on AWS. Design cloud platforms supporting AI/ML workloads, data pipelines, real-time APIs, and high-concurrency backend systems.
  • Hands-on expertise with key AWS services including EC2, ECS/EKS, Lambda, RDS, DynamoDB, S3, VPC, CloudFront, IAM, CloudWatch, and GPU-enabled instances.
  • Build and maintain Infrastructure-as-Code (IaC) using Terraform, CloudFormation, or AWS CDK.
  • Design multi-AZ and multi-region architectures for high availability and disaster recovery (HA/DR).
  • Build reusable platform templates and shared infrastructure modules. 


AI/ML Infrastructure & MLOps 

  • Build and maintain infrastructure for LLM applications, AI inference workloads, model serving platforms, vector databases, and feature stores.
  • Support GPU-based workloads and optimize compute/storage usage.
  • Enable scalable deployment patterns for AI applications using Kubernetes/EKS. Collaborate with Data Science and ML Engineering teams on model deployment, training/tuning of models, CI/CD for ML systems, experiment environments, and reproducibility.
  • Support orchestration and deployment of AI workflows and inference services while implementing observability and reliability for AI pipelines. 


CI/CD, Automation & Developer Productivity 

  • Build and maintain CI/CD pipelines using GitHub Actions, GitLab CI, Jenkins, or AWS CodePipeline.
  • Automate deployments, environment provisioning, and release workflows.
  • Build self-service developer platforms, preview environments, and reusable deployment workflows to improve developer productivity.
  • Implement automated patching, scaling, backups, cleanup workflows, and drift detection. 


Containers, Kubernetes & Platform Reliability

  • Manage Docker-based environments, containerized applications, and optimize workloads using Kubernetes (EKS) or ECS/Fargate.
  • Manage autoscaling, cluster health, node pools, ingress, service mesh, and workload isolation.
  • Optimize infrastructure for performance, resilience, and cost-efficiency.
  • Implement progressive deployment strategies including blue/green, canary, and rolling deployments. 


Observability, Incident Response & SRE Practices

  • Implement observability stacks using CloudWatch, Prometheus, Grafana, ELK, Datadog, OpenTelemetry, or New Relic.
  • Build actionable dashboards and intelligent alerting systems while defining and tracking SLIs, SLOs, and SLAs.
  • Lead incident response, root cause analysis, and blameless postmortems to reduce operational toil and improve MTTR.

FinOps, Cost Governance & Security

  • Continuously monitor and optimize cloud costs (compute utilization, storage lifecycle, GPU usage, and data transfer) using AWS Cost Explorer, Budgets, Trusted Advisor, CloudHealth, or Kubecost.
  • Implement AWS security best practices for IAM, VPCs, security groups, NACLs, encryption, and manage secrets using KMS, SSM Parameter Store, or Vault.
  • Build secure CI/CD pipelines with automated security checks, least-privilege access, audit logging, and ensure compliance readiness for ISO 27001, SOC2, and GDPR.

Collaboration, Leadership & Platform Culture

  • Work closely with engineering, AI/ML, QA, product, and operations teams to drive a DevOps, SRE, GitOps, and automation-first culture.
  • Mentor junior DevOps and Platform Engineers while creating and maintaining detailed runbooks, architecture diagrams, and platform documentation.

Skills & Qualifications


Must-Have:

  • 7+ years of experience in DevOps, SRE, Platform Engineering, or Cloud Infrastructure Engineering.
  • Strong expertise in AWS cloud architecture, services, and deep understanding of Kubernetes (EKS), containers, and cloud-native systems.
  • Strong Infrastructure-as-Code expertise using Terraform, CloudFormation, or CDK. Strong Linux administration, networking, DNS, routing, and load balancing knowledge. Strong scripting/programming experience in Python, Bash, or Go (preferred). Experience with CI/CD automation, GitOps workflows, and observability platforms supporting scalable production systems.


Preferred / Nice-to-Have:

  • Experience with AI/ML infrastructure, MLOps, model serving, vector databases, GPU orchestration, and inference optimization.
  • Familiarity with Kafka, Redis, SQS, and event-driven systems.
  • Exposure to platform engineering, internal developer platforms, and tools like ArgoCD, Flux, Helm, and OpenTelemetry.
  • AWS Certifications: Solutions Architect, DevOps Engineer, or SysOps Administrator. Knowledge of distributed systems and large-scale platform operations. 


Preferred / Nice-to-Have:

  • Experience with AI/ML infrastructure, MLOps, model serving, vector databases, GPU orchestration, and inference optimization.
  • Familiarity with Kafka, Redis, SQS, and event-driven systems.
  • Exposure to platform engineering, internal developer platforms, and tools like ArgoCD, Flux, Helm, and OpenTelemetry.
  • AWS Certifications: Solutions Architect, DevOps Engineer, or SysOps Administrator. Knowledge of distributed systems and large-scale platform operations. 


Here are answers to some questions you may have

Where is your office?

Chennai (Velachery)

Work Model

Work from Office – because great stories are built in person!

Do you have an online presence?

https://amura.ai (we are @AmuraHealth on all social media)


Read more
company logo
Ashish Singh
Posted by Ashish Singh
Remote only
0 - 1 yrs
₹12000 - ₹18000 / mo
CI/CD
DevOps

Build production-grade cloud infrastructure that powers enterprise applications with cutting-edge DevOps practices.


What you'll do:

  • Design CI/CD pipelines (GitHub Actions, Jenkins, GitLab CI)
  • Containerize apps with Docker, deploy on Kubernetes clusters
  • Manage infrastructure as code (Terraform, CloudFormation)
  • Set up monitoring (Prometheus, Grafana, ELK Stack)
  • Cloud migrations (AWS EC2, EKS, RDS → GCP equivalent)
  • Optimize costs and performance for live production systems

What we need:

  • Basic Python/Bash scripting
  • Docker basics, Git workflows
  • Cloud exposure (AWS/GCP/Azure free tier projects)
  • Problem-solving mindset, eagerness to learn

Real impact:

  • Deploy apps used by 1000+ daily users
  • Work with senior DevOps engineers on client deliverables
  • Build portfolio for FAANG-level interviews


Read more
Why apply to jobs via Cutshort
people_solving_puzzle
Personalized job matches
Stop wasting time. Get matched with jobs that meet your skills, aspirations and preferences.
people_verifying_people
Verified hiring teams
See actual hiring teams, find common social connections or connect with them directly.
ai_chip
Move faster with AI
We use AI to get you faster responses, recommendations and unmatched user experience.
Did not find a job you were looking for?
icon
Search for relevant jobs from 10000+ companies such as Google, Amazon & Uber actively hiring on Cutshort.
companies logo
companies logo
companies logo
companies logo
companies logo
Get to hear about interesting companies hiring right now
Company logo
Company logo
Company logo
Company logo
Company logo
Linkedin iconFollow Cutshort
Users love Cutshort
Read about what our users have to say about finding their next opportunity on Cutshort.
Shubham Vishwakarma's profile image

Shubham Vishwakarma

Full Stack Developer - Averlon
I had an amazing experience. It was a delight getting interviewed via Cutshort. The entire end to end process was amazing. I would like to mention Reshika, she was just amazing wrt guiding me through the process. Thank you team.
Companies hiring on Cutshort
companies logos