AWS FinOps Specialist/Engineer at NeoGenCode Technologies Pvt Ltd · Gurugram · 3 - 8 years · ₹5L - ₹16L / yr · Raised funding · Posted 24 Dec 2024

Job Title : AWS FinOps Specialist/Engineer
Experience : 3+ Years
Location : Gurgaon WFO (5 Days)
Summary :
We are seeking an experienced AWS FinOps Specialist/Engineer to manage and optimize cloud financial operations across our AWS environment. The ideal candidate will have strong analytical skills, deep expertise in AWS cost management tools. He would work closely with engineering, finance, and leadership teams to provide visibility into cloud spending and implement strategies to maximize ROI from AWS investments.
Key Responsibilities :
1. Cost Optimization :
- Analyze AWS usage and costs to identify cost-saving opportunities.
- Implement AWS Reserved Instances, Savings Plans, and Spot Instances to reduce costs.
- Right-size AWS resources based on application needs.
- Propose architectural changes to improve cost efficiency.
2. Cloud Financial Governance :
- Set up budgets and alerts in AWS Cost Explorer and Billing Dashboard.
- Define and enforce tagging strategies for cost allocation and resource tracking.
3. Reporting and Analysis :
- Create detailed cost reports and dashboards using AWS tools, such as Cost Explorer, AWS Budgets, and QuickSight.
- Provide actionable insights to stakeholders for better decision-making.
Track key KPIs, including cost per service, cost per account, and overall cloud ROI.
4. Collaboration :
- Work with teams to forecast cloud resource usage and budget accordingly.
Collaborate with finance to align cloud spending with organizational goals.
Partner with vendors and AWS account managers to optimize enterprise agreements.
5. Automation and Tooling :
Automate cost management workflows using AWS tools like Cost and Usage Reports (CUR) and third-party FinOps platforms.
Key Requirements :
- Experience: 3-5 years in AWS cloud operations, cost management, or a related field.
Certifications : AWS Certified Cloud Practitioner or AWS Certified Solutions Architect (Associate or Professional) preferred.
Technical Skills :
- Proficiency in AWS Billing, Cost Explorer, and AWS Cost and Usage Reports (CUR).
- Familiarity with AWS services (e.g., EC2, S3, RDS, Lambda, EKS) and their cost structures.
- Experience with third-party FinOps tools (e.g., CloudHealth, Spot.io, CloudCheckr).

About NeoGenCode Technologies Pvt Ltd
About
Welcome to Neogencode Technologies, an IT services and consulting firm that provides innovative solutions to help businesses achieve their goals. Our team of experienced professionals is committed to providing tailored services to meet the specific needs of each client. Our comprehensive range of services includes software development, web design and development, mobile app development, cloud computing, cybersecurity, digital marketing, and skilled resource acquisition. We specialize in helping our clients find the right skilled resources to meet their unique business needs. At Neogencode Technologies, we prioritize communication and collaboration with our clients, striving to understand their unique challenges and provide customized solutions that exceed their expectations. We value long-term partnerships with our clients and are committed to delivering exceptional service at every stage of the engagement. Whether you are a small business looking to improve your processes or a large enterprise seeking to stay ahead of the competition, Neogencode Technologies has the expertise and experience to help you succeed. Contact us today to learn more about how we can support your business growth and provide skilled resources to meet your business needs.
Candid answers by the company
IT & Engineering Talent Staffing
- Provides full-time and contract-based hiring, delivering handpicked, pre‑screened developers across tech stacks—ranging from web, mobile, AI/ML, Web3/blockchain.
- Maintains a bench o vetted candidates, offering fast delivery of interview-ready profiles—often within 24 hours.
- Offers payroll management, handling compliance, tax, attendance, and documentation for both contractors and full-time employees.
2. End-to-End Project Delivery
- Delivers full-stack development solutions: web, mobile, cloud, AI/ML, Blockchain/Web3.
- Manages entire project lifecycle—requirements gathering, design (UI/UX), development, deployment, and ongoing support .
3. Additional Offerings
- Expands into cybersecurity consulting, digital marketing, and cloud platform services (like AWS, GCP, Azure) .
- Provides strategic IT consulting to align technology solutions with business objectives
Similar jobs (10)
Job Summary
We are looking for an experienced AWS Cloud Engineer with strong expertise in AWS infrastructure, deployment, migration, and cloud operations. The candidate should have hands-on experience managing AWS services such as EC2, EBS, S3, EFS, and FSx, along with infrastructure provisioning, migration, troubleshooting, and optimization.
Key Responsibilities
- Design, deploy, configure, and manage AWS infrastructure environments.
- Perform application and infrastructure migration to AWS.
- Provision and manage EC2 instances, including configuration, scaling, patching, and troubleshooting.
- Manage EBS volumes, snapshots, backups, and storage performance.
- Configure and administer S3 buckets, storage policies, lifecycle management, and access controls.
- Manage EFS for scalable shared file storage.
- Implement and manage Amazon FSx file systems based on application requirements.
- Monitor AWS infrastructure performance, availability, and capacity.
- Troubleshoot infrastructure, networking, storage, and deployment-related issues.
- Implement security best practices including IAM, security groups, encryption, and access controls.
- Support backup, disaster recovery, high availability, and business continuity requirements.
- Optimize AWS resources for performance, scalability, reliability, and cost.
Mandatory Skills
- Strong hands-on experience in AWS Cloud Infrastructure.
- Expertise in EC2, EBS, S3, EFS, and FSx.
- Experience in AWS deployment and migration projects.
- Strong knowledge of AWS networking concepts such as VPC, Subnets, Route Tables, Security Groups, and Load Balancers.
- Experience with IAM and AWS security best practices.
- Good knowledge of AWS monitoring and troubleshooting.
- Experience with cloud infrastructure automation using Terraform or CloudFormation is preferred.
- Strong Linux administration and troubleshooting skills.
- Good understanding of backup, disaster recovery, and high-availability concepts
Amura’s Vision
We believe that the most under-appreciated route to releasing untapped human potential is to build a healthier body, and through which a better brain. This allows us to do more of everything that is important to each one of us.
Billions of healthier brains, sitting in healthier bodies, can take up more complex problems that defy solutions today, including many existential threats, and solve them in just a few decades.
Billions of healthier brains will make the world richer beyond what we can imagine today. The surplus wealth, combined with better human capabilities, will lead us to a new renaissance, giving us a richer and more beautiful culture.
These healthier brains will be equipped with deeper intellect, be less acrimonious, more magnanimous, and have a kinder outlook on the world, resulting in a world that is better than any previous time.
We find this vision of the future exhilarating. Our hopes and dreams are to create this future as quickly as possible and ensure that it is widely distributed and optimized to maximize all forms of human excellence.
Role Overview
We are looking for a highly skilled Senior DevOps Engineer (AI-Native Infrastructure & Platform Engineering) with deep expertise in AWS cloud infrastructure, automation, AI infrastructure operations, and modern DevOps/SRE practices.
This role goes beyond traditional DevOps and requires a seasoned specialist capable of building and operating AI-ready infrastructure platforms that support high-throughput APIs, LLM/AI workloads, GPU-based compute, data-intensive systems, real-time inference pipelines, and scalable ML platforms.
You will be responsible for architecting, automating, securing, and optimizing highly scalable and cost-efficient cloud environments that enable high-velocity engineering and AI teams. This is an ideal position for someone who combines technical ownership, an automation-first mindset, and a passion for developer productivity and platform reliability.
Key Responsibilities
Cloud Infrastructure & Platform Engineering (AWS)
- Architect, deploy, and manage highly scalable and secure infrastructure on AWS. Design cloud platforms supporting AI/ML workloads, data pipelines, real-time APIs, and high-concurrency backend systems.
- Hands-on expertise with key AWS services including EC2, ECS/EKS, Lambda, RDS, DynamoDB, S3, VPC, CloudFront, IAM, CloudWatch, and GPU-enabled instances.
- Build and maintain Infrastructure-as-Code (IaC) using Terraform, CloudFormation, or AWS CDK.
- Design multi-AZ and multi-region architectures for high availability and disaster recovery (HA/DR).
- Build reusable platform templates and shared infrastructure modules.
AI/ML Infrastructure & MLOps
- Build and maintain infrastructure for LLM applications, AI inference workloads, model serving platforms, vector databases, and feature stores.
- Support GPU-based workloads and optimize compute/storage usage.
- Enable scalable deployment patterns for AI applications using Kubernetes/EKS. Collaborate with Data Science and ML Engineering teams on model deployment, training/tuning of models, CI/CD for ML systems, experiment environments, and reproducibility.
- Support orchestration and deployment of AI workflows and inference services while implementing observability and reliability for AI pipelines.
CI/CD, Automation & Developer Productivity
- Build and maintain CI/CD pipelines using GitHub Actions, GitLab CI, Jenkins, or AWS CodePipeline.
- Automate deployments, environment provisioning, and release workflows.
- Build self-service developer platforms, preview environments, and reusable deployment workflows to improve developer productivity.
- Implement automated patching, scaling, backups, cleanup workflows, and drift detection.
Containers, Kubernetes & Platform Reliability
- Manage Docker-based environments, containerized applications, and optimize workloads using Kubernetes (EKS) or ECS/Fargate.
- Manage autoscaling, cluster health, node pools, ingress, service mesh, and workload isolation.
- Optimize infrastructure for performance, resilience, and cost-efficiency.
- Implement progressive deployment strategies including blue/green, canary, and rolling deployments.
Observability, Incident Response & SRE Practices
- Implement observability stacks using CloudWatch, Prometheus, Grafana, ELK, Datadog, OpenTelemetry, or New Relic.
- Build actionable dashboards and intelligent alerting systems while defining and tracking SLIs, SLOs, and SLAs.
- Lead incident response, root cause analysis, and blameless postmortems to reduce operational toil and improve MTTR.
FinOps, Cost Governance & Security
- Continuously monitor and optimize cloud costs (compute utilization, storage lifecycle, GPU usage, and data transfer) using AWS Cost Explorer, Budgets, Trusted Advisor, CloudHealth, or Kubecost.
- Implement AWS security best practices for IAM, VPCs, security groups, NACLs, encryption, and manage secrets using KMS, SSM Parameter Store, or Vault.
- Build secure CI/CD pipelines with automated security checks, least-privilege access, audit logging, and ensure compliance readiness for ISO 27001, SOC2, and GDPR.
Collaboration, Leadership & Platform Culture
- Work closely with engineering, AI/ML, QA, product, and operations teams to drive a DevOps, SRE, GitOps, and automation-first culture.
- Mentor junior DevOps and Platform Engineers while creating and maintaining detailed runbooks, architecture diagrams, and platform documentation.
Skills & Qualifications
Must-Have:
- 7+ years of experience in DevOps, SRE, Platform Engineering, or Cloud Infrastructure Engineering.
- Strong expertise in AWS cloud architecture, services, and deep understanding of Kubernetes (EKS), containers, and cloud-native systems.
- Strong Infrastructure-as-Code expertise using Terraform, CloudFormation, or CDK. Strong Linux administration, networking, DNS, routing, and load balancing knowledge. Strong scripting/programming experience in Python, Bash, or Go (preferred). Experience with CI/CD automation, GitOps workflows, and observability platforms supporting scalable production systems.
Preferred / Nice-to-Have:
- Experience with AI/ML infrastructure, MLOps, model serving, vector databases, GPU orchestration, and inference optimization.
- Familiarity with Kafka, Redis, SQS, and event-driven systems.
- Exposure to platform engineering, internal developer platforms, and tools like ArgoCD, Flux, Helm, and OpenTelemetry.
- AWS Certifications: Solutions Architect, DevOps Engineer, or SysOps Administrator. Knowledge of distributed systems and large-scale platform operations.
Preferred / Nice-to-Have:
- Experience with AI/ML infrastructure, MLOps, model serving, vector databases, GPU orchestration, and inference optimization.
- Familiarity with Kafka, Redis, SQS, and event-driven systems.
- Exposure to platform engineering, internal developer platforms, and tools like ArgoCD, Flux, Helm, and OpenTelemetry.
- AWS Certifications: Solutions Architect, DevOps Engineer, or SysOps Administrator. Knowledge of distributed systems and large-scale platform operations.
Here are answers to some questions you may have
Where is your office?
Chennai (Velachery)
Work Model
Work from Office – because great stories are built in person!
Do you have an online presence?
https://amura.ai (we are @AmuraHealth on all social media)
About Shopalyst:
Shopalyst offers a Discovery Commerce platform for digital marketers. Combining data, AI and deep integrations with digital media and e-commerce platforms, Shopalyst connects people with products they love. More than 500 marquee brands leverage our SaaS platform for data driven marketing and sales in 30 countries across Asia, Europe and Americas. We have offices in Fremont CA, Bangalore, and Trivandrum. Our company is backed by Kalaari Capital.
About the Role:
- This position is responsible for end to end management of Shopalyst's cloud infrastructure platform (AWS and Google Cloud).
- Candidate should have a bachelor's degree in Computer Science and Engineering (or equivalent) plus minimum 3+ years of experience in managing AWS cloud infrastructure.
Core responsibilities:
Infra management
- Review infrastructure change requests, suggest & implement optimal solution in terms of performance, maintainability & cost
- Keep the infrastructure updated/upgraded. This includes servers, operating system, packages and application software
- Maintain all infrastructure configuration as IaC (Infrastructure as Code)
Infra monitoring and site reliability
- Continuous monitoring of infrastructure and ensuring 24x7 availability
- Implement/maintain automated alerting mechanisms for handling infra outages
- Document and maintain infra recovery procedures, ensure successful backup of critical data
- Continuously monitor infra utilization and alert on when to scale up/scale down instances
Infra Security, Compliance & Certifications
- Infrastructure access management - ensure least privilege as the default policy. Review infra access periodically. Manage access to critical systems like AWS/Open VPN
- Adherence to CIS/AWS security benchmarks on AWS Security Hub
- Implement controls required for different attestations/certifications like SOC 2 Type 2,ISO 27xxx, PCI DSS etc
DevOps
- Enable CI/CD process for application deployment
Infra cost monitoring and optimisation
- Services-wise infrastructure cost monitoring analysis and optimisation
Infra trends and new services
- Evaluate new cloud services, analyse feasibility of adoption
Required Skills and Experience:
- AWS: 3+ years experience with using a broad range of AWS technologies (e.g. EC2,S3, ELB, VPC,Route 53, IAM, CloudWatch, CloudFront, RDS, Lambda, Glacier) to develop and maintain an Amazon AWS based cloud solution
- DevOps: Solid experience as a DevOps Engineer in a 24x7 uptime Amazon AWS environment, including automation experience with configuration management tools.
- Infra automation: Experience in Terraform/Ansible
- Scripting Skills: Strong scripting (e.g. Python) and automation skills.
- Operating Systems: Linux system administration and strong shell scripting skills
- Monitoring Tools: Experience with web servers (e.g. Nginx).
- Problem Solving: Ability to analyze and resolve complex infrastructure resource and application deployment issues.
Desired Skills (Not essential but beneficial to have):
- DB Skills: Basic DB administration experience - Cassandra, RDS, MySQL
- Search engines: Experience with search engines such as Apache Solr/Elastic Search
- Version Control: Experience administrating version control systems such as Git
We are looking for a hands-on Senior AWS Cloud Engineer to lead the infrastructure build, optimization, automation, and production deployment of a Multi-Agent AI Chatbot Platform hosted on AWS. The development environment is already in place, and the successful candidate will drive the solution through testing, integrations, and production go-live.
Key Responsibilities
- Review, validate, and optimize existing Terraform code and AWS infrastructure.
- Establish and manage integrations with enterprise platforms such as ServiceNow, Workday, and other third-party systems.
- Design, build, and support secure, scalable, and highly available AWS environments.
- Implement and automate CI/CD pipelines and Infrastructure-as-Code practices.
- Lead infrastructure testing, performance tuning, and production readiness activities.
- Drive deployment and operationalization of the platform in the Production environment.
- Implement cloud governance, security, monitoring, and FinOps best practices.
- Troubleshoot and resolve complex cloud infrastructure issues.
Required Skills & Experience
- 10+ years of IT experience with strong expertise in AWS Cloud Engineering.
- Proven experience designing, deploying, and managing AWS production environments.
- Strong hands-on experience with Terraform and Infrastructure-as-Code.
- Experience with CI/CD pipeline automation and DevOps practices.
- Expertise in AWS services including VPC, IAM, EC2, S3, Lambda, CloudWatch, and networking.
- Experience in performance optimization, reliability, and cloud cost management (FinOps).
- Strong scripting and automation skills.
- Experience integrating enterprise applications through APIs and secure connectivity patterns.
DevOps & Cloud Security Specialist to architect enterprise-grade AWS networking, implement strict IAM security boundaries (HIPAA/GDPR compliance), automate infrastructure via IaC, and maintain crystal-clear technical documentation.
1. Primary Must-Have Competencies (Core Priorities)
A. AWS Networking & VPC Topology (Top Priority)
- Advanced VPC Architecture: Mastery in designing multi-VPC topologies, isolated subnets, custom Route Tables, NAT Gateways, Transit Gateways, and Cross-Region VPC Peering.
- Private Network Security: Extensive experience using VPC Endpoints (Gateway & Interface/PrivateLink) to keep internal AWS traffic completely off the public internet.
- Traffic Ingestion & Edge Security: Expertise in AWS WAF (custom rules, bot control, rate limiting), API Gateway throttling, ALB configuration, and Route 53 global routing.
B. AWS Security, IAM Architecture & Compliance
- Enterprise IAM Governance: Expertise in AWS Organizations, IAM Identity Center (SSO), Permission Boundaries, Service Control Policies (SCPs), and temporary role assumption across multi-account setups.
- Data Protection & Key Management: Deep knowledge of AWS KMS (customer-managed keys, envelope encryption at rest and in transit) and AWS Secrets Manager.
- Audit & Compliance: Setting up centralized logging pipelines (CloudTrail, GuardDuty, AWS Config, CloudWatch Audit Logs) for strict HIPAA/GDPR compliance.
C. Infrastructure as Code (IaC) & Containerization
- Terraform / AWS CDK: Must write production-grade, modular IaC templates from scratch—enforcing network topology and security guardrails directly in code.
- Container Orchestration: Hands-on setup and management of AWS ECS (Fargate) or EKS (Kubernetes) and automated CI/CD pipelines (GitHub Actions, GitLab CI).
D. Architecture Documentation & Systems Mapping
- Technical Documentation: Ability to author clean, standardized architecture diagrams (e.g., C4 model, Draw.io, Lucidchart) and maintain comprehensive runbooks, incident response plans, and compliance documentation.
2. Secondary Competency (Strong Advantage, Not Mandatory)
- Backend Software Development: Hands-on experience or a background in writing/debugging backend code in Node.js, Python, or Go.
- Note: The primary responsibility is cloud infrastructure, security, and automation. However, the ability to read backend code, debug API bottlenecks, or assist developers with microservice integrations is a major bonus.
About the Role
As the Product Manager at CloudKeeper, you will lead the strategy, roadmap, and execution of our real-time Cloud usage optimization and recommendation platforms.
As we expand our portfolio and deepen capabilities across cost optimization, observability, automation, and cloud management, we are seeking a Product Manager who can drive the vision, roadmap, and execution for one or more products within the CloudKeeper Platform Suite such as Lens, Tuner, Commit, Gen AI and Check
You will work cross-functionally across engineering, data science, FinOps SME, go-to-market and customer success to drive product growth, adoption, and customer value.
Key Responsibilities
- Define the vision and product strategy for CloudKeeper Platform Suite, aligned with overall cloud cost & usage optimization goals and the broader CloudKeeper solutions.
- Own the product roadmap: prioritize features, enhancements, integrations, and user experience improvements
- Own end-to-end product lifecycle: concept → PRDs → UX/design → development → launch → iteration
- Work with UX/design to craft intuitive workflows for engineering/DevOps users, ensuring that insights embed into their flow of work
- Engage with customers (particularly engineering and FinOps teams) to gather feedback, identify pain points, and validate product direction.
- Drive annual planning, release cycles, and sprint delivery (Agile methodology) while ensuring high quality and timely launch of features.
- Ensure appropriate documentation, training materials, and enablement for internal teams and external customers.
Desired Qualifications
- Bachelor’s degree in Computer Science, Engineering, Business or related field (Master’s preferred).
- 5+ years of product management experience in SaaS, preferably in cloud infrastructure, FinOps, cloud cost optimisation, or dev tools.
- Experience in B2B / Enterprise SaaS product development is a must (B2C product experience will not be considered)
- Strong understanding of AWS or other public-cloud: you should know at least the major compute / storage / services / networking footprint and usage patterns.
- Proven ability to translate customer needs into product features, write product requirements (PRDs), and work with engineering teams on execution.
- Proficiency with product management, collaboration, and analytics tools (e.g., Jira, Confluence, Figma, Productboard, BI dashboards, PostHog)
Preferred
- Prior experience in cloud cost optimization, cost-management tools, or FinOps practices is a strong plus
- Exposure to FinOps platforms or cloud cost optimization tools (e.g., AWS Cost Explorer/CUR, CloudHealth or equivalents)
What You’ll Work On
- Lead development and launch of next-gen recommendation engine enhancements (e.g., smarter detection of idle/over-provisioned resources).
- Evolve CloudKeeper’s industry-leading Guaranteed Savings model and its underlying pricing, automation, and governance engines
- Design conversational, AI-powered experiences that guide engineers through optimization opportunities, troubleshooting, and architectural decisions
- Define and refine metrics around savings impact, recommendation adoption, user engagement, churn, and product-market fit.
- Explore spaces like multi-cloud optimization, FinOps governance frameworks, cloud budgeting & forecasting, or cloud-native developer tools.
- Identify gaps in the cloud management ecosystem and work with leadership to incubate entirely new product lines, from concept to MVP to full-scale launch
What We Offer
- The opportunity to own a high-impact product in the rapidly growing FinOps/cloud cost optimisation space.
- A collaborative, innovative environment with talented teams in engineering, data science, UX, and product.
- The chance to make a measurable difference: help customers reduce cloud cost while maintaining performance and accelerate their cloud usage optimisation journey.
As a DevOps Engineer at YOYO, you'll own the infrastructure and delivery backbone that keeps our platform running as we grow. You'll build the CI/CD, cloud infrastructure, and observability that let a small, fast-moving team ship confidently and you'll keep our AI and data workloads reliable and affordable at scale. This is a hands-on role with real ownership: you won't be maintaining someone else's setup, you'll be shaping ours. You'll work closely with the backend, AI/ML, and data teams to make deployment boring, incidents rare, and scaling a non-event.
If you are Interested DM me on LinkedIn - Saquib Mundagnur
What You'll Own
CI/CD & developer experience - Build and maintain fast, reliable CI/CD pipelines so engineers ship multiple times a day with confidence. - Make the path from commit to production simple, safe, and repeatable, with sensible automated testing, rollbacks, and release controls.Cloud infrastructure & IaC - Own our cloud infrastructure (AWS/GCP) end to end, managed as code (Terraform or similar) — no click-ops. - Design for scale and cost-efficiency as store and conversation volumes grow.
Containers & orchestration - Run our services on containers/Kubernetes: deployments, autoscaling, networking, and resource management. - Support the specific needs of AI/ML workloads, including GPU-backed inference and batch processing for the speech pipeline.
Reliability & observability (SRE) - Own uptime, performance, and incident response — monitoring, logging, tracing, alerting, on-call, and blameless postmortems. - Define and defend SLOs; keep the platform dependable as it scales across clients.
Data & pipeline infrastructure - Support the infrastructure behind large-scale, edge-to-cloud data movement and processing (audio ingestion, ASR/AI pipelines, analytics). - Keep data workloads reliable, performant, and cost-aware.
Security & compliance - Bake security into the platform: secrets management, IAM/least-privilege, encryption in transit and at rest, network hardening, and vulnerability management. - Support compliance readiness (including India's DPDP Act and enterprise-client security requirements) for a product that handles sensitive customer conversations.
Cost & scale - Own cloud cost visibility and optimization; make scaling decisions that balance reliability and spend.
What You'll Bring - 6+ years in DevOps, SRE, platform, or infrastructure engineering, running production systems at meaningful scale.
- Strong hands-on experience with a major cloud provider (**AWS or Azure or GCP**) and Infrastructure-as-Code (**Terraform** or equivalent).
- Solid experience with **containers and Kubernetes** in production. - Experience building and owning **CI/CD** pipelines (e.g. GitHub Actions, GitLab CI, Jenkins, Argo, or similar).
- Comfort with a scripting/automation language (Python, Go, or Bash) and a strong automation-first mindset.
- Real experience with **observability** (Prometheus/Grafana, ELK, Datadog, OpenTelemetry, or similar) and running incident response / on-call.
- A security-conscious approach — secrets, IAM, encryption, and least-privilege as defaults. - Startup temperament: ownership, pragmatism, and a bias to automate and ship. - Based in or willing to relocate to Bangalore, and up for an onsite/hybrid, in-person team
The Role
As a **DevOps Engineer** you'll own the infrastructure and delivery backbone that
keeps our platform running as we grow. You'll build the CI/CD, cloud infrastructure, and
observability that let a small, fast-moving team ship confidently — and you'll keep our AI and
data workloads reliable and affordable at scale.
This is a hands-on role with real ownership: you won't be maintaining someone else's setup,
you'll be shaping ours. You'll work closely with the backend, AI/ML, and data teams to make
deployment boring, incidents rare, and scaling a non-event. ---
What You'll Own
**CI/CD & developer experience**
- Build and maintain fast, reliable CI/CD pipelines so engineers ship multiple times a day with
confidence. - Make the path from commit to production simple, safe, and repeatable, with sensible
automated testing, rollbacks, and release controls.
**Cloud infrastructure & IaC** - Own our cloud infrastructure (AWS/GCP) end to end, managed as code (Terraform or
similar) — no click-ops. - Design for scale and cost-efficiency as store and conversation volumes grow.
**Containers & orchestration** - Run our services on containers/Kubernetes: deployments, autoscaling, networking, and
resource management. - Support the specific needs of AI/ML workloads, including GPU-backed inference and batch
processing for the speech pipeline.
**Reliability & observability (SRE)** - Own uptime, performance, and incident response — monitoring, logging, tracing, alerting,
on-call, and blameless postmortems. - Define and defend SLOs; keep the platform dependable as it scales across clients.
**Data & pipeline infrastructure** - Support the infrastructure behind large-scale, edge-to-cloud data movement and
processing (audio ingestion, ASR/AI pipelines, analytics). - Keep data workloads reliable, performant, and cost-aware.
**Security & compliance** - Bake security into the platform: secrets management, IAM/least-privilege, encryption in
transit and at rest, network hardening, and vulnerability management. - Support compliance readiness (including India's DPDP Act and enterprise-client security
requirements) for a product that handles sensitive customer conversations.
**Cost & scale** - Own cloud cost visibility and optimization; make scaling decisions that balance reliability
and spend. ---
What You'll Bring - 6+ years in DevOps, SRE, platform, or infrastructure engineering, running production
systems at meaningful scale. - Strong hands-on experience with a major cloud provider (**AWS or Azure or GCP**) and
Infrastructure-as-Code (**Terraform** or equivalent). - Solid experience with **containers and Kubernetes** in production. - Experience building and owning **CI/CD** pipelines (e.g. GitHub Actions, GitLab CI,
Jenkins, Argo, or similar).
- Comfort with a scripting/automation language (Python, Go, or Bash) and a strong
automation-first mindset. - Real experience with **observability** (Prometheus/Grafana, ELK, Datadog,
OpenTelemetry, or similar) and running incident response / on-call. - A security-conscious approach — secrets, IAM, encryption, and least-privilege as defaults. - Startup temperament: ownership, pragmatism, and a bias to automate and ship. - Based in or willing to relocate to Bangalore, and up for an onsite/hybrid, in-person team.
Bonus Points - Experience running **ML/AI or GPU workloads** in production (inference serving, batch
pipelines, model deployment). - Experience with data-intensive infrastructure — streaming/queues (Kafka, SQS), data
pipelines, or large object/audio storage. - Exposure to **edge devices / IoT fleets**, OTA updates, or high-volume device-to-cloud
ingestion. - Experience with compliance/security frameworks (SOC 2, ISO 27001, DPDP). - FinOps / cloud cost-optimization experience. - Early-stage startup experience. ---
Why Join - Own infrastructure that's already live with leading retail brands and growing fast — real
scale, real impact. - Work across genuinely interesting workloads: speech AI, GPU inference, large-scale data,
and edge-to-cloud ingestion. - Small team, high ownership, direct line to engineering leadership — your decisions ship. - Build the platform foundation of a category-defining product from an

Job Title: TechOps Engineer
Location: Bengaluru, India (Hybrid)
Employment Type: Full-time
Experience: 6 Month-2 years (Excluding Internship)
Shift Timing: 2 PM to 11 PM IST
Role Overview
We are excited to find a highly engaged engineer who is obsessed with technology that wants to be a part of a “world class” platform SRE team. Engineers must possess an "automation first" mindset, with a relentless focus on documentation, quality, scalability, and reliability using Infrastructure as Code tools. This position will be part of a platform team that is developing exciting products and solutions and playing a key part in driving forward the electrification of transportation.
What you’ll do:
- Ensure system reliability, uptime, and performance of global platform.
- Conduct real-time surveillance of our EV charging systems to proactively identify and mitigate performance issues and anomalies near 24/7 basis. As such, you collaborate with IDT and FMC players to ensure incident detection also happens outside office hours (monitoring shifts among team members subject to duty schedule)
- Deliver on change & releases like firmware changes and drive insights & intelligence back into testing processes and tech discussions with the wider organization.
- Successfully deliver and project manage first time right commissioning activities alongside our Engineering Procurement Contract Management (EPCM) partners to successfully bring charge points onto our Charge Point Management System (CPMS).
- End-to-end EV charger lifecycle management, including deployment, commissioning, monitoring, maintenance, and decommissioning activities.
- Provide technical guidance and support to DC specialists during the commissioning of EV charging solutions.
- Work closely with Shell, Engineering, and IT colleagues to ensure projects are completed on time and to specification.
- Act as a liaison with the Engineering Procurement Contract Management (EPCM) partner to manage projects from start to finish, ensuring charge points are successfully onboarded on the Charge Point Management System (CPMS).
- Collaborate with development, operations and support teams to build scalable and resilient systems.
- Contribute to incident response, root-cause analysis, and post-mortem reviews, driving continuous improvement.
- Participate in capacity planning, performance tuning, and resource optimization.
- Integrate security and compliance best practices into all infrastructure operations.
- Stay current with emerging SRE tools, frameworks, and cloud technologies to continuously improve reliability practices.
- Participate in and lead on-call rotations and incident response, conducting detailed postmortems and RCA reports.
- Flexible to resolve blocking issues during off hours or weekends if required.
What We’re Looking For:
Basic Qualifications and skills
- Bachelor’s degree in Engineering , Electrical, ECE, Computer Science, Information Technology, or related field.
- Overall 1 years of experience as a Site Reliability Engineer, Technical project coordinator role.
- Proven experience of SRE or Technical Project Coordination with IoT or connected devices based platforms.
- Experience with incident management and on-call best practices. Provide support to on call engineers.
- Excellent analytical and problem-solving skills with a proactive mindset.
- Hands-on experience with AWS Cloud and IaC tools such as Terraform or Ansible.
- Expertise with monitoring and observability tools (Dynatrace,Prometheus, Grafana, Zabbix, etc.).
- Proactively monitor the network, triage performance outliers, and coordinate correction actions to ensure optimal system functionality.
- Fluency in English (spoken and written).
- Successfully recommission or decommission chargers following changes in our network.
- Responsible for the go-live of the chargers on Shell’s public network following commissioning attempts.
Note: This role involves managing infrastructure for a global platform operating in over ten countries, requiring effective communication and collaboration across regions. Strong verbal and written communication skills, along with availability and flexibility to resolve blocking issues, are essential to support On-call Engineers. This role may involve EU or US time‑zone shifts based on business requirements. The shift timing will be 2 PM IST to 11 PM IST.
What We Offer
- Work with some of the brightest minds in the emerging EV industry.
- Make a tangible impact in reducing carbon emissions and enabling sustainable energy.
- Freedom to suggest, implement, and innovate on systems, processes, and technologies.
- Daily ownership in a high-growth, challenging environment.
- Flexible work environment with hybrid schedules and virtualization options.
- Competitive pay and benefits including health coverage, innovative PTO program, and performance bonuses.
We are seeking a highly skilled Senior DevOps Engineer with 8+ years of professional experience to join our team. In this role, you will design, implement, and optimize cloud infrastructure and CI/CD processes.
You will collaborate closely with development, QA, and operations teams to deliver scalable, secure, automated, and reliable solutions on AWS.
The ideal candidate will have strong hands-on experience with AWS, Terraform, Git/GitHub, Jenkins, PowerShell, Python, AWS Systems Manager (SSM) Documents, and Active Directory (AD) administration, along with a passion for automation, efficiency, and operational excellence.
Key Responsibilities
- Design, build, and maintain scalable cloud infrastructure on AWS.
- Develop and manage Infrastructure as Code (IaC) using Terraform.
- Build, maintain, and optimize CI/CD pipelines using Jenkins and GitHub.
- Use AWS Systems Manager (SSM) to support operational automation and system administration.
- Automate system tasks and administrative workflows using PowerShell and other scripting languages.
- Create, maintain, and execute custom SSM Documents for configuration management, patching, automation, and troubleshooting.
- Manage and administer Active Directory (AD), including users, groups, permissions, policies, authentication, and integration with AWS services.
- Implement and manage version-control workflows in Git and GitHub.
- Ensure infrastructure and deployment processes follow best practices for security, reliability, scalability, and cost optimization.
- Monitor and troubleshoot production systems to ensure high availability, performance, and reliability.
- Collaborate with development teams to improve software delivery processes and release management.
- Mentor junior engineers and contribute to the development of DevOps standards and best practices.
Required Skills and Experience
- 8+ years of professional experience, including at least 4 years in DevOps or Site Reliability Engineering (SRE) roles.
- Strong hands-on experience with AWS services, including EC2, VPC, IAM, S3, EKS, Lambda, and related services.
- Proven expertise in Terraform for Infrastructure as Code.
- Experience administering both Linux and Windows operating systems.
- Experience designing and managing CI/CD pipelines using Jenkins and GitHub.
- Proficiency with Git workflows and source-code management best practices.
- Strong PowerShell scripting skills; familiarity with Python and/or Bash is a plus.
- Experience with AWS Systems Manager (SSM), including creating and managing SSM Documents.
- Hands-on experience managing Active Directory, including users, groups, policies, authentication, permissions, and AWS integration.
- Solid understanding of cloud networking, security, monitoring, and troubleshooting.
- Excellent problem-solving, communication, collaboration, and decision-making skills.
Nice-to-Have Skills
- Experience with containerization and orchestration technologies, such as Docker, Kubernetes, and Amazon EKS.
- Knowledge of monitoring and observability tools, such as Amazon CloudWatch, New Relic, and Sumo Logic.
- An AWS certification, such as AWS Certified DevOps Engineer – Professional or AWS Certified Solutions Architect.
Who You Are
- You are eager to learn new technologies and continuously improve your skills.
- You make sound decisions and take ownership of your work.
- You are proactive, self-motivated, and comfortable taking initiative.
- You are a strong communicator who enjoys collaborating with cross-functional teams.
- You are committed to improving processes, automation, reliability, and operational efficiency.











