50+ DevOps Jobs in India
Apply to 50+ DevOps Jobs on CutShort.io. Find your next job, effortlessly. Browse DevOps Jobs and apply today!
🚀 Job Title : DevOps Engineer / Site Reliability Engineer (SRE)
Experience Level : 5+ Years
Location : Gurugram Sector 39, Haryana (On-site)
Employment Type : Full Time Opportunity
About the Role :
We are looking for a proactive DevOps / Site Reliability Engineer (SRE) with around 5 years of hands-on experience designing, automating, and scaling cloud infrastructure and CI/CD delivery pipelines.
In this role, you will bridge the gap between development and operations. You will be responsible for orchestrating containerized applications, automating infrastructure via Code (IaC), establishing SRE best practices (SLIs, SLOs, SLAs), and ensuring maximum uptime, resiliency, and operational efficiency across multi-cloud environments (AWS/Azure/GCP).
Mandatory Skills :
AWS, Kubernetes, Docker, Terraform, Ansible, Jenkins, GitLab CI/CD, GitHub Actions, Python, Bash, CI/CD, Infrastructure as Code (IaC), Grafana, Prometheus, ELK, New Relic, CloudWatch, SRE, SLI/SLO/SLA, Linux
Key Responsibilities :
1. Cloud Infrastructure & Infrastructure as Code (IaC) :
- Provision, configure, and maintain scalable, high-availability infrastructure on multi-cloud platforms, primarily AWS (EC2, VPC, IAM, S3, RDS, Route53, ALB/ASG, Lambda, EBS).
- Build, deploy, and manage Infrastructure as Code (IaC) using Terraform, Ansible, and CloudFormation to enforce consistency and eliminate configuration drift.
- Execute disaster recovery (DR) planning, automated failover / failback mechanisms, and chaos engineering exercises to validate system resiliency.
2. CI/CD, Automation & Development :
- Design, end-to-end maintain, and optimize robust CI/CD pipelines using Jenkins, GitLab CI, and GitHub Actions.
- Automate release pipelines, versioning, branching strategies, and approval gates using Groovy, Python, and Bash scripting. Integrate automated code quality and security scanning tools (SonarQube, Black Duck, or Fortify) directly into delivery pipelines.
- Develop custom tools, scripts, or microservices (e.g., Python / Node.js) to automate manual operational tasks and operational toil.
3. Containerization & Orchestration :
- Onboard and orchestrate containerized microservices utilizing Docker and Kubernetes (including Helm charts).
- Ensure high availability, auto-scaling, resource management, and fault tolerance for Kubernetes pod deployments.
4. Observability, SRE & Incident Management :
- Drive Site Reliability Engineering (SRE) maturity by establishing, tracking, and reporting SLIs, SLOs, and SLAs with cross-functional engineering teams.
- Build, configure, and manage full-stack observability tools : Grafana, Prometheus, New Relic, Elasticsearch / Logstash / Kibana (ELK), Sentry, and AWS CloudWatch.
- Set up real-time alerting, custom metric dashboards, and automated log rotation / pruning scripts.
- Handle production incidents, lead Root Cause Analysis (RCA) investigations, and implement preventive measures to reduce Mean Time to Resolution (MTTR).
Required Qualifications & Skills :
- Education : Bachelor’s Degree in Electronics and Communication Engineering, Computer Science, or a related technical field.
- Experience : ~5 years of experience in DevOps, SRE, or Cloud System Administration roles.
- Cloud & Infrastructure : Hands-on experience with AWS (Core services like EC2, S3, VPC, RDS, IAM, Lambda, Auto Scaling) and exposure to Azure / GCP.
- CI/CD & Version Control : Proficiency with Jenkins, GitLab CI, GitHub Actions, and Git workflows.
- Containerization : Core proficiency in Docker and Kubernetes cluster management / onboarding.
- Infrastructure as Code : Expertise in Ansible, Terraform, or AWS CloudFormation.
- Scripting & Languages : Strong hands-on automation skills with Python, Bash, and foundational knowledge of Node.js, Java or C++.
- Observability & Logging : Strong experience with Grafana, Prometheus, New Relic, ELK stack, or Splunk.
- Database & SQL : Familiarity with relational databases (MySQL, RDS) for monitoring setup and operational analytics.
Soft Skills & Competencies :
- Excellent problem-solving, root-cause identification, and chaos engineering mindsets.
- Strong written and verbal communication skills in English.
- Comfortable working in Agile cross-functional environments and collaborating across development, security, and operations teams.
- Innate drive to reduce manual toil and automate repetitive processes.
Amura’s Vision
We believe that the most under-appreciated route to releasing untapped human potential is to build a healthier body, and through which a better brain. This allows us to do more of everything that is important to each one of us.
Billions of healthier brains, sitting in healthier bodies, can take up more complex problems that defy solutions today, including many existential threats, and solve them in just a few decades.
Billions of healthier brains will make the world richer beyond what we can imagine today. The surplus wealth, combined with better human capabilities, will lead us to a new renaissance, giving us a richer and more beautiful culture.
These healthier brains will be equipped with deeper intellect, be less acrimonious, more magnanimous, and have a kinder outlook on the world, resulting in a world that is better than any previous time.
We find this vision of the future exhilarating. Our hopes and dreams are to create this future as quickly as possible and ensure that it is widely distributed and optimized to maximize all forms of human excellence.
Role Overview
We are looking for a highly skilled Senior DevOps Engineer (AI-Native Infrastructure & Platform Engineering) with deep expertise in AWS cloud infrastructure, automation, AI infrastructure operations, and modern DevOps/SRE practices.
This role goes beyond traditional DevOps and requires a seasoned specialist capable of building and operating AI-ready infrastructure platforms that support high-throughput APIs, LLM/AI workloads, GPU-based compute, data-intensive systems, real-time inference pipelines, and scalable ML platforms.
You will be responsible for architecting, automating, securing, and optimizing highly scalable and cost-efficient cloud environments that enable high-velocity engineering and AI teams. This is an ideal position for someone who combines technical ownership, an automation-first mindset, and a passion for developer productivity and platform reliability.
Key Responsibilities
Cloud Infrastructure & Platform Engineering (AWS)
- Architect, deploy, and manage highly scalable and secure infrastructure on AWS. Design cloud platforms supporting AI/ML workloads, data pipelines, real-time APIs, and high-concurrency backend systems.
- Hands-on expertise with key AWS services including EC2, ECS/EKS, Lambda, RDS, DynamoDB, S3, VPC, CloudFront, IAM, CloudWatch, and GPU-enabled instances.
- Build and maintain Infrastructure-as-Code (IaC) using Terraform, CloudFormation, or AWS CDK.
- Design multi-AZ and multi-region architectures for high availability and disaster recovery (HA/DR).
- Build reusable platform templates and shared infrastructure modules.
AI/ML Infrastructure & MLOps
- Build and maintain infrastructure for LLM applications, AI inference workloads, model serving platforms, vector databases, and feature stores.
- Support GPU-based workloads and optimize compute/storage usage.
- Enable scalable deployment patterns for AI applications using Kubernetes/EKS. Collaborate with Data Science and ML Engineering teams on model deployment, training/tuning of models, CI/CD for ML systems, experiment environments, and reproducibility.
- Support orchestration and deployment of AI workflows and inference services while implementing observability and reliability for AI pipelines.
CI/CD, Automation & Developer Productivity
- Build and maintain CI/CD pipelines using GitHub Actions, GitLab CI, Jenkins, or AWS CodePipeline.
- Automate deployments, environment provisioning, and release workflows.
- Build self-service developer platforms, preview environments, and reusable deployment workflows to improve developer productivity.
- Implement automated patching, scaling, backups, cleanup workflows, and drift detection.
Containers, Kubernetes & Platform Reliability
- Manage Docker-based environments, containerized applications, and optimize workloads using Kubernetes (EKS) or ECS/Fargate.
- Manage autoscaling, cluster health, node pools, ingress, service mesh, and workload isolation.
- Optimize infrastructure for performance, resilience, and cost-efficiency.
- Implement progressive deployment strategies including blue/green, canary, and rolling deployments.
Observability, Incident Response & SRE Practices
- Implement observability stacks using CloudWatch, Prometheus, Grafana, ELK, Datadog, OpenTelemetry, or New Relic.
- Build actionable dashboards and intelligent alerting systems while defining and tracking SLIs, SLOs, and SLAs.
- Lead incident response, root cause analysis, and blameless postmortems to reduce operational toil and improve MTTR.
FinOps, Cost Governance & Security
- Continuously monitor and optimize cloud costs (compute utilization, storage lifecycle, GPU usage, and data transfer) using AWS Cost Explorer, Budgets, Trusted Advisor, CloudHealth, or Kubecost.
- Implement AWS security best practices for IAM, VPCs, security groups, NACLs, encryption, and manage secrets using KMS, SSM Parameter Store, or Vault.
- Build secure CI/CD pipelines with automated security checks, least-privilege access, audit logging, and ensure compliance readiness for ISO 27001, SOC2, and GDPR.
Collaboration, Leadership & Platform Culture
- Work closely with engineering, AI/ML, QA, product, and operations teams to drive a DevOps, SRE, GitOps, and automation-first culture.
- Mentor junior DevOps and Platform Engineers while creating and maintaining detailed runbooks, architecture diagrams, and platform documentation.
Skills & Qualifications
Must-Have:
- 7+ years of experience in DevOps, SRE, Platform Engineering, or Cloud Infrastructure Engineering.
- Strong expertise in AWS cloud architecture, services, and deep understanding of Kubernetes (EKS), containers, and cloud-native systems.
- Strong Infrastructure-as-Code expertise using Terraform, CloudFormation, or CDK. Strong Linux administration, networking, DNS, routing, and load balancing knowledge. Strong scripting/programming experience in Python, Bash, or Go (preferred). Experience with CI/CD automation, GitOps workflows, and observability platforms supporting scalable production systems.
Preferred / Nice-to-Have:
- Experience with AI/ML infrastructure, MLOps, model serving, vector databases, GPU orchestration, and inference optimization.
- Familiarity with Kafka, Redis, SQS, and event-driven systems.
- Exposure to platform engineering, internal developer platforms, and tools like ArgoCD, Flux, Helm, and OpenTelemetry.
- AWS Certifications: Solutions Architect, DevOps Engineer, or SysOps Administrator. Knowledge of distributed systems and large-scale platform operations.
Preferred / Nice-to-Have:
- Experience with AI/ML infrastructure, MLOps, model serving, vector databases, GPU orchestration, and inference optimization.
- Familiarity with Kafka, Redis, SQS, and event-driven systems.
- Exposure to platform engineering, internal developer platforms, and tools like ArgoCD, Flux, Helm, and OpenTelemetry.
- AWS Certifications: Solutions Architect, DevOps Engineer, or SysOps Administrator. Knowledge of distributed systems and large-scale platform operations.
Here are answers to some questions you may have
Where is your office?
Chennai (Velachery)
Work Model
Work from Office – because great stories are built in person!
Do you have an online presence?
https://amura.ai (we are @AmuraHealth on all social media)
You have strong experience supporting cloud environments within the AWS platform.
* You have hands-on experience with cloud platforms, Docker, Kubernetes, Jenkins, Terraform, Artifactory, infrastructure automation, CI/CD pipelines, monitoring, and operational support in enterprise environments.
* You have hands-on experience with infrastructure automation, scripting, and configuration management.
* You understand monitoring, logging, alerting, and incident management practices in cloud operations.
* You have experience working with containers, orchestration platforms, and deployment pipelines.
* You are comfortable supporting Linux and/or Windows server environments in cloud-based ecosystems.
* You understand identity and access management, security controls, patching, and operational compliance practices.
* You have experience with disaster recovery, backup processes, and platform resilience planning.
* You have the technical depth required to support and guide Cloud Operations activities while contributing to the overall effectiveness of the function from the India side.
You have experience supporting DevOps or Site Reliability Engineering practices.
* You have exposure to cost management and cloud optimization tools.
* You are comfortable presenting operational updates, recommendations, and risk considerations to stakeholders.
* You work well in team-based environments and contribute to shared goals.
* You have effective documentation skills for processes, runbooks, and support knowledge.
* You have experience working in regulated or enterprise-scale environments.
* You have familiarity with service management tools and structured change management processes.
* You demonstrate the potential to take on broader operational ownership and support the leadership of the Cloud Operations function from the India side.
* You are comfortable coordinating with multiple stakeholders and contributing to operational decision-making in a global delivery model.
Job Title : DevOps Engineer / Site Reliability Engineer (SRE)
Experience : 5+ Years
Location : Gurugram, Haryana
Work Mode : On-site (Full-time)
About the Role :
We are looking for a skilled DevOps Engineer with 5+ years of experience in cloud infrastructure, CI/CD, automation, Kubernetes, and Site Reliability Engineering (SRE). The ideal candidate will be responsible for building scalable cloud infrastructure, automating deployments, improving system reliability, and ensuring high availability across production environments.
Mandatory Skills :
AWS, Terraform, Ansible, CloudFormation, Jenkins, GitLab CI, GitHub Actions, Docker, Kubernetes, Helm, Python, Bash, Grafana, Prometheus, ELK Stack, CloudWatch, New Relic, SRE, CI/CD, Infrastructure as Code (IaC), Linux
Key Responsibilities :
- Design, deploy, and manage cloud infrastructure primarily on AWS (EC2, VPC, IAM, S3, RDS, Route53, ALB, Auto Scaling, Lambda).
- Build and maintain Infrastructure as Code (IaC) using Terraform, Ansible, and CloudFormation.
- Develop and optimize CI/CD pipelines using Jenkins, GitLab CI, and GitHub Actions.
- Deploy and manage containerized applications using Docker, Kubernetes, and Helm.
- Implement monitoring and observability using Grafana, Prometheus, ELK Stack, CloudWatch, and New Relic.
- Drive SRE practices by defining SLIs, SLOs, SLAs, handling production incidents, conducting RCA, and improving system reliability.
- Automate operational tasks using Python, Bash, and Groovy scripting.
- Collaborate with Development, QA, Security, and Operations teams to ensure reliable and secure software delivery.
Required Skills & Qualifications :
- Bachelor's degree in Computer Science, IT, Electronics, or a related field.
- 5+ years of experience in DevOps, SRE, or Cloud Infrastructure.
- Strong expertise in AWS, with exposure to Azure/GCP.
- Hands-on experience with Terraform, Ansible, CloudFormation, Docker, Kubernetes, Helm, Jenkins, GitLab CI, GitHub Actions, and Git.
- Strong scripting skills in Python and Bash.
- Experience with monitoring tools such as Grafana, Prometheus, ELK Stack, CloudWatch, and New Relic.
- Good understanding of Linux, networking, SQL, and cloud security best practices.
Preferred Skills :
- Experience with multi-cloud environments and DevSecOps practices.
- Knowledge of disaster recovery, automation, and microservices architecture.
- Strong troubleshooting, communication, and problem-solving skills.
We're hiring a Cloud Architect (Contract) to work with our Equity Partners who builds profitable growth by acquiring and operating enterprise software companies. Refining a proprietary operating model across 40+ acquisitions and two decades of hands-on experience, now supercharged by our patented agentic AI platform . In this role, you'll take full architectural control of our CI/CD, observability, and event streaming infrastructure, build the standards every new acquisition plugs into, and use AI-assisted automation to keep 20+ products reliable without proportionally scaling headcount.
Job title: Cloud/Platform Architect (SRE)
Type: Global Remote | Contract
What You Bring
- 8–12 years in platform engineering, DevOps, or SRE, with growing ownership over time
- Deep Terraform experience across multi-account, multi-env setups
- Real production experience with event streaming at scale
- Hands-on Grafana, Prometheus, Loki, and strong AWS depth (ECS, EKS, IAM, VPC, RDS)
- SRE fundamentals: SLOs, error budgets, on-call design, post-mortems
- Bonus: acquisition or greenfield platform-building experience
Roles and Responsibilities
- Own everything outside core AWS infra: CI/CD, observability, event streaming, deployment, incidents
- Define the standards every future acquisition will plug into
- Keep 20+ enterprise products running at serious scale (millions–billions of requests)
- Build self-service tooling so product teams never wait on you
- Use AI/automation to kill toil — not to replace engineering judgement
Ready to build the platform that scales an entire portfolio? — let's connect.
Applicants must read the JD and understand the JD throughly
Job Title: Database & DevOps Engineer
Location: Chennai, Tamil Nadu (On-site/Hybrid)
Experience: 2+ years -
Job Summary
We are looking for a skilled Database & DevOps Engineer with 2+ years of hands-on experience in database administration, cloud infrastructure, CI/CD pipelines, automation, and AI application deployment. The role involves managing database and infrastructure environments, automating deployments, supporting AI/LLM workloads, and ensuring high availability, security, performance, and scalability.
Key Responsibilities
Database Management
• Install, configure, monitor, and maintain SQL and NoSQL databases.
• Manage database backup, recovery, replication, and disaster recovery.
• Optimise performance through indexing, query tuning, and monitoring.
• Ensure database security, integrity, and availability.
• Troubleshoot database issues and provide timely resolutions.
DevOps, Infrastructure & AI Operations
• Design, implement, and maintain CI/CD pipelines.
• Automate infrastructure provisioning using Infrastructure as Code tools.
• Deploy and manage applications across cloud and on-premises environments.
• Configure and maintain containerised applications using Docker and Kubernetes.
• Deploy and support AI/ML, LLM, RAG, and agentic AI applications.
• Configure and manage GPU-enabled infrastructure for AI workloads.
• Support model serving, vector databases, and AI application APIs.
• Monitor application, infrastructure, and AI-service performance, including availability, latency, resource utilization, and failures.
• Maintain Linux servers and automate routine operational tasks.
• Implement secure management of credentials, secrets, model endpoints, and data access.
• Collaborate with development and AI teams to streamline deployment and release processes.
Mandatory Skills
• 2+ years of hands-on experience in Database Administration and DevOps.
• Strong knowledge of MySQL, PostgreSQL, Microsoft SQL Server, and MongoDB.
• Experience with database backup, restoration, replication, indexing, query optimization, and performance monitoring.
• Hands-on experience with Git and CI/CD tools such as Jenkins, GitLab CI/CD, or Azure DevOps.
• Experience with Docker and Kubernetes.
• Experience with AWS, Azure, or GCP.
• Knowledge of Terraform and Ansible.
• Strong Linux administration and shell-scripting skills.
• Experience with Prometheus, Grafana, ELK Stack, or Nagios.
• Working knowledge of networking, security, SSL/TLS, and load balancing.
• Understanding of LLMs, RAG, embeddings, vector databases, and agentic AI architecture.
• Experience deploying Python-based AI/ML services and APIs.
• Familiarity with model-serving tools such as llama.cpp, vLLM, Ollama, Hugging Face, or equivalent.
• Knowledge of GPU-enabled environments, NVIDIA drivers, CUDA, and containerized AI deployment.
• Understanding of AI/LLM observability, model versioning, security, scaling, and rollback.
• Strong troubleshooting, analytical, and problem-solving skills.
Preferred Qualifications
• Bachelor’s degree in Computer Science, Information Technology, Engineering, or a related discipline.
• Cloud, DevOps, Linux, database, or Kubernetes certifications are an advantage.
Technical Stack
• Databases: MySQL, PostgreSQL, Microsoft SQL Server, MongoDB and vector databases
• Cloud: AWS, Azure, GCP
• CI/CD: Jenkins, GitLab CI/CD, Azure DevOps
• Containers and IaC: Docker, Kubernetes, Terraform, Ansible
• Monitoring: Prometheus, Grafana, ELK Stack, Nagios
• AI Infrastructure: Python APIs, LLM serving, RAG, embeddings, vector search, GPU/CUDA
• Operating Systems: Linux—Ubuntu, CentOS, Red Hat
Soft Skills
• Excellent communication and collaboration skills.
• Strong ownership and accountability.
• Ability to work effectively in a fast-paced Agile environment.
• Proactive approach to troubleshooting and continuous improvement.
• Ability to manage multiple priorities and deliver on time.
Very strong infra engineer to administer Linux based VMs (and few Windows VMs as well).
Must have strong knowledge and hands-on on docker, Kubernetes
Must be good at monitoring (AppDynamics, Prometheus, Grafana) and logging (Splunk)
Develop automation scripts using Shell and Python. Must have good knowledge in Ansible.
Should be good at optimize/finetune server configurations.
Should be good in troubleshoot issues and fix them.
Software Development Engineer (1–3 Years Experience)
Overview
We are looking for a proactive Full Stack Engineer with 1–3 years of core backend experience in .NET / .NET Core or Java (Spring Boot), alongside modern frontend skills in TypeScript. In this role, you will help design, build, and maintain scalable web applications and APIs while working across the full software stack.
If your primary backend experience is in Java/Spring Boot, we welcome your application—provided you are eager to adapt and pick up .NET/C# in our environment.
Key Responsibilities
- Backend Development: Build, optimize, and maintain reliable backend services and RESTful APIs using .NET / .NET Core (C#) or Java (Spring Boot).
- Frontend Integration: Develop responsive, clean frontend components using TypeScript and modern JavaScript frameworks (React, Angular, or Vue).
- Engineering Standards: Write clean, testable code, participate in peer code reviews, and maintain unit test coverage.
- Collaboration: Partner with cross-functional team members to implement feature requirements and improve system architecture.
- Performance & Security: Help monitor, debug, and optimize application performance, responsiveness, and data security.
Requirements
- Experience: 1–3 years of hands-on software development experience.
- Backend: Strong foundation in C# / .NET (.NET Core) OR Java / Spring Boot, with willingness to work across frameworks as needed.
- Frontend: Proficiency in TypeScript and at least one modern frontend library/framework (React, Angular, or Vue).
- Core Fundamentals: Solid understanding of REST APIs, relational databases (e.g., PostgreSQL, MySQL, SQL Server), and version control (Git).
- Problem-Solving: Good analytical skills and clarity in technical communication.
Nice to Have
- Experience with cloud platforms (Azure or AWS).
- Familiarity with containerization tools (Docker) and modern CI/CD pipelines.
- Exposure to NoSQL databases or caching mechanisms (Redis).
About the opportunity:
This role is with Sporty Group. CareersInCloud shares this opportunity to help professionals discover relevant technology careers.
Hiring Company: Sporty Group
Sporty Group is looking for an experienced DevOps Engineer to join its remote engineering team.
This role involves managing AWS cloud infrastructure, Kubernetes environments, CI/CD pipelines, monitoring, automation, and improving system reliability.
Key Responsibilities:
• Manage AWS infrastructure and cloud operations
• Maintain Kubernetes clusters and Docker environments
• Build and improve CI/CD pipelines
• Monitor production systems using Grafana, Prometheus, and CloudWatch
• Troubleshoot Linux, networking, and infrastructure issues
Requirements:
• 3+ years of DevOps experience
• Strong AWS and Kubernetes experience
• Experience with Docker, Linux, CI/CD tools
• Good understanding of cloud infrastructure and automation
Apply here:
https://careersincloud.com/jobs/sporty-group-hiring-devops-engineer-remote-experienced
Apply directly at VotersAI website (https://votersai.com/careers), a silicon valley, USA based start-up building a patent-pending, multi-tenant intelligence platform on AWS. We are hiring a DevOps / platform engineer to own our deployment pipeline, cloud infrastructure, observability, and reliability.
What you will do
- Build and operate CI/CD that deploys server-less services to AWS through GitHub Actions and OIDC (no stored credentials), with safe one-at-a-time deploys, verification, and fast rollback.
- Operate and harden a production AWS estate: Lambda, API Gateway, SQS with dead-letter queues, DynamoDB, RDS PostgreSQL, VPC networking, IAM, S3, CloudFront, Route 53, Cognito, Secrets Manager, SES.
- Build out observability — CloudWatch metrics, alarms, dashboards, structured logging, and alerting — and an operations and monitoring surface for the team.
- Engineer for resilience and scale: load and chaos testing, Step Functions self-healing and auto-remediation, capacity and concurrency tuning, and queue-based admission control for high-burst workloads.
- Own least-privilege IAM, secret rotation, and infrastructure security.
- Carry on-call; drive incident response and mean-time-to-resolve.
Required experience
- Minimum 5 years of professional experience is required, in DevOps, SRE, or platform engineering, with deep hands-on AWS: Lambda, API Gateway, SQS, DynamoDB, IAM, VPC, CloudWatch, EventBridge, Step Functions, S3, CloudFront, Cognito, Secrets Manager.
- Infrastructure-as-code: Terraform, CloudFormation, SAM, or CDK.
- GitHub Actions CI/CD, including OIDC federation to cloud roles.
- Strong in BOTH NoSQL and relational databases: DynamoDB (single-table design, capacity and throughput modeling) AND PostgreSQL (RDS operations, tuning, backups).
- REST API operations behind API Gateway: authorizers, throttling, CORS, staged
deployments.
- Observability and incident response: metrics, alarms, dashboards, alerting, SLOs, and MTTR discipline.
- Confident Python and Bash scripting; strong Linux
Preferred
- Capacity planning and queuing theory; chaos and load testing. High-volume messaging: SMS 10DLC and email deliverability.
- Multi-tenant SaaS, and operating under regulated-communications compliance such as TCPA.
- ClickHouse or another columnar / OLAP database in production.
Other Skills:
- Fluent in English
- Extremely focused work habits.
- You will be working on patent pending technology, strong Math background preferred.
- You will have the lifetime opportunity to learn cutting edge technology from Silicon Valley experts.
Location:
You should be able to commute to Aundh Area of Pune.
We are hiring an experienced Azure DevOps Architect to design and drive enterprise-scale cloud engineering, DevOps transformation, and automation initiatives.
Responsibilities:
- Design and implement scalable Azure cloud architectures and DevOps solutions.
- Build and optimize CI/CD pipelines using Azure DevOps / GitHub Actions.
- Drive DevSecOps, Infrastructure-as-Code (Terraform), and cloud automation practices.
- Work with Kubernetes (AKS), Docker, and cloud-native platforms.
- Define engineering standards around security, governance, monitoring, and platform engineering.
- Explore AI-enabled engineering capabilities using Azure AI Foundry, Azure OpenAI, and MCP.
Qualifications:
- 10+ years of experience in Azure Cloud, DevOps, or Cloud Architecture.
- Strong hands-on expertise in Azure, Azure DevOps, Terraform, CI/CD, Kubernetes, and DevSecOps.
- Experience designing enterprise-grade cloud platforms and automation solutions.
- Exposure to AI/GenAI integration in engineering workflows is a plus.
Principal Architect.
Architecture: Microservices, Event-Driven, Serverless, Service Mesh, Multi-tenancy, API Gateway.
DevOps/IaC: Terraform, Kubernetes (EKS/AKS/GKE), Helm, Jenkins/GitHub Actions, ArgoCD, MLOps.
Monitoring/Performance: OpenTelemetry, Datadog/New Relic, Chaos Engineering, Load Testing (k6/JMeter).
Networking/Security: VPC, Zero Trust, IAM, WAF, OAuth2/OIDC, DevSecOps.
Job Description:
An Azure Administrator is a specialized IT professional responsible for implementing, monitoring, and maintaining Microsoft Azure solutions, including major computing, storage, network, and security services. They ensure that Azure services are running optimally and securely, aligning with the organization's requirements and industry best practices.
Azure stands out as the top cloud computing platform, empowering businesses globally to optimize their operations. Beyond Azure services, it delivers cutting-edge AI solutions, powerful data analytics, top-tier security measures, machine learning advancements, natural language processing, and seamless IoT management, among other innovative features
Roles and Responsibility ::
- Managing Azure Subscriptions and Resources
- Implementing and Managing Storage Solutions
- Configuring and Managing Virtual Networks
- Managing Identities and Access
- Deploying and Managing Azure Compute Resources
- Monitoring and Maintaining Azure Resources
Roles and Responsibilities:
· Experience Required 3+ yrs in Devops.
· Position - 1
· Hands-on experience on Azure.
. Hands on experience on IIS
· Ability to design a good cloud solution.
· Strong Linux troubleshooting, Shell Scripting, Kubernetes, Docker, Ansible, Jenkins Skills.
· Design and implement the CI/CD pipeline following the best industry practices using open-source tools.
· Use knowledge and research to constantly modernize our applications and infrastructure stacks.
· Be a team player and strong problem-solver to work with a diverse team.
· Having good communication skills.
Primary Skills:
Observability - ELK (Elactic/Kibana), Prometheus, Grafana, PromQL
Software and automation - Java, Python/Shell/Bash, Rest-SOAP API, docker containerization, Kubernetes, Kafka
Reliability and DR engineering - Distributed architecture and distributed system fundamentals, micro services, and event-driven architecture.
Cross-team coordination, incident triage and resolution, leadership and stakeholder management.
Secondary Skills:
Lang-chain, Langraph, RAG, MCP
Experience with working on LLM's and integrating with the existing applications
Python - FastAPI
Cache - Redis
Program Details:
Design, build, and ship LLM-powered and agentic product features that enhance the team efforts and outcomes.
Build agentic AI systems that reason over context, invoke tools, take real actions, and recover gracefully from failure.
Work on integrating the existing AI tools and should know major AI frameworks and libraries.
Own service reliability and operational governance by defining SLA's, managing error budgets, and reporting reliability (MTTD, MTTR) to leadership for prioritization, risk decisions and planning.
Architect and continuously optimize the observability of platform using Kibana/Elastic (ELF) along with other observability tools like Prometheus, Grafana (dashboards, metrics, alert lifecycle), improving detection quality, reducing noise/toil, and enabling faster triage and measurable uptime improvements.
Engineer advance alerting and automation capabilities with Kibana alerting and anomaly detections and integrating response workflows (routing, runbooks, remediation scripts) to standardize on-call execution and accelerate restoration of services.
Lead incident response for customer-impacting issues across teams-coordination, communications, service restoration, and blameless RCA-then corrective actions that prevent recurrence and reduce operational risk.
Design, automate and validate Disaster Recovery and failover for critical services/journeys (RTO/RPO alignment, DR Drills), ensuring resiliency under failure scenarios and improving recovery.
Consult and partner with application teams by providing production readiness inputs (Resiliency patterns, availability, performance/capacity considerations) and driving platform enhancements that improve stability while optimizing infrastructure and observability spend.
If you are interested for this role, kindly acknowledge this email with your interest.
Senior DevOps Engineer
Location: India, Remote
Type: Contract
Department: Engineering
About Us
OpenAssets is building open standards for the next generation of capital markets, providing the trusted foundation for real-world asset tokenization and digital currencies.
We specialize in alternative asset digitization, offering regulatory-conscious tools to design, issue, and manage digital assets in complex financial environments. We also develop loyalty solutions that turn customer engagement into programmable digital assets, helping brands create personalized experiences that drive growth.
OpenAssets operates at the intersection of finance, technology, and regulation focused on building secure, scalable, and open infrastructure for the future of digital markets. Join OpenAssets and be part of shaping the future of digital finance!
Overview
We are seeking a Senior DevOps Engineer to lead the evolution of our cloud infrastructure and deployment ecosystems.
This role acts as the foundation of our engineering reliability, bridging the gap between high-performance blockchain development and robust, secure cloud operations. You will be responsible for architecting scalable AWS environments, automating mission-critical CI/CD pipelines, and ensuring the absolute integrity of our security and observability frameworks.
What You’ll Do
Cloud Infrastructure & Architecture
- Design, provision, and maintain OpenAssets’ AWS infrastructure: including EC2, ECS/Fargate, Lambda, S3, RDS, VPC, IAM, CloudFront, SQS, ElastiCache, and CloudWatch.
- Architect for high availability, fault tolerance, and cost efficiency across production and non-production environments.
- Manage networking, security groups, subnets, load balancers, and access controls to ensure a secure and well-segmented cloud environment.
- Scale infrastructure dynamically to handle platform growth, partner onboarding, and peak load events.
- Deploy, configure, and manage blockchain node infrastructure supporting on-chain integrations and smart contract interactions.
CI/CD & Automation
- Implement and maintain CI/CD pipelines using GitHub Actions and/or AWS CodeBuild for frontend, backend, and smart contract deployments.
- Manage infrastructure as code using Terraform (or OpenTofu) and/or CloudFormation to ensure reproducible, auditable environments.
- Automate build, test, and deployment processes to minimize manual effort and reduce release risk.
- Develop and maintain automation scripts in Python and Bash for operational and provisioning workflows.
- Administer and scale Kubernetes clusters (EKS) for deploying services and blockchain node infrastructure.
Security & Compliance
- Implement and enforce security best practices across infrastructure: TLS, VPNs, IAM roles, private key management, secrets management, and patch management.
- Work with the CISO to conduct infrastructure security audits and maintain compliance with data handling, uptime, and audit logging requirements.
- Manage secrets and environment configuration securely using AWS Secrets Manager or HashiCorp Vault.
- Ensure encryption at rest and in transit across all critical data stores and communication channels.
- Apply knowledge of blockchain security, including consensus mechanisms, smart contract interactions, and Layer 2 infrastructure, to harden on-chain components.
Observability & Incident Response
- Set up and maintain monitoring, alerting, and dashboards for key platform metrics using CloudWatch, Prometheus, Grafana, or equivalent.
- Establish structured logging and distributed tracing practices (OpenTelemetry or similar) across services.
- Lead incident response for production issues, triaging, communicating, resolving, and writing blameless postmortems.
- Design, implement, and regularly test backup and disaster recovery strategies to minimize data loss and downtime.
Collaboration & Enablement
- Support engineering teams in troubleshooting environment-specific issues across development, staging, and production.
- Document infrastructure architecture, runbooks, and operational procedures — keeping them current and actionable.
- Train developers on deployment tooling, CI/CD workflows, and on-call processes.
- Evaluate and introduce new tooling and practices to improve DevOps maturity across the organization.
Qualifications & Experience
- 5+ years of experience in DevOps, SRE, or cloud infrastructure engineering.
- Deep hands-on expertise with AWS: EC2, ECS/Fargate, Lambda, RDS, S3, VPC, IAM, CloudWatch, SQS, CloudFront, and ElastiCache.
- Strong experience with infrastructure as code using Terraform, OpenTofu, and/or CloudFormation.
- Proficiency with containerization and orchestration (Docker, Kubernetes / EKS).
- Experience designing and maintaining CI/CD pipelines using GitHub Actions, AWS CodeBuild, GitLab CI, or similar.
- Solid Linux systems administration and scripting skills (Bash, Python).
- Strong understanding of cloud networking and security: TLS, VPCs, security groups, IAM, private key management, secrets management.
- Experience with monitoring and observability tooling: CloudWatch, Prometheus, Grafana, ELK stack, or equivalent.
- Familiarity with distributed tracing (OpenTelemetry, Jaeger, or similar).
- Working knowledge of blockchain technologies: consensus mechanisms, smart contract infrastructure, or Layer 2 scaling solutions.
- Calm, systematic problem-solving approach to production incidents and outages.
- Strong communication skills, able to explain infrastructure concerns clearly to both technical and non-technical stakeholders.
- Fluid in English
Nice to Have
- AWS certifications (Solutions Architect, DevOps Engineer Professional, or equivalent).
- Experience deploying and operating blockchain node infrastructure (Ethereum, Polygon, or similar).
- Background in fintech, capital markets, payments, or regulated infrastructure environments.
- Experience with security frameworks and compliance standards (SOC 2, ISO 27001).
- Multi-tenant SaaS infrastructure experience.
- Experience with event-driven architectures (Kafka, SQS, BullMQ).
- Cost optimization practices: reserved instances, Savings Plans, right-sizing.
- Infrastructure as code with Pulumi or CDK.
- Open-source contributions or technical writing.
We are committed to providing equal opportunity for qualified applicants to contract positions, regardless of race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or veteran status. This is a contract opportunity, not a direct employment role.
Full-Stack Developer (Angular & Node.js) - Contract-Based
Role Overview: We are looking for a highly skilled Full-Stack Developer to join our team for a high-impact, 6-month contract engagement. This is an onsite, contract-based role based in Andheri, Mumbai. You will be responsible for building, scaling, and maintaining robust web applications, contributing directly to our core product architecture.
Key Responsibilities:
- Develop and maintain scalable, high-performance web applications using Angular and Node.js.
- Architect and build modular backend services, secure RESTful APIs, and manage database schemas.
- Collaborate with the internal product team to deliver new features and optimize existing workflows.
- Manage complex database query optimization and indexing for high availability.
- Maintain clean, maintainable, and well-documented code that adheres to industry standards.
Technical Requirements:
- Experience: 6–8+ years of professional full-stack development experience.
- Frontend Stack: Strong proficiency with Angular, including component architecture, RxJS, and state management.
- Backend Stack: Strong backend experience with Node.js (Express/NestJS or similar framework) and robust REST API design.
- Languages: Solid command of JavaScript and TypeScript across the entire stack.
- Database & Search: Hands-on experience with PostgreSQL (including schema design and query optimization) and Elasticsearch (indexing, querying, and performance tuning).
- Devops & Security: Sound understanding of security best practices (especially important in a fintech/PE context), alongside familiarity with Git, CI/CD pipelines, and modern testing practices.
- Soft Skills: Ability to operate independently in a fast-paced environment and excellent problem-solving skills.
Contract Details:
- Type: Onsite, Contract-based (6 Months, extendable).
- Budget: 90k / Months
- Location: Andheri, Mumbai. (Open exclusively to Mumbai-based candidates as no relocation assistance will be provided).
- Availability: Immediate joiners preferred.
Senior Cloud Site Reliability Engineer (CSRE) – Azure
About Searce:
Searce is an AI-native, engineering-led modern technology consultancy that empowers
clients to futurify their businesses by delivering real, intelligent business outcomes. As a
trusted partner for over 3,000 clients globally, Searce specializes in cloud modernization,
data engineering, applied AI, and robust cloud platform security. Driven by a "HAPPIER"
cultural mindset and our proprietary evlos problem-solving framework, we eliminate
bureaucratic fluff to build working prototypes fast and scale enterprise production
environments intelligently. We don't just fix systems; we leverage multi-cloud technologies
to transform client operations into distinct competitive advantages.
Position Overview:
We are looking for a high-caliber Senior or Lead Cloud Site Reliability Engineer (CSRE) to
architect, secure, and stabilize next-generation hybrid and multi-cloud environments.
Operating at the intersection of infrastructure design, security compliance, and production
operations, you will serve as the technical Subject Matter Expert (SME) across GCP, Azure,
and AWS.
Whether optimizing a microservice mesh on GKE, tuning autoscaling on AKS, or driving a
massive disaster recovery drill across AWS regions, your focus will be absolute reliability. For
the Lead path, you will couple this deep engineering toolkit with stakeholder management
and mentorship to drive an elite operational culture.
Experience & Level Expectation:
Years of Experience: 3 to 10 years of intensive, hands-on production operations
experience in a dedicated DevOps, Cloud Platform Engineering, or SRE role.
Associate level (3-5 Years): Expected to show flawless execution of IaC, advanced
triaging of infrastructure failures, and ownership of the CI/CD and deployment
lifecycles.
Intermediate level (5-10 Years): Expected to take architectural ownership, serve as
primary Incident Commander for complex outages, design cross-cloud governance
frameworks, and act as a reliable bridge between technical teams and client
leadership.
Key Responsibilities & Role Expectations:
Multi-Cloud Platforms & Orchestration: Design, configure, and maintain
production-grade Kubernetes clusters across major platforms (AKS).
Manage advanced network routing, service meshes (e.g., Istio), and multi-tenant
isolation.
Infrastructure as Code (IaC) & GitOps: Build declarative, enterprise-grade, reusable
infrastructure components using Terraform or Crossplane. Standardize automated
environment provisioning to eliminate configuration drift across multi-branch
environments.
Incident Management & Reliability (SRE): Own and optimize the production on-call
rotation. Lead rapid mitigation strategies for Sev-1/Sev-2 system outages, reducing
Mean Time to Recovery (MTTR) through centralized log and metric correlation.
Root Cause Analysis (RCA): Facilitate rigorous, blameless post-incident reviews to
identify core architectural vulnerabilities and establish long-term fixes preventing
recurrence.
Lifecycle, Patching & Upgrades: Plan and execute zero-downtime cluster upgrades,
operating system patching strategies (Linux/Windows), database lifecycle updates,
and multi-region Disaster Recovery (DR) failover drills.
Core Core Operations & Legacy Integration: Manage enterprise-level hybrid
networking architecture (VPCs, Firewalls, Load Balancers, DNS routing, and DHCP
configurations) while effectively connecting cloud native services to legacy
infrastructures like Active Directory.
Security & Governance: Embed Zero Trust policies, secure secrets management
(Secrets Manager/Key Vault), and continuous vulnerability patching into the
automated SDLC pipeline.
Required Technical Skills:
- Microsoft Azure: Azure Virtual Machines, Virtual Networks, Azure Active Directory, Azure Update Management.
- Containers & Orchestration
- Production-level management of GKE, AKS, and EKS.
- Advanced mastery of Docker, Helm, Kubernetes StatefulSets, Pod Disruption
At BigThinkCode, our technology solves complex problems. We are looking for talented Azure Cloud Devops engineer to join our Cloud Infrastructure team at Chennai.
Please find below our job description, if interested apply / reply sharing your profile to connect and discuss.
Company: BigThinkCode Technologies
URL: https://www.bigthinkcode.com/
Role: Devops Engineer
Experience required: 3–5 years
Work location: Chennai
Joining time: Immediate – 2 weeks
Work Mode: Work from office (Hybrid)
About the Role:
We are looking for a DevOps Engineer with 3+ years of hands-on experience to build and manage CI/CD pipelines, cloud infrastructure, containerized deployments, and operational monitoring. The engineer will work closely with Architects, Developers, QA, and Integration teams to deliver secure, scalable, and production-ready applications across AWS and Microsoft Azure environments.
Key Responsibilities
· Design, implement, and maintain CI/CD pipelines using GitHub Actions, Azure DevOps, Jenkins, or similar tools.
· Deploy, manage, and monitor applications on AWS and/or Microsoft Azure.
· Build and manage containerized applications using Docker.
· Provision cloud infrastructure using Terraform or Infrastructure as Code (IaC).
· Configure IAM, Key Vault/Secrets Manager, networking, SSL, and secure deployments.
· Implement monitoring, logging, alerting, backup, and disaster recovery solutions.
· Collaborate with development teams to troubleshoot build, deployment, and production issues.
· Maintain deployment documentation, runbooks, and operational procedures.
Required Skills & Qualifications
· 3+ years of experience in DevOps / Cloud / Infrastructure Engineering.
· Strong understanding of DevOps principles and CI/CD practices.
· Hands-on experience with AWS and/or Microsoft Azure.
· Experience with GitHub Actions, Azure DevOps, Jenkins, or similar CI/CD tools.
· Strong Linux administration skills.
· Hands-on experience with Docker and container lifecycle management.
· Working knowledge of Terraform or other Infrastructure as Code tools.
· Experience with cloud services such as:
· AWS: EC2, IAM, S3, VPC, RDS, CloudWatch, ECR
· Azure: App Services, Container Apps, Key Vault, Azure Monitor, Application Insights, Azure Container Registry
· Good understanding of networking concepts (DNS, SSL, Load Balancers, Firewalls, Virtual Networks).
· Basic operational knowledge of PostgreSQL/MySQL.
· Strong troubleshooting and production support skills.
Good to Have
· Kubernetes (AKS/EKS) or Azure Container Apps.
· Monitoring tools such as Prometheus, Grafana, ELK, CloudWatch, or Azure Monitor.
· Experience with GitHub Enterprise, Branch Protection, and CI/CD quality gates.
· Exposure to security tools such as CodeQL, Dependabot, Trivy, or SonarQube.
· Shell, Bash, PowerShell, or Python scripting.
· Experience with Microsoft Entra ID (Azure AD), OAuth, or SSO.
· Exposure to enterprise integration environments (Oracle, Kafka, REST APIs, OIC).
Soft Skills
· Strong ownership and problem-solving mindset.
· Good communication and documentation skills.
· Ability to work collaboratively in Agile teams.
· Passion for automation, security, and continuous improvement.
On the Job
● Build and maintain scalable cloud infrastructure on AWS
● Improve CI/CD pipelines, deployment systems, and developer workflows
● Automate infrastructure provisioning using Infrastructure-as-Code (Terraform/Pulumi)
● Manage and optimize Kubernetes clusters and containerized workloads
● Improve observability across systems through monitoring, logging, and alerting
● Drive reliability initiatives including incident response, root cause analysis, and operational
improvements
● Collaborate with engineering teams to improve service scalability, performance, and security
● Implement IAM, secrets management, and infrastructure security best practices
● Optimize infrastructure costs and improve resource efficiency
● Actively leverage AI tools and workflows to improve engineering productivity and automation
Must Haves
● 3–6 years of experience in DevOps, Platform Engineering, or SRE roles
● Strong hands-on experience with AWS infrastructure and services
● Good understanding of Kubernetes, Docker, and container orchestration
● Experience with Infrastructure-as-Code tools like Terraform or Pulumi
● Strong scripting/coding skills in Python, Go, or Bash
● Experience building and maintaining CI/CD pipelines (GitHub Actions, Jenkins, GitLab CI,
etc.)
● Understanding of networking fundamentals, Linux systems, and cloud security practices
● Familiarity with monitoring and observability tools like Prometheus, Grafana, ELK, Datadog,
etc.
● Strong debugging and problem-solving skills
● Ability to work independently in a fast-moving environment
Good To Haves
● Experience working in fintech or high-scale startup environments
● Exposure to service mesh, zero-trust security, or secrets management systems
● Experience with multi-cluster Kubernetes environments
● Familiarity with incident management, SLOs, and reliability engineering practices
● Experience building internal developer platforms or automation tooling
Data Engineer – Splunk & ELK Stack
Job Summary
We are seeking a skilled Data Engineer with hands-on experience in Splunk, the ELK Stack (Elasticsearch, Logstash, Kibana), and modern data engineering practices. The ideal candidate will design, build, and maintain scalable data pipelines, log analytics platforms, and monitoring solutions to support business intelligence, security, and operational excellence.
Key Responsibilities
- Design, develop, and maintain scalable data ingestion and ETL/ELT pipelines.
- Configure, administer, and optimize Splunk environments for log collection, indexing, searching, and reporting.
- Develop and maintain ELK Stack solutions using Elasticsearch, Logstash, Kibana, and Beats.
- Build dashboards, visualizations, alerts, and reports for infrastructure, application, and security monitoring.
- Integrate data from multiple structured and unstructured sources into centralized analytics platforms.
- Optimize Elasticsearch clusters for performance, scalability, and high availability.
- Troubleshoot data pipeline failures, indexing issues, and system performance bottlenecks.
- Automate deployment and configuration using scripting and Infrastructure as Code where applicable.
- Collaborate with DevOps, Security, Cloud, and Application teams to implement observability and monitoring solutions.
- Ensure data quality, governance, and compliance with organizational standards.
- Document technical designs, operational procedures, and best practices.
Required Skills
- Strong experience with Splunk Enterprise administration and development.
- Hands-on experience with the ELK Stack:
- Elasticsearch
- Logstash
- Kibana
- Beats (Filebeat, Metricbeat, Winlogbeat, etc.)
- Experience building ETL/ELT pipelines and data integration workflows.
- Strong SQL skills and experience with relational databases.
- Experience with Python, Shell scripting, or Java for automation.
- Understanding of log management, monitoring, and observability concepts.
- Experience working with Linux environments.
- Knowledge of REST APIs and data ingestion techniques.
- Familiarity with cloud platforms such as AWS, Azure, or Google Cloud.
- Experience with Git and CI/CD pipelines.
Preferred Skills
- Experience with Kafka, Spark, or other streaming technologies.
- Knowledge of Docker and Kubernetes.
- Experience with Terraform, Ansible, or other Infrastructure as Code tools.
- Understanding of SIEM concepts and security analytics.
- Experience with Prometheus, Grafana, or OpenTelemetry.
- Exposure to big data technologies and distributed systems.
We are seeking a highly experienced Azure AI, AIOps & MLOps Architect to lead enterprise-scale AI platform engineering, cloud modernization, DevSecOps transformation, and intelligent automation initiatives.
The ideal candidate should possess deep expertise in Microsoft Azure, Azure AI Foundry, Azure OpenAI, Azure Machine Learning, Kubernetes, Terraform, Azure DevOps, and enterprise observability platforms. The role will focus on designing scalable AI platforms, implementing MLOps and AIOps capabilities, enabling Agentic AI architectures, and driving cloud-native engineering practices across the organization.
Key Responsibilities
Cloud Architecture & Engineering
• Design and implement scalable, secure, and highly available solutions on Microsoft Azure.
• Define cloud architecture standards, reference architectures, and best practices.
• Lead cloud migration and modernisation initiatives across enterprise workloads.
• Implement multi-region disaster recovery and business continuity strategies.
• Oversee Azure networking, identity, security, and governance frameworks.
DevOps & CI/CD
• Architect and implement end-to-end CI/CD pipelines using Azure DevOps or GitHub Actions.
• Drive DevSecOps culture — embedding security scanning, quality gates, and compliance into the delivery pipeline.
• Champion Infrastructure-as-Code (IaC) practices using Terraform, Bicep, or ARM templates.
• Establish branching strategies, release management, and environment promotion standards.
• Define and enforce platform engineering standards and internal developer tooling.
AI & Machine Learning Integration
• Architect AI/ML solutions leveraging Azure AI services — Azure OpenAI, Azure Machine Learning, Azure AI Foundry, and Cognitive Services.
• Design intelligent automation and agentic workflows integrated into enterprise DevOps processes.
• Implement AI-powered capabilities such as code review assistance, anomaly detection, predictive analytics, and natural language automation.
• Define AI governance frameworks: model evaluation, prompt management, responsible AI, and cost controls.
• Design and implement enterprise MLOps frameworks.
• Build automated model training, validation, deployment, and monitoring pipelines.
• Establish model governance and lifecycle management.
Generative AI & Agentic AI
- Design enterprise GenAI solutions using Azure OpenAI.
- Build AI Agents using Azure AI Foundry.
- Develop Agent-to-Agent communication patterns.
- Implement Retrieval Augmented Generation (RAG) architectures.
- Build enterprise Knowledge Management and AI Skill Registry platforms.
- Design multi-agent orchestration frameworks.
Leadership & Stakeholder Engagement
• Serve as the technical authority and subject matter expert for Azure AI and DevOps practices.
• Mentor and guide junior architects, developers, and DevOps engineers.
• Collaborate with business stakeholders, product owners, and vendors to translate requirements into technical solutions.
• Produce architecture documentation, decision records (ADRs), and roadmaps.
• Represent the technology function in enterprise architecture forums and governance boards.
Required Qualifications
• Bachelor's or Master's degree in Computer Science, Information Technology, or a related field.
• 10+ years of experience in cloud engineering and architecture.
• 5+ years of hands-on experience with Microsoft Azure across compute, networking, storage, identity, and data services.
• Proven experience designing and implementing enterprise-grade CI/CD pipelines.
• Strong hands-on expertise with Infrastructure-as-Code (Terraform, Bicep, or ARM).
• Demonstrated experience architecting and deploying AI/ML solutions on Azure (Azure OpenAI, Azure ML, AI Foundry).
• Deep knowledge of DevSecOps principles, tools, and practices.
• Experience with containerisation and orchestration: Docker, Kubernetes (AKS).
• Proficiency in scripting and development: Python, PowerShell, Bash.
• Excellent communication and stakeholder management skills.
Preferred Qualifications
• Microsoft Certified: Azure Solutions Architect Expert.
• Microsoft Certified: DevOps Engineer Expert.
• Microsoft Certified: Azure AI Engineer Associate.
• Experience with Azure API Management (APIM), Event Grid, and Azure Functions.
• Familiarity with Datadog, Prometheus, or equivalent observability platforms.
• Experience in the real estate, retail, or enterprise industry sector.
• Knowledge of agentic AI frameworks and LLM orchestration patterns (LangChain, Semantic Kernel, MCP).
• Background in building Internal Developer Platforms (IDP).
Technical Skills
Domain
Technologies / Tools
Cloud Platform
Microsoft Azure (UAE North / Global)
DevOps & CI/CD
Azure DevOps, GitHub Actions, Jenkins
IaC
Terraform, Bicep, ARM Templates
AI / ML
Azure OpenAI, Azure AI Foundry, Azure ML, Cognitive Services
Containers
Docker, Kubernetes (AKS), Azure Container Apps
Identity & Security
Microsoft Entra ID, Azure Policy, Defender for Cloud
Observability
Datadog, Azure Monitor, Log Analytics, Application Insights
Databases
Azure SQL, PostgreSQL, Cosmos DB
Languages
Python, PowerShell, Bash, YAML
Job Title : Mobile DevOps Engineer
Experience : 5+ Years
Location : Nashik / Pune (Remote – Initial 1 Week Onsite in Nashik)
Job Summary :
We are looking for a Mobile DevOps Engineer with expertise in building and maintaining CI/CD pipelines for iOS and Android applications. The ideal candidate will have hands-on experience with mobile build automation, release management, Azure DevOps, and cloud-based DevOps practices to enable faster, secure, and reliable mobile application delivery.
Mandatory Skills :
Azure DevOps, Mobile CI/CD (iOS & Android), Fastlane/Bitrise/Codemagic, GitHub, Xcode, Gradle, App Store & Google Play release automation, YAML Pipelines, GitHub Actions, Docker, Shell/Python scripting, Mobile DevSecOps, Firebase Crashlytics/Sentry, Agile/Scrum.
Key Responsibilities :
- Design, implement, and maintain CI/CD pipelines for iOS and Android applications.
- Automate build, testing, code signing, and deployment processes.
- Manage app releases on Google Play Store and Apple App Store.
- Configure and maintain Azure DevOps pipelines, GitHub workflows, and release automation.
- Integrate automated testing, security scanning, and quality gates into CI/CD pipelines.
- Monitor build performance, troubleshoot pipeline issues, and optimize release processes.
- Collaborate with Mobile, QA, Backend, and Product teams to streamline delivery.
- Maintain DevOps documentation, release procedures, and deployment best practices.
- Explore AI-powered development tools to improve automation and developer productivity.
Required Skills :
- 5+ years of experience in DevOps, Infrastructure, or Software Engineering.
- 2+ years of hands-on experience with Mobile CI/CD (iOS & Android).
- Strong experience with Azure DevOps, GitHub, Fastlane, Bitrise, Codemagic, or similar tools.
- Solid knowledge of Xcode, Gradle, code signing, provisioning profiles, and keystores.
- Experience with Google Play Store and Apple App Store release automation.
- Familiarity with YAML pipelines, GitHub Actions, Shell/Python scripting, and Docker.
- Knowledge of mobile testing frameworks, DevSecOps practices, and monitoring tools such as Firebase Crashlytics, Sentry, or Datadog.
- Understanding of Agile/Scrum methodologies and strong collaboration skills.
Preferred Skills :
- Experience with Flutter, React Native, or native mobile development.
- Exposure to Azure Cloud services.
- Familiarity with AI-assisted development tools and automation.
Strong Cloud Automation Architect Profile
Mandatory (Experience 1) – Must have 8+ years of total experience with the last 6–7 years focused continuously on hands-on infrastructure automation in enterprise environments
Mandatory (Experience 2) – Must have expert-level Ansible engineering, including advanced Jinja2 templating, dynamic inventory, custom module development, and designing enterprise-grade playbooks, roles, and collections across AWS, Azure, VMware, and REST-integrated environments
Mandatory (Experience 3) – Must have strong Kubernetes and Helm expertise, including custom Helm chart authoring, multi-environment lifecycle management (upgrades, rollbacks, canary releases), namespace management, RBAC, secrets, and network policies — with GitOps adoption via ArgoCD or Flux
Mandatory (Experience 4) – Must have deep Terraform knowledge, including reusable module design covering AWS networking, IAM, EKS, ECS, RDS, and security controls — with CI/CD integration (GitHub Actions or Azure DevOps) including automated testing, linting, drift detection, and policy enforcement
Mandatory (Experience 5) – Must have solid AWS architecture knowledge, including serverless and event-driven automation (Lambda, Step Functions, Event Bridge, SNS/SQS, S3 triggers, Systems Manager), IAM design, cross-account automation across AWS Organizations, and multi-account networking aligned with CIS, NIST, and AWS Well-Architected standards
Mandatory (Experience 6) – Must have experience with security and compliance automation including CIS benchmark enforcement, vulnerability remediation, infrastructure hardening, certificate lifecycle management, and integration with CyberArk, AWS Secrets Manager, and SSM Parameter Store
🚀 We're Hiring | Contract Role
Position: .NET & SQL Developer
📍 Location: Indore
💼 Experience: 6–10 Years
🕒 Shift: 11:30 AM – 8:30 PM IST
If you're available or have suitable consultants on your bench, we'd love to connect.
📩 DM me or share relevant profiles.
#Hiring #DotNet #SQL #ContractJobs #IndoreJobs #ITHiring #HiringNow #BenchSales #TechJobs
Supercharge Your Career as a AI DevOps Engineer at Technoidentity!
At Technoidentity, we're a Data & AI product engineering company with over 15 years of expertise in building durable digital products, intelligent enterprise solutions, and scalable Data & AI platforms. As we continue expanding globally, it's the perfect time to join our team of tech innovators and make a lasting impact.
What’s in it for You?
We are looking for an AI DevOps Engineer with 0–3 years of experience who is passionate about AI, Cloud, DevOps, and Automation. The role involves building, deploying, and managing AI-powered applications, LLM solutions, and cloud-native platforms while ensuring reliability, scalability, security, and observability.
What Will You Be Doing?
- Develop and deploy AI/ML and Generative AI solutions using Python.
- Build applications leveraging LLMs, RAG, and AI agents.
- Create and maintain CI/CD pipelines for AI applications.
- Deploy and manage workloads using Docker and Kubernetes.
- Support cloud platforms (AWS, Azure, or GCP).
- Implement Infrastructure as Code (Terraform) and automation workflows.
- Monitor applications using observability tools such as Prometheus, Grafana, and logging platforms.
- Collaborate with engineering teams to ensure system reliability, performance, and security.
- Contribute to MLOps practices, AI accelerators, and reusable frameworks.
Requirements
What Makes You the Perfect Fit?
- Python programming (mandatory)
- Understanding of Machine Learning, LLMs, Prompt Engineering, and RAG
- Experience with OpenAI, LangChain, LlamaIndex, or Hugging Face
- Docker, Kubernetes, Git, and CI/CD tools
- AWS, Azure, or GCP
- PostgreSQL; MongoDB and Vector Databases are a plus
- Basic knowledge of MLOps, Terraform, and workflow orchestration tools (Airflow/Temporal)
- Familiarity with observability and monitoring tools
Qualifications
- Bachelor's degree in Computer Science, AI, Data Science, IT, or related field
- 0–3 years of experience in AI/ML, Software Engineering, Cloud, DevOps, or related areas
Nice to Have
- Experience with Agentic AI frameworks
- Knowledge of MLOps and AI platform operations
- Exposure to enterprise-grade monitoring, reliability engineering, and security best practices
Principal DevOps Engineer
Note - Screening Requirement: Please note that this position requires a minimum of 8+ years of hands-on Architecting DevOps/SRE/Platform Engineering experience specializing in AWS, EKS/Kubernetes, Terraform, Python, Jenkins, and AI workflows (Mandatory skills) from scratch.
We are looking for an absolute builder who has a proven track record of personally architecting, designing and setting up production-grade AWS EKS clusters entirely from scratch. If your experience is primarily limited to managing, maintaining, or optimizing pre-existing environments that were already stood up by another team, this is not the right opportunity for you.
Who are we?
Securin is an AI-driven cybersecurity company focused on proactive, adversarial exposure and vulnerability management. Our mission is to help organizations reduce cyber risk by identifying, prioritising, and remediating the issues that matter most. Powered by a seasoned team of threat researchers and status as a Certified Naming Authority (CNA), Securin combines artificial intelligence / machine learning, threat intelligence, and deep vulnerability research (including the Dark Web) to deliver an adversarial approach to cyber defense. We help enterprises shift from reactive patching to strategic, risk-based exposure and vulnerability management – driving smarter security decisions and faster remediation.
What do we promise?
We are a highly effective tech-enabled cybersecurity solutions provider and promise continual security posture improvement, enhanced attack surface visibility, and proactive prioritized remediation for every one of our client businesses.
What do we provide?
● A chance to be on the leading edge of cybersecurity and AI
● Ability to have direct impact on company growth and revenue strategy
● An opportunity to mentor and be mentored by experts in multiple disciplines
What do we deliver?
Securin helps organizations to identify and remediate the most dangerous exposures, vulnerabilities, and risks in their environment. We deliver predictive and definitive intelligence and facilitate proactive remediation to help organizations stay a step ahead of attackers. By utilising our cybersecurity solutions, our clients can have a proactive and holistic view of their security posture and protect their assets from even the most advanced and dynamic attacks.
Securin has been recognized by national and international organizations for its role in accelerating innovation in offensive and proactive security. Our combination of domain expertise, cutting-edge technology, and advanced tech-enabled cybersecurity solutions has made Securin a leader in the industry.
Core Technology Stack
AWS , EKS / Kubernetes, Jenkins , Python ,Terraform ,CI/CD , AI
Key Responsibilities:
● Architect and manage the end-to-end SaaS platform infrastructure on AWS, including EKS cluster design, VPC networking, IAM, and multi-region availability.
● Build, maintain, and optimize Jenkins-based CI/CD pipelines and develop Python automation scripts for provisioning, deployments, and runbook automation.
● Define and enforce platform SLOs/SLAs; own the observability strategy across logging, metrics, and tracing.
● Manage and participate in the on-call rotation; act as escalation point for P1/P2 incidents and drive post-incident reviews.
● Drive Infrastructure-as-Code (IaC) practices with Terraform/CloudFormation and champions a culture of automation and operational excellence.
● Collaborate cross-functionally with product, security, and engineering teams to align infrastructure roadmap with business goals.
● Identify opportunities to leverage AI to automate operational and DevOps workflows.
● Design and implement AI-assisted solutions for incident triaging, root cause analysis, log analysis, and performance optimization.
● Drive the adoption of AI-powered tools for infrastructure management, deployment automation, monitoring, and troubleshooting.
● Build intelligent workflows that reduce manual effort in release management, capacity planning, and operational support.
● Integrate AI capabilities into CI/CD pipelines to improve code quality, deployment reliability, and operational efficiency.
● Collaborate with engineering teams to automate repetitive tasks and improve developer productivity.
● Define best practices and governance for the safe and effective use of AI across DevOps processes.
● Measure and report on productivity gains, operational improvements, and cost savings achieved through AI adoption.
Requirements:
● 8+ years of experience in DevOps, SRE, or cloud infrastructure engineering roles.
● Deep hands-on expertise with AWS services (EC2, EKS, RDS, S3, IAM, VPC, CloudFront, Route53, Lambda, etc).
● Strong Kubernetes experience: cluster management, Helm, autoscaling (HPA/KEDA).
● Proficiency with Jenkins for complex CI/CD pipeline design and maintenance.
● Solid Python scripting skills for automation, tooling, and infrastructure management tasks.
● Experience with Infrastructure-as-Code using Terraform and/or AWS Cloud Formation.
● Proven track record of architecting and managing end-to-end SaaS products in a cloud-native environment.
● Strong understanding of networking fundamentals, security best practices, and compliance frameworks (SOC 2, ISO 27001 a plus).
● Hands-on experience with on-call processes and incident management frameworks.
Preferred Qualifications:
● AWS certifications: Solutions Architect Professional, DevOps Engineer Professional, or equivalent.
● Familiarity with service mesh, secrets management (Vault, AWS Secrets Manager), and zero-trust security models.
● Experience with multi-tenant SaaS architectures and tenant isolation strategies.
● Knowledge of FinOps principles and AWS cost management tooling.
● Experience with database DevOps: RDS, Aurora schema migrations, and backup strategies.
Core Competencies:
● Strategic Thinking – ability to translate business goals into scalable technical architecture.
● Operational Excellence – strong bias for reliability, automation, and continuous improvement.
● Communication – ability to clearly articulate complex technical topics to non-technical stakeholders.
● Ownership Mindset – proactively identifies and resolves risks without waiting to be asked.
● Resilience Under Pressure – calm and decisive during incidents; leads by example in high-stress situations.
Why should we connect?
We are a bunch of passionate cybersecurity professionals who are building a culture of security. Today, cybersecurity is no more a luxury but a necessity with a global market value of $150 billion.
At Securin, we live by a people-first approach. We firmly believe that our employees should enjoy what they do. For our employees, we provide a hybrid work environment with competitive best-in-industry pay, while providing them with an environment to learn, thrive, and grow. Our hybrid working environment allows employees to work from the comfort of their homes or the office if they choose to. For the right candidate, this will feel like your second home.
If you are passionate about cybersecurity just as we are, we would love to connect and share ideas
About Ritually
Ritually is building the definitive process discovery platform for back office work. Our product fuses underutilized system telemetry with computer vision to help large enterprise and scaling mid-market companies deeply understand and reimagine their highest value and most repetitive processes for a world where humans and agents work together. We're based in New York and Denver.
We believe deeply in trust (of our customers and each other), craft, customer obsession, and speed.
You'll be joining an AI-native, fast-moving, and repeat founding team. Ritually's founders previously built and exited a startup (Involvio) to Cisco. The company is funded and working with design partners.
The Role
This is a founding full-stack role. You'll be building our application with the founding team from 0-1 You'll own features and infra end to end and across technologies. You'll also own how we build: ensuring our AI development workflow is fast, dependable and secure.
What You'll Do
- Build features end to end, across our desktop client, web app, and the backend services that tie them together.
- Go deep in our cross-platform Electron desktop client the native OS integration that powers on-device capture, and the on-device privacy and redaction that runs before data leaves the machine.
- Move fluidly across technologies and layers, picking up whatever a problem needs rather than staying in one corner of the stack.
- Own the infrastructure and deployment path CI/CD, releases, and the cloud services everything runs on so shipping stays routine and reliable.
- Own how we build: keep our AI development workflow fast, dependable, and secure and establish the conventions, tooling, and guardrails that let a small team build like a much larger one.
- Own observability and on-call basics logging, error tracking, and alerting so we catch issues before our customers do.
- Help set technical direction and keep the codebase legible as we scale the engineering team.
What We're Looking For
- 1-4 years building software across the stack frontend, backend, and the glue between them (internships and meaningful side projects count).
- Comfort working across multiple languages and technologies, and a genuine willingness to learn whatever a problem requires.
- Solid fundamentals in shipping production software version control, testing, debugging, and deployment.
- Experience with, or strong interest in, owning infrastructure and devops: CI/CD, cloud services, and reliability.
- Genuine excitement about AI-native development you've used AI coding agents and want to push how far they can go.
- Comfort with ambiguity and a willingness to own problems end to end in a small team.
Nice to Have
- Experience building desktop applications (Electron or native).
- Experience building internal tooling or automation that made a team measurably faster.
- Experience building applications for on-premise deployment.
- Familiarity with managed backend platforms and AI agent tooling.
- Any prior early-stage startup experience.
- Degree in computer science or a related field.
About the Internship
SkillSecure X is looking for motivated DevOps with AI Interns who are interested in cloud technologies, automation, CI/CD, containerization, and AI-powered DevOps practices. This internship provides hands-on experience through real-world projects and guided mentorship.
Responsibilities
- Assist in building and maintaining CI/CD pipelines.
- Learn and work with Docker, Git, and cloud platforms.
- Support deployment automation and infrastructure management.
- Explore AI tools to improve DevOps workflows and monitoring.
- Monitor application performance and troubleshoot basic issues.
- Complete weekly assignments and document project progress.
Required Skills
- Basic knowledge of Linux and Git.
- Understanding of DevOps concepts and software development lifecycle.
- Familiarity with cloud platforms (AWS, Azure, or Google Cloud) is a plus.
- Basic scripting knowledge (Python, Bash, or Shell).
- Strong problem-solving and analytical skills.
Eligibility
- Undergraduate or postgraduate students.
- Recent graduates.
- Students pursuing Computer Science, Information Technology, Software Engineering, Cloud Computing, DevOps, or related disciplines.
Benefits
- Hands-on DevOps and AI project experience.
- Mentor-guided learning.
- Internship Completion Certificate.
- Letter of Recommendation (based on performance).
- Portfolio-building opportunities.
- Flexible remote internship.
Job Title : React Native SDE 3
Experience : 8+ Years
Employment Type : Contract to Hire (C2H)
Location : Bangalore (Hybrid – 3 Days WFO)
Notice Period : Immediate to 15 Days
Job Summary :
We are looking for an experienced React Native Developer to join its mobile engineering team and help build scalable, high-performance applications used by millions of customers.
This is a founding React Native team opportunity where you'll play a key role in defining architecture, best practices, and driving the migration from native Android/iOS applications to React Native.
Mandatory Skills :
Strong React Native development, JavaScript/TypeScript, Mobile Application Architecture, GraphQL API Integration, Android & iOS Platform Knowledge, Mobile Performance Optimization, Accessibility Standards, CI/CD & DevOps Practices, Agile Methodologies (TDD/BDD), System Design, and experience building scalable consumer-facing mobile applications.
Required Skills :
- Strong hands-on experience in React Native development.
- Experience building and architecting mobile applications.
- Good understanding of Android and iOS platforms.
- Experience with GraphQL APIs.
- Strong knowledge of JavaScript/TypeScript.
- Experience with CI/CD, Agile, and DevOps practices.
- Knowledge of mobile performance optimization and accessibility standards.
- Strong problem-solving and communication skills.
Key Responsibilities :
- Develop and maintain React Native applications.
- Work closely with Product, Backend, QA, and Mobile teams.
- Design scalable and maintainable mobile solutions.
- Participate in coding, testing, deployment, and support activities.
- Contribute to architecture decisions and engineering best practices.
Interview Process :
- Technical Interview (Geektrust)
- Client Coding Round
- Client System Design Round
- Managerial Round (Optional)
About the Role
We are looking for passionate and driven interns across multiple technology domains including Frontend Development, Backend Development, DevOps, AI/ML, and Data Engineering. This internship offers hands-on experience in real-world projects, collaboration with cross-functional teams, and exposure to modern tools and technologies.
Domains & Responsibilities
Frontend Development
- Build responsive and user-friendly web interfaces
- Translate UI/UX designs into functional applications
- Optimize performance and ensure cross-browser compatibility
Backend Development
- Develop APIs and server-side logic
- Work with databases and data storage solutions
- Ensure application security and performance
DevOps
- Assist in CI/CD pipeline setup and automation
- Manage deployments and cloud infrastructure
- Monitor system performance and reliability
AI / Machine Learning
- Develop and train ML models
- Work on NLP, automation, or AI-driven features
- Analyze datasets and evaluate model performance
Data Engineering
- Build and maintain data pipelines (ETL/ELT)
- Ensure data quality and availability
- Work with large datasets and optimize data workflows
Required Skills (Any Domain)
- Frontend: HTML, CSS, JavaScript, React/Vue/Angular
- Backend: Node.js / Python / Java / PHP, APIs, databases
- DevOps: Linux, Git, CI/CD basics, cloud fundamentals
- AI/ML: Python, ML basics, TensorFlow/PyTorch/Scikit-learn
- Data Engineering: SQL, Python, data processing concepts
Good to Have
- Knowledge of Git and version control
- Basic understanding of cloud platforms (AWS/Azure/GCP)
- Problem-solving mindset and willingness to learn
- Exposure to real-world or academic projects
Who Should Apply
- Students or recent graduates in Computer Science, IT, or related fields
- Candidates with strong interest in any of the above domains
- Self-learners with project experience are highly encouraged
Internship Details
- Duration: 3–6 months
- Mode: Remote
- Certificate + PPO (Pre-Placement Offer) based on performance
What You’ll Gain
- Hands-on experience with real projects
- Mentorship from experienced professionals
- Exposure to industry tools and workflows
- Opportunity to convert to a full-time role
DevOps Engineer
AWS Infrastructure, CI/CD & Production Operations
Mumbai (On-site) | Full-time | 2-4 years
About the role:
Unico Connect is an AI-first technology partner that builds custom mobile, web, and AI products for clients across multiple geographies. We are hiring a DevOps Engineer who will own day-to-day cloud infrastructure, deployment automation, and production operations across active customer engagements.
The mandatory requirement for this role is hands-on production experience on AWS, with infrastructure as code, container orchestration, and CI/CD pipelines owned end to end on at least one live customer workload. The role is hands-on. Expect to operate Kubernetes clusters, build CI/CD pipelines, automate environment provisioning, manage TLS and DNS, set up observability, and partner with backend and AI engineers to ship reliably. A typical week includes a Terraform refactor, a deployment pipeline build for a new service, an incident response on a production cluster, and a cost review.
Responsibilities:
- AWS infrastructure: Design and operate production infrastructure on AWS using EC2, EKS or ECS, S3, RDS, IAM, VPC, CloudFront, and Route53. Own configuration, networking, and cost.
- Infrastructure as code: Write and maintain Terraform or Pulumi modules. Drive consistency across environments and tenants through IaC rather than manual configuration.
- Kubernetes and containers: Operate production EKS clusters. Manage Helm charts, Ingress, autoscaling, secrets, and workload isolation.
- CI/CD pipelines: Build and maintain pipelines using GitHub Actions, GitLab CI, or equivalent. Include automated tests, security scans, and rollback paths.
- TLS, DNS, and CDN automation: Automate domain provisioning, TLS issuance (Let's Encrypt, cert-manager, ACM), and CDN configuration (CloudFront, Cloudflare).
- Observability and incident response: Set up monitoring, logging, and alerting using Prometheus, Grafana, ELK, Loki, or CloudWatch. Lead incident response and write postmortems.
- Secrets and security: Manage secrets through Vault, AWS Secrets Manager, or KMS. Apply least-privilege IAM and review access regularly.
- Cost monitoring: Track and optimise AWS spend across environments. Surface waste and propose remediations.
Requirements:
- Hands-on AWS production experience (mandatory). Must have personally operated production workloads on AWS, with responsibility for IaC, deployments, and incident response on at least one live customer or internal-platform deployment. POCs and lab environments do not qualify.
- 2 to 4 years of hands-on DevOps or infrastructure experience. Candidates with slightly less experience but strong demonstrated ownership are welcome to apply.
- AWS depth. Hands-on with EC2, S3, IAM, VPC, EKS or ECS, RDS, CloudFront, and Route53. Working knowledge of CloudWatch and AWS cost tooling.
- Kubernetes in production. Hands-on operation of EKS or equivalent. Comfort with Helm, Ingress controllers, autoscaling, and resource quotas.
- Infrastructure as code. Strong with Terraform (preferred) or Pulumi. Modular code, state management, and review discipline.
- CI/CD pipelines. Production experience with GitHub Actions, GitLab CI, or equivalent. Comfort with multi-environment pipelines and release strategies.
- Scripting and automation. Strong Bash and Python (or Go) for tooling. Linux fluency at the command line.
- Observability stack. Hands-on with Prometheus, Grafana, ELK or Loki, and at least one APM tool (Datadog, New Relic, or equivalent).
- Networking, TLS, and security fundamentals. Comfortable with DNS, TLS certificate lifecycle, VPC peering, and security groups.
Nice to have: multi-tenant SaaS infrastructure experience; service mesh (Istio, Linkerd); GitOps (ArgoCD, Flux); sandboxed execution environments (Firecracker, gVisor); exposure to platform engineering or developer-platform teams.
Mobile Automation Test Engineer
Experience 5 – 12 Years
Location
Pune | Hyderabad | Vizag
About the Role
We are seeking an experienced Mobile Automation Testing professional with strong hands-on expertise in building and executing automation frameworks for mobile applications. This role is ideal for candidates who enjoy working in fast-paced environments and collaborating closely with cross-functional teams to ensure high-quality mobile solutions.
Key Responsibilities
- Design, develop, and execute mobile automation test scripts using Appium
- Build and maintain automation frameworks using Java and Selenium
- Perform automated testing for Android and/or iOS applications
- Integrate automation scripts with CI/CD pipelines in Agile or DevOps environments
- Analyze test results, identify defects, and work with development teams for resolution
- Participate in sprint planning, daily stand-ups, and other Agile ceremonies
- Ensure test coverage across functional, regression, and integration testing
Must-Have Skills
- 5+ years of hands-on experience in Mobile Automation Testing
- Strong programming experience in Java
- Solid expertise in Selenium and Appium
- Experience testing mobile applications (Android / iOS)
- Good communication and coordination skills
- Practical exposure to Agile and DevOps methodologies
Additional Skills (Good to Have)
- Experience with CI/CD tools like Jenkins, Git, Maven/Gradle
- Knowledge of mobile device farms or cloud testing tools
- Exposure to API testing tools (Postman, REST Assured)
- Understanding of mobile test strategies and best practices
Who Should Apply
- Automation testers looking to specialize in mobile platforms
- Candidates comfortable working in enterprise-scale testing environments
- Professionals who can independently handle automation deliverables
Senior MLOps Engineer
LLM Operations, Observability & Eval Infrastructure
📍 Mumbai (On-site) | Full-time | 5-7 years
About the Role:
Unico Connect is an AI-first technology partner that builds custom mobile, web, and AI products for clients across multiple geographies.
We are hiring a Senior MLOps Engineer for a dedicated client engagement focused on building an AI-powered application builder platform. The platform consumes LLMs at scale through provider APIs.
This role owns the operational discipline around production LLM consumption - increasingly called LLMOps - covering observability, evaluation infrastructure, model lifecycle, cost operations, prompt deployment, and agent run reliability.
The mandatory requirement is hands-on production experience operating LLM-backed systems, with a strong DevOps or SRE foundation. This is not a model training or ML science role.
The work is making the system around the AI engineer's designs observable, controlled, reliable, and economically accountable. You will pair daily with the Senior AI Engineer, who designs prompts, evals, and agent behaviour - you operationalise those systems for production.
A typical week includes a tracing audit on a degraded agent run, an eval pipeline build for a new model release, a cost attribution review, and a staged prompt rollout.
Responsibilities:
Observability and Tracing
Build and own end-to-end tracing for agent runs: every prompt, response, tool call, token count, latency, and cost, linked to user session and project.
Stand up and operate LLM observability tooling (Langfuse, LangSmith, Braintrust, or Arize Phoenix).
Make debugging a single bad agent run among thousands a routine workflow through searchable traces, failure taxonomies, and dashboards segmented by task type.
Evaluation Infrastructure as a Production System
Operationalise the eval suite designed by the Senior AI Engineer: automated execution in CI on every prompt or model change, with results stored and trended over time.
Implement regression gates that block quality-degrading changes from shipping.
Build production sampling to continuously score a sample of real agent runs and catch quality drift that offline evals miss.
Model Lifecycle Management
Pin model versions, never "latest".
Own the upgrade process: run the eval suite against new model releases and manage eval-gated migrations.
Maintain fallback chains across providers for graceful degradation or queueing during outages.
Track provider deprecation schedules and plan migrations ahead of forced cutoffs.
Cost Operations
Implement per-user and per-task cost attribution - token spend is the platform's largest variable cost and requires the same rigour as cloud cost management.
Set up budget alerts and anomaly detection so a single user or bug cannot burn significant spend overnight.
Monitor prompt cache hit rates and quantify savings.
Manage capacity planning around provider rate limits, including quota negotiation and throughput tiering.
Prompt and Configuration Deployment
Treat prompts as production artifacts: version control for prompts and agent configurations, staged rollout infrastructure (deploy a prompt change to a percentage of traffic before full rollout), A/B testing infrastructure, instant rollback, and audit history covering which prompt version served which user and when.
Reliability Engineering for Agent Runs
Agent runs are long, stateful, and failure-prone.
Own retry and resume semantics so a run that fails mid-way does not restart from scratch.
Implement timeouts and circuit breakers on provider calls, dead-letter handling for failed runs, and queue and concurrency management for agent workloads.
SLO Ownership and Incident Response
Define and track SLOs for agent run latency and completion rates.
Lead incident response when SLOs are breached.
Write postmortems.
Surface reliability risks proactively before they reach users.
Safety and Compliance Operations
Run the moderation pipeline (prompt and output classification) in production.
Monitor for abuse patterns and own incident response when the agent misbehaves at scale.
Maintain audit logs and implement data retention and residency policies for prompts and generated code as enterprise requirements emerge.
AI-Assisted Engineering Discipline
Use Claude, Cursor, and similar tools day to day for infrastructure code, scripts, and pipelines.
Set the team standard for safe use, review, and validation of AI-generated infrastructure before it ships.
Requirements:
Hands-on production ownership of LLM-backed systems in operation (mandatory).
Must have personally shipped and operated at least one LLM-powered system in production, with operational responsibility including oncall, incident response, and reliability ownership.
Alternatively: strong DevOps or SRE background with demonstrated hands-on familiarity with LLMOps tooling (Langfuse, LangSmith, Braintrust, Arize, or equivalent).
POCs and lab work do not qualify.
5+ years of overall engineering experience
With at least 2 years in DevOps, SRE, platform engineering, or LLM operations roles.
This is not an ML science role.
A DevOps or SRE background with a substantive pivot into LLMOps is a strong qualification.
Observability and Tracing Depth
Production experience with LLM observability tooling - Langfuse, LangSmith, Braintrust, or Arize Phoenix.
Comfortable instrumenting with OpenTelemetry, Prometheus, and Grafana.
Able to build and search trace pipelines, define failure taxonomies, and surface quality signals from production traffic.
CI/CD and Quality Gate Experience
Strong with GitHub Actions or GitLab CI.
Experience building automated quality gates: eval-gated pipelines, regression enforcement, or coverage gates that block degrading changes from shipping.
Cost Management and Attribution for Usage-Based Services
Experience owning cost attribution for cloud API spend or equivalent.
Comfortable with budget alerts, anomaly detection, and per-user or per-task cost breakdowns.
Reliability Engineering for Long-Running, Stateful Workloads
Experience with queues, retry patterns, idempotency, and failure recovery on asynchronous or multi-step workloads.
Comfortable defining SLOs and being accountable for them on production systems.
Multi-Provider API Management
Familiarity with LLM provider rate limits, version pinning, fallback chains, and quota management across OpenAI, Anthropic, Google, or equivalent.
Infrastructure as Code and Deployment Automation
Hands-on with Terraform or Pulumi and Docker.
AWS working knowledge (EC2, S3, IAM, EKS or ECS).
Strong with CI/CD for deploying services and configuration changes safely.
Nice to Have
- Experience with prompt A/B testing or staged rollout infrastructure
- Workflow orchestration (BullMQ, Temporal, Celery)
- Content moderation pipeline experience
- Data residency and compliance requirements for AI systems
- Kubernetes (EKS) in production
- AWS certifications
About CNH
Connect and Heal (CNH) is a healthcare organization focused on transforming healthcare delivery through technology-enabled solutions, clinical excellence, and exceptional patient experiences. We build scalable digital platforms that support healthcare operations, patient engagement, telehealth, and enterprise healthcare services.
Role Summary
We are seeking a highly motivated Engineering Lead to drive the design, development, and delivery of scalable technology solutions that support CNH's business growth and healthcare operations. The role requires a strong blend of technical expertise, people leadership, architecture oversight, and execution excellence.
The Engineering Lead will lead engineering teams, collaborate with Product, Operations, Clinical, and Business stakeholders, and ensure delivery of high-quality, secure, and scalable products.
Key Responsibilities
Technical Leadership
- Lead the design, development, and deployment of scalable applications and platforms.
- Drive engineering best practices, coding standards, architecture reviews, and technical governance.
- Ensure system reliability, performance, security, and scalability.
- Evaluate and implement emerging technologies to improve product capabilities and engineering efficiency.
Team Leadership
- Build, mentor, and manage high-performing engineering teams.
- Conduct performance reviews, coaching, and career development discussions.
- Foster a culture of accountability, innovation, collaboration, and continuous learning.
- Support hiring and talent development initiatives.
Product & Delivery Management
- Partner with Product Managers and business stakeholders to translate business requirements into technical solutions.
- Drive sprint planning, prioritization, estimation, and delivery management.
- Ensure timely and high-quality execution of engineering roadmaps.
- Manage risks, dependencies, and technical debt effectively.
Stakeholder Management
- Collaborate with Operations, Clinical, Customer Success, and Business teams.
- Communicate project progress, risks, and outcomes to leadership.
- Support strategic initiatives through technology-driven solutions.
Quality & Compliance
- Ensure adherence to information security, data privacy, and healthcare compliance requirements.
- Establish monitoring, testing, and release management processes.
- Drive automation, observability, and operational excellence.
Qualifications
- Bachelor's or Master's degree in Engineering, Computer Science, or related field.
- 8–12+ years of software engineering experience.
- 3–5+ years of experience leading engineering teams.
- Experience building and scaling enterprise-grade applications.
Technical Skills & Expertise
- Strong proficiency in modern programming languages such as Java, Python, Node.js, Golang, or .NET.
- Hands-on experience in designing and managing RESTful APIs, API Gateways, and Microservices Architecture.
- Strong knowledge of CI/CD Automation using tools such as Jenkins, GitHub Actions, GitLab CI/CD, Azure DevOps, or similar platforms.
- Experience in DevOps practices, including Infrastructure as Code (IaC), containerization (Docker), orchestration (Kubernetes), and cloud-native deployments.
- Expertise in cloud platforms such as AWS, Azure, or GCP.
- Experience implementing monitoring, logging, and observability solutions.
- Familiarity with modern databases, distributed systems, and high-availability architectures.
- Exposure to AI/ML tools, Generative AI solutions, LLMs, AI-assisted development tools (GitHub Copilot, Cursor, ChatGPT, etc.), and their integration into enterprise applications.
- Understanding of healthcare interoperability standards, APIs, and data security principles is desirable.
Additional Responsibilities
- Drive adoption of AI-powered engineering practices to improve developer productivity, code quality, and operational efficiency.
- Establish and maintain robust CI/CD pipelines to enable faster and more reliable software releases.
- Lead DevOps transformation initiatives, promoting automation, infrastructure scalability, and engineering excellence.
- Define and govern enterprise-wide API strategy, standards, security, and lifecycle management.
- Evaluate emerging technologies and AI innovations to support CNH's digital healthcare roadmap.
Preferred Candidate Profile
- 10–15 years of experience in software engineering, with 4–6 years in engineering leadership roles.
- Proven experience leading teams building cloud-native, API-first, AI-enabled platforms.
- Strong background in DevOps, CI/CD automation, platform engineering, and digital transformation initiatives.
- Experience managing engineering teams delivering mission-critical products at scale.
CNH offers an opportunity to build impactful healthcare technology solutions that improve access, quality, and patient outcomes while working with a passionate and mission-driven team.
Full Stack Engineers
Location: Chennai
Experience: 3+ Years
Required Skills:
- Strong knowledge of C#, .NET Core, Azure DevOps
- Working knowledge of the JS frameworks – React, Angular etc
- Experience in container-based development, AKS, Service Fabric etc
- Experience in messaging queue like RabbitMQ, Kafka
- Experience in Azure Services like Azure Logic Apps, Azure Functions
- Experience in databases like SQL Server, PostgreSQL
Job Description:
We are looking for experienced product development engineers/experts who could join our cloud product engineering team to build the next gen applications for our global customers. If you are a technology enthusiast and have passion to develop enterprise cloud products with quality, security, and performance, we are eager to discuss with you about the potential role.
Responsibilities:
- Understand the business requirements and technical constraints and architect/design/develop.
- Participate in the complete development life cycle.
- Review the architecture/design/code of self and others.
- Develop enterprise application features/services using Azure cloud services, C# .NET Core, ReactJS etc, implementing DevSecOps principles.
- Own and be accountable for the Quality, Performance, Security and Sustenance of the respective product deliverables.
- Strive for self-excellence along with enabling success of the team/stakeholders.
Requirements:
- 3 to 12 years of experience in developing enterprise software products
- Knowledge of reporting solutions like PowerBI, Apache SuperSet etc
- Knowledge of Micro-Services and/or Micro-Frontend architecture
- Knowledge of Code Quality, Code Monitoring, Performance Engineering, Test Automation Tools
Role: Azure Devops Consultant
Location: Bangalore
Fulltime
Work Mode: 5 Days Work From Office.
Responsibilities:
· Design and architect CI/CD pipelines and development automation solutions.
· Lead implementation of build versioning, packaging, and dependency
management strategies
· Define standards for code quality, security scanning, and automated testing
integration
· Create reference architectures for development workflows and automation
patterns
· Collaborate with security teams to implement "shift-left" security practices
· Guide implementation of A/B testing and feature flagging architectures
Skills:
· Strong expertise in Azure DevOps,PAAS, Git, CI/CD practices
· Deep understanding of SDLC and automation
· Containerization (Docker, Kubernetes)
· Security scanning tools (Veracode, SonarQube)
· Build tools and artifact management
· IaC (Terraform, ARM templates)
Key Focus Areas:-
• Scalable SaaS and distributed systems
• Microservices architecture
• Backend performance and API optimization
• Cloud infrastructure and DevOps
• Multi-tenant platforms
• Reliability, scalability, and engineering governance
Main Responsibilities:-
• Design fault-tolerant backend systems
• Lead architecture decisions and technical planning
• Improve system performance, scalability, and deployment workflows
• Guide senior developers and enforce engineering standards
• Reduce technical debt and solve complex production issues
• Collaborate with DevOps on CI/CD, monitoring, and AWS infrastructure.
• You will work on complex, enterprise-scale systems including large SaaS platforms, distributed backend architectures, pricing intelligence engines, CRM solutions, 3D configuration ecosystems, route optimization modules, payment integrations, high-volume APIs, and modern cloud infrastructure.
• We are not looking for a “people manager who used to code.
Required Skills:-
• Backend: Node.js, TypeScript, REST APIs, event-driven systems
• Frontend Understanding: React.js architecture concepts
• Databases: PostgreSQL, MySQL, Redis, query optimization
• Cloud/DevOps: AWS, Docker, CI/CD, Kubernetes (preferred)
• System Design: Microservices, distributed systems, high-availability SaaS platforms
Preferred Experience:-
• SaaS product companies
• Enterprise CRM systems
• Large-scale production platforms
• Pricing engines, WebGL/3D systems, AI-assisted workflows
• U.S.-based product engineering environments
Experience Needed:-
• 10–15 years in software engineering
• 3+ years in architect/principal engineer/engineering lead roles
• Proven experience building scalable production systems
Ideal Candidate:-
A technically strong, ownership-driven architect who:
• Thinks strategically
• Stays calm under pressure
• Solves complex engineering problems
• Mentors teams effectively
• Builds long-term scalable systems rather than quick fixes
Required Skills
● Experience: Minimum of 5 years of professional experience in a DevOps Engineer
role
● Cloud Proficiency: Proven experience with at least one major cloud provider (AWS,
Azure, or GCP).
● Scripting & Programming: Strong scripting skills in languages such as Bash, Python,
or Go.
● IaC Tools: Hands-on experience with Terraform.
● Container Technology: Expertise in Docker and Kubernetes.
● CI/CD Tools: Proficient with CI/CD platforms like Jenkins, GitLab CI, or Travis CI.
● Configuration Management: Experience with configuration management tools like
Ansible, Chef, or Puppet.
● Version Control: Strong knowledge of Git and branching strategies.
● Problem-Solving:Excellent problem-solving abilities and a commitment to automation
and continuous improvement.
Job Title: Sr. AI Ops Engineer
Experience: 8+ Years
Location: Remote
We’re looking for a hands-on Lead AI Ops Engineer to drive AI-powered automation across enterprise infrastructure. This role involves working closely with senior stakeholders to identify opportunities and implement Agentic / AIOps solutions at scale.
What You’ll Do
- Identify automation opportunities across Cloud & Enterprise Infrastructure
- Design & build AI/Agentic solutions aligned with AIOps frameworks
- Collaborate with Principal Engineers / Directors on high-impact initiatives
- Own end-to-end delivery: Problem → Solution → Deployment
Tech Stack
- Must: Python
- Good to have: JavaScript, Java, Scala, R
- AI/ML: Databricks, MLflow
- Frameworks: LangChain, LangGraph, LLaMA, Cohere, DBRX
- Cloud: AWS, Azure, GCP
- Others (Plus): Google ADK, MCP, A2A, Argo
Ideal Profile
- Strong engineering maturity & ownership
- Experience in AI-driven automation / AIOps
- Ability to work directly with senior technical stakeholders
Why Apply?
Work on cutting-edge AI Ops & Agentic automation in a large-scale enterprise environment with high visibility.

It’s a global digital engineering and technology (MNC)
Job Details:
- Role: Principal Engineer, Salesforce Health cloud
- Experience: 9-11 Years
- Employment Type: Full-time
- Work Mode: Remote
REQUIREMENTS:
- Should be able to lead end-to-end Salesforce Health Cloud solution architecture and delivery for large, enterprise healthcare implementations.
- Deep expertise in Health Cloud features, including Patient/Member Profiles, Care Plans, Care Programs, Care Teams, Health Timelines, Assessments, Care Gaps, and Utilization Management.
- Strong leadership in translating complex healthcare business requirements into scalable Salesforce Health Cloud solutions.
- Should be able to drive architectural decisions covering Salesforce core, Health Cloud extensions, integrations, data models, and automation.
- Should have hands-on leadership experience with Apex, Lightning Web Components (LWC), Flows, Triggers, Batch Apex, and Salesforce APIs.
- Should be able to define and govern integration architecture using REST/SOAP APIs, middleware, and healthcare standards such as FHIR and HL7.
- Must have strong understanding of EHR/EMR integrations, payer/provider workflows, and patient/member lifecycle management.
- Should be able to lead all phases of the delivery lifecycle requirements, solution design, development, testing, deployment, and post-go-live support.
- Provide technical governance, conduct design/code reviews, and manage architectural risks and technical debt.
- Should be able to lead and mentor multi-shore Salesforce teams, including developers, admins, QA, and offshore delivery teams.
- Should have hands-on experience with DevOps and CI/CD tools such as Copado, Git, Azure DevOps, Jenkins.
- Should be able to collaborate closely with program managers to manage scope, timelines, risks, and delivery commitments.
- Should be able to drive innovation and continuous improvement, including PoCs, adoption of new Salesforce Health Cloud features, and platform enhancements.
- Should be able to participate in roadmap planning, release strategy, and long-term platform evolution discussions.
- Must have strong experience working in Agile / Scaled Agile environments, supporting sprint planning and release management.
- Excellent communication, leadership, decision-making, and stakeholder management skills.
- Mentoring the team members to meet the client's needs and holding them accountable for high standards of delivery.
- Being able to understand and relate technology integration scenarios and be able to apply these learnings in complex troubleshooting scenarios.
RESPONSIBILITIES:
- Understanding the client’s business use cases and technical requirements and be able to convert them into technical design which elegantly meets the requirements.
- Mapping decisions with requirements and be able to translate the same to developers.
- Identifying different solutions and being able to narrow down the best option that meets the client’s requirements.
- Own overall Health Cloud technical strategy, ensuring scalability, security, and compliance with healthcare regulations (HIPAA/PHI).Showcasing a consulting mindset by acting as a solution provider rather than an order taker.
- Act as a trusted advisor for business stakeholders, architects, and senior leadership on Health Cloud capabilities and best practices.
- Review and approve high-level and low-level technical designs, ensuring alignment with enterprise architecture standards.
- Ensure adoption of Salesforce best practices, governor limit optimization, performance tuning, and code quality standards.
- Guide teams on Salesforce security model, data access, sharing rules, consent management, and compliance needs.
- Identifying project/service stakeholders at an early stage and working with them to ensure that the
- Act as a bridge between business stakeholders and Salesforce technical teams.
- Collaborate with Salesforce Architects and Developers to design scalable, best-practice solutions.
- Validate solution feasibility against Salesforce platform capabilities and project constraints.
- Prepare and maintain process flows, business process models, functional diagrams, and documentation.
- Participate actively in Agile/Scrum ceremonies including sprint planning, backlog grooming, daily stand-ups, sprint reviews, and retrospectives.
- Act as a trusted advisor to users by explaining Salesforce capabilities and solution trade-offs.
- Support user adoption through training materials, user guides, and release notes.
- Analyze existing Salesforce implementations and recommend process improvements and optimization opportunities.
- Defining guidelines and benchmarks for NFR considerations during project implementation
- Writing and reviewing design document explaining overall architecture, framework, and high-level design of the application for the developers
- Reviewing architecture and design on various aspects like extensibility, scalability, security, design patterns, user experience, NFRs, etc., and ensure that all relevant best practices are followed.
- Developing and designing the overall solution for defined functional and non-functional requirements; and defining technologies, patterns, and frameworks to materialize it
- Understanding and relating technology integration scenarios and applying these learnings in projects
- Resolving issues that are raised during code/review, through exhaustive systematic analysis of the root cause, and being able to justify the decision taken.
- Carrying out POCs to make sure that suggested design/technologies meet the requirements.
Qualifications
Bachelor’s or master’s degree in computer science, Information Technology, or a related field.
We are seeking a highly skilled and motivated Business Analyst to join our team delivering bespoke
software solutions to our clients. In this role, you will translate business needs into clearly written
user stories with testable acceptance criteria that act as the contract between client intent and what
the development team builds. You will be the bridge between Clients, Developers and the Project
Manager — and the person accountable for ensuring that what gets delivered is what the business
actually asked for.
Experience of documenting key projects artefacts and have a good understanding of SDLC. Basic
understanding of the Microsoft power platform tools.
Responsibilities:
• Being part of a team with a strong sense of product ownership and commitment to build
scalable, extensible and robust software and reports.
• Delivering outcomes that are clearly defined, using discretion over how to achieve them.
• Making suggestions for improvements to the work of the team, based on previous
experience and knowledge of similar situations.
• Involved in the development of complete solutions, from initiation to handover, participating
in a cross-functional team across requirements gathering, data analysis, process mapping,
user-story authoring, and user acceptance testing.
• Writing clear, testable acceptance criteria for every user story - using Given/When/Then or
equivalent structured formats - and owning them as the definition of 'done' that the
development team builds against and the client signs off on.
• Building and maintaining strong relationships with key stakeholders/clients
Requirements:
• Proven experience in a business analyst or product analyst role delivering bespoke software
solutions
• Demonstrable track record of writing high-quality acceptance criteria - specific, testable, and
traceable back to business intent - using Given/When/Then or equivalent. Able to show
examples of stories where their criteria caught gaps before development started.
• Experience of quickly building trusted stakeholder relationships at all levels of the business
• Experience of working with both waterfall, agile and hybrid methodologies
• Experience of technical writing and presentation skills using tools such as: PowerPoint, Word,
Visio
• Strong understanding and applied knowledge of User Experience and Customer Journey
concepts.
• Confident working with ambiguity
• Ability to work independently and be self-motivated.
• Experience in conceptual modelling; ability to see the big picture and envision possible
solutions.
• Outstanding verbal and non-verbal communication skills
• Ability to facilitate a team to consensus on scope and requirement decisions.
• Strong grasp of definition-of-done and requirements traceability - able to demonstrate how
each delivered feature maps back to a documented business need and its acceptance
criteria.
• Collaborate with cross-functional systems teams to document systems requirements and
understand their impact on other teams.
• Work with development teams to translate business operational and functional
requirements.
• Produce comprehensive documentation for systems and processes, using Business Process
Modelling techniques
• Experience in UML Modelling (Use Case Diagrams, Sequence Diagrams, Activity Diagrams,
Class Diagrams etc.)
Other:
• Bachelor’s degree: A degree in Business studies, Information Systems or a related field is
preferred but not essential
• Exposure to Backlog user requirement management, tools such as Jira, Azure devOps, would
be beneficial
• Business Acumen: Ability to understand and translate business requirements into data driven solutions. Strong analytical and problem-solving skills with a keen attention to detail.
• Communication Skills: Excellent verbal and written communication skills. Ability to effectively
communicate complex data concepts to both technical and non-technical stakeholders.
• Team Player: Capable of working collaboratively in a team environment. Willingness to share
knowledge and assist colleagues.
• Time Management: Strong organisational skills and the ability to manage multiple projects
simultaneously. Proven ability to meet deadlines and deliver high-quality work.
Join our dynamic team and help drive data-centric decision-making across the organisation.
Associate Principal Engineer, Linux Administrator
Location: Bengaluru, India (Hybrid)
Employment Type: Full-time
Experience:9-11 years
Job description
REQUIREMENTS:
- Strong experience in DevOps, Platform Engineering, and Infrastructure Automation
- Deep hands-on expertise in Linux Administration (RHEL, CentOS, Ubuntu) – OS hardening, security, patching, and performance management (Must Have)
- Strong experience with Cloud Technologies – Public & Private Cloud environments (Must Have)
- Hands-on experience with Infrastructure as Code (IaC) using Terraform (Must Have)
- Strong automation expertise using Ansible for configuration management and infrastructure provisioning (Must Have)
- Experience building and managing CI/CD pipelines and end-to-end deployment automation
- Strong experience with Kubernetes administration, orchestration, and cluster management (Must Have)
- Hands-on experience with Docker containerization and Helm package management
- Experience managing large-scale development and infrastructure environments
- Strong understanding of Networking concepts, connectivity, design, troubleshooting, and network automation
- Experience with Observability & Monitoring tools and best practices
- Experience with Proxmox virtualization platform administration and management
- Knowledge of Edge Technologies and distributed infrastructure environments
- Basic understanding and administration of Active Directory (AD)
- Experience implementing AI-driven Automation solutions and operational efficiencies
- Strong understanding of infrastructure security, compliance, and governance
- Experience working in Agile/Scrum environments
- Strong troubleshooting, analytical, and problem-solving skills
- Excellent communication and stakeholder management skills
RESPONSIBILITIES:
- Design, build, and manage scalable infrastructure platforms across cloud and on-premise environments
- Administer and maintain Linux servers including security hardening, patching, performance tuning, and troubleshooting
- Develop and manage Infrastructure as Code (IaC) solutions using Terraform
- Automate infrastructure provisioning, configuration management, and operational tasks using Ansible
- Design, implement, and maintain CI/CD pipelines for application and infrastructure deployments
- Deploy, manage, and optimize Kubernetes clusters and containerized workloads
- Manage Docker environments and Helm-based application deployments
- Design and implement network solutions ensuring security, reliability, and scalability
- Monitor infrastructure health, performance, and availability using observability and monitoring tools
- Manage and support Proxmox virtualization environments
- Implement AI-driven automation initiatives to improve operational efficiency and reduce manual effort
- Support edge infrastructure deployments and distributed computing environments
- Collaborate with development, security, and operations teams to deliver reliable platform services
- Troubleshoot production incidents and perform root cause analysis
- Define infrastructure standards, automation frameworks, and operational best practices
- Ensure high availability, scalability, security, and reliability of infrastructure platforms
- Mentor junior engineers and provide technical leadership on DevOps and platform engineering initiatives
- Participate in Agile ceremonies and contribute to continuous improvement initiatives
- Work closely with stakeholders to understand infrastructure requirements and deliver optimal solutions
Qualifications
Bachelor’s or master’s degree in computer science, Information Technology, or a related fields
Remote
3 - 6 yrs
Full-time Contract Role
We are looking for a skilled Python Developer with strong experience in Django, Flask, REST APIs, and scalable backend development. The ideal candidate should have hands-on experience building enterprise applications, API-driven systems, AI integrations, automation tools, and cloud-based solutions.
The candidate should be passionate about clean architecture, problem-solving, and developing scalable web applications with modern technologies.
Key Responsibilities
- Design, develop, and maintain scalable backend applications using Python and Django/Flask.
- Build and optimize RESTful APIs for web and mobile applications.
- Work on database design, query optimization, and performance tuning.
- Integrate third-party APIs such as payment gateways, APIs.
- Develop AI/ML-based solutions including RAG applications and automation systems.
- Collaborate with frontend developers using React/Angular integration.
- Implement authentication, RBAC, and security best practices.
- Debug, troubleshoot, and optimize existing applications.
- Participate in code reviews and technical architecture discussions.
- Prepare technical documentation and deployment workflows.
- Work with asynchronous processing, queues, sockets, and real-time systems.
Required Skills
- Strong experience in Python development using Django, Flask, FastAPI, REST API development, SQLAlchemy, and scalable backend architecture.
- Experience with frontend technologies, UI integration, and responsive web development.
- Strong knowledge of MySQL, PostgreSQL, SQLite, database design, query optimization, and performance tuning.
- Experience working with LangChain, RAG applications, LLM integrations, Scikit-learn, and machine learning algorithms.
- Experience in backend integrations, automation, real-time systems, and web scraping.
- Knowledge of Azure cloud services, SSO implementation, cloud deployment, and scalable application architecture.
Preferred Qualifications
- Experience working on SaaS products and enterprise platforms.
- Experience with cloud deployment and scalable architecture.
- Familiarity with CI/CD pipelines and DevOps practices.
- Experience handling production-level applications.
- Strong analytical and debugging skills.
Soft Skills
- Strong communication skills.
- Team collaboration and leadership abilities.
- Problem-solving mindset.
- Client handling experience.
- Ability to work independently.
Education
Bachelor’s Degree in Computer Applications, Computer Science, IT, or related field.
Ideal Candidate Profile
- We are looking for a strong Python developer with full stack development experience and strong Django backend expertise in scalable API-driven architecture.
- The ideal candidate should be passionate about building scalable and high-performance applications, possess strong analytical and debugging skills, and be capable of working in a fast-paced and dynamic environment.
Job Summary
Contribute to the development and optimization of web applications by combining DevOps Solutions, React.js, and full stack Java skills. The role involves designing, implementing, and maintaining software solutions that meet the business requirements efficiently. (1.) Key Responsibilities
1. Collaborate with cross functional teams to define, design, and ship new features.
2. Build and optimize web applications using react.js and full stack java technologies.
3. Implement devops practices to automate processes and improve overall software development efficiency.
4. Troubleshoot and debug applications to ensure optimal performance and reliability.
5. Conduct code reviews and provide constructive feedback to team members.
6. Stay updated on industry trends and new technologies to continuously improve development processes.
Skill Requirements
1. Proficient in devops solutions to streamline software development and deployment processes.
2. Strong experience in react.js for building interactive user interfaces.
3. Expertise in full stack java development, including backend and frontend technologies.
4. Solid understanding of database management and integration.
5. Ability to work in an agile environment and deliver high-quality code within agreed timelines.
6. Excellent problem-solving skills and attention to detail.
7. Good communication skills to collaborate effectively with team members and stakeholders.
Certifications: Relevant certifications in DevOps, React.js, or Java technologies are a plus.
Skill (Primary):
Modern Application Development-Full Stack Development-Java Full Stack
Removal Date
29-May-2027
Strong Principal DevOps Engineer Profile
Mandatory (Experience 1): Must have 10+ years in DevOps / SRE / Infrastructure roles with hands-on experience (clear scale signals like traffic, uptime, latency, infra size should be mentioned) in B2B SAAS companies
Mandatory (Experience 2): Must have worked in Principal / Staff / Lead DevOps / SRE / Platform Engineer role and demonstrated org-level ownership - setting infra roadmap, defining DevOps charter, or structuring the platform function not just domain-level technical ownership
Mandatory (Experience 3): Must show evidence of strategic authorship, defined multi-year infra/platform strategy, drove company-wide architectural shifts as an initiator (not implementer), or directly interfaced with VP Eng / CTO / product leadership on infra direction
Mandatory (Experience 4): Must have B2B SaaS company experience with multi-tenant architecture OR multiple production stacks (multi-env / multi-client systems)
Mandatory (Tech Skills 1 - Cloud & Infra): AWS (VPC, EKS, EC2, RDS, networking), Kubernetes (EKS) at scale, Designing high availability, multi-region systems
Mandatory (Tech Skills 2 - Automation & IaC): Terraform (must-have), Helm / GitOps, Strong scripting (Python / Go / Bash)
Mandatory (Tech Skills 4 - Reliability & Observability): SRE principles (SLOs, SLIs, error budgets), Monitoring tools (Prometheus, Grafana, Datadog), Alerting, on-call, incident management
Mandatory (Leadership): Must demonstrate leadership experience in an individual contributor capacity having mentored senior engineers, driven cross-team technical alignment, or anchored org-wide initiatives without having moved into a people management or engineering manager role
Mandatory (Company): Strong B2B SaaS product companies only
Role & Responsibilities
1. Ansible Automation Design and maintain enterprise-grade playbooks, roles, and collections. Automate OS patching, configuration drift correction, security hardening, and compliance enforcement across AWS, Azure, VMware, and REST-integrated environments. Combine Ansible with Terraform for seamless post-provisioning configuration.
2. AWS Cloud Automation Architect serverless and event-driven automation using Lambda, Step Functions, EventBridge, SNS/SQS, S3 triggers, and Systems Manager. Build scalable, cross-account automation across AWS Organizations with proper IAM boundaries. Align all implementations with CIS, NIST, and AWS Well-Architected standards.
3. Kubernetes & Helm Develop and own custom Helm charts for multi-environment Kubernetes deployments. Manage the full Helm lifecycle including upgrades, rollbacks, and canary releases. Drive GitOps adoption using ArgoCD or Flux. Automate namespace management, RBAC, secrets, and network policies.
4. Infrastructure as Code & CI/CD Build reusable, versioned Terraform modules covering AWS networking, IAM, EKS, ECS, RDS, and security controls. Implement CI/CD pipelines for IaC using GitHub Actions or Azure DevOps — with automated testing, linting, drift detection, and policy enforcement baked in.
5. Security & Compliance Automation Automate CIS benchmark enforcement, vulnerability remediation, and infrastructure hardening. Integrate with CyberArk, AWS Secrets Manager, and SSM Parameter Store. Implement policy-as-code using OPA, Gatekeeper, or Conftest. Automate certificate lifecycle management end to end.
6. Observability & Auto-Remediation Connect automation workflows to New Relic and LogicMonitor for telemetry-driven triggers. Build self-healing routines that detect, diagnose, and resolve incidents automatically. Convert operational runbooks into fully automated diagnostic workflows to cut MTTR significantly.
7. ITSM Integration Integrate automation into SymphonyAI-driven ticketing, approval flows, CMDB updates, and change management processes. Build operational runbooks that map directly to ITSM workflows.
Ideal Candidate
- Strong Cloud Automation Architect Profile
- Mandatory (Experience 1) – Must have 8+ years of total experience with the last 6–7 years focused continuously on hands-on infrastructure automation in enterprise environments
- Mandatory (Experience 2) – Must have expert-level Ansible engineering, including advanced Jinja2 templating, dynamic inventory, custom module development, and designing enterprise-grade playbooks, roles, and collections across AWS, Azure, VMware, and REST-integrated environments
- Mandatory (Experience 3) – Must have strong Kubernetes and Helm expertise, including custom Helm chart authoring, multi-environment lifecycle management (upgrades, rollbacks, canary releases), namespace management, RBAC, secrets, and network policies — with GitOps adoption via ArgoCD or Flux
- Mandatory (Experience 4) – Must have deep Terraform knowledge, including reusable module design covering AWS networking, IAM, EKS, ECS, RDS, and security controls — with CI/CD integration (GitHub Actions or Azure DevOps) including automated testing, linting, drift detection, and policy enforcement
- Mandatory (Experience 5) – Must have solid AWS architecture knowledge, including serverless and event-driven automation (Lambda, Step Functions, Event Bridge, SNS/SQS, S3 triggers, Systems Manager), IAM design, cross-account automation across AWS Organizations, and multi-account networking aligned with CIS, NIST, and AWS Well-Architected standards
- Mandatory (Experience 6) – Must have experience with security and compliance automation including CIS benchmark enforcement, vulnerability remediation, infrastructure hardening, certificate lifecycle management, and integration with CyberArk, AWS Secrets Manager, and SSM Parameter Store
- Mandatory (Skill) – Must have proficiency in Python, Bash, or PowerShell for automation scripting, with proven ability to build production-grade automation frameworks end-to-end
- Preferred (Skill 1) – Experience with policy-as-code (OPA, Gatekeeper, Conftest) and IaC testing tools (Molecule, Terratest, Checkov, tfsec)
- Preferred (Skill 2) – Experience with multi-account AWS automation, service mesh (Istio/Linkerd), and ITSM integration (SymphonyAI or equivalent — ticketing, approval flows, CMDB updates, change management)
- Preferred (Skill 3) – Experience with AWS cost optimization automation and converting operational runbooks into self-healing diagnostic workflows
About the Role
We are looking for a Senior DevOps Engineer to lead the design, automation, and scaling of our hybrid cloud infrastructure spanning public cloud and private/on-premises environments. You will partner closely with software engineering, security, and product teams to build reliable, secure, and high-performance systems that support rapid product delivery. This is a hands-on role with significant influence over our infrastructure strategy, deployment workflows, and engineering culture.
Key Responsibilities
- Architect, deploy, and maintain scalable, highly available infrastructure across both public cloud (AWS, Azure, GCP) and private cloud platforms (OpenStack, VMware vSphere/Tanzu, Nutanix, or similar).
- Operate and maintain on-premises infrastructure: hypervisors, compute, storage (Ceph, NetApp, SAN/NAS), networking (SDN, VLANs, BGP, MPLS), and hardware capacity planning, alongside their public cloud equivalents.
- Design and own CI/CD pipelines that deploy seamlessly across public and private environments.
- Implement and manage Infrastructure as Code (Terraform, Ansible, Pulumi) with strong version control and review practices, using providers for both public and private cloud platforms.
- Manage container orchestration (Kubernetes, ECS, OpenShift, Rancher) across managed cloud services and self-managed/bare-metal clusters, including upgrades, autoscaling, and workload reliability.
- Build observability into all systems through logging, metrics, tracing, and alerting (Prometheus, Grafana, Datadog, ELK, or similar) with unified visibility across hybrid environments.
- Champion security best practices: secrets management, IAM hardening, network segmentation, vulnerability scanning, and compliance (SOC 2, ISO 27001, HIPAA, or data-sovereignty requirements).
- Lead incident response, root-cause analysis, and post-mortems; drive long-term reliability improvements and SLO/SLA adherence.
- Optimize cost, capacity, and resource utilization across public cloud spend and on-premises hardware without compromising performance or availability.
- Partner with data center operations and network providers on hardware provisioning, firmware management, MPLS circuit management, and lifecycle planning.
- Mentor junior DevOps and software engineers; promote DevOps culture, automation-first thinking, and shared ownership of production.
- Evaluate and introduce new tools, platforms, and processes that improve developer productivity and system reliability.
Required Qualifications
- 5+ years of experience in DevOps, SRE, or Platform Engineering roles, with at least 2 years at a senior level.
- Deep expertise with at least one major public cloud provider (AWS, Azure, or GCP) in production.
- Hands-on experience operating private cloud or virtualization platforms (OpenStack, VMware, Nutanix, or equivalent) in production.
- Strong experience with virtualization, storage systems, and enterprise networking in on-premises environments.
- Strong hands-on experience with Kubernetes in production, including both managed cloud and self-managed/bare-metal clusters.
- Proficiency in Infrastructure as Code (Terraform and Ansible strongly preferred).
- Solid scripting and programming skills in Python, Go, Bash, or similar.
- Experience designing and operating CI/CD pipelines using tools such as GitHub Actions, GitLab CI, Jenkins, CircleCI, or ArgoCD.
- Strong Linux systems administration and networking fundamentals (TCP/IP, DNS, load balancing, VPNs, firewalls, routing, MPLS).
- Experience with monitoring and observability stacks (Prometheus, Grafana, Datadog, New Relic, ELK, or OpenTelemetry).
- Proven track record of leading incident response and improving system reliability.
- Excellent communication skills and the ability to collaborate across engineering, security, infrastructure, and product teams.
Preferred Qualifications
- Experience designing hybrid and multi-cloud architectures, including secure connectivity (Direct Connect, ExpressRoute, MPLS, VPN, SD-WAN) between public and private environments.
- Familiarity with service meshes (Istio, Linkerd), API gateways, and GitOps workflows (ArgoCD, Flux).
- Background in security-focused or regulated environments and exposure to compliance frameworks.
- Experience with database administration (PostgreSQL, MySQL, Redis, MongoDB) in cloud-managed and self-hosted setups.
- Contributions to open-source DevOps or cloud infrastructure tooling.
- Relevant certifications (AWS Solutions Architect / DevOps Engineer, Azure Administrator, CKA, CKAD, RHCE, VMware VCP, OpenStack Certified Administrator, HashiCorp Terraform Associate).
We are seeking a highly skilled and motivated Business Analyst to join our team delivering bespoke
software solutions to our clients. In this role, you will translate business needs into clearly written
user stories with testable acceptance criteria that act as the contract between client intent and what
the development team builds. You will be the bridge between Clients, Developers and the Project
Manager — and the person accountable for ensuring that what gets delivered is what the business
actually asked for.
Experience of documenting key projects artefacts and have a good understanding of SDLC. Basic
understanding of the Microsoft power platform tools.
Responsibilities:
• Being part of a team with a strong sense of product ownership and commitment to build
scalable, extensible and robust software and reports.
• Delivering outcomes that are clearly defined, using discretion over how to achieve them.
• Making suggestions for improvements to the work of the team, based on previous
experience and knowledge of similar situations.
• Involved in the development of complete solutions, from initiation to handover, participating
in a cross-functional team across requirements gathering, data analysis, process mapping,
user-story authoring, and user acceptance testing.
• Writing clear, testable acceptance criteria for every user story - using Given/When/Then or
equivalent structured formats - and owning them as the definition of 'done' that the
development team builds against and the client signs off on.
• Building and maintaining strong relationships with key stakeholders/clients
Requirements:
• Proven experience in a business analyst or product analyst role delivering bespoke software
solutions
• Demonstrable track record of writing high-quality acceptance criteria - specific, testable, and
traceable back to business intent - using Given/When/Then or equivalent. Able to show
examples of stories where their criteria caught gaps before development started.
• Experience of quickly building trusted stakeholder relationships at all levels of the business
• Experience of working with both waterfall, agile and hybrid methodologies
• Experience of technical writing and presentation skills using tools such as: PowerPoint, Word,
Visio
• Strong understanding and applied knowledge of User Experience and Customer Journey
concepts.
• Confident working with ambiguity
• Ability to work independently and be self-motivated.
• Experience in conceptual modelling; ability to see the big picture and envision possible
solutions.
• Outstanding verbal and non-verbal communication skills
• Ability to facilitate a team to consensus on scope and requirement decisions.
• Strong grasp of definition-of-done and requirements traceability - able to demonstrate how
each delivered feature maps back to a documented business need and its acceptance
criteria.
• Collaborate with cross-functional systems teams to document systems requirements and
understand their impact on other teams.
• Work with development teams to translate business operational and functional
requirements.
• Produce comprehensive documentation for systems and processes, using Business Process
Modelling techniques
• Experience in UML Modelling (Use Case Diagrams, Sequence Diagrams, Activity Diagrams,
Class Diagrams etc.)
Other:
• Bachelor’s degree: A degree in Business studies, Information Systems or a related field is
preferred but not essential
• Exposure to Backlog user requirement management, tools such as Jira, Azure devOps, would
be beneficial
• Business Acumen: Ability to understand and translate business requirements into data driven solutions. Strong analytical and problem-solving skills with a keen attention to detail.
• Communication Skills: Excellent verbal and written communication skills. Ability to effectively
communicate complex data concepts to both technical and non-technical stakeholders.
• Team Player: Capable of working collaboratively in a team environment. Willingness to share
knowledge and assist colleagues.
• Time Management: Strong organisational skills and the ability to manage multiple projects
simultaneously. Proven ability to meet deadlines and deliver high-quality work.
Join our dynamic team and help drive data-centric decision-making across the organisation.
Role Overview
We’re hiring a Senior Backend Engineer who can build high-performance backend systems using Node.js. You’ll work on complex engineering problems, optimize for scale, and own features end-to-end.
Key Responsibilities
● Develop backend services, APIs, and microservices using Node.js frameworks (Express, Nest.js).
● Own modules end-to-end: design → implementation → testing → deployment.
● Write high-quality, testable, maintainable code.
● Optimize performance, scalability, and reliability across systems.
● Troubleshoot production issues, perform RCA, and improve stability.
● Work with databases (SQL & NoSQL), caching, and async processing.
● Collaborate with DevOps for CI/CD, deployments, and automation.
● Contribute to improving engineering standards, documentation, and processes
What we’re looking for:
● 4 to 7 years of hands-on backend development experience.
● Strong in Node.js, async patterns, TypeScript (preferred), and REST APIs.
● Solid understanding of: Data modeling Queues & background jobs Caching (Redis) API optimization
● Experience with relational DBs (PostgreSQL/MySQL) and NoSQL (DynamoDB/MongoDB/ Redis).
● Experience with messaging/queues: Kafka, RabbitMQ, SQS, or equivalent.
● Familiarity with Docker, Git, and CI/CD pipelines.
● Exposure to AWS or cloud infrastructure.
● Strong debugging and performance tuning skills.
● Ability to work in a fast-paced, startup-like environment.
Here are answers to some questions you may have
Where is your office?
Chennai (Velachery)
Work Model
Work from Office – because great stories are built in person!
Do you have an online presence?
https://amura.ai (we are @AmuraHealth on all social media)
For better insights:
https://drive.google.com/file/d/11UHgXqYc7XSJ3IAdFyEm-EurgOzqsmzA/view?usp=sharing




























