
Senior Cloud & Devops Engineer
at provides mobile application development & support services.
We are looking for a self motivated and goal oriented candidate to lead in architecting, developing, deploying, and maintaining first class, highly scalable, highly available SaaS platforms.
This is a very hands-on role. You will have a significant impact on Wenable's success.
Technical Requirements:
8+ years SaaS and Cloud Architecture and Development with frameworks such as:
- AWS, GoogleCloud, Azure, and/or other
- Kafka, RabbitMQ, Redis, MongoDB, Cassandra, ElasticSearch
- Docker, Kubernetes, Helm, Terraform, Mesos, VMs, and/or similar orchestration, scaling, and deployment frameworks
- ProtoBufs, JSON modeling
- CI/CD utilities like Jenkins, CircleCi, etc..
- Log aggregation systems like Graylog or ELK
- Additional development tools typically used in orchestration and automation like Python, Shell, etc...
- Strong security best practices background
- Strong software development a plus
Leadership Requirements:
- Strong written and verbal skills. This role will entail significant coordination both internally and externally.
- Ability to lead projects of blended teams, on/offshore, of various sizes.
- Ability to report to executive and leadership teams.
- Must be data driven, and objective/goal oriented.

Similar jobs
About the company
The client is a US-based cybersecurity startup prevents, detects, and responds to software supply chain attacks by analyzing behavior across the full software development lifecycle for both developers and AI coding agents. They are building a vertical AI agent for supply chain security across three pillars: securing AI agents on developer machines, OSS package security, and CI/CD security, covering the entire agentic pipeline from dev environment to cloud.
Founded by ex-Microsoft, 21 years & ex-Uber, Microsoft, Plaid founders, they are a 16-person team working on hard problems at the intersection of security, AI, and open source.
Why this role is exciting
They are at the forefront of supply chain security research and product development. They were the first to detect several major supply chain attacks in 2025 and 2026, including the axios npm compromise and tj-actions. Their research is regularly cited by Bloomberg, TechCrunch, Hacker News, and Dark Reading. The US Cybersecurity and Infrastructure Security Agency (CISA) has published advisories citing the company.
Beyond their enterprise customers, the company has been adopted by more than 15,000 open-source projects, including projects from Microsoft, Google, Amazon, and Datadog.
This is a zero-to-one platform role. Today the AWS cost, service health visibility, and their release process are owned in pieces by engineers who are also shipping product. You will own all three end to end, as the first dedicated platform hire, with meaningful early-stage equity upside.
Their stack
GitHub.com for code repositories. GitHub Actions for CI/CD pipelines. AWS serverless for hosting our services: Lambda, API Gateway, DynamoDB, and S3. Backend services are written in Golang.
What you'll do
- Own AWS cost. Build real visibility into where our cloud spend goes, then drive it down. Put guardrails and anomaly alerting in place so cost regressions get caught in days, not at the end of the month.
- Build service health visibility. Define and instrument the metrics that tell us whether the platform is healthy. Build the dashboards, set the SLOs, and wire up alerting that pages on real problems and stays quiet otherwise.
- Own the release process. Design and implement how code gets from a pull request to production, including staged rollouts, rollback, release approvals, and change management that satisfies our enterprise customers and our ISO 27001 obligations without slowing engineers down.
- Improve stability. Run infrastructure as code, tighten our incident response and on-call practices, and drive postmortems to actual fixes.
- Secure our own pipelines. We sell CI/CD and cloud security, so our own build and deploy path should be the best reference implementation we have. You will run it that way.
What we're looking for
- 2 to 5 years of experience with strong engineering fundamentals.
- Demonstrated AWS cost reduction. We want to hear what the spend was, what you did, and what it became.
- Hands-on depth with AWS serverless in production, not just in theory.
- Experience implementing metrics, dashboards, and alerting for service health, and evidence that it changed how the team operated.
- Experience building or significantly improving CI/CD and release processes, ideally with GitHub Actions.
- Comfortable writing code. Golang or Python, plus infrastructure as code.
- An AI-native mindset, with prior hands-on experience building with or operating AI agents.
- Early-stage startup experience. You are extremely hands-on and can drive a problem end to end without a team behind you.
- A security background is a plus but not required.
At Quantiphi, you will be a key leader within our Platform Engineering team, responsible for defining and driving the architectural vision for cloud-based solutions. You will lead the design and implementation of robust pipeline infrastructure, orchestrating integration flows among upstream RPA processes, the AR tool backend, and downstream systems. You will work closely with stakeholders, cross-functional teams, and clients to ensure that solutions are scalable, secure, and aligned with business objectives.
Must have skills:
- Strong expertise in designing and architecting scalable, secure, and highly available AWS platforms using services such as API Gateway, Lambda, EC2, ECS/EKS, VPC, IAM, S3, CloudWatch, CloudTrail, ECR, SNS, SQS, EventBridge, Systems Manager (SSM), KMS, and Secrets Manager.
- Proven technical experience in designing and overseeing end-to-end platform integration architectures, orchestrating data and process flows between upstream systems, third-party applications (e.g., ERP systems such as BaaN, SAP, Oracle), internal services, and downstream consumer applications.
- Strong expertise in API architecture, including RESTful APIs, API Gateway, authentication/authorization (OAuth, JWT, API Keys), API lifecycle management, integration patterns, and secure third-party connectivity.
- Extensive experience designing AWS IAM security governance, including IAM Roles, Policies, Cross-Account Access, Permission Boundaries, Service Control Policies (SCPs), and enterprise identity and access management strategies.
- Strong experience designing AWS infrastructure using Infrastructure as Code (Terraform, AWS CloudFormation, or AWS CDK), with reusable platform modules, reference architectures, and standardized deployment frameworks.
- Expertise in designing and implementing enterprise CI/CD and DevSecOps pipelines using GitHub Actions, Jenkins, GitLab CI/CD, or AWS CodePipeline, integrating automated testing, security scanning, and deployment governance.
- Strong knowledge of enterprise networking, including VPC architecture, Transit Gateway, Direct Connect, VPN, Load Balancers, Route 53, DNS, PrivateLink, and hybrid connectivity.
- Experience implementing enterprise monitoring and observability using Amazon CloudWatch, CloudTrail, dashboards, centralized logging, and alerting frameworks to ensure platform reliability and operational excellence.
- Proficiency in automation using Python, Bash, PowerShell, and AWS SDK (Boto3) to streamline platform operations and deployments.
- Excellent technical solution design, stakeholder management, technical leadership, documentation, and mentoring skills, with the ability to drive architecture reviews and establish engineering best practices.
Good to have skills:
- Experience designing enterprise integration platforms using event-driven and microservices architectures.
- Knowledge of AWS Organizations, Control Tower, Landing Zone, governance frameworks, and FinOps best practices.
- Experience with container platforms such as Docker, Kubernetes, Amazon ECS/EKS, and service mesh technologies.
- Familiarity with enterprise observability platforms such as Grafana, Prometheus, Datadog, Splunk, or OpenSearch.
- Experience with enterprise messaging and integration technologies such as Kafka, Amazon MSK, MQ, or EventBridge.
- Knowledge of disaster recovery, business continuity, and high-availability architecture.
- AWS Certifications such as Solutions Architect – Professional, DevOps Engineer – Professional, Security Specialty, or Advanced Networking – Specialty.
- Experience working in Agile, DevOps, and Platform Engineering environments.
Required Skills
● Experience: Minimum of 5 years of professional experience in a DevOps Engineer
role
● Cloud Proficiency: Proven experience with at least one major cloud provider (AWS,
Azure, or GCP).
● Scripting & Programming: Strong scripting skills in languages such as Bash, Python,
or Go.
● IaC Tools: Hands-on experience with Terraform.
● Container Technology: Expertise in Docker and Kubernetes.
● CI/CD Tools: Proficient with CI/CD platforms like Jenkins, GitLab CI, or Travis CI.
● Configuration Management: Experience with configuration management tools like
Ansible, Chef, or Puppet.
● Version Control: Strong knowledge of Git and branching strategies.
● Problem-Solving:Excellent problem-solving abilities and a commitment to automation
and continuous improvement.

Location: Bangalore
Experience: 2–5 years
Type: Full-time | On-site
Start: Immediate
Why this role exists
Most systems don’t fail because of one big outage.
They fail because reliability is treated as an afterthought.
Right now, uptime depends too much on individual heroics.
That doesn’t scale.
This role exists to build a reliability system where:
- Uptime is predictable
- Failures are contained
- Escalations don’t depend on leadership
What you’ll do
You will not just monitor systems.
You will own reliability as a product.
1. Drive uptime to production-grade reliability
- Improve system uptime to 99.9% customer-facing SLA within 4 months
- Define and track:
- SLAs / SLOs / error budgets
- Ensure reliability is measured from the customer’s perspective, not internal metrics
2. Build incident response as a system
- Set up a 24/7 incident response rotation across 3 engineers
- Eliminate dependency on leadership (no single escalation point)
- Define:
- Incident severity levels
- Response playbooks
- Escalation protocols
- Ensure fast detection → containment → resolution
3. Contain and fix erratic system behavior
- Identify and resolve:
- Latency spikes
- Downtime incidents
- Integration failures
- Build guardrails to prevent recurrence
- Focus on root cause elimination, not temporary fixes
4. Create continuous reliability feedback loops
- Work closely with engineering teams to:
- Surface recurring failure patterns
- Improve build quality
- Reduce production bugs
- Ensure learnings from incidents directly improve future releases
5. Improve observability and monitoring
- Build dashboards and alerts for:
- System health
- Performance metrics
- Failure signals
- Ensure issues are detected before customers report them
6. Reduce operational fragility
- Remove single points of failure (people, systems, workflows)
- Improve system resilience across:
- Deployments
- Integrations
- Runtime environments
What success looks like
- Uptime reaches 99.9%+ reliably
- Incidents are:
- Detected early
- Contained quickly
- Resolved permanently
- No dependency on a single individual for escalation
- System behavior becomes predictable and stable
- Engineering teams ship with higher reliability confidence
Who you are
- You have 2-5 years of experience in SRE / DevOps / backend systems
- You have worked on production systems with real uptime expectations
- You think in:
- Systems
- Failure modes
- Trade-offs
- You are comfortable debugging live, high-pressure environments
What will make you stand out
- Experience with:
- Distributed systems
- Cloud infrastructure (AWS / Azure / GCP)
- Monitoring & alerting tools
- Have built or improved:
- Incident response systems
- Reliability frameworks
- Strong debugging skills across:
- Infra
- Application
- Integrations
Compensation
₹60,000/month (fixed)
(Aligned with role scope and impact expectations)
Why join
- You will define reliability standards for a production AI platform
- Your work directly impacts:
- Customer trust
- Product performance
- Enterprise readiness
- You will move the system from reactive → predictable
What this role is not
- Not just monitoring dashboards
- Not limited to handling tickets
- Not dependent on escalation to leadership
What this role is
- A builder of reliability systems
- A guardian of uptime and performance
- A multiplier of engineering quality
One question to self-evaluate
Can you build a system where downtime is rare, predictable, and never dependent on a single person?
Greetings!
Wissen Technology is hiring for Kubernetes Lead/Admin.
Required:
- 7+ years of relevant experience in Kubernetes
- Must have hands on experience on Implementation, CI/CD pipeline, EKS architecture, ArgoCD & Statefulset services.
- Good to have exposure on scripting languages
- Should be open to work from Chennai
- Work mode will be Hybrid
Company profile:
Company Name : Wissen Technology
Group of companies in India : Wissen Technology & Wissen Infotech
Work Location - Bangalore
Website : www.wissen.com
Wissen Thought leadership : https://www.wissen.com/articles/
LinkedIn: https://www.linkedin.com/company/wissen-technology

- Proficiency in Python , Django and Other Allied Frameworks;
- Expert in designing UI/UX interfaces;
- Expert in testing, troubleshooting, debugging and problem solving;
- Basic knowledge of SEO;
- Good communication;
- Team building and good acumen;
- Ability to perform;
- Continuous learning
Required qualifications and must have skills
-
5+ years of experience managing a team of 5+ infrastructure software engineers
-
5+ years of experience in building and scaling technical infrastructure
-
5+ years of experience in delivering software
-
Experience leading by influence in multi-team, cross-functional projects
-
Demonstrated experience recruiting and managing technical teams, including performance management and managing engineers
-
Experience with cloud service providers such as AWS, GCP, or Azure
-
Experience with containerization technologies such as Kubernetes and Docker
Nice to have Skills
-
Experience with Hadoop, Hive and Presto
-
Application/infrastructure benchmarking and optimization
-
Familiarity with modern CI/CD practices
-
Familiarity with reliability best practices

Java with cloud
|
Core Java, SpringBoot, MicroServices |
|
- DB2 or any RDBMS database application development |
|
- Linux OS, shell scripting, Batch Processing |
|
- Troubleshooting Large Scale application |
|
- Experience in automation and unit test framework is a must |
|
- AWS Cloud experience desirable |
|
- Agile Development Experience |
|
- Complete Development Cycle ( Dev, QA, UAT, Staging) |
|
- Good Oral and Written Communication Skills |
Should be open to embracing new technologies, keeping up with emerging tech.
Strong troubleshooting and problem-solving skills.
Willing to be part of a high-performance team, build mature products.
Should be able to take ownership and work under minimal supervision.
Strong Linux System Administration background (with minimum 2 years experience), responsible for handling/defining the organization infrastructure(Hybrid).
Working knowledge of MySQL databases, Nginx, and Haproxy Load Balancer.
Experience in CI/CD pipelines, Configuration Management (Ansible/Saltstack) & Cloud Technologies (AWS/Azure/GCP)
Hands-on experience in GitHub, Jenkins, Prometheus, Grafana, Nagios, and Open Sources tools.
Strong Shell & Python scripting would be a plus.
Required Skills and Experience
- 4+ years of relevant experience with DevOps tools Jenkins, Ansible, Chef etc
- 4+ years of experience in continuous integration/deployment and software tools development experience with Python and shell scripts etc
- Building and running Docker images and deployment on Amazon ECS
- Working with AWS services (EC2, S3, ELB, VPC, RDS, Cloudwatch, ECS, ECR, EKS)
- Knowledge and experience working with container technologies such as Docker and Amazon ECS, EKS, Kubernetes
- Experience with source code and configuration management tools such as Git, Bitbucket, and Maven
- Ability to work with and support Linux environments (Ubuntu, Amazon Linux, CentOS)
- Knowledge and experience in cloud orchestration tools such as AWS Cloudformation/Terraform etc
- Experience with implementing "infrastructure as code", “pipeline as code” and "security as code" to enable continuous integration and delivery
- Understanding of IAM, RBAC, NACLs, and KMS
- Good communication skills
Good to have:
- Strong understanding of security concepts, methodologies and apply them such as SSH, public key encryption, access credentials, certificates etc.
- Knowledge of database administration such as MongoDB.
- Knowledge of maintaining and using tools such as Jira, Bitbucket, Confluence.
- Work with Leads and Architects in designing and implementation of technical infrastructure, platform, and tools to support modern best practices and facilitate the efficiency of our development teams through automation, CI/CD pipelines, and ease of access and performance.
- Establish and promote DevOps thinking, guidelines, best practices, and standards.
- Contribute to architectural discussions, Agile software development process improvement, and DevOps best practices.










