Engineering Lead - workflow Orchestration at Tech AI startup in Bangalore · Remote only · 7 - 10 years · ₹15L - ₹25L / yr · Remote only · Posted 21 Nov 2025

Engineering Lead - workflow Orchestration
at Tech AI startup in Bangalore
We are seeking a seasoned Engineering Lead with deep expertise in workflow orchestration systems, stateful execution engines, and distributed task runtimes. You will architect the next generation of our DAG-based runtime, develop the infrastructure for agentic workflow composition, and drive execution excellence across engineering. This role combines hands-on systems design with technical leadership and team mentorship.
Key Responsibilities
1. Architect & Build the Swarm Runtime
- Design and implement a DAG-based orchestration engine using Temporal, Argo Workflows, or equivalent event-driven runtimes.
- Build a scalable primitive registry for tasks, operators, guards, and computational nodes.
- Architect a robust scheduler capable of handling event triggers, retries, backoffs, and distributed coordination.
2. Develop the Workflow Composition Layer
- Define and build a YAML/JSON-based DSL for describing agentic workflows, dependencies, and execution semantics.
- Create a schema-driven rules engine ensuring validations, model calls, parallelism, conditional branching, and approval gates are seamlessly integrated.
3. Orchestration Logic & Runtime Intelligence
- Implement orchestration logic that coordinates:
- Validation layers
- Model calls (LLMs, embedding engines, external APIs)
- Human-in-the-loop approval gates
- Stateful transitions and checkpointing
- Ensure deterministic execution, traceability, and safe rollback mechanisms.
4. Workflow Certification & Automated Testing
- Define certification standards for every workflow type, including:
- Functional correctness
- Latency and concurrency thresholds
- Error-handling expectations
- Observability and trace coverage
- Build automated regression test suites validating workflow integrity before deployment.
5. Engineering Leadership & Delivery
- Lead, mentor, and grow a team of backend and systems engineers.
- Drive sprint planning, reviews, engineering discipline, and roadmap execution.
- Own runtime delivery deadlines, cross-team coordination, and release quality.
6. Observability, Reliability & Production Excellence
- Instrument the runtime with observability hooks: metrics, tracing, structured logs, and execution heatmaps.
- Build robust retry logic, distributed locks, idempotency guards, and failover strategies.
- Improve runtime stability, throughput, and scale characteristics.
Requirements
Must-Have
- 8+ years of backend engineering experience building high-scale systems.
- 3+ years leading teams focused on workflows, automation, orchestration, or distributed runtimes.
- Deep understanding of:
- Stateful orchestration engines (Temporal, Step Functions, Argo, Airflow)
- Message queues, pub/sub systems, and event-driven patterns
- Retry logic, compensating transactions, and idempotent operations
- Distributed tracing, observability pipelines, and health checks
- Strong background in concurrent programming, async task management, and execution models.
- Hands-on experience with at least one systems language or backend stack (Python, Go, Rust, Node).
Nice-to-Have
- Experience building workflow DSLs or schema-driven interpreters.
- Familiarity with LLM pipelines, agentic runtimes, or AI-driven workflow automation.
- Knowledge of Kubernetes-based runtime environments and workflow controllers.
- Experience with pluggable architecture design, sandboxing, or execution policies.
What This Role Offers
- Ownership of the core execution engine powering Perceive Now’s intelligent automation platform.
- A high-impact leadership position shaping architectural strategy and engineering culture.
- The opportunity to solve cutting-edge problems at the intersection of orchestration, distributed systems, and AI.

Similar jobs (10)
Supercharge Your Career as a Principal Lead Developer at Technoidentity!
At Technoidentity, we're a Data & AI product engineering company with over 15 years of expertise in building durable digital products, intelligent enterprise solutions, and scalable Data & AI platforms. As we continue expanding globally, it's the perfect time to join our team of tech innovators and make a lasting impact.
What’s in it for You?
Technoidentity is building a high-impact Temporal Capability Center to design, deliver, and scale resilient workflow-orchestration solutions for enterprise customers. We are seeking a hands-on, technically strong Principal Lead Developer who combines deep software-engineering expertise with solution-architecture leadership, distributed-systems knowledge, and an ownership mindset.
This role is suited to a resourceful engineer who can independently take multi-faceted business and technical requirements from discovery through design, implementation, rollout, and operational improvement. You will work directly with overseas clients, shape architecture decisions, lead multiple technical workstreams, and mentor junior developers while remaining actively involved in code, design reviews, and delivery.
The ideal candidate has strong practical experience in Python, Java, and/or TypeScript, has experience with Cloud Platforms, understands modern AI-assisted engineering practices. Experience with distributed workflow solutions using orchestration platforms such as Temporal, Cadence, Camunda, or similar technologies will be preferred, but is not a requirement.
Requirements
What Will You Be Doing?
- Lead the end-to-end technical delivery of workflow-orchestration and distributed-systems solutions for enterprise clients.
- Design robust, scalable, secure, and observable architectures using durable execution, asynchronous messaging, APIs, event-driven patterns, and cloud-native components.
- Build production-grade services and workflow implementations using Python, Java, and/or TypeScript.
- Design and implement long-running business processes using orchestration frameworks such as Temporal, Cadence, Camunda, or equivalent platforms.
- Apply distributed-systems principles to address reliability, idempotency, retries, timeouts, compensating transactions, eventual consistency, failure recovery, and data consistency.
- Design and implement Saga patterns, including orchestration-based and choreography-based approaches, for multi-service business transactions.
- Lead architecture discovery sessions, technical workshops, solution presentations, and design reviews with overseas clients and internal stakeholders.
- Translate business requirements into clear solution architectures, implementation plans, technical specifications, estimates, and delivery milestones.
- Drive multiple concurrent technical streams, identify dependencies and risks early, and maintain delivery quality under changing priorities.
- Establish engineering standards for workflow design, API contracts, error handling, versioning, testing, deployment, observability, and operational readiness.
- Use AI-assisted SDLC practices responsibly to improve developer productivity, code quality, test coverage, documentation, and delivery velocity.
- Design and contribute to agentic orchestration solutions, including coordination of AI agents, tool/API integrations, workflow state management, guardrails, human-in-the-loop controls, and evaluation approaches where relevant.
- Perform hands-on development of critical components, prototypes, integrations, and reference implementations.
- Conduct code reviews and architectural reviews, ensuring solutions are maintainable, performant, secure, and aligned with engineering best practices.
- Mentor and guide junior developers; support technical growth through pairing, reviews, reusable patterns, technical sessions, and constructive feedback.
- Collaborate with delivery, product, platform, QA, DevOps, security, and client teams to ensure successful implementation and production adoption.
- Contribute reusable assets, accelerators, templates, documentation, and best practices to the Temporal Capability Center.
Benefits
What Makes You the Perfect Fit?
- Bachelor’s or Master’s degree in Computer Science, Engineering, or a related technical discipline, with 6+ years of relevant work experience.
- Significant professional software-development experience, including experience leading technical delivery and mentoring engineering teams.
- Strong hands-on proficiency in at least one of the following languages:
o Python
o Java
o TypeScript
- Demonstrated experience designing and building distributed, scalable, highly available, or event-driven systems.
- Strong understanding of microservices architecture, REST and/or asynchronous APIs, message-driven systems, and integration patterns.
- Practical expertise with distributed-systems concerns, including:
o Idempotency and duplicate-message handling
o Retries, backoff, timeouts, and circuit-breaking strategies
o Failure handling and recovery
o Eventual consistency and data synchronization
o State management for long-running processes
o Observability, logging, metrics, tracing, and operational debugging
- Strong understanding and practical application of the Saga pattern and compensating-transaction design.
- Experience designing solution architectures and communicating technical trade-offs to both engineering and business stakeholders.
- Ability to lead junior developers through technical direction, task decomposition, code review, and mentorship.
- Strong written and verbal communication skills in English.
- High ownership, self-motivation, adaptability, and the ability to make sound decisions in ambiguous and fast-moving environments.
Preferred Qualifications
- Production experience with Temporal, Cadence, Camunda, AWS Step Functions, Azure Durable Functions, Netflix Conductor, Apache Airflow, or comparable orchestration technologies..
- Experience migrating from legacy schedulers, BPM platforms, or synchronous microservice flows to durable workflow orchestration.
- Experience building agentic or AI-enabled applications using LLMs, tool calling, agent frameworks, retrieval-augmented generation, MCP, guardrails, or evaluation frameworks.
- Practical use of AI coding assistants and AI-aided SDLC workflows for planning, implementation, testing, code review, documentation, and troubleshooting.
- Cloud-native development experience on AWS, Azure, Google Cloud Platform, or a hybrid-cloud environment.
- Experience with Kubernetes, Docker, CI/CD pipelines, Infrastructure as Code, and container-based deployments.
- Experience with event-streaming and messaging platforms such as Kafka, RabbitMQ, cloud queues, or pub/sub systems.
- Experience with relational and NoSQL databases, data-modeling strategies, and transactional/outbox patterns.
- Experience with API gateways, security controls, OAuth/OIDC, secrets management, and secure service-to-service communication.
- Familiarity with domain-driven design, event sourcing, CQRS, and enterprise integration patterns.
- Experience in technical consulting, pre-sales support, solution accelerators, or client-facing technical leadership.
Leadership Expectations
As a Principal Lead Developer, you will be expected to:
- Act as a trusted technical advisor to clients and internal delivery teams.
- Balance architecture leadership with hands-on engineering execution.
- Proactively identify problems, propose practical solutions, and drive them to completion without waiting for detailed direction.
- Manage and prioritize several workstreams while communicating progress, risks, decisions, and dependencies clearly.
- Raise engineering quality through reusable standards, design patterns, code reviews, and mentoring.
- Build confidence with clients through technical depth, clarity, responsiveness, and reliable delivery.
- Foster a collaborative, accountable, learning-oriented engineering culture within the Temporal Capability Center.
Success Measures
Success in this role will be measured by:
- Delivery of reliable, scalable, and maintainable orchestration solutions that solve real client problems.
- Quality of solution architecture, code, technical documentation, and operational readiness.
- Effective application of workflow-orchestration and Saga patterns to complex distributed business processes.
- Ability to independently lead client-facing technical discussions and turn them into executable delivery plans.
- Delivery predictability across multiple concurrent technical initiatives.
- Measurable improvement in engineering productivity and quality through AI-assisted SDLC practices.
- Growth, engagement, and technical effectiveness of junior developers under your guidance.
- Contribution of reusable Temporal Capability Center assets, reference architectures, and best practices.
Why Join Technoidentity’s Temporal Capability Center?
- Work on complex, high-value enterprise workflow and distributed-systems challenges.
- Help shape a specialized capability center focused on durable execution, orchestration, AI-enabled delivery, and modern architecture.
- Collaborate with overseas clients and multidisciplinary global teams.
- Influence technical standards, reusable accelerators, and the future direction of orchestration solutions at Technoidentity.
- Take on a role with meaningful architecture ownership, technical leadership, and hands-on engineering impact.

Required Experience: 10–15 years (with at least 3–5 years in leadership roles)
● 10–15 years of overall experience in backend engineering, with strong exposure to
Python and/or Golang.
● 3–5 years of experience managing engineering teams.
● Proven experience delivering large-scale, distributed systems in production
environments.
● Strong understanding of microservices, cloud-native architecture, and DevOps
practices.
● Hands-on background in backend engineering (able to guide teams technically, even
if not coding daily).
● Familiarity with CI/CD pipelines, observability, and performance optimization.
● Experience in financial services or high-transaction domains is a plus.
● Experience leading teams that have utilized AI-driven development practices (e.g.,
agentic coding, LLM integration) to improve productivity and innovation is a
significant advantage.
Skills
● Excellent leadership and people management abilities.
● Strong communication and stakeholder management skills.
● Ability to balance technical depth with business priorities.
● Problem-solving mindset with a focus on delivery and impact.
● Passion for building engineering culture and improving developer experience.
We are seeking a Senior Full Stack Engineer to join our team in a long-term contractor capacity to continue development of a production-grade platform hosted on AWS.
This application supports policy processing, third-party integrations, compliance workflows, reporting, and intelligent automation capabilities. The ideal candidate is a strong software engineer first, capable of contributing across the full stack while helping scale and evolve the platform.
This role requires someone who can step into an existing system, understand complex workflows quickly, and independently deliver high-quality solutions.
Responsibilities
• Design, develop, and maintain full-stack application features across frontend and backend systems
• Build and support integrations with third-party systems and APIs
• Develop workflow-driven processes using Temporal
• Build scalable APIs and backend services using Python
• Maintain and optimize relational databases using PostgreSQL
• Develop reporting and analytics capabilities using charting libraries
• Contribute to AI-enabled features and integrations within the platform
• Improve CI/CD pipelines and deployment processes
• Participate in architecture discussions and help shape long-term technical direction
• Work closely with business and technical stakeholders to deliver production-ready solutions
Required Qualifications
Engineering
• at least 7+ years of full stack software engineering experience
• Strong proficiency in Python
• Strong frontend development experience with modern web frameworks
• Strong backend API development experience
• Experience designing and building scalable applications
• Strong understanding of software architecture and best practices
• Experience working in complex, integrated systems
Workflow and Orchestration
• Proven experience with Temporal
• Experience building and managing workflow orchestration patterns
• Familiarity with asynchronous processing and event-driven systems
Database and Reporting
• Strong experience with PostgreSQL
• Strong SQL and data modeling experience
• Experience building reporting dashboards and analytics features
• Experience with charting libraries such as Chart.js, D3.js, or Plotly
DevOps
• Experience with CI/CD pipelines
• Familiarity with containerized deployments
• Experience with cloud environments and modern development workflows
Preferred Qualifications
• Experience in insurance, surplus lines, or compliance-based applications
• Experience integrating with third-party vendors and external APIs
• Experience with AI tooling, LLM integrations, and context engineering
• Experience building intelligent automation features
What We Are Looking For
• Self-driven and highly autonomous
• Strong problem-solving ability
• Comfortable with ownership and accountability
• Able to contribute with minimal supervision
• Strong communication skills in English
• Comfortable working U.S.-based business hours
Ideal Candidate
A senior full stack engineer who can quickly contribute to an active production system, own features end-to-end, and help expand a platform that sits at the center of complex business workflows and integrations. Send resume with projects and contact information.
Job Title: Platform Engineer
Location: Bangalore(Onsite)
Experience Level: 3-8
Salary Range: 20-30LPA
Description:
Join a team building an AI-native enterprise platform that helps businesses make faster, smarter and more consistent operational decisions using AI, enterprise data and workflow automation.
Design and build the core platform for enterprise decision workflows. Develop reusable workflow and decision runtimes. Build scalable, cloud-native distributed systems and event-driven architectures. Design enterprise-grade APIs and platform services. Build integrations with systems such as SAP and Oracle. Develop secure multi-tenant services with authentication and RBAC. Build and manage AWS cloud infrastructure and deployment systems. Implement monitoring and observability for production systems. Support both cloud and on-premise deployments. Enable faster onboarding and deployment of new enterprise workflows.
Requirements:
- Strong hands-on experience with Python
- FastAPI
- PostgreSQL
- Docker
- AWS
- Practical experience with Redis
- Kafka/event streaming
- REST APIs
- CI/CD
- Git
- Good understanding of Kubernetes
- Distributed systems
- Event-driven architecture
- Enterprise SaaS
- Microservices
- Strong backend engineering fundamentals
- Ability to design scalable, reliable and production-ready systems
The recruiter has not been active on this job recently. You may apply but please expect a delayed response.
Strong Tech Lead / Staff Engineer / Lead Engineer Profiles
2
Mandatory (Experience 1) - Must have minimum 6+ years of overall Software Engineering/Development experience building production backend systems.
3
Mandatory (Experience 2) - Must have Tech Lead ownership, with experience setting technical direction, architecture, frameworks, or engineering standards used by other engineers.
4
Mandatory (Experience 3) - Must have strong hands-on Backend development experience with deep proficiency in at least one backend programming language. (Python, Java, Golang etc)
5
Mandatory (Experience 4) - Must have strong System design/distributed systems experience, including architectural trade-offs, scalability, failure modes, reliability, and long-term maintainability.
6
Mandatory (Experience 5) - Must have strong CS fundamentals in Data Structures & Algorithms, Operating Systems, and Networking, with the ability to apply them to complex backend/system problems
7
Mandatory (Experience 6) - Must have owned complex, ambiguous, or business-critical engineering problems end-to-end, demonstrating clear technical judgment, trade-offs, problem-solving, and measurable technical/business impact; not limited to feature implementation.
8
Mandatory (Experience 7) - Must have formal technical leadership experience, including mentoring Senior Engineers, conducting design/code reviews, driving engineering quality, and leading technical decisions across a team.
9
Mandatory (Experience 8) - Must be a hands-on IC + Tech Lead, with approximately 70–80% hands-on coding/engineering involvement, not a pure people-management profile.
10
Mandatory (Company) - Product Companies / Startups / B2B SaaS
11
Mandatory (Education) - CS degree or equivalent (B.Tech/B.E./B.S.)
12
Preferred (Skills) - Experience with Kafka/NATS/RabbitMQ, AWS, Docker/Kubernetes, CI/CD, observability, security/compliance, multi-tenant systems, and financial/billing platforms.
About the Role
We are hiring Staff / Principal Engineers to take full, hands-on ownership of Blitzy's most critical production-grade systems and to deliver high-leverage features that materially improve customer outcomes and engineering velocity. This is the most senior individual contributor role at the company today.
This is not a Senior-plus role, an architecture-only role, or a promotion-track role. We are looking for someone who has already operated at Principal / Staff+ scope in a highly technical environment and expects to spend their time writing, reviewing, and shipping production code.
This role is 100% hands-on. Leverage comes from system ownership, execution quality, and durable technical decisions — not people management or process.
Responsibilities
- Own mission-critical production systems end-to-end, ensuring correctness, scalability, performance, reliability, and operational excellence.
- Design, build, and ship high-impact backend systems and features that improve product reliability, performance, and customer value.
- Architect scalable services and cloud infrastructure using technologies such as Python, REST, gRPC, Kubernetes, and Terraform.
- Identify and resolve complex technical bottlenecks that limit engineering quality, system performance, or organizational velocity.
- Build and operate LLM-powered systems and validation loops that evaluate correctness, consistency, durability, and production performance.
- Design and evolve data architectures incorporating relational, NoSQL, graph, and vector databases to support complex enterprise applications and semantic retrieval.
- Modernize and improve complex enterprise systems while balancing reliability, maintainability, scalability, and delivery speed.
- Set and uphold engineering quality standards through hands-on technical leadership, sound technical judgment, and ownership of long-term technical decisions.
Qualifications
- Direct experience with Python as a primary programming language, backend frameworks, and microservices architectures.
- Expertise in REST and gRPC, with proficiency in Node.js and JavaScript.
- Proficiency in GCP, along with experience using at least one additional cloud platform such as AWS or Azure.
- Advanced knowledge of Kubernetes and Terraform in production environments.
- Experience operating highly available production systems, including monitoring, scalability, reliability, performance optimization, and operational tooling.
- Strong knowledge of SQL and NoSQL databases, including PostgreSQL, MySQL, MongoDB, Cassandra, or DynamoDB.
- Familiarity with graph databases such as Neo4j and vector databases or embedding infrastructure for semantic search and retrieval.
- Hands-on experience building and operating LLM-powered systems in production, including evaluation, validation, regression testing, tracing, and failure analysis.
- Working knowledge of LangSmith or comparable LLM observability and evaluation tools; familiarity with OpenAI, Anthropic, or similar model providers is a plus.
- Ability to contribute across the full stack, with a strong understanding of frontend architecture and the ability to debug, design, and ship across frontend, backend, infrastructure, and AI systems.
- Understanding of large-scale enterprise software systems, including architecture, integration, deployment, modernization, and long-term maintainability.
- Proven track record of operating at Staff+, Principal Engineer, or equivalent level, independently driving complex technical initiatives and delivering high-impact outcomes with minimal supervision.
Blitzy is a Cambridge, MA based AI software development platform on a mission to revolutionize the software development life cycle by autonomously building custom software to unlock the next industrial revolution. We're transforming how enterprises build software, turning enterprise requirements into enterprise grade code with an agentic software development platform that can autonomously execute 80% of the quantum of software development work. We're backed by multiple tier 1 investors, and have proven success as founders of previous start-ups.
Our Culture
Who we are:
Led by two pioneering co-founders we are one of the fastest growing companies in the U.S., creating our own category of enterprise autonomous software development. We automate thousands of hours of software development for our customers, which includes strong representation within the Fortune 500.
How we work:
- We move Blitzy Fast: Time is both our company’s and our clients’ most precious asset. We move quickly and decisively to innovate internally and deliver exceptional software externally.
- Championship Mindset: We operate like a professional sports team. We win as a team by holding ourselves and each other to high standards, collaborating in-person, and remaining focused on the mission.
- Passion for Invention: We’re pushing the frontier of what’s possible, requiring constant innovation and iteration.
- We Work for the Customer: We focus on delivering outsized value to the customers we work with and expanding those relationships into deep, meaningful partnerships.
- We believe in being ‘everyday athletes’: taking care of ourselves so we can bring our best minds to work. We promote great sleep, movement, and restorative activities for
Blitzy is an equal opportunity employer committed to building a diverse and inclusive team. We believe different perspectives make us stronger.
About the Role
We are seeking a hands-on Tech Lead to design, build, and integrate AI-driven systems that automate and enhance real-world business workflows. This is a high-impact role for someone who enjoys full-stack ownership — from backend AI architecture to frontend user experiences — and can align engineering decisions with measurable product outcomes.
You will begin as a strong individual contributor, independently architecting and deploying AI-powered solutions. As the product portfolio scales, you will lead a distributed team across India and Australia, acting as a System Integrator to align engineering, data, and AI contributions into cohesive production systems.
Example Project
Design and deploy a multi-agent AI system to automate critical stages of a company’s sales cycle, including:
- Generating client proposals using historical SharePoint data and CRM insights
- Summarizing meeting transcripts
- Drafting follow-up communications
- Feeding structured insights into dashboards and workflow tools
The solution will combine RAG pipelines, LLM reasoning, and React-based interfaces to deliver measurable productivity gains.
Key Responsibilities
- Architect and implement AI workflows using LLMs, vector databases, and automation frameworks
- Act as a System Integrator, coordinating deliverables across distributed engineering and AI teams
- Develop frontend interfaces using React/JavaScript to enable seamless human-AI collaboration
- Design APIs and microservices integrating AI systems with enterprise platforms (SharePoint, Teams, Databricks, Azure)
- Drive architecture decisions balancing scalability, performance, and security
- Collaborate with product managers, clients, and data teams to translate business use cases into production-ready systems
- Mentor junior engineers and evolve into a broader leadership role as the team grows
Ideal Candidate Profile
Experience Requirements
- 5+ years in full-stack development (Python backend + React/JavaScript frontend)
- Strong experience in API and microservice integration
- 2+ years leading technical teams and coordinating distributed engineering efforts
- 1+ year of hands-on AI project experience (LLMs, Transformers, LangChain, OpenAI/Azure AI frameworks)
- Prior experience in B2B SaaS environments, particularly in AI, automation, or enterprise productivity solutions
Technical Expertise
- Designing and implementing AI workflows including RAG pipelines, vector databases, and prompt orchestration
- Ensuring backend and AI systems are scalable, reliable, observable, and secure
- Familiarity with enterprise integrations (SharePoint, Teams, Databricks, Azure)
- Experience building production-grade AI systems within enterprise SaaS ecosystems
Role & Responsibilities
Responsibilities
• Business: Immerse in operations until you think like an insider.
Rapidly acquire domain expertise through direct observation, translate between business and engineering seamlessly, and mentor engineers in your area on immersion. Influence senior stakeholders effectively, manage complex stakeholder landscapes with competing agendas, and build trust rapidly with new stakeholders.
• Delivery: Lead rapid delivery initiatives across teams in your area, coach on prototype-first approaches, and establish trust through consistent fast delivery. Build complete applications rapidly across any technology stack, select the right tools for each problem, and define clear criteria for prototype-to-production transitions.
• Generative AI: Architect RAG systems for complex use cases across teams, implement advanced techniques (hybrid search, reranking, query expansion), mentor engineers on RAG best practices, and establish RAG standards. Lead evaluation strategy across teams, establishing annotation guidelines, training human-calibrated LLM judges, and building evaluation pipelines that connect tracing to datasets to experiments.
• People: Build high-performing teams across your area, navigate complex interpersonal dynamics, foster psychological safety, and create environments where diverse perspectives are valued. Influence through communication at all levels — from frontline to executive. Handle difficult conversations skilfully and train engineers in your area on effective communication.
• AI-Augmented Development: Optimise AI tool usage across teams in your area, train engineers on AI-augmented and agentic engineering workflows, evaluate new AI development tools, and establish practices that balance AI speed with verification rigour.
• Scale: Design complex multi-component systems end-to-end, evaluate architectural options for large initiatives across teams, guide technical decisions for your area, and mentor engineers on architecture. Create debt reduction strategies across teams, influence roadmap decisions to include debt work, and teach engineers when to accept debt for speed versus when to invest in quality.
• Documentation: Define documentation standards across teams in your area, create documentation systems and templates, train engineers on spec-driven development, and ensure documentation quality across projects. Lead pattern generalization initiatives, defining criteria for when to generalize versus keep custom.
• Reliability: Define reliability standards across teams in your area, drive post-incident improvements systematically, design capacity planning processes, andmentor engineers on SRE practices.
Ideal Candidate
- Strong Staff Software Engineer / FDE profile (full-stack + production GenAI, multi-team technical leadership)
- Mandatory (Experience 1) – Must have 7+ years of relevant professional software engineering experience, with demonstrated full-stack delivery across backend and frontend.
- Mandatory (Experience 2) – Must have deep production experience with Python AND JavaScript/TypeScript, working comfortably across the full stack.
- Mandatory (Experience 3) – Must have 2+ years of experience in generative AI applications developement — LLM integrations, vector databases, RAG systems, and evaluation pipelines
- Mandatory (Experience 4) – Must have strong experience with modern frontend frameworks (Next.js / React) and backend API development.
- Mandatory (Experience 5) – Must have extensive experience with cloud platforms (AWS preferred; Azure/GCP valued), including infrastructure-as-code (CloudFormation / Terraform).
- Mandatory (Experience 6) – Must have working knowledge of multiple database paradigms — relational (PostgreSQL), document, and key-value (Redis) — with ability to select the right storage per problem.
- Mandatory (Experience 7) – Must have strong experience with CI/CD pipelines (e.g. GitHub Actions), containerization, and production deployment strategies.
- Mandatory (Experience 8) – Must have demonstrable fluency with AI coding tools (Claude Code, Cursor, GitHub Copilot, or similar) and proven ability to design agentic engineering workflows and train teams on them
- Preferred (Experience) – Advanced RAG techniques — hybrid search, reranking, query expansion — and establishing RAG standards across teams

What You Will Own
Platform Architecture
- Full architectural ownership of the non-AWS toolchain: CI/CD, observability, event streaming, automation, secrets, and deployment infrastructure
- Define, build, and enforce platform standards across portfolio products
- Terraform IaC for all infrastructure — nothing provisioned manually, everything versioned and reviewed
- Self-service developer platform so product teams ship without waiting on platform
Event Streaming & Pipeline Infrastructure
- Own the event streaming architecture, operational standards, and health monitoring across all products using real-time pipelines
- Design and maintain batch processing infrastructure alongside live event flows
- Ensure pipeline reliability, throughput, and cost are actively managed at scale
CI/CD & Deployment
- Build and maintain CI/CD pipelines (GitHub Actions) across all portfolio products
- Automate triage and retry logic for known failure classes — flaky tests, dependency timeouts, OOM kills — so engineers are only paged for genuinely novel failures
- Deployment standards: release management, rollback mechanisms, canary and blue-green patterns where justified
Observability & Reliability
- Own the full observability stack: Grafana, Prometheus, and Loki across all products
- SLOs and error budgets defined per product; reliability tracked consistently
- Build alerting that correlates signals and surfaces diagnostic context alongside notifications — so on-call engineers arrive at an incident with hypotheses, not a blank screen
- Incident response: on-call design, escalation playbooks, post-mortem facilitation
- Automated remediation scoped to a defined set of safe, idempotent actions — container restarts, ECS task scaling, known rollback patterns. Novel or ambiguous failures escalate to a human with full context attached
Acquisition Onboarding
- Platform audit and gap analysis for every new acquisition — assessing CI/CD maturity, IaC coverage, observability gaps, and security posture
- Migration plan and execution for each portfolio company joining the platform.
- Target: full platform integration within a defined window per acquisition
What We're Looking For
Experience & Background
- 8–12 years in platform engineering, DevOps, or SRE — with clear evidence of increasing ownership over time
- Strong Terraform depth across multi-environment, multi-account setups
- CI/CD ownership across a multi-product environment with GitHub Actions
- Experience with event streaming infrastructure at production scale — design, operations, reliability, and cost management
- Hands-on Grafana, Prometheus, and Loki in production
- AWS operational depth: ECS, EKS, RDS, IAM, VPC, CloudWatch, Cost Explorer
- SRE fundamentals: SLOs, error budgets, on-call design, post-mortem culture
- Acquisition or greenfield platform integration experience strongly preferred
This is a remote position.
About Leegality:
Leegality works with large Indian businesses to digitally transform critical compliance processes in a fast, easy and secure way.
We have multiple products across 2 categories:
Document Infrastructure:
Products that help businesses build paperless processes at scale:
- Document Execution Workflow: A unified platform for businesses to digitally execute (eSign, eStamp, Template Pre-fill, Document Fraud Prevention etc.) agreements, forms and other documents in a compliant way. Currently in use by 2000+ Indian businesses from giants like HDFC and SBI Cards to high-growth disruptors like goDigit and Cars24.
- Contract Management: An AI-powered platform for businesses to quickly review, negotiate and take action on contracts
- Signstation: A simple platform for businesses to digitally sign simple documents like invoices, policies and letters in a cost effective manner
Consent Infrastructure:
- Consentin: An end-to-end DPDP and Privacy compliance platform for Indian businesses
- Consentin Lens: A data discovery platform for businesses to identify the personal data they collect and store.
If you’re interested in building mission critical software that operates at population scale (75 million + Indians have signed at least one document through Leegality) then join Leegality.
Curious about our impact? Explore our customer success stories: leegality.com/case-studies
Our Culture
At Leegality, trust, ownership, transparency, and having fun while doing meaningful work are core to how we operate — not just values on paper. Our team rated us an incredible 97 eNPS for FY 2023–24 — the highest among 175+ startups surveyed.
We focus deeply on helping our people grow and stay motivated. Some of the perks you’ll enjoy:
- Flexible working hours
- Hybrid work setup
- Bi-annual performance appraisals
- A culture that rewards initiative, curiosity, and impact
If you're looking for a place where you can make a real difference while working with smart, driven, and genuinely nice people, welcome to Leegality.
Location: Hybrid
Role Overview
We are looking for a Technical Lead – Python to lead the design, development, and evolution of scalable backend systems that power Leegality's products. This role combines hands-on backend development with technical leadership, enabling you to influence architecture, mentor engineers, and drive engineering excellence across the team.
You will work closely with Product, Design, QA, DevOps, and other engineering teams to build reliable, secure, and high-performance applications while fostering a culture of ownership, collaboration, and continuous improvement.
Key Responsibilities
- Lead the design, development, and maintenance of scalable backend applications using Python.
- Own the technical architecture of backend systems, ensuring scalability, reliability, maintainability, and security.
- Design, develop, and optimize RESTful APIs and microservices for high-performance applications.
- Drive architectural decisions, establish engineering best practices, and promote clean, maintainable code.
- Conduct code reviews and mentor engineers through technical guidance, pair programming, and knowledge sharing.
- Lead technical estimation, sprint planning, solution design, and execution for engineering initiatives.
- Optimize application performance through database tuning, caching, asynchronous processing, and efficient system design.
- Ensure high code quality through unit testing, integration testing, CI/CD pipelines, and automated deployment practices.
- Troubleshoot and resolve complex production issues while driving root cause analysis and long-term improvements.
- Evaluate and adopt new technologies, frameworks, and engineering practices to continuously improve the platform.
- Build and foster a high-performing engineering culture focused on collaboration, accountability, innovation, and continuous learning.
Desired Skills
- 7–11 years of experience in backend software development, with at least 2–4 years in a technical leadership role.
- Strong proficiency in Python and object-oriented programming concepts.
- Hands-on experience with Python frameworks such as Django, Flask, or FastAPI.
- Strong understanding of RESTful API design and development, authentication mechanisms (JWT/OAuth), and API security best practices.
- Experience designing scalable backend architectures, distributed systems, and microservices.
- Strong knowledge of relational databases such as MySQL or PostgreSQL, with experience in query optimization, indexing, and schema design.
- Exposure to NoSQL databases such as MongoDB and caching technologies like Redis.
- Experience with cloud platforms such as AWS, containerization using Docker, orchestration using Kubernetes, and CI/CD pipelines.
- Strong understanding of software engineering best practices, including design patterns, SOLID principles, code reviews, testing, and documentation.
- Experience with version control systems such as Git and modern development workflows.
- Excellent analytical, problem-solving, communication, and stakeholder management skills.
- Ability to balance hands-on development with technical leadership and delivery ownership in a fast-paced product environment.
- Experience working in SaaS or product-based organizations is preferred.
- Exposure to AI-assisted development tools such as GitHub Copilot, Cursor, or ChatGPT is an added advantage.













