Staff Software Engineer, AI Agents at Asha Health (YC F24) · Bengaluru (Bangalore) · 3 - 7 years · ₹100L - ₹100L / yr · Raised funding · Posted 21 Jun 2026

About Asha Health
Asha Health helps medical practices launch their own AI clinics. We're backed by Y Combinator, General Catalyst, 186 Ventures, Reach Capital and many more. We recently raised an oversubscribed seed round from some of the best investors in Silicon Valley. Our team includes AI product leaders from companies like Google, physician executives from major health systems, and more.
About the Role
We're looking for a top 0.01% Software Engineer to join our engineering team in our Bangalore office.
4.6 fundamentally changed the game, which means that high intelligence, high agency engineers can now do the work of 10+ good engineers. It doesn't make sense to have anyone but the best on the team.
Since low level coding has become easier, what we expect from engineers on our team has expanded. Engineers on our team are expected to:
- Ship features end to end at a rapid pace
- Deeply research the domain and be their own product managers
- Ensure reliability and quality is best-in-class
- Design robust eng architecture, and develop testing and observability tools for each feature pre-launch
- Build each feature with deep customer empathy, meaning planning out and building stellar UX yourself
This means to thrive in a startup environment like ours, you not only need to be a stellar engineer, but you need to:
- Be super adept with AI development tools and building the AI systems that build your features for you (Conductor, Browser agents, QA agents, Ralph loops, adverserial agents, and more).
- Have exceptional product and UX taste, meaning you can ship features that are more effective than those historically designed by teams of product managers and designers.
- Take the highest level of ownership around feature outcomes, reliability, and observability.
Other Points to Note
- We are growing rapidly, our work has impact on tens of thousands of patients if not more.
- On our team, everything you do is on the bleeding edge of applied AI.
- We expect a high level of commitment from everyone on the team, most folks work 6 days and lead every project with intensity. It's a high ask, and we only bring on the best people. We compensate significantly above market, accordingly.

About Asha Health (YC F24)
About
Asha Health is a Y Combinator backed AI healthcare startup. We help medical practices spin up their own AI clinic. We've raised an oversubscribed seed round backed by top Silicon Valley investors, and are growing rapidly. Our team consists of AI product experts from companies like Google, as well as senior physician executives from major health systems.
Tech stack
Candid answers by the company
We help medical practices spin up their own AI clinic.
Similar jobs (10)
About Us
We believe the future of software development is AI-native — where engineers operate at a higher level of abstraction and quality remains non-negotiable.
Incubyte is a software craft consultancy where the “how” of building software matters as much as the “what”.
We partner with companies of all sizes, from helping enterprises build, scale, and modernize to early-stage founders bring their ideas to life.
Our engineers operate in an AI-native development model, using AI as a collaborator across the SDLC to accelerate development while upholding the discipline of software craftsmanship. Guided by Software Craftsmanship and Extreme Programming practices, we build reliable, maintainable, and scalable systems with speed, without compromising quality. If this way of building software resonates with you, we’d like to talk.
Our Guiding Principles
These principles define how we work at Incubyte. They are non-negotiable.
Relentless Pursuit of Quality with Pragmatism
We build high-quality systems without losing sight of delivery.
Extreme Ownership
We take responsibility end-to-end for decisions, execution, and outcomes.
Proactive Collaboration
We collaborate closely, challenge each other, and solve problems together.
Active Pursuit of Mastery
We continuously improve our craft and raise our bar.
Invite, Give, and Act on Feedback
We seek, give, and act on feedback to get better every day.
Ensuring Client Success
We act as trusted partners and focus on real outcomes, not just output.
Experience Level
This role is ideal for engineers with total 3+ years of experience with a proven track record of shipping complex projects successfully.
An experienced individual contributor and leader who thrives in large, complex projects with widespread impact.
What You’ll Do as a Software Craftsperson
- Design and build high-quality, maintainable systems using disciplined engineering practices such as TDD, continuous refactoring, and pair programming
- Operate in an AI-native development model, using AI as a collaborator to explore architecture and design, accelerate development, and continuously improve systems while applying strong judgment to ensure that speed never compromises quality
- Take end-to-end ownership of outcomes from problem understanding and system design to implementation, deployment, and operation in production
- Make thoughtful design decisions that balance simplicity, scalability, and long-term maintainability in real-world systems
- Maintain a high bar for engineering quality through rigorous testing, code reviews, and continuous feedback
- Investigate and resolve production issues, and implement systemic improvements to prevent recurrence
- Work directly with clients, navigate ambiguity, and translate business problems into well-designed technical solutions
- Contribute to improving team practices, tooling, and systems to raise the overall quality and effectiveness of engineering
Requirements
What You’ll Bring
- 3+ years of experience building high-quality, production systems (flexible based on demonstrated capability)
- Strong fundamentals in software engineering, including object-oriented design, system design, and testing practices such as TDD
- Demonstrated ability to build simple, maintainable, and scalable systems with a focus on long-term reliability
- Proficiency in one or more modern technologies, Python, PHP, JavaScript, or TypeScript, with the ability to learn new technologies quickly
- Deep experience working with Git in collaborative environments, including managing shared codebases, conducting code reviews, and maintaining a high bar for quality
- Ability to operate effectively in an AI-native workflow using AI as a collaborator to explore solutions and accelerate development, while applying strong judgment to ensure correctness, quality, and maintainability
- Clear thinking and strong problem-solving ability, with the capacity to break down complex problems into simple, well-structured solutions
- A strong sense of ownership — you take responsibility for outcomes, care deeply about quality, and are not comfortable shipping work that does not meet your standards.
Benefits
Life at Incubyte
We are a remote-first company with structured flexibility. Teams commit to shared rhythms during core hours, ensuring smooth collaboration while maintaining autonomy. Twice a year, we come together in person for a co-working sprint and once a year for a retreat - with all travel expenses covered.
Our environment is built for crafters: pairing, refactoring, experimenting with AI, and pushing the boundaries of software excellence. We are all lifelong learners, and our work is our passion.
Perks
- Dedicated learning & development budget.
- Sponsorship for conference talks.
- Comprehensive medical & term insurance.
- Employee-friendly leave policies.
- Home Office fund
- Medical Insurance
Product Engineer — Role Summary
We are looking for a Product Engineer to build user-facing products that combine AI capabilities with practical applications. You will work at the intersection of software engineering and AI, developing autonomous, agent-driven systems that solve complex educational and research problems.
Key Responsibilities:
- Product Development: Design, build, and deploy production-ready applications powered by LLMs and AI agents.
- Data Engineering: Build scalable ETL/ELT pipelines to process structured and unstructured data, including text and audio, for RAG and model fine-tuning.
- Agentic Workflows: Develop multi-step AI agents with tool calling, APIs, databases, search, reasoning, and memory.
- Rapid Prototyping: Turn ideas and research concepts into interactive, production-ready applications.
- AI Integration: Use frameworks such as LangChain, LlamaIndex, AutoGen, or custom orchestrators to integrate AI into scalable systems.
- User Experience: Transform raw AI outputs into reliable, intuitive, and responsive user experiences.
- Collaboration: Work closely with ML researchers and data engineers to integrate custom and fine-tuned models.
- Observability: Monitor agent behavior, manage edge cases, reduce hallucinations, and improve reliability in production.
The ideal candidate combines strong software engineering, AI/LLM expertise, data engineering, and product thinking, with the ability to take an AI concept from prototype to production.
We are looking for a Senior Full Stack Engineer to join our lean, high-impact engineering team. This is a key founding-team role where you will work closely with the CTO and help build the engineering foundation of the company.
The ideal candidate is a backend-strong full-stack engineer with solid Python and React experience, a strong understanding of AI/LLMs, and the ability to independently take ownership of projects from idea to production.
What You'll Do:
- Build and scale applications across the backend and frontend
- Develop backend services and APIs using Python
- Build and maintain frontend applications using React
- Design and implement AI/LLM-powered applications and features
- Work with AI services and integrate them into real-world applications
- Contribute to system design and architecture
- Work with cloud infrastructure and services
- Take ownership of features and initiatives from ideation to production
- Identify problems and proactively drive solutions without requiring constant direction
- Work closely with the CTO and founding team to shape engineering practices and future systems
- Communicate technical concepts clearly to technical and non-technical stakeholders
What We're Looking For:
- 3–6 years of software engineering experience
- Strong experience with Python
- Good hands-on experience with React
- Strong backend development experience with full-stack exposure
- Practical understanding of AI, LLMs and AI services
- Experience building at least one meaningful AI project beyond basic API/chatbot integrations
- Experience taking an AI/ML application or feature into production is preferred
- Exposure to cloud platforms
- Good understanding of system design and software architecture
- Strong problem-solving and ownership mindset
- Excellent communication skills and ability to explain complex technical concepts in simple language
- Genuine hands-on experience with the technologies mentioned on the resume
What Makes This Role Different:
This is not a typical execution-focused engineering role. We are building a lean "A-team" of experienced engineers who will eventually help shape the team and take on senior/leadership responsibilities.
We're looking for someone who:
- Takes initiative rather than waiting for tasks
- Can work independently in an early-stage environment
- Is comfortable exploring and building with AI
- Can go deep into the work they've done and explain the technical decisions behind it
- Wants significant ownership and influence over the product and engineering direction
Location & Compensation
- Work Mode: Remote initially
- Future Location: Expected relocation to Bangalore in approximately 1 year
- Working Hours: Regular IST hours
- Travel: Potential opportunities to travel to Germany
Interview Process:
3 Rounds
- Technical & Project Deep Dive – Conversational discussion around past projects, AI experience and technical depth
- System Design & Technical Round – Deeper technical and architecture discussion
- Founder Round – Communication, culture fit, ownership and initiative
AI Engineer
LLMs, Agents & AI Services
📍 Mumbai (On-site) | Full-time | 2-4 years
About the Role:
Unico Connect is an AI-first technology partner that builds custom mobile, web, and AI products for clients across multiple geographies.
AI is core to how we design, deliver, and scale software for our customers.
We are hiring an AI Engineer for a dedicated client engagement building a complex production AI platform, working on the AI capabilities and agentic features at the core of the product.
The mandatory requirement for this role is at least one AI feature personally shipped to production for real users, with operational ownership.
The role suits someone who thinks quickly on solutioning, can take an ambiguous problem to a working prototype in days, and has the discipline to carry it through to production with predictable economics.
You will work alongside the Senior AI Engineer and the wider pod, with ownership of parts of the AI surface area of the product.
Responsibilities:
Solutioning and POCs
Translate ambiguous customer problems into working POCs at speed.
Pick the right model, framework, and architecture, and demonstrate value early before scaling investment.
LLM Application Development
Build AI features and services using LLM APIs from OpenAI, Anthropic, Google, and self-hosted open-weight models (Llama, Qwen, Mistral).
Choose the right model per use case based on cost, latency, capability, and context-window trade-offs.
Agentic System Design
Design and implement agentic workflows using LangGraph, CrewAI, AutoGen, LlamaIndex Agents, or custom orchestration.
Cover tool use, planning, memory, and multi-step reasoning appropriate to the problem.
API and Service Development
Build production AI services and APIs using Python and FastAPI.
Handle streaming responses, async processing, structured outputs, retries, and graceful degradation when models or tools fail.
Retrieval and Tool Integration
Implement RAG pipelines with vector databases (Pinecone, Weaviate, Qdrant, pgvector, Chroma), embeddings, chunking strategies, hybrid search, and reranking.
Integrate external tools, internal APIs, and document sources through tool-calling and MCP-style patterns.
Cost Analysis and Unit Economics
Model the per-request and per-user cost of every AI feature before it ships.
Track token usage, prompt caching, batching, and model-routing strategies.
Drive measurable improvements in unit economics.
Production Hardening
Add observability and tracing (LangSmith, Langfuse, OpenTelemetry), guardrails, content safety checks, prompt injection defences, and fallback behaviour.
Prompt Engineering and Evaluation
Design, test, and iterate prompts with measured outcomes.
Build evaluation harnesses for accuracy, hallucination, latency, and cost.
Run benchmarks across models and prompt variants before locking in a design.
Requirements:
AI Feature Shipped to Production (Mandatory)
Must have personally built and shipped at least one AI feature that runs in production for real users, with operational ownership.
POCs, internal demos, and one-off scripts do not qualify.
2 to 4 Years of Professional Software or AI Engineering Experience
With at least one production AI feature owned end to end.
Strong Python Proficiency and API Development with FastAPI
Comfort with type hints, async, packaging, testing, streaming responses, and authentication.
Production-grade Python, not notebook-only code.
Hands-on Depth Across the LLM and Agent Stack
Working experience with at least two of OpenAI, Anthropic Claude, Google Gemini, or self-hosted open-weight models (vLLM, Ollama, Together, Replicate).
Working familiarity with at least one agent framework (LangGraph, CrewAI, AutoGen, LlamaIndex Agents) or hand-rolled equivalent.
Working knowledge of RAG, embeddings, and vector databases (Pinecone, Weaviate, Qdrant, pgvector, Chroma).
Solutioning Speed and POC Velocity
Demonstrated ability to move from a fuzzy problem to a working prototype in days.
Strong instinct for what to build first, what to defer, and what to throw away.
Cost Discipline for Production AI
Ability to calculate, monitor, and optimise the cost of LLM APIs, tokens, embeddings, vector store usage, and infrastructure.
Treats unit economics as a first-class concern.
AWS Familiarity
Working knowledge of EC2, S3, IAM, and at least one of Bedrock, SageMaker, or equivalent.
Comfortable in a Fast-Moving Environment
Self-directed, comfortable with ambiguity, takes ownership without being asked, and ships under shifting priorities.
Strong Written and Spoken English Communication
Able to explain trade-offs to non-AI engineers, designers, product managers, and clients in plain language.
Nice to Have
- fine-tuning or LoRA, QLoRA, PEFT exposure
- MCP server authoring
- eval framework experience (LangSmith, Promptfoo, Ragas, DeepEval)
- open-source AI contributions
- multi-modal models (vision, audio)
About Us
We believe the future of software development is AI-native — where engineers operate at a higher level of abstraction and quality remains non-negotiable.
Incubyte is a software craft consultancy where the “how” of building software matters as much as the “what”.
We partner with companies of all sizes, from helping enterprises build, scale, and modernize to early-stage founders bring their ideas to life.
Our engineers operate in an AI-native development model, using AI as a collaborator across the SDLC to accelerate development while upholding the discipline of software craftsmanship. Guided by Software Craftsmanship and Extreme Programming practices, we build reliable, maintainable, and scalable systems with speed, without compromising quality. If this way of building software resonates with you, we’d like to talk.
Our Guiding Principles
These principles define how we work at Incubyte. They are non-negotiable.
Relentless Pursuit of Quality with Pragmatism
We build high-quality systems without losing sight of delivery.
Extreme Ownership
We take responsibility end-to-end for decisions, execution, and outcomes.
Proactive Collaboration
We collaborate closely, challenge each other, and solve problems together.
Active Pursuit of Mastery
We continuously improve our craft and raise our bar.
Invite, Give, and Act on Feedback
We seek, give, and act on feedback to get better every day.
Ensuring Client Success
We act as trusted partners and focus on real outcomes, not just output.
Job Description
This is a remote position.
This is a remote position.
Experience Level
This role is ideal for engineers with 3–6 years of experience and a strong background in building scalable, production-grade software systems.
We are looking for hands-on Software Engineers with deep expertise in Python and TypeScript, with exposure to AI/LLM systems and modern infrastructure tooling.
What You’ll Do as a Software Craftsperson
• Take full ownership of the software development lifecycle for complex, cross-functional initiatives — from design through production readiness.
• Build and maintain robust, scalable backend systems and APIs using Python and TypeScript, following clean code and software craftsmanship principles.
• Design and deliver features end-to-end, balancing scope, quality, and long-term maintainability.
• Identify technical and product issues beyond the immediate scope of work, proactively raising risks and driving solutions.
• Shape team practices around code quality, testing, tooling, and continuous improvement using DevEx and DORA principles.
• Collaborate closely with clients and internal teams to understand requirements, clarify priorities, and align on outcomes.
• Mentor and guide fellow engineers to raise overall team performance and promote a culture of learning.
• Leverage AI tools (LLMs, agentic frameworks, etc.) to accelerate design, development, testing, and delivery where applicable.
Requirements
What You’ll Bring
3–6 years of overall software engineering experience with a strong track record of owning and delivering complex production systems.
Must-Have Skills
• Python (must-have): Deep expertise in writing idiomatic, testable, production-grade Python — including advanced OOP, data structures, algorithms, and software engineering best practices.
• TypeScript (must-have): Strong proficiency in building and maintaining type-safe, scalable applications across frontend and/or backend TypeScript codebases.
• Strong system design skills: ability to architect scalable, maintainable, and observable systems with a focus on reliability and long-term operability.
• Solid engineering practices: experience with TDD, CI/CD, code reviews, refactoring, and continuous deployment in Agile or eXtreme Programming environments.
• Working knowledge of relational databases, web server ecosystems, REST/gRPC APIs, and performance optimisation.
• Experience with source control, bug tracking, user story writing, and maintaining clear technical documentation.
Good-to-Have Skills
• AI / LLM experience: hands-on exposure to building applications using LLMs (GPT, Claude, Gemini, or similar) for tasks such as document classification, entity extraction, or structured data generation.
• LLM orchestration frameworks: familiarity with LangChain, LlamaIndex, Mastra, Agno, or similar agentic frameworks.
• Prompt engineering: experience refining prompts and orchestration patterns to improve response accuracy, consistency, and structured outputs.
• Vector databases and observability tooling: exposure to tools like Pinecone, PgVector, Qdrant, LangSmith, or DeepEval.
• Container orchestration and infrastructure-as-code: experience with Kubernetes, Terraform, and Docker for deploying and managing production workloads.
Benefits
Life at Incubyte
We are a remote-first company with structured flexibility. Teams commit to shared rhythms during core hours, ensuring smooth collaboration while maintaining autonomy. Twice a year, we come together in person for a co-working sprint and once a year for a retreat — with all travel expenses covered.
Our environment is built for crafters: experimenting with real-world systems, solving complex infrastructure challenges, and contributing to cutting-edge AI initiatives. We are all lifelong learners, and our work is our passion.
Perks
• Dedicated learning & development budget
• Sponsorship for conference talks
• Comprehensive medical & term insurance
• Employee-friendly leave policies
• Home Office fund
• Medical Insurance
About Us
We believe the future of software development is AI-native — where engineers operate at a higher level of abstraction and quality remains non-negotiable.
Incubyte is a software craft consultancy where the “how” of building software matters as much as the “what”.
We partner with companies of all sizes, from helping enterprises build, scale, and modernize to early-stage founders bring their ideas to life.
Our engineers operate in an AI-native development model, using AI as a collaborator across the SDLC to accelerate development while upholding the discipline of software craftsmanship. Guided by Software Craftsmanship and Extreme Programming practices, we build reliable, maintainable, and scalable systems with speed, without compromising quality. If this way of building software resonates with you, we’d like to talk.
Our Guiding Principles
These principles define how we work at Incubyte. They are non-negotiable.
Relentless Pursuit of Quality with Pragmatism
We build high-quality systems without losing sight of delivery.
Extreme Ownership
We take responsibility end-to-end for decisions, execution, and outcomes.
Proactive Collaboration
We collaborate closely, challenge each other, and solve problems together.
Active Pursuit of Mastery
We continuously improve our craft and raise our bar.
Invite, Give, and Act on Feedback
We seek, give, and act on feedback to get better every day.
Ensuring Client Success
We act as trusted partners and focus on real outcomes, not just output.
Job Description
This is a remote position.
Experience Level
This role is ideal for engineers with total 5+ years of experience with a proven track record of shipping complex projects successfully.
An experienced individual contributor and leader who thrives in large, complex projects with widespread impact.
What You’ll Do as a Software Craftsperson
- Design and build high-quality, maintainable systems using disciplined engineering practices such as TDD, continuous refactoring, and pair programming
- Operate in an AI-native development model, using AI as a collaborator to explore architecture and design, accelerate development, and continuously improve systems while applying strong judgment to ensure that speed never compromises quality.
- Take end-to-end ownership of outcomes from problem understanding and system design to implementation, deployment, and operation in production
- Make thoughtful design decisions that balance simplicity, scalability, and long-term maintainability in real-world systems
- Maintain a high bar for engineering quality through rigorous testing, code reviews, and continuous feedback
- Investigate and resolve production issues, and implement systemic improvements to prevent recurrence
- Work directly with clients, navigate ambiguity, and translate business problems into well-designed technical solutions
- Contribute to improving team practices, tooling, and systems to raise the overall quality and effectiveness of engineering
Requirements
What You’ll Bring
- 5+ years of experience building high-quality, production systems (flexible based on demonstrated capability)
- Strong fundamentals in software engineering, including object-oriented design, system design, and testing practices such as TDD
- Demonstrated ability to build simple, maintainable, and scalable systems with a focus on long-term reliability
- Proficiency in one or more modern technologies Python, React, AI, JavaScript, or TypeScript, with the ability to learn new technologies quickly
- Deep experience working with Git in collaborative environments, including managing shared codebases, conducting code reviews, and maintaining a high bar for quality
- Ability to operate effectively in an AI-native workflow using AI as a collaborator to explore solutions and accelerate development, while applying strong judgment to ensure correctness, quality, and maintainability
- Clear thinking and strong problem-solving ability, with the capacity to break down complex problems into simple, well-structured solutions
- A strong sense of ownership — you take responsibility for outcomes, care deeply about quality, and are not comfortable shipping work that does not meet your standards.
Benefits
Life at Incubyte
We are a remote-first company with structured flexibility. Teams commit to shared rhythms during core hours, ensuring smooth collaboration while maintaining autonomy. Twice a year, we come together in person for a co-working sprint and once a year for a retreat - with all travel expenses covered.
Our environment is built for crafters: pairing, refactoring, experimenting with AI, and pushing the boundaries of software excellence. We are all lifelong learners, and our work is our passion.
Benefits
- Dedicated learning & development budget.
- Sponsorship for conference talks.
- Comprehensive medical & term insurance.
- Employee-friendly leave policies.
- Home Office fund
- Medical Insurance
We are looking for an Engineering Lead to own the entire technology stack — from onboarding and underwriting to disbursals, repayments, and collections — and to build the engineering function into something genuinely AI-native.
What You'll Own
● Full tech stack: backend, frontend, infrastructure, integrations, and data pipelines
● Real-time underwriting and decisioning systems
● LOS/LMS architecture — onboarding, disbursals, repayments, and collections
● Integrations with bureaus, KYC providers, account aggregators, and payment gateways
● Reconciliation systems — disbursement, repayment, and NACH reconciliation end-to-end
● AWS infrastructure: scaling, reliability, uptime, and cloud cost ownership ● Data infrastructure for the credit and risk team — feature pipelines, model serving, experiment infrastructure
● Engineering leadership: hiring, sprint planning, code reviews, and execution standards
● Compliance systems: RBI guidelines, DPDP, KYC/AML, e-NACH, e-sign
AI-Native Engineering
This is a core part of the role, not a bonus. You will build a machine-readable knowledge base of the entire codebase — architecture, data models, service contracts, coding standards, decision history — so that AI agents working on code have the context to produce accurate, consistent output. You will build skills for code review, developer onboarding, and recurring engineering workflows. You will build a code review pipeline where agents do the first pass on every pull request. The knowledge base and the skills improve over time as the team grows and the product evolves.
What We're Looking For
● 7+ years in software engineering, with at least 2 years leading teams or architecture
● Strong hands-on experience with Python, Django, and React Native
● Deep expertise in AWS and cloud-native architecture
● Experience with both SQL and NoSQL databases
● Strong understanding of distributed systems, microservices, and API design
● Experience owning reconciliation or payment flow infrastructure in a lending or payments context
● Prior experience in fintech / NBFC / digital lending — mandatory
● Strong understanding of the full loan lifecycle — mandatory
● You have used LLMs seriously as engineering tools and have strong opinions about what makes AI-assisted development produce good output versus mediocre output
Bonus: Kubernetes / Kafka, AI/ML-driven underwriting, Account Aggregator framework, e-NACH / e-Sign / Video KYC integrations
What Success Looks Like
● scales with strong uptime, performance, and reliability
● Reconciliation runs cleanly — no financial discrepancies surface late ● A new engineer joins and is writing standard, correct code within their first week
● The credit team is never blocked on an engineering dependency
● Engineering health metrics are tracked and visibly improving
● AI agents are doing the structured first pass on code reviews, and the system gets smarter over time
The Role
You own AI systems end to end. From the speech-to-text models that turn audio into text, to the diarization that separates and identifies speakers, to the agentic layer that turns conversation into memory and action, to the observability and evaluation that keep all of it honest in production. This is a wide role by design. You will own model selection, serving, and production reliability. If you want to tune one model and ignore the system around it, this is not the role.
What You Will Own
• Speech-to-text. Evaluate, integrate, and optimize STT models across cloud and self-hosted. Drive accuracy and cost trade-offs with ground-truth metrics.
• Speaker diarization and identification. Push accuracy on hard, real-world, multi-speaker audio.
• Agentic AI. Build the memory and retrieval pipeline, LLM orchestration, and the agent workflows that sit on top of captured conversation.
• Model serving and infrastructure. Stand up and optimize self-hosted serving (vLLM, Triton class). Own latency, throughput, and cost per user.
Observability
An always-on wearable means models run in production every second, on messy real-world audio. You own the visibility into that.
• Instrument the full audio-to-memory pipeline: STT, diarization, retrieval, and LLM calls.
• Define and track model-quality SLOs in production: transcription drift, diarization error over time, retrieval relevance, latency, throughput, and cost per user.
• Build dashboards and alerting so model degradation is caught before users feel it.
• Trace failures across a distributed, always-on system using metrics, logs, and traces.
• Close the loop. Production signals feed back into evaluation and model selection.
Evaluation
We do not ship what we cannot measure. You own the systems that prove a model is actually better, not just newer.
• Build and own ground-truth evaluation harnesses for every model in the stack.
• Measure with real metrics: WER for transcription, DER for diarization, Recall and F1 for retrieval and speaker identification.
• Build and maintain labeled benchmark datasets that reflect real, messy, multi-speaker audio.
• Run regression and A/B evaluations on every model swap, prompt change, or pipeline update. Nothing ships on a vibe.
• Reject anecdotal proxies, single confidence scores, and cherry-picked examples as evidence of quality.
What We Are Looking For
• 3 to 5 years as an AI/ML engineer with production systems behind you. Engineering and production experience is non-negotiable.
• Depth across the modern AI stack: LLMs, speech models, vector retrieval, model serving.
• Strong software engineering. You write code that ships and survives contact with real users.
• Fluency in Python and the production ML ecosystem.
• Comfort with cloud infrastructure (GCP a plus) and containerized deployment on Kubernetes.
• A working command of observability and evaluation. You measure first and trust metrics over intuition.
• First-principles reasoning and metric discipline.
Nice to Have
• Research background or publications. A strong signal, not a substitute for production work.
• Audio and speech ML experience (STT, diarization, voice).
• Experience self-hosting and optimizing open models.
• Experience with LLM gateway and agent orchestration patterns.
• Experience building eval harnesses or production model-monitoring systems.
Requirements
Agentic work is must. Audio is good to have
. Self hosting models is a must
Experience with LLM gateway and agent orchestration is a must have
Full-Stack Software Engineer | Remote
We're hiring a product-focused Full-Stack Engineer to join a small, fast-moving tech team. This is a hands-on role you'll build complete features across frontend, backend, APIs, databases, deployment, and increasingly, AI-assisted workflows.
What You'll Do:
- Build and ship full-stack features using modern JS/TypeScript, plus Java or Node.js on the backend
- Work across frontend, backend, APIs, databases, integrations, and deployment
- Apply secure engineering practices (auth, input validation, secrets, dependencies)
- Use AI tools for coding, debugging, testing, and code review
- Explore AI agents, tool calling, and MCP-based integrations
What We Need:
- 3+ years of professional software engineering experience
- Strong JS/TypeScript skills; experience with Angular or React
- Backend experience with Java, Node.js, or similar
- Comfortable with REST APIs, databases, Git, cloud environments
- Awareness of OWASP principles and application security
- Interest/experience in AI dev tools, Docker, CI/CD
Job Description – AI Engineer (End-to-End Development & Deployment)
Role Summary
We are looking for an AI Engineer with hands-on experience in designing, developing, deploying, and maintaining Generative/Agentic AI solutions in production. The ideal candidate should have end-to-end ownership of AI applications, from development to deployment, monitoring, and optimization.
Key Responsibilities
● Design, build, and deploy Generative/Agentic AI solutions.
● Develop applications using LLMs, RAG, AI agents, and vector databases.
● Build scalable APIs and integrate AI solutions with enterprise applications.
● Implement CI/CD pipelines, containerization, and MLOps best practices.
● Monitor, optimize, and maintain production AI systems.
● Collaborate with cross-functional teams to deliver business-driven AI solutions.
Required Skills
● Strong programming skills in Python.
● Experience with vector databases (e.g., Pinecone, FAISS, ChromaDB) and graph memory systems
● Knowledge of atleast one agent development framework: Google ADK (preferred), LangChain/LangGraph/LlamaIndex, CrewAI
● Experience with LLMs, RAG, GenAI, AgenticAI Agents
● Hands-on experience with FastAPI, and REST APIs.
● Knowledge of Docker, Kubernetes, Git, CI/CD.
● Experience with AWS, Azure, or GCP.
● Experience with security compliance, monitoring and observability tools such as AWS CloudWatch, Azure Monitor, Google Cloud Monitoring.










