AI/Full-Stack Engineer – Help Us Transform Home Loan Automation at InvestPulse · Remote only · 2 - 5 years · ₹3L - ₹6L / yr · Bootstrapped · Remote only · Posted 27 Jun 2025

AI/Full-Stack Engineer – Help Us Transform Home Loan Automation
at InvestPulse
LendFlow is an AI-powered home loan assessment platform that helps mortgage brokers and lenders save hours by automating document analysis, income validation, and serviceability assessment. We turn complex financial documents into clear insights—fast.
We’re building a smart assistant that ingests client docs (bank statements, payslips, loan summaries) and uses modular AI agents to extract, classify, and summarize financial data in minutes, not hours. Think OCR + AI agents + compliance-ready outputs.
🛠️ What You’ll Be Building
As part of our early technical team, you’ll help us develop and launch our MVP. Key modules include:
- Document ingestion and OCR processing (Textract, Document AI)
- AI agent workflows using LangChain or CrewAI
- Serviceability calculators with business rule engines
- React + Next.js frontend for brokers and analysts
- FastAPI backend with PostgreSQL
- Security, encryption, audit logging (privacy-first design)
🎯 We’re Looking For:
Must-Have Skills:
- Strong experience with Python (FastAPI, OCR, LLMs, prompt engineering)
- Familiarity with AI agent frameworks (LangChain, CrewAI, Autogen, or similar)
- Frontend skills in React.js / Next.js
- Experience with PostgreSQL and cloud storage (AWS/GCP)
- Understanding of financial documents and data privacy best practices
Bonus Points:
- Experience with OCR tools like Amazon Textract, Tesseract, or Document AI
- Building ML/NLP pipelines in real-world apps
- Prior work in fintech, lending, or proptech sectors

About InvestPulse
About
A centralised platform designed to help property investors manage their portfolios. Provides real-time insights on rental income, expenses, and key financial metrics. Helps streamline finances, reduce reliance on spreadsheets, and support data-driven decision-making.
Tech stack
Candid answers by the company
At InvestPulse, we're on a mission to empower real estate investors with the tools and insights they need to make informed decisions and maximise the potential of their investment properties.
Product showcase
Company social profiles
Similar jobs (10)
We are looking for an experienced Software Engineer to join an AI engineering startup developing a document collection platform for accountants and professional services firms.
Preference to candidates from Kerala, India.
The product eliminates the friction involved in gathering client files by automating document requests, centralising their collection, and organising incoming documents according to each organisation’s preferred folder structure.
The ideal candidate will be able to take ownership of work from start to finish, communicate clearly, and deliver high-quality solutions within tight timeframes.
What You’ll Work On
You’ll work with Python and Django daily, including models, views, templates, background jobs, and the wider product around them.
The frontend uses Django templates with HTMX and Alpine.js, built with Vite, TypeScript, and Tailwind CSS. The stack runs in Docker using PostgreSQL, Redis, RabbitMQ, and Celery.
You may also assist with ancillary projects, including custom integrations.
Project-based training will be provided.
Technology Stack
Backend: Python, Django, PostgreSQL, Celery, Redis and RabbitMQ
Frontend: Django Templates, HTMX, Alpine.js, Vite, TypeScript and Tailwind CSS
Infrastructure: Docker
Must Have
- Strong Python and Django skills
- Comfortable working with Docker
- Fluent written and spoken English
- Clear communication skills, including providing concise updates, asking honest questions, and writing information that others can act on
- Evidence of exceptional ability—not simply a list of tools, but something challenging you have built or solved
Preference will be given to candidates with at least three years of relevant professional experience.
Nice to Have
- Frontend experience with HTML, CSS and JavaScript
- Experience with HTMX, Alpine.js, TypeScript or Tailwind CSS
- Knowledge of PostgreSQL, Celery or pytest
- Experience with integrations, including APIs, OAuth and cloud storage
- Basic accounting knowledge
How to Apply
Please do not send a generic CV alone. Your application must include:
- Evidence of exceptional ability: Describe a project, open-source contribution, production system or challenging problem you solved. Include a link to the repository, write-up or demo where possible. Focus on your personal contribution by detailing the specific parts of the project where you played a critical role and explaining precisely what you built or solved.
- What you accomplished: Provide a short explanation of the outcome in your own words.
- The hardest part: Explain the hardest part of the problem and how you dealt with it.
- Your use of AI: Explain whether you use AI in your work and, if so, how you use it.
- Your professional experience and interests: Include a brief paragraph summarising your professional experience and general interests.
Applications that do not include the above Croissant details above will not be considered.
What We Offer
- For the right candidate, salary will not be a constraint
- Project-based training
- A rewarding career with genuine opportunities for professional growth
- The opportunity to work on an innovative AI-driven product
- A remote, full-time position
Job Details and Application Submission
Location: Remote
Employment Type: Full-time
Contract: One-year contract, with the possibility of extension based on satisfactory performance
Probationary Period: Six months
Preferred Experience: Three or more years
About Us
We believe the future of software development is AI-native — where engineers operate at a higher level of abstraction and quality remains non-negotiable.
Incubyte is a software craft consultancy where the “how” of building software matters as much as the “what”.
We partner with companies of all sizes, from helping enterprises build, scale, and modernize to early-stage founders bring their ideas to life.
Our engineers operate in an AI-native development model, using AI as a collaborator across the SDLC to accelerate development while upholding the discipline of software craftsmanship. Guided by Software Craftsmanship and Extreme Programming practices, we build reliable, maintainable, and scalable systems with speed, without compromising quality. If this way of building software resonates with you, we’d like to talk.
Our Guiding Principles
These principles define how we work at Incubyte. They are non-negotiable.
Relentless Pursuit of Quality with Pragmatism
We build high-quality systems without losing sight of delivery.
Extreme Ownership
We take responsibility end-to-end for decisions, execution, and outcomes.
Proactive Collaboration
We collaborate closely, challenge each other, and solve problems together.
Active Pursuit of Mastery
We continuously improve our craft and raise our bar.
Invite, Give, and Act on Feedback
We seek, give, and act on feedback to get better every day.
Ensuring Client Success
We act as trusted partners and focus on real outcomes, not just output.
Experience Level
This role is ideal for engineers with total 3+ years of experience with a proven track record of shipping complex projects successfully.
An experienced individual contributor and leader who thrives in large, complex projects with widespread impact.
What You’ll Do as a Software Craftsperson
- Design and build high-quality, maintainable systems using disciplined engineering practices such as TDD, continuous refactoring, and pair programming
- Operate in an AI-native development model, using AI as a collaborator to explore architecture and design, accelerate development, and continuously improve systems while applying strong judgment to ensure that speed never compromises quality
- Take end-to-end ownership of outcomes from problem understanding and system design to implementation, deployment, and operation in production
- Make thoughtful design decisions that balance simplicity, scalability, and long-term maintainability in real-world systems
- Maintain a high bar for engineering quality through rigorous testing, code reviews, and continuous feedback
- Investigate and resolve production issues, and implement systemic improvements to prevent recurrence
- Work directly with clients, navigate ambiguity, and translate business problems into well-designed technical solutions
- Contribute to improving team practices, tooling, and systems to raise the overall quality and effectiveness of engineering
Requirements
What You’ll Bring
- 3+ years of experience building high-quality, production systems (flexible based on demonstrated capability)
- Strong fundamentals in software engineering, including object-oriented design, system design, and testing practices such as TDD
- Demonstrated ability to build simple, maintainable, and scalable systems with a focus on long-term reliability
- Proficiency in one or more modern technologies, Python, PHP, JavaScript, or TypeScript, with the ability to learn new technologies quickly
- Deep experience working with Git in collaborative environments, including managing shared codebases, conducting code reviews, and maintaining a high bar for quality
- Ability to operate effectively in an AI-native workflow using AI as a collaborator to explore solutions and accelerate development, while applying strong judgment to ensure correctness, quality, and maintainability
- Clear thinking and strong problem-solving ability, with the capacity to break down complex problems into simple, well-structured solutions
- A strong sense of ownership — you take responsibility for outcomes, care deeply about quality, and are not comfortable shipping work that does not meet your standards.
Benefits
Life at Incubyte
We are a remote-first company with structured flexibility. Teams commit to shared rhythms during core hours, ensuring smooth collaboration while maintaining autonomy. Twice a year, we come together in person for a co-working sprint and once a year for a retreat - with all travel expenses covered.
Our environment is built for crafters: pairing, refactoring, experimenting with AI, and pushing the boundaries of software excellence. We are all lifelong learners, and our work is our passion.
Perks
- Dedicated learning & development budget.
- Sponsorship for conference talks.
- Comprehensive medical & term insurance.
- Employee-friendly leave policies.
- Home Office fund
- Medical Insurance
This is a remote position.
About Leegality:
Leegality works with large Indian businesses to digitally transform critical compliance processes in a fast, easy and secure way.
We have multiple products across 2 categories:
Document Infrastructure:
Products that help businesses build paperless processes at scale:
- Document Execution Workflow: A unified platform for businesses to digitally execute (eSign, eStamp, Template Pre-fill, Document Fraud Prevention etc.) agreements, forms and other documents in a compliant way. Currently in use by 2000+ Indian businesses from giants like HDFC and SBI Cards to high-growth disruptors like goDigit and Cars24.
- Contract Management: An AI-powered platform for businesses to quickly review, negotiate and take action on contract
- Signstation: A simple platform for businesses to digitally sign simple documents like invoices, policies and letters in a cost effective manner
Consent Infrastructure:
- Consentin: An end-to-end DPDP and Privacy compliance platform for Indian businesses
- Consentin Lens: A data discovery platform for businesses to identify the personal data they collect and store.
If you’re interested in building mission critical software that operates at population scale (75 million + Indians have signed at least one document through Leegality) then join Leegality.
Curious about our impact? Explore our customer success stories: leegality.com/case-studies
Our Culture
At Leegality, trust, ownership, transparency, and having fun while doing meaningful work are core to how we operate — not just values on paper. Our team rated us an incredible 97 eNPS for FY 2023–24 — the highest among 175+ startups surveyed.
We focus deeply on helping our people grow and stay motivated. Some of the perks you’ll enjoy:
- Flexible working hours
- Hybrid work setup
- Bi-annual performance appraisals
- A culture that rewards initiative, curiosity, and impact
If you're looking for a place where you can make a real difference while working with smart, driven, and genuinely nice people, welcome to Leegality.
Location: Hybrid
Job Brief:
- As a Machine Learning Engineer specializing in Computer Vision (CV) and Natural Language Processing (NLP), you will develop solutions to interesting technical problems, exploring exciting growth opportunities and having a real impact on our product, particularly focusing on document and content intelligence.
- To ensure success, you should demonstrate solid data science knowledge and experience in a related ML, CV, or NLP role. A first-class engineer will be someone whose expertise enhances our systems for document intelligence and content processing
Responsibilities:
- Designing machine learning systems, self-running artificial intelligence (AI) software, and specialized models for Computer Vision and Natural Language Processing applications.
- Transforming data science prototypes and applying appropriate deep learning algorithms and tools to text and image/document data.
- Solving complex CV and NLP problems with multi-layered data types, such as image/document classification, information extraction, semantic search, and object detection.
- Optimizing existing machine learning models, with a focus on high-performance model deployment for CV and NLP tasks.
- Developing ML algorithms (including large language models/LLMs and computer vision models) to analyze huge volumes of historical text, image, and document data to make predictions and automate workflows.
- Running tests, performing statistical analysis, and interpreting test results for CV/NLP model performance.
- Documenting machine learning processes, model architectures, and data pipelines.
- Keeping abreast of developments in machine learning, Computer Vision, and Natural Language Processing.
Requirements:
- 3+ years of relevant experience in Machine Learning Engineering, with a strong focus on Computer Vision and/or Natural Language Processing.
- Advanced proficiency with Python.
- Extensive knowledge of ML frameworks, libraries (e.g., PyTorch, Transformers), data structures, data modeling, and software architecture.
- Experience with building and maintaining scalable RESTful APIs (e.g., FastAPI).
- In-depth knowledge of mathematics, statistics, deep learning (CNNs, RNNs, Transformers), and algorithms.
- Superb analytical and problem-solving abilities, especially for unstructured data challenges.
- Great communication and collaboration skills.
- Excellent time management and organizational abilities.
- Experience with cloud platforms (e.g., AWS) for model deployment and MLOps.
Recruitment Process:
- Our hiring process combines AI-powered evaluations with structured interviews to ensure a fair and seamless experience.
- You will be contacted via email with the next steps upon being shortlisted.
- The process may include Assessments, AI-enabled interviews, and In-Person Interviews with our team.
- Final selection and CTC will be based on your overall performance and experience.
Apply directly through our career page: https://careers.leegality.com/jobs/Careers
For more information about us please visit our:
Our Company and Culture: https://bit.ly/3Iqm5SB
Our Website: www.leegality.com/
Our LinkedIn Page: www.linkedin.com/company/leegality/
Leegality's Privacy Notice: https://www.leegality.com/employee-privacy-notice
About the Role
We are seeking a hands-on Tech Lead to design, build, and integrate AI-driven systems that automate and enhance real-world business workflows. This is a high-impact role for someone who enjoys full-stack ownership — from backend AI architecture to frontend user experiences — and can align engineering decisions with measurable product outcomes.
You will begin as a strong individual contributor, independently architecting and deploying AI-powered solutions. As the product portfolio scales, you will lead a distributed team across India and Australia, acting as a System Integrator to align engineering, data, and AI contributions into cohesive production systems.
Example Project
Design and deploy a multi-agent AI system to automate critical stages of a company’s sales cycle, including:
- Generating client proposals using historical SharePoint data and CRM insights
- Summarizing meeting transcripts
- Drafting follow-up communications
- Feeding structured insights into dashboards and workflow tools
The solution will combine RAG pipelines, LLM reasoning, and React-based interfaces to deliver measurable productivity gains.
Key Responsibilities
- Architect and implement AI workflows using LLMs, vector databases, and automation frameworks
- Act as a System Integrator, coordinating deliverables across distributed engineering and AI teams
- Develop frontend interfaces using React/JavaScript to enable seamless human-AI collaboration
- Design APIs and microservices integrating AI systems with enterprise platforms (SharePoint, Teams, Databricks, Azure)
- Drive architecture decisions balancing scalability, performance, and security
- Collaborate with product managers, clients, and data teams to translate business use cases into production-ready systems
- Mentor junior engineers and evolve into a broader leadership role as the team grows
Ideal Candidate Profile
Experience Requirements
- 5+ years in full-stack development (Python backend + React/JavaScript frontend)
- Strong experience in API and microservice integration
- 2+ years leading technical teams and coordinating distributed engineering efforts
- 1+ year of hands-on AI project experience (LLMs, Transformers, LangChain, OpenAI/Azure AI frameworks)
- Prior experience in B2B SaaS environments, particularly in AI, automation, or enterprise productivity solutions
Technical Expertise
- Designing and implementing AI workflows including RAG pipelines, vector databases, and prompt orchestration
- Ensuring backend and AI systems are scalable, reliable, observable, and secure
- Familiarity with enterprise integrations (SharePoint, Teams, Databricks, Azure)
- Experience building production-grade AI systems within enterprise SaaS ecosystems
We are building an advanced, AI-driven multi-agent software system designed to revolutionize task automation and code generation. This is a futuristic AI platform capable of:
✅ Real-time self-coding based on tasks
✅ Autonomous multi-agent collaboration
✅ AI-powered decision-making
✅ Cross-platform compatibility (Desktop, Web, Mobile)
We are hiring a highly skilled **AI Engineer & Full-Stack Developer** based in India, with a strong background in AI/ML, multi-agent architecture, and scalable, production-grade software development.
### Responsibilities:
- Build and maintain a multi-agent AI system (AutoGPT, BabyAGI, MetaGPT concepts)
- Integrate large language models (GPT-4o, Claude, open-source LLMs)
- Develop full-stack components (Backend: Python, FastAPI/Flask, Frontend: React/Next.js)
- Work on real-time task execution pipelines
- Build cross-platform apps using Electron or Flutter
- Implement Redis, Vector databases, scalable APIs
- Guide the architecture of autonomous, self-coding AI systems
### Must-Have Skills:
- Python (advanced, AI applications)
- AI/ML experience, including multi-agent orchestration
- LLM integration knowledge
- Full-stack development: React or Next.js
- Redis, Vector Databases (e.g., Pinecone, FAISS)
- Real-time applications (websockets, event-driven)
- Cloud deployment (AWS, GCP)
### Good to Have:
- Experience with code-generation AI models (Codex, GPT-4o coding abilities)
- Microservices and secure system design
- Knowledge of AI for workflow automation and productivity tools
Join us to work on cutting-edge AI technology that builds the future of autonomous software.
Job Title: Full Stack AI Engineer
Location: Remote/Hyderabad
Experience Level: 3-5
Salary Range: 12-18LPA
Application Link:https://beyond.ciltriq.com/apply/BUILD
Description:
Join a team building AI-powered systems that solve complex business problems and automate operational workflows across document processing, voice agents, enterprise integrations, workflow automation, and multi-agent systems.
Strong full-stack foundations: frontend state management, asynchronous user experiences and performance; backend API design, authentication, data modelling, databases, queues and distributed systems.
Strong coding ability in Python and JavaScript or TypeScript, with practical experience in modern frontend frameworks and backend services.
Requirements:
- Design and build complete systems: frontend applications, backend services, APIs, databases, data pipelines and integrations with customer systems.
- Build multi-agent workflows with clear agent responsibilities, tool access, shared state, context management, routing, handoffs and coordination across sequential and parallel tasks.
- Make agent execution dependable through durable state, checkpoints, retries, timeouts, idempotency, recovery and human approval or review where needed.
- Deliver document-processing pipelines, voice agents and retrieval-based AI applications, connecting model outputs to useful actions in real business workflows.
- Own quality in production: automated tests, AI evaluations, guardrails, observability, access controls, deployments, incident response and clear documentation.
- Choose where AI adds value and where deterministic software is the better fit. Balance accuracy, latency, cost, security and maintainability.
- Improve reusable engineering foundations, review code and help other engineers grow as the team expands.
- A solid understanding of tool calling, structured outputs, retrieval, context and memory management, model selection and evaluation.
- Practical cloud and deployment experience, including containers, CI/CD, secrets management, logging, monitoring and production debugging.
- Ability to reason from first principles, investigate failures across system boundaries and communicate technical decisions clearly to customers and teammates.
- Useful additional experience: Document AI and OCR, real-time voice systems, enterprise integrations, agent protocols such as MCP, and orchestration frameworks.
- Useful additional experience: Mentoring engineers or building reusable platforms.
About Us
We believe the future of software development is AI-native — where engineers operate at a higher level of abstraction and quality remains non-negotiable.
Incubyte is a software craft consultancy where the “how” of building software matters as much as the “what”.
We partner with companies of all sizes, from helping enterprises build, scale, and modernize to early-stage founders bring their ideas to life.
Our engineers operate in an AI-native development model, using AI as a collaborator across the SDLC to accelerate development while upholding the discipline of software craftsmanship. Guided by Software Craftsmanship and Extreme Programming practices, we build reliable, maintainable, and scalable systems with speed, without compromising quality. If this way of building software resonates with you, we’d like to talk.
Our Guiding Principles
These principles define how we work at Incubyte. They are non-negotiable.
Relentless Pursuit of Quality with Pragmatism
We build high-quality systems without losing sight of delivery.
Extreme Ownership
We take responsibility end-to-end for decisions, execution, and outcomes.
Proactive Collaboration
We collaborate closely, challenge each other, and solve problems together.
Active Pursuit of Mastery
We continuously improve our craft and raise our bar.
Invite, Give, and Act on Feedback
We seek, give, and act on feedback to get better every day.
Ensuring Client Success
We act as trusted partners and focus on real outcomes, not just output.
Job Description
This is a remote position.
Experience Level
This role is ideal for engineers with total 5+ years of experience with a proven track record of shipping complex projects successfully.
An experienced individual contributor and leader who thrives in large, complex projects with widespread impact.
What You’ll Do as a Software Craftsperson
- Design and build high-quality, maintainable systems using disciplined engineering practices such as TDD, continuous refactoring, and pair programming
- Operate in an AI-native development model, using AI as a collaborator to explore architecture and design, accelerate development, and continuously improve systems while applying strong judgment to ensure that speed never compromises quality.
- Take end-to-end ownership of outcomes from problem understanding and system design to implementation, deployment, and operation in production
- Make thoughtful design decisions that balance simplicity, scalability, and long-term maintainability in real-world systems
- Maintain a high bar for engineering quality through rigorous testing, code reviews, and continuous feedback
- Investigate and resolve production issues, and implement systemic improvements to prevent recurrence
- Work directly with clients, navigate ambiguity, and translate business problems into well-designed technical solutions
- Contribute to improving team practices, tooling, and systems to raise the overall quality and effectiveness of engineering
Requirements
What You’ll Bring
- 5+ years of experience building high-quality, production systems (flexible based on demonstrated capability)
- Strong fundamentals in software engineering, including object-oriented design, system design, and testing practices such as TDD
- Demonstrated ability to build simple, maintainable, and scalable systems with a focus on long-term reliability
- Proficiency in one or more modern technologies Python, React, AI, JavaScript, or TypeScript, with the ability to learn new technologies quickly
- Deep experience working with Git in collaborative environments, including managing shared codebases, conducting code reviews, and maintaining a high bar for quality
- Ability to operate effectively in an AI-native workflow using AI as a collaborator to explore solutions and accelerate development, while applying strong judgment to ensure correctness, quality, and maintainability
- Clear thinking and strong problem-solving ability, with the capacity to break down complex problems into simple, well-structured solutions
- A strong sense of ownership — you take responsibility for outcomes, care deeply about quality, and are not comfortable shipping work that does not meet your standards.
Benefits
Life at Incubyte
We are a remote-first company with structured flexibility. Teams commit to shared rhythms during core hours, ensuring smooth collaboration while maintaining autonomy. Twice a year, we come together in person for a co-working sprint and once a year for a retreat - with all travel expenses covered.
Our environment is built for crafters: pairing, refactoring, experimenting with AI, and pushing the boundaries of software excellence. We are all lifelong learners, and our work is our passion.
Benefits
- Dedicated learning & development budget.
- Sponsorship for conference talks.
- Comprehensive medical & term insurance.
- Employee-friendly leave policies.
- Home Office fund
- Medical Insurance
Position: Senior/Lead Full Stack Engineer – Gen AI / Agentic AI
Experience: 7+ Years
Employment: Permanent Position
Location: Banglore / Hyderabad
Job Summary
We are looking for a Senior/Lead Full Stack Engineer – Gen AI / Agentic AI with strong hands-on experience in Python, React.js, MongoDB, Java/Spring Boot and Generative AI/Agentic AI.
The candidate should have experience designing and developing scalable enterprise applications and implementing production-grade LLM, RAG, AI Agent and multi-agent solutions.
Key Responsibilities
- Design, develop and maintain scalable full-stack applications using Python, React.js, MongoDB and Java/Spring Boot.
- Build production-grade Generative AI and Agentic AI applications using LLMs and modern AI frameworks.
- Develop RAG pipelines, AI agents, tool calling, memory management, planning and agent orchestration.
- Work with LangChain, LangGraph, MCP, vector databases and semantic search.
- Develop Python-based APIs, microservices and asynchronous applications using FastAPI/Flask/Django.
- Build REST APIs and event-driven microservices with focus on scalability, performance and resilience.
- Integrate LLMs, embeddings, vector stores and external enterprise tools/services.
- Implement prompt engineering, LLM evaluation, guardrails and AI observability.
- Develop responsive front-end applications using React.js.
- Work with MongoDB, SQL and hybrid data models.
- Implement CI/CD pipelines and support cloud/OCP deployments.
- Follow secure coding, testing, code quality and performance best practices.
- Participate in architecture, technical design, code reviews and mentoring of team members.
- Collaborate with business and technical stakeholders to translate requirements into scalable solutions.
Mandatory Skills
- Python
- React.js
- Gen AI / Agentic AI
- RAG + LLM
- LangChain / LangGraph
- MongoDB
- Java + Spring Boot
- REST APIs / Microservices
- Vector Databases / Embeddings
- MCP / AI Agent orchestration
Good to Have
- FastAPI / Flask / Django
- Kafka / Solace
- Docker / Kubernetes
- AWS / Azure / GCP / OCP
- CI/CD – Jenkins / GitHub Actions
- LLMOps / AI evaluation / observability
- ELK / Grafana / Splunk / AppDynamics
- SQL / NoSQL
- Agile/Scrum
About the role
We are building AI systems that read, understand and act on real business documents, bank statements, financial reports, policy documents and forms and putting them into production where accuracy and cost both matters.
This is not a research role and it is not a prompt-writing role. You will own features end to end: pick and deploy open-source models, build the pipelines around them, measure whether they actually work on our documents, drive the cost per document down, and keep the whole thing running in production.
You will work closely with the engineering and product teams, and your work will be directly used by business users from day one.
What you will do
Deploy and evaluate open-source models
- Select, deploy and benchmark open-source LLMs and vision-language models for specific, narrow use cases not general chat.
- Build evaluation sets from real documents and define what "good" means numerically (field-level accuracy, extraction recall, hallucination rate) before shipping.
- Run structured comparisons between models and approaches, and write up the trade-offs so the team can make a decision.
- Apply quantization, batching and other optimizations to fit models into a sensible GPU budget.
Build and optimize AI orchestration
- Design multi-step pipelines that combine deterministic code, ML models and LLM calls and know when not to use an LLM.
- Optimize for latency, cost and reliability: caching, batching, request routing, fallback tiers, retries and graceful degradation.
- Instrument pipelines so failures are visible and traceable rather than silent.
Ship to production
- Package models and services with Docker, expose them behind clean APIs, and deploy them to our GPU and CPU infrastructure.
- Handle the unglamorous production concerns: cold starts, timeouts, concurrency limits, versioning, rollback and monitoring.
- Own on-call-style responsibility for the AI features you build, including cost tracking.
Must-have skills
Programming & engineering
- Strong Python: type hints, async/await, dataclasses/Pydantic, clean module design, testing.
- REST API development with FastAPI (or Flask/Django with a willingness to move to FastAPI).
- Git, code review discipline, and the ability to write code someone else can maintain.
- Comfortable in Linux and on the command line.
Machine learning fundamentals
- Working knowledge of PyTorch and the Hugging Face ecosystem (transformers, tokenizers, accelerate).
- Understanding of inference-time concepts: tokenization, context windows, batching, precision (FP16/BF16/INT8), memory footprint.
- Ability to read a model card and a paper well enough to judge whether a model fits a use case.
Document processing
- Hands-on experience with at least two of: pypdfium2, PyMuPDF, pdfplumber, pdfminer.six, Docling, Unstructured, Surya, DocTR, LayoutLM family.
- Practical OCR experience (Tesseract, PaddleOCR, or a cloud OCR) and an understanding of when OCR is the wrong tool.
- Experience extracting tables from PDFs and dealing with merged cells, multi-line rows, and inconsistent column layouts.
Strongly preferred
You will be a much stronger candidate with any of these. We do not expect all of them.
Model serving & optimization
- vLLM, TGI, Ollama, llama.cpp, or Triton Inference Server.
- Quantization formats and tooling: GGUF, AWQ, GPTQ, bitsandbytes, ONNX Runtime, INT8 export.
- Serverless GPU platforms: Modal, RunPod, Replicate, Baseten including cold-start and container-lifecycle management.
- LoRA / QLoRA fine-tuning with PEFT for narrow, task-specific improvements.
Vision-language models
- Practical use of open VLMs: Qwen2.5-VL, InternVL, Granite Vision, Molmo, Phi-Vision, or similar.
- Awareness of where VLMs hallucinate especially on numeric and financial content and patterns for constraining them (using the model for layout only, sourcing values from the text layer, constrained decoding).
Orchestration & pipelines
- Workflow orchestration: Dagster, Airflow, Prefect, or Temporal.
- Async job patterns: Celery, RQ, or platform-native spawn/poll patterns.
- LLM orchestration frameworks (LangGraph, LlamaIndex, Haystack) with the judgement to know when plain Python is a better answer.
- Structured output enforcement: Instructor, Outlines, XGrammar, JSON schema / tool-use modes.
Evaluation & observability
- Building golden datasets and regression suites for extraction tasks.
- Eval tooling: promptfoo, DeepEval, Ragas, or in-house harnesses.
- LLM tracing and monitoring: Langfuse, Arize Phoenix, LangSmith, OpenTelemetry.
Nice extras
- Rule engines and policy evaluation (Open Policy Agent / Rego, Drools, rule-engine).
- Experience in fintech, lending, insurance or accounting documents.
- Handling of PII and data-security practices in document pipelines.
- Contributions to open-source ML or document-processing projects.
Why join us
- Real production ownership from month one your work goes to actual users, not a demo.
- Genuinely hard technical problems in document AI, not wrappers over an API.
- Small team, short decision cycles, direct access to leadership.
- Budget and freedom to evaluate and adopt new open-source models as they land.
To apply: send your CV along with a short note on one AI system you have taken to production what it did, what the accuracy was, and what broke.
AI Engineer
LLMs, Agents & AI Services
📍 Mumbai (On-site) | Full-time | 2-4 years
About the Role:
Unico Connect is an AI-first technology partner that builds custom mobile, web, and AI products for clients across multiple geographies.
AI is core to how we design, deliver, and scale software for our customers.
We are hiring an AI Engineer for a dedicated client engagement building a complex production AI platform, working on the AI capabilities and agentic features at the core of the product.
The mandatory requirement for this role is at least one AI feature personally shipped to production for real users, with operational ownership.
The role suits someone who thinks quickly on solutioning, can take an ambiguous problem to a working prototype in days, and has the discipline to carry it through to production with predictable economics.
You will work alongside the Senior AI Engineer and the wider pod, with ownership of parts of the AI surface area of the product.
Responsibilities:
Solutioning and POCs
Translate ambiguous customer problems into working POCs at speed.
Pick the right model, framework, and architecture, and demonstrate value early before scaling investment.
LLM Application Development
Build AI features and services using LLM APIs from OpenAI, Anthropic, Google, and self-hosted open-weight models (Llama, Qwen, Mistral).
Choose the right model per use case based on cost, latency, capability, and context-window trade-offs.
Agentic System Design
Design and implement agentic workflows using LangGraph, CrewAI, AutoGen, LlamaIndex Agents, or custom orchestration.
Cover tool use, planning, memory, and multi-step reasoning appropriate to the problem.
API and Service Development
Build production AI services and APIs using Python and FastAPI.
Handle streaming responses, async processing, structured outputs, retries, and graceful degradation when models or tools fail.
Retrieval and Tool Integration
Implement RAG pipelines with vector databases (Pinecone, Weaviate, Qdrant, pgvector, Chroma), embeddings, chunking strategies, hybrid search, and reranking.
Integrate external tools, internal APIs, and document sources through tool-calling and MCP-style patterns.
Cost Analysis and Unit Economics
Model the per-request and per-user cost of every AI feature before it ships.
Track token usage, prompt caching, batching, and model-routing strategies.
Drive measurable improvements in unit economics.
Production Hardening
Add observability and tracing (LangSmith, Langfuse, OpenTelemetry), guardrails, content safety checks, prompt injection defences, and fallback behaviour.
Prompt Engineering and Evaluation
Design, test, and iterate prompts with measured outcomes.
Build evaluation harnesses for accuracy, hallucination, latency, and cost.
Run benchmarks across models and prompt variants before locking in a design.
Requirements:
AI Feature Shipped to Production (Mandatory)
Must have personally built and shipped at least one AI feature that runs in production for real users, with operational ownership.
POCs, internal demos, and one-off scripts do not qualify.
2 to 4 Years of Professional Software or AI Engineering Experience
With at least one production AI feature owned end to end.
Strong Python Proficiency and API Development with FastAPI
Comfort with type hints, async, packaging, testing, streaming responses, and authentication.
Production-grade Python, not notebook-only code.
Hands-on Depth Across the LLM and Agent Stack
Working experience with at least two of OpenAI, Anthropic Claude, Google Gemini, or self-hosted open-weight models (vLLM, Ollama, Together, Replicate).
Working familiarity with at least one agent framework (LangGraph, CrewAI, AutoGen, LlamaIndex Agents) or hand-rolled equivalent.
Working knowledge of RAG, embeddings, and vector databases (Pinecone, Weaviate, Qdrant, pgvector, Chroma).
Solutioning Speed and POC Velocity
Demonstrated ability to move from a fuzzy problem to a working prototype in days.
Strong instinct for what to build first, what to defer, and what to throw away.
Cost Discipline for Production AI
Ability to calculate, monitor, and optimise the cost of LLM APIs, tokens, embeddings, vector store usage, and infrastructure.
Treats unit economics as a first-class concern.
AWS Familiarity
Working knowledge of EC2, S3, IAM, and at least one of Bedrock, SageMaker, or equivalent.
Comfortable in a Fast-Moving Environment
Self-directed, comfortable with ambiguity, takes ownership without being asked, and ships under shifting priorities.
Strong Written and Spoken English Communication
Able to explain trade-offs to non-AI engineers, designers, product managers, and clients in plain language.
Nice to Have
- fine-tuning or LoRA, QLoRA, PEFT exposure
- MCP server authoring
- eval framework experience (LangSmith, Promptfoo, Ragas, DeepEval)
- open-source AI contributions
- multi-modal models (vision, audio)






