Applied AI & Data Engineer at Morganstek · Remote only · 6 - 10 years · ₹24L - ₹36L / yr · Remote only · Posted 29 Sep 2026

The Role
We’re looking for a Senior Applied AI & Data Engineer to become our first dedicated AI and data engineer.
You’ll build conversational AI experiences across web, mobile, and in-store channels while developing the data foundation behind them. You’ll make key technical decisions and own your work through to production.
What You’ll Do
• Build AI assistants using tool calling to work with real product, search, and order systems
• Design guardrails and evaluation sets to ensure AI responses are accurate and safe
• Build real-time and voice-enabled AI experiences
• Improve product data quality through AI-assisted enrichment and review workflows
• Build data pipelines, analytics, and personalisation systems
• Work closely with web and mobile developers and help guide technical implementation
What You’ll Need
• 6+ years of experience building and running production backend systems
• Strong Python skills, plus experience with JavaScript/TypeScript backends
• Experience shipping at least one LLM-powered feature to real users
• Experience with search and relevance
• Experience building data pipelines and analytics stores
• Comfortable deploying and monitoring services on a major cloud platform
• Strong communication skills and the ability to work independently
Nice to Have
• Experience with speech or voice AI
• E-commerce or retail technology experience
• Experience building multilingual products

Similar jobs (10)
Location: Hyderabad, India (home base), deployed at client sites in India. Occasional Middle East exposure possible.
About the Role
You will work as a senior AI engineer who embeds inside a customer's business. Your job is to learn how the business makes money, find the highest value problem, and build a working system that solves it.
Four behaviors define this role:
- Go where the work happens. You work onsite with the customer, in the room where decisions are made.
- Show working software early. You build a prototype in days, not a document in weeks.
- One person owns the outcome. You are the single point of accountability for the result.
- Stay after go-live. You keep running and improving the system after launch.
You are the single point of accountability. You are not a solo builder. A full KnackLabs engineering team in Hyderabad builds and runs the production systems behind you.
This role involves extended onsite deployments at client locations in other cities, sometimes up to six months at a stretch. Please apply only if you are ready for this way of working.
What you'll own
- Discovery - Learn how the customer makes money. Find the highest value problem to solve first.
- The prototype - Build a working prototype fast, using real or sample data, to prove the idea.
- The roadmap - Decide what to build, in what order, and set clear success measures tied to business outcomes.
- The build - Design and ship the production system with the Hyderabad engineering team. This includes data integration, agents, retrieval, and evaluations.
- The client relationship - Be the trusted technical contact for the customer, from engineers to senior leaders.
- Go live and after - Deploy the system, watch how it performs, fix problems, and improve it over time.
- Feedback to the product - Share what you learn in the field so the vendor's product and our internal tools get better.
What we are looking for
- Around 4 or more years of software engineering experience, including customer-facing or client delivery work.
- Strong programming skills in Python. Working knowledge of TypeScript or JavaScript.
- A full-stack development experience with strength in backend technologies.
- Production experience with large language models, including prompt engineering and agent development.
- You build with AI coding tools like Claude Code or Codex as your default way of working, and you have shipped real apps or agents this way.
- Experience building retrieval-augmented generation (RAG) systems: chunking, embeddings, vector databases, retrieval, and reranking.
- Experience building and deploying AI systems.
- Experience integrating with APIs and enterprise systems.
- Experience with at least one cloud platform (AWS, Azure, or GCP).
- Clear communication. You can explain a technical choice to an engineer and to a business leader.
- High ownership and comfort with ambiguity. You can take an unclear problem and turn it into a plan.
- Willingness to work onsite at client locations in India for extended periods, and to travel as the work needs.
Nice to have
- Experience with on-premises or private cloud (VPC) deployments.
- Experience with observability and tracing tools such as LangSmith or Braintrust.
- Experience with data engineering and pipelines.
- A history of side projects, open source contributions, or products you shipped end-to-end.
- Experience in embedded or forward-deployed roles before.
- Experience working at a consulting or professional services firm in a client-facing delivery role.
Stack and tools
- Languages: Python and TypeScript.
- Models: Claude and other frontier or open-source models, chosen to fit the customer.
- AI patterns: RAG, agents, prompt engineering, and evaluations.
- Vector and retrieval: vector databases and retrieval pipelines.
- Cloud: AWS, Azure, or GCP, on public or private cloud.
- Integration: REST APIs and enterprise system connectors.
Location: Hyderabad, India (home base), deployed at client sites in India. Occasional Middle East exposure possible.
About the Role
You will work as a senior AI consultant who embeds inside a customer's business. Your job is to learn how the business makes money, find the highest value problem, and build a working system that solves it.
You will not hand over a document and walk away. You will show working software early, own the roadmap, own the client relationship, and stay after go-live to run and improve the system.
Four behaviors define this role:
- Go where the work happens. You work onsite with the customer, in the room where decisions are made.
- Show working software early. You build a prototype in days, not a document in weeks.
- One person owns the outcome. You are the single point of accountability for the result.
- Stay after go-live. You keep running and improving the system after launch.
You are the single point of accountability. You are not a solo builder. A full KnackLabs engineering team in Hyderabad builds and runs the production systems behind you.
This role involves extended onsite deployments at client locations in other cities, sometimes up to six months at a stretch. Please apply only if you are ready for this way of working.
What you'll own
- Discovery - Learn how the customer makes money. Find the highest value problem to solve first.
- The prototype - Build a working prototype fast, using real or sample data, to prove the idea.
- The roadmap - Decide what to build, in what order, and set clear success measures tied to business outcomes.
- The build - Design and ship the production system with the Hyderabad engineering team. This includes data integration, agents, retrieval, and evaluations.
- The client relationship - Be the trusted technical contact for the customer, from engineers to senior leaders.
- Go live and after - Deploy the system, watch how it performs, fix problems, and improve it over time.
- Feedback to the product - Share what you learn in the field so our platform and internal tools get better.
What we are looking for
- Around 7 or more years of software engineering experience, including customer-facing or client delivery work.
- Experience working at a consulting or professional services firm in a client-facing delivery role.
- A full-stack development experience with strength in backend technologies.
- Strong programming skills in Python. Working knowledge of TypeScript or JavaScript.
- Production experience with large language models, including prompt engineering and agent development.
- You build with AI coding tools like Claude Code as your default way of working. You have built real apps and agents this way, not just used it for document generation or review.
- Experience building retrieval-augmented generation (RAG) systems: chunking, embeddings, vector databases, retrieval, and reranking.
- Experience building and deploying AI systems.
- Experience integrating with APIs and enterprise systems.
- Experience with at least one cloud platform (AWS, Azure, or GCP).
- Experience building evaluations to measure accuracy, safety, latency, and cost.
- Clear communication. You can explain a technical choice to an engineer and to a business leader.
- High ownership and comfort with ambiguity. You can take an unclear problem and turn it into a plan.
- Willingness to work onsite at client locations in India for extended periods, and to travel as the work needs.
Nice to have
- Experience deploying AI systems in regulated industries such as insurance, banking, or the public sector.
- Experience with on-premises or private cloud (VPC) deployments.
- Experience with observability and tracing tools such as LangSmith or Braintrust.
- Experience with data engineering and pipelines.
- A history of side projects, open source contributions, or products you shipped end-to-end.
- Experience in embedded or forward-deployed roles before.
Stack and tools
- Languages: Python and TypeScript.
- Models: Claude and other frontier or open source models, chosen to fit the customer.
- AI patterns: RAG, agents, prompt engineering, and evaluations.
- Vector and retrieval: vector databases and retrieval pipelines.
- Cloud: AWS, Azure, or GCP, on public or private cloud.
- Integration: REST APIs and enterprise system connectors.
Product Engineer — Role Summary
We are looking for a Product Engineer to build user-facing products that combine AI capabilities with practical applications. You will work at the intersection of software engineering and AI, developing autonomous, agent-driven systems that solve complex educational and research problems.
Key Responsibilities:
- Product Development: Design, build, and deploy production-ready applications powered by LLMs and AI agents.
- Data Engineering: Build scalable ETL/ELT pipelines to process structured and unstructured data, including text and audio, for RAG and model fine-tuning.
- Agentic Workflows: Develop multi-step AI agents with tool calling, APIs, databases, search, reasoning, and memory.
- Rapid Prototyping: Turn ideas and research concepts into interactive, production-ready applications.
- AI Integration: Use frameworks such as LangChain, LlamaIndex, AutoGen, or custom orchestrators to integrate AI into scalable systems.
- User Experience: Transform raw AI outputs into reliable, intuitive, and responsive user experiences.
- Collaboration: Work closely with ML researchers and data engineers to integrate custom and fine-tuned models.
- Observability: Monitor agent behavior, manage edge cases, reduce hallucinations, and improve reliability in production.
The ideal candidate combines strong software engineering, AI/LLM expertise, data engineering, and product thinking, with the ability to take an AI concept from prototype to production.
About LeadSquared
LeadSquared is a leading sales execution and marketing automation platform trusted by 2,000+ businesses globally, including healthcare, education, financial services, and real estate. Headquartered in Bengaluru with offices across the US, UK, UAE, and Southeast Asia, we empower sales teams to close faster, smarter, and at scale.
Our AI team is at the forefront of integrating cutting-edge large language model capabilities into enterprise workflows — building intelligent agents, copilots, and automation systems that redefine how businesses operate.
Role Overview
We are looking for a Senior AI Engineer with hands-on experience building LLM-powered agents and agentic AI systems. You will design, develop, and deploy autonomous AI pipelines that solve complex, multi-step business problems — from lead qualification and follow-up automation to intelligent CRM workflows and beyond.
This role is ideal for someone who is deeply excited about the frontier of AI, can move fast, and wants their work to directly impact millions of sales professionals worldwide.
Key Responsibilities
•
Design and build LLM-powered agentic systems using frameworks such as LangChain, LlamaIndex, AutoGen, or CrewAI to automate complex, multi-step workflows.
•
Develop and maintain Retrieval-Augmented Generation (RAG) pipelines with vector databases (Pinecone, Weaviate, Chroma, pgvector) for domain-specific knowledge grounding.
•
Build and integrate tool-use and function-calling capabilities into AI agents, enabling dynamic interaction with internal APIs, databases, and third-party services.
•
Implement prompt engineering strategies including chain-of-thought, few-shot prompting, and structured output parsing to ensure reliable agent behavior.
•
Design evaluation frameworks and observability pipelines (LangSmith, Helicone, custom metrics) to monitor agent performance, accuracy, and cost.
•
Collaborate with product, sales, and domain teams to translate business requirements into AI-driven solutions and features.
•
Optimize LLM inference for latency and cost using techniques like caching, model distillation, quantization, and batching.
•
Stay current with the rapidly evolving LLM ecosystem and proactively propose improvements and new approaches.
•
Contribute to internal best practices, documentation, and knowledge-sharing across the engineering org.
Required Qualifications
Experience
•
2–4 years of professional software engineering experience, with at least 1–2 years focused on LLM/AI systems.
•
Proven experience shipping LLM-based products or agentic AI systems into production environments.
Technical Skills
•
Strong proficiency in Python and familiarity with async programming patterns for AI pipelines.
•
Hands-on experience with LLM APIs: OpenAI (GPT-4o), Anthropic (Claude), Google (Gemini), or open-source models (Llama, Mistral).
•
Experience with agentic frameworks: LangChain, LangGraph, LlamaIndex, AutoGen, CrewAI, or similar.
•
Solid understanding of RAG architectures, embedding models, and semantic search.
•
Experience with vector databases and similarity search infrastructure.
•
Knowledge of REST APIs, microservices architecture, and containerization (Docker/Kubernetes).
Problem-Solving & Mindset
•
Strong ability to decompose ambiguous, open-ended problems into structured AI system designs.
•
Experience with prompt debugging, LLM evaluation, and iterative refinement workflows.
•
Ability to balance research exploration with engineering pragmatism to ship reliable systems.
Preferred Qualifications
•
Experience with multi-agent orchestration and agent memory systems (short-term and long-term).
•
Familiarity with fine-tuning or RLHF workflows for domain adaptation.
•
Background in NLP, information retrieval, or conversational AI.
•
Prior experience in B2B SaaS or CRM domain is a plus.
•
Contributions to open-source AI/ML projects or published research/blogs.
•
Experience with cloud platforms: AWS, GCP, or Azure — particularly AI/ML services
We are building an advanced, AI-driven multi-agent software system designed to revolutionize task automation and code generation. This is a futuristic AI platform capable of:
✅ Real-time self-coding based on tasks
✅ Autonomous multi-agent collaboration
✅ AI-powered decision-making
✅ Cross-platform compatibility (Desktop, Web, Mobile)
We are hiring a highly skilled **AI Engineer & Full-Stack Developer** based in India, with a strong background in AI/ML, multi-agent architecture, and scalable, production-grade software development.
### Responsibilities:
- Build and maintain a multi-agent AI system (AutoGPT, BabyAGI, MetaGPT concepts)
- Integrate large language models (GPT-4o, Claude, open-source LLMs)
- Develop full-stack components (Backend: Python, FastAPI/Flask, Frontend: React/Next.js)
- Work on real-time task execution pipelines
- Build cross-platform apps using Electron or Flutter
- Implement Redis, Vector databases, scalable APIs
- Guide the architecture of autonomous, self-coding AI systems
### Must-Have Skills:
- Python (advanced, AI applications)
- AI/ML experience, including multi-agent orchestration
- LLM integration knowledge
- Full-stack development: React or Next.js
- Redis, Vector Databases (e.g., Pinecone, FAISS)
- Real-time applications (websockets, event-driven)
- Cloud deployment (AWS, GCP)
### Good to Have:
- Experience with code-generation AI models (Codex, GPT-4o coding abilities)
- Microservices and secure system design
- Knowledge of AI for workflow automation and productivity tools
Join us to work on cutting-edge AI technology that builds the future of autonomous software.
Role: AI Developer
Experience: 3–4 Years
Employment Type: Full-Time
Location: Goregaon, Mumbai
About the Role
We are looking for an experienced AI Developer with 3–4 years of software development experience and strong hands-on exposure to Generative AI, AI Agents, Copilots, and AI-powered application development.
The candidate will be responsible for building production-ready AI solutions, developing agentic workflows, modernizing legacy applications, and integrating LLM capabilities into enterprise applications.
Key Responsibilities
- Design, develop, and deploy AI Agents and agentic workflows for enterprise use cases.
- Build AI Copilots and LLM-powered applications using modern AI frameworks and APIs.
- Develop RAG-based applications using embeddings, vector databases, and enterprise data.
- Work on legacy application migration and modernization, leveraging AI-assisted development and code transformation techniques.
- Analyze legacy codebases and design strategies for AI-driven migration, refactoring, and modernization.
- Integrate LLMs with enterprise applications, APIs, databases, and third-party systems.
- Implement tool calling, function calling, multi-agent workflows, and workflow automation.
- Perform prompt engineering, context optimization, model evaluation, and AI application testing.
- Take ownership of AI solutions from POC and prototyping through production deployment.
- Collaborate with product managers, architects, and engineering teams to convert business requirements into scalable AI solutions.
- Stay updated with emerging technologies in Generative AI, Agentic AI, LLMs, and AI-assisted software development.
Required Skills
- 3–4 years of professional software development experience.
- Strong proficiency in Python and/or JavaScript/TypeScript.
- Hands-on experience developing Generative AI / LLM-based applications.
- Strong understanding of AI Agents, RAG, Prompt Engineering, LLM APIs, and embeddings.
- Experience with frameworks such as LangChain, LangGraph, Semantic Kernel, AutoGen, or equivalent.
- Experience working with REST APIs, databases, Git, and cloud environments.
- Hands-on experience with vector databases such as Pinecone, Weaviate, Chroma, FAISS, or equivalent.
- Good understanding of software architecture, debugging, testing, and deployment practices.
Good to Have
- Experience with Microsoft Copilot / Copilot Studio.
- Experience working with Claude, OpenAI, Gemini, Azure OpenAI, or open-source LLMs.
- Experience in legacy application migration, modernization, or code conversion.
- Knowledge of Azure AI / AWS / Google Cloud AI services.
- Experience with MCP, multi-agent systems, tool calling, and AI orchestration.
- Experience building enterprise-grade AI solutions with focus on security, scalability, and performance.
Job Title: Full Stack AI Engineer
Location: Remote/Hyderabad
Experience Level: 3-5
Salary Range: 12-18LPA
Application Link:https://beyond.ciltriq.com/apply/BUILD
Description:
Join a team building AI-powered systems that solve complex business problems and automate operational workflows across document processing, voice agents, enterprise integrations, workflow automation, and multi-agent systems.
Strong full-stack foundations: frontend state management, asynchronous user experiences and performance; backend API design, authentication, data modelling, databases, queues and distributed systems.
Strong coding ability in Python and JavaScript or TypeScript, with practical experience in modern frontend frameworks and backend services.
Requirements:
- Design and build complete systems: frontend applications, backend services, APIs, databases, data pipelines and integrations with customer systems.
- Build multi-agent workflows with clear agent responsibilities, tool access, shared state, context management, routing, handoffs and coordination across sequential and parallel tasks.
- Make agent execution dependable through durable state, checkpoints, retries, timeouts, idempotency, recovery and human approval or review where needed.
- Deliver document-processing pipelines, voice agents and retrieval-based AI applications, connecting model outputs to useful actions in real business workflows.
- Own quality in production: automated tests, AI evaluations, guardrails, observability, access controls, deployments, incident response and clear documentation.
- Choose where AI adds value and where deterministic software is the better fit. Balance accuracy, latency, cost, security and maintainability.
- Improve reusable engineering foundations, review code and help other engineers grow as the team expands.
- A solid understanding of tool calling, structured outputs, retrieval, context and memory management, model selection and evaluation.
- Practical cloud and deployment experience, including containers, CI/CD, secrets management, logging, monitoring and production debugging.
- Ability to reason from first principles, investigate failures across system boundaries and communicate technical decisions clearly to customers and teammates.
- Useful additional experience: Document AI and OCR, real-time voice systems, enterprise integrations, agent protocols such as MCP, and orchestration frameworks.
- Useful additional experience: Mentoring engineers or building reusable platforms.
Position Overview
We are seeking a versatile Senior Full Stack & AI Agent Developer to architect, build, and
maintain end-to-end software solutions spanning web platforms, desktop applications, and
autonomous AI agents capable of interacting with and controlling these software systems.
The ideal candidate will bridge traditional engineering software with cutting-edge artificial
intelligence to automate data processing and enhance operational decision-making. While
not strictly required, a background or strong interest in the energy sector—specifically
drilling and completion operations—is highly desirable.
Key Responsibilities
• Full Stack Development: Design, develop, and deploy robust web applications and
native desktop software utilized by engineering and operational teams.
• AI Agent Engineering: Build, train, and integrate autonomous AI agents and LLM-
driven workflows capable of interpreting data, executing commands, and safely
controlling desktop and web-based software.
• Workflow Automation: Translate complex workflows into intuitive software features
and autonomous agent actions, minimizing manual data entry and operational
bottlenecks.
• Data Integration: Handle high-frequency data streams and integrate them seamlessly
into user interfaces and backend AI models.
• Architecture & Scalability: Ensure high performance, security, and scalability across
cloud infrastructure (AWS/Azure), local desktop environments, and potential edge
computing setups.
• Cross-Functional Collaboration: Work closely with domain experts and end-users to
translate field challenges into technical product requirements.
Required Qualifications & Experience
• Experience: Minimum of 5 years of professional software development experience,
with a proven track record of delivering production-ready web and desktop
applications.
• Programming Languages: Strong proficiency in Python, JavaScript/TypeScript, and at
least one compiled language (C#, C++, or Java).• Web & Desktop Frameworks: Hands-on experience with modern frontend
frameworks (React, Angular, or Vue.js), Node.js, and desktop application development
(Electron, WPF, Qt, or Tauri).
• AI & Agent Tooling: Demonstrated experience building AI agents using LLM APIs
(OpenAI, Anthropic), open-source models (Hugging Face), LangChain, LlamaIndex,
AutoGPT, or custom agent architectures.
• Automation & UI Control: Expertise in software control mechanisms using tools like
Selenium, Playwright, PyAutoGUI, Appium, or computer vision-based GUI automation to
allow AI agents to navigate software.
• Cloud, DevOps & Databases: Experience with Git, Docker, CI/CD pipelines, cloud
platforms (AWS/Azure/GCP), RESTful APIs, GraphQL, and relational/NoSQL databases.
Preferred Qualifications (Strong Plus)
• Industry Domain Expertise: Prior hands-on development experience within the oil and
gas sector, specifically focused on drilling, completions, rig operations, or subsurface
engineering software.
• Data & Protocols: Familiarity with oilfield data standards (e.g., WITSML, OPC-UA) and
time-series databases.
• Experience deploying AI models and agents in edge or low-connectivity environments
(such as offshore rigs or remote drilling sites).
• Familiarity with safety-critical software design and cybersecurity standards in
industrial control systems (ICS/SCADA).
• Degree in Computer Science, Software Engineering, Petroleum Engineering, or a related technical discipline.
The recruiter has not been active on this job recently. You may apply but please expect a delayed response.
KEY RESPONSIBILITIES:
•Build agents with persistent context & memory
•Design self-learning feedback loops
•Implement RAG pipelines for domain knowledge
•Manage conversation state & orchestration
•Integrate with LLM APIs (OpenAI, Claude, open-source)
Iterate fast — ship daily, measure weekly
MUST-HAVE SKILLS
•Python / TypeScript proficiency
•LangChain, CrewAI, AutoGen or custom frameworks
•Experience with vector DBs (Pinecone, Weaviate, Qdrant)
•Prompt engineering & evaluation pipelines
•Understanding of agent architectures (ReAct, tool-use)
Git, CI/CD, containerization basics
AI Engineer
LLMs, Agents & AI Services
📍 Mumbai (On-site) | Full-time | 2-4 years
About the Role:
Unico Connect is an AI-first technology partner that builds custom mobile, web, and AI products for clients across multiple geographies.
AI is core to how we design, deliver, and scale software for our customers.
We are hiring an AI Engineer for a dedicated client engagement building a complex production AI platform, working on the AI capabilities and agentic features at the core of the product.
The mandatory requirement for this role is at least one AI feature personally shipped to production for real users, with operational ownership.
The role suits someone who thinks quickly on solutioning, can take an ambiguous problem to a working prototype in days, and has the discipline to carry it through to production with predictable economics.
You will work alongside the Senior AI Engineer and the wider pod, with ownership of parts of the AI surface area of the product.
Responsibilities:
Solutioning and POCs
Translate ambiguous customer problems into working POCs at speed.
Pick the right model, framework, and architecture, and demonstrate value early before scaling investment.
LLM Application Development
Build AI features and services using LLM APIs from OpenAI, Anthropic, Google, and self-hosted open-weight models (Llama, Qwen, Mistral).
Choose the right model per use case based on cost, latency, capability, and context-window trade-offs.
Agentic System Design
Design and implement agentic workflows using LangGraph, CrewAI, AutoGen, LlamaIndex Agents, or custom orchestration.
Cover tool use, planning, memory, and multi-step reasoning appropriate to the problem.
API and Service Development
Build production AI services and APIs using Python and FastAPI.
Handle streaming responses, async processing, structured outputs, retries, and graceful degradation when models or tools fail.
Retrieval and Tool Integration
Implement RAG pipelines with vector databases (Pinecone, Weaviate, Qdrant, pgvector, Chroma), embeddings, chunking strategies, hybrid search, and reranking.
Integrate external tools, internal APIs, and document sources through tool-calling and MCP-style patterns.
Cost Analysis and Unit Economics
Model the per-request and per-user cost of every AI feature before it ships.
Track token usage, prompt caching, batching, and model-routing strategies.
Drive measurable improvements in unit economics.
Production Hardening
Add observability and tracing (LangSmith, Langfuse, OpenTelemetry), guardrails, content safety checks, prompt injection defences, and fallback behaviour.
Prompt Engineering and Evaluation
Design, test, and iterate prompts with measured outcomes.
Build evaluation harnesses for accuracy, hallucination, latency, and cost.
Run benchmarks across models and prompt variants before locking in a design.
Requirements:
AI Feature Shipped to Production (Mandatory)
Must have personally built and shipped at least one AI feature that runs in production for real users, with operational ownership.
POCs, internal demos, and one-off scripts do not qualify.
2 to 4 Years of Professional Software or AI Engineering Experience
With at least one production AI feature owned end to end.
Strong Python Proficiency and API Development with FastAPI
Comfort with type hints, async, packaging, testing, streaming responses, and authentication.
Production-grade Python, not notebook-only code.
Hands-on Depth Across the LLM and Agent Stack
Working experience with at least two of OpenAI, Anthropic Claude, Google Gemini, or self-hosted open-weight models (vLLM, Ollama, Together, Replicate).
Working familiarity with at least one agent framework (LangGraph, CrewAI, AutoGen, LlamaIndex Agents) or hand-rolled equivalent.
Working knowledge of RAG, embeddings, and vector databases (Pinecone, Weaviate, Qdrant, pgvector, Chroma).
Solutioning Speed and POC Velocity
Demonstrated ability to move from a fuzzy problem to a working prototype in days.
Strong instinct for what to build first, what to defer, and what to throw away.
Cost Discipline for Production AI
Ability to calculate, monitor, and optimise the cost of LLM APIs, tokens, embeddings, vector store usage, and infrastructure.
Treats unit economics as a first-class concern.
AWS Familiarity
Working knowledge of EC2, S3, IAM, and at least one of Bedrock, SageMaker, or equivalent.
Comfortable in a Fast-Moving Environment
Self-directed, comfortable with ambiguity, takes ownership without being asked, and ships under shifting priorities.
Strong Written and Spoken English Communication
Able to explain trade-offs to non-AI engineers, designers, product managers, and clients in plain language.
Nice to Have
- fine-tuning or LoRA, QLoRA, PEFT exposure
- MCP server authoring
- eval framework experience (LangSmith, Promptfoo, Ragas, DeepEval)
- open-source AI contributions
- multi-modal models (vision, audio)






