Urgent Opening for Sr. AI Engineer at Techwize · Ahmedabad · 4 - 13 years · ₹2L - ₹15L / yr · Raised funding · Posted 29 Jul 2025

We are looking for skilled AI Engineer for our organization. Please have a look to below details and revert accordingly if we can discuss further
Company Name: TechWize (A Business unit of Mangalam Information Technologies Pvt Ltd.)
Our Accreditations -
- 25 years of industry presence
- Salesforce Partner
- ISO 27001:2019 certified
- Great Place to Work certified
- HIPAA Compliant
- SOC2 Compliant
- NASSCOM Member
Our EVP (Employee Value proposition)
- We are a Great place to work certified company.
- 30 Earned Leaves during calendar Year
- Career progression and continuous Learning & Development (Technical, Soft skills, Communication, Leadership)
- Performance bonus & Loyalty Bonus Benefits
- 5 Days working
- Rewards and Recognition programs
- Standard Salary as per market norms
- Equal career opportunities, No discrimination
- Magnificent & Dynamic Culture
- Festival celebrations & fun events
Explore more : https://techwize.com/, https://mangalaminfotech.com/
Position: AI/Sr. AI Engineer
Job location: Ahmedabad
Experience: 4+ years
Job Overview:
We are seeking a highly experienced and innovative Senior AI Engineer with a strong background in Generative AI, including LLM fine-tuning and prompt engineering. This role requires hands-on expertise across NLP, Computer Vision, and AI agent-based systems, with the ability to build, deploy, and optimize scalable AI solutions using modern tools and frameworks.
Key Responsibilities:
- Design, fine-tune, and deploy generative AI models (LLMs, diffusion models, etc.) for real-world applications.
- Develop and maintain prompt engineering workflows, including prompt chaining, optimization, and evaluation for consistent output quality.
- Build NLP solutions for Q&A, summarization, information extraction, text classification, and more.
- Develop and integrate Computer Vision models for image processing, object detection, OCR, and multimodal tasks.
- Architect and implement AI agents using frameworks such as LangChain, AutoGen, CrewAI, or custom pipelines.
- Collaborate with cross-functional teams to gather requirements and deliver tailored AI-driven features.
- Optimize models for performance, cost-efficiency, and low latency in production.
- Continuously evaluate new AI research, tools, and frameworks and apply them where relevant.
- Mentor junior AI engineers and contribute to internal AI best practices and documentation.
Required Skills & Qualifications:
- Bachelor’s or Master’s in Computer Science, AI, Machine Learning, or related field.
- 5+ years of hands-on experience in AI/ML solution development.
- Proven expertise in fine-tuning LLMs (e.g., LLaMA, Mistral, Falcon, GPT-family) using techniques like LoRA, QLoRA, PEFT.
- Deep experience in prompt engineering, including zero-shot, few-shot, and retrieval-augmented generation (RAG).
- Proficient in key AI libraries and frameworks:
- LLMs & GenAI: Hugging Face Transformers, LangChain, LlamaIndex, OpenAI API, Diffusers
- NLP: SpaCy, NLTK.
- Vision: OpenCV, MMDetection, YOLOv5/v8, Detectron2
- MLOps: MLflow, FastAPI, Docker, Git
- Familiarity with vector databases (Pinecone, FAISS, Weaviate) and embedding generation.
- Experience with cloud platforms like AWS, GCP, or Azure, and deployment on in house GPU-backed infrastructure.
- Strong communication skills and ability to convert business problems into technical solutions.

Similar jobs (10)
Senior Generative AI Engineer
Employment Type: Permanent with VDart Digital
Work Location: Marathalli, Bengaluru
Job Description
We are seeking a highly skilled Senior Generative AI Engineer with strong expertise in designing, developing, and deploying enterprise-scale AI solutions using Large Language Models (LLMs) and modern Generative AI frameworks. The ideal candidate should have hands-on production experience building scalable GenAI applications, AI agents, autonomous workflows, and Retrieval-Augmented Generation (RAG) systems in cloud-native environments.
This role requires deep technical expertise in LLM orchestration, AI application architecture, prompt engineering, vector databases, MLOps, and production deployment of AI systems. Candidates should have proven experience delivering real-world AI solutions in enterprise environments with strong exposure to cloud platforms and DevOps practices.
Key Responsibilities
- Design, build, and deploy enterprise-grade Generative AI applications using Large Language Models (LLMs).
- Develop intelligent AI agents and autonomous workflows using frameworks such as LangChain, CrewAI, LangGraph, AutoGen, or similar agentic AI frameworks.
- Implement and optimize Retrieval-Augmented Generation (RAG) pipelines using vector databases and semantic search technologies.
- Work extensively on prompt engineering, tool calling, memory management, agent orchestration, and multi-agent systems.
- Integrate and manage LLMs such as OpenAI, Azure OpenAI, Claude, Gemini, Llama, Mistral, or similar foundation models.
- Develop scalable AI services and APIs using Python and FastAPI.
- Build production-ready AI solutions with high availability, scalability, monitoring, and observability.
- Deploy and manage AI applications in cloud-native environments using Docker and Kubernetes.
- Collaborate with Data Science, ML Engineering, and DevOps teams to operationalize AI solutions.
- Implement CI/CD pipelines and automated deployment processes for AI workloads.
- Monitor model performance, latency, reliability, and operational efficiency in production environments.
- Ensure AI solutions follow enterprise security, governance, and responsible AI standards.
- Evaluate and adopt emerging Generative AI tools, frameworks, and models.
Required Skills
Generative AI & LLM Expertise
- Strong hands-on experience with Generative AI and Large Language Models (LLMs).
- Production-level experience building and deploying GenAI applications.
- Expertise in LangChain, CrewAI, LangGraph, AutoGen, or similar frameworks.
- Experience with AI agents, autonomous workflows, and multi-agent architectures.
- Strong understanding of prompt engineering, embeddings, model evaluation, and LLM orchestration.
- Experience integrating OpenAI, Azure OpenAI, Claude, Gemini, Llama, Mistral, or similar models.
RAG & Vector Databases
- Strong experience implementing RAG pipelines and semantic retrieval systems.
- Experience with vector databases such as Pinecone, Weaviate, ChromaDB, FAISS, or Milvus.
- Understanding of chunking strategies, embeddings, indexing, reranking, and retrieval optimization.
Python & AI Development
- Strong proficiency in Python.
- Experience with FastAPI for AI service and API development.
- Experience with AI/ML libraries and data processing tools such as Pandas and NumPy.
Cloud & Production Deployment
- Mandatory production experience on at least one cloud platform:
- Microsoft Azure
- Experience deploying scalable AI applications in enterprise production environments.
- Hands-on experience with Docker, Kubernetes, Jenkins, Terraform, and CI/CD pipelines.
- Strong understanding of MLOps, AI deployment lifecycle, monitoring, and observability.
Engineering & Operational Excellence
- Strong understanding of software engineering best practices.
- Experience with Git, version control, automated testing, and release management.
- Experience building secure, scalable, and high-performance AI solutions.
- Ability to troubleshoot production AI systems and optimize performance.
Preferred Skills
- Experience with AI observability and evaluation frameworks.
- Exposure to fine-tuning, PEFT, LoRA, or model optimization techniques.
- Experience with enterprise AI governance and responsible AI practices.
- Knowledge of distributed AI systems and scalable inference architectures.
- Familiarity with AI security and compliance standards.
Qualifications
- Bachelor’s or Master’s degree in Computer Science, Artificial Intelligence, Data Science, or a related field.
- 3–8 years of overall software engineering experience.
- Minimum 3+ years of hands-on experience in Generative AI and LLM-based application development,
- Proven track record of delivering enterprise-scale AI solutions in production environments.
- Strong communication and stakeholder management skills.
Key Responsibilities
• Design, build, and deploy machine learning and AI models that power Transient.AI's core products (research
automation, document intelligence, investor matching, and workflow orchestration).
• Work on applied NLP/LLM systems, including retrieval-augmented generation, structured extraction from
unstructured financial documents, and model evaluation pipelines.
• Partner closely with product and founding engineers to translate capital markets workflows into scalable AI
systems.
• Own model performance, reliability, and cost — from experimentation through production deployment.
• Build and maintain data pipelines, feature stores, and evaluation frameworks to support rapid iteration.
• Ensure systems meet the compliance, auditability, and security standards required in regulated financial
environments.
What We're Looking For
• 5+ years of experience building and deploying machine learning or AI systems in production.• Strong hands-on experience with Python and modern ML/AI frameworks (PyTorch, TensorFlow, Hugging Face,
LangChain, or equivalent).
• Experience with LLMs — fine-tuning, prompt engineering, RAG architectures, or agentic systems — is highly
valued.
• Solid grounding in data structures, distributed systems, and MLOps practices (model serving, monitoring,
versioning).
• Prior experience at a strong product company, high-growth startup, or a top-tier engineering background
• Comfort operating in an early-stage, high-ownership environment with limited process and high ambiguity.
• Exposure to fintech, capital markets, or other regulated industries is a plus, though not mandatory
We are seeking Generative AI Developers with strong Python programming and AI/ML expertise to build, deploy, and optimize LLM-powered applications. The role involves developing RAG solutions, AI agents, and enterprise GenAI applications while collaborating with cross-functional teams.
Key Responsibilities
- Develop and enhance Generative AI applications using LLMs and AI frameworks.
- Build and optimize RAG pipelines, vector search, and AI-powered workflows.
- Design effective prompts and fine-tune models using techniques such as LoRA and QLoRA.
- Develop REST APIs and integrate AI capabilities into enterprise applications.
- Deploy, monitor, and maintain AI solutions in cloud and containerized environments.
- Ensure code quality through testing, debugging, documentation, and code reviews.
- Follow Responsible AI, security, and data governance practices.
Required Technical Skills
- Strong proficiency in Python, OOP, APIs, debugging, and software development best practices.
- Good understanding of Data Structures & Algorithms, complexity analysis, and problem-solving.
- Hands-on experience with LLMs, Prompt Engineering, RAG, AI Agents, and embeddings.
- Experience with LangChain, LangGraph, LlamaIndex, Hugging Face, or similar frameworks.
- Knowledge of vector databases, semantic/hybrid search, and retrieval architectures.
- Experience with PyTorch, TensorFlow, or Keras.
- Familiarity with Docker, Git, CI/CD, and cloud platforms (Azure/AWS/GCP).
- Understanding of AI governance, data privacy, and Responsible AI principles.
Preferred Skills
- Experience with Agentic AI frameworks (CrewAI, AutoGen, Semantic Kernel).
- Exposure to Azure AI Foundry, Databricks, or enterprise AI platforms.
- Knowledge of multimodal AI applications.
Qualifications
- Bachelor's or Master's degree in Computer Science, AI, Data Science, or a related field.
- 5 years of software development experience, including AI/ML or Generative AI projects.
- Experience building and deploying production-grade AI solutions.
Assessment Focus Areas
Candidates will be evaluated on:
- Python coding and problem-solving
- Data Structures & Algorithms
- LLMs, RAG, and Agentic AI concepts
- API development and system design
- Cloud deployment and AI solution architecture
🔹 Key Responsibilities
• Design, develop, and deploy production-grade AI/ML and Generative AI solutions
• Work on GEO, AEO, and SGE initiatives to improve visibility across AI-driven search platforms
• Optimize content and digital experiences for conversational queries and LLM-based search
• Develop solutions using LLMs, NLP, embeddings, semantic search, RAG, and vector databases
• Analyze search intent, AI-generated responses, citations, retrieval patterns, and content discoverability
• Build frameworks to measure GEO/AEO strategies and AI-search performance
• Collaborate with Product, Engineering, Content, SEO, Marketing, and Business teams
• Improve solution accuracy, relevance, latency, and user experience
🔹 Mandatory Requirements
✅ 1–4 years of professional experience
✅ Minimum 1 year of hands-on experience in GEO, AEO, or SGE
✅ Experience with prompt engineering, embeddings, vector search, or RAG systems
✅ Understanding of semantic search and entity-based optimization
✅ Exposure to ChatGPT, Google Gemini, or similar LLM platforms
✅ Knowledge of schema, context building, content structuring, and knowledge representation
🎓 Preferred Education
B.Tech, M.Tech, Integrated M.Sc., or MS from a Tier-1 engineering institute such as IIT, NIT, BITS, VIT, DTU, or NSUT.
Location: Jaipur (Work From Office)
Employment Type: Full-Time
We're looking for a GenAI Engineer (LLM Engineer) to build scalable AI-powered SaaS applications using Large Language Models (LLMs). You'll develop intelligent AI workflows, integrate LLMs into production systems, and build secure, high-performance AI solutions.
Key Responsibilities
- Integrate LLM APIs (OpenAI, Claude, Hugging Face) into production applications.
- Design and optimize RAG pipelines and prompt engineering workflows.
- Build and manage Vector Databases (Pinecone, Weaviate, pgvector).
- Optimize AI performance, latency, and operational cost.
- Ensure secure, scalable AI architecture.
- Collaborate with Product and Engineering teams to deliver AI-powered features.
Requirements
- 3+ years of backend development using Python, Go, or Node.js.
- Hands-on experience with LLMs, LangChain or LlamaIndex.
- Strong understanding of RAG, Prompt Engineering, and Vector Databases.
- Experience with AWS, GCP, or Azure.
- Knowledge of APIs, Microservices, and AI application development.
Preferred: Experience in SaaS/FinTech, LLMOps, or Model Fine-tuning.
Education: B.Tech, BCA, or equivalent technical qualification.
Apply Now
Application Form: https://zfrmz.com/pAKb2ynfomIsuNwRfRbV?utm_source=cutshort
About the Role:
We are looking for an ideal candidate with 5+ years of experience in Data Science / Machine Learning, with strong hands-on experience in Generative AI, Large Language Models (LLMs), NLP, and AI-powered applications. The candidate should be comfortable working across the complete AI lifecycle—from understanding business requirements and experimenting with models to building, evaluating, deploying, and monitoring production-grade GenAI solutions.
The role requires a combination of strong technical expertise, business understanding, problem-solving ability, and stakeholder management skills.
Key Responsibilities:
Generative AI & LLM
· Design, develop, and deploy Generative AI and LLM-based solutions for enterprise use cases.
· Work with models such as OpenAI, Azure OpenAI, Llama, Mistral, Gemini, or equivalent LLM platforms.
· Develop applications using prompt engineering, structured outputs, function/tool calling, and LLM orchestration.
· Design and implement Retrieval-Augmented Generation (RAG) solutions.
· Work with vector databases and semantic search for enterprise knowledge retrieval.
· Develop and evaluate AI agents and multi-step AI workflows.
· Implement techniques such as prompt optimization, context management, grounding, and hallucination reduction.
· Develop AI solutions for text classification, summarization, information extraction, question answering, document intelligence, and other enterprise use cases.
Machine Learning & Data Science
· Develop and optimize traditional Machine Learning and statistical models where appropriate.
· Perform data exploration, feature engineering, model selection, training, validation, and evaluation.
· Apply appropriate ML and statistical techniques to solve business problems.
· Work with structured, unstructured, and semi-structured data.
· Develop scalable data pipelines to support AI/ML solutions.
· Collaborate with Data Engineers to prepare and manage data for AI applications.
AI Evaluation & Productionization
· Design evaluation frameworks to measure LLM accuracy, relevance, groundedness, toxicity, latency, and cost.
· Implement guardrails and responsible AI practices.
· Monitor model and application performance in production.
· Identify model/data drift and implement appropriate improvement strategies.
· Optimize AI solutions for performance, scalability, reliability, and cost.
· Support deployment and productionization of AI/ML solutions.
· Client & Delivery Responsibilities
· Work closely with the CEO, Delivery team, Solution Architects, Engineering teams, and clients to understand business problems and identify AI opportunities.
· Translate business requirements into practical AI/ML solutions.
· Participate in client discussions, solution presentations, technical workshops, and POCs.
· Develop rapid prototypes and demonstrate the feasibility of GenAI solutions.
· Convert successful POCs into scalable, production-ready applications.
· Provide technical guidance and contribute to AI solution architecture.
· Prepare technical documentation, solution approaches, and project estimates where required.
· Stay current with developments in Generative AI, LLMs, Agentic AI, and AI engineering.
Required Skills:
· 5+ years of hands-on experience in Data Science, Machine Learning, AI, or a related field.
· Strong practical experience in Generative AI and LLM-based applications.
· Strong proficiency in Python.
· Strong understanding of Machine Learning and statistical concepts.
· Hands-on experience with:
o LLMs
o Prompt Engineering
o RAG
o Vector Databases
o Embeddings
o Semantic Search
o LLM Evaluation
o AI Guardrails
· Experience with frameworks/tools such as LangChain, LangGraph, LlamaIndex, or equivalent.
· Experience with APIs and integrating LLMs into enterprise applications.
· Strong SQL and data handling skills.
· Experience working with large and complex datasets.
· Strong understanding of NLP concepts.XX
Technical Skills:
· Experience with OpenAI / Azure OpenAI / AWS Bedrock / Google Vertex AI.
· Experience with vector databases such as Pinecone, Weaviate, Milvus, FAISS, or equivalent.
· Experience with Databricks, Snowflake, or cloud data platforms.
· Experience with Docker and CI/CD.
· Exposure to AWS, Azure, or GCP.
· Experience with ML/AI deployment and MLOps.
· Knowledge of AI security, data privacy, governance, and responsible AI.
· Experience building AI Agents / Agentic AI workflows.
· Experience with multimodal AI is an added advantage
Key Competencies
· Strong analytical and problem-solving ability.
· Ability to translate business problems into practical AI solutions.
· Strong communication and presentation skills.
· Ability to interact confidently with senior stakeholders and clients.
· Strong ownership and delivery mindset.
· Ability to work independently in a fast-paced environment.
- Strong experimentation and innovation mindset.
- Ability to balance technical feasibility, business value, scalability, and cost.
Required Education & Experience:
· Bachelor's or Master's degree in Computer Science, Data Science, Artificial Intelligence, Statistics, Mathematics, Engineering, or a related discipline
Principal Software Engineer
Company Summary :
As the recognized global standard for project-based businesses, Deltek delivers software and information solutions to help organizations achieve their purpose. Our market leadership stems from the work of our diverse employees who are united by a passion for learning, growing and making a difference. At Deltek, we take immense pride in creating a balanced, values-driven environment, where every employee feels included and empowered to do their best work. Our employees put our core values into action daily, creating a one-of-a-kind culture that has been recognized globally. Thanks to our incredible team, Deltek has been named one of America's Best Midsize Employers by Forbes, a Best Place to Work by Glassdoor, a Top Workplace by The Washington Post and a Best Place to Work in Asia by World HRD Congress. www.deltek.com
Position Responsibilities :
About the Role
We are seeking a highly motivated AI Solutions Engineer to join Deltek’s growing AI Center of Excellence team to design, develop, deploy, and optimize internal Artificial Intelligence and Machine Learning solutions that solve complex business challenges. The ideal candidate combines deep expertise in AI, machine learning, Generative AI, Large Language Models (LLMs), SLMs, software engineering, cloud computing, and MLOps/LLMOps to build scalable, production-grade AI applications.
The AI Solutions Engineer will collaborate with AI data scientists, architects, and engineering teams to deliver innovative AI-driven solutions while ensuring security, scalability, governance, and operational excellence. This role reports to the Senior AI Solutions Architect.
Key Responsibilities
AI & Machine Learning Development
- Design, build, train, evaluate, and deploy machine learning and deep learning models.
- Develop Generative AI solutions using Large Language Models (LLMs) such as GPT, Claude, Gemini, Llama, and Mistral.
- Implement Retrieval-Augmented Generation (RAG), prompt engineering, fine-tuning, and AI agent frameworks.
- Build NLP, recommendation systems, forecasting, predictive analytics, and intelligent automation solutions.
- Optimize model performance, scalability, latency, and cost.
Software Engineering & Solution Development
- Develop production-grade AI applications using Python and modern software engineering practices.
- Build APIs, microservices, and AI-powered enterprise applications.
- Integrate AI services with enterprise systems, business applications, and data platforms.
- Apply coding standards, automated testing, CI/CD, and version control best practices.
MLOps & AI Operations
- Design and implement MLOps pipelines for model development, deployment, monitoring, and lifecycle management.
- Automate model training, validation, testing, and deployment processes.
- Monitor model performance, data drift, hallucinations, and operational metrics.
- Support continuous improvement and reliability of AI platforms.
Cloud & Platform Engineering
- Develop AI solutions on Azure, AWS, or Google Cloud platforms.
- Leverage cloud-native AI services, containerization, Kubernetes, and serverless technologies.
- Build scalable architectures supporting enterprise AI workloads and real-time inference.
AI Governance & Security
- Ensure compliance with Responsible AI, security, privacy, and regulatory requirements.
- Implement model governance, explainability, bias mitigation, and risk management practices.
- Maintain standards for secure design, deployment, and operation of AI solutions.
Required Qualifications
Education
- Bachelor's or Master's degree in Computer Science, Artificial Intelligence, Data Science, Engineering, or a related technical field.
Experience
- 5+ years of software engineering or machine learning development experience.
- 2+ years of hands-on experience developing and deploying Agentic AI, Generative AI or AI/ML solutions in production environments.
Technical Skills
Programming & Engineering
- Strong expertise in Python.
- Experience with Java, ReactJS, JavaScript, or similar programming languages.
- Solid understanding of algorithms, data structures, APIs, and software design principles.
Artificial Intelligence & Machine Learning
- Machine Learning and Deep Learning concepts and frameworks.
- Model training, evaluation, optimization, and deployment.
Generative AI
- Large Language Models (LLMs) & SLMs
- Prompt Engineering
- Retrieval-Augmented Generation (RAG)
- AI Agents and Agentic Workflows
- Fine-tuning and model customization
- Vector embeddings and semantic search
Frameworks & Tools
- PyTorch, TensorFlow, Scikit-learn
- LangChain, LlamaIndex, Semantic Kernel, MCP, A2A and Transformers
- FastAPI, Flask
Data & Analytics
- SQL and NoSQL databases
- Data pipelines, ETL, and data modeling
- Experience with AWS, Azure and Google
MLOps & DevOps
- MLflow, Kubeflow, Azure ML, SageMaker
- Docker and Kubernetes
- Git, GitHub, Azure DevOps, Jenkins
- CI/CD automation and model monitoring
Cloud Platforms
- AWS (preferred)
- AWS Bedrock or Azure OpenAI Service
- AWS SageMaker
- Google Vertex AI
Preferred Qualifications
- Experience designing enterprise-scale AI platforms and products.
- Knowledge of multi-agent architectures and autonomous AI systems.
- Experience with vector databases such as Pinecone, Snowflake Cortex, Pgvector, Weaviate, Chroma, or Azure AI Search.
- Understanding of AI governance, compliance, and Responsible AI frameworks.
- Relevant certifications in Azure AI, AWS Machine Learning, or Google Cloud AI.
About the role
We are seeking an AI Engineer to build and implement AI systems for content production at scale. You'll work at the intersection of engineering and content designing prompt pipelines, integrating generative models, and building the tooling that turns source material into finished creative output. The ideal candidate is technically strong but also has taste: someone who understands story and craft, and can tell the difference between output that's technically correct and output that's actually good.
Responsibilities
- Build and iterate on prompt pipelines and multi-agent workflow components
- Design and integrate agentic workflows orchestrate multi-step, tool-using agents that plan, call models, and hand off between stages in production
- Deploy and serve open-source models set up inference endpoints, manage GPU compute, and optimize for latency and cost
- Write evals compare outputs against references, quantify quality, and feed results back into the pipeline
- Work on data pipelines: structured extraction from messy source text, localization, similarity/dedup
- Debug and maintain pipeline stages in production
What you bring:
- (1+/3+) years of engineering experience, or a strong portfolio of shipped projects
- Solid Python fundamentals clean, working, readable code
- Hands-on experience with LLM APIs and prompt engineering (personal projects count)
- Comfort with Git, REST APIs, and working in a Linux environment
- A feel for content and narrative you can judge whether generated output is actually good, not just valid
- Curiosity and clear communication you ask good questions and don't stay stuck silently
Preferred
- Exposure to agent/orchestration frameworks (LangGraph, LangChain, CrewAI)
- Familiarity with vector databases, embeddings, or RAG (Qdrant, pgvector)
- Hands-on work with open-source generative media models Flux, LTX, Wan, or similar
- Experience deploying open-source models for inference (vLLM, ComfyUI, Replicate/Cog, Docker + GPU)
- Experience writing evals or LLM-as-judge scoring
- Node.js and Fastapi familiarity, or experience deploying on AWS
About the role
We are seeking an AI Engineer to build and implement AI systems for content production at scale. You'll work at the intersection of engineering and content designing prompt pipelines, integrating generative models, and building the tooling that turns source material into finished creative output. The ideal candidate is technically strong but also has taste: someone who understands story and craft, and can tell the difference between output that's technically correct and output that's actually good.
Responsibilities
- Build and iterate on prompt pipelines and multi-agent workflow components
- Design and integrate agentic workflows orchestrate multi-step, tool-using agents that plan, call models, and hand off between stages in production
- Deploy and serve open-source models set up inference endpoints, manage GPU compute, and optimize for latency and cost
- Write evals compare outputs against references, quantify quality, and feed results back into the pipeline
- Work on data pipelines: structured extraction from messy source text, localization, similarity/dedup
- Debug and maintain pipeline stages in production
What you bring:
- (1+/3+) years of engineering experience, or a strong portfolio of shipped projects
- Solid Python fundamentals clean, working, readable code
- Hands-on experience with LLM APIs and prompt engineering (personal projects count)
- Comfort with Git, REST APIs, and working in a Linux environment
- A feel for content and narrative you can judge whether generated output is actually good, not just valid
- Curiosity and clear communication you ask good questions and don't stay stuck silently
Preferred
- Exposure to agent/orchestration frameworks (LangGraph, LangChain, CrewAI)
- Familiarity with vector databases, embeddings, or RAG (Qdrant, pgvector)
- Hands-on work with open-source generative media models Flux, LTX, Wan, or similar
- Experience deploying open-source models for inference (vLLM, ComfyUI, Replicate/Cog, Docker + GPU)
- Experience writing evals or LLM-as-judge scoring
- Node.js and Fastapi familiarity, or experience deploying on AWS
Job Description:
We are seeking a highly skilled Machine Learning Engineer to join our team. The ideal candidate will have a strong background in Natural Language Processing (NLP), Large Language Models (LLMs), and Python programming.
You will work closely with data scientists, product managers, and data engineers to design, develop, and deploy high-performance AI/ML models and integrate generative AI solutions into existing workflows.
Your responsibilities will include:
- Collaborating with cross-functional teams to design and deliver high-performance AI models, including NLP, computer vision, semantics engines, linguistic analysis, risk management, and time-series prediction models. Integrating generative AI solutions into existing workflow systems.
- Developing and maintaining the ML Operations CI/CD pipeline for seamless deployment and monitoring. Training, tuning, and optimizing AI models and algorithms for enhanced performance.
- Implementing complex real-time data and AI/ML applications to capture knowledge and automate decision-making processes.
- Creating ML/AI models for business teams and establishing metrics to track their accuracy and performance. Overseeing the full lifecycle of algorithm development, from ideation to deployment and monitoring. Evaluating and ranking ML algorithms based on their potential success in solving specific problems.
- Serving as an internal resource for AI/ML needs, providing guidance and insights to stakeholders during strategic discussions.
Required Experience and Skills:
Machine Learning:
- Proficient in generative AI techniques, prompt engineering, and Retrieval-Augmented Generation (RAG) (3+ years).
- Experience with Large Language Models (LLMs) such as OpenAI, Gemini, LLAMA, and other state-of-the-art models (3+ years).
- Expertise in using ML/AI libraries such as Pandas, NumPy, PyTorch, TensorFlow, Keras, BERT, LayoutLM, and traditional ML algorithms (5+ years).
- Experience with distributed ML/AI training libraries/models: Koalas, Horovod, DDP.
Python Programming and Software Engineering:
- Expertise in Pythonic clean coding practices, including the use of decorators, generators, and descriptors (5+ years).
- Strong understanding of software design principles such as DRY, OAOO, YAGNI, KIS, EAFP/LBYL, and defensive programming (2+ years).
- Proficient in software design concepts focusing on cohesion and coupling (2+ years). Knowledge of SOLID principles (2+ years).
Education and Experience:
- Minimum Bachelor's degree or foreign equivalent in Computer Science, Electrical Engineering, or a closely related field.
- At least 5 years of experience as a software engineer and 5 years of ML-related programming.






