Data Scientist - R & D at Acuity Knowledge Partners · Bengaluru (Bangalore) · 3 - 9 years · ₹35L - ₹50L / yr · Profitable · Posted 21 Feb 2025

Data Scientist is responsible to discover the information hidden in vast amounts of data, and help us make smarter decisions to deliver better products. Your primary focus will be in applying Machine Learning and Generative AI techniques for data mining and statistical analysis, Text analytics using NLP/LLM and building high quality prediction systems integrated with our products. The ideal candidate should have a prior background in Generative AI, NLP (Natural Language Processing), and Computer Vision techniques. Additionally, experience in working with current state of the art Large Language Models (LLMs), and Computer Vision algorithms.
Job Responsibilities:
» Building models using best in AI/ML technology.
» Leveraging your expertise in Generative AI, Computer Vision, Python, Machine Learning, and Data Science to develop cutting-edge solutions for our products.
» Integrating NLP techniques, and utilizing LLM's in our products.
» Training/fine tuning models with new/modified training dataset.
» Selecting features, building and optimizing classifiers using machine learning techniques.
» Conducting data analysis, curation, preprocessing, modelling, and post-processing to drive data-driven decision-making.
» Enhancing data collection procedures to include information that is relevant for building analytic systems
» Working understanding of cloud platforms (AWS).
» Collaborating with cross-functional teams to design and implement advanced AI models and algorithms.
» Involving in R&D activities to explore the latest advancements in AI technologies, frameworks, and tools.
» Documenting project requirements, methodologies, and outcomes for stakeholders.
Technical skills
Mandatory
» Minimum of 5 years of experience as Machine Learning Researcher or Data Scientist.
» Master's degree or Ph.D. (preferable) in Computer Science, Data Science, or a related field.
» Should have knowledge and experience in working with Deep Learning projects using CNN, Transformers, Encoder and decoder architectures.
» Working experience with LLM's (Large Language Models) and their applications (For e.g., tuning embedding models, data curation, prompt engineering, LoRA, etc.).
» Familiarity with LLM Agents and related frameworks.
» Good programming skills in Python and experience with relevant libraries and frameworks (e.g., PyTorch, and TensorFlow).
» Good applied statistics skills, such as distributions, statistical testing, regression, etc.
» Excellent understanding of machine learning and computer vision based techniques and algorithms.
» Strong problem-solving abilities and a proactive attitude towards learning and adopting new technologies.
» Ability to work independently, manage multiple projects simultaneously, and collaborate effectively with diverse stakeholders.
Nice to have
» Exposure to financial research domain
» Experience with JIRA, Confluence
» Understanding of scrum and Agile methodologies
» Basic understanding of NoSQL databases, such as MongoDB, Cassandra
Experience with data visualization tools, such as Grafana, GGplot, etc.

About Acuity Knowledge Partners
About
Acuity Knowledge Partners (Acuity) is a leading provider of bespoke research, analytics and technology solutions to the financial services sector, including asset managers, corporate and investment banks, private equity and venture capital firms, hedge funds and consulting firms. Its global network of over 6,000 analysts and industry experts, combined with proprietary technology, supports more than 500 financial institutions and consulting companies to operate more efficiently and unlock their human capital, driving revenue higher and transforming operations. Acuity is headquartered in London and operates from 10 locations worldwide.
We EMPOWER our clients to drive revenues higher. We INNOVATE using our proprietary technology and automation solutions. Finally, we enable our clients to TRANSFORM their operating model and cost base.
Company video
Candid answers by the company
We EMPOWER our clients to drive revenues higher. We INNOVATE using our proprietary technology and automation solutions. Finally, we enable our clients to TRANSFORM their operating model and cost base.
Similar jobs (10)
About the Role:
We are looking for an ideal candidate with 5+ years of experience in Data Science / Machine Learning, with strong hands-on experience in Generative AI, Large Language Models (LLMs), NLP, and AI-powered applications. The candidate should be comfortable working across the complete AI lifecycle—from understanding business requirements and experimenting with models to building, evaluating, deploying, and monitoring production-grade GenAI solutions.
The role requires a combination of strong technical expertise, business understanding, problem-solving ability, and stakeholder management skills.
Key Responsibilities:
Generative AI & LLM
· Design, develop, and deploy Generative AI and LLM-based solutions for enterprise use cases.
· Work with models such as OpenAI, Azure OpenAI, Llama, Mistral, Gemini, or equivalent LLM platforms.
· Develop applications using prompt engineering, structured outputs, function/tool calling, and LLM orchestration.
· Design and implement Retrieval-Augmented Generation (RAG) solutions.
· Work with vector databases and semantic search for enterprise knowledge retrieval.
· Develop and evaluate AI agents and multi-step AI workflows.
· Implement techniques such as prompt optimization, context management, grounding, and hallucination reduction.
· Develop AI solutions for text classification, summarization, information extraction, question answering, document intelligence, and other enterprise use cases.
Machine Learning & Data Science
· Develop and optimize traditional Machine Learning and statistical models where appropriate.
· Perform data exploration, feature engineering, model selection, training, validation, and evaluation.
· Apply appropriate ML and statistical techniques to solve business problems.
· Work with structured, unstructured, and semi-structured data.
· Develop scalable data pipelines to support AI/ML solutions.
· Collaborate with Data Engineers to prepare and manage data for AI applications.
AI Evaluation & Productionization
· Design evaluation frameworks to measure LLM accuracy, relevance, groundedness, toxicity, latency, and cost.
· Implement guardrails and responsible AI practices.
· Monitor model and application performance in production.
· Identify model/data drift and implement appropriate improvement strategies.
· Optimize AI solutions for performance, scalability, reliability, and cost.
· Support deployment and productionization of AI/ML solutions.
· Client & Delivery Responsibilities
· Work closely with the CEO, Delivery team, Solution Architects, Engineering teams, and clients to understand business problems and identify AI opportunities.
· Translate business requirements into practical AI/ML solutions.
· Participate in client discussions, solution presentations, technical workshops, and POCs.
· Develop rapid prototypes and demonstrate the feasibility of GenAI solutions.
· Convert successful POCs into scalable, production-ready applications.
· Provide technical guidance and contribute to AI solution architecture.
· Prepare technical documentation, solution approaches, and project estimates where required.
· Stay current with developments in Generative AI, LLMs, Agentic AI, and AI engineering.
Required Skills:
· 5+ years of hands-on experience in Data Science, Machine Learning, AI, or a related field.
· Strong practical experience in Generative AI and LLM-based applications.
· Strong proficiency in Python.
· Strong understanding of Machine Learning and statistical concepts.
· Hands-on experience with:
o LLMs
o Prompt Engineering
o RAG
o Vector Databases
o Embeddings
o Semantic Search
o LLM Evaluation
o AI Guardrails
· Experience with frameworks/tools such as LangChain, LangGraph, LlamaIndex, or equivalent.
· Experience with APIs and integrating LLMs into enterprise applications.
· Strong SQL and data handling skills.
· Experience working with large and complex datasets.
· Strong understanding of NLP concepts.XX
Technical Skills:
· Experience with OpenAI / Azure OpenAI / AWS Bedrock / Google Vertex AI.
· Experience with vector databases such as Pinecone, Weaviate, Milvus, FAISS, or equivalent.
· Experience with Databricks, Snowflake, or cloud data platforms.
· Experience with Docker and CI/CD.
· Exposure to AWS, Azure, or GCP.
· Experience with ML/AI deployment and MLOps.
· Knowledge of AI security, data privacy, governance, and responsible AI.
· Experience building AI Agents / Agentic AI workflows.
· Experience with multimodal AI is an added advantage
Key Competencies
· Strong analytical and problem-solving ability.
· Ability to translate business problems into practical AI solutions.
· Strong communication and presentation skills.
· Ability to interact confidently with senior stakeholders and clients.
· Strong ownership and delivery mindset.
· Ability to work independently in a fast-paced environment.
- Strong experimentation and innovation mindset.
- Ability to balance technical feasibility, business value, scalability, and cost.
Required Education & Experience:
· Bachelor's or Master's degree in Computer Science, Data Science, Artificial Intelligence, Statistics, Mathematics, Engineering, or a related discipline
Strong Data Scientist / AI Engineer / Generative AI Engineer profile.
2
Mandatory (Experience 1) - Must have 3+ years of hands-on experience in Data Science, Artificial Intelligence, Machine Learning, Deep Learning, NLP, or Generative AI application development.
3
Mandatory (Experience 2) - Must have strong hands-on experience in Python programming, backend development, API development, and production-grade application support.
4
Mandatory (Experience 3) - Must have experience working with Machine Learning and Deep Learning frameworks such as PyTorch, TensorFlow, Keras, or Scikit-learn.
5
Mandatory (Experience 4) - Must have hands-on experience in NLP use cases such as text classification, sentiment analysis, entity recognition (NER), semantic search, embeddings, or document understanding.
6
Mandatory (Experience 5) - Must have experience working with Large Language Models (LLMs) such as GPT, LLaMA, Mistral, Phi, Claude, Gemini, or similar models.
7
Mandatory (Experience 6) - Must have hands-on experience building or implementing Retrieval Augmented Generation (RAG) solutions, vector search, semantic search, or knowledge-based AI applications.
8
Mandatory (Experience 7) - Must have experience with Prompt Engineering and Generative AI frameworks such as LangChain, LangGraph, AI Agents, Azure OpenAI, or similar technologies.
9
Mandatory (Experience 8) - Must have experience developing, consuming, or integrating APIs using Python frameworks such as FastAPI, Flask, or similar technologies.
10
Mandatory (CTC) - The CTC breakup offered will be 75% fixed + 25% variable, as per company policy.
11
Preferred (Experience 1) - Experience with LLMOps/MLOps tools for monitoring, evaluation, experimentation, and versioning of AI models.
12
Preferred (Experience 2) - Exposure to Azure OpenAI, Azure Kubernetes Service (AKS), Kubernetes, cloud-native AI deployments, or distributed systems.
13
Preferred (Experience 3) - Experience working with PostgreSQL, MongoDB, Redis, Kafka, or large-scale data platforms.
14
Preferred (Experience 4) - Familiarity with Docker, Kubernetes, cloud platforms, and scalable deployment architecture.
15
Preferred (Company) - Candidates from AI-first startups, product companies, SaaS organizations, fintech, or data-driven technology companies.
16
Mandatory ( Age ) - Candidate Should be Below 28 Years.
17
Mandatory ( Pedigree) - B.TECH / M.TECH from Tier 1 Colleges (IIT's, NIT's, BITS) are Considered.
Strong Data Scientist / AI Engineer / Generative AI Engineer profile.
2
Mandatory (Experience 1) - Must have 3+ years of hands-on experience in Data Science, Artificial Intelligence, Machine Learning, Deep Learning, NLP, or Generative AI application development.
3
Mandatory (Experience 2) - Must have strong hands-on experience in Python programming, backend development, API development, and production-grade application support.
4
Mandatory (Experience 3) - Must have experience working with Machine Learning and Deep Learning frameworks such as PyTorch, TensorFlow, Keras, or Scikit-learn.
5
Mandatory (Experience 4) - Must have hands-on experience in NLP use cases such as text classification, sentiment analysis, entity recognition (NER), semantic search, embeddings, or document understanding.
6
Mandatory (Experience 5) - Must have experience working with Large Language Models (LLMs) such as GPT, LLaMA, Mistral, Phi, Claude, Gemini, or similar models.
7
Mandatory (Experience 6) - Must have hands-on experience building or implementing Retrieval Augmented Generation (RAG) solutions, vector search, semantic search, or knowledge-based AI applications.
8
Mandatory (Experience 7) - Must have experience with Prompt Engineering and Generative AI frameworks such as LangChain, LangGraph, AI Agents, Azure OpenAI, or similar technologies.
9
Mandatory (Experience 8) - Must have experience developing, consuming, or integrating APIs using Python frameworks such as FastAPI, Flask, or similar technologies.
10
Mandatory (CTC) - The CTC breakup offered will be 75% fixed + 25% variable, as per company policy.
11
Preferred (Experience 1) - Experience with LLMOps/MLOps tools for monitoring, evaluation, experimentation, and versioning of AI models.
12
Preferred (Experience 2) - Exposure to Azure OpenAI, Azure Kubernetes Service (AKS), Kubernetes, cloud-native AI deployments, or distributed systems.
13
Preferred (Experience 3) - Experience working with PostgreSQL, MongoDB, Redis, Kafka, or large-scale data platforms.
14
Preferred (Experience 4) - Familiarity with Docker, Kubernetes, cloud platforms, and scalable deployment architecture.
15
Preferred (Company) - Candidates from AI-first startups, product companies, SaaS organizations, fintech, or data-driven technology companies.
16
Mandatory ( Age ) - Candidate Should be Below 28 Years.
17
Mandatory ( Pedigree) - B.TECH / M.TECH from Tier 1 Colleges (IIT's, NIT's, BITS) are Considered.
Sr.Data Scientist,Python, AI ML
We are looking for a skilled Data Scientist to analyze complex datasets, develop predictive models, and generate actionable insights that support business decisions. The ideal candidate should have strong statistical, analytical, and programming skills, along with hands-on experience in machine learning.
Sr Engineer – Artificial Intelligence
Job Summary
As an AI Engineer at Emerson, you will be responsible for analysing complex data sets to
identify trends, develop predictive models, and provide actionable insights. You will work closely
with cross-functional teams to understand business needs and deliver data-driven solutions that
enhance decision-making and drive business growth.
In This Role, Your Responsibilities Will Be:
Analyze large, complex data sets using statistical methods and machine learning
techniques to extract meaningful insights.
Develop and implement predictive models and algorithms to solve business problems
and improve processes.
Create visualizations and dashboards to effectively communicate findings and insights to
stakeholders.
Work with data engineers, product managers, and other team members to understand
business requirements and deliver solutions.
Clean and preprocess data to ensure accuracy and completeness for analysis.
Prepare and present reports on data analysis, model performance, and key metrics to
stakeholders and management.
Participate in regular Scrum events such as Sprint Planning, Sprint Review, and Sprint
Retrospective
Stay updated with the latest industry trends and advancements in data science and
machine learning techniques.
For This Role, You Will Need:
Bachelor’s degree in computer science, Data Science, Statistics, or a related field or a
master's degree or higher is preferred.
Total 5-7 years of industry experience
More than 3 years of experience in a data science or analytics role, with a strong track
record of building and deploying models.
Proficiency in programming languages such as Python or R, and experience with data
manipulation libraries (e.g., pandas, NumPy).
Excellent understanding of Agentic Frameworks like Microsoft Agent Framework.
Experience with NLP, NLG, and Large Language Models Open Source as well as Cloud
based models.
Experience with SQL and NoSQL databases such as MongoDB, Cassandra, Vector
databases
Experience with Dockers, Asynchronous Data Orchestrators, environments etc.
Strong analytical and problem-solving skills, with the ability to work with complex data
sets and extract actionable insights.
Excellent verbal and written communication skills, with the ability to present complex
technical information to non-technical stakeholders.
Preferred Qualifications that Set You Apart:
Prior experience in engineering domain would be nice to have
Prior experience in working with teams in Scaled Agile Framework (SAFe) is nice to
have
Possession of relevant certification/s in data science from reputed universities
specializing in AI.
Familiarity with cloud platforms, Microsoft Azure is preferred
Ability to work in a fast-paced environment and manage multiple projects simultaneously.
Strong analytical and troubleshooting skills, with the ability to resolve issues related to
model performance and infrastructure.
Roles & Responsibilities
- Design and develop intelligent AI-based applications using advanced NLP and LLM techniques to solve real-world business challenges in financial services.
- Build and optimize Retrieval-Augmented Generation (RAG) pipelines leveraging structured and unstructured financial data.
- Integrate and orchestrate LLMs/SLMs for question-answering, summarization, semantic search, and document understanding.
- Develop and maintain RESTful APIs (sync and async) to serve NLP models and chatbot interfaces using frameworks like FastAPI, Flask, etc.
- Should have knowledge of advanced prompting techniques.
- Implement semantic search, hybrid search, and text retrieval systems using Elasticsearch and vector databases (e.g., FAISS, Pinecone, Weaviate).
- Perform NLP tasks such as entity recognition, text classification, intent detection, embedding generation, and sentiment analysis where required.
- Monitor and fine-tune LLM/SLM performance with real-world user data to improve relevance, latency, and accuracy.
- Exposure to LLMOps tools for monitoring, evaluation, and versioning of AI models in production.
- Build, train, and evaluate deep learning models for NLP tasks including classification, NER, summarization, and embedding generation.
- Develop traditional machine learning models (e.g., regression, decision trees, clustering) for structured data analysis and prediction tasks.
- Interact with cross-functional teams to understand system issues and follow up with respective teams to get them fixed.
- Understand and identify areas of improvement across businesses and participate in solution identification and implementation.
- Should be able to work as an Individual Contributor on new and existing projects.
- Positive and problem-solving attitude, must work as an independent contributor.
Ideal Candidate
1.Strong Data Scientist / AI Engineer / Generative AI Engineer profile.
2.Mandatory (Experience 1) - Must have 3+ years of hands-on experience in Data Science, Artificial Intelligence, Machine Learning, Deep Learning, NLP, or Generative AI application development.
3.Mandatory (Experience 2) - Must have strong hands-on experience in Python programming, backend development, API development, and production-grade application support
4.Mandatory (Experience 3) - Must have experience working with Machine Learning and Deep Learning frameworks such as PyTorch, TensorFlow, Keras, or Scikit-learn.
5.Mandatory (Experience 4) - Must have hands-on experience in NLP use cases such as text classification, sentiment analysis, entity recognition (NER), semantic search, embeddings, or document understanding.
6.Mandatory (Experience 5) - Must have experience working with Large Language Models (LLMs) such as GPT, LLaMA, Mistral, Phi, Claude, Gemini, or similar models
.7.Mandatory (Experience 6) - Must have hands-on experience building or implementing Retrieval Augmented Generation (RAG) solutions, vector search, semantic search, or knowledge-based AI applications.
8.Mandatory (Experience 7) - Must have experience with Prompt Engineering and Generative AI frameworks such as LangChain, LangGraph, AI Agents, Azure OpenAI, or similar technologies.
9.Mandatory (Experience 8) - Must have experience developing, consuming, or integrating APIs using Python frameworks such as FastAPI, Flask, or similar technologies.
10.Mandatory (CTC) - The CTC breakup offered will be 75% fixed + 25% variable, as per company policy.
11.Preferred (Experience 1) - Experience with LLMOps/MLOps tools for monitoring, evaluation, experimentation, and versioning of AI models.
12.Preferred (Experience 2) - Exposure to Azure OpenAI, Azure Kubernetes Service (AKS), Kubernetes, cloud-native AI deployments, or distributed systems.
13.Preferred (Experience 3) - Experience working with PostgreSQL, MongoDB, Redis, Kafka, or large-scale data platforms.
14.Preferred (Experience 4) - Familiarity with Docker, Kubernetes, cloud platforms, and scalable deployment architecture.
15.Preferred (Company) - Candidates from AI-first startups, product companies, SaaS organizations, fintech, or data-driven technology companies.
16.Mandatory ( Age ) - Candidate Should be Below 28 Years.
17.Mandatory ( Pedigree) - B.TECH / M.TECH from Tier 1 Colleges (IIT's, NIT's, BITS) are Considered.
About the Role
We are looking for a highly skilled Data Scientist with strong expertise in Machine Learning, MLOps, and Generative AI. The ideal candidate will have hands-on experience in building scalable ML models, deploying them in production, and working with modern AI frameworks, including GenAI technologies.
Key Responsibilities
· Design, develop, and deploy machine learning models for real-world business problems
· Work on end-to-end ML lifecycle: data preprocessing, model building, evaluation, deployment, and monitoring
· Implement and manage MLOps pipelines for scalable and reproducible workflows
· Utilize tools like MLflow for experiment tracking, model versioning, and lifecycle management
· Develop and integrate Generative AI (GenAI) solutions such as LLM-based applications
· Collaborate with cross-functional teams (engineering, product, business) to translate requirements into AI solutions
· Optimize model performance and ensure production stability
· Stay updated with the latest advancements in AI/ML and GenAI ecosystems
Required Skills & Qualifications
· 4+ years of experience in Data Science / Machine Learning
· Strong programming skills in Python
· Hands-on experience with ML modeling techniques (supervised, unsupervised, NLP, etc.)
· Solid understanding of MLOps practices and tools
· Experience with MLflow or similar model lifecycle tools
· Practical experience in Generative AI (GenAI), including working with LLMs
· Experience with libraries/frameworks like Scikit-learn, TensorFlow, PyTorch
· Strong understanding of data structures, algorithms, and statistics
· Experience with cloud platforms (AWS/GCP/Azure) is a plus
Good to Have
· Experience with LLM fine-tuning, prompt engineering, or RAG pipelines
· Exposure to Docker, Kubernetes, and CI/CD pipelines
· Knowledge of data engineering workflows
Role Overview
As a Data Scientist, you will work with business stakeholders, AI engineers, and domain experts to transform data into actionable insights and intelligent solutions. You will develop machine learning models, perform statistical analysis, and contribute to AI-driven products that create measurable business impact.
Key Responsibilities
Data Science & Machine Learning
- Analyze structured and unstructured data to identify patterns, trends, and business opportunities.
- Perform exploratory data analysis (EDA), feature engineering, and data preparation.
- Develop, evaluate, and optimize machine learning models for prediction, classification, clustering, and forecasting.
- Apply statistical techniques to solve business problems and validate model performance.
- Design and execute experiments to improve model accuracy and business outcomes.
AI Solution Development
- Collaborate with AI Engineers, Data Engineers, and domain experts to build AI-powered solutions.
- Translate business requirements into scalable data science approaches.
- Contribute to Generative AI and advanced analytics initiatives where applicable.
- Document methodologies, model performance, and key findings.
Required Technical Skills
- Strong programming skills in Python and SQL for data analysis, feature engineering, and machine learning.
- Strong understanding of Statistics, Probability, Linear Algebra, and Calculus as applied to machine learning and data science.
- Experience with Exploratory Data Analysis (EDA), data preprocessing, feature engineering, feature selection, and handling missing or imbalanced data.
- Good understanding of Supervised, Unsupervised, and Ensemble Machine Learning algorithms, including their assumptions, strengths, limitations, and appropriate use cases.
- Strong knowledge of Regression, Classification, Clustering, Time Series Forecasting, Dimensionality Reduction, Recommendation Systems, and Anomaly Detection techniques.
- Experience with Model Evaluation, Cross-Validation, Hyperparameter Optimization, Bias-Variance Trade-off, Feature Importance, Explainable AI (XAI), and Performance Metrics.
- Understanding of Statistical Inference, Hypothesis Testing, Probability Distributions, Sampling Techniques, Confidence Intervals, and A/B Testing.
- Experience translating business problems into analytical approaches and developing scalable, data-driven solutions.
- Working knowledge of Generative AI, Large Language Models (LLMs), Prompt Engineering, and Retrieval-Augmented Generation (RAG) is preferred.
Preferred Qualifications
- Bachelor's or master's degree in computer science, Artificial Intelligence, Data Science, Statistics, Mathematics, Engineering, or a related field.
- 2–4 years of experience developing machine learning or data science solutions.
- Experience working on end-to-end data science projects in a business environment.
Nice to Have
- Exposure to Generative AI, LLMs, RAG, or Agentic AI.
- Experience with Computer Vision or Natural Language Processing (NLP).
- Familiarity with cloud-based AI platforms.
- Knowledge of construction, engineering, manufacturing, or industrial domains.
- Participation in hackathons, research, Kaggle competitions, or open-source projects.
Soft Skills
Strong analytical and problem-solving skills, effective communication and collaboration, ownership mindset, adaptability, continuous learning, and a passion for innovation.
POSITION OVERVIEW
We are seeking an experienced Senior Data Scientist & Generative AI Specialist on a contractual basis to support a premier Germany-based chemical manufacturing enterprise. In this role, you will lead the end-to-end design, development, and deployment of production-grade GenAI applications, multi-modal LLM workflows, and advanced retrieval platforms tailored to complex industrial and enterprise data ecosystems.
Working closely with cross-functional global teams, you will build robust backend microservices, implement state- of-the-art RAG/GraphRAG architectures, and leverage cloud-native AI infrastructure (Azure, Vector DBs, Knowledge Graphs) to drive operational efficiency and data-driven innovation.
KEY RESPONSIBILITIES
- GenAI & LLM System Engineering: Design, build, and deploy production-grade multi-modal GenAI applications processing text, structured technical documentation, images, and telemetry data.
- Advanced RAG & Graph Architecture: Implement cutting-edge Retrieval-Augmented Generation (RAG) and GraphRAG pipelines using document parsing frameworks, custom embeddings, vector databases, and knowledge graphs to capture complex domain relationships.
- Scalable Backend Development: Architect high-throughput, low-latency microservice APIs using Python, FastAPI, and Flask, leveraging asynchronous programming (asyncio) and strict type validation (Pydantic) for long-running LLM processes.
- Agentic Systems & Azure Ecosystem: Build autonomous agent systems using modern frameworks (MCP, A2A) and orchestrate enterprise workflows across the Microsoft Azure AI ecosystem (Azure AI Foundry, AI Search, Document Intelligence, Databricks).
- Model Optimization & Evaluation: Execute systematic LLM fine-tuning, prompt optimization, and rigorous evaluation frameworks to assess AI output accuracy, reliability, and business impact against industrial requirements.
- Data Layer Management: Architect and maintain enterprise database layers combining SQL (PostgreSQL) for structured transactional data with specialized vector search engines and graph stores.
- Rapid Prototyping: Utilize AI-assisted development tools (Copilot, Claude Code) to accelerate delivery timelines and rapidly build functional UI prototypes for client feedback.
TECHNICAL QUALIFICATIONS
Core Development & Backend:
• Python Mastery: Deep expertise in writing clean, production-ready Python using asynchronous programming (asyncio), strict type-hinting (Pydantic), and automated testing patterns.
• Backend Microservices: Hands-on experience building microservices with FastAPI and Flask structured to handle asynchronous, long-running AI background tasks.
• Database Engineering: Strong command of PostgreSQL, relational schema design, vector indexing, and knowledge graph paradigms.
Machine Learning & AI Infrastructure:
• Model Expertise: Hands-on experience with leading multi-modal LLM architectures (OpenAI, Anthropic, Google) and domain-specific AI workflows.
• Retrieval & Parsing: Proven track record with document extraction frameworks, embedding models, vector search engines, and GraphRAG architectures.
• Cloud Infrastructure: Strong proficiency with Azure AI infrastructure (Foundry, Databricks, AI Search, Document Intelligence).
• Agentic Frameworks: Practical experience with open-source agent protocols (MCP, A2A), parameter-efficient fine-tuning (PEFT/LoRA), and model evaluation methodology.
CONTRACT & REMOTE REQUIREMENTS
• Contract Engagement: Contractual structure tailored to project milestones and deliverables.
• 100% Remote Setup: Fully equipped home office with high-speed, secure internet infrastructure.
• Timezone Overlap: Guaranteed 4-hour daily overlap with Central European Time (CET/CEST - Germany) to ensure smooth collaboration with enterprise stakeholders.
• Communication: Fluent professional English communication skills (written and spoken) for asynchronous and real-time technical coordination.
Job Summary:
We are looking for a skilled Data Scientist with strong expertise in demand forecasting, predictive analytics, and emerging Generative AI technologies. The ideal candidate should have hands-on experience in machine learning, deep learning, NLP, and LLM-based solutions, along with proficiency in Python, SQL, Power BI, and advanced Excel. This role involves building scalable forecasting models and leveraging AI/GenAI to deliver actionable business insights.
Key Responsibilities:
- Develop and deploy demand forecasting models using machine learning and deep learning techniques.
- Analyze historical data to identify trends, seasonality, and demand patterns.
- Build predictive models to improve supply chain and inventory planning.
- Work with large datasets using Python and SQL for data extraction, transformation, and analysis.
- Design dashboards and reports using Power BI for business stakeholders.
- Utilize advanced Excel techniques (Pivot Tables, Power Query, formulas) for analysis and reporting.
- Build and integrate NLP-based solutions for text data analysis and insights.
- Develop and implement LLM-based applications using Generative AI frameworks.
- Design and deploy RAG (Retrieval-Augmented Generation) pipelines for intelligent data retrieval and response generation.
- Collaborate with cross-functional teams (operations, finance, product) to align forecasting and AI solutions.
- Continuously improve model accuracy and performance through experimentation and optimization.
Required Skills:
- Strong proficiency in Python (Pandas, NumPy, Scikit-learn, TensorFlow/PyTorch).
- Solid understanding of machine learning & deep learning algorithms.
- Experience in demand forecasting / time-series analysis (ARIMA, Prophet, LSTM, etc.).
- Hands-on experience with NLP techniques and libraries (NLTK, SpaCy, Transformers).
- Experience working with LLMs and Generative AI frameworks (OpenAI, Hugging Face, LangChain, etc.).
- Strong understanding of RAG architectures and vector databases (FAISS, Pinecone, etc.).
- Advanced knowledge of SQL for data manipulation.
- Hands-on experience with Power BI for visualization and reporting.
- Expertise in advanced Excel (Power Query, dashboards, data modeling).
- Strong analytical and problem-solving skills.
Preferred Qualifications:
- Experience in supply chain, logistics, or e-commerce forecasting.
- Knowledge of cloud platforms (AWS, Azure, or GCP).
- Familiarity with data pipelines and ETL processes.
- Understanding of business metrics and KPIs related to demand planning.






