This repository contains the evaluation criteria, definitions, and scales for Hubble_S's LLM-powered chatbot and recommendation system.
Our evaluation framework consists of two stages:
Quality Evaluation: Assesses must-haves and quality criteria for both the chatbot and recommendation system.
Packaging Evaluation: Separate criteria for chatbot and recommendations (emails).