87,782 context-grounded question-answer pairs scraped and generated
from the official Stevens Institute of Technology website, used to
fine-tune LLaMA-2-7B for domain-specific academic advising.
context: Source URL + scraped web content
question: Natural language question answerable from context
answer: Concise, grounded answer
Total pairs: 87,782… See the full description on the dataset page:
https://huggingface.co/datasets/chauben/stevens-qa-finetuning-87k.