This dataset is a combination of Stanford's Alpaca (
https://github.com/tatsu-lab/stanford_alpaca) and FiQA (
https://sites.google.com/view/fiqa/) with another 1.3k pairs custom generated using GPT3.5
Script for tuning through Kaggle's (
https://www.kaggle.com) free resources using PEFT/LoRa:
https://www.kaggle.com/code/gbhacker23/wealth-alpaca-lora
GitHub repo with performance analyses, training and data generation scripts, and inference notebooks:
https://github.com/gaurangbharti1/wealth-alpaca… See the full description on the dataset page:
https://huggingface.co/datasets/xxkkkkkk/finance-alpaca.