Views
No views yet
unsloth/Llama-3.1-8B-Instruct via TIES-Merge (sign-consensus + magnitude-pruning).KayaTechAI/SFT-Llama-3.1-8B-Financial-Instruct (LoRA adapter)
base_model_name_or_path: KayaTechAI/Merged-IPT-Llama-3.1-8B-Financial-Instruct
KayaTechAI/SFT-Llama-3.1-8B-Financial-Instruct-Sentiment (LoRA adapter)
base_model_name_or_path: KayaTechAI/SLERP-IPT-Llama-3.1-8B-Financial-Instruct
(itself a mergekit SLERP of KayaTechAI/Merged-IPT-Llama-3.1-8B-Financial-Instruct + unsloth/Meta-Llama-3.1-8B-Instruct)unsloth/Llama-3.1-8B-Instruct, which would have been numerically invalid
since LoRA deltas are only meaningful relative to the exact base they were
fit against). Only after baking were the two resulting full checkpoints
delta-merged onto unsloth/Llama-3.1-8B-Instruct.PeftModel.merge_and_unload() for each adapter against its
verified base, producing two full dense checkpoints.unsloth/Llama-3.1-8B-Instruct.
KayaTechAI/sp500-summary-dataset,
KayaTechAI/FIN-Sentiment-SFT) before full production rollout.