This dataset is a reformatted version of the original HackMentor research dataset, structured specifically in ChatML format for seamless integration with SFTTrainer and ChatML-based instruct models.
Original Paper: HackMentor: Fine-tuning Large Language Models for Cybersecurity
Format: ChatML (messages column containing role and content keys) / Converted from ShareGPT format. (90%, 10% and 316 test samples)
Task: Fine-tuning… See the full description on the dataset page:
https://huggingface.co/datasets/madox81/HackMentorDS_Chat.