Pretokenized SFT data for slot-wise agent-context compaction, built from
scatyf3/speccompact-rollouts-deepseekv4
by src/train/scripts/build_llamafactory_slot_sft.py (speculative-compaction
repo, llama-factory branch). Slot targets were generated by a DeepSeek-V4
teacher; sequences are tokenized with the Qwen/Qwen3-0.6B tokenizer (shared
by all Qwen3 sizes), chat template applied with enable_thinking=False.… See the full description on the dataset page:
https://huggingface.co/datasets/sy128/SpecComp-slot-sft-pretokenized-deepseekv4.