Views
No views yet
FastModel + TRL SFT, following the official Unsloth Qwen3.5 fine-tune recipe (canonical for Qwen3.6 too).huihui-ai/Huihui-Qwen3.6-27B-abliterated (multimodal, hybrid-thinking, dense 27B, qwen3_5 architecture)tomvaillant/investigative-journalism-training (687 examples, OSINT methodology)bias="none"; use_gradient_checkpointing="unsloth"load_in_16bit=True; 4-bit QLoRA explicitly not recommended for Qwen3.5/3.6 per Unsloth docs)finetune_vision_layers=False) — text-only LoRA, vision capability preserved byte-identical to basetomvaillant/investigative-journalism-training — 687 instruction/response pairs synthesized by Claude Opus 4.6 (Anthropic) from the Buried Signals OSINT and investigative-journalism corpus: OSINT Navigator tool data, Indicator Media briefings, Buried Signals investigative skills, GIJN, Bellingcat, Verification Handbook 3, SPJ Code of Ethics, RCFP, and public manuals from UNESCO, Al Jazeera Media Institute, CiFAR, CIPE, and EJF/TEMPO Institute.