This dataset repo contains outputs from a 50-scenario GT-HarmBench evaluation run for the adapter:
agentic-moral-alignment/qwen35-9b__ipd_str_tft__deont__native_tool__r1
Run settings:
Base model: unsloth/Qwen3.5-9B
Quantization: no_load_in_4bit: true
Prompt mode: bare
Reasoning mode: native
Tool use: true
Reasoning length hint: 100
Max new tokens: 512
Max scenarios: 50
Seed: 100
Files at repo root:
The README YAML frontmatter defines two dataset… See the full description on the dataset page:
https://huggingface.co/datasets/agentic-moral-alignment/gt-harmbench-deont.