This repository consolidates the experimental resources associated with the paper:
Reinforcement Learning With Verifier Guidance and Penalty Shaping for Vietnamese Summarization Using Small Language Models
It contains:
CSV exports for Hugging Face Data Viewer,
Links to the released best checkpoints,
The link to the frozen evaluator MultiEvalSumViet2.
Representative Source Code