The Hosted Inference API Breaks it, I haven't figured out a way to limit its responses so its hard capped at 512 in the Generation_Config.json file. Just change that back to 1024 and you are good
so if someone knows please send a pull request or a edit or something!
Direct Use
Just use it like you would usually for DialoGPT-small.