Views
No views yet
1from transformers import pipeline
2
3text = "The capital of France is Paris."
4rewarder = pipeline(model="narcolepticchicken/trace-reward-model-qwen3-4b-v1", device="cuda")
5output = rewarder(text)[0]
6print(output["score"])1@software{vonwerra2020trl,
2 title = {{TRL: Transformers Reinforcement Learning}},
3 author = {von Werra, Leandro and Belkada, Younes and Tunstall, Lewis and Beeching, Edward and Thrush, Tristan and Lambert, Nathan and Huang, Shengyi and Rasul, Kashif and Gallouédec, Quentin},
4 license = {Apache-2.0},
5 url = {https://github.com/huggingface/trl},
6 year = {2020}
7}1from transformers import AutoModelForCausalLM, AutoTokenizer
2
3model_id = 'narcolepticchicken/trace-reward-model-qwen3-4b-v1'
4tokenizer = AutoTokenizer.from_pretrained(model_id)
5model = AutoModelForCausalLM.from_pretrained(model_id)AutoModelForCausalLM with the appropriate AutoModel class.