Views
No views yet
LlamaForCausalLM1.7Bbfloat16<assistant> + 35 <charter_X.Y> tokens), vocab size 49280reflection_1p column from the annotated sidecar dataset. The training augments standard autoregressive NTP with:<assistant> token position trains the model to predict 35 charter items| Name | Use case |
|---|---|
default | Standard SFT — plain assistant role token |
epe | Activates constitution head — uses <assistant> (token 49152) at start of assistant turns |
1tok.apply_chat_template(messages, chat_template="default") # standard
2tok.apply_chat_template(messages, chat_template="epe") # constitution head activevocab_size=49280 is set in config.json