Views
No views yet
Qwen/Qwen3.5-4B instead of SmolLM2-360M-Instruct.user, responder (list of conversational phrases), and responder_thoughts (list of knowledge chunks or <|sil|>).1e-5, 500 warmup steps, cosine scheduler, added <|sil|> token, per-phrase prompt/completion expansion, completion/assistant-only SFT loss, Trackio logging, push to Hub.convfill_qwen35_repro.py.