Views
No views yet
| Metric | Value |
|---|---|
| Training Step | 571 |
| Epoch | 1 |
1vllm serve llm-model-lab/thinking-v2-3epoch \
2 --enable-lora \
3 --lora-modules epoch1=llm-model-lab/gemma3-27b-orpo-thinking-v2-checkpoint-571 \
4 --max-lora-rank 128 \
5 --dtype bfloat16adapter_model.safetensors: LoRA weightsadapter_config.json: PEFT configurationtokenizer.json, tokenizer_config.json: Tokenizerchat_template.jinja: Gemma3 chat template