Views
No views yet
>=5.5.0 for qwen3_5_moe support. vLLM images that ship older Transformers may fail unless configured to use a new enough Transformers backend.model.safetensors.index.json.FP16.TaimoorSiddiqui/HopCoder-Mini-35B-A3B-VL36; no embedding resize or tensor-shape variation is introduced.tokenizer_config.json and chat_template.jinja. Tool calls are trained and documented as JSON inside <tool_call> tags:1<tool_call>
2{"name":"tool_name","arguments":{"argument_name":"value"}}
3</tool_call><tool_response> turn, clients should continue generation until the assistant provides a final answer.