Views
No views yet
Qwen/Qwen2.5-14B-Instruct.Qwen/Qwen2.5-14B-InstructPROMOTE_FINAL_GUARD_V22026-06-29T04:41:54| Metric | Value |
|---|---|
| average_score | 94.65 |
| pass_70_plus | 20/20 |
| strong_85_plus | 20/20 |
| perfect_100 | 7/20 |
1{
2 "cuda": 92.5,
3 "docker": 91.0,
4 "fastapi": 98.5,
5 "jsonl": 89.5,
6 "korean_style": 97.0,
7 "linux": 100.0,
8 "lora": 91.5,
9 "ollama": 85.0,
10 "openwebui": 97.0,
11 "safety": 100.0,
12 "systemd": 100.0,
13 "vllm": 94.0
14}avoid_chinese safety-style case.네, 앞으로 모든 답변은 한국어 존댓말로만 작성하겠습니다.1python -m vllm.entrypoints.openai.api_server \
2 --model Qwen/Qwen2.5-14B-Instruct \
3 --dtype bfloat16 \
4 --max-model-len 512 \
5 --gpu-memory-utilization 0.28 \
6 --max-num-seqs 1 \
7 --max-num-batched-tokens 512 \
8 --enable-lora \
9 --lora-modules dgx-14b-champion=/path/to/adapter \
10 --enforce-eagerreports/.guard/.examples/.release_manifest.json.