Views
No views yet
✅ Overall Evaluation Conclusion: After balancing multiple dimensions (performance, reasoning quality, resource consumption), the 4B-parameter base model demonstrates the best overall performance and is recommended as the default choice.
| Model Type | Parameters | Description |
|---|---|---|
| Base Models | 0.6B,1.7B,4B,8B,14B | Standard text generation models for general dialogue and instruction following |
| Thinking Model | 14B | Enables "Chain-of-Thought" capability, suitable for complex reasoning tasks |
| Embedding Model | 0.6B | Used for vector retrieval in RAG (sentence embedding) |
| Reranker Model | 0.6B | Used for re-ranking in RAG (cross-encoder style reranking) |
rank (r): 8alpha: 256target_modules: ["q_proj", "k_proj", "v_proj", "o_proj", "gate_proj", "up_proj", "down_proj"]transformers + peft + bitsandbytes