Qwen/Qwen3-Embedding-4B.bge-m3-ko-h100, it is the better model to highlight when you want the strongest general Korean retrieval story.Qwen/Qwen3-Embedding-4B with a contrastive (in-batch-negative) objective.Qwen/Qwen3-Embedding-4Bnvidia/Nemotron-Personas-Korea — synthetic Korean personas (demographic / geographic / personality), used as (query, positive-document) pairssentence-transformers MultipleNegativesRankingLoss (in-batch negatives, InfoNCE-style)111677_20260506_114341_both_2gpuMIRACLRetrievalko| cutoff | Precision | Recall | F1 | mAP | mRR | NDCG |
|---|---|---|---|---|---|---|
| @1 | 0.46479 | 0.29452 | 0.32978 | 0.29452 | 0.464789 | 0.46479 |
| @3 | 0.25665 | 0.42305 | 0.28522 | 0.42305 | 0.577318 | 0.45190 |
| @5 | 0.18310 | 0.47493 | 0.23726 | 0.47493 | 0.577318 | 0.45852 |
| @10 | 0.11737 | 0.58359 | 0.17922 | 0.58359 | 0.577318 | 0.49227 |
| @20 | 0.07254 | 0.67332 | 0.12345 | 0.67332 | 0.577318 | 0.52325 |
| @100 | 0.02028 | 0.83695 | 0.03891 | 0.83695 | 0.577318 | 0.56629 |
| @1000 | 0.00246 | 0.95250 | 0.00491 | 0.95250 | 0.577318 | 0.58808 |
output/111677_20260506_114341_both_2gpu/qwenbenchmark_results/autorag_benchmark.jsonbenchmark_results/qwen_miracl_fast4/miracl_benchmark.txtbenchmark_results/qwen_miracl_fast4/miracl_benchmark.jsondragonkue/snowflake-arctic-embed-l-v2.0-ko 0.740433, dragonkue/BGE-m3-ko 0.729993, nlpai-lab/KURE-v1 0.727739, and nlpai-lab/KoE5 0.711356 on the Korean retrieval leaderboard cited by the model card0.7484