This dataset contains full benchmark results (aggregate scores + per-sample prompts and model responses) for evaluating the official CodeV paper fine-tune yang-z/CodeV-QC-7B on the VerilogEval v1 benchmark.
This is an independent reproduction of the "CodeV-QC" row in Table III of the CodeV paper (arXiv:2407.10424), on the same Qwen2.5-Coder-7B base used by my own fine-tune. Four out of six metrics exceed… See the full description on the dataset page:
https://huggingface.co/datasets/muratkarahan/verilogeval-v1-codev-qc-7b-results.