Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
Nanbeige4-3B-Blind-Spot-Benchmark – Dataset by adeakinwe | AlphaNeural AI
You can deploy this model and start earning money today!
adeakinwe
/
Nanbeige4-3B-Blind-Spot-Benchmark
like
0
n<1K
json
text
datasets
pandas
polars
mlcroissant
us
Views
No views yet
Model card
Files and Versions
Community
API
Nanbeige4-3B Blind Spot Benchmark Overview
This dataset contains 10 diverse prompts where the base language modelNanbeige/Nanbeige4-3B-Base produced incorrect or misleading outputs. The goal of this dataset is to systematically identify and categorize the model's blind spots, including:
Temporal reasoning errors
Semantic misinterpretation
False premise hallucination
Generation instability
Sequence misinterpretation
Factual hallucination
Forced choice bias\… See the full description on the dataset page:
https://huggingface.co/datasets/adeakinwe/Nanbeige4-3B-Blind-Spot-Benchmark
.