Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
introspect-ai-benchmark – Dataset by Aurther-Nadeem | AlphaNeural AI
You can deploy this model and start earning money today!
Aurther-Nadeem
/
introspect-ai-benchmark
like
0
text-generation
en
mit
100K<n<1M
json
tabular
text
datasets
pandas
polars
mlcroissant
us
introspection
llm-evaluation
activation-steering
mechanistic-interpretability
Views
No views yet
Model card
Files and Versions
Community
API
IntrospectAI Benchmark Dataset
Empirical benchmark for measuring introspection in Large Language Models through activation steering experiments.
Dataset Description
This dataset contains trial results from 4 experiments (A, B, D, E) testing whether LLMs can monitor, report, and control their own internal states.
Experiments
Experiment Name Question
detection Injected Thoughts Can the model detect when an external thought is injected?
attribution… See the full description on the dataset page:
https://huggingface.co/datasets/Aurther-Nadeem/introspect-ai-benchmark
.