Beta
Explore
Marketplace
Neural Labs
Playground
Wallet
Docs
code-alpaca-instruct-unfiltered – Dataset by ewof | AlphaNeural AI
You can deploy this model and start earning money today!
ewof
/
code-alpaca-instruct-unfiltered
like
0
10K<n<100K
json
text
datasets
pandas
mlcroissant
polars
us
Views
No views yet
Model card
Files and Versions
Community
API
This dataset is HuggingFaceH4/CodeAlpaca_20K unfiltered, removing 36 instances of blatant alignment. 19986 instructions remain.
https://huggingface.co/datasets/HuggingFaceH4/CodeAlpaca_20K/blob/29ba7b7fdf0c55e5435c848cf6bbf9782fef62a6/data/test-00000-of-00001.parquet
https://huggingface.co/datasets/HuggingFaceH4/CodeAlpaca_20K/blob/a123ae447f02484d83c3457438b4422cd8417ad5/data/train-00000-of-00001.parquet
i combined all of these files above into code_alpaca_data.jsonl with parquet2json and ran… See the full description on the dataset page:
https://huggingface.co/datasets/ewof/code-alpaca-instruct-unfiltered
.