This is branch of
https://huggingface.co/datasets/ekacare/NidaanKosha-100k-V1.0 .Here the files have been classified on the basis of 15 categories that clusters all the different varied tests, so as to facilitate the ease of working on this for ML tasks :
cbc_keywords = [
"haemoglobin", "hemoglobin", "hb", "rbc", "red blood cell", "wbc", "white blood cell", "tlc", "total leucocyte",
"platelet", "pcv", "hematocrit", "hct"… See the full description on the dataset page:
https://huggingface.co/datasets/shreyanbr/ekacare-NidaanKosha-100k-segmented.