m1: Unleash the Potential of Test-Time Scaling for Medical Reasoning in Large Language Models
A simple test-time scaling strategy, with minimal fine-tuning, can unlock strong medical reasoning within large language models.
Hi! Welcome to the huggingface repository for m1 (Github, Paper)!
m1 is a medical LLM designed to enhance reasoning through efficient test-time scaling. It enables lightweight models to match or exceed the performance of much larger… See the full description on the dataset page:
https://huggingface.co/datasets/UCSC-VLAA/m1k-tokenized.