Views
No views yet
meta-llama/Meta-Llama-3.2-1B trained on all sections from the United States Code of Federal Regulations (CFR). The goal: provide a specialized assistant for navigating and answering questions about U.S. federal regulations.Hardware/Environment:
Training was conducted on Modal using a single NVIDIA H200 GPU.
Training speed: ~1.10 steps/sec, 35 samples/sec.
Note: This loss is typical for a Llama-3 1B model on legal/complex text. For comparison: random output would yield >2.0; perfect memorization of a small dataset would yield <1.0. This is in the “actually learned something useful” range for this setup.
transformers library.