MoreThought/Fable-5.1-Max-Reasoning-Filtered-10000x Viewer • Updated about 2 hours ago • 15k • 5.17k • 295
SLMFineTrain Collection A collection of datasets made to train general-purpose SLMs under 200M parameters • 6 items • Updated 3 days ago • 1
SLMFineTrain Collection A collection of datasets made to train general-purpose SLMs under 200M parameters • 6 items • Updated 3 days ago • 1
OpenFineTrain Collection Very fine 55 trillion+ tokens worth of ungated huggingface datasets for LLM training across 80+ datasets. • 92 items • Updated 3 days ago • 11