Updated • 6.37k
• 196
Viewer
• Updated • 170M • 23.3k
• 90
Viewer
• Updated • 621M • 11.9k
• 87
Locutusque/UltraTextbooks
Viewer
• Updated • 5.52M • 673
• 198
PrimeIntellect/StackV1-popular
Viewer
• Updated • 93M • 914
• 2
Viewer
• Updated • 11.7M • 49
• 5
EleutherAI/the_pile_deduplicated
Viewer
• Updated • 134M • 21.3k
• 110
HIT-TMG/KaLM-embedding-pretrain-data
Viewer
• Updated • 23.7M • 1.82k
• 20
suriyagunasekar/stackoverflow-with-meta-data
Viewer
• Updated • 19.9M • 218
• 12
Viewer
• Updated • 13.6M • 1.13k
• 5
Viewer
• Updated • 3.71M • 1.02M
• 650
Viewer
• Updated • 474M • 64
• 4
EleutherAI/deep-ignorance-annealing-mix
Viewer
• Updated • 89M • 72
• 1
Viewer
• Updated • 10.2M • 50
• 5
Viewer
• Updated • 1.76M • 21.1k
• 403
Viewer
• Updated • 167M • 3.83k
• 68
Locutusque/deeplm-training-data
Viewer
• Updated • 2.17M • 133
• 3
nvidia/Llama-Nemotron-Post-Training-Dataset
Viewer
• Updated • 3.91M • 4.03k
• 645
Updated • 52.2k
• 248
EssentialAI/essential-web-v1.0
Preview
• Updated • 47.2k
• 218