Hugging Face
Models
Datasets
Spaces
Posts
Docs
Enterprise
Pricing
Log In
Sign Up
mpasila
's Collections
Finnish fine-tunes
Japanese2English datasets
ExLlamaV2 quantizations
Finnish Instruct Datasets
Pre-training dataset prep
Magnum used datasets
Pre-training dataset prep
updated
27 days ago
Some datasets I should probably use.
Upvote
-
JeanKaddour/minipile
Viewer
•
Updated
Jun 20, 2023
•
1.01M
•
1.67k
•
115
wikimedia/wikipedia
Viewer
•
Updated
Jan 9
•
61.6M
•
58.5k
•
609
neuralwork/arxiver
Viewer
•
Updated
21 days ago
•
63.4k
•
3.77k
•
349
ohsuz/tiny-textbooks-edu
Viewer
•
Updated
Jun 11
•
3.31M
•
48
•
1
ohsuz/tiny-code-textbooks-edu
Viewer
•
Updated
Jun 11
•
1.84M
•
51
•
2
Upvote
-
Share collection
View history
Collection guide
Browse collections