loubnabnl HF staff commited on
Commit
5cee0f2
1 Parent(s): 9299928

Update README.md

Browse files
Files changed (1) hide show
  1. README.md +1 -4
README.md CHANGED
@@ -8,10 +8,7 @@ pinned: false
8
  ---
9
 
10
  # HuggingFaceTB
11
- This is the home for smol models (SmolLM) and high quality pre-training datasets.
12
-
13
-
14
- We released:
15
 
16
  - [FineWeb-Edu](https://huggingface.co/datasets/HuggingFaceFW/fineweb-edu): a filtered version of FineWeb dataset for educational content, paper available [here](https://huggingface.co/papers/2406.17557).
17
  - [Cosmopedia](https://huggingface.co/datasets/HuggingFaceTB/cosmopedia): the largest open synthetic dataset, with 25B tokens and more than 30M samples. It contains synthetic textbooks, blog posts, stories, posts, and WikiHow articles generated by Mixtral-8x7B-Instruct-v0.1. Blog post available [here](https://huggingface.co/blog/cosmopedia).
 
8
  ---
9
 
10
  # HuggingFaceTB
11
+ This is the home for smol models (SmolLM) and high quality pre-training datasets. We released:
 
 
 
12
 
13
  - [FineWeb-Edu](https://huggingface.co/datasets/HuggingFaceFW/fineweb-edu): a filtered version of FineWeb dataset for educational content, paper available [here](https://huggingface.co/papers/2406.17557).
14
  - [Cosmopedia](https://huggingface.co/datasets/HuggingFaceTB/cosmopedia): the largest open synthetic dataset, with 25B tokens and more than 30M samples. It contains synthetic textbooks, blog posts, stories, posts, and WikiHow articles generated by Mixtral-8x7B-Instruct-v0.1. Blog post available [here](https://huggingface.co/blog/cosmopedia).