allenai / dolma

Data and tools for generating and inspecting OLMo pre-training data.
https://allenai.github.io/dolma/
Apache License 2.0
894 stars 90 forks source link

[EXPERIMENT ONLY, NOT FOR MERGING] Exporting First 200 Text #159

Closed power10dan closed 1 month ago

undfined commented 1 month ago

Obsolete, closing.