allenai / dolma

Data and tools for generating and inspecting OLMo pre-training data.
https://allenai.github.io/dolma/
Apache License 2.0
909 stars 94 forks source link

Disambiguating that the repo is for the dolma toolkit in various docs #104

Closed arnavic closed 7 months ago

arnavic commented 7 months ago

@soldni -- took a stab at some clarifying language for the Dolma docs, can you take a look and see if they make sense?