issues
search
allenai
/
dolma
Data and tools for generating and inspecting OLMo pre-training data.
https://allenai.github.io/dolma/
Apache License 2.0
972
stars
107
forks
source link
Allow specifying different bins for visualization and computation.
#190
Open
soldni
opened
2 months ago