allenai / dolma

Data and tools for generating and inspecting OLMo pre-training data.
https://allenai.github.io/dolma/
Apache License 2.0
910 stars 95 forks source link

Add attribute correlations #68

Closed Muennighoff closed 10 months ago