allenai / dolma

Data and tools for generating and inspecting OLMo pre-training data.
https://allenai.github.io/dolma/
Apache License 2.0
909 stars 94 forks source link

Fix issue in getting started tutorial using wikipedia data #117

Closed RohitRathore1 closed 5 months ago

RohitRathore1 commented 6 months ago

Hello! Thank you for your PR. Left a comment to explain the changes. A better title for this PR would also help.

Hi, sorry for the delay in reply. I had missed this.