The KL3M Data Project: Copyright-Clean Training Resources for Large Language Models

paper
First page of The KL3M Data Project: Copyright-Clean Training Resources for Large Language Models
authors:Bommarito II, M. J., Bommarito, J., & Katz, D. M.
year:2025
venue:arXiv preprint
details:arXiv preprint arXiv:2504.07854

pdf preview

related projects and datasets

citation

Bommarito II, M. J., Bommarito, J., & Katz, D. M. (2025). The KL3M Data Project: Copyright-Clean Training Resources for Large Language Models. arXiv preprint. arXiv preprint arXiv:2504.07854.