Codeberg Proposes ToS Update to Ban LLM Training Data Extraction
Codeberg, an open-source code hosting platform, is considering an extension to its Terms of Use that would explicitly prohibit the extraction of data for training large language models. The proposed change is outlined in a pull request on Codeberg's own repository, inviting community review and discussion. The move reflects growing concern among open-source communities about AI companies scraping hosted code without contributor consent. If adopted, the updated terms would add a new layer of protection for developers who host their projects on the platform.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.
Discussion (0)
Log in to join the discussion and vote.
Log in