I have used the new template below to show how it looks. We can add more details to the linked CONTRIBUTING.md in the pop repo.
I have not included any LLM (also known as AI) generated content in...
IBM Granite (EDIT: and Apertus) do disclose all their training data and claim that all their training data is effectively free of copyright (highly permissively licensed). I have not been able to verify that, due to my lack of skills with the conventions and tools of LLM / Agent training and publishing.
The Apertus Swiss AI unfortunately doesn’t seem to live up to its claims, and they have been silent on the issue brought up there. I honestly suspect the same of IBM’s Granite, but have not investigated.
Thank you for the link! It does look like Apertus itself might be Free Software (the U.S. copyright office says training can infringe, but is usually fair use), but it can still output derivative works of copyrighted inputs that might prevent them from being distributed as-is (for example, requiring attribution) – at all, much less under a strong copyleft.
IBM Granite (EDIT:
and Apertus) do disclose all their training data and claim that all their training data is effectively free of copyright (highly permissively licensed). I have not been able to verify that, due to my lack of skills with the conventions and tools of LLM / Agent training and publishing.So, yeah, probably (EDIT:
twoone).for those who said the deets, thank you!
The Apertus Swiss AI unfortunately doesn’t seem to live up to its claims, and they have been silent on the issue brought up there. I honestly suspect the same of IBM’s Granite, but have not investigated.
Thank you for the link! It does look like Apertus itself might be Free Software (the U.S. copyright office says training can infringe, but is usually fair use), but it can still output derivative works of copyrighted inputs that might prevent them from being distributed as-is (for example, requiring attribution) – at all, much less under a strong copyleft.