The GNU Compiler Collection will reject copyright-significant contributions derived from LLM output, with limited exceptions for trivial changes and test cases.
I don’t understand specifically the copyright issue with respect to “trained on x data”
It’s a question of whether the output of the LLM is a derivative work of the training data. From a copyright standpoint, if the work is derivative (and not fair use), then it needs to follow the terms set by any licenses you have to its parent works.
It’s a question of whether the output of the LLM is a derivative work of the training data. From a copyright standpoint, if the work is derivative (and not fair use), then it needs to follow the terms set by any licenses you have to its parent works.