view the rest of the comments
LocalLLaMA
Welcome to LocalLLaMA! Here we discuss running and developing machine learning models at home. Lets explore cutting edge open source neural network technology together.
Get support from the community! Ask questions, share prompts, discuss benchmarks, get hyped at the latest and greatest model releases! Enjoy talking about our awesome hobby.
As ambassadors of the self-hosting machine learning community, we strive to support each other and share our enthusiasm in a positive constructive way.
Rules:
Rule 1 - No harassment or personal character attacks of community members. I.E no namecalling, no generalizing entire groups of people that make up our community, no baseless personal insults.
Rule 2 - No comparing artificial intelligence/machine learning models to cryptocurrency. I.E no comparing the usefulness of models to that of NFTs, no comparing the resource usage required to train a model is anything close to maintaining a blockchain/ mining for crypto, no implying its just a fad/bubble that will leave people with nothing of value when it burst.
Rule 3 - No comparing artificial intelligence/machine learning to simple text prediction algorithms. I.E statements such as "llms are basically just simple text predictions like what your phone keyboard autocorrect uses, and they're still using the same algorithms since <over 10 years ago>.
Rule 4 - No implying that models are devoid of purpose or potential for enriching peoples lives.
I was trying to reply to this, bringing up the controversy that erupted when image generators were able to use prompts with artist names to simulate their style, and how that shows how blurry the copyright line is. As I wrote it, I realized how complicated the issue really is, and how little I understand current copyright laws. Some may argue there isn't any vagueness, an artist's work in any format is theirs and shouldn't be copied for other use, training or otherwise. But we all use other people's work in our own, don't we? There is no solid line where one is ethical and one isn't. Plagiarism without proper attribution has always been wrong, but if you take the same content and tweak it a bit, it's not. LLMs can and do that, through statistical randomness, while using someone else's material, just like human brains would.
The compensation issue is definitely a problem and where the ethics come into play. Even if many people freely read a book in a library, it has been bought at some point. But how do you determine who was unfairly "stolen" from? Isn't it basically everyone at this point?
Compensation is actually the only issue. If they were fairly paid, most artists would be ok with people enjoying their works in any way they can.
I know I'm way too nitpicky... But I feel I should point out there's some confusion hidden here...
There's stuff which is moral to do. And stuff which is legal to do. Those aren't the same! When talking about who can mention what, what copyright law demands people to do and those blurry lines, we're concerned with legality. That doesn't have anything to do with ethics. (At least not directly.)