17
How much can AI companies control the AI?
(sh.itjust.works)
There is no such thing as a Stupid Question!
Don't be embarrassed of your curiosity; everyone has questions that they may feel uncomfortable asking certain people, so this place gives you a nice area not to be judged about asking it. Everyone here is willing to help.
Reminder that the rules for lemmy.ca still apply!
Thanks for reading all of this, even if you didn't read all of this, and your eye started somewhere else, have a watermelon slice 🍉.
To answer the other side of this: LLMs are the sum of their parts and weights. If the weighting algorithm is closed, they could very easily have encoded preferences into it. If the data it is trained on is skewed, that will also show up in the results.
We regularly hear how LLMs are “trained on the Internet”. But we know that’s not really the case; they’re trained on specific Reddit subreddits, collections of books, collections of images, stack exchange, and a number of other unnamed sources. In the case of Gemini, I’m pretty sure it’s trained on Google’s search index data itself.
So which sources are part of the core model and what weight they’re given in final results will definitely impact how the model responds to a prompt.