this post was submitted on 22 Dec 2024
1467 points (97.4% liked)

Technology

60055 readers
3378 users here now

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related content.
  3. Be excellent to each another!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, to ask if your bot can be added please contact us.
  9. Check for duplicates before posting, duplicates may be removed

Approved Bots


founded 2 years ago
MODERATORS
 

It's all made from our data, anyway, so it should be ours to use as we want

you are viewing a single comment's thread
view the rest of the comments
[–] NoForwardslashS@sopuli.xyz 3 points 1 day ago (1 children)

So you're saying the data wouldn't exist anywhere in the source code, but it would still be able to answer questions based on the data it has previously seen?

[–] stephen01king@lemmy.zip 16 points 1 day ago (1 children)

That is how LLM works, they don't store the data as data, but as weight values.

[–] NoForwardslashS@sopuli.xyz 1 points 23 hours ago (1 children)

So then why, if it were all open sourced, including the weights, would the AI be worthless? Surely having an identical but open source version, that would strip profitability from the original paid product.

[–] Bronzebeard@lemm.ee 3 points 21 hours ago (1 children)

It wouldn't be. It would still work. It just wouldn't be exclusively available to the group that created it-any competitive advantage is lost.

But all of this ignores the real issue - you're not really punishing the use of unauthorized data. Those who owned that data are still harmed by this.

[–] stephen01king@lemmy.zip 2 points 20 hours ago (1 children)

It does discourages the use of unauthorised data. If stealing doesn't give you competitive advantage, it's not really worth the risk and cost of stealing it in the first place.

[–] Bronzebeard@lemm.ee 1 points 7 hours ago (1 children)

If you can still use it after you stole it, as opposed to not being able to use it at all... Then it does give you an incentive

[–] stephen01king@lemmy.zip 1 points 6 hours ago

If you did all the work and potentially criminal collection of data, but everyone else gets the benefit as well, that is not an incentive. You underestimate how selfish corporations can be.

OpenAI wouldn't stay at the forefront of LLM if every competitor gets to use the model they spent money on training.