this post was submitted on 13 Jul 2026
614 points (98.3% liked)

Technology

87685 readers
2726 users here now

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related news or articles.
  3. Be excellent to each other!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, this includes using AI responses and summaries. To ask if your bot can be added please contact a mod.
  9. Check for duplicates before posting, duplicates may be removed
  10. Accounts 7 days and younger will have their posts automatically removed.

Approved Bots


founded 3 years ago
MODERATORS
 

For years, tech giants have argued that if information is available on the internet, it can be used for AI model development and outputs. They call it fair use. Content owners have tried to prevent this, with no success.

Now Anthropic, OpenAI, and Google are discovering what the rest of the internet has already learned through painful experience: once you put something online, people will find ways to use it in ways you don't like and can't stop.

The latest flashpoint is something called "distillation," using the outputs of one AI model to improve another. Anthropic says competitors are harvesting its outputs at scale, turning billions of dollars of research into a shortcut for rivals. OpenAI and Google have made similar warnings recently.

Remove Paywalls Link

you are viewing a single comment's thread
view the rest of the comments
[–] uriel238@lemmy.blahaj.zone 46 points 1 month ago (9 children)

This is how we know AI should be a collectivist project, one that isn't owned by large corporations but is funded by taxes and developed in academies, and all IP derived from it falls into the public domain.

Besides which a lot of artists mind a lot less when their material is borrowed by a non-profit, or to serve a public works project. (There are exceptions. Disney is notoriously litigious about murals in nurseries.)

PS: Development of a robust public domain is the only reason that intellectual property should exist at all. Also it's not property so much as a licensed temporary monopoly.

PSS: History has already shown us that people will invent stuff and do fabulous art simply by being allowed to live in a state other than desperation. Public welfare programs beget art booms. The most recent example of this was during the COVID-19 lockdown which came with extended unemployment and stimulus checks, resulting in the Great Resignation in which a lot of people turned their hobbies into something lucrative.

[–] qaz@lemmy.world 19 points 1 month ago (4 children)

I agree. The worst part about GitHub training LLM's on my FOSS code without permission for me is that they then keep the models to themselves. Like if you're going to use all my code without permission, at least allow me to run the model locally.

My personal opinion is that all models trained on copyleft code should be open-weights, most FOSS licenses didn't account for this specific possibility, but this is the only way to follow them in spirit.

[–] vala@lemmy.dbzer0.com 1 points 1 month ago

Not only should the models be made open, their output should inherently be GPL license since it's the only real way to avoid violating GPL (which they are doing now IMO).

load more comments (3 replies)
load more comments (7 replies)