this post was submitted on 16 May 2026
130 points (97.8% liked)

Fuck AI

7560 readers
2389 users here now

"We did it, Patrick! We made a technological breakthrough!"

A place for all those who loathe AI to discuss things, post articles, and ridicule the AI hype. Proud supporter of working people. And proud booer of SXSW 2024.

AI, in this case, refers to LLMs, GPT technology, and anything listed as "AI" meant to increase market valuations.

founded 2 years ago
MODERATORS
top 14 comments
sorted by: hot top controversial new old
[–] cypherpunks@lemmy.ml 30 points 2 months ago (4 children)

identify AI that has used copyrighted material

but, that is basically all modern "AI".

(the only LLM i've heard of which actually claims that its training corpus is freely licensed is Apertus...)

[–] youcantreadthis@quokk.au 8 points 2 months ago* (last edited 2 months ago)

We callin it Plagarized Information Stochastic Stupidity now the only PISS you've heard of

[–] Hackworth@piefed.ca 4 points 2 months ago (1 children)

Adobe claims to only train their image generator, Firefly, on images from their stock library.

[–] cloudskater@piefed.blahaj.zone 3 points 2 months ago (1 children)

Even if it didn't use copyrighted stuff, the concept of "generative" AI is fascist to its very core.

[–] HaraldvonBlauzahn@feddit.org 1 points 2 months ago

the concept of “generative” AI is fascist to its very core.

Can you explain? I might miss the connection.

[–] YourMomsTrashman@lemmy.world 2 points 2 months ago (1 children)

Traditionally, with machine learning, it is standard practice to mention what datasets and/or pretrains were used, so that the results are transparent and can be replicated. With GPT-2, it was "the common crawl and our own crawled 8 million web pages", and since then I feel it's mostly left out, falling back on (easily manipulated) benchmarks instead 😬

[–] cypherpunks@lemmy.ml 2 points 2 months ago* (last edited 2 months ago) (1 children)

Yep. But just providing a list of millions of URLs and saying "we trained on this" as some models in the past have done also didn't make it possible to replicate; by the time anyone re-fetches them all, many of the URLs will inevitably have changed or disappeared.

[–] YourMomsTrashman@lemmy.world 1 points 2 months ago

That's exactly why projects like the common crawl exist though !

[–] very_well_lost@lemmy.world 14 points 2 months ago* (last edited 2 months ago) (1 children)

People have actually been doing this to catch plagiarism for centuries, long before LLMs were a thing.

See trap streets for one of the better known examples.

[–] LodeMike@lemmy.today 3 points 2 months ago

I learned about that from Doctor Who!

[–] sidefaceturdtalker@leminal.space 4 points 2 months ago (1 children)

One way to push back forsure but we need to refresh the tree of liberty asap

[–] ZDL@lazysoci.al 1 points 2 months ago

Interestingly, literally zero of the people I've seen who word things this way ever seem to volunteer to be the ones doing the watering. Are you going to break the losing streak or are you going to continue confirming my belief that it's only chicken hawks who say this?

[–] ExtremeDullard@piefed.social 2 points 2 months ago* (last edited 2 months ago)