lemm.ee

Local All Communities Log in Sign up

Local All Communities

1.7K

Stack Overflow bans users en masse for rebelling against OpenAI partnership — users banned for deleting answers to prevent them being used to train ChatGPT(www.tomshardware.com)

posted 2 months ago

by

misk@sopuli.xyz

in

technology@lemmy.world

Sort:

Hot Top Controversial New Old

You are viewing a single thread.

View all comments View context

[ +- ]

Thomas@discuss.tchncs.de

78 points

2 months ago

Those would be harvested to train LLMs even without asking first. 😐

report

reply

[ +- ]

sramder@lemmy.world

45 points

2 months ago

At this point I’m assuming most if not all of these content deals are essentially retroactive. They already scrapped the content and found it useful enough to try and secure future use, or at least exclude competitors.

report

reply

[ +- ]

rickyrigatoni@lemm.ee

13 points

2 months ago

They scraped the content, liked the results, and are only making these deals because it’s cheaper than getting sued.

report

reply

[ +- ]

AeroLemming@lemm.ee

3 points

2 months ago

Can they really sue (with a chance of winning) if you scrape content that’s submitted by users? That’s insane.

report

reply

[ +- ]

linearchaos@lemmy.world

34 points

2 months ago

Honestly? I’m down with that. And when the LLM’s end up pricing themselves out of usefulness, we’ll still have the fediverse version. Having free sites on the net with solid crowd-sourced information is never a bad thing even if other people pick up the data and use it.

It’s when private sites like Duolingo and Reddit crowd source the information and then slowly crank down the free aspect that we have the problems.

The Ad sponsored web model is not viable forever.

report

reply

[ +- ]

bort@sopuli.xyz

18 points

2 months ago

The Ad sponsored web model is not viable forever.

a thousand times this

report

reply

[ +- ]

danc4498@lemmy.world

27 points

2 months ago

I’d rather the harvesting be open to all than only the company hosting it.

report

reply

[ +- ]

mox@lemmy.sdf.org

10 points

2 months ago

Assuming the federated version allowed contributor-chosen licenses (similar to GitHub), any harvesting in violation of the license would be subject to legal action.

Contrast that with Stack Exchange, where I assume the terms dictated by Stack Exchange deprive contributors of recourse.

report

reply

[ +- ]

chameleon@kbin.social

7 points

2 months ago

SO already was. Not even harvested as much as handed to them. Periodic data dumps and a general forced commitment to open information were a big part of the reason they won out over other sites that used to compete with them. SO most likely wouldn’t have existed if Experts Exchange didn’t paywall their entire site.

As with everything else, AI companies believe their training data operates under fair use, so they will discard the CC-SA-4.0 license requirements regardless of whether this deal exists. (And if a court ever finds it’s not fair use, they are so many layers of fucked that this situation won’t even register.)

report

reply

[ +- ]

Rolando@lemmy.world

2 points

2 months ago

But users and instances would be able to state that they do not want their content commercialized. On StackOverflow you have no control over that.

report

reply

[ +- ]

ArbitraryValue@sh.itjust.works

5 points

2 months ago

You can state what you don’t want, but no one will be paying attention. Except maybe the LLM reading your posts…

report

reply

[ +- ]

pivot_root@lemmy.world

1 point

2 months ago

Yup. Laws are only suggestions until you get caught.

report

reply

[ +- ]

ArbitraryValue@sh.itjust.works

3 points

2 months ago

*

I suspect it isn’t even illegal, but I’m not an expert.

report

reply

Show more comments

Technology

!technology@lemmy.world

This is a most excellent place for technology news and articles.

Our Rules

Follow the lemmy.world rules.
Only tech related content.
Be excellent to each another!
Mod approved content bots can post up to 10 articles per day.
Threads asking for personal tech support may be deleted.
Politics threads may be removed.
No memes allowed as posts, OK to post as comments.
Only approved bots from the list below, to ask if your bot can be added please contact us.
Check for duplicates before posting, duplicates may be removed

Approved Bots

@L4s@lemmy.world
@autotldr@lemmings.world
@PipedLinkBot@feddit.rocks
@wikibot@lemmy.world

Community stats

17K
Monthly active users
10K
Posts
466K
Comments

Community moderators

L3s@lemmy.world
enu@lemm.ee
fry@fry.gs
L3s@fry.gs
enu@lemmy.world
L4sBot@fry.gsB
L4sBot@lemmy.worldB

modlog legal instances join-lemmy.org

lemmy-ui-next v0.11.0 (github)lemmy v0.19.5 (github)