Asking ChatGPT to Repeat Words ‘Forever’ Is Now a Terms of Service Violation(www.404media.co)

posted 1 year ago

misk@sopuli.xyz

technology@lemmy.world

234 commentshide report

Sort:

Hot Top Controversial New Old

You are viewing a single thread.

View all comments

[ - ]

MNByChoice@midwest.social

27 points

1 year ago

Any idea what such things cost the company in terms of computation or electricity?

permalink

report

[ - ]

merc@sh.itjust.works

-3 points

1 year ago

Essentially nothing. Repeating a word infinite times (until interrupted) is one of the easiest tasks a computer can do. Even if millions of people were making requests like this it would cost OpenAI on the order of a few hundred bucks, out of an operational budget of tens of millions.

The expensive part of AI is training the models. Trained models are so cheap to run that you can do it on your cell phone if you’re interested.

permalink

report

parent

[ - ]

apinanaivot@sopuli.xyz

4 points

1 year ago

GPT4 definitely isn’t cheap to run.

permalink

report

parent

[ - ]

merc@sh.itjust.works

3 points

1 year ago

Depends how you define “cheap”. They’re orders of magnitude cheaper to run than they are to train.

permalink

report

parent

[ - ]

ExLisper@linux.community

7 points

1 year ago

What? They are not just generating this word in a loop. The model still calculates probability for each repetition, just like for any other query. It’s as expensive as other queries which is definitely not free.

permalink

report

parent

[ - ]

merc@sh.itjust.works

-2 points

1 year ago

The model still calculates probability for each repetition

Which is very cheap.

as expensive as other queries which is definitely not free

It’s still very cheap, that’s why they allow people to play with the LLMs. It’s training them that’s expensive.

permalink

report

parent

[ - ]

ExLisper@linux.community

2 points

1 year ago

Yes, it’s not expensive but saying that it’s ‘one of the easiest tasks a computer can do’ is simply wrong. It’s not like it’s concatenates strings, it’s still performing complicated calculations using on of the most advanced AI techniques known today and each query can be 1000x times more expensive than a google search. It’s cheap because a lot of things at scale are cheap but pretty much any other publicly available API on the internet is ‘easier’ than this one.

report

[ - ]

3 points

1 year ago

Well it depends what user experience and quality you are after. Some of Meta’s Llama 2 models require several GBs of GPU ram to run and be responsive.

permalink

report

parent

[ - ]

kromem@lemmy.world

3 points

1 year ago

You’re correct.

While costs are tracked per token, behind the scenes the longer the response the more it costs to continue generating, so millions of users suddenly thinking they are clever replicating what they read getting it to max output tokens is a substantial increase in underlying costs.

The DeepMind researchers were likely doing that by API call, which they were at least paying for on a per token basis.

And the terms hasn’t been updated to prevent it, they’ve always had this item as prohibited:

Attempt to or assist anyone to reverse engineer, decompile or discover the source code or underlying components of our Services, including our models, algorithms, or systems (except to the extent this restriction is prohibited by applicable law).

permalink

report

parent

[ - ]

Daxtron2@startrek.website

64 points

1 year ago

That’s not the reason, it’s because it was seemingly outputting training data (or at least data that looks like it could be training data)

permalink

report

parent

[ - ]

regbin_@lemmy.world

1 point

1 year ago

It’s definitely cost. There are other ways to make it generate text that is similar to training data without needing it to endlessly repeat words so I doubt OpenAI cares in that aspect.

permalink

report

parent

[ - ]

Daxtron2@startrek.website

1 point

11 months ago

It doesn’t endlessly repeat, there’s a cap on token generation per request. It absolutely is because of the recent “exploit”

permalink

report

parent

[ - ]

regbin_@lemmy.world

1 point

11 months ago

I don’t think they would care if it didn’t get popular and having thousands of people trying it out, eating up huge amount of compute resources.

It’s a known quirk of LLMs.

permalink

report

parent

Show more comments

[ - ]

MNByChoice@midwest.social

20 points

1 year ago

Sure, but this cannot be free.

Edit: oh, are you suggesting it is the normal cost? Nuts, chathpt is not repeating forever.

permalink

report

parent

[ - ]

nickwitha_k (he/him)@lemmy.sdf.org

2 points

11 months ago

I think that they were referring to the exploit that was recently published. Google researchers were able to reliably get the LLM to output training data verbatim, including PII.

To me, this reads as damage control for that. Especially as they are being sued for copyright infringement, which they and their proponents have been claiming is impossible (clearly, they were either wrong or lying).

permalink

report

parent

Technology

!technology@lemmy.world

Create post

This is a most excellent place for technology news and articles.

Our Rules

Follow the lemmy.world rules.
Only tech related content.
Be excellent to each another!
Mod approved content bots can post up to 10 articles per day.
Threads asking for personal tech support may be deleted.
Politics threads may be removed.
No memes allowed as posts, OK to post as comments.
Only approved bots from the list below, to ask if your bot can be added please contact us.
Check for duplicates before posting, duplicates may be removed

Approved Bots

Community stats

17K
Monthly active users
12K
Posts
557K
Comments

Our Rules

Approved Bots

Community stats

Community moderators