887
submitted 10 months ago by misk@sopuli.xyz to c/technology@lemmy.world
you are viewing a single comment's thread
view the rest of the comments
[-] MNByChoice@midwest.social 26 points 10 months ago

Any idea what such things cost the company in terms of computation or electricity?

[-] Daxtron2@startrek.website 63 points 10 months ago

That's not the reason, it's because it was seemingly outputting training data (or at least data that looks like it could be training data)

[-] MNByChoice@midwest.social 19 points 10 months ago* (last edited 10 months ago)

Sure, but this cannot be free.

Edit: oh, are you suggesting it is the normal cost? Nuts, chathpt is not repeating forever.

[-] nickwitha_k@lemmy.sdf.org 2 points 10 months ago

I think that they were referring to the exploit that was recently published. Google researchers were able to reliably get the LLM to output training data verbatim, including PII.

To me, this reads as damage control for that. Especially as they are being sued for copyright infringement, which they and their proponents have been claiming is impossible (clearly, they were either wrong or lying).

[-] regbin_@lemmy.world 1 points 10 months ago* (last edited 10 months ago)

It's definitely cost. There are other ways to make it generate text that is similar to training data without needing it to endlessly repeat words so I doubt OpenAI cares in that aspect.

[-] Daxtron2@startrek.website 1 points 10 months ago

It doesn't endlessly repeat, there's a cap on token generation per request. It absolutely is because of the recent "exploit"

[-] regbin_@lemmy.world 1 points 10 months ago

I don't think they would care if it didn't get popular and having thousands of people trying it out, eating up huge amount of compute resources.

It's a known quirk of LLMs.

load more comments (11 replies)
this post was submitted on 04 Dec 2023
887 points (97.9% liked)

Technology

58665 readers
3438 users here now

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related content.
  3. Be excellent to each another!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, to ask if your bot can be added please contact us.
  9. Check for duplicates before posting, duplicates may be removed

Approved Bots


founded 1 year ago
MODERATORS