82
submitted 10 months ago by corbin@infosec.pub to c/technology@beehaw.org
you are viewing a single comment's thread
view the rest of the comments
[-] acastcandream@beehaw.org 33 points 10 months ago

CC BY-NC-SA 4.0

Why are you putting a CC license on your comments?

[-] BotCheese@beehaw.org 21 points 10 months ago

From what I understand it is some thing for AI, to stop them from harvesting or to poison the data, by having it repeating therefore more likely to show up.

[-] beefcat@beehaw.org 60 points 10 months ago

Sounds an awful lot like that thing boomers used to do on Facebook where they would post a message on their wall rescinding Facebook's rights to the content they post there. I'm sure it's equally effective.

[-] Bene7rddso@feddit.de 4 points 10 months ago

Sure, the fun begins when it starts spitting out copyright notices

[-] t3rmit3@beehaw.org 2 points 10 months ago

That would require a significant number of people to be doing it, to 'poison' the input pool, as it were.

[-] corbin@infosec.pub 42 points 10 months ago

It seems pretty well established at this point that AI training models don't respect copyright.

[-] mozz@mbin.grits.dev 20 points 10 months ago* (last edited 10 months ago)

I would be extremely extremely surprised if the AI model did anything different with "this comment is protected by CC license so I don't have the legal right to it" as compared with its normal "this comment is copyright by its owner so I don't have the legal right to it hahaha sike snork snork snork I absorb" processing mode.

[-] Max_P@lemmy.max-p.me 13 points 10 months ago

No but if they forget to strip those before training the models, it's gonna start spitting out licenses everywhere, making it annoying for AI companies.

It's so easily fixed with a simple regex though, it's not that useful. But poisoning the data is theoretically possible.

[-] t3rmit3@beehaw.org 1 points 10 months ago

Only if enough people were doing this to constitute an algorithmically-reducible behavior.

If you could get everyone who mentions a specific word or subject to put a CC license in their comment, then an ML model trained on those comments would likely output the license name when that subject was mentioned, but they don't just randomly insert strings they've seen, without context.

[-] peter@feddit.uk 19 points 10 months ago

That seems stupid

[-] acastcandream@beehaw.org 12 points 10 months ago

Interesting. Feels like that thing people used to add to FB comments back in the day that did nothing but in the case of AI I could see it maybe doing something. I’ll be looking into it - thanks!

[-] conciselyverbose@kbin.social 19 points 10 months ago* (last edited 10 months ago)

To turn every comment, no matter how on topic, into obnoxious spam.

this post was submitted on 15 Feb 2024
82 points (100.0% liked)

Technology

37830 readers
752 users here now

A nice place to discuss rumors, happenings, innovations, and challenges in the technology sphere. We also welcome discussions on the intersections of technology and society. If it’s technological news or discussion of technology, it probably belongs here.

Remember the overriding ethos on Beehaw: Be(e) Nice. Each user you encounter here is a person, and should be treated with kindness (even if they’re wrong, or use a Linux distro you don’t like). Personal attacks will not be tolerated.

Subcommunities on Beehaw:


This community's icon was made by Aaron Schneider, under the CC-BY-NC-SA 4.0 license.

founded 2 years ago
MODERATORS