• EnderMB@lemmy.world
    link
    fedilink
    arrow-up
    8
    ·
    4 months ago

    I’d be very surprised if comments weren’t versioned in some way, so even if you delete or rewrite that data, it’s probably still there and a part of training data.

    • athos77@kbin.social
      link
      fedilink
      arrow-up
      3
      ·
      4 months ago

      They said years ago that they only kept one previous version, which is why everyone overwrote and then deleted their stuff.

      It’s possible that reddit changed that, but honestly? That requires a level of foresight that I believe is entirely beyond spez. He didn’t foresee AI products, he literally paid all the bandwidth for them to harvest the data, he didn’t foresee changes to API pricing, he didn’t foresee the protests, how long they’d last, or how many people just walked away.

      Hell, in the previous big “closed subs” protest they’d never even considered a moderator rebellion: once the mods took the subs private, the admins were accidentally locked out as well - they had to negotiate to get them re-opened while they worked on backdoor changes that wouldn’t break reddit.

      I just don’t see them having the foresight to add in preservation code, nor to allocate the database and storage space to keep up with it. I think if you overwrote and then deleted your stuff, reddit doesn’t have it anymore. Of course, it’s still out there, in Google’s cache and the internet archive and all the other snapshots she preservation schemes and the data already harvested for the various AIs, but at least it’s no longer indeed reddit’s control, and they won’t be able to profit from it.