• Dave.@aussie.zone
    link
    fedilink
    arrow-up
    44
    ·
    7 days ago

    At this point, I’m just hoping that when we are sifting through the smoking wreckage of the AI crash we’ll find a hundred million books that we can add to Anna’s Archive.

    And some RAM. RAM would be nice.

    • Einskjaldi@lemmy.world
      link
      fedilink
      arrow-up
      5
      ·
      7 days ago

      They aren’t never seen before unique rare books, they’re just out of print and not lots of copies. They’re destroying them to sidestep copyright concerns.

      • MonkeMischief@lemmy.today
        link
        fedilink
        arrow-up
        2
        ·
        7 days ago

        They are ripping them from their spines and pulping them once the have what they want.

        They treat their books like they treat their employees.

      • BigPotato@lemmy.world
        link
        fedilink
        arrow-up
        11
        ·
        7 days ago

        They mean the PDFs of them or ePub or whatever they’re scanning them into. Yeah, the original is gone but imagine if the Library of Alexandria went up in flames but every ash contained a complete work for free to everyone to read at their leisure.

        Oh, and Ram.

        • Eldritch@piefed.world
          link
          fedilink
          English
          arrow-up
          5
          ·
          7 days ago

          Storage space is valuable for their AI stuff. Why would they keep around copies like that on valuable storage space if they don’t think they’re going to be using them again once they’ve trained their model? I really hope they are scanning them to a format or something like that that. But I doubt it. These aren’t the most forward looking or intelligent people you’ll find. I wouldn’t be at all surprised to find out that they never did anything more than scan it into train the model and then flush it from the system. It’s 100% on brand for them.

          • Dran@lemmy.world
            link
            fedilink
            arrow-up
            7
            ·
            7 days ago

            You always keep training data because you never know what you need for the next generation. Ironically they might be preserved (privately though unfortunately) pretty well