Authors using a new tool to search a list of 183,000 books used to train AI are furious to find their works on the list.

  • ThrowawayOnLemmy@lemmy.world
    link
    fedilink
    English
    arrow-up
    40
    arrow-down
    11
    ·
    edit-2
    1 year ago

    That’s an interesting take, I didn’t know software could be inspired by other people’s works. And here I thought software just did exactly as it’s instructed to do. These are language models. They were given data to train those models. Did they pay for the data that they used to train for it, or did they scrub the internet and steal all these books along with everything everyone else has said?

    • lloram239@feddit.de
      link
      fedilink
      English
      arrow-up
      2
      arrow-down
      6
      ·
      1 year ago

      And here I thought software just did exactly as it’s instructed to do.

      AI isn’t software. Everything the AI knows is from the books. There is no human instructing the AI what to do. All the human does is build the scaffolding to let the AI learn, everything else is in the data.