‘Impossible’ to create AI tools like ChatGPT without copyrighted material, OpenAI says::Pressure grows on artificial intelligence firms over the content used to train their products

  • AnneBonny@lemmy.dbzer0.com
    link
    fedilink
    English
    arrow-up
    14
    arrow-down
    9
    ·
    11 months ago

    I don’t understand why people are defending AI companies sucking up all human knowledge by saying “well, yeah, copyrights are too long anyway”.

    Would you characterize projects like wikipedia or the internet archive as “sucking up all human knowledge”?

    • MBM@lemmings.world
      link
      fedilink
      English
      arrow-up
      17
      arrow-down
      2
      ·
      11 months ago

      Does Wikipedia ever have issues with copyright? If you don’t cite your sources or use a copyrighted image, it will get removed

    • dhork@lemmy.world
      link
      fedilink
      English
      arrow-up
      16
      arrow-down
      2
      ·
      11 months ago

      In Wikipedia’s case, the text is (well, at least so far), written by actual humans. And no matter what you think about the ethics of Wikipedia editors, they are humans also. Human oversight is required for Wikipedia to function properly. If Wikipedia were to go to a model where some AI crawls the web for knowledge and writes articles based on that with limited human involvement, then it would be similar. But that’s not what they are doing.

      The Internet Archive is on a bit less steady legal ground (see the resent legal actions), but in its favor it is only storing information for archival and lending purposes, and not using that information to generate derivative works which it is then selling. (And it is the lending that is getting it into trouble right now, not the archiving).

      • phillaholic@lemm.ee
        link
        fedilink
        English
        arrow-up
        4
        ·
        11 months ago

        The Internet Archive has no ground to stand on at all. It would be one thing if they only allowed downloading of orphaned or unavailable works, but that’s not the case.

      • randon31415@lemmy.world
        link
        fedilink
        English
        arrow-up
        2
        ·
        11 months ago

        Wikipedia has had bots writing articles since the 2000 census information was first published. The 2000 census article writing bot was actually the impetus for Wikipedia to make the WP:bot policies.

    • assassin_aragorn@lemmy.world
      link
      fedilink
      English
      arrow-up
      9
      arrow-down
      1
      ·
      11 months ago

      Wikipedia is free to the public. OpenAI is more than welcome to use whatever they want if they become free to the public too.

      • afraid_of_zombies@lemmy.world
        link
        fedilink
        English
        arrow-up
        1
        arrow-down
        4
        ·
        11 months ago

        It is free. They have a pair model with more stuff but the baseline model is more than enough for most things.

          • afraid_of_zombies@lemmy.world
            link
            fedilink
            English
            arrow-up
            1
            arrow-down
            5
            ·
            11 months ago

            There also shouldn’t be goal post moving in lemmy threads but yet here we are. Can you move the goalposts back into position for me?

            • assassin_aragorn@lemmy.world
              link
              fedilink
              English
              arrow-up
              3
              arrow-down
              1
              ·
              11 months ago

              My position has always been that OpenAI can either pay for training materials or make money solely on advertisements. Having a paid version is completely unacceptable if they aren’t paying for training.

              • afraid_of_zombies@lemmy.world
                link
                fedilink
                English
                arrow-up
                1
                arrow-down
                3
                ·
                11 months ago

                OpenAI is more than welcome to use whatever they want if they become free to the public too.

                My position has always been

                Left the goalposts and went on to gaslighting