cross-posted from: https://infosec.pub/post/52443941

Newly unsealed court filings show Microsoft privately called OpenAI’s data practices “theft” while both companies scraped paywalled Times content, built datasets from it, and warned internally it would gut publishers.

  • medem@lemmy.wtf
    link
    fedilink
    arrow-up
    6
    ·
    2 days ago

    If what you actually mean is that ideas can’t be copyrighted, then yes, you are right.

    Content, or simply works, is different: in most jurisdictions, whenever you create anything, you own its copyright whether you like it or not. In Europe, IIRC, there is not even a way of waiving that right.

    The point here is that AS crawlers not only use but also reproduce information (sometimes verbatim) they don’t own the copyright for and is therefore not theirs to publish. As anyone who went to the university knows, presenting the work of others as yours is the one thing you’re not allowed to do.