• onlinepersona@programming.dev
    link
    fedilink
    arrow-up
    2
    ·
    19 hours ago

    SerpiApi doesn’t give two shits about the openweb. They are just here to exploit it with their scraping. It’s very convenient to call a search engine the openweb, an action I fully support, but it’s not because they are some grand protectors or something. They just want to make money.

    For Google, it could be “very dangerous” to argue that the knowledge panel is “chock full of copyrighted material,” Rose suggested. Since the search giant doesn’t license all the content in the knowledge box, Google could risk future lawsuits if the act of algorithmically creating the knowledge box without licenses suddenly becomes viewed as infringement, Rose said.

    If Google comes out of this a winner, it’ll be because the judge was paid off or there was some obscure law they used, that made no sense but was somehow applicable here, in a twisted manner.

  • Tollana1234567@lemmy.today
    link
    fedilink
    arrow-up
    1
    ·
    17 hours ago

    google and reddit, openai are all joined at the hip. most of the training data comes from them, plus scraping WIKIPEDIA TOO.

  • AJMaxwell@lemmy.world
    link
    fedilink
    arrow-up
    28
    ·
    2 days ago

    Google getting a taste of their own scraping. Decades of scraping content for “search visibility” and when someone scrapes them, they cry all the way to court.