A frog who wants the objective truth about anything and everything.

Admin of SLRPNK.net

XMPP: prodigalfrog@slrpnk.net

Alt lemmy account: Cafefrog@lemmy.cafe

  • 344 Posts
  • 842 Comments
Joined 3 years ago
cake
Cake day: July 4th, 2023

help-circle





  • Ah man, that absolutely sucks. I went and read all of their responses. You did a fantastic job presenting a compelling case to them, but for the most part they seemed to not want to engage with the meat of your points, or acknowledge your distinction between code output and AI bug searching. And odd of one to equate AI code as equivalent to code coming to them in a dream, yet never address the copyright angle. Thanks for taking the time to engage with them to find all that out.

    I’m starting to see that avoiding slop even in the FLOSS world isn’t going to be trivial, and more are jumping on board than I anticipated… I’m thinking it may be necessary and worthwhile to probe as many widely used FLOSS projects as possible on their stance to then compile the results into a list on codeberg (so others can contribute their findings as well), to make it easier for people to find software that refuses to use AI.

    Maybe calling it the Slopless Software Project?










  • I haven’t heard of it happening and until it does it’s basically a non issue.

    Projects that operate on that assumption are essentially sticking their head in the sand and hoping for the best. LLM output is on particularly shaky ground regarding copyright and it’s vulnerability to being sued for plagiarism, see here: https://slrpnk.net/comment/23533549

    I think it’s a particularly dangerous gamble for projects based in the US, where the court system is extremely pro-corporation, making it likely to be much harder for a small FLOSS project to defend itself successfully.

    Using the internet in any capacity directly helps the financi situations of the big tech companies

    If someone gets their internet in the US, it is likely supplying a mega corporation with a monthly revenue stream, as internet providers are an oligopoly here. In other countries, this is often not the case.

    Internet access in the modern day is, for the most part, extremely difficult to avoid needing in order to function easily in society.

    In comparison, Corporate LLMs are extremely easy to avoid, and all of society got along fine without it just a few short years ago.

    Is it not worthwhile to lower your fossil fuel use where possible, despite many in the US needing a car due to terrible public transport? Is it not also worth lowering animal meat consumption where possible for both animal welfare concerns and environmental concerns? I think there’s equal merit in depriving as much capital (and thus power) from pro-fascist tech corporations as possible, even if it is not total.


  • The reasons it's a problem (spoiler, click to expand)

    The Licensing Problem:

    Unknowingly using potentially copyrighted code from an LLM in a FLOSS project opens it up to being sued for infringement, which most FLOSS devs can’t afford to fight, especially in the US’s current extremely pro-corporate courts.

    It’s putting a target on your back for down the road when it becomes profitable for patent trolls to use AI to try to scan for copyrighted code on public code bases. Big tech companies could do the same to squash an open source competitor.

    Both Haiku OS and BSD are not allowing AI code contributions for this very reason.

    The Ethical problem:

    Simply using an AI that’s run on a corporate data centers encourages the construction of yet more data centers, with all of the environmental/climate negatives they bring, as well as local harms they induce on the people living near them, such as increased electricity rates.

    Using corporate AI directly helps the financial situations of the big tech companies that host them (by boosting usage/user numbers, they are able to attract more investment capital), most of which are owned by right-wing CEO’s who are more than willing to collaborate with and fund fascist governments to ensure that they are not regulated in search of both maximum profits. Some of these companies, such as Nvidia, Palantir and Oracle, genuinely appear to be seeking to use these tools for what would previously be considered crackpot conspiracy theory levels of public control and surveillance.

    But if someone is absolutely determined to use them regardless, I would hope they either use a locally run LLM, or use a distributed model like AI Horde, as at the very least that does not encourage the construction of corporate data centers nor boost their financial prospects.



  • The problem is both using LLM code in an open-source project due to the threat that poses to the project if they unknowingly introduce copyright infringement, as as well as the very real ethical concerns inherent to using corporate owned and hosted LLM models.

    The problems in detail (spoiler, click to expand)

    The Licensing Problem:

    Unknowingly using potentially copyrighted code from an LLM in a FLOSS project opens it up to being sued for infringement, which most FLOSS devs can’t afford to fight, especially in the US’s current extremely pro-corporate courts.

    It’s putting a target on your back for down the road when it becomes profitable for patent trolls to use AI to try to scan for copyrighted code on public code bases. Big tech companies could do the same to squash an open source competitor.

    The Ethical problem:

    Simply using an AI that’s run on a corporate data centers encourages the construction of yet more data centers, with all of the environmental/climate negatives they bring, as well as local harms they induce on the people living near them, such as increased electricity rates.

    Using corporate AI directly helps the financial situations of the big tech companies that host them (by boosting usage/user numbers, they are able to attract more investment capital), most of which are owned by right-wing CEO’s who are more than willing to collaborate with and fund fascist governments to ensure that they are not regulated in search of both maximum profits. Some of these companies, such as Nvidia, Palantir and Oracle, genuinely appear to be seeking to use these tools for what would previously be considered crackpot conspiracy theory levels of public control and surveillance.

    But if someone is absolutely determined to use them regardless, I would hope they either use a locally run LLM, or use a distributed model like AI Horde, as at the very least that does not encourage the construction of corporate data centers nor boost their financial prospects.