The research community is in uproar after OpenAI released a trove of more than 700 mathematical preprints entirely generated by AI on 6 October. The San Francisco, California-based maker of ChatGPT posted the preprints on the software repository Github.

Although some mathematicians celebrated the solution of longstanding problems, others took to social media to complain about being scooped. Some were incensed at what one physicist called a ‘slopocalypse’, even if the mathematical content could end up being formally correct.

  • givesomefucks@lemmy.world
    link
    fedilink
    English
    arrow-up
    113
    ·
    edit-2
    11 hours ago

    I saw the Wired article yesterday, it’s worse than it sounds.

    For at least one of the problems an OpneAI researcher initiated a partnership with a mathametician, loaded all of his work into a chatbot without his knowledge, and then published the result as 100% from the chatbot in this “trove”.

    When the mathematician confronted the OpenAI employee, he was told sharing credit was “too complicated” so they wouldn’t acknowledge he had already mostly had it done. When he tried to push back, OpenAI then threatened his career if he tried to tell people what had happened.

    It’s is very very unlikely that was a single case. If that happened once, they definitely at least tried it more than once.

    Even if it actually solved anything, taking in years of human work and then spitting out one last step is nowhere near what they’re portraying this as. They’re intentionally making it sound like they just gave the chat or the problem and it solved what a human couldn’t.

    All it did was steal human work and rush publication

    • SaveTheTuaHawk@lemmy.ca
      link
      fedilink
      English
      arrow-up
      13
      ·
      7 hours ago

      OpenAI then threatened his career if he tried to tell people what had happened.

      and thus ended academic collaborations with OpenAI.

    • [object Object]@lemmy.ca
      link
      fedilink
      English
      arrow-up
      17
      ·
      8 hours ago

      The whole threatening careers thing happened with navier stokes too.

      These people are completely rotten to the core.

    • yeahiknow3@lemmy.world
      link
      fedilink
      English
      arrow-up
      2
      ·
      edit-2
      3 hours ago

      Solving math proofs with AI is like doing art with AI. Pointless. There are practical problems and calculations computers could solve instead, but they’re targeting the mathematical community directly to undermine the discipline as a whole. This means people will stop sharing work and insights with each other, which has never been the case before.

      • givesomefucks@lemmy.world
        link
        fedilink
        English
        arrow-up
        1
        ·
        5 hours ago

        I don’t know what you’re trying to say, and I feel like you don’t either.

        This means people will stop sharing work and insights with each other, which has never before been the case.

        That part was clearly, but obviously false on multiple levels

        • yeahiknow3@lemmy.world
          link
          fedilink
          English
          arrow-up
          2
          ·
          edit-2
          2 hours ago

          I’m actually paraphrasing Terence Tao and some of the top mathematicians on the planet: pre-AI, math researchers discussed and shared ideas about their work. Now they are increasingly afraid of getting scooped, and you can see why.

          Non-mathematicians don’t realize this, but proofs aren’t intrinsically valuable (useful in any way whatsoever) outside of what they mean about the community capable of creating them. Destroying that community is game over.

          Scientific discoveries without scientists could still be useful to society. Not so about mathematical proofs without mathematicians.

          • givesomefucks@lemmy.world
            link
            fedilink
            English
            arrow-up
            1
            ·
            3 hours ago

            I’m actually paraphrasing Terence Tao

            That doesn’t even mean you’re saying the same thing bud…

            Non-mathematicians don’t realize this, but proofs aren’t intrinsically valuable outside of what they mean about the community capable of creating them.

            Do you think “non-mathematicians” think solving a math problem magically gets someone a check? Like, if anythi g you have it backwards

            You’re way off on everything though.

            You should paraphrase less, wait till after you understand what you’re reading.

            • yeahiknow3@lemmy.world
              link
              fedilink
              English
              arrow-up
              2
              ·
              edit-2
              2 hours ago

              I provided a link. Are you… not able to read?

              What do you think I mean by “valuable” and why would that equal “getting a check”?

              Fucking bizarre. This might be the dumbest conversation I’ve participated in this week.

    • HubertManne@piefed.social
      link
      fedilink
      English
      arrow-up
      10
      ·
      10 hours ago

      it might not even be the last step in a way. A person may have it solved but worries about publishing until they check and double check and run it down a few times to be sure and then of course there is the writing. So really it might just be acting as a technical writer for the researcher at best.

    • MountingSuspicion@reddthat.com
      link
      fedilink
      English
      arrow-up
      10
      ·
      10 hours ago

      All it did was steal human work and rush publication

      Ok, but literally all it does is steal human work and provide an output.

      I wonder how the mathematician expected it to go. Was he expecting to be directly provided the AI output so he could publish it? Was he going to care if the AI used a novel approach he hadn’t employed and was unable to identify where it got it from? We have seen since the beginning that AI has an attribution problem, so it seems like his actual concern was that HE wasn’t credited.

      I’m not an AI supporter, but this whole “it did just a small part” complaint is weird to me. I believe I read that most were solved in 3 hours of compute. I’m sure the mathematician wasn’t just 3 hours from solving the problem. It did it faster than humans did. Isn’t that the point? Does it matter how little the step was if it hadn’t been taken yet? It could’ve taken mathematics years to make that last step. Shouldn’t we be excited to see this development? If we’re going to have AI, isn’t the point that it accelerates problem solving. If AI sped up life saving medical research by years or even just months by doing one last step, would people be having the same reaction?

      I’ve seen the same people complaining about this regularly talk about using AI for coding. Why is it ok that so many companies have put programmers in this position, but all the sudden it’s bad for mathematicians?

      • demonsword@lemmy.world
        link
        fedilink
        English
        arrow-up
        15
        ·
        10 hours ago

        Why is it ok that so many companies have put programmers in this position, but all the sudden it’s bad for mathematicians?

        the former has just been normalized and tech bros are hard at work trying to normalize the latter, but it’s not ok in either case

  • Zwuzelmaus@feddit.org
    link
    fedilink
    English
    arrow-up
    18
    ·
    9 hours ago

    What we should do IMHO:

    Make the owner of an AI legally responsible for the actions of their AI system, but not legal owners of the AI system’s results/products.

  • cannedtuna@lemmy.world
    link
    fedilink
    English
    arrow-up
    39
    ·
    11 hours ago

    “Why is OpenAI trying to solve hundreds of math problems to be released all at once?”, he added.

    even if the mathematical content could end up being formally correct.

    It should be noted that none of what they released has been verified. Throwing 700 papers out there is intentional to generate buzz, “hey look how great we are”, and to slow attempts at proving or disproving their results. That way even if a few get disproven early they can still say “well that’s just a couple out of hundreds” because it’s gonna take forever to work through the slop dump. It could be a long time before we get a percentage of accuracy and by then most people will have forgotten or lost interest since the headline will be stale by then.

    some mathematicians who might find their research projects pre-empted by the OpenAI release. “We should put that aside for today though. … he wrote, because it shows how quickly the technology has become useful to mathematicians.”

    Yeah this fucks the researchers. Funding could get pulled from related projects leaving their work in hiatus until someone can check the slop for accuracy. Till then someone can just point to this and say “oh it’s already been solved” like it’s a done deal already.

    “Look how useful we are!!!” Turns in unverified crap

    • schipelblorp@sh.itjust.works
      link
      fedilink
      English
      arrow-up
      22
      ·
      edit-2
      11 hours ago

      Throwing 700 papers out there is intentional to generate buzz, “hey look how great we are”, and to slow attempts at proving or disproving their results.

      The Gish gallop is a rhetorical technique in which a person in a debate attempts to overwhelm an opponent by presenting an excessive number of arguments, without regard for their accuracy or strength, with a rapidity that makes it impossible for the opponent to address them in the time available. Gish galloping prioritizes the quantity of the galloper’s arguments at the expense of their quality.

      The term “Gish gallop” was coined in 1994 by the anthropologist Eugenie Scott, who named it after the creationist Duane Gish, described by Scott as the technique’s “most avid practitioner”.[1][2]

      • HubertManne@piefed.social
        link
        fedilink
        English
        arrow-up
        5
        ·
        10 hours ago

        thanks because you know I have seen this with trump and the right in the us but did not think of it in terms of something recognized.

        • schipelblorp@sh.itjust.works
          link
          fedilink
          English
          arrow-up
          3
          ·
          edit-2
          10 hours ago

          You’re more patient than I am; I tune that nonsense right out.

          But I bet if you isolated RW talking points, you’d find they’re not even proper arguments so much as emotional triggers. Like a common argument is “That’s socialism!” A good ghish gallop’s arguments at least have internal validity.

          • HubertManne@piefed.social
            link
            fedilink
            English
            arrow-up
            2
            ·
            9 hours ago

            well what I mean is they will throw out things and when challenged or fact checked then go to a new thing. sometimes not new but proved false long enough ago that people forgot. but yeah a lot is poorly done to begin with but its using large amounts and going to a new one if its disproved and actually ignoring that they ever used it. Oh and then attacking fact checking places as being political for pointing out facts or yelling at news entities for doing it.

  • Zwuzelmaus@feddit.org
    link
    fedilink
    English
    arrow-up
    26
    ·
    12 hours ago

    Stealing from all the world and then bragging boldly about it.

    Is that the new accepted behaviour?

    The next ruler of the world?

  • lepinkainen@lemmy.world
    link
    fedilink
    English
    arrow-up
    11
    ·
    edit-2
    11 hours ago

    Can someone please explain to a non-mathematician what the issue is specifically?

    We all want the problems solved, right? Is there a wrong way to solve them? Or a process that people are supposed to follow not to “scoop” others? A honor system of sorts?

    If a human dropped the same 700 solutions online, would it still be bad?

    • Imgonnatrythis@sh.itjust.works
      link
      fedilink
      English
      arrow-up
      4
      ·
      7 hours ago

      There’s an onus now on those that care about math to check the proofs. If errors are found in one their might be similar errors in another. A serial release would have been cleaner and allowed commentary that might help improve the accuracy or perhaps even the way the model conveyed the data to make it more understandable / digestable.

      Openai dug the problems up from somewhere. They could have collaborated with some of the people that defined the problems in the first place. This is just perpetuation of their culture of theft.

    • CelloMike@lemmy.world
      link
      fedilink
      English
      arrow-up
      32
      ·
      11 hours ago

      There are two issues really - one is a perception that they’re not vetting any of this output, just dumping it out on the internet and expecting real mathematicians to actually check it all, so yes if a human was doing the same thing it’d probably still be suspect.

      The other is the widely held suspicion that much of the “breakthrough” mathematics here has been stolen from real researchers who’ve been using chatgpt to develop their theories by talking to it

    • [deleted]@piefed.world
      link
      fedilink
      English
      arrow-up
      16
      ·
      11 hours ago

      It is a lot of stuff from an unreliable source that basically vomits out stuff that looks correct but is often not and that is a ton of human work to verify that humans were already in the process of doing.

      It is most likely just the stuff humans were close to finishing being scooped up and spit out and not reliably correct.

    • zener_diode@feddit.org
      link
      fedilink
      English
      arrow-up
      9
      ·
      11 hours ago

      I think there are three main issues (though I might be overlooking something):

      1. AI taking over peoples work is as much an issue for mathematicians here, as it is for programers, artists, writers, etc. elsewhere.
      2. We still need to check if these 700 papers are correct (AI makes mistakes at least as often as humans, usually more often). This involves a huge amount of effort. And if it turns out that they contain a bunch of mistakes, OpenAI is unlikely to care. To them, these papers have already served their purpose (generating headlines).
      3. Solving a problem in mathematics means being able to (formally) convince other people that your opinion is correct (massively simplified, but imo that is what mathematics boils down to), by getting them to understand why your opinion is correct. Having a pile of linear algebra spit out 700 texts about certain math problems removes the whole “human understanding” part of that.

      However, it’s not entirely unprecedented for a computer to provide a proof that we don’t completely understand. I forgot the details (and I’ll edit them in if I can find it), but there was at least one problem that was solved using a computer and brute force trying millions of cases. That produces a proof much longer than any human could read in a lifetime, and iirc there where quite a few mathematicians unhappy about it at the time too.

      • Eq0@literature.cafe
        link
        fedilink
        English
        arrow-up
        8
        ·
        11 hours ago

        I can add about your last paragraph(it’s literally the core of what I do). There is a field of computer assisted proofs, that is mathematical proofs that need a computer to be completed.

        The first and most well known example is the four color theorem. How many colors do you need to color “a map”. Answer: 4. Proof: very very long. By hand it is possible to prove that there are only 1834 options, and a computer was used to color all these explicit options. At the time, it was a scandal. Nowadays, other computer assisted proofs are accepted, such as the ones relying on validated computing: if a computer (with some restrictions and guardrails) can show that a certain value is over/under a given threshold then something else is true.

        Then, there is the validated proofs approach. This is where Lean comes into play, if you have heard about it. You can ask a computer to check your proof. This is helpful for confusing, long proofs (most of math). You input all the logical steps you took and Lean confirms that all is logically sound. Many AI proofs are “Lean verified”, but there is controversy if they are proving what they claim they are proving. It’s also a massive chunk of code that nobody can understand.

        • Grimy@lemmy.world
          link
          fedilink
          English
          arrow-up
          2
          ·
          10 hours ago

          Have you looked at some of the problem solved? I’m way out of my depth here, so I’m having trouble understanding how much 700 is. If it is proven and understandable, how much of an advancement would this represent? Sorry if it’s a random question, but you seem more knowledgeable than most.

          • Eq0@literature.cafe
            link
            fedilink
            English
            arrow-up
            4
            ·
            9 hours ago

            Just to give a comparison, a high-output mathematician publishes between 2 and 5 papers a year (depending on branch of math and dividing by co-authors).

            Most are incremental work, so a little step towards solving a problem or a conjecture. A lot is just a “hey, look at this neat trick”. Solving “big problems” is usually the work of a decade or more, in which the mathematician is working on other stuff as well. So let’s roughly say that solving a big problem usually takes some 10 years of work and some 10-30 papers (assuming working roughly half time on it).

            So on one hand, 1 paper per big problem is too little to actually understand what’s going on, on the other 700 papers are the output of more than a hundred mathematicians over a year. A math department in a university is around 50 mathematicians, I would say? So two medium-sized departments.

            The other part of your comment. “If its proven and understandable”

            At the moment, the preprints are assumed to be a shit sandwich. (Aka “would you eat a sandwich is there could be shit in it?” ). Using the results is just too risky without understanding if they are correct and how they are build.

            The “understandable” bit is also really hard. I picked a random one in my field (nothing I directly worked on) and it was unreadable - mostly because of notation used without defining it and no explanation of what is going on, no overview or intuition. So to me it seems more shit than sandwich. As a reviewer, I would never accept such a paper.

            Then finally, let is assume it’s all perfect and good. The goal of proving something in math is to develop understanding and a method to apply to other cases. So once all these papers are studied and understood, poop discarded, rest of the sandwich saved, ideally we will have new understandings of whole sections of math, new connection between items we’re weren’t aware of. What I think AI did in this context is to chain things that were already known, but there was no one whose knowledge spanned wide enough to know that all the pieces were already laid out. So it could be groundbreaking - once we remove the shit. How much shit there is is anyone’s guess.

            • Grimy@lemmy.world
              link
              fedilink
              English
              arrow-up
              2
              ·
              edit-2
              8 hours ago

              Thanks! That really puts things into perspective. It’s less impressive than I first thought, especially since it seems like spaghetti that needs a lot of untangling. 700 struck me as a huge number at first.

    • Zwuzelmaus@feddit.org
      link
      fedilink
      English
      arrow-up
      6
      ·
      11 hours ago

      We all want the problems solved, right? Is there a wrong way to solve them?

      Science is not mainly about solving problems (that would be engineering for example).

      But more specifically to your question:

      There are many wrong ways of doing scientific work, and what current AI does is one of them: reading it all, rephrasing it all a little and then already declare everything as done & good, even when it isn’t.

      A scientist needs proof, and peer reviews, etc. before the work can be called good.

      Never forget AI isn’t creating really new things, it can only chew up and spit out what has been there before. This is especially true in science.

  • ExLisper@lemmy.curiana.net
    link
    fedilink
    English
    arrow-up
    3
    ·
    edit-2
    8 hours ago

    Here’s what the field should do: ignore those paper and never publish them. Then have people verify the math and publish the proofs as their own (whoever can do it first gets the credit).

    • HubertManne@piefed.social
      link
      fedilink
      English
      arrow-up
      4
      ·
      10 hours ago

      that makes a lot of sense. they are possible proofs till someone actually verifies them. heck even publishing is about more verification.

    • Eq0@literature.cafe
      link
      fedilink
      English
      arrow-up
      2
      ·
      9 hours ago

      I still don’t want a “first pass the post” kind of math, but that would be a sensible solution