Among the many recent controversies around AI, there is the claim that to solve a longstanding mathematical problem, OpenAI stole the work of the New York University mathematician Tristan Buckmaster and a colleague, Levent Alpöge, who also is a researcher for Anthropic. I have no ability to assess the validity of the claim, but as I understand it, the argument is that OpenAI had access to work done by Buckmaster and Alpöge. It then built on this work to quickly rush ahead and find a solution, which also had a $1 million prize attached to it.

Insofar as that accurately describes what OpenAI did, this seems like it would be equivalent to having a professor doing a seminar for colleagues and then having one of the attendees do a few additions and claim the work as their own. This would not be viewed as very collegial behavior in a normal academic department.

I’m not sure what claim a professor would have against a colleague under such circumstances. I suppose if their work was properly cited, there would not be a claim of plagiarism. (OpenAI’s work did not directly cite the work of Buckmaster and Alpöge, from my understanding.) I suppose there could be a basis for disputing rights to the million-dollar prize. But the problem seems much greater.

If AI can effectively steal the work of others, and claim it as its own, then it undermines the practice of openness in science that has allowed for enormous progress for centuries. This practice depended on norms, whereby scientists would properly credit the work of their peers.

In the seminar scenario I described, if one of the attendees listened to the presentation and realized the one or two additional steps needed to get over the goal line, they would be expected to raise the issue with the presenter. Assuming the additions were correct, the presenter would finish the paper, acknowledging the contribution or possibly adding the attendee as a co-author, if their contribution was sufficiently important.

But that relies on norms. If the owner of the AI doesn’t share those norms and can freely scrape scientific work wherever it finds it, and then take credit for the product, it will threaten the longstanding practice of openness in science. With the risk of having years of work stolen without getting any credit, researchers will likely keep their work closely held until they have a finished paper.1 They would only share it with a small group of trusted colleagues. Posting a working paper, or having an open seminar, would put control over their work at risk.

However, it is important to recognize the villain in this story. It is Sam Altman and his company, not the AI. Altman and the company are the ones responsible for what their AI does. If it did something they disapprove of, it is their responsibility to correct it. However adept AI may become, it is not the active player; it is the individual or company that set it in motion.

This issue applies more generally. There has been a social media mini firestorm over Texas Attorney General and Republican Senate candidate Ken Paxton using AI to have his opponent, James Talarico, making a series of outlandish statements that he never said. Much of the outrage seems to be over Paxton using AI, rather than the fact that he is lying about what Talarico said.

The issue here is that Paxton is defaming Talarico; the AI is secondary. The AI presumably makes the defamation more effective, but if Paxton were an exceptionally gifted mimic, he could likely also trick some people into believing that he was speaking in Talarico’s voice when making outlandish assertions. If that were the case, it would be every bit as bad as using AI to accomplish the job.

There were also the claims that the girls’ school in Iran, which was destroyed by the US military on the first day of the war, killing 120 children, was selected as a target by AI. That may be the case, but the important point was that the military did not have a system in place to properly review targets. It doesn’t matter whether or not an AI system identified the target. The problem was that there was no system that ensured the places designated were proper military targets.

It is important to keep our eyes on the ball. The bad actors are those who do bad things with AI, not the AI. This is like the husband who gets drunk and beats his wife and then says the whiskey did it. No one should ever accept that story. In the same vein, if a theft of mathematical work did occur, it was Sam Altman who did it and it was Ken Paxton who libeled his political opponent. The AI is beside the point.

top 50 comments

sorted by: hot top controversial new old
[–] 10 points 9 hours ago
[–] 1 point 7 hours ago
[–] 4 points 11 hours ago*

Not your AI, not your data.

  • source
  • [–] -5 points 10 hours ago (19 children)

    LLMs clap are clap not clap AI clap.

    Try to get that into you sycophant brain, and you will understand why this bullshit happens in the first place.

    Plagiarism is the core and only mechanism of LLMs that produces anything that people get impressed by. The pre- and postprocessing logic for grammar rules is not the stuff that makes gullible people believe there is "intelligence" involved.

  • source
  • hideshow 19 child comments
  • [–] 2 points 1 hour ago

    Your statement ignores the meaning of AI or the decades of work that have gone into the field.

    They're not sentient. They're not sapient. They're not conscious, nor do they have feelings, volition, desire, or anything that can be described as an internal life. They can't even think in they sense that anyone means by it.

    None of those things are intelligence though.

    An LLM posesses the ability to respond to external stimuli in a way that is meaningful. They also have the ability to learn, which is the act of a system transforming information or stimuli in a way that is applied to later stimuli.

    My air conditioner learns and is intelligent. It measures the environment, takes actions based on that stimulus, and when I make adjustments it records when I like different settings and applies that to the temperatures it's recorded in the past to adjust early. An LLM may be more sophisticated than my air conditioner but it's in the same category.

    The mistake is thinking that AI in the real world is what AI is on TV.

  • source
  • parent
  • [–] 4 points 6 hours ago (1 child)

    In a way, LLMs are kind of strange medium for texts, as are books. Books are a medium for written speech; they transport words written by intelligent beings. Nobody who understands books would claim that books themselves are intelligent. Or that a photograph of a persons robs them of their soul.

    But what is characteristic of LLMs is that it disowns the intelligent content from its creators. This is what makes coorporations salivate avout money. The theft is built-in and foundational how this technology works. Compare to an online encyclopedia.

  • source
  • parent
  • hideshow 1 child comment
  • [–] 2 points 6 hours ago

    The theft is built-in and foundational how this technology works

    Aye, this is one of the key takeaways.

    Another being that the technology is good enough to mass-manipulate people. And for statistically "helpful" surveillance tasks in a totalitarian police state.

  • source
  • parent
  • [–] 10 points 8 hours ago (4 children)

    LLMs clap are clap not clap AI clap.

    Yes they are.

    They're built on artificial neural networks that share a basis (a basis, not a replica) with biological neurons that produce probabilistic results (not deterministic).

    Plagiarism is the core and only mechanism of LLMs that produces anything that people get impressed by.

    As opposed to what? Don't get me wrong, the blatant violation of copyright is a massive issue, and in a logical world would bury each and every LLM company out there. But what other mechanism to "learn" do you think exists?

    gullible people believe there is "intelligence" involved.

    If you mean "intelligence" in the colloquial sense, then sure. But that's not the same word that researchers use. Intelligence is simply the ability to make a decision based on input data. And by that definition even an if-else statement is a form of intelligence (which it is). So then the discussion becomes the degree or "amount" intelligence.

    LLMs are absolutely intelligent. The issue is that most people conflate "intelligence" with "sentience" and "sapience", which they are not. They aren't "clever" or "creative". They are extremely good at taking large swaths of information and distilling it, finding hidden patterns, of finding specific data points that are otherwise deeply buried. Good examples are finding bugs in code or novel and massively beneficial uses for existing medication (which has already happened multiple times)

    I'm not defending the current state of AI. Energy use is a problem. Water use is a problem. The grift is a problem. The copyright violations are a problem.

    But this "LLMs are not AI" rhetoric is just pure ignorance.

  • source
  • parent
  • hideshow 4 child comments
  • [+] -7 points 6 hours ago (3 children)

    with biological neurons that produce probabilistic results

    Ah, so you understood neurons better than neuroscientists, eh? You have missed the part where no scientist worth their name claims to be able to fully explain how consciousness works and whether or not it affects our decision making.

    LLMs are absolutely intelligent.

    Okay, I think we're done here. You clearly have no idea what you are talking about.

  • source
  • parent
  • hideshow 3 child comments
  • [–] 5 points 3 hours ago* (last edited 2 hours ago) (2 children)

    so you understood neurons better than neuroscientists

    No, but I have no idea what that statement has to do with artificial neurons.

    You have missed the part where no scientist worth their name claims to be able to fully explain how consciousness works and whether or not it affects our decision making.

    And you have missed the part where I explicitly mentioned that LLMs are NOT "sapient" or "sentient", which would preclude "consciousness" also.

    Okay, I think we're done here. You clearly have no idea what you are talking about.

    Says the person using colloquial definitions and applying them to a field of active research.

    The facts don't care about your opinions. A basic program with hardcoded decision trees is "intelligent". An LLM can do more at a dynamic (and probabilistic) level. Therefore, LLMs fit the academic definition of intelligence.

    Even the general word definition:

    the ability to learn or understand things or to deal with new or difficult situations.

    I can point an LLM to my code that's never been seen by an LLM and it can read it, explain to me exactly what it does, point out bugs, and even give suggestions on areas that can be improved.

    Is it "smart"? Debatable.

    Is it conscious, sapient, sentient, or "alive"? Hard no. And that was never the argument.

    Edit: autocorrect speaking mistake

  • source
  • parent
  • hideshow 2 child comments
  • [–] -3 points 2 hours ago (1 child)

    And you have missed the part where I explicitly mentioned that LLMs are NOT “sapient” or “sentient”, which would preclude “consciousness” also.

    I am inclined to believe that people who use the word "sapi*"anything are just very very gullible and/or Dunning-Kruger-victims.

    You completely missed the point I made that links understanding and conscience in humans.

  • source
  • parent
  • hideshow 1 child comment
  • [–] 1 point 27 minutes ago

    am inclined to believe that people who use the word "sapi*"anything

    You mean the word "sapient"? Because sentient doesn't even have the letter "p" in it.

    just very very gullible and/or Dunning-Kruger-victims.

    A Dunning-Kruger "victim".....

    You completely missed the point

    Say that again into a mirror and then you'll actually be making a real and coherent point.

  • source
  • parent
  • [–] 1 point 7 hours ago (4 children)

    LLM'S do not represent and cannot be Artificially Sapient, and that I will grant you. I do not consider them intelligent because intelligence involves understand and the ability to understand context specifically.

    But that doesn't mean LLM'S don't fit some of the technical definitions of AI as seen in the other person's comment who responded to you. We have been calling game NPC's driving algorithms AI for decades. We have to accept that the term has multiple meanings and specify that what we mean when we say LLM's lack intelligence and or sapience.

  • source
  • parent
  • hideshow 4 child comments
  • [–] -1 points 6 hours ago (3 children)

    But that doesn’t mean LLM’S don’t fit some of the technical definitions of AI

    When faced with slop vendors arguing in bad faith, we shouldn't get caught up on technicalities. The "technically it is (a subset of) AI" argument is used to sell gullible people solutions that are already getting people killed - and not necessarily those who were gullible enough to believe the "AI" bullshit.

    So for LLMs, it should be common sense to call them slop / dangerous bullshit generators and refuse to talk with anyone who sells them as "AI".

  • source
  • parent
  • hideshow 3 child comments
  • [–] 3 points 6 hours ago (2 children)

    Two things.

    1. When we aren't specific in our arguments about this technology it allows people to blow past facts to suit personal interests and insist their viewpoint is correct because their perception trumps facts that aren't clearly defined.

    2. Pretending the symantics of a term aren't integral to the argument being made against it does us no favors when trying to educate people about the topic.

    The algorithm in my game analogy does what it does because the parameters for its instructions are finite and well thought out and well programmed to give the same result based on pre-defined instructions and parameters.

    LLM's do that same thing, but with a metric fuck ton more parameters and fewer defining instructions.

    The reason they keep failing with guardrails is because they put too many unvetted parameters into the algorithm and concise and colloquial instructions in spoken language cannot account for all the variables included in those parameters.

    It's a fatal design flaw but it's important to people's understanding that at the end of the day an LLM is a computer program and it does absolutely nothing unless it is instructed to do something. It won't perform tasks of it's own volition.

    Pretending that we can take shortcuts in linguistics to get people to understand the technical facts doesn't really work here. I'm not saying we can't call them slop/BS.

    Gullible people don't become less gullible because you changed some words.

  • source
  • parent
  • hideshow 2 child comments
  • [–] -1 points 5 hours ago (1 child)

    My main point is that the use of the term "intelligence" to describe sometbing that is absolutely without any understanding, by people who don't try to sell the snake oil and should have the interests of others in mind, is a disservice to educating people about the dangers.

  • source
  • parent
  • hideshow 1 child comment
  • [–] 3 points 5 hours ago

    Part of language is about whether or not there is an ability to infer what is being talked about without specifically naming it.

    Because LLM's have existed longer than the current technology of "AI LLM's" we have to specify what we're talking about. Which is why we use the term we do the same way you would possibly use the term Xerox or Hoover or Scotch tape.

    The difference is those terms I listed are a shorthand for other words that get the idea across without misinterpretation, but can easily be swapped for photocopy, vacuum, and cellophane tape.

    The technical term given to Generative AI LLM's that differentiates it from the LLM's that came before it is in fact Generative AI LLM's and is shortened to Generative AI or Gen AI
    When you use the term slop, that can be linguistically confusing because to a non-native English speaker or a farmer slop has a different connotation altogether.

    This does not negate your argument, because I generally agree that we are attributing intelligence to this tech that it does not have and that this was an intentional choice by the Tech Bros who coined the term. I even agree that it adds further confusion because it's being used for medical tech that doesn't use an LLM at all.

    But if we're going to change the name it needs to be a new term that represents what this tech actually is instead of using a more confusing/less specific slang that has other connotations.

    Doctors don't call vaginas pussies.

  • source
  • parent
  • [–] 1 point 9 hours ago (5 children)

    They are AI the same way ELIZA is.

  • source
  • parent
  • hideshow 5 child comments
  • [–] 5 points 9 hours ago (4 children)

    Nah, Eliza had a programmed logic, all phrases she reacted to had been implemented deterministically by a human. There was no training data.

  • source
  • parent
  • hideshow 4 child comments
  • [–] 2 points 7 hours ago* (3 children)

    At first I thought that maybe the "no match" responses, like "Please go on" were randomly selected from a list, but even there there was a pointer and iterator. Truly no randomness at all!

    (If there was any randomness I was going to cheekily argue that it was a "probabilistic model" :-) )

  • source
  • parent
  • hideshow 3 child comments
  • [–] 2 points 6 hours ago (2 children)

    The randomness in Eliza (I never reviewed the code) would have been of the sorts "out of x possible answers indexed starting with zero, do a truncate( ran( 1.0 ) * X ) and select the answer with that index". But yeah, apart from that it was deterministic. Same input history = same output.

  • source
  • parent
  • hideshow 2 child comments
  • [–] 2 points 5 hours ago* (1 child)

    No, my point is there is no rand call at all, check the source code here https://blog.adafruit.com/2025/01/21/the-code-for-eliza-the-original-1960s-chatbot-found/

    Edit: there is a line ELIZA MAD (which is just language module, but still I love it lol)

  • source
  • parent
  • hideshow 1 child comment
  • [–] 40 points 1 day ago (2 children)

    I also read that the AI used about $25 millions worth of tokens, to solve the problem (with a "prize" of $1 mil), not counting a lot of actual mathematicians supervising the works. Can you imagine if that sort of money was available to academics. I know a couple of math post grads and I feel they spend more time trying to secure grants than doing actual work, or at least that's what they are always talking about

  • source
  • hideshow 2 child comments
  • [–] 2 points 7 hours ago*

    to solve the problem

    Just a reminder that the NS problem is composed from 4 statements, AB making core, CD being variant that a lot mathematicians did not really care about; CD are the variants that OpenAi stole the potential solution to (as it is not verified yet).

  • source
  • parent
  • [–] 172 points 1 day ago (7 children)

    If AI can effectively steal the work of others, and claim it as its own, then it undermines the practice of openness in science that has allowed for enormous progress for centuries.

    Open source developers have been saying this for years, ever since the first ChatGPT was released. It was laundering open source code, claiming the work of others as it's own, ignoring the license.

  • source
  • hideshow 7 child comments
  • Thank you! It's so weird to see some pro AI people act like this somehow crossed a line. It's my understanding that part of the issue is that the AI used its chat logs to get some of this information, but I've seen no one explain how that is any different from what was done to people who believed their work was private or protected under a license but ended up having it in a training set anyway. I don't believe there's an expectation of privacy when using the plagiarism bot, and my understanding is it does reserve the right to train off your chats.

    The concern regarding attribution is also strange to me because it's possible that the attribution for any specific thing is completely lost during the training process. When people have it generate images or stories or even just a response, it's not going line by line explaining all the places it stole from. When people talked about AI bootstrapping itself there was no discussion of how we were going to keep track of all the people whose hard work it was building off of.

    I'm against AI in general for a bunch of reasons, but the fact they're upset they got scooped is so minimal in my opinion. Wasn't AI supposed to solve all these problems for us, faster than we could alone? It's doing exactly what they said it would, and since they never bothered to care about attribution before I'm having a hard time sympathizing with them now.

  • source
  • parent
  • [–] 20 points 1 day ago* (4 children)

    I stopped contributing to open source code because of it. I consider it stupid now to contribute to public work because it will be almost instantly stolen and laundered through an AI that breaches the license its published under.

  • source
  • parent
  • hideshow 4 child comments
  • load more comments (1 reply)
    [–] 27 points 1 day ago

    A similar take: https://www.thebignewsletter.com/p/is-artificial-intelligence-going

    I realized it was more realistic for [us] to imagine that bots would take over our civilization and supplant humanity than to imagine that we might apply the rule of law to Sam Altman. And that, to me, proved that we don’t have a crisis with technological advances, we have a crisis of the rule of law.

  • source
  • [–] 40 points 1 day ago (2 children)

    simple reminder: capitalism will do this, every single chance it gets. this isn’t an ai thing or a sam altman thing, it’s a capitalist thing. since forever.

  • source
  • hideshow 2 child comments
  • load more comments (2 replies)
    [–] 50 points 1 day ago (1 child)

    Why are we acting like it's news that generative AI companies are building their core products of the theft and exploitation of other people's intellectual labor?

    Get with the program, it's their entire business model.

  • source
  • hideshow 1 child comment
  • [–] 13 points 1 day ago

    The captured SCOTUS said its okay and no further discussion needed as they needed to hit the tarmac asap because they already booked a spin on their new private jet/gift but absolutely not a bribe.

  • source
  • parent
  • [–] 22 points 1 day ago (1 child)

    if corporations are people - time to hold them liable to the law - especially when they’re committing espionage and cyber crimes, and theft

  • source
  • hideshow 1 child comment
  • [–] 25 points 1 day ago* (last edited 1 day ago) (2 children)

    The metaphor of the open seminar is misleading. In an open seminar, the presenter stake a claim to the work, even if incomplete. In this case, the work was in private repositories. A more precise methodology would be the AI overhearing the authors discussing in their office and using that as base for a proof. This is even sneakier because there has been no public claim of the method or results by the authors.

    Edited after a commenter pointed out that no sleep + autocorrect = random words

  • source
  • hideshow 2 child comments
  • load more comments (2 replies)
    [–] 5 points 1 day ago (4 children)

    There's no low Altman won't sink to

  • source
  • hideshow 4 child comments
  • [–] [S] 14 points 1 day ago (3 children)

    The way they handled the situation is honestly bad, I wonder what the heck is their PR department is doing:

    https://cims.nyu.edu/%7Etristanb/statement.pdf

    https://xcancel.com/OpenAI/status/2097375276384567642

  • source
  • parent
  • hideshow 3 child comments
  • [–] 17 points 1 day ago (20 children)

    "guns don't kill people. people kill people."

  • source
  • hideshow 20 child comments
  • [–] 27 points 1 day ago (1 child)

    I think OP is trying to say that we cannot let "AI" launder the responsibility that real humans have for their actions.

    The point (as I understand it) of "guns don't kill people" is that it is patently absurd to imply that "tools" are wholly neutral implements. The quote is trying to hold the "tool" accountable. While it's making a seemingly contradictory point to OP, I think it's more that both can be true at once--OP is trying to highlight the specific problem with "the AI is the one doing all the naughty things", because Sam Altman et al. would certainly like everyone to think so. I totally agree that "AI" is not merely a "tool" (in the sense of the neutrality so-implied), but that is a separate point.

    (re-reading what I just wrote... it kinda sounds like clanker speak?? I swear I'm a real human)

  • source
  • parent
  • hideshow 1 child comment
  • load more comments (1 reply)
  • load more comments (1 reply)
    [–] 3 points 1 day ago
    load more comments
    view more: next ›