15 comments

  • gwern 25 minutes ago
    This has been a perennial problem with my own generated comics with both Nano Banana Pro and ChatGPT (all generations). I often have to put in an extra edit to erase the false signature. It is annoying and I'm unsurprised most users don't bother.
    • bulder 5 minutes ago
      Despite what all the clickwrap warnings and "AI can make mistakes" subtitles might lead you to believe, the service offering of AI is explicitly designed to be as "one and done" as possible. The inherent nature of these tools is to service laziness, and disincentivize too much scrutiny.
    • johnnyanmac 15 minutes ago
      >I'm unsurprised most users don't bother.

      And then people wonder why the default mood of AI is so pessimistic. It's just revealing all of society's broken windows and adding a few more in the process.

  • MarkusQ 11 minutes ago
    Plagiarism as a Service.

    We can all try really hard to pretend that's not the business model, but that's totally the business model.

  • dormento 49 minutes ago
    The copyright washing machine strikes again.
    • Joel_Mckay 11 minutes ago
      https://en.wikipedia.org/wiki/Suchir_Balaji

      Whistleblower mysteriously found dead after exposing industrial scale Copyright abuse. Isomorphic plagiarism of $9Tn in FOSS and user content is the product. =3

      • joegibbs 7 minutes ago
        Why would they kill him for this? Everyone already knew they were doing it, where else would the data be coming from?
  • nyorai 7 minutes ago
    It seems like all the criticisms of Gen AI and LLM seems to concentrate on OpenAI and their products over products from anthropic and others. I am pretty sure that this faking of signature can be done by Gemini, claude as easily as chatgpt.
    • heartbreak 3 minutes ago
      Claude doesn’t do image generation?
  • WD-42 20 minutes ago
    Engineering manager at my company put a comic at the end of our sprint demo that was signed bloper. Except it wasn’t funny at all, and kind of weird. I asked him, sure enough it was ChatGPT and he didn’t notice the signature.
  • deforestgump 4 minutes ago
    Butlerian Jihad yesterday
  • ThrowawayR2 17 minutes ago
    Nice. This demolishes the "LLMs can reason" (but not enough to avoid this sort of basic error) and "humans make mistakes too" (not like this) talking points from the LLM promoters.
  • pollefeys 12 minutes ago
    signatures in, signatures out
  • baubino 11 minutes ago
    > “It’s like somebody attributed a quote to me that I didn’t say.”

    > Katzenstein considers the reproduction of his signature by ChatGPT to be more than just a violation of intellectual property; to him, it’s closer to false impersonation. “[ChatGPT] is attaching my name to work that I do not endorse or like. It’s slop, and unlike the other slop that I’ve encountered, this is slop that’s pretending to be me.”

    > “I’ve had people hack my credit card,” said Joe Dator, a New Yorker contributor for the past 20 years. “That feels like less of a violation than this. When they hacked my credit card, they didn’t dress up like me.”

    So this has morphed from plagiarism and copyright infringement (bad) to impersonation (also bad, arguably worse, and maybe more provable in court). It’s chilling to think of the implications of having one’s signature attached to a document or to words that are not one’s own.

  • dyauspitr 30 minutes ago
    Makes sense. Pretty much every example it is trained on has a signature.
  • zzzeek 32 minutes ago
    I follow anti-LLM discourse quite a lot, and across the main bulletpoints: energy/carbon emissions, content worker harm, job displacement, deskilling, mental health effects, and copyright/plagiarism, the plagiarism one seems to have the most attention, and it's also the most solvable, if there were only more serious effort on ethically sourced models that can actually do the real science / math / code work that is what LLMs are best at. The whole world of LLMs to create videos/books/art/literature is where most of the offense is (the video/imagery side of it is where most of the energy/carbon emissions problems are too. and content worker harm).

    I really wish there'd be a split among these disciplines (science/math/code vs. videos/art/literature) - one is vastly more problematic than the other.

    • UqWBcuFx6NV4r 25 minutes ago
      Yep. I’d probably be a lot less chastised in some circles for using Claude Code at work if it wasn’t misconstrued as being in support of, I don’t know, encroaching on the hypothetical commissions of a chronically online instagram furry artist or something.

      It is very tiring to say “I don’t necessarily disagree with you about AI ‘art’, but in my field—which you do not understand, and in which the underlying build process is often not the creative output—AI presents very real productivity gains” for the umpteenth time.

      I am skeptical of there being sufficient data to build “ethical” training datasets, and I’m confident that much of the same contingent will (somewhat rightfully) argue that ‘second-generation’ copyrighted AI material has already irreversibly made its way into every modern dataset.

      • forthegains 4 minutes ago
        Yeah and it only took stealing all the intellectual output of everyone ever made.

        But sure, there's um, an ethical way of doing that?

      • sublinear 16 minutes ago
        > has already irreversibly made its way into every modern dataset

        The "gray goo" scenario finally happens... for AI. That's actually the good ending for humanity. I love it! Poetic and believable. Data doesn't "heal" like nature. :D

      • johnnyanmac 3 minutes ago
        [delayed]
    • altermetax 17 minutes ago
      I don't really see the difference, code is protected by copyright (or copyleft) as much as art is, and yet the LLM scrapers use it without scruples. Same goes for math and science publications.
    • vouaobrasil 18 minutes ago
      > I really wish there'd be a split among these disciplines (science/math/code vs. videos/art/literature) - one is vastly more problematic than the other.

      I disagree that they can be separated. Practically, I think they can't. Because the mere invention of new tools inspires even more AI advancement and that in turn will cause the other side (artistic side) to degenerate even more.

      I'm anti-LLM all the way, 100%, no exceptions. Zero tolerance.

    • PunchyHamster 14 minutes ago
      They can't be ethically sourced and good at the same time.

      The current models intelligence depends on massive training dataset of essentially stolen data

    • asa123 27 minutes ago
      i think math is close to art (just to be a contrarian, but kind of really)

      its somewhat funny that math people are in a conundrum as to support or not support but this might partially be because some wish to believe that math itself is and can be useful and therefore accelerating is good

      but the art people have no such delusions so they’re just strictly against

      imo proof writing is more akin to art than coding/tech but…

    • singpolyma3 19 minutes ago
      Models aren't in school so plagiarism doesn't really apply
      • Loughla 7 minutes ago
        I can't tell if you're serious or not.
  • skybrian 45 minutes ago
    Whatever the original intention, this is clearly a bug and should be fixed. But should ChatGPT sign its cartoons with its own name or leave them unsigned?
    • kzsh 37 minutes ago
      I think what you're seeing is the probability of a particular signature or style of signature appearing on a particular style of cartoon, not an intent to sign.
      • ctippett 1 minute ago
        This. It's also while you'll sometimes get a mangled Getty Images watermark on some image generations, or a logo in the bottom left corner. If it's a prominent feature in the training dataset it'll show up, exactly how these models are supposed to work.

        The 'bug' here is whatever post-processing step or system prompt is in place to steer the model away from doing this.

    • DonsDiscountGas 12 minutes ago
      SynthID watermark but no visible signature
    • saalweachter 38 minutes ago
      I mean, the bug is, "it generated the most likely cluster of pixels in the corner of a New Yorker cartoon".
    • johnnyanmac 13 minutes ago
      We call it a "watermark" when a machine "signs" its work. And I wouldn't be surprised if its required as a part of AI disclosure in the coming years.
    • Forgeties79 41 minutes ago
      No one should be allowed to claim they drew something when they didn’t draw any part of it and LLM’s are not people/can’t work without a person. We don’t credit pens and paintbrushes after all.

      One could argue nobody should be allowed to claim it. It just exists.

    • dorkwood 17 minutes ago
      Why is it a bug? If other parts of the generated illustration are similarly taken from an artist, why not the signature as well? Why is a signature crossing the line but the rest of the image isn't?
      • DonsDiscountGas 11 minutes ago
        For the same reason I'm allowed to draw, paint, or write things very similar to what others have drawn, painted, or written but I have to sign my own name not theirs.
        • pessimizer 5 minutes ago
          You are a person, LLMs are not. You know this, which is why you know that if you signed someone else's name it would be forgery, but when you see the machine do it you call it a bug.

          If the machine is like you, the machine is a forger. The machine is not like you, it is simply blending the work of others to order. Adding someone else's signature is simply part of that statistical process.

  • parineum 11 minutes ago
    This is a pretty solid argument against people who argue that LLMs are more than just (very massive) next token predictors.

    If there was any thought or underlying thought going on here not putting a signature (at least a real one) would be the right move, despite it being less likely. It would realize, while generating the pixels that eventually became a signature, that it shouldn't do that.

  • aaron695 27 minutes ago
    [dead]