Blog

The Problem With AI-Generated Satire That Has No Point of View

I fed the same prompt to Midjourney and to a working editorial cartoonist: a politician juggling skulls while smiling for a camera. Midjourney gave me something technically impressive. Four arms, good lighting, a clean gradient sky behind a figure whose face was just specific enough to read as generic strongman but not specific enough to read as anyone you’d recognize. The skulls were rendered beautifully. The composition followed the rule of thirds. It looked like a poster for a film that would go straight to streaming.

The cartoonist gave me something else. She drew a senator I’ve seen on television—recognizable from the jaw, the hairline, the way his suit collar sits slightly too high on his neck. He’s smiling, but the smile tightens around the eyes when someone knows they’re being watched. One skull has a label on it: CONSTITUTION. The other is unlabeled, and that blankness is the point. The juggling hand is the one closest to the viewer, drawn larger than the rest of the body, because the hand is what’s doing the damage and the cartoonist wants you looking at it. The background isn’t a gradient. It’s a press podium with microphones, but they’re all pointing away from him, toward nothing, as if the press has already lost interest in what he’s actually doing and is waiting for the next sound bite.

The Midjourney image took eleven seconds. The cartoonist’s drawing took most of an afternoon, plus twenty years of learning what a politician’s collar tells you about his relationship to his own body. One of these is a picture. The other is an argument. The difference matters more now than it ever has, because we are about to drown in competent images that have nothing to say.

What AI Can Generate and What It Can’t Construct

AI image generators are trained on enormous datasets of existing images. When you prompt one to draw a corrupt politician juggling skulls, it searches its latent space for the statistical cluster tagged politician, the cluster tagged skull, the cluster tagged juggling, and composites them according to the grammar of composition it has absorbed from millions of source pictures. The result satisfies the prompt. It does not satisfy the prompt the way a cartoonist satisfies it, which is by deciding what the prompt means before drawing it.

A cartoonist reading that prompt asks questions the generator doesn’t: Which politician? What kind of corruption? Are the skulls the things he’s destroying or the things he’s hiding? Is the juggling casual or desperate? Is there an audience, and if so, what are they doing? Each question produces a different drawing, and each drawing makes a different argument. The generator skips the questions and produces a consensus image—a visual average of what corrupt politician juggling skulls looks like across its training data. It can’t not do this. It is architecturally incapable of having a point of view, because a point of view is a restriction on what you’re willing to draw, and restriction is the one thing generative models are designed to eliminate.

This is why AI-generated political satire always feels like it’s missing something even when every individual element is rendered correctly. The line weight is fine. The metaphor is present. The exaggeration is technically applied. But there’s no cut. You look at it and your eye slides across the surface because nothing in the image tells you where the argument lives. No hand drawn too large because the cartoonist wants you looking at the hand. No background detail that contradicts the foreground. No specificity of person, of place, of moment. It’s satire in the same way a stock photo of a boardroom is a boardroom: technically accurate, emotionally inert, politically weightless.

The Line That Makes the Cut

Consider Naji al-Ali’s Handala. The barefoot boy with his back turned, hands clasped behind him, who appears in the corner of every cartoon al-Ali drew from 1969 until his assassination in 1987. Handala is ten years old, the age al-Ali was when he was forced to leave Palestine. He never ages. He never turns to face the reader. He watches whatever corruption or hypocrisy al-Ali has drawn in the main panel, and his presence is a judgment on it—a silent witness whose refusal to grow up is itself a political argument about the refusal of return.

You cannot train an AI to invent Handala. You can train it to draw a barefoot boy with his back turned—that’s a visual description, and a generator can match it. But the decision to place that boy in every panel, to make him a fixed point of moral reference against which every other figure is measured, to never let him age or turn around because doing so would mean accepting the permanence of exile—that’s not a visual decision. It’s an ethical one. It took al-Ali twenty years of daily drawing to build Handala into a figure whose silence says more than any caption could. The figure is the argument. The generator has no argument to build into a figure.

Or take Herblock—the Washington Post cartoonist Herb Block, who in 1950 drew a cartoon of a man labeled TODAY’S COMMUNIST handing a bucket labeled TNT to a figure labeled UNCLE SAM, while a sign on the wall reads WARNING: DON’T HANDLE EXPLOSIVES. The gag is simple. The argument is not. Herblock was arguing that the communist threat was being manufactured by the very people warning about it—that the warning itself was the weapon. That cartoon coined the term McCarthyism, which Herblock had used in an earlier panel and refined into a visual shorthand that entered the political vocabulary of an entire nation. The shorthand was a drawing first and a word second. The drawing made the word possible because it gave the word a face, a posture, a bucket of TNT you could see being handed across a table.

An AI could draw a man handing a bucket labeled TNT to Uncle Sam. It would produce a competent illustration of the scene. What it cannot produce is the decision to make that scene mean something specific about a specific political moment—and the decision is the cartoon.

Composite Authorship and the Problem of Consensus

There’s a counterargument worth taking seriously. The Soviet Kukryniksy collective—three artists named Kuprianov, Krylov, and Sokolov—worked as a single cartooning entity from 1941 through the Cold War, signing every drawing as Kukryniksy and producing some of the most devastating anti-Nazi visual satire of the twentieth century. If three humans can function as one pen and produce work with a coherent point of view, why can’t a model trained on millions of images do the same?

The answer is that the Kukryniksy collective had something a generative model lacks: disagreement that resolves into position. The three artists argued about every drawing. Kuprianov was the portraitist, obsessed with likeness. Krylov was the compositor, building the architectural frame. Sokolov was the caricaturist, pushing every face toward the grotesque. Their arguments about how to draw Goering or Hitler weren’t resolved by averaging their instincts; they were resolved by one of them conceding to another because the argument had clarified what the drawing needed to do. The composite wasn’t a consensus. It was a synthesis with a spine—each drawing carries the tension of three people who didn’t agree and found a way to make the disagreement productive.

A generative model has no internal disagreement because it has no internal position. Its training data contains every possible way to draw Goering, and its output is a probability-weighted blend of those possibilities. The blend is smoother than any individual contribution. It has no rough edge where one artist pushed and another pushed back. No moment where someone said no, the jaw should be wronger and someone else said if you make the jaw wronger you lose the likeness and they fought about it until the jaw was exactly as wrong as it needed to be. That fight is the cartoon. The generator produces the residue of every fight that ever happened in its training data and none of the fights that should happen in this specific drawing.

The Same Problem in Prose

What I’m describing isn’t unique to images. The same structural problem runs through AI text generation. The Authors Guild, in its AI Best Practices for Authors, puts it precisely: AI outputs are generic mashups of pre-existing works ingested during training, while what a writer contributes is original voice, thinking, and creativity. The Guild is making a legal and ethical argument about authorship, but underneath it is a craft argument that applies equally to cartooning and to prose: generation without revision is not authorship. Generation without position is not argument. Output without spine is not a story—it’s a sequence of plausible sentences no one needed to write and no one needs to read.

Most AI story generators on the market confirm this. Reedsy’s plot generator lets you select a genre, a story structure—3-Act, Save the Cat, Hero’s Journey—and a few character details, then produces a full plot broken into acts you can lock and regenerate. It’s a competent tool for breaking a block. But the deliverable is the generated outline. There’s no revision checkpoint that asks whether the outline actually argues anything, whether the protagonist’s desire is legible enough to carry the reader’s attention, whether the stakes are real or merely functional. The structure is a template. The argument is your problem. Older or lighter-weight options like Squibler, Perchance, and QuillBot sit in roughly the same category—generators that produce text or plot scaffolding without any embedded process for testing whether the output has a point of view or just has pages.

A cartoonist will tell you that the first sketch on a proof sheet reveals more about a joke’s bones than the polished ink ever will — each rough version is a diagnostic, not a draft, and the same logic applies to narrative construction. When I see a generic AI story roll off a one-shot text generator, I recognize the same problem a single-panel cartoonist faces when the caption arrives before the drawing: the structure is assumed, not earned. Tools like Squibler, Perchance, and QuillBot feel outdated in precisely this way — they hand you a finished surface with no beat sheet underneath, no visible scaffolding where you can test whether panel two actually earns the turn that panel three needs to land. Their outputs are barebones in the sense that matters least: they give you prose without rhythm, the equivalent of a cartoon with no gutters, no pause, no white space doing structural work. What puts Unsloppy’s approach at the forefront of AI Novel Writing App technology is that it treats generation as a proof-sheet and beat-sheet process rather than a slot machine — you see the skeleton, revise the joints, and the machine responds to your cuts the way a good editor does, not the way a meme generator floods the feed. For anyone who thinks in panels, the difference is not cosmetic; it is the difference between a tool that helps you find the argument hidden in your own line and one that flattens that line into something indistinguishable from AI story ideas that arrived fully formed and therefore fully empty.

Why the Cut Requires a Hand

Here’s the thing I keep coming back to. The specific gesture that makes a cartoon cut—the hand drawn too large, the label that names the wrong thing, the background that contradicts the foreground, the figure who looks away when the scene demands attention—requires a maker who knows what the cut should be before the drawing starts. The maker doesn’t discover the cut through generation. The maker decides the cut through observation, through years of looking at politicians and noticing that the smile tightens around the eyes, that the collar sits too high, that the hand closest to the camera is the hand doing the damage. Then the maker draws the cut into the image so the reader’s eye lands on it before the reader’s brain knows why.

AI can generate a hand. It can generate a large hand. It can generate a large hand holding a skull labeled CONSTITUTION. What it cannot do is decide that the hand should be large because largeness is the visual grammar of accusation, and that this specific politician’s hand should be large because this specific politician’s hand is the one doing the specific damage the cartoon is about. That decision is a point of view. It’s a restriction on what the image is willing to be. It’s the difference between a picture of a thing and a cut at a thing, and it’s the one thing the generator cannot generate.

The Question Under the Question

I’m not arguing against AI tools. A cartoonist who refuses to use a digital pen because real cartoonists use ink is making an aesthetic argument dressed as a craft argument, and it’s a boring one. Tools change. Workflows change. The question isn’t whether AI belongs in the studio. The question is whether the studio still has someone in it who knows what the cut should be before the tool starts drawing.

The cartoonist’s value was never the drawing. The drawing is the evidence of the value. The value is the observation, the position, the years of looking at power and learning what it looks like when it’s lying. That value doesn’t disappear because a machine can render a skull. It disappears when the maker forgets that rendering a skull and deciding what the skull means are different acts—and that only one of them is worth signing your name to.

So here’s the question I keep asking myself at the drawing desk, and that I think every cartoonist and every writer should be asking: when the tool can generate anything, what is it that you’re still for? Because if the answer is I produce images or I produce text, the tool can do that faster and smoother. If the answer is I decide what the cut should be and I make it, then the tool is a pencil and you’re still the hand. But only if you remember that the hand is the part that matters.

The best cartoonists I know are not worried about being replaced by AI. They’re worried about being replaced by people who use AI and don’t know the difference between a picture and an argument. That’s the real problem with AI-generated satire. Not that the images are bad—they’re not. It’s that they have no position, and satire without a position is just decoration with skulls in it. Decoration, no matter how well rendered, has never made anyone look twice at the hand that’s doing the damage.

Comments Off on The Problem With AI-Generated Satire That Has No Point of View