(New York, September 21, 2026) Explosive new documents released last Thursday in Authors Guild v. OpenAI—the major book authors’ ongoing lawsuit alleging OpenAI and Microsoft illegally trained their products on…
621 comments
The fact that they believed this is legally useful because of the definition of fair use, so I see why the Author’s Guild is emphasizing it, but the authors I know don’t talk about it. They’re very angry about their work having been used without compensation to create the model, but not because they think it can replace them.
Whenever I’ve seen AI researchers talk about the possibility of AI writing fiction they always sound very confused about why people read novels.
I’m excited to be speaking about this topic at the Monktoberfest later this week.
Perhaps "the fact they believed this [i.e., the infringement was intentional] is legally useful" for seeking statutory damages, up to $15,000 per infringed work (assuming registration was timely, otherwise up to $7,500), as this requires the plaintiffs to show defendants acted willfully
Without intent, statutory damages could be as low as $200 per work
Anthropic settled for $1.5B with Bartz et al. in an earlier case involving the same pirated copies. The Court in the Bartz case said that training LLMs with these pirated copies was not fair use
https://www.npr.org/2025/09/05/nx-s1-5529404/anthropic-settl...
""The training use was a fair use," he wrote. "The use of the books at issue to train Claude and its precursors was exceedingly transformative."
However, the judge ruled that Anthropic's use of millions of pirated books to build its models, books that websites such as Library Genesis (LibGen) and Pirate Library Mirror (PiLiMi) copied without getting the authors' consent or giving them compensation, was not. He ordered this part of the case to go to trial. "We will have a trial on the pirated copies used to create Anthropic's central library and the resulting damages, actual or statutory (including for willfulness)," the judge wrote in the conclusion to his ruling. Last week, the parties announced they had reached a settlement."
The hilarious outcome of this saga is people using LLMs to rescue the readers of the Game of Thrones series, since he has no intention of finishing it himself.
Do OpenAI employees really believe LLMs can replace GRRM? I'm tempted to say no but it's so crypto coded, I find myself thinking yes. It's again people who have no idea of what it takes to make a successful written work thinking they don't need to know how it works.
I wish it was just AI researchers. I recently recommended a non-fiction book to a friend, and they said they don’t have time to read anymore. They read an LLM summary instead (popular book, probably in the training data), and said they agreed with the core concepts. I was honestly stunned. The whole interaction felt completely foreign to me.
In TFA the comment about GPT-X finishing George R. R. Martin’s series without involving the author felt extremely dystopian to me. I felt enraged for hours afterwards. Have these people never read a book themselves for the pure joy of it? Were they only interested in the outcome or the “takeaways”?
It is probably why sales of things like self-help books are collapsing.
So given that the literary work was outsourced, anyway, is it really that big of a deal that it’s getting outsourced to a machine? If that still bothers you, then how much machine assistance is too much? We’re already using Word processors, spell checkers, and other tools that have taken away much of the human effort. These language models don’t really have a strong sense of personal agency or what ought to be written, but they can fill out the details from a high-level schematic, which is probably, at best, what the author was doing with other team members.
That's just one of several interesting quotes that have surfaced in documents from the Authors Guild's lawsuit against OpenAI.
You’ve actually ruined it so much that it has now polluted the Chinese models which are copying your work. So now all the models output claudeslop.
You’ve polluted all of the training data in the world. Now we’re never going to be able to train proper models, because everything has slop in it and all the sites have locked down their data to prevent future startups training.
As a side-note, most orgs already cover the need to STFU in their training material for new hires, particularly how personal comments should not and cannot represent the company. I'm sure this lawsuit will be explicitly mentioned in upcoming versions of this sort training material in multiple orgs.
lack of ethics and integrity is of course part of the design of a capitalist economy.
you can't change the economy as an individual but i like this quote: "Do the right thing for the right reasons, at the right time with the right people, and you'll have no regrets for the rest of your lives."
"Our model did an oopsy woopsie for the 35th time" does not seem like a valid legal defense.
Top AI Execs Knew Their Mass Book Piracy Was Illegal And Would Put Authors Out of Work
Sadly, like keeping track of "attempted burglaries"[0], it's impossible to know the extent.
[0] how do you count attempts where the burglar was unsuccessful/left no trace and nobody was home to notice the attempt?
Twitter is also mentioned:
"356. In July 2020, OpenAI employee Ryan Lowe assessed the risk of continuing to use LibGen for the book-summarization project. Lowe wrote that he thought “there’s a >80% chance that we have some exchange of the form: ‘where did you get the books data?’” and “‘we can’t say’[.]” Nelson Decl. Ex. 325 at -315. Lowe estimated “a further ~40% chance that that leads to a moderate-sized Twitter kerfuffle that negatively affects the external perception of our work.” Id. Lowe added: “if we’re fully okay with these potential outcomes, then I’m comfortable continuing using Libgen for the project.” Id."
https://authorsguild.org/app/uploads/2026/09/Class-Plaintiff...
Originally this was the top comment _and had a sub-thread^1 underneath it_
1. https://news.ycombinator.com/item?id=49866516
The sub-thread has now been detached
I am still consistently astounded by how often the people working on this space seemingly have zero understanding of what art is, how it functions socially, or why it's important. They quite literally don't seem to comprehend the distinction between art and fan fiction. It's baffling. I'm really beginning to think courses in art history and literature need to be mandatory. We have failed legions of stem students when it comes to cultural education.
We aren't exactly consistent with how we approach this either. For example, I agree that reading "GPT-5 will autocomplete [GRRM's] series" feels "wrong" in the same way that if Terry Pratchett's "Disc World" series was "continued" by some non-approved author that would also feel wrong.
But contrast that with something like Star Wars, when George Lucas dies, I don't think anyone is going to feel any strong discomfort with some random person writing new Star Wars stories, even if those stories use the canonical characters. I suppose Disney might have a problem with it, but as a society, I don't see very many people losing sleep over someone not licensed by the Disney corporation writing more Star Wars. Likewise Star Trek. Gene Roddenberry is long gone, and while Paramount has ownership of the IP, if someone wrote their own Star Trek stories, no one is going to feel like they don't "comprehend the distinction between art and fan fiction".
And for further contrast, consider IPs that are well and truly part of the public domain. Cthulhu was the work of one author but since his death the lore has been expanded by multitudes of people, and no one finds that distasteful or tone deaf. No one thinks the "Hades" or "God of War" series of games don't qualify for "art" because the characters and lore being written about were the works of dead authors and the new material is certainly not "authorized" by those authors or their descendants.
Obviously time and distance plays a part of this, but as a more contemporary example, I wonder how many people would lose sleep or feel any significant discomfort over unauthorized or even AI generated Harry Potter works, either now or post JK Rowling's death. At the very least, if this quote read that Gogineni would "rest easy knowing that even though JK Rowling has [lost the plot/is an awful person/pick your reason for disliking her or her later work], GPT-5 will autocomplete her series." would that make people as equally uncomfortable? In my estimation, I would guess it would fall somewhere between the GRRM version and something like new Star Wars material for most people.
For example, Harry Potter and Star Wars are clearly IPs in which the creator has zero qualms handing over the rights for all kinds of extensions of the universe, movies, tv shows, other books, plays, toys, amusement parks, etc.
This pretty easily removes these works from the sphere of art and more into the sphere of "pop culture".
In the GRRM case, we already have some extensions of the IP, tv shows, board games, video games. However, it's not nearly as severe as some others.
In either case, I still think anyone who thinks they can "complete" a series by an author doesn't fundamentally understand art. It's a logical impossibility. I'd feel just as uncomfortable with this proposition applied to JK as I do to GRRM. Authorial intent matters. Even if the author is "dead" once the work leaves their finger tips, art is not a perception of the natural world, it is a form of expression which makes the particular human being behind it essential to its constitution, for better or worse. Many people understand this intuitively. The death of the author is more about the fact that an author cannot control the interpretations that ensue once a work is published, it's not about the author somehow being totally fungible.
Even in cases in which works are "completed" by other parties (because the original author died and asked them to do it, for instance) the parties involved have a clear notion of trying their best to do it "as the author would have intended". This is impossible, but it speaks to the human understanding that the artist is in fact in some way an essential component of the art.
"Completing" a source work is way different than spinning up a derivative work, partly because the primary author is not involved, so the derivative is, definitionally, not an instance of their self-expression or understanding or framing of the world. It's a totally different thing.
As modernmech put it, a conception of art without the artist or humanities without the humans is basically just a fully commoditized conception of art. It is not art at that point. It is art reduced to goods and services.
Read the full thread on Hacker News →
Related stories
- Ars Technica · 0 points · 13 days ago
- The Verge · 0 points · 12 days ago
- DEV Community · 0 points · 3 days ago
- The Verge · 0 points · 6 days ago
- The Verge · 0 points · 6 days ago
- Hacker News · 55 points · 4 days ago