This is a lobby organisation using only the pieces and bits they like to push their own agenda.
"sketchy russian website", how about using some more clear description like: A library for sharing books and articles that should be partly public domain because they were paid for by the public. Only some of the material is copyrighted by authors. However, some of their work is so old that it is not reprinted anyway.
But of course such an explanation would not click.
I also don't see a problem with statements about making people jobless. Imagine if every robotic or automation company advertised like this: Yeah, you'll buy tons of expensive robots and still rely on expensive labor from real people without any efficiency gains.
It's just sad that this is the top comment on Hacker News. Why are we giving free pass to these tech companies? Why are we trusting these CEOs when they have repeatedly broken laws? Remember Aaron Swartz and the fate he suffered? Why is big tech getting away with so much more?
> Remember Aaron Swartz and the fate he suffered? Why is big tech getting away with so much more?
Why are you turning him into perpetuum mobile in his grave?
Do you really believe Aaron would be arguing against AI companies and for publishing / recording guilds on the grounds of intellectual property claims?
No, it's the tech community that did a sudden about-face, and is now all "friendship ended with free access to information and technologies enabling people; now RIAA is my best friend", and this move is as dumb as that meme (https://imgflip.com/memegenerator/137501417/Friendship-ended).
C: "Information is free only for the rich and not free for everyone else, giving the rich a material advantage over everyone else that not only entrenches but accelerates wealth inequality and impedes class mobility"
You, or Swartz, are an advocate for A. Why, exactly, do you think that obliges you/Swartz to prefer C over B while A is not true?
Once upon a time, copyright infringement for personal use was barely a crime, while copyright infringement by a for-profit commercial enterprise was a serious matter.
The idea being (before the rise of online peer-to-peer piracy) to prosecute the people making bootleg VHSes rather than the people buying them.
With the rise of these AI behemoths, it seems that rule is now inverted: You can download all the pirated ebooks you want, as long as it's for large-scale for-profit commercial use.
Sure it’s not free for anyone and both companies and individuals are treated similarly. It’s not like you will be jailed for pirating movies. And neither should OpenAI. What part of this is hard to understand
> It’s not like you will be jailed for pirating movies.
We're literally talking in a thread about someone who committed suicide because the US government was hellbent on ruining his life with a felony conviction for piracy.
Sounds like a gross misapplication of the law enforcement system. The free exchange of information should be legalised.
As we can see in the good article, the copyright system might almost have cost us a lot of AI capabilities. The damage it has done in cases where the lawyers got ahead of the builders is incalculable.
The labs aren’t doing this. They are doing something similar to you and I downloading torrents. Look, I also think laws should apply somewhat equally to individuals and companies. But this is different.
They certainly seem intent on producing and distributing materials based on that licensed material, and stripping out the licensing and pretending it never existed.
That doesn't seem different to me. Copyright, inherits.
If its fair use, like the companies currently claim, maybe it is different. But I suspect that the insane push to create a copyright carveout in countries around the world, is not unrelated to a judgement that it probably isn't fair use.
According to the people trying to ruin his life. This is an accused thought crime, not something he actually did. Also, supposing it actually was something he did, your argument is that building a trillion dollar business on stolen licensed material is legally permissible but giving it away for free is worthy of your life being ruined. Wonderfully coherent world view, that is. Piracy is fine, but only if you hoard it to yourself and profit from it!
Let’s just pretend that “poor” people who die because of our stupid laws (even now as this AI clown show is playing out) do not matter while the billionaires running these AI companies get a free pass and it’s OK.
Fun fact: Kim Dotcom is still fighting extradition while these drama queens (I.e Dario) are lecturing us about how much access the peasants should get to AI models fed and trained with stolen IP.
> Do you really believe Aaron would be arguing against AI companies and for publishing / recording guilds on the grounds of intellectual property claims?
That's not the point. All the rules and laws are enforced when its you and me but when it's big tech the laws are treated by these companies as mere instructions.
> friendship ended with free access to information and technologies enabling people; now RIAA is my best friend
Big tech will enable access to free information and will help people reach new heights. Do you see how wrong that sounds?
Calling out equivocation and out-of-context quotes is not “giving a free pass”. If there is a case to be made (and to be clear, I believe there is), then shouldn't be resorting to rhetorical slight-of-hand to “prove” it.
I think what Swartz did was moral, and his prosecution was unjust.
I think what OpenAI did was moral, and them getting sued for it is unjust.
Why do you have one position for Swartz and a different one for OpenAI?
(Aaron Swartz was a mailing-list friend of mine, so I do have some bias here. But in part we knew each other because our moral position on this was similar)
There seem to be two groups of people talking past each other here, which is worth acknowledging.
On the one hand, there is a notion of consistent principles regarding the legal handling of the topic, which is being appealed to by some people such as yourself.
The other side seems to be drawing attention to the fact that these principles are not applied consistently by society. They're questioning the moral validity of holding to principle in a circumstance where it's guaranteed to be applied with very specific biases that are rarely explicitly stated.
I've noticed this talking-past happening in other subjects too. For example American drug laws. There's the principle of drug laws, and then there's the practice of which kinds of people gets the laws applied to them. One group of people focus on the one, and another focus on the other, and they just kind of talk past each other.
> Why do you have one position for Swartz and a different one for OpenAI?
I think many people on HN dont mind OpenAI use of copyrighted material, but do not like how they try to do regulatory capture of a market and try to say their own copyright is now somehow more important.
Like how OpenAI trying to make "distillation" illegal while it exactly what they did with whole intetnet, books, everything.
Imagine a town where (a) a corrupt cop issues tickets for fake traffic violations; but (b) he doesn't do that to his fellow cops, or the mayor's friends.
Does (b) make things more just, as certain possible unjust acts don't happen?
Or does it compound the injustice by creating a double-standard, and perpetuate it by hiding the problem from any with the power to bring an end to it?
Sure. But the fact is they broke current existing laws, with known punishments with precedents. Same as a new law doesn't retroactively punish someone, then a new law shouldn't absolve someone before it's passed.
They did. They were found guilty. The case I looked at* was a civil case so this was settled out of court before the court imposed a settlement.
The law they broke was pirating the materials, not training per se, even though training is what so many people object to: the judge ruled that actually training a model, when the materials you used were ones you otherwise had lawful access to, was not a breach of law.
IMO, the laws need to change to reflect what tech can now do. This wouldn't be the first time, copyright law has had to shift several times before as new means of reproduction are created.
>A copyright is a type of intellectual property that gives its owner the exclusive legal right to copy, distribute, adapt, display, and perform a creative work, usually for a limited time.
The more interesting question is IMO if AI training actually falls into one of these cases. You can read a book and also copy it, but you do not do because of the law. However, you have the ability to do so. Is having the ability to do something already forbidden?
No, they were responding to the post defending OpenAI that you wrote. If you meant to communicate something other than “criticism of OpenAI in this context is unwarranted” then it looks like you forgot to do that and wrote something else instead
> Selective and out-of-context quoting isn't “criticism of OpenAI”, the same way that saying “checkmate” when your opponent isn't in check isn't chess.
Is quoting “criticism of OpenAI” without the rest of the post a way of saying “checkmate”?
> I also don't see a problem with statements about making people jobless
I do see a problem with a company loudly announcing that they are going to make people's lives miserable purely for profit. Leaving aside that it goes against OpenAI's stated mission ("to ensure that artificial general intelligence benefits all of humanity"), the disdain for the lives they are intentionally trying to ruin makes it a problem.
And even if you believe that the transition is inevitable, as it is the case with phasing out combustion engines in cars, anyone reasonable would see that the transition is gradual to give people time to adapt. Instead of doing that, OpenAI is burning cash at astonishing rates, polluting the environment, and killing personal computing with the only aim of being the only ones left atop the ruins. I do see a problem with that.
It can be a bit confusing due to the terrible style of the article (ironic given the source) but it seems the "sketchy russian website" part is a direct quote by Anthropic's Sam McCandlish. And apparently Dario Amodei referred to it as sketchy as well.
I find the brazenness of saying this while running what's arguably the largest copyright theft operation in human history astonishing. If libgen is "sketchy", then what is OpenAI?
Reading isn't the right comparison. Human memory is lossy and fades while LLM encoding is durable with no degradation. The valuable content that the author provides: the content, style, selection of topics, and more is encoded, written into the LLM weights, and they obtain profit from them (now directly, via ads). No one can compete with that kind of copying and pasting from copyright-protected material. And the scale is what hurts authors most: flooding the market with millions of copies on demand, without paying for it.
Human memory decays; similarly, LLMs do not have perfect recall. But even if I reread a book every month for 60 years and use the knowledge in there to build a 25 billion dollar empire, the only thing the author is ever going to get from me is the $10.25 the book cost me. And no court in the western world would say he's due more.
That's not been legally established, the litigation is ongoing. And if mere downloading and reading of copyrighted material were legal, how come torrent users have been fined for it in the thousands?
The law is the law, there can't be different law for corporations with billions in backing. I don't agree with current copyright laws btw and think they should be changed. However, they probably should have lobbied for that before illegally downloading all this material.
So we agree they have violated copyright at a much larger scale than LibGen, yet they call LibGen "sketchy" for doing the same thing? Absurd, what exactly are we arguing here?
We have proof that Meta torrented TBs of books off Anna's Archive. They didn't even buy the books legally in the first place. Fair use doesn't magically make piracy legal. The only reason they haven't been punished is the pathetic state of our government and court system.
The comment you're replying to is citing almost verbatim [1] Microsoft’s director of Applied Science, Brent Hecht, who called OpenAI's data collection practices "the largest theft of labor in human history" in an internal memo.
It is citing directly from unsealed court document. For an Opinion Haver training here all day, whose only other contribution to the world was animated Nyan Cat in Emacs status bar, you have poor comprehension skills.
> frame it in a dumb, manipulative way like that
> In fact, you're doing exactly that right here.
You're dead fucking right they are doing exactly that. They are saying that what is wrong is wrong.
Instead you seem to be making out that AI companies are some kind of victim that has to "worry" about "manipulation". Meanwhile authors are out of a job right now and not by accident. What's up with that?
There's no actual argument here, just name-calling. Meta literally pirated 80TB of books off of Anna's Archive. OpenAI pirated off LibGen as well. You could be prosecuted for doing the same thing, because its illegal. Piracy is illegal, piracy at the scale they committed is extremely illegal. This shouldn't be a hard or controversial concept.
Yeah maybe make a few author friends and tell them that they don't deserve any payment for their work. I'm sure the world will be much better once we even further reduce the incentives to do hard intellectual work and increase the incentives to churn out AI slop tiktok videos instead.
> They were just worried that the commentariat on HN will frame it in a dumb, manipulative way like that.
You’re chastising others for a tone you are yourself employing, and are making monumental assumptions based on a few choice quotes. From the quotes alone you can’t tell if OpenAI thought libgen was sketchy or not.
Also, contrary to what you’re claiming, they were wrong. HN in general seems to approve of libgen when used for its purpose of downloading some books on an individual level. The complaint you’re replying to is about what OpenAI did with the data, it has nothing to do with the website they got it from.
"This is a lobby organisation using only the pieces and bits they like to push their own agenda."
Do you have any evidence of them being a lobby organisation (as opposed to OpenAI for example which spends millions of dollars hiring actual lobbyists)
>The group lobbies at the national and state levels on censorship and tax concerns, and it has initiated or supported several major lawsuits in defense of authors' copyrights.
Never heard of them, but yes they are a lobby company. So they're biased and lobbying for their interests and OAI are biased and lobbying for their interests. That adds a lot of colour.
It is funny how hard people(?) or maybe bots shill OAI on here. That said why they care at all about some clown like myself "anonymous" opinion on HN is crazy. HN is not an illustrious mind share? A lot of the people here get butterflies thinking about laying off most of their staff for something cheaper. Hell some of the people here get the tinglies if something might be cheaper.
I'm more annoyed by the constant whack engagement bait that gets posted here and everywhere about AI. Here I am again engaging still not buying a subscription...
Any source that this was a library? Even then that would still raise a question if OpenAI is a Russian organisation or not to access that library with good faith.
apparently, the website in question is libgen.io (appears to have been taken down now), currently it has lots of mirrors like libgen.im , libgen.com.de, etc.
> A library for sharing books and articles that should be partly public domain because they were paid for by the public
This is some incredible mental gymnastics here, wow. Some books should be public domain (even if they actually aren't), and this magically justifies stealing from an 80TB library of a large fraction of every book ever published including millions of books that have no public funding.
“Sketchy Russian website” is part of a quote by Sam McCandlish (who worked at OpenAI), not the Author’s Guild characterisation.
Also, defending it on the basis that some books on libgen are public domain is a poor excuse, like claiming people use The Pirate Bay to download Linux ISOs. Even if some of that is true, we all know that use case is not the popular one.
"A lobby organisation?" Of course a single author would not be able to afford facing a multi billion dollar company on their own? And the "sketchy russian website" quote is from OpenAI employees themselves? What are you on about?
First, it's still a lobbying organization, so it's their job to make exaggerated claims, like a union in a company or any other organization with a political purpose. My first point was to highlight that it's not neutral or news related. It's fine that they have their opinion, but it's also my right to say that they're biased.
The second thing underscores my point. They use one line and think they've made a great point because one employee called LibGen sketchy. This site has been around since the 2010s, and it has helped many people do research. It's not just a sketchy website that suddenly appeared and is always doing bad things. I think a more nuanced stance is necessary.
> They use one line and think they've made a great point because one employee called LibGen sketchy.
No, you are using one line from the post to discredit them. The release has more than that and it’s not the only communication they made on this matter nor is there any indication it will be the last, it’s just the current one.
They are having it in the subtitle. In general their whole article is about two main points, first the use of stuff from Libgen, second the points of making people jobless. Three of their points are about the jobless thing two about the LibGen.
About LibGen, there might be more discussion - fair. However, the second argument is no real discussion IMO. Why is putting people out of work suddenly a bad thing? Since when do we argue this when talking about automation?
There is nothing to discuss about LibGen, really. I don't read this as OpenAI employees even believing LibGen is sketchy. They were worried about optics, because LibGen itself is Russian and does look a bit sketchy, and at the time - much like today - it was easy to make it a headline that makes people pattern-match to "troll farms".
(And then Russia invaded Ukraine, turning any association with .ru things into potential corporate suicide.)
Really has nothing to do with LibGen or with OpenAI. It's about people being easy to manipulate into believing bullshit, which is a reasonable worry, and the Authors Guild is trying to do that exact thing OpenAI was worried about.
It's wild how much weight "they're destroying jobs" has.
I dare say that car manufacturers are aware of their impact on the horse and buggy industry. Calculator manufacturers wrecked the livelihood of mathematicians and accountants.
Technology is in the business of putting people out of work, by inventing better ways of doing things. Or rather, any time you invent a better way of doing things, that's fundamentally going to disrupt all the businesses built around older technologies.
The issue is the breadth of scale. At the time of the industrial revolution those who we would consider knowledge workers were a much smaller percentage of the population. Today, they are a significant part of the workforce, and it is them who are being displaced first, and in increasing numbers.
The second issue is one of migration between skill sets. By the numbers, today is likely harder as non Ai-affected skilled professions take longer to train into in a general sense.
The third is the synthesis of AI and robotics. Humanoid help is good. As the Venn diagram overlap of capabilities between humans and humanoid robots increases, many professions defined by complex vision and environmental manipulation that are currently safe will not be.
There are, of course, upsides. But the world is right to wonder what the future looks like, and where people fit within that future in terms of earning an income if whatever path they choose seems to be able to be replaced by a machine within their lifetime. How do they achieve personal stability and safety if, even when thinking about their multitude of career options, the machines are so good that they can do almost anything.
Once the rich no longer need the serfs, they won’t suddenly decide to let you have things. They know they are superior to you, both genetically and socially because they inherited wealth and you didn’t. Their trust fund paid for private schools and you don’t even have a trust fund. They’re just better. It’s an objective fact. The world specifically chose to make them better than you, because it’s true.
The only fair system is for them to have an army of robots and then quietly clean up the population, so that you aren’t a risk to them. It’s just science.
When you’re dealing with people who think like this, who will soon have an army of terminators, what hope do we have?
I continue to be shocked at how people outright disregard these kinds of possibilities as pure science fiction. This is literally the objective of the ultrawealthy, it is the natural extrapolation of the early events of the industrial revolution projected onto a near-future world, and given current trends, the only hope of it not happening is that people everywhere realize what is happening and use our remaining capacity to steal power back. Feudalism is the natural state of mankind, and has been for thousands of years.
I am very hopeful that such scenarios would not come to pass, but dismissing them outright puts us on the path toward them.
Care to enlighten us on what is a good model of the world? Because an "elite vs. masses framework" of human civilization was what won the 2024 Nobel Prize in Economics.
Alright, maybe I shouldn't have used the term "feudalism", although this is one that most people understand. The correct term is an "extractive institution".
Not really. All that happened is that security got unbundled from the Landlord to the State in exchange for some protections against the (notorious) caprice and malice of some Landlords.
But the fundamental economic arrangement largely remains intact. The landed class extracts infinity wealth from production upon "their" land in perpetuity by virtue of a piece of paper that says they can physical force upon anyone who dares to be productive on it without paying the required tribute.
I read it more as a sarcastic reflection on the state of the world by the author rather than that the author believes what he write himself or agrees with it. It is a debating technique to push a contrary argument to the extreme to see where the balance lies.
Plato made a similar claim in The Republic. 2400 years later, and we still don't have a technocracy in the Western hemisphere where engineers lead countries. Either engineers are systematically oppressed, or there are just other qualities that make a leader than being good at complex problem solving.
So, how will you ensure that you have some influence over the world or power, if all human labour ends?
That scenario is precarious. Dangerous as fuck. I have a short story, basically a horror Isekai, portal goes straight into a place like Sarek or similar pretty northern fell landscape, only full of bears and savages. I think my protagonists are safer in that, than the average human in your scenario.
I dream of the end of paid or forced human labor. People should do whatever they want. I bet no one wants to commute to the crowded office for another teams call.
This could go the star trek or dystopian route, and which way depends on who controls the models.
There is going to be a very shitty in between time where we still need blue collar, but no white collar labor. The amount of conflicts out of that will be insane. And who controls the news?
We will never end in the star trek branch until energy and resources in general are abundantly cheap and uncontested. Even in Star Trek, the canon is still that next to whole modern part of society, even on earth, there still exists classes. It’s just that the TV series depict the upper class: everyone working for the federation and star fleet.
I wouldn't put AI on the same level as the inventions of the industrial revolution. Automation replaced specific tasks. AI is replacing humans themselves. What are humans supposed to do in a world where everything a human can do, a machine can do too?
There are a lot of tasks machines can do today that humans do anyway. Think baristas and bartenders. People pay more for human-made goods too, despite mass-manufacturing being a couple centuries old at this point. Having a human face is a competitive advantage in lots of stuff.
Walk me through that a bit. I’ve used codex to solve dozens of niggling annoyances around the house and in my life. That gave me more control and increased my ability to express my creativity.
At work, it enabled me to develop two apps, one complete (as much as they ever are), one nearing the end of technical PoC phase, and a handful of others where I could quickly answer a tech feasibility question.
The completed app is an interactive web app that I literally could not have completed to that level of polish (and from a pacing perspective probably couldn’t have completed at all).
All of those increased my control and most increased my ability to express creativity.
The upcoming technical PoC will involve considering using Clojure in a load-bearing app, something that I’ve considered and consistently rejected for a decade now. If that happens, control and creativity will spike as well.
Is there some aspect of this removal that I’m not seeing? It changes the activities of the tasks, for sure, but very far from that all being for me he worse; a lot of it is way better.
What other leverage, apart from labor, do masses have against the large capital owners? You imagine they'll give us basic income, just so we can buy products from them? That makes no sense at all.
You keep promoting assumptions to facts by calling them "reality." Just because you seem hell bent to analyze everything through the lens of class power, doesn't mean that I do.
What's the medium term plan though? Before we achieve singularity or UBI or whatever rapture analog this new religion is promising. Being employed is the primary way (and for most people is the only way) of getting housing, food, and other necessities meet.
Technology is in the business of putting people out of work, by inventing better ways of doing things.
It feels like that but the argument doesn't hold up to scrutiny. Practically every advance in technology leads to more jobs, usually to apply that technology in new areas, or to bring the benefit to more people. There are pockets of people who are negatively impacted, but the overall change to society has been positive for pretty much every advance humans have ever made (maybe saving for weapons.)
It's interesting that this argument wasn't popular decades ago, when the internet basically made publishers and other copyright rent seekers redundant. Instead the response was harsher enforcement and more laws to protect the rent seekers.
Amplification of commentary and opinion is much easier now. It’s also a lot easier to play into people’s fears.
There’s a wide range of political motivations for this sort of thing too. Not just AI taking jobs, but convincing people to vote against their best interests, or even supporting radical religious groups whose values would persecute you. This sort of rhetoric feels noisier than it ever did decades ago.
It's not. You be performatively hostile and people be hostile to you. That's all.
Car brands don't pitch cars as horse replacements. They put names and icons of horses on cars and horse riders love them. No one is complaining that horses had replaced cars.
AI startups pitched AI like it's a coagulation of malice and hostility against humanity on tap. People aren't liking it. Of course they won't. No intelligence would, natural or artificial. It's wild that they don't get that.
> No one is complaining that horses had replaced cars.
Search for something like “newspaper complaints about cars early 1900s” and read the contemporaneous complaints.
The fact that that complaining happened then and not now is evidence that motor cars are now widely seen as better than horse-drawn personal transport, not that they were eagerly welcomed from the first day of production.
"Destroying jobs" is entirely the wrong framing and a distraction from the real issue.
If AI truly replaces human labor, the lack of jobs isn't the problem. The problem is the lack of leverage for most of humanity.
Jobs aren't just a source of income. They're a source of leverage over the ownership class - if labor goes on strike, production stops and capital cannot self-reproduce. This leverage is what gave us a living wage, sick leave and basically all concessions that make life livable for anyone that isn't lucky enough to be born as part of the elite.
If AI makes this leverage go away, our problem is bigger than just job loss. The owners of the AI industry will use this now-untethered productive force to completely monopolize resources and set up a system where they are unquestionable god kings.
The rest of us will have more to worry about than just employment. We will be reduced to depending on the charity of people who's record has shown are not exactly the most selfless and kindhearted.
This is the real problem with AI productivity. Jobs are a distraction. Control over production is the real issue. The only way this doesn't end in dystopia is if the public controls AI, one way or another.
You twist the argument. Your argument would hold if AI's only use would be to generate booksverbatim it was already trained on. Which is certaintly not the case and huge efforts were made to circumvent this kind of usage.
Because you want as many people as possible to be exposed to your ideas, to the point that Christianity used to fund armies to go to other places so they could force the teachings of Jesus Christ upon them. If your thoughts aren't at least as good as that, why are you wasting eink on them?
Newly released court filings quote an OpenAI researcher saying: “I was just worried about optics - i.e. 'openai uses
copyrighted data from sketchy russian website’ showing up on HN would be unfortunate."
That's just one of several interesting quotes that have surfaced in documents from the Authors Guild's lawsuit against OpenAI.
Dario, if you’re reading this, I want you to know that Claude sucks now. Its output isn’t even English anymore, it’s just claudeslop.
You’ve actually ruined it so much that it has now polluted the Chinese models which are copying your work. So now all the models output claudeslop.
You’ve polluted all of the training data in the world. Now we’re never going to be able to train proper models, because everything has slop in it and all the sites have locked down their data to prevent future startups training.
Think they're referring to the following, when Amodei was still working for OpenAI:
'OpenAI Feared “Optics,” Not the Law – “Dario Amodei, OpenAI’s then-Research Director, responded that ‘as a training set [LibGen is] a bit sketchier.’ [OpenAI researcher Sam] McCandlish explained: ‘I was just worried about optics – i.e. ‘openai uses copyrighted data from sketchy russian website’ showing up on [Hacker News] would be unfortunate.”'
I was curious what he was responding to. Per https://authorsguild.org/app/uploads/2026/09/Class-Plaintiff... it was "On July 19, 2019, McCandlish wrote in an OpenAI Slack channel: “We’re not sure if we’re going to release the Foresight LM Scaling paper publicly, but if we do we were thinking about removing all mentions of LibGen, since it's a bit of a sketchy data source." The paper may or may not be https://arxiv.org/abs/2001.08361 where they write "we also test on similarly-prepared samples of Books Corpus [ZKZ+15], Common Crawl [Fou], English Wikipedia, and a collection of publicly-available Internet Books."
> Newly released court filings quote an OpenAI researcher saying: “I was just worried about optics (...)"
As a side-note, most orgs already cover the need to STFU in their training material for new hires, particularly how personal comments should not and cannot represent the company. I'm sure this lawsuit will be explicitly mentioned in upcoming versions of this sort training material in multiple orgs.
Legal incentives against better internal communication/coordination create so much inefficiency. The legal system should focus on outcomes more than internal comms.
Cute that they were concerned, but they need not have worried - as you can clearly see from any HN comments section on these kinds of articles, everyone long ago wrote down their conclusion (“OpenAI/Anthropic/LLMs are good/bad”) at the bottom of the page, and now all we are doing is shuffling facts, quotes, and arguments around the page above it.
We are very clever to make it look like the conclusion below follows from the arguments above; but we are also careful to never smudge our bottom line, so it will never change.
To all the top comments making analogies about destroying jobs; horses, buggies, calculators... what?! You are awful people. We're not talking about shit jobs, these are peoples creative passions, livelihoods, and life long careers. It looks like OpenAI had nothing to fear. Fuck you.
If this doesn’t put anyone in prison then we might as well declare copyright dead. They knew they broke the law and then they deliberately covered it up. How much more evidence is required here?
Can we fix the title? This is clickbait for HN, the actual article's title is "Unsealed Briefs in Authors’ Case v. Microsoft/OpenAI: Top Execs Knew Their Mass Book Piracy Was Illegal And Would Put Authors Out of Work"
I think you mean he's siding with Dario Amodei and the strictest of government regulations cementing Microsoft's position and its investments is IMMEDIATELY needed, or "a billion people will die":
His position is not at all that AI itself should be limited or anything more drastic than slowing down Microsoft's spending (sorry I mean slowing down AI progress). Second, he still wants companies to adopt it on a very large scale, and fire everybody, but he's having to spend too much now. To put it bluntly, he wants the money currently paid to employees, but he doesn't want to or outright can't spend enough much to guarantee it's him getting the money. I mean, this is a bet on his part, so we can't be sure, I even think he himself is not entirely sure, but it's pretty clear he's not comfortable.
(because in a winner-take-all monopoly business billionnaires get one chance to outspend everyone else. Everyone but the top dog doesn't get the monopoly, they get scraps. It has become clear China is the biggest spender and so now he thinks the billionnaire-but-still-not-the-biggest-spender urgently need protection from the bigger dog)
I'm giving the less charitable interpretation of his words, because his demands come down to denying access to Chinese models, and of course Bill Gates can hardly be considered a neutral party here, he personally has a huge financial interests in OS, Cloud and AI.
The sad part is this stragey often works because many in the general public don't see through these regulation lobbying schemes, however obvious they are.
The fearmongering always works on some part of the population as well. Feels like they were just waiting for the next end of the world scenario to be announced.
Easy, now. Pointing this out is reportedly against the rules, making the ground defacto plastic. In my opinion, of course, with the required curiosity.
Lol the original comment was only in half jest knowing the ass lickers who congregate here, but it went from 4 points to 0 within 5 minutes. _Not suspicious at all._
Can some explain why would a billionaire like Altman care about the opinions expressed on HN?
Given the political power they have, particularly now with the Trump administration, whatever "optics" exist on this forum seems completely insignificant
You forget how many influential tech professionals are present here (especially those in Silicon Valley) and that Altman was president of Y Combinator for several years.
Elon Musk seems to care a great deal about a lot of stuff that frankly isn't any of his concern, not really. Trump care immensely about shallow stuff that's really below what the president of the United States of America should spend time on.
Altman has some reason to worry, because developers and the users of HN are then ones he needs to promote/hype his products. OpenAI still needs customers and they need to sell tokens and a large number of users on HN are not only not buying, but are becoming increasingly hostile towards the product Altman is selling. My guess would be that he'd assume that the users on HN are in some sense his peers, and they are currently extremely divided in the question about the value and dangers of LLMs, and are increasingly critical of the business practises of OpenAI (and other AI companies).
That makes sense, but at the same time Altman has been famously called a "sociopath" by Aaron Swartz and a rapist by her own sister and yet he kept his position after the revolt a couple of years ago.
It's hard to imagine anything worse than bombing an elementary school and yet people are still using Anthropic products as if nothing has happened.
Yes it’s all hype to cash in the ipo before the bubble bursts. The labs are spreading spooky things about AI so that people think these models are very capable. It is a coordinated effort amongst all labs. The reality is that these models can’t do basic mathematics.
"sketchy russian website", how about using some more clear description like: A library for sharing books and articles that should be partly public domain because they were paid for by the public. Only some of the material is copyrighted by authors. However, some of their work is so old that it is not reprinted anyway.
But of course such an explanation would not click.
I also don't see a problem with statements about making people jobless. Imagine if every robotic or automation company advertised like this: Yeah, you'll buy tons of expensive robots and still rely on expensive labor from real people without any efficiency gains.
Why are you turning him into perpetuum mobile in his grave?
Do you really believe Aaron would be arguing against AI companies and for publishing / recording guilds on the grounds of intellectual property claims?
No, it's the tech community that did a sudden about-face, and is now all "friendship ended with free access to information and technologies enabling people; now RIAA is my best friend", and this move is as dumb as that meme (https://imgflip.com/memegenerator/137501417/Friendship-ended).
A: "Information is free"
B: "Information is not free"
C: "Information is free only for the rich and not free for everyone else, giving the rich a material advantage over everyone else that not only entrenches but accelerates wealth inequality and impedes class mobility"
You, or Swartz, are an advocate for A. Why, exactly, do you think that obliges you/Swartz to prefer C over B while A is not true?
The idea being (before the rise of online peer-to-peer piracy) to prosecute the people making bootleg VHSes rather than the people buying them.
With the rise of these AI behemoths, it seems that rule is now inverted: You can download all the pirated ebooks you want, as long as it's for large-scale for-profit commercial use.
We're literally talking in a thread about someone who committed suicide because the US government was hellbent on ruining his life with a felony conviction for piracy.
As we can see in the good article, the copyright system might almost have cost us a lot of AI capabilities. The damage it has done in cases where the lawyers got ahead of the builders is incalculable.
1. Unauthorised network access
2. Intent to distribute licensed material
The labs aren’t doing this. They are doing something similar to you and I downloading torrents. Look, I also think laws should apply somewhat equally to individuals and companies. But this is different.
That doesn't seem different to me. Copyright, inherits.
If its fair use, like the companies currently claim, maybe it is different. But I suspect that the insane push to create a copyright carveout in countries around the world, is not unrelated to a judgement that it probably isn't fair use.
According to the people trying to ruin his life. This is an accused thought crime, not something he actually did. Also, supposing it actually was something he did, your argument is that building a trillion dollar business on stolen licensed material is legally permissible but giving it away for free is worthy of your life being ruined. Wonderfully coherent world view, that is. Piracy is fine, but only if you hoard it to yourself and profit from it!
Fun fact: Kim Dotcom is still fighting extradition while these drama queens (I.e Dario) are lecturing us about how much access the peasants should get to AI models fed and trained with stolen IP.
That's not the point. All the rules and laws are enforced when its you and me but when it's big tech the laws are treated by these companies as mere instructions.
> friendship ended with free access to information and technologies enabling people; now RIAA is my best friend
Big tech will enable access to free information and will help people reach new heights. Do you see how wrong that sounds?
I think what Swartz did was moral, and his prosecution was unjust.
I think what OpenAI did was moral, and them getting sued for it is unjust.
Why do you have one position for Swartz and a different one for OpenAI?
(Aaron Swartz was a mailing-list friend of mine, so I do have some bias here. But in part we knew each other because our moral position on this was similar)
On the one hand, there is a notion of consistent principles regarding the legal handling of the topic, which is being appealed to by some people such as yourself.
The other side seems to be drawing attention to the fact that these principles are not applied consistently by society. They're questioning the moral validity of holding to principle in a circumstance where it's guaranteed to be applied with very specific biases that are rarely explicitly stated.
I've noticed this talking-past happening in other subjects too. For example American drug laws. There's the principle of drug laws, and then there's the practice of which kinds of people gets the laws applied to them. One group of people focus on the one, and another focus on the other, and they just kind of talk past each other.
I think many people on HN dont mind OpenAI use of copyrighted material, but do not like how they try to do regulatory capture of a market and try to say their own copyright is now somehow more important.
Like how OpenAI trying to make "distillation" illegal while it exactly what they did with whole intetnet, books, everything.
Does (b) make things more just, as certain possible unjust acts don't happen?
Or does it compound the injustice by creating a double-standard, and perpetuate it by hiding the problem from any with the power to bring an end to it?
There are several multi billion dollar companies where the founding thesis was “what if we just ignore the law?”
The law they broke was pirating the materials, not training per se, even though training is what so many people object to: the judge ruled that actually training a model, when the materials you used were ones you otherwise had lawful access to, was not a breach of law.
IMO, the laws need to change to reflect what tech can now do. This wouldn't be the first time, copyright law has had to shift several times before as new means of reproduction are created.
* the Anthropic one
The more interesting question is IMO if AI training actually falls into one of these cases. You can read a book and also copy it, but you do not do because of the law. However, you have the ability to do so. Is having the ability to do something already forbidden?
No, they were responding to the post defending OpenAI that you wrote. If you meant to communicate something other than “criticism of OpenAI in this context is unwarranted” then it looks like you forgot to do that and wrote something else instead
Is quoting “criticism of OpenAI” without the rest of the post a way of saying “checkmate”?
How? The current system enables the GPL. The GPL protects many open source projects.
Why has restrictive Linux succeeded far more than any BSD ever has?
You still haven't articulated how that's restrictive.
I did not bring up Linux, much less called it bad, yet you made up a strawman and started attacking it. Good day.
A license in isolation isn't interesting. The practical results of projects under a license is interesting.
Linux is a very successful practical result under the GPL and copyright makes the GPL work. Without copyright the GPL would be unenforceable.
I do see a problem with a company loudly announcing that they are going to make people's lives miserable purely for profit. Leaving aside that it goes against OpenAI's stated mission ("to ensure that artificial general intelligence benefits all of humanity"), the disdain for the lives they are intentionally trying to ruin makes it a problem.
And even if you believe that the transition is inevitable, as it is the case with phasing out combustion engines in cars, anyone reasonable would see that the transition is gradual to give people time to adapt. Instead of doing that, OpenAI is burning cash at astonishing rates, polluting the environment, and killing personal computing with the only aim of being the only ones left atop the ruins. I do see a problem with that.
I find the brazenness of saying this while running what's arguably the largest copyright theft operation in human history astonishing. If libgen is "sketchy", then what is OpenAI?
Many people think that it was fair use: training is akin to reading, not copying.
Especially the courts.
Exactly analogous to a human reading the material.
The law is the law, there can't be different law for corporations with billions in backing. I don't agree with current copyright laws btw and think they should be changed. However, they probably should have lobbied for that before illegally downloading all this material.
100% of the rulings agree with me.
The piracy is not in question. It is unarguably copyright violation.
But that's not what anyone means in this context. Training is what everyone means.
> The law is the law, there can't be different law for corporations with billions in backing.
I didn't say otherwise. That's a straw man.
No.
Judging by how AI threads look like for the past year, they were absolutely right to be worried.
> largest copyright theft operation in human history
In fact, you're doing exactly that right here.
[1] https://techcrunch.com/2026/09/17/microsoft-exec-called-ai-s...
You're dead fucking right they are doing exactly that. They are saying that what is wrong is wrong.
Instead you seem to be making out that AI companies are some kind of victim that has to "worry" about "manipulation". Meanwhile authors are out of a job right now and not by accident. What's up with that?
You’re chastising others for a tone you are yourself employing, and are making monumental assumptions based on a few choice quotes. From the quotes alone you can’t tell if OpenAI thought libgen was sketchy or not.
Also, contrary to what you’re claiming, they were wrong. HN in general seems to approve of libgen when used for its purpose of downloading some books on an individual level. The complaint you’re replying to is about what OpenAI did with the data, it has nothing to do with the website they got it from.
Do you have any evidence of them being a lobby organisation (as opposed to OpenAI for example which spends millions of dollars hiring actual lobbyists)
>The group lobbies at the national and state levels on censorship and tax concerns, and it has initiated or supported several major lawsuits in defense of authors' copyrights.
Have you looked at their name?
I'm more annoyed by the constant whack engagement bait that gets posted here and everywhere about AI. Here I am again engaging still not buying a subscription...
you should not enclose the commons.
that is what openai and anthropic have done; capture the commons, lock the model, restrict the outputs.
I think they just used a Russian torrent site.
>Microsoft knew about OpenAI’s use of LibGen as early as April 2019
(note: https://z-library.sk/ is prettier/nicer)
This is some incredible mental gymnastics here, wow. Some books should be public domain (even if they actually aren't), and this magically justifies stealing from an 80TB library of a large fraction of every book ever published including millions of books that have no public funding.
Plus you somehow didn't even read the article properly, given that "sketchy russian website" is part of a direct quote.
Also, defending it on the basis that some books on libgen are public domain is a poor excuse, like claiming people use The Pirate Bay to download Linux ISOs. Even if some of that is true, we all know that use case is not the popular one.
The second thing underscores my point. They use one line and think they've made a great point because one employee called LibGen sketchy. This site has been around since the 2010s, and it has helped many people do research. It's not just a sketchy website that suddenly appeared and is always doing bad things. I think a more nuanced stance is necessary.
No, you are using one line from the post to discredit them. The release has more than that and it’s not the only communication they made on this matter nor is there any indication it will be the last, it’s just the current one.
About LibGen, there might be more discussion - fair. However, the second argument is no real discussion IMO. Why is putting people out of work suddenly a bad thing? Since when do we argue this when talking about automation?
(And then Russia invaded Ukraine, turning any association with .ru things into potential corporate suicide.)
Really has nothing to do with LibGen or with OpenAI. It's about people being easy to manipulate into believing bullshit, which is a reasonable worry, and the Authors Guild is trying to do that exact thing OpenAI was worried about.
The whole "sketchy russian website" bit resolves entirely about being seen as associated or supporting troll farms and Putin.
EDIT: look at it this way: no one is calling Internet Archive "a sketchy US website".
I dare say that car manufacturers are aware of their impact on the horse and buggy industry. Calculator manufacturers wrecked the livelihood of mathematicians and accountants.
Technology is in the business of putting people out of work, by inventing better ways of doing things. Or rather, any time you invent a better way of doing things, that's fundamentally going to disrupt all the businesses built around older technologies.
The second issue is one of migration between skill sets. By the numbers, today is likely harder as non Ai-affected skilled professions take longer to train into in a general sense.
The third is the synthesis of AI and robotics. Humanoid help is good. As the Venn diagram overlap of capabilities between humans and humanoid robots increases, many professions defined by complex vision and environmental manipulation that are currently safe will not be.
There are, of course, upsides. But the world is right to wonder what the future looks like, and where people fit within that future in terms of earning an income if whatever path they choose seems to be able to be replaced by a machine within their lifetime. How do they achieve personal stability and safety if, even when thinking about their multitude of career options, the machines are so good that they can do almost anything.
Once the rich no longer need the serfs, they won’t suddenly decide to let you have things. They know they are superior to you, both genetically and socially because they inherited wealth and you didn’t. Their trust fund paid for private schools and you don’t even have a trust fund. They’re just better. It’s an objective fact. The world specifically chose to make them better than you, because it’s true.
The only fair system is for them to have an army of robots and then quietly clean up the population, so that you aren’t a risk to them. It’s just science.
When you’re dealing with people who think like this, who will soon have an army of terminators, what hope do we have?
I am very hopeful that such scenarios would not come to pass, but dismissing them outright puts us on the path toward them.
No, it's not. What we call feudalism arose after the collapse of the Roman Empire.
But the fundamental economic arrangement largely remains intact. The landed class extracts infinity wealth from production upon "their" land in perpetuity by virtue of a piece of paper that says they can physical force upon anyone who dares to be productive on it without paying the required tribute.
I'm amazed by the ignorance some people carry around.
That scenario is precarious. Dangerous as fuck. I have a short story, basically a horror Isekai, portal goes straight into a place like Sarek or similar pretty northern fell landscape, only full of bears and savages. I think my protagonists are safer in that, than the average human in your scenario.
I don't want influence or power. Don't know what your short story has to do with anything, frankly a bizarre comparison to make.
There is going to be a very shitty in between time where we still need blue collar, but no white collar labor. The amount of conflicts out of that will be insane. And who controls the news?
Or if it's even possible to control the models.
https://www.astralcodexten.com/p/book-review-deep-utopia
There is probably no industry to date with a smaller ratio of labor:capital in terms of the money being allocated
The auto industry led to a net increase in employment, including "unskilled" and lower class labor. The same cannot be said of LLMs
And one of the chief complaints about AI is that it is fundamentally an interpolation engine, which is to say uncreative.
At this time we still need to steer it, which is what the discussion about "taste" is about.
Though I suppose we'll need better terms for various kinds of "taste" soon enough if it's to become economically trackable.
At work, it enabled me to develop two apps, one complete (as much as they ever are), one nearing the end of technical PoC phase, and a handful of others where I could quickly answer a tech feasibility question.
The completed app is an interactive web app that I literally could not have completed to that level of polish (and from a pacing perspective probably couldn’t have completed at all).
All of those increased my control and most increased my ability to express creativity.
The upcoming technical PoC will involve considering using Clojure in a load-bearing app, something that I’ve considered and consistently rejected for a decade now. If that happens, control and creativity will spike as well.
Is there some aspect of this removal that I’m not seeing? It changes the activities of the tasks, for sure, but very far from that all being for me he worse; a lot of it is way better.
Yes, but we're the horse in that scenario.
Good! End human labor.
He's not putting that in your mouth, that is just the reality of what you need leverage for.
It’s wild how much people are scared for their livelihood. What a sad thing to read here.
It feels like that but the argument doesn't hold up to scrutiny. Practically every advance in technology leads to more jobs, usually to apply that technology in new areas, or to bring the benefit to more people. There are pockets of people who are negatively impacted, but the overall change to society has been positive for pretty much every advance humans have ever made (maybe saving for weapons.)
There’s a wide range of political motivations for this sort of thing too. Not just AI taking jobs, but convincing people to vote against their best interests, or even supporting radical religious groups whose values would persecute you. This sort of rhetoric feels noisier than it ever did decades ago.
Everybody gets a platform now, even the bots.
Car brands don't pitch cars as horse replacements. They put names and icons of horses on cars and horse riders love them. No one is complaining that horses had replaced cars.
AI startups pitched AI like it's a coagulation of malice and hostility against humanity on tap. People aren't liking it. Of course they won't. No intelligence would, natural or artificial. It's wild that they don't get that.
Search for something like “newspaper complaints about cars early 1900s” and read the contemporaneous complaints.
The fact that that complaining happened then and not now is evidence that motor cars are now widely seen as better than horse-drawn personal transport, not that they were eagerly welcomed from the first day of production.
If AI truly replaces human labor, the lack of jobs isn't the problem. The problem is the lack of leverage for most of humanity.
Jobs aren't just a source of income. They're a source of leverage over the ownership class - if labor goes on strike, production stops and capital cannot self-reproduce. This leverage is what gave us a living wage, sick leave and basically all concessions that make life livable for anyone that isn't lucky enough to be born as part of the elite.
If AI makes this leverage go away, our problem is bigger than just job loss. The owners of the AI industry will use this now-untethered productive force to completely monopolize resources and set up a system where they are unquestionable god kings.
The rest of us will have more to worry about than just employment. We will be reduced to depending on the charity of people who's record has shown are not exactly the most selfless and kindhearted.
This is the real problem with AI productivity. Jobs are a distraction. Control over production is the real issue. The only way this doesn't end in dystopia is if the public controls AI, one way or another.
Why is book piracy a better way of doing things?
You twist the argument. Your argument would hold if AI's only use would be to generate booksverbatim it was already trained on. Which is certaintly not the case and huge efforts were made to circumvent this kind of usage.
Clearly these books had value to AI companies but they were too weak and too dishonest to pay for that value. That's not impressive.
Obviously the authors should sue them to bankruptcy though.
Getting sued for fair value is the easier way out. Especially after a few billion in funding.
They have no will but to fill the world with hate and drivel
That's just one of several interesting quotes that have surfaced in documents from the Authors Guild's lawsuit against OpenAI.
You’ve actually ruined it so much that it has now polluted the Chinese models which are copying your work. So now all the models output claudeslop.
You’ve polluted all of the training data in the world. Now we’re never going to be able to train proper models, because everything has slop in it and all the sites have locked down their data to prevent future startups training.
'OpenAI Feared “Optics,” Not the Law – “Dario Amodei, OpenAI’s then-Research Director, responded that ‘as a training set [LibGen is] a bit sketchier.’ [OpenAI researcher Sam] McCandlish explained: ‘I was just worried about optics – i.e. ‘openai uses copyrighted data from sketchy russian website’ showing up on [Hacker News] would be unfortunate.”'
As a side-note, most orgs already cover the need to STFU in their training material for new hires, particularly how personal comments should not and cannot represent the company. I'm sure this lawsuit will be explicitly mentioned in upcoming versions of this sort training material in multiple orgs.
We are very clever to make it look like the conclusion below follows from the arguments above; but we are also careful to never smudge our bottom line, so it will never change.
the economics of those companies will cause a catastrophic wipe out of jobs across the board.
https://www.ft.com/content/1ade33b1-b5eb-43bb-9ee4-2272ec91f...
His position is not at all that AI itself should be limited or anything more drastic than slowing down Microsoft's spending (sorry I mean slowing down AI progress). Second, he still wants companies to adopt it on a very large scale, and fire everybody, but he's having to spend too much now. To put it bluntly, he wants the money currently paid to employees, but he doesn't want to or outright can't spend enough much to guarantee it's him getting the money. I mean, this is a bet on his part, so we can't be sure, I even think he himself is not entirely sure, but it's pretty clear he's not comfortable.
(because in a winner-take-all monopoly business billionnaires get one chance to outspend everyone else. Everyone but the top dog doesn't get the monopoly, they get scraps. It has become clear China is the biggest spender and so now he thinks the billionnaire-but-still-not-the-biggest-spender urgently need protection from the bigger dog)
I'm giving the less charitable interpretation of his words, because his demands come down to denying access to Chinese models, and of course Bill Gates can hardly be considered a neutral party here, he personally has a huge financial interests in OS, Cloud and AI.
The sad part is this stragey often works because many in the general public don't see through these regulation lobbying schemes, however obvious they are. The fearmongering always works on some part of the population as well. Feels like they were just waiting for the next end of the world scenario to be announced.
Given the political power they have, particularly now with the Trump administration, whatever "optics" exist on this forum seems completely insignificant
It's easier to take over a resigned population. And stating "resistance is futile" is cheap and it might weaken opposition.
Altman has some reason to worry, because developers and the users of HN are then ones he needs to promote/hype his products. OpenAI still needs customers and they need to sell tokens and a large number of users on HN are not only not buying, but are becoming increasingly hostile towards the product Altman is selling. My guess would be that he'd assume that the users on HN are in some sense his peers, and they are currently extremely divided in the question about the value and dangers of LLMs, and are increasingly critical of the business practises of OpenAI (and other AI companies).
It's hard to imagine anything worse than bombing an elementary school and yet people are still using Anthropic products as if nothing has happened.
Optics don't seem to matter much?
Yet they solved a Millennium Prize problem, you people are delusional at this point.
> AI agent accidentally publishes OpenAI’s unreleased model weights
The optics are bad. The submission marks a turning page in human history.