this post was submitted on 10 Sep 2026
278 points (90.6% liked)
Technology
88303 readers
2906 users here now
This is a most excellent place for technology news and articles.
Our Rules
- Follow the lemmy.world rules.
- Only tech related news or articles.
- Be excellent to each other!
- Mod approved content bots can post up to 10 articles per day.
- Threads asking for personal tech support may be deleted.
- Politics threads may be removed.
- No memes allowed as posts, OK to post as comments.
- Only approved bots from the list below, this includes using AI responses and summaries. To ask if your bot can be added please contact a mod.
- Check for duplicates before posting, duplicates may be removed
- Accounts 7 days and younger will have their posts automatically removed.
Approved Bots
founded 3 years ago
MODERATORS
you are viewing a single comment's thread
view the rest of the comments
view the rest of the comments
I don't recall that one...but in fairness...AI generates a metric shit ton more code than humans. We'd have to normalize the results. Interestingly, looking it up now, someone DID normalize for that very case. Bug rate per commit for the AI-assisted versions landed within normal historical range. A pre-AI release had more regressions. The 3.4.3 regressions were primarily from the CVE security patches, not the AI work. Zero CVEs from the Claude-assisted commits.
EDIT: Correct URL https://alexispurslane.github.io/rsync-analysis/
The author of that article admits that the sample size is not large enough to draw meaningful conclusions.
But besides that, I believe LLM code generators can be a useful tool, provided you are willing to go over their output with a fine-tooth comb and assume it is broken until you have proven otherwise, because the hallucination problem is inherent to the technology and they're never going to completely solve it, and are willing to overlook the myriad ethical issues with all major LLMs in existence today.
Hey, you brought it up - I just pulled the thread. Not my fault it cuts against your argument.
So, exactly like a junior dev?
That's a different claim than the one you opened with though. Most of those objections have documented counter-arguments, btw:
https://blog.andymasley.com/p/a-cheat-sheet-for-conversations-about
https://aicentral.substack.com/p/why-anthropic-burned-the-books
Data centres were polluting long before LLMs arrived. Crypto, cloud storage, Netflix, YouTube, AWS - AI isn't all data centre load. Blaming the latter for the former is like blaming sunscreen for melanoma.
I'd really like to see a source for the first guy's numbers, especially how he accounts for things like training and water usage at the power plant, and as for the second guy, Anthropic is not preserving shit. The Internet Archive preserves books. Anthropic scans them and doesn't publish the scans so that they can train an LLM that might or might not be able to regurgitate some fragments of that text, and publish that. That's preservation in the same way that painting a picture of you is keeping you alive forever.
Also, neither of those address the effects on creatives' livelihoods or the mental health of LLM users. Chatbot psychosis is real. People who routinely use LLMs to do things provably become worse at doing them without them. Students use LLMs to make an end run around having to learn everything they need to know to be effective members of a society, like how to articulate their points, how not to fall for rhetorical traps, what history was really like, and between that and the disastrous effects of No Child Left Behind, teachers are quitting in droves and there's a literacy crisis.
The citations are inline hyperlinks throughout the piece - IEA, Lawrence Berkeley National Lab, MIT Technology Review, Google's own efficiency data.
They're there. Click them.
We've now gone from "AI can't code," to "AI code is a malicious risk," to "humans would never do that," to "that's a disengenous example" to "ok, but what about this, this and this"
We're verging on a gish gallop at this point, so I demur. Let's stick to the claims at hand instead of litigating shifting goal posts.
On the topic of the second article -
What the judge ruled was that it qualifies as fair use specifically because they destroyed the copies - the destruction is what made the scanning fair use, not what made it wrongful.
Retaining the digital files without destroying the physical copies would have been the violation.
Meaning the law perversely incentivised this behavior. (If you want the actual court reporting, the Ars Technica piece the article quotes is the cleaner source).
https://arstechnica.com/ai/2025/06/anthropic-destroyed-millions-of-print-books-to-build-its-ai-models/
So, it's more complicated than "Anthropic are book burners."