this post was submitted on 01 Oct 2026
147 points (96.2% liked)

Showerthoughts

43884 readers
324 users here now

A "Showerthought" is a simple term used to describe the thoughts that pop into your head while you're doing everyday things like taking a shower, driving, or just daydreaming. The most popular seem to be lighthearted clever little truths, hidden in daily life.

Here are some examples to inspire your own showerthoughts:

Rules

  1. All posts must be showerthoughts
  2. The entire showerthought must be in the title
  3. No politics
    • If your topic is in a grey area, please phrase it to emphasize the fascinating aspects, not the dramatic aspects. You can do this by avoiding overly politicized terms such as "capitalism" and "communism". If you must make comparisons, you can say something is different without saying something is better/worse.
    • A good place for politics is c/politicaldiscussion
  4. Posts must be original/unique
  5. Adhere to Lemmy's Code of Conduct and the TOS

If you made it this far, showerthoughts is accepting new mods. This community is generally tame so its not a lot of work, but having a few more mods would help reports get addressed a little sooner.

Whats it like to be a mod? Reports just show up as messages in your Lemmy inbox, and if a different mod has already addressed the report, the message goes away and you never worry about it.

founded 3 years ago
MODERATORS
 

It's hard to put this into words succinctly. But when I was a kid before the Internet, the library was the main source of human knowledge, and it was very organized through the Dewey Decimal System which imposed a kind of tree-like hierarchy across a wide spectrum of topics. And while I imagined that one day this might all wind up served by computers, I thought this organization would survive the transition. But it didn't.

I suppose in the early days of the Internet, they tried? Yahoo was kind of an Internet directory in the beginning, and the Whole Internet catalog was another such effort. But these eventually fell apart and we had to resort to search engines to find anything. Frankly, it's a bit like the early days of personal computing when we used flat file systems instead of hierarchical ones with proper directory trees. When you were limited to what could fit on a floppy, this wasn't such a big burden. But the Internet as it stands might as well be a giant flat file system.

Today, even the search engines are failing us, and we are turning to AI. In the pre-Internet era, we had sort of an equivalent to this also. They were called librarians. But a librarian's job was made easier by the fact that all the books were carefully organized by topic. For AI, it's as though the library were just a massive pile of books tossed around haphazardly and the librarian had to make sense of it enough to pull whatever you're looking for out of the chaos. Is it at all surprising, then, that it takes a huge amount of energy and computing resources to make any of this work?

you are viewing a single comment's thread
view the rest of the comments
[–] GreenShimada@lemmy.world 6 points 1 day ago (3 children)

I do some freelancing that involves online desk research, and for some topics, it's like pulling teeth to find a source written by a real human. Even non-Google search engines can't help but give me dozens of obviously AI-written pages. Which is why I as a human am getting paid to do real research. It's literally wading through slop to find anything a human did.

The thing is, AI companies are terrified of model collapse because if they train AI systems on AI output, the model becomes useless. But the rate at which websites on any topic are just slop being scraped to train a new model seems to be either unsustainable for using the open internet for that training, or an assured way to reach model collapase quickly.

[–] Eq0@literature.cafe 2 points 1 day ago

There are different theories around model collapse, so there might be a way to avoid it even if most of your training data is AI generated. The jury is still out though

[–] ricdeh@lemmy.world 1 points 1 day ago

The thing is, AI companies are terrified of model collapse because if they train AI systems on AI output, the model becomes useless.

That's actually not a fact, more of a possibility. As far as am I aware, reinforcement learning from AI feedback (RLAIF) is currently employed by some of the most capable frontier AI companies; Anthropic openly admits to it, and their models are arguably the best.

A term that is often thrown around nowadays is recursive self-improvement (RSI), and it seems that there are at this time more knowledgeable proponents of RSI for intelligence development than opponents who raise the risk of model collapse.

But yes, training on today's sloppified public internet would probably only have downsides.

[–] tunetardis@piefed.ca 2 points 1 day ago

That is just so depressing. It's eating its own garbage to create more content. Like some image that's been run through multiple passes of lossy compression with different algorithms.