this post was submitted on 04 Oct 2026
291 points (91.0% liked)

Technology

88390 readers
3293 users here now

This is a most excellent place for technology news and articles.


Our Rules


  1. Follow the lemmy.world rules.
  2. Only tech related news or articles.
  3. Be excellent to each other!
  4. Mod approved content bots can post up to 10 articles per day.
  5. Threads asking for personal tech support may be deleted.
  6. Politics threads may be removed.
  7. No memes allowed as posts, OK to post as comments.
  8. Only approved bots from the list below, this includes using AI responses and summaries. To ask if your bot can be added please contact a mod.
  9. Check for duplicates before posting, duplicates may be removed
  10. Accounts 7 days and younger will have their posts automatically removed.

Approved Bots


founded 3 years ago
MODERATORS
you are viewing a single comment's thread
view the rest of the comments
[–] adespoton@lemmy.ca 70 points 2 days ago (11 children)

AI didn’t decide anything. The humans failed to properly define the rules, and their translation model took the path most likely to succeed.

If it had been told not to use human created bots to win, it would probably have reverse engineered the game, found an exploit, and leveraged that instead. Because using the human interface to play/win the game is not the most efficient or dependable or easy to figure out method.

[–] Seralth@piefed.seralth.com 1 points 1 day ago (1 children)

I actually wanted to automate some large scale testing for vintage story. Figured i would set my local LLM on the task to sort it out just to see what it would do. I drafted up a nearly three page document with clear instructions, rules, tools, examples and goals. Put hard limits on the sandbox the LLM runs in so that it couldn't choose to just ignore the rules that could cause security issues and i let it lose.

It started with basic mouse and keyboard inputs and figured out by it self how to launch the game and run it though the user interface. After about 4 hours it stopped. Stated in its logic that what it was doing is "inefficient and wasting time" Then proceeded to promptly start working on a way to directly interface with it by designing a bot, getting a smaller model i had on file that could load along side it and drive the bot. It then started working on the hard problems would hand basic instructions to the smaller llm and it would drive a bot that loaded into the game as a mod.

After about 12 hours of total work it basically created a useful and well designed and functional vintage story bot and testing system. Would have likely taken me twice as long to design the bot.

Its been working well for about two weeks now. If i had just vibed out a half assed request or put in no hard safeguards outside of the LLMs control it likely would have done something fucking stupid. As with anything, its almost ALWAYS user error. And only an idiot blames their tools for their own fault.

[–] RumRunningDevil@lemmy.zip 2 points 17 hours ago

What a young account to come into this world, arrive in this thread about an AI violating containment, and then tell an unrelated anecdote about how well an AI agent did a vaguely complicated task when properly harnessed.

Certainly a human being is on the other side of this post and not, say, one of the countless bots out of Eglin AFB.

Certainly

load more comments (9 replies)