this post was submitted on 24 Jun 2026
139 points (80.9% liked)

Selfhosted

60951 readers
586 users here now

A place to share alternatives to popular online services that can be self-hosted without giving up privacy or locking you into a service you don't control.

Rules:

Detailed Rules Post

  1. Be civil.

  2. No spam.

  3. Posts are to be related to self-hosting.

  4. Don't duplicate the full text of your blog or readme if you're providing a link.

  5. Submission headline should match the article title.

  6. No trolling.

  7. Promotion posts require active participation, with an account that is at least 30 days old. F/LOSS without a paywall has exceptions, with requirements. See the rules link for details. Tags [CBH] or [AIP] are required, see the links in Rule 8 for details.

  8. AI-related discussions and AI-involved promotional posts have additional requirements for tagging, as noted in Rule 7 and the AI & Promotional Post Expanded Rules post, and find example disclosures here.

Resources:

Any issues on the community? Report it using the report flag.

Questions? DM the mods!

founded 3 years ago
MODERATORS
 

Do you host your own ML / AI / LLM? What do you use, and what do you use it for?

(page 4) 50 comments
sorted by: hot top controversial new old
[–] Reygle@lemmy.world 1 points 1 month ago

I prefer my critical faculties completely intact and un-altered, thank you very much.
I do not require or desire a 400 watt bullshit-artist yes-man or vulnerability coder cooking my GPU.

[–] ccunning@lemmy.world 1 points 1 month ago

I’ve got ollama setup with whisper and piper and a HA voice PE, but I honestly haven’t gotten around to configuring much yet. Most notable thing was being able to use the wake word to start a timer, but it was pickier than old Siri about the precise wording.

[–] namelivia@lemmy.world 1 points 1 month ago

No, too expensive. I wish I could but it doesn't make sense financially for me right now, it is much cheaper to buy openrouter credits from time to time

[–] daniskarma@lemmy.dbzer0.com 1 points 1 month ago* (last edited 1 month ago)

Yeah, mostly for translation purposes.

I think I currently have gemma 4 set up.

[–] alexcleac@szmer.info 1 points 1 month ago

I've been running ministral on CPU on a home-server: works pretty nicely, not very performant for everyday tasks and the savings were not sufficient for it to make sense. It still was cheaper and faster to just use Mistral API and get better models.

[–] irmadlad@lemmy.world 1 points 1 month ago (8 children)

I've tried just about most of the small models. Tried NanoClaw. I just don't have the equipment necessary to pull that off and make it a worthwile, in house tool rather than an in house oddity. I really, really want to tho. So much so that I have been looking at what it would take to accomplish that, which seems to be at the $4k to $5k USD range. The sweet spot for GPUs seems to be at the 32 gb level. It is pricey, but hell, at my age, I figure wtf....I should treat myself. Whats wrong with that? If I do pull the trigger, I want it to be a LTS type computer like I built 15 years ago and is still running like a champ today tho it's probably worth less than a quarter of what I had invested. So, I'd probably overstock it to the max.

load more comments (8 replies)
[–] Egonallanon@feddit.uk 1 points 1 month ago

I've fiddled around with a few models on ollama and opencode but more for the sake of seeing what I can run as ive yet to really find a use for it in my home usage.

[–] Bluefruit@lemmy.world 1 points 1 month ago

I'm still messing around with self hosting llm, rn ive settled on using lumo from proton if I use an llm.

When I have run llm, I used koboldcpp. Works pretty well, depends on what you are doing and what models you use. Forget which models ive been using off the top of my head

[–] dotAlexX@lemmy.world 1 points 1 month ago

I would love to run and host a local LLM on my phone just to tinker and learn. I found a tutorial on setting DeepSeek on your Android phone using Termux but it is a year old. I'm sure there are better more efficient LLMs that can run on a phone now.

[–] vegetaaaaaaa@lemmy.world 1 points 4 weeks ago
[–] PapaSkwat@lemmy.wtf 0 points 1 month ago* (last edited 1 month ago)

I host my own AI, mostly for testing and because I wanted something that was mine and mine alone. I use Ollama and run models like Llama, Mistral, and Qwen. I honestly don’t use it much, but I wanted to have my own setup just in case online services go down or become less available. It’s part of my whole “own everything I use” mantra that I’ve been on lately.

load more comments
view more: ‹ prev next ›