Self-hosting LLMs

posted 2 months ago

I’d like to self host a large language model, LLM.

I don’t mind if I need a GPU and all that, at least it will be running on my own hardware, and probably even cheaper than the $20 everyone is charging per month.

What LLMs are you self hosting? And what are you using to do it?

Sort:

Hot Top Controversial New Old

[ - ]

Karna@lemmy.ml

4 points

2 months ago

My (docker based) configuration:

Linux > Docker Container > Nvidia Runtime > Open WebUI > Ollama > Llama 3.1

Docker: https://docs.docker.com/engine/install/

Nvidia Runtime for docker: https://docs.nvidia.com/datacenter/cloud-native/container-toolkit/latest/install-guide.html

Open WebUI: https://docs.openwebui.com/

Ollama: https://hub.docker.com/r/ollama/ollama

permalink

report

[ - ]

𝕸𝖔𝖘𝖘@infosec.pub

1 point

2 months ago

GPT4All and Jan.AI are good places to start.

permalink

report

[ - ]

Nexy@lemmy.sdf.org

2 points

2 months ago

I run locally mistral-nemo in my 1070-ti

permalink

report

[ - ]

astrsk@fedia.io

4 points

2 months ago

If you don’t need to host but can run locally, GPT4ALL is nice, has several models to download and plug and play with different purposes and descriptions, and doesn’t require a GPU.

permalink

report

[ - ]

theshatterstone54@feddit.uk

2 points

2 months ago

I second that. Even my lower-midrange laptop from 3 years ago (8GB RAM, Integrated AMD GPU) can run a few of the smaller LLMs, and it’s true that you don’t even need a GPU as they can run in RAM. And depending on how much RAM you have and what GPU, you might find models performing better in RAM instead of on the GPU. Just keep in mind that when a model says, for example, 8GB Memory required, if you have 8GB RAM, you can’t run it cuz you also have your operating system and other applications running. If you have 8GB video memory on your GPU though, you should be golden (I think).

permalink

report

parent

[ - ]

InverseParallax@lemmy.world

7 points

2 months ago

Ollama, llama3.2, deepcode and a bunch of others.

Using a GPU but man they’re picky, they mostly want Nvidia gpus.

Do NOT be afraid to run on the cpu. It’s slow, but for 1 user it’s actually mostly fine.

permalink

report

Selfhosted

!selfhosted@lemmy.world

Create post

A place to share alternatives to popular online services that can be self-hosted without giving up privacy or locking you into a service you don’t control.

Rules:

Be civil: we’re here to support and learn from one another. Insults won’t be tolerated. Flame wars are frowned upon.
No spam posting.
Posts have to be centered around self-hosting. There are other communities for discussing hardware or home computing. If it’s not obvious why your post topic revolves around selfhosting, please include details to make it clear.
Don’t duplicate the full text of your blog or github here. Just post the link for folks to click.
Submission headline should match the article title (don’t cherry-pick information from the title to fit your agenda).
No trolling.

Resources:

selfh.st Newsletter and index of selfhosted software and apps
awesome-selfhosted software
awesome-sysadmin resources
Self-Hosted Podcast from Jupiter Broadcasting

Any issues on the community? Report it using the report flag.

Questions? DM the mods!

Community stats

3.7K
Monthly active users
2K
Posts
23K
Comments

Community stats

Community moderators