Self-hosted AI: local LLMs with Ollama + Open WebUI #33
Labels
No labels
area/ci
area/media
area/network
area/observability
area/platform
area/security
area/storage
area/web
type/project
type/task
No milestone
No project
No assignees
1 participant
Notifications
Due date
No due date set.
Dependencies
No dependencies set.
Reference
itguyeric/infra#33
Loading…
Add table
Add a link
Reference in a new issue
No description provided.
Delete branch "%!s()"
Deleting a branch is permanent. Although the deleted branch may continue to exist for a short time before it actually gets removed, it CANNOT be undone in most cases. Continue?
Run local models on the fleet instead of reaching for a cloud API. Ollama to manage and serve models, Open WebUI as the chat front end. Uses the RTX for inference, so plan GPU allocation since Plex and a possible Immich ML container also want it.
Doubles as the local engine for other services: Paperless-AI extraction (#28) and anything we wire up later. New bootc image with GPU passthrough like the Plex box, behind SWAG, LAN-only or over Tailscale.
Blocked by: Need AI-capable server