Local LLM stack that doesn't fight you
Running Qwen + Llama locally on a 4090. Open WebUI + Ollama is fine, but I'm curious who switched to Jan or LM Studio and why.
Also: anyone got a sane setup for sharing a local model across a small team?
Running Qwen + Llama locally on a 4090. Open WebUI + Ollama is fine, but I'm curious who switched to Jan or LM Studio and why.
Also: anyone got a sane setup for sharing a local model across a small team?
ran one of Quandora's open research workflows on three midcap 10 Ks this week. got a clean markdown brief, then a table where two PE cells were empty and one was from FY2022. no error toast. just qui…

Official now: https://www.cnbc.com/2026/09/03/nvidia agrees to buy hugging face for almost 13 billion ai expansion.html $12.9 billion. Second biggest Nvidia deal ever after the Groq assets. Delangue…
LangChain + Mem0 + vLLM behind Open WebUI. Sounds like alphabet soup but it runs our internal docs bot. Happy to share the boring parts (evals, chunking, failure modes) if useful.

3 comments
Join the discussion
Log in to comment.
Open WebUI + vLLM for the shared box. Ollama for solo laptops. Split saved us headaches.
Jan is nicer offline. Open WebUI wins for multi-user.
Tailscale + one beefy machine running Open WebUI. Works for our 4-person team.