Openwebui + open terminal

Reddit r/LocalLLaMA News

Summary

The author describes how integrating Open Terminal with Open WebUI enables a self-hosted AI agent capable of searching large legal documents and generating notes via shell commands, using Qwen 27B on Ollama. They highlight that terminal access allows the model to reason step-by-step and extract relevant sections from documents far exceeding the context window.

Context: I don't code. My use is document research and document creation (mainly for legal search) searching inside large documents like a tax code (500+ pages) and building notes or pptx from what comes back. I've been running Open WebUI for a while on my Unraid box, pointed at the API of my inference machine (5060 Ti + 5070 Ti). I tinkered a lot. I tried Hermes on my main machine against the same API. It worked well but it was complex, and a bare-metal install made me uneasy. I also tried LM Studio Bionic with good results, but it didn't fit how I wanted inference organised (using ollama on the inference box). What I actually wanted was a self-hosted agent that works with Open WebUI while keeping things safe and under control. At one point I considered installing a harness like Hermes or Pi on each client and just connecting to the API instead. In the end I gave Open Terminal a shot. It's the companion container from the Open WebUI project that gives the model a shell — you run it as its own container and connect it through Integrations, so it isn't installed inside Open WebUI itself. Mine runs unprivileged, on bridge, with appdata mounted at /home/user. The model gets a shell in a box, not on the host. That was the part I cared about. It has enhanced Open WebUI a lot. It now reasons step by step, and with the terminal it reliably locates and extracts the right sections from documents far larger than the context window — list the folder, grep, read only what matters. Then it uses those results to build a document, the way another agent would. Setup: Qwen 27B Q4_K_M on Ollama, 100k context configured. On a ~35k token prompt I measure roughly 1,050 t/s prompt processing and ~46 t/s generation. Prefill speed is the number that matters for this use case — it's what makes chewing through a large document bearable. I was about to give up on Open WebUI. If your use case looks like mine, don't sleep on Open Terminal.
Original Article

Similar Articles

Simpler self hosted alt to Open WebUI

Reddit r/LocalLLaMA

OvertChat is a simpler, polished self-hosted alternative to Open WebUI for local AI models, featuring single Docker compose setup, built-in web search, and Kokoro TTS, all MIT licensed.

Open WebUI Desktop Released!

Reddit r/LocalLLaMA

Open WebUI Desktop launches as a native app letting users run local LLMs or connect to remote servers without Docker or terminal setup, featuring offline operation, system-wide voice input, and floating chat overlay.

OpenComputer | An Open Source Computer Built For Agents.

Reddit r/LocalLLaMA

OpenComputer is an open-source virtual machine environment for AI agents that provides a human-accessible computer interface, allowing agents to safely operate while users can observe and collaborate. It runs locally with small context models and avoids screenshot-based navigation for efficiency.

Open Browser Use

Product Hunt

Open Browser Use is an open-source tool for browser automation designed for local AI agents.