d8cf8dd80a62e6a58c7ec8c290a65bcb4fc44983
Hermes Ollama Models
Native-like Hermes dashboard plugin for local Ollama model management and chat.
Chat capabilities
- Select an installed Ollama model and load it into memory
- Chat through Ollama's native
/api/chatendpoint - Attach screenshots and JPEG/PNG/WebP images for vision-capable models
- Attach text PDFs; PDF text is extracted with
pypdf - Add public HTTP/HTTPS URLs for HTML/text, images, or PDFs
- View live host RAM and swap statistics
- View Ollama's loaded-model memory split: total, GPU VRAM, and normal RAM/offload
- View NVIDIA GPU telemetry when
nvidia-smiis available
The chat UI polls runtime memory once per second and keeps a short in-browser history for the current page session.
Security limits
- Uploaded files are limited to 20 MiB each
- Fetched URLs are limited to 15 MiB and a 30-second timeout
- Private, loopback, link-local, reserved, multicast, and unspecified URL targets are blocked, including redirect destinations
- Remote documents are inserted as untrusted content, not system instructions
- Only locally installed models can be selected or loaded; chat does not download models
Dependency
Install the plugin's Python dependency in the Hermes runtime environment:
pip install -r requirements.txt
pypdf is required for text-based PDF extraction. Scanned/image-only PDFs need OCR and are not converted to text by this plugin.
Languages
Python
97.6%
Shell
2.4%