VRAM Orchestrator
Monitor GPU memory in real time. The system unloads inactive text models before heavy image, video, or 3D generation tasks.
Control Ollama, Stable Diffusion Forge, ComfyUI, and Kokoro TTS with automated VRAM orchestration.
Local LLM Server Manager is a cross-platform orchestrator for local artificial intelligence engines. The application operates on Windows and Linux. The system coordinates Large Language Models, image generation, 3D mesh reconstruction, video synthesis, and speech generation.

| Layout Section | Primary Function | Active Elements |
|---|---|---|
| Telemetry Header | Live Hardware Monitoring | GPU name, total/used VRAM bar, service health indicator, and refresh button. |
| Primary Navigation | Workspace Switcher | Tabs for My Models, Hugging Face Hub, CivitAI, Multimodal Studio, AI Assistant, and Settings. |
| Main Content Canvas | Engine Interaction | Model cards, KV cache context calculator, download progress, WebGL canvas, and video player. |
| Companion Windows | Floating Multi-Window | Detachable Documentation and AI Assist windows with magnetic flank docking. |
5246).Explore the documentation guides to set up and operate your local AI stack: