A production-grade, privacy-first desktop AI workstation for local LLMs.
Engineered by Kapil Kumar Yadav
LocalLLMMind is a full-featured desktop AI workstation that interfaces directly with Ollama to run LLMs 100% locally. Zero data leaves your machine. It combines the polish of modern AI products with hardware-optimized streaming, live benchmarking, reasoning model introspection, multimodal vision, RAG document chat, and a complete model management toolkit.
- Paste, drag-and-drop, or browse images directly into the chat composer.
- Automatic Base64 conversion to Ollama's
/api/chatimagesarray. - Click-to-zoom lightbox for high-resolution image inspection.
- Supports
llama3.2-vision,llava,moondream, and all Ollama vision models.
- Tokens are buffered on a 35ms micro-interval (~30fps), reducing React re-renders by ~75%.
- Jitter-free auto-scroll during active generation.
- Live generation metrics badge: speed (
β‘ 54.2 tok/s), token count, duration, and active model.
- Built-in
<think>...</think>tag parser for DeepSeek-R1, Qwen 2.5, and reasoning models. - Interactive collapsible thought accordion with live pulsing indicator during chain-of-thought.
- Drag-and-drop or browse 40+ file formats: source code (
.py,.js,.ts,.rs,.go,.java,.cpp), configs (.json,.yaml,.toml,.env), data (.csv,.sql), and docs (.md,.txt). - Client-side text extraction and vector chunking sandbox β nothing leaves your machine.
- 500KB safety guard to prevent context window overflows.
- View installed models with disk size, parameter count, quantization format, and model family.
- Pull new models with real-time progress bars and cancel support.
- 1-click quick-pull chips for popular models (
llama3.2,deepseek-r1:8b,phi4,gemma3).
- Edit any user message β truncates the conversation and streams a fresh response from that point.
- Regenerate assistant responses with full version history.
- Version navigation β browse between regenerated responses with
β 2/3 βΆarrows.
- 6 built-in system personas: General Assistant, Software Architect, Bug Hunter, Concise & Direct, Academic Researcher, Creative Writer.
- Switch personas per-conversation from the chat header.
- Run two models simultaneously on the same prompt.
- Compare responses side-by-side with live generation speeds (
tok/s) and vote for the better output.
- Built-in prompt template library with CRUD, categories, and import/export.
- Type
/in the composer for instant slash-command autocomplete.
- Full-text search across all conversations with highlighted match snippets.
- Results grouped by conversation with 1-click jump to the exact message.
- After the first exchange, the LLM automatically generates a descriptive 3β5 word title for the conversation.
- Create color-coded project folders to organize conversations.
- Collapsible sidebar tree with chat count badges.
- Move conversations between projects.
- Speak prompts using Web Speech Recognition with live interim feedback.
- Listen to assistant responses with markdown-sanitized TTS output.
- Live header chip showing estimated tokens vs. context limit (e.g.
~1.2k / 4k tok). - Color-coded thresholds: blue β amber (65%+) β red (85%+).
- Choose between browser memory or a local directory on your hard drive.
- Auto-sync conversations to disk as JSON + readable Markdown files (Obsidian-compatible).
- Full JSON backup import/export with data portability.
- Export as PNG image, Markdown, plain text, or PDF (print-optimized A4).
- Native OS Web Share API integration.
- Ground LLM responses with real-time web search context.
- Source citations displayed as clickable chips on messages.
- Interactive sandbox drawer for rendering HTML/CSS/JS artifacts generated by the LLM.
- Full shortcuts manager with conflict detection and custom keybinding creation.
- Bind custom key combos to prompt templates and quick actions.
| Shortcut | Action |
|---|---|
| Cmd+K | New Conversation |
| Cmd+B | Toggle Sidebar |
| Cmd+F | Find in Chat |
| Cmd+Shift+F | Global Search |
| Cmd+Shift+M | Model Manager |
| Cmd+, | Settings |
| Cmd+/ | Shortcuts Manager |
| Escape | Stop Generation / Dismiss |
- Installable as a standalone desktop app via Chrome/Edge.
- Offline caching service worker for app shell and static assets.
- 8 Curated Designer Presets: Modern Dark, Clean Light, Midnight Obsidian, Cyberpunk Neon, Tokyo Night, Forest Emerald, Sunset Glow, and Royal Amethyst.
- Custom Accent Color Swatches: Pick from curated neon/pastel accents or custom hex values with instant live preview.
- Message Bubble Roundness: Granular slider from sharp (4px) to pill (24px) bubble shapes.
- Typography Scale Adjuster: Adjust font scaling (Small, Medium, Large, Extra Large) across all chat and workspace interfaces.
- Persistent Preferences: Theme configuration syncs across restarts via local storage.
- Implicit Live Web Search: The agent queries DuckDuckGo for live facts, current documentation, and news without requiring manual approval clicks, citing sources directly.
- Real-Time Weather Integration: Live weather conditions and forecasts via Open-Meteo.
- Interactive Agent Activity Log: View step-by-step tool traces, execution times, token counts, and export trace JSONs.
- Collapsible Activity Feed: Quick "Hide activity" / "Show activity" toggle with persistent visibility preference.
- Human-in-the-Loop Safe Execution: Destructive actions (modifying files and executing shell commands) require explicit approval.
- First-Class Ollama Integration: 100% private, native local execution.
- Top Cloud Providers: Seamlessly toggle between OpenAI, Anthropic Claude, Google Gemini, and xAI Grok.
- TypeSafe Jev Integration: Native adapter for deterministic and structured LLM decision workflows.
- Custom Endpoints: Connect any OpenAI-compatible API gateway with custom base URLs and headers.
- Deterministic Evaluation Workflows: Model selection, structured state parameters, and decision criteria matrix.
- Strict Output Validation: Ensure schema compliance and structured JSON guarantees.
- Live Test Suite: Execute multi-attribute scoring and benchmark decision matrices before production deployment.
- Curated glassmorphism design with fluid gradients and subtle micro-animations.
- Fully responsive β works on desktop, tablet, and mobile.
- Install Node.js (v18+).
- Install and run Ollama:
ollama serve
- Pull a model:
ollama pull llama3.2 # For vision: ollama pull llama3.2-vision:11b
git clone https://github.com/kapilyadav22/local_llm_ui.git
cd local_llm_ui
npm install
npm run devOpen http://localhost:2210/ in your browser.
npm run build
npm run previewPull and run the pre-built, lightweight Alpine image directly from Docker Hub:
docker run -d \
--name localllmmind \
-p 3000:80 \
--add-host=host.docker.internal:host-gateway \
-e OLLAMA_URL=http://host.docker.internal:11434 \
--restart unless-stopped \
kapilyadav22/localllmmind:latestOpen http://localhost:3000 in your browser.
π¦ Docker Hub Repositories:
kapilyadav22/localllmmind(Official)kapilyadav22/runlocalllm(Mirror)
# Start LocalLLMMind β http://localhost:3000
docker compose up -dAutomatically connects to host Ollama at
http://host.docker.internal:11434.
docker compose --profile with-ollama up -ddocker build -t localllmmind .
docker run -d \
-p 3000:80 \
--add-host=host.docker.internal:host-gateway \
-e OLLAMA_URL=http://host.docker.internal:11434 \
--name localllmmind \
localllmmind| Variable | Default | Description |
|---|---|---|
PORT |
80 |
Nginx listening port inside container |
OLLAMA_URL |
http://host.docker.internal:11434 |
Ollama API endpoint on host gateway |
localllmmind/
βββ src/
β βββ components/
β β βββ Chat/ # ChatView, MessageBubble, MessageInput, WelcomeScreen,
β β β # ArenaMessageBubble, ContextMeter, PersonaDialog,
β β β # PromptLibraryDialog, ShareChatDialog, ArtifactSandboxDrawer
β β βββ Layout/ # AppLayout, Sidebar, GlobalSearchDialog, ProjectDialog
β β βββ Settings/ # SettingsDialog, ModelManagerDialog
β β βββ common/ # MarkdownRenderer, ShortcutsManager, AppLogo, ModelSelector
β βββ constants/ # App branding, models, personas, shortcuts
β βββ hooks/ # useAudio (Speech-to-Text & TTS)
β βββ services/ # ollamaService, webSearchService
β βββ store/ # React Context + useReducer state management
β βββ utils/ # Storage, documentUtils, pdfExport, dialogService
β βββ theme.js # MUI Dark & Light theme tokens
β βββ main.jsx # Entrypoint with PWA service worker registration
βββ public/ # PWA manifest, service worker, icons
βββ Dockerfile # Multi-stage Node + Nginx Alpine build
βββ docker-compose.yml # Production orchestration with Ollama profiles
βββ nginx.conf.template # Streaming reverse proxy for LLM tokens
βββ vite.config.js # Vite config with Ollama API proxy
- Antigravity-Style Agent Chat: Pinned bottom composer for continuous follow-ups, auto-scrolling tool streams, and 1-click file proposal application.
- GitHub Sync & Gists: Link your GitHub account via Personal Access Token. Push commits to existing or new repositories, or export code to Gists.
- Project Import & Folders: Connect local folders directly or import ZIP archives (preserves relative paths, auto-skips
node_modulesand binaries). - Persistent Terminal & Runner: Full interactive terminal sessions with persistent state across reloads, command history, and background processes.
- Code Formatting & Git: Syntax-aware formatting (Prettier & Black), file staging, line review comments, and unified diff inspection.
- Model Manager: Inspect, warm, or unload active models in Ollama memory without deleting downloaded weights.
- Autonomous Agent Mode: Tool-calling agent with implicit web search, document retrieval, and collapsible step-by-step activity traces.
- Code Agent Workspace: Autonomous repository scanning, file diagnosis, refactoring, and code proposal generation.
- Agent Traces & Activity: Real-time duration metrics, token counters, and downloadable JSON execution traces.
- Playground & Evaluations: Custom prompt tuning, parameter presets, and automated fixture evaluations with assertion scoring.
- Interactive Walkthrough Video:
docs/assets/demo_video.mp4(Full desktop walkthrough: Developer Chat, Autonomous Agent Mode, GitHub linking & remote sync, Theme Studio, and Multimodal Vision)
| # | Scenario / Feature | Real Screenshot | Description |
|---|---|---|---|
| 1 | Antigravity IDE Agent Chat | Inspect Preview | Multi-turn developer agent chat with continuous auto-scrolling, code proposals, and sticky bottom prompt composer. |
| 2 | GitHub Account Linking | Inspect Preview | Git Source Control sidebar with GitHub PAT token connection, account badge, and remote repository sync options. |
| 3 | Push Project to GitHub | Inspect Preview | Connected GitHub account profile with 1-click Push to remote repository, branch target, and commit authoring. |
| 4 | Multimodal Vision | Inspect Preview | Upload images with real-time thumbnail preview, size badge, and visual layout/color analysis prompt. |
| 5 | Code Studio Refactor | Inspect Preview | Full project file explorer with Python/JS syntax highlighting and AI Assistant refactoring prompt. |
| 6 | Local RAG & Knowledge Base | Inspect Preview | Local RAG document dialog showing chunking statistics, token counts, and semantic retrieval sandbox. |
| 7 | Extensive Code Review & Comments | Inspect Preview | Multi-tab code workspace with file tree, editor, and dedicated line-targeted review comments panel. |
| 8 | AI Providers Configuration | Inspect Preview | Unified provider settings: OpenAI, Anthropic Claude, Google Gemini, xAI Grok, Jev, Ollama, and Custom endpoints. |
| 9 | TypeSafe Jev Decision Studio | Inspect Preview | Deterministic decision studio modal with model selector, state schema editor, and evaluation criteria matrix. |
| 10 | Model Comparison & Arena | Inspect Preview | Side-by-side responses (Llama 3.2 vs DeepSeek-R1) with live tok/s benchmarks, token counts, and voting chips. |
| 11 | Chat Dashboard | Inspect Preview | Full workstation home showing sidebar, project folders, model selector, context meter, and quick prompts. |
| 12 | Autonomous Agent Mode (Active) | Inspect Preview | Step-by-step tool trace, live implicit DuckDuckGo search execution, and collapsible activity panel. |
| 13 | Autonomous Agent Mode (Clean) | Inspect Preview | Streamlined view with activity collapsed for distraction-free reading of grounded agent answers. |
| 14 | Theme Customizer & Swatches | Inspect Preview | 8 curated theme presets, custom accent color swatches, bubble roundness slider, and font scale tuner. |
| 15 | Evaluations & Benchmark Lab | Inspect Preview | System prompt tuning, fixture test runs, pass/fail assertion metrics, and model latency comparisons. |
Kapil Kumar Yadav β Senior Software Engineer
If you find LocalLLMMind valuable for your workflow, consider sponsoring or supporting this project:
- π Sponsor on GitHub Sponsors
- β Star this repository to help others discover it
- π Report bugs or submit features to help improve the project
Licensed under the Apache License 2.0.
Copyright Β© 2026 Kapil Kumar Yadav. All rights reserved.








