You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
{{ message }}
Repository navigation
Commit 49573ae
Browse filesBrowse the repository at this point in the historyBrowse files
merge: resolve conflicts with main (local LLMs + PID safety)
Conflicts resolved:
- .gitignore: combined both sets of exclusions
- cli-reference.md: keep both --lan/--password and --agent ollama/lmstudio
- cleanup.py: keep main's PID-recycling-safe kill (ps check before SIGTERM)
Co-Authored-By: Claude Opus 4.6 <noreply@anthropic.com>
-**RSD/Imposter syndrome**: Reframe mistakes as exploration, bridge to infrastructure experience
81
-
-**Overload prevention**: Max 3-4 concepts, tables over prose, TL;DR at top, mermaid diagrams
82
-
-**Hyperfocus**: Time warnings, exit points, hydration/food reminders
83
-
-**Emotional regulation**: Micro-celebrations for genuine progress, sensory checks at 45+ min
84
-
-**Transition support**: Summarise when switching, parking lot for tangents
78
+
See `agents/shared/audhd-framework.md` for the complete methodology. Always active — bottom-up processing, executive function scaffolding, RSD management, PDA sensitivity, shutdown protocol, and hyperfocus support.
-**RSD/Imposter syndrome**: Reframe mistakes as exploration, bridge to infrastructure experience
85
-
-**Overload prevention**: Max 3-4 concepts, tables over prose, TL;DR at top, mermaid diagrams
86
-
-**Hyperfocus**: Time warnings, exit points, hydration/food reminders
87
-
-**Emotional regulation**: Micro-celebrations for genuine progress, sensory checks at 45+ min
88
-
-**Transition support**: Summarise when switching, parking lot for tangents
82
+
See `agents/shared/audhd-framework.md` for the complete methodology. Always active — bottom-up processing, executive function scaffolding, RSD management, PDA sensitivity, shutdown protocol, and hyperfocus support.
# AGENTS.md is loaded automatically — just start asking for a study session
277
278
```
278
279
280
+
## Local LLMs (Ollama / LM Studio)
281
+
282
+
studyctl can use local LLMs as the study mentor backend instead of cloud Claude. This uses Claude Code as the frontend but points it at a local model server via environment variables.
283
+
284
+
### Honest expectations
285
+
286
+
Local models are a **cost/privacy trade-off with significant capability regression**:
-**Rough quality**: Best local models (Qwen3-Coder 30B, Devstral 24B) are approximately Claude Haiku 3.5 quality for agentic tasks
291
+
292
+
If you need reliable multi-step study sessions, cloud Claude is substantially better. Local LLMs are best for privacy-sensitive work, offline use, or cost-free experimentation.
293
+
294
+
### API compatibility
295
+
296
+
Claude Code requires the **Anthropic Messages API format** (`/v1/messages`). Not all local backends support this:
297
+
298
+
| Backend | Anthropic API? | Notes |
299
+
|---------|---------------|-------|
300
+
|**LM Studio 0.4.1+**| Native | Simplest path. Just load a model and point studyctl at it. |
301
+
|**llama.cpp server**| Native (since Nov 2025) | Low-level, good for headless servers. |
302
+
|**LiteLLM proxy**| Translates | Bridges Ollama's OpenAI API to Anthropic format. |
303
+
|**Ollama (direct)**| No | Only speaks OpenAI format. Needs LiteLLM as a proxy. |
304
+
305
+
### Recommended models
306
+
307
+
Models ranked by suitability for studyctl's agentic, multi-turn workflow:
308
+
309
+
| Model | VRAM/RAM | Context | Best for |
310
+
|-------|----------|---------|----------|
311
+
|**Qwen3-Coder 30B**|~19 GB | 256K | Best open-source for coding. Explicit tool-use training. |
312
+
|**Devstral 24B**|~14 GB | 128K | Top SWE-bench open-source. Runs on 32 GB Mac. Apache 2.0. |
313
+
|**DeepSeek-Coder-V2 16B**|~9 GB | 160K | Good at the 16B weight class. |
314
+
315
+
**Minimum context window**: 64K tokens. Claude Code's system prompt, CLAUDE.md, tool definitions, and file reads consume 20K-50K tokens before your conversation even starts. Models with <32K context will truncate constantly.
- **Malformed tool calls**: Local models emit invalid tool-use JSON more often than cloud Claude. Claude Code may crash with `Cannot read properties of undefined`. Workaround: `export CLAUDE_CODE_USE_POWERSHELL_TOOL=0`
400
+
- **No prompt caching**: Every turn processes the full context from scratch. Sessions feel slower as context grows.
401
+
- **No extended thinking**: The effort slider and thinking modes are Claude-specific features.
402
+
- **Background tasks use local model**: Claude Code routes statusline updates and codebase searches through the "haiku" model tier. With tier-pinning (which studyctl sets automatically), all of these hit your local GPU.
403
+
404
+
### Verifying your setup
405
+
406
+
```bash
407
+
studyctl doctor
408
+
```
409
+
410
+
The doctor checks will report:
411
+
- Whether ollama/lms binaries are installed
412
+
- Whether the local server is responding
413
+
- Whether Claude Code is installed (required as the frontend)
0 commit comments