docs: establish Ollama as the documentation engine
Route all prose generation (session logs, commit messages, Syncro comments, client notes, code docs) through Ollama qwen3:14b by default. Claude reviews output and owns verbatim-accuracy sections (credentials, IPs, command outputs). GrepAI context lookups keep the Ollama service warm, eliminating the 30-50s cold-start in normal workflow. Updates: OLLAMA.md (documentation engine scope + warm-start note), CLAUDE.md (Ollama section), save.md (narrative drafting), checkpoint.md (commit message body drafting). Co-Authored-By: Claude Sonnet 4.6 <noreply@anthropic.com>
This commit is contained in:
@@ -241,11 +241,13 @@ Vault structure: `infrastructure/`, `clients/`, `services/`, `projects/`, `msp-t
|
||||
|
||||
## Local AI (Ollama)
|
||||
|
||||
Tier 0 — use before spawning Haiku agents for low-stakes tasks (free, fast, private):
|
||||
Tier 0 — **Ollama is the documentation engine.** Route all prose generation through it: session log narratives, commit messages, ticket comments, client notes, code docs. Claude reviews output, owns credentials/facts/execution.
|
||||
|
||||
- **DESKTOP-0O8A1RL:** `http://localhost:11434`
|
||||
- **Other machines:** `http://100.92.127.64:11434` (Tailscale required)
|
||||
- **Models:** `qwen3:14b` (summarize/classify/draft), `codestral:22b` (code suggestions — always review)
|
||||
- **Full reference:** `.claude/OLLAMA.md` (connection examples, model selection, review policy)
|
||||
- **Models:** `qwen3:14b` (all documentation/prose), `codestral:22b` (code suggestions — always review)
|
||||
- **Warm-start:** GrepAI keeps the Ollama service running; qwen3 VRAM swap is ~5s worst case, not 50s
|
||||
- **Full reference:** `.claude/OLLAMA.md` (documentation engine scope, model selection, review policy)
|
||||
|
||||
### GrepAI (Semantic Code Search)
|
||||
|
||||
|
||||
Reference in New Issue
Block a user