docs: update .env.example Ollama endpoint/model to match the working config

Root-caused a real bug while fixing this: mint_profile was silently
falling back to the tiny offline template pool on nearly every call,
producing near-identical spirits (same 3-4 persona openings verbatim,
compound-noun names from a 5x5 pool). Traced it to the old Ollama
endpoint being unreachable/overloaded and, separately, a model-tag
mismatch (minicpm-v4.5:latest was requested but only :8b exists on the
working box). Verified 6 candidate models against the real mint task on
the new endpoint — only minicpm-v4.5:8b reliably returns valid JSON;
the others (qwen3.5, ministral-3, granite4.1, lfm2.5, ornith) output
conversational prose instead of the structured schema. Confirmed live
through the actual guest summon path post-restart: genuinely distinct,
well-written personas now, not fallback template text.

.env itself is gitignored and was updated locally; this just keeps the
example honest.

Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
This commit is contained in:
Indiana
2026-07-29 19:07:58 +00:00
parent 94c283634f
commit b68e5d7fe5

View File

@@ -1,8 +1,8 @@
DATABASE_URL=postgresql+asyncpg://quantumancy:quantumancy@localhost:5432/quantumancy DATABASE_URL=postgresql+asyncpg://quantumancy:quantumancy@localhost:5432/quantumancy
OLLAMA_BASE_URL=http://10.30.20.107:11434 OLLAMA_BASE_URL=http://10.30.20.186:11434
PORT=7777 PORT=7777
# LLM tiers on the Ollama box (fast = fragments/ambient, chat = direct contact/minting) # LLM tiers on the Ollama box (fast = fragments/ambient, chat = direct contact/minting)
OLLAMA_FAST_MODEL=granite4.1:3b OLLAMA_FAST_MODEL=granite4.1:3b
OLLAMA_CHAT_MODEL=minicpm-v4.5:latest OLLAMA_CHAT_MODEL=minicpm-v4.5:8b
LLM_MAX_CONCURRENCY=2 LLM_MAX_CONCURRENCY=2
LLM_MAX_QUEUE_DEPTH=8 LLM_MAX_QUEUE_DEPTH=8