File size: 1,406 Bytes
ca3d977
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
2
3
4
5
6
7
8
9
10
11
12
13
14
15
16
17
18
19
20
21
22
23
24
25
26
27
28
29
30
31
32
33
34
# Runtime instructions

Planner Cache does not distribute base-model weights or llama.cpp. Create the
Python environment, provide local model paths, and use the included launchers.
The exact Gemma tokenizer and metadata bundle is also an external upstream
requirement. Its contents are not redistributed in these publication packs.

```bash
python -m venv .venv
.venv/bin/pip install -e '.[dev,publishing]'

export LLAMA_CPP_DIR=/path/to/llama.cpp
export GEMMA_MODEL=/path/to/compatible-gemma.gguf
export GEMMA_TOKENIZER_BUNDLE=/path/to/matching-gemma-tokenizer-bundle
./run-gemma.sh

export PYTHIA_MODEL=/path/to/pythia-1.4b
export REVIEW_MODEL="$GEMMA_MODEL"
./run-pythia.sh
```

The launchers resolve the project root from their own location. Optional paths
include `PYTHIA_TTL`, `GEMMA_LTL`, `GEMMA_TOKENIZER_BUNDLE`, `ROUTER_PATH`,
`PPKG_PATH`, `SESSION_ROOT`, `LLAMA_WEB_UI`, `WEB_HOST`, and `WEB_PORT`.

Gemma uses llama.cpp and the active `.ltl`. Pythia uses the semantic `.ttl` and
uses the configured frozen GGUF as a CPU structured reviewer by default. The
review request is separate from visible generation and uses neither TTL nor LTL.

The browser is the primary conversation interface. Terminal commands `/state`,
`/personality`, `/events`, `/save`, and `/quit` are secondary diagnostics.
Every session records transcript, events, metadata, P-cache state, and final
state under `sessions/`.