Skip to content

Commit 3b718f9

Browse files
committed
docs(compaction): correct ollama ttl identity vs table key
The table key is bare ollama; the harness stamps slash-form on sourceId. Stamp a production-looking Codex model on the governor fixture instead of the helper default.
1 parent f0057e4 commit 3b718f9

2 files changed

Lines changed: 6 additions & 2 deletions

File tree

‎docs/ARCHITECTURE.md‎

Lines changed: 1 addition & 1 deletion
Original file line numberDiff line numberDiff line change
@@ -202,7 +202,7 @@ When a cycle's input tokens cross a threshold, the director compacts the inferen
202202
| `ollama` | never | Local inference has no remote cache to expire |
203203
| anything else | 10 min | Assumed OpenAI-style in-memory economics; tune per upstream |
204204

205-
Production identity for the `ollama` row is `sourceId` `ollama` / `ollama/<instance>` via `isOllamaProviderId` (`src/provider/ollama.ts`), not a bare model id: Ollama is `buildOpenAISource` (`provider: openai-compatible`, model `llama3` / `qwen3`). Slash-form `ollama/…` is a table key, not what the harness stamps.
205+
The `ollama` table key is the bare provider segment; the harness stamps slash-form on `sourceId` (`ollama/default` / `ollama/<instance>`), with `provider: openai-compatible` (`buildOpenAISource`) and a bare model (`llama3` / `qwen3`). Idle-recompress disable matches via `isOllamaProviderId` on that `sourceId` (`src/provider/ollama.ts`), not by looking up slash-form as a table key.
206206

207207
- **Overflow recovery** — A `context_overflow` inference error would otherwise become a terminal error reply; the governor compacts and retries instead, bounded so a history the compactor cannot shrink does not loop forever. Overflow ignores hysteresis for the compact itself.
208208

‎src/agent/compaction.test.ts‎

Lines changed: 5 additions & 1 deletion
Original file line numberDiff line numberDiff line change
@@ -891,7 +891,11 @@ describe("provider-aware idle recompress (CL-8745)", () => {
891891
);
892892
codex.noteInferenceDone(
893893
ttlInferenceDone(
894-
{ sourceId: "codex/work", provider: "codex-responses" },
894+
{
895+
sourceId: "codex/work",
896+
provider: "codex-responses",
897+
model: "gpt-5.6-luna",
898+
},
895899
false,
896900
),
897901
tenTurns,

0 commit comments

Comments
 (0)