原始内容
ytdb-slate
Slate is a thread-weaving orchestration extension for the pi coding agent.
The orchestrator (your main pi session) dispatches bounded actions to persistent worker threads. Each completed action is compressed by an LLM into an episode — a durable, structured record (intent, actions, findings, artifacts, open issues, handoff notes) that the orchestrator composes into further dispatches instead of re-reading raw transcripts. On top of that, Slate injects a mandatory workflow doctrine: research before design, design review, adversarial review, and track review — with optional umbrella draft-PR publishing for tracks. An opt-in model-failover map adds high availability: when a model API fails, the orchestrator, worker threads, and episode compression each retry once on a configured equal-quality alternative.
Why Slate?
Long-horizon agentic work fails on context management, not model capability. Existing architectures each solve a piece of the problem and trade away the rest. Compaction is unpredictably lossy; naive subagents isolate context but hand back only a single response string; markdown plans go stale and get under-executed. Rigid task trees can't adapt to information discovered mid-task, and planner/executor stacks synchronize through compress-and-return boundaries that risk dropping critical state.
Slate's answer is the thread: the orchestrator dispatches one bounded action at a time; the worker executes it and returns an episode. The orchestrator keeps the reactivity of a plain agent loop while gaining the context isolation, compaction, and parallelism that the single-context loop lacks.
| Aspect | ReAct | Markdown plan | Task trees | RLM (Recursive Language Models) | Devin / Manus / Altera | Claude Code / Codex subagents | Slate |
|---|---|---|---|---|---|---|---|
| Planning | implicit | file | explicit tree | REPL | planning agent | plan mode | implicit / adaptive — no upfront static plan |
| Decomposition | none | none | direct tree | REPL functions | task-based | subagent delegation | implicit |
| Synchronization | single thread | single thread | gated steps | REPL return | reduce & return | message passing | episodes |
| Intermediate feedback | per step | per step | on task failure | on execution end | after compress | message passing | per episode |
| Context isolation | none | none | per subtask | per subcall | subagent | subagent | per thread |
| Context compaction | none | none | task-based | REPL slicing | subagent compress | compaction | episode compress |
| Parallel execution | none | none | none | in REPL | Altera only | native | native |
| Expressivity | high | high | low | high | medium | medium | high |
| Adaptability | yes | if plan updated | no | limited — no mid-run course correction | yes | limited by message passing | yes |
Characterizations of third-party systems reflect their publicly described designs at the time of writing. The taxonomy is adapted from the Random Labs technical report introducing the thread-weaving "Slate" architecture (its agent ships as the npm package @randomlabs/slate); ytdb-slate is an independent implementation of that architecture for pi.
Threads are not subagents:
- Per-episode feedback. Each thread executes one bounded action and hands control back. The orchestrator adapts after every episode — reactive like a ReAct loop — instead of firing off a subagent and hoping the result comes back usable.
- Compaction at a chosen boundary. Compression happens at a predictable point — action completion — instead of mid-stream when context overflows. The compression itself is still LLM-performed and lossy, but the boundary is chosen, not forced.
- Episodes compose. One thread's episode can seed another thread's context, so conclusions travel across the system by reference. Subagents pass back a single response string; episodes are durable, structured, reusable records.
Full rationale: docs/design-principles.md, shipped in the package — if this summary and that document disagree, the document wins.
Feature development workflow
In orchestrator mode, Slate injects a track-based development workflow as doctrine — it is not optional; project configuration can extend it (doctrineExtraPath) but never replace it. The flow:
- Research — interactive exploration; a research log opens lazily when defined triggers fire (non-trivial decisions, surprises, risky invariants, session boundaries).
- User design review — the user approves the design direction before any machine review.
- Adversarial review — a fresh-context reviewer attacks the design pre-implementation; findings are triaged with the user.
- Track split approval — the user approves how the change splits into independently reviewable tracks.
- Per-track loop — implement → agent code review → fixes → mandatory user review → marker commit (an empty commit marking the track's boundary in history).
- Delivery — with draft-PR publishing enabled, the squash-merge is performed by the user; the agent never merges. With it disabled (the default), the change lands as a squashed commit that carries the workflow log's key content in its message.
The workflow scales with change size: multi-track changes get the full flow, single-track changes skip the split and the marker commits, and trivial changes shrink the pre-implementation reviews to consent plus a micro-review (skippable with explicit user consent) — the agent code review and the mandatory user review still apply at every tier. Umbrella draft-PR publishing activates only when workflow.draftPRs is true.
This summary is orientation only; the shipped docs listed below are normative.
Model failover
Slate can ride through model API outages. The opt-in modelFailover map in .pi/slate.json (trusted projects only) maps a model (provider/id) to an equal-quality alternative; on an eligible model API failure — not an abort or a context overflow, and for the orchestrator and worker sites only after pi's own retries are exhausted — the affected site (orchestrator, worker thread, or episode compression) retries once on the mapped model (single hop, never chained). An orchestrator failover also persists the mapped model as pi's global default — switch back with /model after the outage. The map is empty by default (feature off) and read once at session start. Full semantics: the modelFailover row in Configuration and the shipped docs/model-failover.md — if they disagree, that document wins.
Install
Install into a project (pinned, project-scoped — recorded in .pi/settings.json):
pi install -l npm:ytdb-slate@<version>
Note on pinning: pinned specs (
@<version>) are deliberately skipped bypi update --extensions/pi update --all. Bumping the pin is a conscious project change — review Slate's shipped workflow docs for changes when you do.
Configuration
Optional config file: slate.json in the project's pi config dir (.pi/slate.json). It is honored only in trusted projects (see Trust).
| Key | Type | Default | Semantics |
|---|---|---|---|
orchestratorModeDefault |
boolean | false |
Start fresh interactive sessions with orchestrator mode ON. |
episodeModel |
string | newest available Anthropic Sonnet, else the worker's own model | Model (provider/id) used to compress a finished action into an episode. |
workerTools |
string[] | ["read", "bash", "edit", "write", "grep", "find", "ls"] |
Tools available to worker threads (an empty list also falls back to the default). |
maxConcurrent |
number | 4 |
Maximum number of worker actions running concurrently (must be ≥ 1 — unenforced: a value of 0 or less silently hangs all dispatches). Excess dispatches wait in a queue; actions on the same thread always run in dispatch order. Default rationale: shipped docs/design-principles.md §5 (repo-local note). |
contextBudget |
number | object | 256000 (Anthropic models: 400000) |
Absolute orchestrator context budget (tokens) at which Slate auto-pauses and prepares a fresh-session handoff — semantics, defaults, per-model overrides, and rationale in docs/context-budget.md. |
orchestratorPromptDocs |
string[] | [] |
Project markdown files (paths relative to the project root) whose contents are appended to the orchestrator system prompt. |
workerPromptDocs |
string[] | [] |
Project markdown files whose contents are appended to every worker-thread system prompt. |
workflow.draftPRs |
boolean | false |
Enable umbrella draft-PR publishing for tracks. |
doctrineExtraPath |
string | — | Project markdown whose content is appended to the orchestrator doctrine (project-specific workflow additions). |
reviewPerspectivesPath |
string | — | Project review charters, each declaring its own finding-ID prefix. The doctrine references this path; the orchestrator reads the file alongside the shipped review rules. |
modelFailover |
object (string → string) | — (empty, failover off) | Map of provider/id → equal-quality alternative model; on a model API failure the affected site retries once on the mapped model (worker/orchestrator sites only after pi's own retries are exhausted; episode compression has no pi retry loop). Read once at session start — see docs/model-failover.md. |
Example .pi/slate.json (the docs/agents/... paths are placeholders — point them at markdown files that actually exist in your project):
{
"orchestratorModeDefault": true,
"episodeModel": "anthropic/claude-sonnet-5",
"maxConcurrent": 4,
"orchestratorPromptDocs": ["docs/agents/orchestrator-guidelines.md"],
"workerPromptDocs": ["docs/agents/thread-guidelines.md"],
"workflow": { "draftPRs": true },
"doctrineExtraPath": "docs/agents/workflow-additions.md",
"reviewPerspectivesPath": "docs/agents/review-perspectives.md",
"modelFailover": { "anthropic/claude-sonnet-5": "openai/gpt-5.2" }
}
Silent skip: the project-file keys fail silently — no error is shown. For the content-injected keys (
orchestratorPromptDocs,workerPromptDocs,doctrineExtraPath) a missing, unreadable, or empty file is skipped and nothing is injected. ForreviewPerspectivesPaththe pointer is omitted only when the file is missing — the file is not read at injection time, so an unreadable or empty file is still cited. Verify your paths after copying the example.
Trust
Slate reads project configuration (.pi/slate.json) and injects project files (orchestratorPromptDocs, workerPromptDocs, doctrineExtraPath, reviewPerspectivesPath) only in trusted projects. In untrusted projects Slate runs with built-in defaults and injects nothing from the working tree.
Shipped docs
In orchestrator mode, Slate appends a short doctrine (a block of numbered rules) to the orchestrator's system prompt each turn. The doctrine does not embed the workflow docs — it cites them by absolute path, resolved inside the installed package (not your project), and the orchestrator reads them on demand:
docs/track-workflow.md— the track-based workflow (research → design review → adversarial review → track review)docs/pr-publishing.md— umbrella draft-PR publishing (cited only whenworkflow.draftPRsistrue)docs/review-rules.md— review discipline and finding rulesdocs/design-principles.md— Slate's own design rationaledocs/model-failover.md— the opt-inmodelFailovermap (reference documentation — unlike the entries above it is not workflow doctrine and is not cited by the doctrine)docs/context-budget.md— the orchestratorcontextBudget: defaults, per-model overrides, the window clamp, and the pricing rationale (also reference documentation, not cited by the doctrine)
Project-specific additions layer on top — they extend, not replace, the shipped doctrine — via two distinct mechanisms:
- Content injection:
doctrineExtraPath(appended to the doctrine itself, re-read at each prompt assembly) andorchestratorPromptDocs/workerPromptDocs(appended to the respective system prompts). - Pointer:
reviewPerspectivesPathis cited by path from the doctrine's review rule and read on demand, like the shipped docs.
License
Apache-2.0 — see LICENSE.