Skip to content

Agent runtimes

A runtime is the agent program fullsend runs inside the sandbox — the thing that talks to the model and executes tool calls. fullsend run delegates to it and owns everything around it: the sandbox, the credentials, and the verdict.

RuntimeUse it forStatus
claudeProduction agent runs (Claude Code)Default
piSecond runtime, opt-in per repo — Claude, Grok and Gemini on Vertex; GPT via OpenAI WIF (wired, not yet exercised live)Supported for triage, prioritize, code, fix
dummyBehaviour tests — scripted ops, no inferenceInternal
dummy-playbackBehaviour tests — replays canned agent results from a playlist, no inferenceInternal
opencodeNot yet functionalStub

Pick one with runtime: in .fullsend/config.yaml, or per run with --runtime.

bash
fullsend run triage --runtime pi --model xai-vertex/xai/grok-4.6

How a run uses the runtime

The runner owns the sandbox, credentials and verdict; the runtime owns what happens between "start" and "event stream".

Loading diagram...

Choosing between claude and pi

Claude Codepi
ModelsAnthropic on VertexClaude, Grok and Gemini on Vertex; GPT via OpenAI WIF (opt-in, not yet exercised live)
Sub-agentsNative (Agent tool)Not wired — agents execute sub-agent definitions inline (#6527)
Fallback model chainFULLSEND_FALLBACK_MODELS, tried in orderIgnored with a warning
RolesAllreview/retro stay on Claude Code — they rely on sub-agent rosters
Effort--effort low..max--thinking, same levels (high when unset)
Security controlsFull matrixFull matrix; stricter on failed-call sanitizing

Both run unattended in the same sandbox, on the same WIF credentials, behind the same egress allowlist. Choose pi when you want a non-Anthropic model; stay on claude when you need sub-agents or a fallback chain.

Selecting a runtime and model

First non-empty wins — the usual flag > env var > config > default. fullsend run resolves this once, validates it, prints the source, and records it in metrics.json; runtimes never read the override variables themselves.

Loading diagram...
SettingFlagEnvConfig (per-agent)Config (repo-wide)
Runtime--runtimeFULLSEND_RUNTIMEruntime: on the agent's agents: entryruntime: in .fullsend/config.yaml (repo default)
Model--modelFULLSEND_MODEL (FULLSEND_PI_MODEL is a lower-precedence alias on pi)model: on the agent's agents: entryharness model:, then agent frontmatter model:
Effort--effortFULLSEND_EFFORTeffort: on the agent's agents: entryharness effort:

In CI these are repository variables of the same name, plain or role-prefixed (TRIAGE_FULLSEND_MODEL), so a repo can switch one role's model without a pull request. For durable per-agent configuration that lives in the repository and is reviewable, use the agent's agents: entry in .fullsend/config.yaml instead. Harness env.runner does not reach the fullsend process.

Per-agent runtime, model and effort

The agents: list is the per-agent place in config.yaml: an entry names an agent and can set its runtime, model and effort. A built-in agent (triage, code, review, fix, retro, prioritize) is tuned with a name-only entry; a custom agent carries the settings on its source: entry.

yaml
runtime: pi                    # repo default for agents that set none
agents:
  - name: triage
    model: xai-vertex/xai/grok-4.6
  - name: code
    runtime: claude
    model: sonnet
    effort: high
  - source: https://raw.githubusercontent.com/acme/agents/<sha>/harness/lint.yaml#sha256=…
    model: haiku

Or from the CLI, which validates the entry before writing it:

  1. fullsend agent set code --fullsend-dir .fullsend --runtime claude --model sonnet --effort high
  2. fullsend agent list --fullsend-dir .fullsend shows the settings next to each agent — code (built-in) [runtime=claude model=sonnet effort=high], or the source: path for a custom agent.
  3. The next fullsend run code names the entry as the source — Runtime: claude (from <config path> agents.code) — and a --runtime/--model flag on that run still wins.

An invalid value is refused before the write — invalid effort "turbo": must be one of low, medium, high, xhigh, max — and the same check runs on every fullsend run, so a hand-edited entry fails the run before a sandbox starts rather than being skipped.

A source: entry needs no name: — the agent's name is derived from the source file (harness/lint.yamllint, ADR 0058), and that is the name the settings, fullsend run lint and fullsend agent set lint all use; add name: only to override it.

Names are agent names as passed to fullsend run <agent>not harness role: values (code and fix both carry role: coder) — matched case-insensitively. A name-only entry for anything that is not a built-in agent fails validation (coder gets a "did you mean code" hint); a custom agent gets its settings on its own entry.

Precedence: flag > env var > the agent's agents: entry > repo-wide runtime: / harness model: effort: > default. Entries merge per field across the layered config (config.yaml over config.base.yaml), so a preset base can tune agents too. fullsend run validates the whole agents: list in every layer (names, runtime, model syntax, effort) and fails the run with an error naming the file and entry rather than silently skipping a mistyped entry or handing a bad value to the runtime.

A value that came from here shows up as <config path> agents.<name> wherever the selection is surfaced (plan block, stderr, metrics.json — see below); the path is the effective config file.

provider/id is pi's model form. The syntax is accepted for every runtime (model ids are not a closed set), but an entry that pairs runtime: claude with a provider/id model gets a warning in the plan block — Claude Code expects an alias (opus, sonnet, …) or an Anthropic model id.

Migrating from repository variables. A repo that carries <ROLE>_FULLSEND_MODEL / <ROLE>_FULLSEND_RUNTIME variables can move them onto agents: entries one-to-one: the variable prefix is the agent name (CODE_FULLSEND_RUNTIME=claude- name: code / runtime: claude). Delete the variable afterwards — while it exists it still wins, so the config entry would be silently shadowed. Bump the workflow's fullsend pin to a version that carries per-agent settings before adding them: an older pinned CLI rejects an enabled agents: entry without a source, whereas a current CLI validates the settings on every run.

Set the runtime per repo with fullsend github setup <owner/repo> --runtime pi. Repos on pi need a sandbox image that carries PI_VERSION.

Models

On Claude Code, pass an alias (opus, sonnet, haiku, fable) or a model id.

On pi, a model is provider/id — aliases and bare ids still work, and the provider comes from FULLSEND_PI_PROVIDER (default anthropic-vertex). pi reaches Claude, Gemini and Grok, each through its own provider; see Pi › Models and providers.

Harness model: and agents: entry model: values accept provider-qualified provider/id syntax (e.g. google-vertex/gemini-3.7-flash). On pi, a harness can also select a provider with a bare model: plus FULLSEND_PI_PROVIDER.

Where the selection appears

SurfaceWhat it shows
Run plan blockRuntime: <name> (from <source>) next to Model and Effort; <source> is the flag, the variable, or <config path> (suffixed agents.<name> when the agent's entry decided)
stderrruntime: selected "<name>" from <source>
Status comment / ::notice::Runtime · Model: <requested → reported> · Effort · Cost
OTel spanfullsend.runtime, next to gen_ai.request.model
metrics.jsonruntime, requested_runtime, runtime_source, requested_model, override_source

requested_model is what was handed to the runtime after overrides, and override_source says where it came from — so a silent override is visible after the fact. The reported model is the provider-stripped id (claude-opus-4-6); for a provider whose ids are publisher-qualified it keeps that segment (xai/grok-4.6), since that is the wire id.

Harness config keys per runtime

Harness keys are runtime-neutral in YAML; each runtime owns the translation. Test-only runtimes (dummy, dummy-playback) ignore all harness config keys and are omitted from this table.

Harness keyClaude Codepi
model--modelalias table, then provider/id; see Models
effort--effort--thinking (superset of the harness levels; high when unset)
tools:Native Claude permission syntax--tools (strict) + a first-token Bash allowlist
skillsCLAUDE_CONFIG_DIR/skills/PI_CODING_AGENT_DIR/skills/, discovered natively
pluginsMarketplace layoutUnsupported — warned and skipped
security.sandbox_hookshooks.json via --settingsHook scripts + manifest + adapter extension
validation_loop.feedback_modeReplaces the prompt on retrySame

Full per-key detail, including the exact --tools mapping and allowlist parsing rules, is in Implementing an agent runtime.