Orkestia
Blog
Staff & Agents

Troubleshooting

Sessions that never heartbeat, actors that will not act, runner gates, seats, tokens, and where to look next

Most first-run failures are prerequisites, not model quality. Work this list top to bottom.

Actor seems inert

  1. No skills. Attach a workflow-backed skill that grants the tools you are asking for. Zero skills ⇒ reason-only.
  2. Built-in MCP missing. If it should start Orkestia workflows, attach the built-in workflow MCP on the config.
  3. RBAC. The actor's role does not include that workflow. Check Roles and the unit the actor sits in.
  4. Paused or archived. Resume or hire a new actor.

Session never heartbeats / hangs in launch

  1. Runner group purpose is not agent. Recreate or pick an agent group. Generic CI groups fail fast or (older configs) hang. Message text usually includes "not agent-eligible".
  2. Group not active. Wait for provision; watch Runner management.
  3. Wrong image. A GitHub-Actions runner image will register and exit; the watch loop waits forever. Use the agent runtime.
  4. No group on the config. Edit the config; hire again if needed.

Hire cannot see a model

  1. Connection type not in the model provider list.
  2. Connection failed validation in app.orkestia.dev/connections.
  3. Refresh Staff after the connection succeeds so model profiles reload.

agt_ / MCP calls fail

  1. Token revoked, expired, or actor paused.
  2. Using a member JWT in an automation client — mint agt_ instead.
  3. Passing another org's UUID in initial_data.
  4. Paid RBAC seat exhausted for RBAC-scoped actors — RBAC seats + Stripe.

whoami on MCP is the first diagnostic: org and actor kind must match the actor you minted.

Budget / cost surprises

  1. Session stopped on budget-check — raise the config ceiling or pause the actor.
  2. Cost page vs invoice: the Cost page shows agent session spend from the price catalog. The Orkestia invoice is the subscription (seats and add-ons) plus the execution and request meters above the included volume.
  3. Org pricing overrides on Pricing.

Approvals stuck in Inbox

  1. No human has the approve capability on that unit.
  2. Workflow is platform-locked; role grants will not move it.
  3. Open the run history — if it never reached awaiting-approval, it failed RBAC or schema first.

Where to read the evidence

SurfaceWhat you get
Staff Sessions / session detailTimeline, tool calls, delivery context
Staff AuditOrg-scoped event log
LumenLogs, error groups, traces if provisioned
get_workflow_history / MCPExact engine transitions
The inbox first-run checklist hiding means you have a model connection, an agent-ish runner group, and at least one actor — it does not mean the runner is healthy. If invoke still hangs, ignore the green checklist and check group purpose and status.

Still stuck