mirror of
https://github.com/rennf93/roboco.git
synced 2026-08-03 07:23:24 +02:00
* feat(goals): company charter singleton — data layer (Business Goals slice 1)
First slice of the company-in-a-box "Business Goals" phase: a single CEO-owned
charter row (north star + objectives + constraints + operating policy) that
will be injected into every agent's context_briefing so all work is goal-aware.
- CompanyGoalsTable: singleton table (all-zeros id), JSON objectives /
constraints / operating_policy, updated_at / updated_by.
- migration 032: create + seed the singleton row (offline-renderable; column
server-defaults fill an INSERT of just the id).
- CompanyGoalsService: get() (empty defaults when unset) + upsert() (singleton,
partial update, caller commits).
- tests: empty defaults, roundtrip, singleton + partial-update preservation.
Next slices (mapped, not yet built): briefing injection (BriefingInputs +
build_context_briefing + EvidenceRepo), API route (GET any / PUT CEO-only),
panel /goals page, and base/Board/PM prompt mentions.
* feat(goals): inject the company charter into every agent briefing (slice 2)
The charter is now goal-aware context for every agent:
- BriefingInputs gains company_goals; build_context_briefing surfaces it.
- EvidenceRepo.company_goals(): single-row lookup returning a COMPACT charter
(north star + objectives + constraints + operating policy; audit columns
dropped, lists capped) or None when unset, so an empty charter never bloats
the per-verb briefing.
- _briefing_for wires it into every context_briefing.
Tests: briefing surfaces company_goals (defaults None); repo returns None for an
absent/empty charter and the compact dict when set.
* feat(goals): company charter API — GET any agent, PUT CEO-only (slice 3)
- routes/company_goals.py: GET returns the charter (any authenticated agent —
it drives every briefing); PUT is CEO-only (403 otherwise), partial update via
model_dump(exclude_unset=True), explicit commit.
- schemas/company_goals.py: response + partial-update models.
- registered at /api/company-goals.
- tests: GET open to any role, CEO update persists + is readable, non-CEO 403.
* feat(goals): make the company charter actionable in agent prompts (slice 5)
Agents already receive company_goals in the briefing (slice 2); now tell them to
act on it:
- base.md: universal "Align with the company charter" section — favour work and
trade-offs that advance the objectives, honour the constraints, flag conflicts;
never a license to leave your role.
- board / main_pm / cell_pm: role-specific lines tying triage / cell-routing /
subtask decomposition to the charter.
Prompts are composed at spawn from base.md + roles/*.md directly (compose_prompt),
so no _generated regeneration is needed.
* feat(goals): company charter panel page (slice 4)
CEO-facing editor for the charter at /company-goals:
- lib/api/company-goals.ts: get / update (PUT) client.
- company-goals-card.tsx: edit north star + constraints (one per line) +
objectives / operating_policy (JSON, parsed + validated with toast errors);
display derives from server state (no set-state-in-effect).
- (dashboard)/company-goals/page.tsx + a "Company Goals" sidebar nav link.
tsc --noEmit + eslint clean. Completes Phase 1 (Business Goals): data, briefing
injection, API, prompts, panel.
* fix(test): make test_app route assertions robust to FastAPI 0.137 _IncludedRouter
FastAPI 0.137 stopped flattening include_router into app.routes — each include is
now an _IncludedRouter (a BaseRoute with no .path), so `{r.path for r in
app.routes}` raised AttributeError and the two router-registration tests failed
(the bump arrived via the claude-agent-sdk update in uv.lock). Add
_registered_paths(): OpenAPI schema paths (the stable public contract) plus each
included router's prefix, which also covers the websocket /ws mount (never in the
schema). Drops the now-incorrect type: ignore[attr-defined].
* feat(research): pluggable web search/fetch for Board + PM agents
Add a provider-agnostic web-research capability so the Board and PMs can
ground decisions in current external evidence the knowledge base can't
answer.
- ResearchService selects a provider adapter from config: Tavily, Brave,
and Exa adapters plus a NullProvider that degrades gracefully when no
key is set. Result count and fetched-content size are clamped to caps.
- /api/research/search and /api/research/fetch: role-gated to Board + PMs
(and the CEO), with a per-agent/day Redis quota that fails open.
- roboco-search MCP server (web_search / web_fetch) calls those routes;
the provider key stays server-side and agent containers never egress.
Mounted per role by the orchestrator, behind a master switch.
- Charter-aware prompt guidance for Board, Main PM, and Cell PM.
Additive: with no key configured it is a no-op and the existing delivery
lifecycle is unchanged.
* feat(pitch): Board pitch -> CEO approve -> auto-provision repos
Add an additive origination path so a product can be proposed, approved,
and stood up without manual repo/Project setup.
- Pitch entity + migration (pitches table); PitchService create/list/
reject/approve.
- GitHubProvisioningService: the one place that creates repos (POST
/orgs/{org}/repos). Server-side token/org; when unconfigured the whole
approve path is inert and nothing is created.
- On approval: provision one repo per target cell, register a Project per
repo, create a Product when multi-cell, and seed one Main-PM delivery
task — all reusing the existing Product / coordination-task machinery.
- /api/pitches: Board authors (PO/HoM), CEO approves/rejects, Board+PM+CEO
view. Errors mapped via a single translator.
Additive: the delivery lifecycle is untouched; with no provisioning token
the capability is a no-op. Agent-facing pitch tool + panel are follow-ups.
* feat(strategy): dormant autonomous strategy engine (engine 2)
Add a second, optional engine that watches the company against its
standing goals and surfaces what needs the CEO — without touching the
delivery lifecycle (engine 1).
- StrategyEngine.assess() reports observations: the company is idle while
goals stand, and tasks stranded in 'blocked' past a threshold.
- run_cycle() notifies the CEO (notify-only; it never spends, builds, or
auto-approves — originating work stays a CEO decision).
- Orchestrator runs it on its own interval, started/stopped with the other
background loops; the loop returns immediately unless enabled.
DORMANT by default (strategy_engine_enabled=False): the loop never runs and
a standard deployment is unchanged. Auto-origination is a further opt-in.
* docs(changelog): record Business Goals, Web Research, Pitch->Provision, and the dormant strategy engine under Unreleased
* feat(secretary): wire the Secretary role end-to-end (foundation)
Add SECRETARY as a distinct role — the CEO's conversational chief-of-staff,
governed separately from the Prompter (which stays read-only/human-only).
This is the role foundation only; authority, the live agent, and the panel
land in following commits.
- foundation/identity: Role.SECRETARY (board level), seeded secretary-1 agent,
role-level mapping.
- journaling read tier (ALL — it advises the CEO), role_config entry,
per-role model (opus), prompt-layer mapping + roles/secretary.md.
- i_am_idle gains SECRETARY so the role has a verb surface.
- migration 034: add 'secretary' to the agentrole enum (mirrors 025).
- Role-registry tests updated for the new role.
Inert by itself (nothing spawns it yet); additive — existing roles unchanged.
* feat(secretary): directives + gate-list authority (backend)
The Secretary acts only under CEO command. Low-risk directives (relay a
dictated message) execute immediately; high-impact ones — charter edits,
task start/cancel/override, pitch approval, announcements — are recorded
pending and run only after the CEO confirms (the gate list).
- secretary_directives table (migration 035) as the command audit + queue.
- SecretaryService: read company state; submit (direct->run, gated->queue +
notify CEO); confirm/reject; execution runs with the CEO as actor through
the existing services (the Secretary never holds CEO authority itself).
- /api/secretary: submit + state/task reads (Secretary or CEO); list/confirm/
reject (CEO only). Writes commit explicitly.
* feat(secretary): live conversational agent (container + bridge)
Stand up the Secretary as a persistent Claude-SDK container the CEO chats
with, mirroring the Intake agent and reusing its driver/session machinery.
- secretary_driver: build_secretary_options exposes read_company_state /
read_task / submit_directive as SDK tools that call /api/secretary/* with
the agent's HMAC token; backend-call logic is module-level + tested.
- secretary_main: container entrypoint (receiver + relay) reusing IntakeDriver.
- orchestrator: start/spawn/reap secretary session + run-cmd builder; no
workspace clone (reads state via API), mints a role=secretary token.
- secretary_live routes: panel <-> container bridge over the live registry.
- agent-secretary image (Dockerfile + compose build service).
Inert until a session is started; additive — intake and all agents unchanged.
* feat(secretary): panel chat + directive confirmation queue
The CEO's Secretary surface: a live chat (SSE) to talk to the Secretary, and
a 'Needs your confirmation' queue listing gated directives the Secretary
proposed — each with Confirm / Reject. Adds the sidebar nav entry.
- lib/api/secretary.ts: live (start/stream/status/send/stop) + directive
(list/confirm/reject) + state clients (all as the CEO).
- hooks/use-secretary.ts: drives one chat, accumulating SSE token deltas.
- secretary page: chat pane + pending-directive cards.
Completes the Secretary end-to-end (role + authority + live agent + panel).
* feat(pitch): agent-facing pitch tool + pitches panel
Complete the pitch path: the Board can now author pitches through the gateway,
and the CEO reviews/approves them in the panel.
- content_actions.pitch (Board-only) -> PitchService.create, returning an
Envelope; wired as a do-tool (do_server + /api/v1/do/pitch + schema) and
added to the Board's do-tools.
- Panel /pitches page: lists pitches with CEO Approve & provision / Reject;
sidebar nav entry.
Pitch (Phase 4) is now end-to-end: author -> CEO approve -> auto-provision.
* feat(cockpit): read-only 'is the business winning?' summary
A pure aggregation for the CEO over existing data — no new state, no writes.
- CockpitService.summary(): charter north-star/objectives, delivery counts
(in-flight/blocked/awaiting-CEO), 30-day spend vs the charter's budget cap,
pending pitches, and the strategy engine's signals (what needs you). Stamped
basis='proxy' — performance is a proxy until real launches.
- GET /api/cockpit/summary (CEO / Board / Main PM / Secretary).
- Panel /cockpit page + sidebar nav.
Reuses goals + usage + StrategyEngine.assess(); reads only.
* docs(changelog): add the Secretary and Cockpit to Unreleased
* fix(test): isolate the company-goals empty-defaults test from committed state
The shared test DB persists committed writes across tests; a route test
commits a charter, so the unit test's 'unset' assertion must establish its
own clean precondition rather than assume global emptiness.
* fix(gateway): lower evidence_repo complexity to rank A (xenon gate)
company_goals()'s 4-way `or` emptiness check tipped the module average to
rank B; `any(...)` is equivalent and keeps the module under the gate's A bar.
* chore(compose): mirror agent-secretary-image build into docker-compose.yaml
Both compose files are byte-identical and tracked; .yaml carries the same
agent-secretary-image build service already present in docker-compose.yml.
* chore(lifecycle): regenerate artifacts for secretary i_am_idle
The secretary role gained i_am_idle in the lifecycle spec; regenerate the
generated prompt/doc/json artifacts so foundation-check stays green.
* docs(changelog): cut the company-in-a-box phases to 0.4.0
Label the six additive phases (business goals, web research, pitch-provision,
strategy engine, secretary, cockpit) as 0.4.0; tag v0.4.0 is held until the
branch merges to master so it points at the release commit.
---------
Co-authored-by: Renn F <rennf93@users.noreply.github.com>
612 lines
24 KiB
Python
612 lines
24 KiB
Python
"""
|
|
RoboCo Configuration
|
|
|
|
Environment-based settings using Pydantic Settings.
|
|
"""
|
|
|
|
from functools import lru_cache
|
|
|
|
from pydantic import Field, computed_field
|
|
from pydantic_settings import BaseSettings, SettingsConfigDict
|
|
|
|
|
|
class Settings(BaseSettings):
|
|
"""
|
|
Application settings loaded from environment variables.
|
|
|
|
Environment variables are prefixed with ROBOCO_ by default.
|
|
"""
|
|
|
|
model_config = SettingsConfigDict(
|
|
env_prefix="ROBOCO_",
|
|
env_file=".env",
|
|
env_file_encoding="utf-8",
|
|
case_sensitive=False,
|
|
extra="ignore",
|
|
)
|
|
|
|
# ==========================================================================
|
|
# Application
|
|
# ==========================================================================
|
|
app_version: str = "0.2.0"
|
|
debug: bool = False
|
|
environment: str = Field(
|
|
default="development", pattern="^(development|staging|production)$"
|
|
)
|
|
|
|
# ==========================================================================
|
|
# API Server
|
|
# ==========================================================================
|
|
host: str = Field(default="127.0.0.1", description="Use 0.0.0.0 for containers")
|
|
port: int = 8000
|
|
api_url: str | None = Field(
|
|
default=None,
|
|
description="Override API URL for containerized agents (e.g., http://roboco-orchestrator:8000)",
|
|
)
|
|
# CORS
|
|
cors_origins: list[str] = Field(
|
|
default=[
|
|
"http://localhost:3000",
|
|
"http://localhost:5173",
|
|
]
|
|
)
|
|
cors_allow_credentials: bool = True
|
|
|
|
@computed_field # type: ignore[prop-decorator]
|
|
@property
|
|
def internal_api_url(self) -> str:
|
|
"""
|
|
Internal API base URL for service-to-service communication.
|
|
|
|
Uses api_url if set (for containerized agents), otherwise builds from host/port.
|
|
Note: 0.0.0.0 is only valid for binding, not connecting - use 127.0.0.1 instead.
|
|
"""
|
|
if self.api_url:
|
|
return f"{self.api_url.rstrip('/')}/api"
|
|
connect_host = "127.0.0.1" if self.host == "0.0.0.0" else self.host # nosec B104
|
|
return f"http://{connect_host}:{self.port}/api"
|
|
|
|
# ==========================================================================
|
|
# Database
|
|
# ==========================================================================
|
|
database_host: str = "localhost"
|
|
database_port: int = 5432
|
|
database_user: str = "roboco"
|
|
database_password: str = "roboco"
|
|
database_name: str = "roboco"
|
|
database_echo: bool = Field(default=False, description="Log SQL queries")
|
|
database_pool_size: int = Field(default=10, ge=1)
|
|
database_max_overflow: int = Field(default=20, ge=0)
|
|
database_pool_timeout: int = Field(default=10, ge=1)
|
|
database_pool_recycle: int = Field(default=1800, ge=60)
|
|
|
|
@computed_field # type: ignore[prop-decorator]
|
|
@property
|
|
def database_url(self) -> str:
|
|
"""Async PostgreSQL connection URL."""
|
|
return (
|
|
f"postgresql+asyncpg://{self.database_user}:{self.database_password}"
|
|
f"@{self.database_host}:{self.database_port}/{self.database_name}"
|
|
)
|
|
|
|
@computed_field # type: ignore[prop-decorator]
|
|
@property
|
|
def database_url_sync(self) -> str:
|
|
"""Sync PostgreSQL connection URL (for Alembic)."""
|
|
return (
|
|
f"postgresql://{self.database_user}:{self.database_password}"
|
|
f"@{self.database_host}:{self.database_port}/{self.database_name}"
|
|
)
|
|
|
|
# ==========================================================================
|
|
# Redis
|
|
# ==========================================================================
|
|
redis_host: str = "localhost"
|
|
redis_port: int = 6379
|
|
redis_db: int = 0
|
|
redis_password: str | None = None
|
|
|
|
@computed_field # type: ignore[prop-decorator]
|
|
@property
|
|
def redis_url(self) -> str:
|
|
"""Redis connection URL."""
|
|
if self.redis_password:
|
|
return f"redis://:{self.redis_password}@{self.redis_host}:{self.redis_port}/{self.redis_db}"
|
|
return f"redis://{self.redis_host}:{self.redis_port}/{self.redis_db}"
|
|
|
|
# ==========================================================================
|
|
# RAG (in-house engine with pgvector)
|
|
# ==========================================================================
|
|
rag_persist_dir: str = ".roboco"
|
|
rag_chunk_strategy: str = Field(
|
|
default="fixed",
|
|
pattern="^(fixed|semantic|hierarchical|contextual)$",
|
|
description="Chunking strategy (fixed recommended, semantic loads extra model)",
|
|
)
|
|
rag_chunk_size: int = Field(default=512, ge=100)
|
|
rag_chunk_size_docs: int = Field(
|
|
default=1536, ge=100, description="Chunk size for docs (larger for 8K context)"
|
|
)
|
|
rag_chunk_size_journals: int = Field(
|
|
default=1024, ge=100, description="Chunk size for journals/reflections"
|
|
)
|
|
rag_chunk_overlap: int = Field(default=128, ge=0)
|
|
rag_auto_update_enabled: bool = Field(default=True)
|
|
rag_auto_update_interval: int = Field(
|
|
default=300, ge=60, description="Seconds between auto-updates"
|
|
)
|
|
|
|
@computed_field # type: ignore[prop-decorator]
|
|
@property
|
|
def rag_store_url(self) -> str:
|
|
"""PostgreSQL connection URL for the in-house vector store."""
|
|
return (
|
|
f"postgres://{self.database_user}:{self.database_password}"
|
|
f"@{self.database_host}:{self.database_port}/{self.database_name}"
|
|
)
|
|
|
|
# ==========================================================================
|
|
# AI/LLM Providers
|
|
# ==========================================================================
|
|
anthropic_api_key: str | None = None
|
|
|
|
# Default models
|
|
default_embedding_model: str = Field(
|
|
default="qwen3-embedding:0.6b",
|
|
description="Embedding model. Qwen3 Embedding for quality + 32K context.",
|
|
)
|
|
embedding_dimensions: int = Field(
|
|
default=1024,
|
|
description="Embedding dimensions (1024 for qwen3-embedding)",
|
|
)
|
|
|
|
# Local LLM for RAG answer synthesis
|
|
local_llm_model: str = Field(
|
|
default="glm-5:cloud",
|
|
description="Local LLM for RAG answer synthesis "
|
|
"(non-thinking models are faster)",
|
|
)
|
|
local_llm_base_url: str = Field(
|
|
default="http://roboco-ollama:11434/v1",
|
|
description="Base URL for local LLM (Ollama OpenAI-compat API)",
|
|
)
|
|
ollama_base_url: str = Field(
|
|
default="http://roboco-ollama:11434",
|
|
description="Base URL for Ollama native API (embeddings, model mgmt)",
|
|
)
|
|
|
|
# ==========================================================================
|
|
# Web Research (pluggable external search/fetch for Board + PM roles)
|
|
# ==========================================================================
|
|
# Calls go agent -> roboco-search MCP -> /api/research/* -> ResearchService
|
|
# -> provider. The provider key lives ONLY in this server-side process; it
|
|
# is never injected into agent containers, and agents never egress — the
|
|
# provider's own API does. Unset key => graceful NullProvider (empty
|
|
# results, no hard fail).
|
|
research_enabled: bool = Field(
|
|
default=True,
|
|
description=(
|
|
"Master switch for the web-research capability. When false the "
|
|
"roboco-search MCP server is not mounted into any agent container."
|
|
),
|
|
)
|
|
research_provider: str = Field(
|
|
default="tavily",
|
|
pattern="^(tavily|brave|exa|null)$",
|
|
description=(
|
|
"Web-search provider adapter. 'tavily' (LLM-native cited results "
|
|
"+ extract), 'brave' (independent index; no fetch), 'exa' "
|
|
"(neural search + contents), or 'null' (always-empty stub). "
|
|
"Swapping providers is a config change only."
|
|
),
|
|
)
|
|
research_api_key: str | None = Field(
|
|
default=None,
|
|
description=(
|
|
"API key for the selected research provider. Server-side only — "
|
|
"never reaches an agent container. Unset => NullProvider."
|
|
),
|
|
)
|
|
research_max_results: int = Field(
|
|
default=5,
|
|
ge=1,
|
|
le=20,
|
|
description="Hard cap on web_search results per call (top-k clamp).",
|
|
)
|
|
research_fetch_max_chars: int = Field(
|
|
default=20000,
|
|
ge=500,
|
|
description="Hard cap on extracted characters returned by web_fetch.",
|
|
)
|
|
research_timeout_seconds: float = Field(
|
|
default=15.0,
|
|
gt=0,
|
|
description="Per-request timeout for outbound provider HTTP calls.",
|
|
)
|
|
research_daily_quota_per_agent: int = Field(
|
|
default=50,
|
|
ge=1,
|
|
description=(
|
|
"Maximum web_search + web_fetch calls per agent per UTC day. "
|
|
"Tracked in Redis; fails open if Redis is unreachable."
|
|
),
|
|
)
|
|
|
|
# ==========================================================================
|
|
# Security
|
|
# ==========================================================================
|
|
encryption_key: str = Field(
|
|
default="",
|
|
description="Fernet encryption key for secrets.",
|
|
)
|
|
|
|
# ==========================================================================
|
|
# GitHub repository provisioning (pitch -> approve -> auto-provision)
|
|
# ==========================================================================
|
|
# The only place that CREATES GitHub repos (vs. clone/branch/PR existing
|
|
# ones). Server-side only; never injected into agent containers. Unset
|
|
# token/org => disabled => the pitch approval path is inert (no repo is
|
|
# created) until the CEO configures it.
|
|
provisioning_enabled: bool = Field(
|
|
default=True,
|
|
description=(
|
|
"Master switch for pitch auto-provisioning. With no token/org set "
|
|
"the capability is inert regardless of this flag."
|
|
),
|
|
)
|
|
provisioning_token: str = Field(
|
|
default="",
|
|
description=(
|
|
"GitHub PAT used to create repos in the provisioning org "
|
|
"(needs repo + org admin scope). Server-side only."
|
|
),
|
|
)
|
|
provisioning_org: str = Field(
|
|
default="",
|
|
description="GitHub organization where new repos are provisioned.",
|
|
)
|
|
github_api_base_url: str = Field(
|
|
default="https://api.github.com",
|
|
description="GitHub REST API base URL (override for GitHub Enterprise).",
|
|
)
|
|
provisioning_timeout_seconds: float = Field(
|
|
default=30.0,
|
|
gt=0,
|
|
description="Per-request timeout for outbound GitHub provisioning calls.",
|
|
)
|
|
provisioning_repo_private: bool = Field(
|
|
default=True,
|
|
description="Whether provisioned repos are created private.",
|
|
)
|
|
|
|
# ==========================================================================
|
|
# Autonomous strategy engine ("engine 2") — DORMANT by default
|
|
# ==========================================================================
|
|
# A separate background loop that watches the company against its standing
|
|
# goals and surfaces drift/idle/stranded work to the CEO (notify-only —
|
|
# never spends or builds). Default OFF: the loop never starts and the
|
|
# existing delivery lifecycle is untouched until the CEO opts in.
|
|
strategy_engine_enabled: bool = Field(
|
|
default=False,
|
|
description=(
|
|
"Master switch for the autonomous strategy engine. OFF by default; "
|
|
"when off the background loop does not run at all."
|
|
),
|
|
)
|
|
strategy_engine_interval_seconds: int = Field(
|
|
default=1800,
|
|
ge=60,
|
|
description="Seconds between strategy-engine assessment passes.",
|
|
)
|
|
strategy_stranded_blocked_minutes: int = Field(
|
|
default=120,
|
|
ge=5,
|
|
description=(
|
|
"A task blocked longer than this is surfaced as stranded "
|
|
"(needs a human decision)."
|
|
),
|
|
)
|
|
|
|
# ==========================================================================
|
|
# Workspaces (Multi-Agent Git)
|
|
# ==========================================================================
|
|
workspaces_root: str = Field(
|
|
default="/data/workspaces",
|
|
description="Root directory for all agent workspaces",
|
|
)
|
|
workspace_auto_clone: bool = Field(
|
|
default=True,
|
|
description="Automatically clone repos when workspace is first accessed",
|
|
)
|
|
workspace_clone_timeout: int = Field(
|
|
default=300,
|
|
ge=30,
|
|
description="Timeout in seconds for git clone operations",
|
|
)
|
|
workspace_refresh_fetch_timeout_seconds: int = Field(
|
|
default=60,
|
|
ge=5,
|
|
description=(
|
|
"Timeout in seconds for the best-effort `git fetch origin` "
|
|
"that runs on every healthy-clone re-entry into "
|
|
"ensure_workspace. Refresh fetches transfer small deltas only "
|
|
"— blocking 300s (the full-clone timeout) on every spawn "
|
|
"against a hung remote is operationally bad."
|
|
),
|
|
)
|
|
workspace_install_dev_deps: bool = Field(
|
|
default=True,
|
|
description=(
|
|
"After cloning an agent workspace, install the project's dev "
|
|
"dependencies into the workspace's own environment so the "
|
|
"`make quality` gate (ruff/mypy/pytest for Python, the lint/"
|
|
"typecheck toolchain for the TS panel) is available without "
|
|
"the agent re-downloading tooling per task. Detects Python "
|
|
"(pyproject.toml → `uv sync`) and Node/TS (package.json → "
|
|
"`pnpm install`/`npm install`). Idempotent; skipped when the "
|
|
"relevant lockfile is unchanged since the last install."
|
|
),
|
|
)
|
|
workspace_dep_install_timeout_seconds: int = Field(
|
|
default=600,
|
|
ge=30,
|
|
description=(
|
|
"Timeout in seconds for the post-clone dev-dependency install "
|
|
"(`uv sync` / `pnpm install`). Cold installs of a large TS "
|
|
"panel or a Python project with native wheels can take several "
|
|
"minutes; the default clone timeout is too short for this."
|
|
),
|
|
)
|
|
|
|
# ==========================================================================
|
|
# Transcript retention (agent Claude Code transcripts under ~/.claude)
|
|
# ==========================================================================
|
|
transcript_retention_days: int = Field(
|
|
default=14,
|
|
ge=1,
|
|
description=(
|
|
"Default retention window, in days, for agent Claude Code "
|
|
"transcripts (the *.jsonl files agents write under "
|
|
"~/.claude/projects). A background sweep prunes agent-owned "
|
|
"transcripts older than this. Panel-editable: a stored "
|
|
"`transcript_retention_days` system setting overrides this default "
|
|
"when present; this is the fallback used before one is set."
|
|
),
|
|
)
|
|
transcript_prune_enabled: bool = Field(
|
|
default=True,
|
|
description=(
|
|
"Whether the orchestrator background sweep prunes old agent "
|
|
"transcripts. Disable to keep every transcript indefinitely."
|
|
),
|
|
)
|
|
transcript_prune_interval_seconds: int = Field(
|
|
default=3600,
|
|
ge=300,
|
|
description=(
|
|
"Minimum seconds between transcript-retention prune passes. The "
|
|
"prune is age-based (days), so it need not run more than hourly."
|
|
),
|
|
)
|
|
|
|
# ==========================================================================
|
|
# Git command execution
|
|
# ==========================================================================
|
|
git_command_timeout_seconds: int = Field(
|
|
default=30,
|
|
ge=5,
|
|
description=(
|
|
"Default timeout in seconds for a single orchestrator-side git "
|
|
"subprocess (status, log, checkout, fetch, push, …). Short by "
|
|
"design — most git operations are sub-second."
|
|
),
|
|
)
|
|
git_commit_timeout_seconds: int = Field(
|
|
default=180,
|
|
ge=30,
|
|
description=(
|
|
"Timeout in seconds for staging + committing a changeset "
|
|
"(`git add` / `git commit`). Large multi-file changesets (e.g. "
|
|
"the Next.js panel) can exceed the 30s default-git timeout while "
|
|
"git hashes every object and the orchestrator re-chowns the "
|
|
"tree, so the commit choreography uses this longer budget."
|
|
),
|
|
)
|
|
git_network_timeout_seconds: int = Field(
|
|
default=120,
|
|
ge=30,
|
|
description=(
|
|
"Timeout in seconds for git ops that talk to origin (fetch / pull "
|
|
"/ push). A push or fetch on a large private monorepo from a "
|
|
"self-hosted runner can far exceed the sub-second local-op "
|
|
"default; short-budgeting it is what made open_pr time out before "
|
|
"the branch reached the remote."
|
|
),
|
|
)
|
|
|
|
session_idle_timeout_seconds: int = Field(
|
|
default=3600,
|
|
ge=30,
|
|
description=(
|
|
"Idle seconds before a messaging session is swept closed. The "
|
|
"previous 300s default was shorter than a human conversation pause, "
|
|
"so a person's chat session expired and reopened between messages."
|
|
),
|
|
)
|
|
|
|
protected_git_urls: list[str] = Field(
|
|
default_factory=list,
|
|
description=(
|
|
"Repo URL substrings a project may not point at (e.g. the roboco "
|
|
"source repo). Blocks agent commits/merges from reaching a protected "
|
|
"repository; set this to sandbox smoke-test projects."
|
|
),
|
|
)
|
|
|
|
# ==========================================================================
|
|
# Agent Guardrails (per-session budgets, loop detection, SLAs)
|
|
# ==========================================================================
|
|
agent_tool_call_warn: int = Field(
|
|
default=50,
|
|
ge=1,
|
|
description="Soft warning threshold for per-session tool calls",
|
|
)
|
|
agent_tool_call_halt: int = Field(
|
|
default=150,
|
|
ge=1,
|
|
description="Hard cap for per-session tool calls; orchestrator stops container",
|
|
)
|
|
agent_loop_threshold: int = Field(
|
|
default=3,
|
|
ge=2,
|
|
description="Identical tool+args repeats in the window that flag a loop",
|
|
)
|
|
agent_loop_window: int = Field(
|
|
default=10,
|
|
ge=2,
|
|
description="How many recent tool calls to inspect for loop detection",
|
|
)
|
|
agent_stop_attempt_allowance: int = Field(
|
|
default=1,
|
|
ge=1,
|
|
description="Stop-without-terminal attempts before auto-substitute",
|
|
)
|
|
|
|
# Per-(role, state) SLAs for stuck-task sweep; seconds.
|
|
agent_sla_developer_in_progress: int = Field(default=2 * 3600, ge=60)
|
|
agent_sla_developer_verifying: int = Field(default=30 * 60, ge=60)
|
|
agent_sla_qa_claimed: int = Field(default=30 * 60, ge=60)
|
|
agent_sla_documenter_claimed: int = Field(default=60 * 60, ge=60)
|
|
agent_sla_cell_pm_claimed: int = Field(default=4 * 3600, ge=60)
|
|
|
|
# ==========================================================================
|
|
# Agent Gateway
|
|
# ==========================================================================
|
|
manifest_host_dir: str = Field(
|
|
default="/app/manifests",
|
|
description=(
|
|
"Orchestrator-side directory where per-agent tool manifests are "
|
|
"written. Must be a path that's bind-mounted from the host "
|
|
"(see docker-compose.yml) so the docker daemon can in turn mount "
|
|
"the file into spawned agent containers as /app/tool-manifest.json."
|
|
),
|
|
)
|
|
public_base_url: str = Field(
|
|
default="http://127.0.0.1:8000",
|
|
description="Public base URL for commit-trailer links",
|
|
)
|
|
|
|
# Gateway coordination thresholds
|
|
# Single source of truth for "claim heartbeat is stale": consumed both by
|
|
# `trigger_filter` (deciding whether to QUEUE a fresh spawn) and by
|
|
# `_reap_stale_claims` (deciding whether to RELEASE the claim back to
|
|
# pending). Keeping them on one field guarantees both layers agree on
|
|
# the same tick — the reaper runs first, releases the row, and the
|
|
# queued spawn finds an unclaimed task. Splitting them into two fields
|
|
# opens a window where trigger_filter queues duplicate spawns against a
|
|
# claim the reaper hasn't yet released — pure dispatcher churn.
|
|
claim_stale_seconds: int = Field(
|
|
default=180,
|
|
ge=60,
|
|
description="Claim heartbeat staleness threshold (seconds)",
|
|
)
|
|
# Reaper window for stale-claim detection. Dogfooding reaped agents at
|
|
# ~180s while they were actively
|
|
# retrying — LLM inference + retry loops routinely exceed 3 min
|
|
# between verb successes. 600s is large enough to accommodate that
|
|
# without letting a genuinely-stuck container linger.
|
|
# Distinct from claim_stale_seconds (which drives trigger_filter
|
|
# spawn queueing); keeping them separate avoids a window where a
|
|
# higher reap threshold would also delay spawn-queue decisions.
|
|
stale_claim_reap_seconds: int = Field(
|
|
default=600,
|
|
ge=60,
|
|
description=(
|
|
"Reaper-only stale claim threshold (seconds); "
|
|
"override via ROBOCO_STALE_CLAIM_REAP_SECONDS"
|
|
),
|
|
)
|
|
# A task left CLAIMED/IN_PROGRESS with an assignee but no running container
|
|
# (e.g. a reassignment that didn't spawn) is invisibly stuck — the heartbeat
|
|
# reaper can't see it because its heartbeat was seeded fresh at claim time.
|
|
# After this short grace window the dispatcher (re)spawns the assignee, or
|
|
# releases the task to pending for re-dispatch. Shorter than the
|
|
# heartbeat reaper window: this is the "no agent at all" case, not the
|
|
# "agent went silent mid-run" case.
|
|
claimed_no_agent_grace_seconds: int = Field(
|
|
default=120,
|
|
ge=30,
|
|
description=(
|
|
"Grace window (seconds) before the orchestrator (re)spawns or "
|
|
"releases a claimed/in_progress task that has no running agent; "
|
|
"override via ROBOCO_CLAIMED_NO_AGENT_GRACE_SECONDS"
|
|
),
|
|
)
|
|
# Pre-gateway parity: PMs wrote a fresh
|
|
# journal:decision around each decision point, not once at task
|
|
# creation. The PM-decision tracing gate (delegate, unblock,
|
|
# escalate_up, escalate_to_ceo) treats decisions older than this
|
|
# window as missing, forcing a new note(scope='decision', ...) on
|
|
# each pass through the gate.
|
|
pm_decision_window_seconds: int = Field(
|
|
default=300,
|
|
ge=1,
|
|
description=(
|
|
"Recency window (seconds) for PM journal:decision to satisfy "
|
|
"gating verbs; override via ROBOCO_PM_DECISION_WINDOW_SECONDS"
|
|
),
|
|
)
|
|
spawn_cooldown_seconds: int = Field(
|
|
default=60,
|
|
ge=1,
|
|
description="Per-task spawn rate cooldown (seconds)",
|
|
)
|
|
role_spawn_rate_per_minute: int = Field(
|
|
default=6,
|
|
ge=1,
|
|
description="Per-role spawn rate limit (per minute)",
|
|
)
|
|
|
|
# Tracing-gate thresholds
|
|
qa_notes_min_chars: int = Field(
|
|
default=80,
|
|
ge=1,
|
|
description="Minimum characters for QA notes",
|
|
)
|
|
docs_notes_min_chars: int = Field(
|
|
default=20,
|
|
ge=1,
|
|
description="Minimum characters for docs notes",
|
|
)
|
|
|
|
# Commit-validator thresholds (wired into the gateway commit() gate)
|
|
commit_subject_min_chars: int = Field(
|
|
default=20,
|
|
ge=1,
|
|
description="Minimum characters for a commit subject",
|
|
)
|
|
commit_banned_words: tuple[str, ...] = Field(
|
|
default=(
|
|
"wip",
|
|
"tmp",
|
|
"asdf",
|
|
"oops",
|
|
"fix",
|
|
"update",
|
|
"change",
|
|
"stuff",
|
|
"things",
|
|
),
|
|
description="Banned single-word commit subjects",
|
|
)
|
|
|
|
|
|
@lru_cache
|
|
def get_settings() -> Settings:
|
|
"""Get cached settings instance."""
|
|
return Settings()
|
|
|
|
|
|
# Global settings instance
|
|
settings = get_settings()
|