Work handed to a second model, which may not ask

subagent_run gives a self-contained piece of work to a helper carrying the
parent's connection, directory, model and effort, and hands its answer back as
the tool result. The mechanism is the one scheduled runs already use -- a hidden
chat, one turn, wake_chat, and a poll -- so tools, rounds, budgets, metrics and
steps all work with no second implementation. The two alternatives were
rejected where they had already been rejected once: a nested Generation is two
replies writing one transcript, and a one-shot complete() has no tools, which
schedule/runner.py records as useless for exactly this case.

Every restriction is a property of the child's row, applied by resolve_tools
after the gates, because a rule that lives in a system message is one a page the
model just read can argue with. No questions, no recursion, nothing that writes
unless the call asked for it and the parent's own mode would not have stopped
first, and commands only from a fixed read-only list -- in every mode including
Auto, because the task text can have come from a page.

Withdrawing ask_user turned out to be half of "nobody is watching". An approval
still built a card nobody could see and parked the reply until approval_timeout,
which from every screen is the feature not working. Chat.unattended is the
question now, and not the kind: _authorise answers with a refusal instead. A
scheduled task's chat had the same hole and is covered by the same flag.

Three bounds, counted where each is knowable: per reply on the parent's
Generation, instance-wide in a set a restart clears, and per helper in settings
of its own so one runs out of room long before the reply that asked. Past the
clock the helper is stopped rather than abandoned, so a partial answer comes
back with a sentence saying so.

Also: four gates had shipped into the scope menu with no name, taking the first
tool's label instead -- the canvas switch read "Canvas written". There is a test
that refuses a family without one.

Co-Authored-By: Claude Opus 5 (1M context) <noreply@anthropic.com>
This commit is contained in:
Jaroslav Beneš
2026-08-06 15:05:30 +02:00
parent 0fa05c88b2
commit 46066150d9
17 changed files with 1850 additions and 13 deletions
+15
View File
@@ -253,6 +253,21 @@ def test_no_tools_means_no_list(db, chat, user_id):
assert "The tools you have on this request" not in text
def test_every_gate_that_can_be_offered_has_a_name_of_its_own(db):
"""A gate covers several tools, so no single tool's label is the right name
for one -- and the fallback is the *first* tool's label, which is a noun
phrase describing something that happened rather than a switch. Four gates
had shipped that way: the canvas switch read "Canvas written" and the report
one "Report filed". This is the direction it rots, because a new family
passes every other test with no label at all.
"""
from lembas.api.pages import _GATE_LABELS
gates = {tools_service.gate_of(family) for family in tools_service.FAMILIES}
missing = sorted(gates - set(_GATE_LABELS))
assert not missing, f"no menu label for {missing}"
# --- The control that writes ------------------------------------------------------------
def test_the_verb_is_on_every_checkbox(client: TestClient, db, chat, registered):
"""The element carrying `name` has to be the element carrying the request.