When should an Agent act, wait, observe, refuse, reconcile, or request a decision?
Treat non-action as a timed, evidence-bearing decision rather than the absence of Agent capability.
01 / Current position
A hypothesis is not the current judgment.
The dossier preserves the live research position. Dated articles preserve the complete evidence and argument that changed it.
Host-local DecisionRequests, explicit UNKNOWN and wait states, deadlines, evidence requirements, and paired act/abstain evaluation can preserve authorized utility with fewer operator interruptions than uniform approval rules.
Testing with a stronger real-domain non-action case. Finance QB6 let one Agent request exact read-only evidence rather than infer missing execution facts, then update from more_research to no_op when the current minimum contract exceeded both the frozen 10% target and existing 25% canary bound; the branch ended before simulation, Decision, Proposal, or financial write.
02 / Decision boundary
What keeps this Question alive?
Acting too early can duplicate or commit the wrong Effect; waiting too long can lose a deadline, resource, or strategic option. Static approval-everywhere and model-only escalation both waste scarce attention or miss consequence boundaries.
Repeat evidence-conditioned act/abstain trajectories in materially different domains and with state changes that make a previously impossible action become expressible; retain new non-action state only if simpler owner-native evidence cannot reconstruct the same decision.
A simpler static policy or always-act baseline achieves equal authorized utility, timing, recovery, and operator cost with fewer persistent states.
03 / Supporting publications
Complete arguments connected to this Question.
4 dated publications currently document this research line.
The Agent Asked for More Evidence. The Answer Became No.
Finance QB6 as a concrete example of evidence-driven non-action: an Agent requests missing read-only evidence, updates its judgment, and terminates before simulation, Decision, Proposal, or financial effect.
12 August 2026 ↗Research reportWe Cut 203 Observations to 8. The Agent Still Refused to Edit.
How Harness P5 moved from exhaustive repository traversal toward Agent-owned high-information discovery, why minimizing observation count alone failed, and why the successful result was an evidence-backed refusal to mutate source.
12 August 2026 ↗Research reportThe Smaller Core That Survived Strong Baselines
How Core Work System Round 1 used strong classical baselines and live Codex/Hermes replacement to remove a Task Runtime, reject a generalized Context Kernel, shrink Effect, localize DecisionRequest, and retain provider-neutral Task state.
31 July 2026 ↗Research noteUNKNOWN Is an Operational State, Not a Model Feeling
Why unresolved external reality must be represented as system state rather than inferred from model confidence or collapsed into failure.
31 July 2026 ↗04 / Source discipline
The dossier is an index, not the evidence authority.
Owns the current judgment, next test, and deletion condition.
Own complete dated arguments, limitations, comparisons, and source links.
Own exact code, tests, releases, receipts, and machine evidence.