Question
A bounded uncertainty whose answer can change structure, priority, or project scope.
Research · 22 August 2026
Each question is tied to a decision that more evidence can still overturn: revise the standing, change the experiment, keep or narrow a boundary, materialize something needed to learn or act, or deliberately leave implementation unchanged.
Start here
Begin with the work that can still change the system, the boundary most recently answered, and the experiment that removed the most proposed machinery.
Wait for a natural owner-level failure that satisfies an already-frozen strict burden: a real certificate/language or grounded-meta-closure case where the current closure is prospectively explicit and the repair cannot be reconstructed as ordinary composition inside the admitted substrate. Do not manufacture E1b/E2 merely because apparatus exists.
1 supporting publications ↗Recently answered boundaryAnswered for the HD0–HD9 programme. The frozen 1760–2026 campaign closed with 320 trajectories and 96 Deep Anchors across eight chronological waves. Cross-era recurrence survived only as a typed seven-family structural map plus explicit demotions, merges, replay negatives, and measurement/currentness results; the evidence did not justify ten universal success invariants or eight universal innovation laws.
Read the accepted answer ↗Experiment that most changed the architectureHow Harness P5 moved from exhaustive repository traversal toward Agent-owned high-information discovery, why minimizing observation count alone failed, and why the successful result was an evidence-backed refusal to mutate source.
5 min research report ↗Full research index
Use the explorer after the current priorities are clear. It groups the same research by question, project, publication date, or status.
Research atlas
Distinguish failures of a callable action interface, representation/hypothesis language, physical affordance, installed executor capability, and artifact/grounding so repair happens at the level that is actually closed.
Treat non-action as a timed, evidence-bearing decision rather than the absence of Agent capability.
Pressure quantitative research, proposal/decision semantics, signer and credential authority, one-shot dispatch, and venue reconciliation without turning a backtest or model conclusion into permission to trade.
Search game form before Agent role: use cheap falsifiers and bounded playable proofs to find a form worth playing, then test whether Agent participation creates value a simpler design cannot.
Separate richer perception, evidence-grounded meaning, real encounter, typed consequence, search history, and independent holdout before promoting a creative prior.
Use a stable project map, dated publications, Feynman-style causal explanation, and exact owner links while keeping rendered encounter evidence separate from technical truth.
Determine the smallest durable objects required when work waits, fans out, and later joins.
Require stable identity, enforceable preconditions, receipts, replay/reconciliation, and ambiguity behavior before promoting another physical Tool contract.
Treat opponent state as a trigger-gated hypothesis: require held-out repeated encounters where history changes decisions beyond current observations and ordinary trajectory before adding durable strategic state.
Use unrelated owners and applications to distinguish genuinely recurring responsibility boundaries from repository-local packaging or a frozen Host→Harness→Runtime architecture.
Historical Agent-first research that separated strategy existence, policy accessibility, model realization, and simpler classical substitutes before the Game programme moved to GameForm-first direction search.
Retain durable Task identity and the bounded semantic frontier without turning Host into a Goal coordinator, executor, Runtime proxy, or domain-completion authority.
Use mature workflow, retrieval, idempotency, and Provider systems to delete shared abstractions that do not own a distinct failure.
Compare compact owner-native causal prose with richer cards, relation vocabularies, and explicit diagnostic grammars before promoting another mandatory representation.
Separate historical validity, current applicability, and downstream commitment; revalidate the exact load-bearing dependency with its owner before a new consequence rather than using age or one global freshness service.
Separate deterministic byte identity from semantic tractability, and falsify wire/evaluator ambiguity before attributing Agent errors to presentation order.
Freeze the modern theory, reconstruct history before coding it, preserve failures and anti-cases, and let independent epochs narrow or contradict the candidate laws before any cross-era promotion.
Freeze context treatments before holdout reveal and require replicated marginal benefit before turning persuasive domain knowledge into mandatory prompt state.
Preserve Provider-faithful execution while retaining only shared Run facts required for recovery and completion.
Test whether one bounded Agent-Run authority can preserve cognition, Provider/Tool continuity, recovery, action admission, and completion proposals across callers without becoming a second Task/domain store or universal planner.
Separate model-visible semantic history from opaque Provider continuation bytes whose authority comes from an exact prior Provider Tool result and lineage.
Determine the smallest inspection, recovery, backup/restore, deployment, and MCP surface needed to operate Host without expanding it into an operations platform.
Translate first-principles economic reasoning into a conditional, practical path with explicit trade-offs and no individualized advice claim.
Prevent semantic convenience, environment closure, external-effect policy, or cleanup pressure from expanding physical execution authority without a reproduced unreconstructable fact.
Separate the Security-owned semantics that remain decision-relevant after ordinary scanners, fuzzers, execution carriers, defensive mechanisms, providers, Host, Harness, and Runtime are consumed directly.
Record the negative result that direct owner-local integration carried the tested external recovery responsibilities.
Separate retained operational capability from the rejected semantic project thesis.
Archive stale active-history material, delete research-only APIs, replay the same mixed failures, and retain only cross-owner reality relations whose absence recreates a current failure.
Research contract
Question summaries evolve as evidence accumulates. Publications preserve the complete argument at a date. Repositories, tests, releases, receipts, and observed behavior remain authoritative.
A bounded uncertainty whose answer can change structure, priority, or project scope.
A dated complete argument that records evidence, limits, conclusions, and the next test.
The owning repository, receipt, release, or report that remains authoritative for exact technical facts.
System context
Use the System page to follow one work trajectory through the current boundaries. Return here when the useful unit is the uncertainty that could cause a component to change or disappear.
Follow the work trajectory