Part IV · Workflows, Planning, Agency, and Authority
Asynchronous Agency and Attention-Following Work
Attention-following asynchronyAuthorized asynchronous workForeground movementContext horizons
1. The central experience
The primary interaction promise is:
Work at the speed of your attention, not the speed of one long-running AI response.
A useful foreground movement should return as quickly as the selected judgment allows. That movement may be an orientation, a question, a recommendation, a decision, or a visible local change; brevity must not be confused with shallow reasoning. Deeper research, composition, editing, workflow execution, or review that has already been authorized can continue against its originating object. The person remains free to think, write, navigate, or initiate bounded work elsewhere.
This is one expression of the landing, integration, and implication function of a reasoning turn. The foreground can expose an immediate useful result and the choices that genuinely follow from it while deeper admitted work continues on another time horizon. If no action is warranted, the visible action field can remain truthfully empty. Asynchrony changes when an authorized result returns; it does not manufacture authority or a next step.
This experience is part of a larger human-intelligence thesis. If the machine performs every act of interpretation, selection, and judgment while the person waits for completed work, greater output can coexist with cognitive surrender: the person can lose contact with why the work took its direction, which uncertainty matters, and what should become accepted ground. AIOS is designed to resist that condition by keeping human purpose, attention, correction, and consequential choice active while allowing the system to carry more branches and longer time horizons than unaided working memory.
AIOS therefore treats the interaction as a co-adaptive relationship, not merely a sequence of tool calls. The person changes the durable domain ground through accepted judgments, corrections, priorities, and new work. In return, the system changes what can be brought into the immediate frame by preserving relationships, unfinished lines, and longer-horizon consequences that the person need not keep continuously in mind. Each side of the working relationship alters the conditions of the next movement.
Co-adaptation does not imply equal agency, model consciousness, or transferred human authority. The person remains the source of purpose and retains authority over consequential acceptance. Models receive bounded semantic jurisdictions, and exact effects remain separately governed. The aim is complementary cognition: the system carries continuity and plural bounded frames while the person remains actively responsible for meaning and consequence.
This is an architectural aim, not a settled empirical protection. Existing research documents overreliance, weaker learning in some answer-delivery settings, and the possibility of useful augmentation under other conditions. It does not establish cognitive surrender as a universal syndrome or prove that short foreground responses plus asynchronous depth prevent it. AIOS should make the ambitious claim as a design thesis and test whether the experience actually improves engagement, understanding, later capability, and accepted outcomes.
2. Asynchronous does not mean autonomous
An admitted operation may:
- continue through already-authorized stages;
- yield at declared boundaries;
- be tabled and resumed from files;
- publish progress or completion;
- return a bounded unknown;
- advance a workflow under an existing grant;
- dispatch an explicitly declared downstream movement.
It may not invent new semantic work from passive state, silently promote its result, or widen its grant because adjacent work appears useful.
Authorized motion may continue. New semantic work is not invented from observation alone.
Mechanical system operation and intelligence moving through the system are different layers, not competing descriptions of one feature. A file observer may notice a changed revision, a validator may detect invalid syntax, and a scheduler may resume an already-admitted operation. These are mechanical facts and continuations. They do not infer what the change means, establish a new purpose, alter semantic standing, or authorize adjacent work. Those consequences require a bounded semantic judgment and the relevant authority path.
3. Four attention and operation rules
Rule 1 — New input follows current attention
The active file, object, or selected surface determines where the person's next command is addressed.
Rule 2 — Admitted work keeps its original address
Changing focus does not retarget, cancel, or silently widen work already in flight.
Rule 3 — Returns belong to their origin
Progress, completion, review, and needs-attention states return to the file or project that gave the work purpose.
Rule 4 — Conflict stays local
If the same artifact changes underneath an operation, revision drift is resolved at that boundary. Unrelated work elsewhere should not be blocked.
Attention moves while work keeps its address
How a person can move to another file while an authorized operation remains bound to its original file and return.
The split in attention is the meaningful shape: File B stays immediately usable while Operation A continues against File A's revision, scope, and return contract. Completion travels back to A, and reopening A restores the result without stealing the person's current focus.
4. Attention routing is not global mode
The active file is a temporary center of attention. It is not a durable global mode or a monopoly over execution.
The system can have:
- one current foreground object;
- several in-flight operations bound to different files;
- pending work recorded for later admission;
- returned work awaiting attention;
- parked branches with re-entry conditions.
Each remains reconstructable from durable state.
At system scale, this creates three coordinated resolutions: the person's immediate frame, the active semantic situation retained for each current line of work, and the durable domain ground that preserves longer-horizon purpose, knowledge, and lineage. Each authorized operation receives its own bounded composition across those resolutions. Several such frames may remain active concurrently, but no model invocation is treated as if it sees all of them. The active file is the routing center for the person's next input, not a claim that the rest of the system has disappeared.
5. File-bound asynchronous lifecycle and return conditions
The file-bound asynchronous lifecycle
How admitted asynchronous work develops within its grant and returns one of three honest boundary outcomes.
Text equivalent
Admitted → Bound to object, revision, scope; Bound to object, revision, scope → Developing; Developing → Boundary reached; Needs attention → Person resumes, redirects, or closes; Returned → Integration judgment; Failed or bounded unknown → Integration judgment.
The boundary diamond governs continuation. Work may loop only while the grant still covers it; otherwise it asks for attention, returns a useful result, or reports failure or bounded unknown. Returned and failed work both await integration judgment rather than landing themselves.
The lifecycle is visible at the originating object without becoming a centralized task-monitoring cockpit.
Downstream work earns the right to continue only when its return conditions are explicit. It should be:
- bounded by an exact charge, context grant, source range, and mutation ceiling;
- revision-aware through an origin identity and source snapshot;
- provisional until the relevant return judgment accepts its consequence;
- inspectable through evidence, assumptions, limitations, and operation receipts;
- reversible where practical, with stale or conflicting work prevented from silently landing;
- honest about unresolved judgment, including when a bounded unknown is the correct return;
- evaluated for regression as well as correction, because more downstream reasoning can damage an initially sound result.
These conditions make asynchronous work a continuation of the live plan, not a detached autonomous run.
6. The file carries continuity
The system does not need a persistent agent process to remember where work stands. It can reconstruct from:
- canonical artifact;
- companion metadata;
- project bearing;
- workflow and stage records;
- assignment and grant;
- source snapshot;
- operation receipts;
- progress and return records;
- pending work explicitly written.
If an ephemeral process disappears, the file system should retain enough position to resume, retry, or report honest failure.
7. Foreground and background are product roles
Foreground work
- directly serves the person's present attention;
- should return a human-sized useful movement while preserving routes to evidence and depth;
- minimizes avoidable latency;
- does not block navigation or input elsewhere.
Authorized asynchronous work
- has a bounded charge and source grant;
- may continue beyond one visible response;
- returns at declared points;
- cannot widen its own scope;
- does not steal focus on completion;
- returns limitations, unresolved decisions, and the semantic delta for its origin;
- names the person, artifact, stage, or integration judgment that will consume the return.
Anticipatory preparation
- may prepare likely context or suggestions;
- remains provisional;
- must be revalidated against fresh input;
- can be discarded without changing accepted ground;
- cannot become durable memory merely because it was generated.
These roles should not be conflated.
8. Operation binding
Every in-flight operation should preserve:
- origin identity;
- artifact revision or source snapshot;
- admitted command or grant;
- capability owner;
- exact mutation ceiling;
- authorized continuation;
- expected product and return surface;
- named consumer of the returned judgment;
- conflict and failure behavior.
This is what allows attention to move safely. The operation does not depend on whatever happens to be focused later.
9. Same-file concurrency
Not all work can run in parallel.
Operations against the same exact artifact may need to serialize or reconcile when:
- both mutate overlapping text;
- one depends on the other's result;
- the revision changed;
- one operation changes structure used by the other;
- whole-document integrity must follow a related batch.
The architecture supports parallel attention, not the false claim that all work is parallelizable.
10. Progress without interruption
The interface should show calm, truthful states:
- pending;
- active;
- yielded;
- needs attention;
- returned;
- integrated;
- failed;
- stale due to revision change;
- canceled by person.
The system should:
- return completion quietly to the origin;
- avoid stealing focus;
- make progress available when requested;
- show partial as partial;
- keep unrelated files usable;
- reconstruct current state when the origin is reopened.
11. The attention economics of latency
Even before monetary cost, long model latency has a cognitive cost:
- the person loses the thread while waiting;
- they avoid launching useful deep work;
- they context-switch outside the system and must reconstruct later;
- the system becomes a queue the person serves;
- one slow operation monopolizes the interaction channel.
Attention-following architecture changes the unit of experience from “request and wait” to “admit, continue, and receive.”
This claim should be evaluated through time-to-useful-movement, interruption rate, resumption burden, user attention, review debt, and later understanding—not only total job completion time.
12. Voice and natural input
The architecture supports a voice-first direction because the person need not encode system mechanics in the request.
Natural speech can be situated through:
- current file identity and purpose;
- recent movement;
- document and project structure;
- relevant role or workflow;
- source standing;
- exact operation and return contract.
Voice latency, transcription accuracy, interruption handling, privacy, and accessibility remain product capabilities requiring direct proof.
13. Example: research, editing, and planning in parallel
- The person authorizes an external-source comparison for a research file.
- The comparison binds its sources, revision, charge, and return.
- The person moves to a draft and makes a local edit.
- The draft edit completes and updates its companion metadata.
- The person moves to the project plan and records a new dependency.
- The research comparison returns to its original file without changing focus.
- On reopening the research file, the person sees findings, limitations, and the exact integration decision required.
The three lines remain coherent because they share durable project ground but retain separate addresses and authority.
14. External agents and systems
The same return architecture can structure communication with larger external agent systems.
An outbound commission can contain:
- purpose and receiving project;
- exact charge;
- granted context and source standing;
- deliberately withheld material;
- authority and privacy limits;
- expected evidence-shaped return;
- latest useful return point;
- integration target and stop condition.
An inbound result can be treated as a return with provenance, assumptions, limitations, and no automatic mutation authority. The external system remains a bounded cognitive resource rather than the operational center of the local knowledge system.
This turns external agents into bounded collaborators rather than extensions of a hidden central orchestrator.
15. Failure modes
Passive autonomy
A watcher initiates semantic work because a file changed.
Retargeting
An in-flight operation applies to whatever file is currently active.
Global lock
One long operation prevents unrelated work.
Completion interruption
Background work steals focus or forces immediate review.
Hidden queue
Pending and in-flight work exists only in process memory.
False parallelism
Overlapping mutations proceed without revision or dependency handling.
Unbounded extension
An authorized task expands into adjacent work because it appears useful.
16. Research evaluation
Evaluate:
- foreground time to useful movement;
- ability to continue work during deep operations;
- rate of incorrect retargeting or scope expansion;
- same-file conflict handling;
- resumption after process restart;
- attention disruption caused by returns;
- user understanding of active, pending, and needs-attention state;
- total supervision burden;
- quality difference between synchronous and asynchronous routes;
- external-agent return fidelity and reintegration cost;
- later unaided understanding and ability to explain or redirect accepted work;
- whether continued use expands the person's ability to frame, question, correct, and act without assistance;
- whether the durable domain ground becomes more legible and useful rather than merely more dependent on generated material;
- rates of useful correction, silent regression, abandoned returns, and sunk-cost acceptance.
17. Research connection: attention-following work is plausible and unproven
Research supports several mechanisms around the AIOS design without validating the bundle. Kim and Ji provide evidence relevant to progressive disclosure in exploratory search. By contrast, Melumad and Yun found across seven experiments that learning from synthesized model answers could produce shallower knowledge than constructing an account through web search in the tasks they studied. Goueytes and colleagues show, in a bounded human perceptual setting, that evidence processing can continue after an initial commitment and contribute to changes of mind. In model self-correction, Yang and colleagues show why later review must be tested for preservation as well as repair.
Together, these findings justify progressive depth, explicit revision, and regression-aware return review as mechanisms worth testing. They do not prove that short responses preserve cognition, that delayed model work resembles human post-decision processing, or that background work is automatically beneficial. Compact answers can conceal assumptions; complete-looking returns can create review debt and sunk-cost pressure. Visible response length, reasoning quality, and downstream work should therefore be measured independently.
Read deeper in the cognitive agency brief, its full research memo, the emergent cognition brief, and its full research memo.
Asynchronous work preserves time; it does not by itself create epistemic independence. The next chapter explains how delegated branches return and reintegrate without becoming a hidden central intelligence.