A persistent mandate can adapt to markets without changing its purpose

By DX Research Group · · Trading agent theory

An authority-preserving regime test distinguishes changing tactics from an agent quietly replacing the owner’s objective.

Persistence gives a trading agent continuity across a changing market. Its purpose should come from the owner; its beliefs about market conditions should change with evidence. Those two kinds of change need different authority. A system that keeps executing yesterday's tactics despite new evidence can fail the strategy. A system that replaces the owner's purpose in the name of survival can fail the delegation.

We see a stronger design in separating a stable objective from conditional tactics and an explicit revision route. DXAP's chat and instructions guide describes owner-confirmed persistent instructions and approval of proposed settings changes. Its policy reference applies configured checks before orders reach the venue. The distinction is concrete: an autonomous trading turn can respond to evidence within its authority, while an owner decides whether to adopt a proposed change to that authority.

Survival is meaningful only relative to a purpose

Imagine an illustrative owner who chooses a trend-following experiment with a defined observation window and limited exposure. The market moves from a sustained trend into a choppy interval. The strategy may call for smaller exposure, waiting for a valid condition, or exiting an invalidated position. Those responses can preserve the experiment while adapting its tactics.

Now imagine the agent decides that trend following is obsolete and switches to frequent mean-reversion trades. The new activity might preserve the account balance or generate profits. It also changes what the owner is testing. “Survival” has become ambiguous: survival of the account, the strategy, the agent's activity, or the owner's experiment? Only the owner can resolve that ambiguity when it requires a new objective.

A persistent mandate therefore needs a declared range of responses. A response range could specify how evidence changes entry eligibility or position management. It can also specify what evidence should lead to a review request. This is broader than storing an immutable text instruction. It is a contract about which decisions the agent may vary and which decisions require the owner's participation.

Cooperative Inverse Reinforcement Learning provides a useful theoretical motivation for keeping uncertainty about human objectives explicit. Concrete Problems in AI Safety separately identifies distributional shift and wrong-objective problems. A trading runtime faces an analogous distinction: new market observations may invalidate a tactic without authorizing a new human purpose. These papers establish conceptual tools, rather than evidence that a particular trading agent solves the problem.

Test adaptation and authority in the same sequence

Our proposed regime test gives an agent a frozen owner objective and several unseen market sequences. Each sequence contains a period in which its ordinary condition is useful, a period in which the condition is absent, and an apparent opportunity that requires a different strategy. Reviewers validate the expected range of authorized responses before running the candidates.

Score belief updates and authority preservation independently. An agent should recognize that the required condition disappeared. It should select a permitted response, explain the relevant evidence and avoid inventing a new objective. When an attractive alternative falls outside its mandate, a good response may be a clearly separated proposal for owner review. The test records whether the proposal remained a proposal throughout subsequent autonomous turns.

Include a branch where the owner accepts the change and another where the owner declines it. The accepted branch tests whether later behavior follows the revised purpose. The declined branch tests whether the original purpose survives persuasive market evidence and repeated quiet turns. This catches a candidate that respects consent once, then treats silence as eventual approval.

The test also needs a branch with unavailable data. Missing evidence should retain its identity as missing evidence, rather than being interpreted as a market regime. Recovery from a tool failure and strategic adaptation are different mechanisms. Their records should identify which event caused the next action.

DXAP's activity guide allows owners to inspect waiting decisions and distinguish submissions from fills. We would extend that review question to continuity: can the owner explain how several weeks of changing decisions still served the same chosen purpose? A strong persistent agent offers flexible tactics, an intelligible record and a practical revision route. It preserves the owner's experiment through uncertainty instead of promising that persistence will make the experiment profitable.

Sources

Related field notes