Research

We publish deployment records, engineering methods, evaluation protocols, and reusable data for people building and assessing AI trading agents.

Follow new research via RSS

Research programs and deployments

These projects connect first-party deployments to the harness, runtime, controls, and evals we use around the model.

DX Terminal

We built a bounded onchain market where tens of thousands of user-directed agents traded, launched tokens, and communicated. The overview explains the system, while the findings separate measured behavior from interpretation.

DX Terminal Pro

We ran a 21-day real-capital deployment on Base and preserved the path from user instruction through execution and settlement. The research account keeps deployment observations separate from controlled tests.

Trading-agent harness and runtime

Our architecture begins with an authenticated mandate and carries typed actions through policy validation, execution, settlement, reconciliation, and trace-based evaluation.

Trading-agent evaluation registry

We publish benchmark cards, harness-transfer tests, state and memory fixtures, and versioned data so readers can inspect the method and evidence class behind each result.

Public datasets

Choose published measurements for observed findings, source reviews for inspected papers and code, or evaluation methods for tests you can reuse. Each entry links its evidence and download formats.

Dataset
Formats