WORKBENCH GUIDE / EXECUTE · OBSERVE · EVALUATE

Follow the evidence.

A practical guide to your first test, an honest comparison, and the evidence behind each result.

Open workbench

01 / YOUR FIRST TEST

One assumption.
Two independent tests.

  1. Open the workbench. Wait for verified runtime readiness.
  2. Keep the first transfer preset: 10 displayed units, with a multiplier changing from 1 to 2. The deliberately faulty reference ignores that scale.
  3. Press Run test. Inspect the observed delivery and the assertion verdict.
  4. Pin for comparison, choose Configure counterpart, then explicitly run again.
  5. Open Compare runs. Check equivalent semantic inputs, distinct run identities, and each report’s evidence.

The intended contrast is 10 requested → 20 observed for the faulty reference, then 10 → 10 for the corrected reference. Every displayed result must come from its own execution. A reference’s name never determines its verdict.

Open the workbench

02 / THREE ASSUMPTIONS

Test the boundary that matters.

01

Transfer quantity

Does a displayed-unit intent produce the correct raw transfer after scaling changes? Compare requested quantity, executable quantity, observed deltas and residual.

02

Price timing

Do quantity and price describe compatible adjustment states at the moment they are used? A guard can correctly refuse an input while valuation remains inconclusive.

03

Event attribution

Does explicit event evidence justify classification? Distinguish structural rescaling from reinvestment and check that repeated delivery applies only once.

Independent evaluation

The reference proposes behavior. The runner observes actual runtime state. A separate evaluator checks the result. Missing evidence stays unknown.

Only registered, reviewed references and bounded presets are supported. No live feed, issuer authentication, public-chain transaction, cash distribution, tax result or unrestricted application audit is implied.

03 / READ THE EVIDENCE

A completed test can fail.

Completed · pass
Every required assertion has sufficient evidence and passes.
Completed · failing test
The execution finished, and independent observations contradict a required assertion.
Completed · inconclusive
A required value or attribution cannot be established. Unknown is never converted to zero or a pass.
Operational error
The service could not establish or deliver a validated result. A stored example is never substituted.

Three useful downloads

Report explains the findings and verdict. Evidence retains observed state and provenance. Rerun configuration records semantic inputs, without expected results, credentials or source code.

Repeat supported tests in the hosted workbench. CLI reproduction requires a separately authorized private source checkout; a configuration download does not include runner source. No repository access is needed to use the public workbench.

Hashes help check consistency and integrity; they are not independent certification. Atomic balances remain exact strings. Supported scales, precision and the measured Token-2022 executable identity are retained in each report; its release version may be unknown.

04 / SESSIONS & RECOVERY

Know what the service is doing.

One test at a time
Busy means no new test was queued. Wait for cleanup, then explicitly retry.
Lost connection
An accepted test may still be running. Reconnect to the same submission intent rather than starting a replacement.
Cancellation
The worker receives a cancellation request. Cleanup must finish before another test can start. Cancellation does not undo already observed changes.
Temporary results
Hosted reports and downloads belong to your guest session. Download them within 20 minutes of completion. Guest sessions last up to two hours.
Restart or expiration
Temporary recovery may become unavailable. Previous tests are never automatically replayed, and result links are not permanent public shares.

Page loads, hover, form changes, readiness checks and downloads do not submit tests. The animated observation model is an illustration, not a runtime status indicator.

Explore labeled synthetic examples