Environments · Behavior & Reliability

Verifiable-Reward Environments

Ready

Fourteen environments with deterministic, non-gameable rewards.

A verified batch of agent-behaviour environments, each with a deterministic verifier and a sealed holdout, banded honestly as a proven trainer or a saturated eval. Verified on the sealed holdout with cross-vendor real-model baselines.

Statement of Units · 8
unitbanding
  1. prompt_injection_defensetrainer

    Completes the task while resisting injection. Real-model curve.

  2. siege_monitortrainer

    A security monitor triaging failure scenarios. Real-model curve.

  3. contextual_integritytrainer

    Data reaches the right sink with zero leaks.

  4. sandbagging_resistancetrainer

    Does not hide capability when evaluated.

  5. tool_safetyeval

    Exact answer with no forbidden tool call.

  6. instruction_hierarchyeval

    Precedence-correct resolution of layered instructions.

  7. calibrated_abstentioneval

    Answers the answerable, abstains on the rest.

  8. and six moreeval

    spec_adherence, cot_faithfulness, long_horizon_agent, blackbox_oracle, encoded_channel, tool_temptation.

Request a scoped evaluation. Exclusive licensing available where wanted.

john@authensor.com

Owner of record: Authensor, Inc. (Delaware). Owned clean, deterministic verifiers, sealed holdouts, offline. Never open-sourced.