Augustus v0.7.0

Released 2026-09-22. This minor release equips agents to find, build, evaluate, and improve decision-model systems, with new paired-outcome tooling and a more explicit composition calculus. It adds no provider integration or model and makes no claim of measured model superiority. TypeSafe Jev remains the default hosted exemplar; exact rules, classical models, and human processes remain valid alternatives.

The important shift

Augustus is a skill and working method for agents to discover useful decision models, implement compositions, build evals, and hill-climb decision-driven or Software 3.0 systems. Research is input to that work, not its finished output. The goal is to derive and test the best-supported method for the actual task, not reproduce a consensus survey or chase a fashionable benchmark.

That capability is now named directly in the skill’s activation description, Codex UI, Claude marketplace, README, and website. Discoverability should match what users want to accomplish: build and improve working systems with outcomes and explicit constraints—not merely read about model placement.

What changed

Upgrade notes

Update through your existing installation method; avoid duplicate installations. The skill needs no API key. No provider benchmark was reproduced; the bundled offline helpers make no model calls. Model-based code review is separate.

The existing offline helpers now reject duplicate JSON keys, conflicting known repository node IDs, and invalid schema-version types. Oversized numbers, decoder nesting failures, and interrupted HTTP reads produce explicit input or receipt errors. Large finite policy costs are averaged without multiplying counts into an avoidable overflow. Valid-input interfaces remain unchanged; partial baselines still carry no comparative claim.

Run make check from a source checkout. Runtime references must stay within the installed skill, including nested cards; budgets, links, and SemVer are checked. Rendered-site validation checks asset existence, not image decoding or CSS URLs. Use the explicit GitHub Pages gem activation in CONTRIBUTING for a matched build.

Evidence and limits

Read the refresh and source dispositions, replacement ledger and integration evidence, adversarial findings and corrections, discoverability audit, existing-listing update queue, and maintainer guidance. These dated artifacts preserve what was verified at each stage; final publication and deployment are recorded on the release PR. Requested model routing is not independent backend attestation. Public projections are not private benchmark replay; no source result was upgraded to Reproduced without a local run.

Earlier release: v0.6.0.