truereward-harwell
The recorded-state reader for truereward, over the Harwell engine.
It grades a change to a codebase that has no unit test suite. A case records what a job leaves behind when the code is right, and claims the parts of it the case is about: data sets, spool files and tables. Grading runs the same job against the changed codebase and reads those claims against the approved baseline.
- A case is a starting state, a run and claims, and nothing unclaimed is compared. A case passes when every claim of it holds, fails when one does not, and is ungradable under a declared code when one could not be evaluated. A claim that names an observable and reads no field of it asserts that observable whole, byte for byte; an observable no claim names is not read at all, so a correct change is never failed on something nobody was asking about.
- A case is data. A directory with
case.json, astart/snapshot and anexpected/snapshot, validated against the JSON Schema shipped in the package. Every field a case reads is resolved to bytes when the case is recorded, so grading never opens the candidate's source to decide what is compared. grade_case(case, codebase)runs the case twice in a contained run, reads the claims over the first end state and the baseline, and answers with atruereward.CaseResult. The two runs are compared with each other, over the same claimed observables, under the author's own mask, which answers the determinism check.record.record(...)records a case: refuse a case that names no claim, run the job twice, refuse a run that did not complete or that varied in something the case claims, keep what it left as the approved baseline, freeze every field map, and read every claim back over the baseline it just approved.selftest.run(...)asks whether a case set would catch a bug: the reference change must score 1.0, the unchanged codebase 0.0, and every seeded defect less than 1.0.
A task's test file:
from functools import partial
from pathlib import Path
import pytest
from truereward_harwell import cases, grade_case
CASES = Path(__file__).parent / "cases"
@pytest.mark.parametrize("case", cases(CASES), ids=lambda c: c.id)
def test_case(case, grade):
grade(case, partial(grade_case, codebase="/grade"))
Run it with pytest -p truereward_harwell, which brings the truereward
plugin with it.
Licensed under Apache-2.0.
Release files for truereward-harwell 0.3.2
For a detailed explanation of source distributions (sdists) and built distributions (wheels), please see the package formats documentation.
Source distribution (sdist)
| File | Size | Uploaded | |
|---|---|---|---|
| truereward_harwell-0.3.2.tar.gz | 42.3 kB | Details |
Built distribution (wheel)
| File | Interpreter | ABI | Platform | Reset |
|---|---|---|---|---|
| truereward_harwell-0.3.2-py3-none-any.whl | Python 3 | none | any | Details |
Total release size: 89.6 kB
Release files / truereward_harwell-0.3.2.tar.gz
| Download URL | truereward_harwell-0.3.2.tar.gz |
|---|---|
| Size | 42.3 kB |
| Tags | Source |
|
SHA-256 checksum How to use checksums |
c02ae1f1ae2b0b9c5048080adb5909be5a2ca11c65b71bbb45232096635c1a00
|
|
BLAKE2b-256 checksum How to use checksums |
4ab15d4be854fa808870f9cc0a970f8d55aaef9c1884e72e42ee24fe58c64a8b
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
Yes |
| Uploaded via |
twine/7.0.0 CPython/3.13.14
|
Release files / truereward_harwell-0.3.2-py3-none-any.whl
| Download URL | truereward_harwell-0.3.2-py3-none-any.whl |
|---|---|
| Size | 47.3 kB |
| Tags | Python 3 |
|
SHA-256 checksum How to use checksums |
ca65b48be7402f6effa26b4d2a908ba54bb692364a42ea0cb664179ccb748a0a
|
|
BLAKE2b-256 checksum How to use checksums |
fc58198435f89eba5c825d056b487a45daba4f3fcd727b4f03c01ca80e8de946
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
Yes |
| Uploaded via |
twine/7.0.0 CPython/3.13.14
|