Agentic framework for integration, end2end, multiuserinteraction, security testing based on usecases.
Find a file
custodian-sync 335c5cff82 chore(consistency): sync task status from DB [auto]
Updated by fix-consistency on 2026-09-29:
  - update .custodian-brief.md for test-driver

Assistant: claude-code
Assistant-Model: sonnet
Assistant-Process: 237582@bnt-lap001
Assistant-Session: f2b3d9f1-8fb9-4b9c-bc2b-837ec5dfc826
2026-09-29 21:43:47 +02:00
crystallized Preserve oracle semantics in generated regression judgments 2026-09-28 15:57:45 +02:00
docs Preserve oracle semantics in generated regression judgments 2026-09-28 15:57:45 +02:00
history Reject lossy observations and validate schedules before execution 2026-09-28 14:50:27 +02:00
lab Complete local generalisation tasks and document remaining experiment blockers 2026-09-28 12:06:24 +02:00
research Preserve oracle semantics in generated regression judgments 2026-09-28 15:57:45 +02:00
scenarios Complete local generalisation tasks and document remaining experiment blockers 2026-09-28 12:06:24 +02:00
src/testdriver Prevent passing run verdicts after failed scheduled realizations 2026-09-28 16:28:05 +02:00
tests Prevent passing run verdicts after failed scheduled realizations 2026-09-28 16:28:05 +02:00
tools Complete local generalisation tasks and document remaining experiment blockers 2026-09-28 12:06:24 +02:00
usecases Complete local generalisation tasks and document remaining experiment blockers 2026-09-28 12:06:24 +02:00
workplans Qualify wait tasks per Task Wait Qualifiers v0.1 (CUST-WP-0074-T05, long tail) 2026-09-29 21:25:00 +02:00
.custodian-brief.md chore(consistency): sync task status from DB [auto] 2026-09-29 21:43:47 +02:00
.gitignore Untrack __pycache__ and ignore Python build artifacts 2026-08-23 00:44:35 +02:00
.repo-classification.yaml Register with State Hub, persist concept assessment, seed first workplan 2026-08-22 22:40:39 +02:00
AGENTS.md AGENTS.md: fill unresolved {CREDENTIAL_ROUTING} placeholder with the credential-routing section 2026-09-22 08:34:20 +02:00
INTENT.md Complete local generalisation tasks and document remaining experiment blockers 2026-09-28 12:06:24 +02:00
pyproject.toml T09: crystallization 2026-08-23 00:21:48 +02:00
README.md Align scope with intent and add durable evidence and bounded variants 2026-09-28 14:31:22 +02:00
SCOPE.md Reject lossy observations and validate schedules before execution 2026-09-28 14:50:27 +02:00
WORK-RECORDS.md Align scope with intent and add durable evidence and bounded variants 2026-09-28 14:31:22 +02:00

test-driver

Agentic framework for integration, end-to-end, multi-user interaction and security testing, driven by use cases.

Tests mature alongside the software they protect: fluid and agentic while behaviour is changing, deterministic once it settles. See INTENT.md for the thesis and SCOPE.md for boundaries.

Status: research prototype. Deterministic claims, heuristic discovery, adaptation classification and crystallization run locally. Three synthetic multi-user use cases and a 38-mutation catalogue exercise the model. Live-model economics, browser-engine coverage and independent authoring-cost measurement remain blocked in workplans/TD-WP-0003-generalise-and-settle.md. See readiness and interface decisions. Opt-in local evidence receipts and explicit actor/argument variants are available; see usage and the scope/intent assessment. An independently owned real-system pilot remains blocked in TD-WP-0004.

Run

python3 -m pytest -q          # the whole suite
python3 -m pytest -q tests/test_reference_scenario.py

No third-party dependencies. Python ≥ 3.11, pytest for the suite.

Layout

src/testdriver/   the kernel — intent, world, actions, drivers,
                  observers, oracles, evidence, runner
usecases/         durable test intent, including not-yet-runnable use cases
lab/              the system under test
scenarios/        reference scenarios
research/         hypotheses, experiments, findings, fitness map
docs/             concept model, improvement loop, milestones, design notes
history/          assessments and completed-work write-ups

Reading order

  1. INTENT.md — the thesis
  2. docs/TestDriverConceptModel.md — canonical concept set
  3. docs/TestDriverClassificationDesign.md — why adaptation cannot normalize a defect, and where model judgment is and is not permitted
  4. docs/TestDriverInitialMilestones.md — canonical milestones M0–M10