Extend the Agent standard matrix with a report-audit scenario so model input, output, handoff, and secret-safety evidence are tested instead of remaining narrative-only.
Constraint: The user requested another test pass and expanded Agent/swarm testing scenarios under docs/.
Rejected: Treating the model I/O report as untested documentation | it would leave the handoff and input/output evidence unguarded.
Confidence: high
Scope-risk: moderate
Directive: Keep model I/O reports under docs/ and redact secret-shaped values during export.
Tested: .venv/bin/python -u -B examples/run_standard_scenario_acceptance.py; .venv/bin/python -B -m unittest discover -s tests; .venv/bin/python -B -m py_compile swarm_minimal/*.py examples/*.py tests/*.py; .venv/bin/python -u -B examples/run_academic_standard_evaluation.py; git diff --check; docs secret-pattern scan.
Not-tested: Large-scale concurrent 3/5/7 worker load and external browser rendering were not run.
Co-authored-by: OmX <omx@oh-my-codex.dev>
Refresh the model/Agnet I/O report from a new S01-S07 standard matrix run and add the scenario coverage summary directly to the report.
Constraint: The linked Gitee docs page should show current evidence and the user asked to supplement the tested scenarios.
Rejected: Leaving the previous live run as latest | it would make the linked report stale after rerunning S07.
Confidence: high
Scope-risk: narrow
Directive: Keep regenerated model I/O reports under docs/ and strip trailing whitespace from model-produced multiline output.
Tested: python -u -B examples/run_standard_scenario_acceptance.py; git diff --check; python -B -m py_compile swarm_minimal/*.py examples/*.py tests/*.py; python -B -m unittest discover -s tests; python -u -B examples/run_academic_standard_evaluation.py
Not-tested: Browser rendering of the internal Gitee page was not verified because the URL is internal; git push updates the same origin/main path.
Co-authored-by: OmX <omx@oh-my-codex.dev>
Define Agent and swarm-specific acceptance evidence, move the reports under docs, and make the homepage point to the current standard, live run, model I/O, and handoff evidence.
Constraint: Agent quality standards are configured from industry AI and agent risk references because there is no single accepted swarm-Agent certification standard.
Rejected: Treating py_compile or unittest as the primary quality standard | they are evidence collection tools, not the Agent quality standard itself.
Confidence: high
Scope-risk: moderate
Directive: Keep future standard reports under docs/ and keep secrets in ignored local .env files only.
Tested: git diff --cached --check; python -B -m py_compile swarm_minimal/*.py examples/*.py tests/*.py; python -B -m unittest discover -s tests; python -u -B examples/run_academic_standard_evaluation.py
Not-tested: Did not rerun the full live Azure/NewAPI S07 scenario after moving docs; previous live run 3e8e58ae4e084bc8b90cf5c46f8992f3 passed before the docs relocation.
Co-authored-by: OmX <omx@oh-my-codex.dev>