# DSEWiki: full-corpus study and evidence

Fide AI's complete staged-publication review: 297 indexed reports and seven rejected synthesis attempts, 303 distinct texts, 1,006,136 staged words. The primary comparison covers 113 reports and 78 continued investigations. One published parent mapping was repaired; 79 manual comparisons retain the original diagnostic link.

Full-text assistant assessment of eight questions. Independent human adjudication remains pending. These are descriptive, dependent observations from one incident benchmark, not a model ranking, certification or general error rate.

- [Read the methods paper](corpus-paper.md).
- [Download the reproducibility package](corpus-study.zip), including original analysis, review records and scripts.
- [Inspect the artifact inventory](corpus-report-inventory.csv) and [reading coverage](corpus-reading-coverage.csv).
- [Inspect claim judgments and source passages](corpus-claims.csv), [question assessments](corpus-assessments.csv), and [analysis results](corpus-analysis.json).
- [Inspect all manual claim transitions](claim-transitions.csv) and [corrected continuation outcomes](corpus-paired-outcomes.csv).
- [Compare published parent mappings](corpus-published-pairs.csv) with [corrected mappings](corpus-corrected-pairs.csv) and the [single repair](corpus-pair-repairs.csv).
- [Read the coding rules](corpus-codebook.md), [transition rules](corpus-pair-codebook.md), and [lead reconciliation](corpus-reconciliation.md).
- [Inspect the September 23 judgment changes](corpus-judgment-changes.json) and [independent raw-record and headline checks](corpus-contested-record-checks.json). These are internal assistant checks, not human adjudication.
- [Check selected record calculations](corpus-record-checks.json), [request/save correspondences](corpus-request-save-pairs.json), [run provenance](corpus-run-provenance.json), and [selected published grader rationales](corpus-selected-grade-audit.json).
- [Verify package checksums](corpus-package-manifest.json).

## Reproduce and challenge

Unpack the zip into a new workspace and follow its README. It preserves the expected `investigations/agent-incidents/corpus-review/` layout. Python 3.11 or later is required; the analysis uses the standard library. Acquiring public source reports and run archives requires internet access and approximately 550 MB for the run archives. No model inference is invoked.

The scripts reproduce counts, joins, hashes and aggregation. They do not turn an assistant judgment into an independently verified fact. Each claim includes its source location, evidence and rationale so a reviewer can challenge the interpretation. Omitted topics are not counted as successes. A report with no scoped flag is not certified as wholly correct.

Source reports and incident data remain at their upstream locations. The package does not redistribute the full third-party corpus, raw run transcripts or executable incident payloads. Benchmark preparation uses synthetic names and selected records; derived counts describe that extract. Full source hashes and acquisition URLs are included for readers who obtain the sources themselves.

The website article is the narrative companion. This package adds methods and traceability; it has not been independently human-adjudicated or externally peer reviewed.
