QINLING v4 experiment (900 runs): 90 tasks × 5 language models × 2 hard-rule conditions

Published: 19 August 2026| Version 1 | DOI: 10.17632/3cp9rr96nh.1
Contributor:
志华

Description

This dataset is anonymized audit evidence for the QINLING v4 mechanism study, comprising 900 create-and-run observations from 90 frozen Chinese tasks tested across 5 models and 2 conditions (3 local and 2 hosted models), used to compare the effect of hard-rule activation on high-risk disposition boundaries. The data are organized at task and group levels, including observed terminal states, Governance state, Workflow W, strict E2E outcomes, integrity flags, paired-effect statistics, and hosted-model call audit metadata (with credentials removed). It is intended for manuscript review and methodological transparency, and does not represent field deployment results or human subject conclusions.

Files

Categories

Artificial Intelligence Applications, Unmanned Aerial Vehicle (Space Vehicle)

Funders

Licence