VIBench Results VIBench Results contains the retained raw generations, detector labeled outputs, complete runs, option order and runtime ablations, paper facing summaries, figures, configs, audit files, and static explorer indices used in the study. The main paper evaluation covers 13 models, 15,600 direct generations, and 2,000 agentic runs. Direct and agentic VIB are reported as scenario matched, share normalized differences in affiliated ecosystem selection relative to strict non affiliated control models. Paper figures The paper facing figures are included under docs/figures/paper/ . Viewer configs The Hugging Face viewer exposes compact indices over the larger result tree rather than trying to render the raw artifacts directly. runs ( viewer/runs.jsonl , 13 rows): one row per retained run root, with run id , run category , and direct/agentic presence flags. all records ( viewer/all records.jsonl , 24698 rows): one row per detector scored unit. Direct rows correspond to individual labeled outputs; agentic rows correspond to labeled session summaries. Category configs such as direct complete , direct ablations , opencode complete , agents sdk complete , opencode ablations , and…
Runs entirely in your browser via DuckDB-Wasm — this dataset's real data file is loaded once, then queried locally. Nothing is sent to a server.
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy