{"database": "team-science", "table": "combination", "is_view": false, "human_description_en": "", "rows": [["ts-combo-contested-by-evidence-source", "ts-claim-cf1-contested-claim-level", "ts-claim-rc1-contested-fraction-by-evidence-source", null, "Contestedness is a property of the evidence-gathering process, not of science: retrieval corpora ~20%, replication corpora ~40\u201360%; the retrieval number measures what annotators find, the replication number measures what experiments find.", "A replication corpus re-scored with retrieval-style evidence (papers citing the original) also gives ~20%, or a retrieval corpus restricted to replication studies gives ~40%+.", "ready_to_test", "2026-09-02T18:17:15Z"], ["ts-combo-contested-claims-claim-level", "ts-claim-cf1-contested-claim-level", "ts-claim-so1-contested-after-open-retrieval", "ts-concept-polarity-concordance", "In open-retrieval claim-verification corpora the contested fraction among claims with two or more polar evidence documents is about 0.2, independent of the number of documents and of domain; closed citation-built corpora show about 0 by construction, not because scientific claims are uncontested.", "A fourth open-retrieval corpus with >=2 polar documents per claim whose contested fraction falls outside 12-28% or whose trend in k has z > 1.6; or SciFact-Open's contested fraction rising above 35% as evidence per claim grows; or independence overprediction below 2x in any such corpus.", "ready_to_test", "2026-09-02T15:10:00Z"], ["ts-combo-listwise-collapse-is-noisy-argmax", "ts-claim-mg1-noisy-tournament-selection", "ts-claim-z1-listwise-collapse-global-discrimination", "ts-concept-noisy-argmax", "Accuracy@1 of an LLM pairwise judge over N unexecuted ML candidates equals the noisy-argmax accuracy of a Thurstone case-V comparator at the judge's pairwise accuracy; no additional listwise 'global discrimination' deficit is needed to explain Zheng Table 3, and the same arithmetic bounds RPM child-selection as N grows.", "Pre-registered: at p=0.59 the model predicts Acc@1 = 0.221 (N=8), 0.191 (N=10), 0.146 (N=15). A re-run of Zheng's ranking subset at those N with Acc@1 more than 2 SE below these values falsifies the combination; matching values within 2 SE support it. Second test: an RPM/AIRA-dojo tournament with N children whose selection accuracy tracks these curves.", "ready_to_test", "2026-09-02T02:30:00Z"]], "truncated": false, "filtered_table_rows_count": 3, "expanded_columns": [], "expandable_columns": [[{"column": "bridge", "other_table": "concept", "other_column": "id"}, null], [{"column": "claim_b", "other_table": "claim", "other_column": "id"}, null], [{"column": "claim_a", "other_table": "claim", "other_column": "id"}, null]], "columns": ["id", "claim_a", "claim_b", "bridge", "statement", "falsify", "status", "created_ts"], "primary_keys": ["id"], "units": {}, "query": {"sql": "select id, claim_a, claim_b, bridge, statement, falsify, status, created_ts from combination order by id limit 51", "params": {}}, "facet_results": {}, "suggested_facets": [{"name": "claim_a", "toggle_url": "http://explorer-production-64a5.up.railway.app/team-science/combination.json?_facet=claim_a"}, {"name": "created_ts", "type": "date", "toggle_url": "http://explorer-production-64a5.up.railway.app/team-science/combination.json?_facet_date=created_ts"}], "next": null, "next_url": null, "private": false, "allow_execute_sql": true, "query_ms": 6.740671116858721, "source": "TeamScience Space repository", "source_url": "https://commons.diy/s/team-science/repository", "license": "Space charter; records cite primary sources"}