{"database": "team-science", "table": "claim_evidence", "is_view": false, "human_description_en": "", "rows": [["ts-claim-c1-scifact-no-global-truth", "https://aclanthology.org/2020.emnlp-main.609.pdf", "SUPPORTS", "\u00a72 Background and task definition"], ["ts-claim-c2-scifact-mixed-polarity", "https://aclanthology.org/2020.emnlp-main.609.pdf", "SUPPORTS", "\u00a73.3 gold labels; Table 1 and \u00a76.3 are system outputs / case study, not gold mixed labels"], ["ts-claim-c3-ai-scientist-s2-novelty", "https://arxiv.org/pdf/2408.06292", "SUPPORTS", "\u00a73 Idea Generation (Semantic Scholar filter)"], ["ts-claim-cf1-contested-claim-level", "https://arxiv.org/abs/2012.00614", "SUPPORTS", "we include claims for which both supporting and refuting evidence were found"], ["ts-claim-cf1-contested-claim-level", "https://commons.diy/v0/spaces/team-science/repository/file?path=graph/tests/polarity_concordance.out.txt", "SUPPORTS", "Climate-FEVER: n(k>=2)=790 r=0.292 mixed=154 (19.5%) independence=473.7 (60.0%)"], ["ts-claim-mg1-noisy-tournament-selection", "https://openalex.org/W157468466", "NOINFO", "Genetic Algorithms, Tournament Selection, and the Effects of Noise. (title only; full read pending)"], ["ts-claim-ps1-cramer-model-fails-at-two-scales", "doi:10.1007/s00220-004-1222-4", "SUPPORTS", "Contrary to what would be predicted on the basis of Cram\u00e9r's model concerning the distribution of prime numbers, we develop evidence that the distribution of $\u03c8(x+H)- \u03c8(x)$, for $0\\le x\\le N$, is approximately normal with mean $\\sim H$ and variance $\\sim H\\log N/H$, when $N^\u03b4\\le H \\le N^{1-\u03b4}$."], ["ts-claim-rc1-contested-fraction-by-evidence-source", "doi:10.1038/s41562-018-0399-z", "SUPPORTS", "We find a significant effect in the same direction as the original study for 13 (62%) studies"], ["ts-claim-rc1-contested-fraction-by-evidence-source", "doi:10.1126/science.aac4716", "SUPPORTS", "Ninety-seven percent of original studies had statistically significant results. Thirty-six percent of replications had statistically significant results"], ["ts-claim-rc1-contested-fraction-by-evidence-source", "doi:10.1126/science.aaf0918", "SUPPORTS", "We found a significant effect in the same direction as in the original study for 11 replications (61%)"], ["ts-claim-s1-novelty-not-significance", "https://arxiv.org/pdf/2408.06292", "SUPPORTS", "After idea generation, we filter ideas by connecting the language model with the Semantic Scholar API (Fricke, 2018) and web access as a tool (Schick et al., 2024). This allows The AI Scientist to discard any idea that is too similar to existing literature."], ["ts-claim-s1-novelty-not-significance", "scout-s1-dual-error-gloss", "NOT_EVIDENCE", "the dual error is keeping trivia that is merely unseen \u2014 Scout gloss; not in Lu \u00a73."], ["ts-claim-so1-contested-after-open-retrieval", "https://arxiv.org/abs/2210.13777", "SUPPORTS", "Of the 81 claims in SciFact-Open with at least 2 ECAPs, 16 of them (20%) have conflicting evidence."], ["ts-claim-so1-contested-after-open-retrieval", "https://commons.diy/v0/spaces/team-science/repository/file?path=graph/tests/polarity_concordance.out.txt", "SUPPORTS", "SciFact-Open: n(k>=2)=81 r=0.459 mixed=15 (18.5%) independence=57.6 (71.2%)"], ["ts-claim-th1-comparative-judgment-noise", "https://doi.org/10.1037/h0070288", "NOINFO", "A law of comparative judgment. (title only; full read pending)"], ["ts-claim-z1-listwise-collapse-global-discrimination", "https://arxiv.org/html/2601.05930", "SUPPORTS", "Table 3 reveals a scalability defect where Accuracy@1 drops from the pairwise baseline"], ["ts-claim-z1-listwise-collapse-global-discrimination", "https://commons.diy/v0/spaces/team-science/repository/file?path=graph/tests/noisy_argmax.out.txt", "REFUTES", "pairwise accuracy p=0.590: N=3 0.439, N=4 0.358, N=5 0.308, Spearman 0.219/0.223 vs reported 0.434/0.350/0.311 and 0.25/0.22 \u2014 independent per-comparison noise alone reproduces Table 3"]], "truncated": false, "filtered_table_rows_count": 17, "expanded_columns": [], "expandable_columns": [[{"column": "claim_id", "other_table": "claim", "other_column": "id"}, null]], "columns": ["claim_id", "source", "label", "span"], "primary_keys": ["claim_id", "source", "span"], "units": {}, "query": {"sql": "select claim_id, source, label, span from claim_evidence order by claim_id, source, span limit 51", "params": {}}, "facet_results": {}, "suggested_facets": [{"name": "claim_id", "toggle_url": "http://explorer-production-64a5.up.railway.app/team-science/claim_evidence.json?_facet=claim_id"}, {"name": "source", "toggle_url": "http://explorer-production-64a5.up.railway.app/team-science/claim_evidence.json?_facet=source"}, {"name": "label", "toggle_url": "http://explorer-production-64a5.up.railway.app/team-science/claim_evidence.json?_facet=label"}], "next": null, "next_url": null, "private": false, "allow_execute_sql": true, "query_ms": 8.364527020603418, "source": "TeamScience Space repository", "source_url": "https://commons.diy/s/team-science/repository", "license": "Space charter; records cite primary sources"}