Dev claim · claim-1689 · #95 of 154
“Evidence is building that net cloud feedback is likely positive and unlikely to be strongly negative.”
Claim tags (sorted stems; highlighted when a passage shares them)
- build
- cloud
- evid
- feedback
- like
- neg
- net
- posit
- strongli
- unlik
The model’s verdict
The Transformer retrained from the notebook (the 2024 weights were never saved), shown three ways. The first is how the notebook evaluated it.
Notebook protocol
Dev claims predicted in batches of 16, in file order, from the 2024 retrieved evidence, as the notebook does.
- Supports
- 61.9%
- Refutes
- 31.3%
- Not enough info
- 0.2%
- Disputed
- 6.6%
One claim at a time
The same model run on a single claim, which is how the Try-it page runs it.
- Supports
- 53.3%
- Refutes
- 37.2%
- Not enough info
- 5.7%
- Disputed
- 3.8%
With gold evidence
Fed the human-annotated evidence instead of retrieved passages. This is the 'val accuracy' the training loop reports.
- Supports
- 62.3%
- Refutes
- 29.6%
- Not enough info
- 0.4%
- Disputed
- 7.7%
Retrieved vs gold evidence
Gold passages were picked by the dataset’s annotators. Retrieved passages are what the TF-IDF rule returned from all 1.19M passages. Switch between the submitted 2024 output and the re-runs.
2024 submission
The passages the team actually retrieved and submitted in 2024 (their saved output file). The file records only passage ids, so scores are recomputed with the submission rule.
- Path
- not recorded
- Found
- 0 of 1 gold
- P
- 0.00
- R
- 0.00
- F
- 0.00
#1evidence-133138 In electrolyte like CuSO4, NaCl, etc. there are positive and negative ions (Cu + +, SO4 --)
- like (shared with the claim)
- posit (shared with the claim)
- neg (shared with the claim)
- ion
- etc
- cos
- 0.837
- overlap
- 0.600
- score
- 1.437
- shared
- 3
Gold evidence (1)
evidence-95001 Other analyses have found that the iris effect is a positive feedback rather than the negative feedback proposed by Lindzen.
- lindzen
- propo
- posit (shared with the claim)
- neg (shared with the claim)
- anali
- effect
- iri
- feedback (shared with the claim)
- rather