Research · 2025
UNH at CheckThat! 2025: Fine-tuning Vs Prompting in Claim Extraction
Extracting one succinct, checkable claim from a messy social post—and measuring the gap between benchmark scores and usefulness.
We compared fine-tuning and prompting setups for claim extraction. FLAN-T5 won on METEOR, while iterative self-refinement sometimes produced claims a fact-checker would actually want to work with. That gap is the whole plot: evaluation has to stay connected to the user and the work the system is meant to support.