MultiFC - Ansatte

MultiFC: A Real-World Multi-Domain Dataset for Evidence-Based Fact Checking of Claims

Publikation: Bidrag til bog/antologi/rapport › Konferencebidrag i proceedings › Forskning › fagfællebedømt

Dokumenter

OA-MultiFC
Forlagets udgivne version, 502 KB, PDF-dokument

Augenstein, Isabelle
Lioma, Christina
Dongsheng Wang
Lucas Chaves Lima
Casper Hansen
Christian Hansen
Simonsen, Jakob Grue

We contribute the largest publicly available dataset of naturally occurring factual claims for the purpose of automatic claim verification. It is collected from 26 fact checking websites in English, paired with textual sources and rich metadata, and labelled for veracity by human expert journalists. We present an in-depth analysis of the dataset, highlighting characteristics and challenges. Further, we present results for automatic veracity prediction, both with established baselines and with a novel method for joint ranking of evidence pages and predicting veracity that outperforms all baselines. Significant performance increases are achieved by encoding evidence, and by modelling metadata. Our best-performing model achieves a Macro F1 of 49.2%, showing that this is a challenging testbed for claim veracity prediction.

Originalsprog	Engelsk
Titel	Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP)
Forlag	Association for Computational Linguistics
Publikationsdato	2019
Sider	4684-4697
DOI	https://doi.org/10.18653/v1/D19-1475
Status	Udgivet - 2019
Begivenhed	2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP) - Hong Kong, Kina Varighed: 3 nov. 2019 → 7 nov. 2019

Konference

Konference	2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP)
Land	Kina
By	Hong Kong
Periode	03/11/2019 → 07/11/2019

Antal downloads er baseret på statistik fra Google Scholar og www.ku.dk

Ingen data tilgængelig

ID: 239563731

Datalogisk Institut

MultiFC: A Real-World Multi-Domain Dataset for Evidence-Based Fact Checking of Claims

Dokumenter

Konference

Antal downloads er baseret på statistik fra Google Scholar og www.ku.dk