Abstrakt
The proliferation of sophisticated misinformation threatens societal trust, yet most AI detectors operate as opaque’black boxes,’ lacking the verifiable reasoning essential for human oversight and adoption in high-stakes domains. This critical transparency gap demands a new paradigm where interpretability is a core design principle, not a post-hoc feature. In this work, we propose the Explainable Fake News Detection (XFND) framework, a human-centered pipeline that marries expert-guided feature space validation with evidence-anchored explanation synthesis using large language models. Our approach demonstrably improves feature space separability before training, increasing the silhouette score on public datasets like LIAR by up to 63% (from 0.19 to 0.31). On established benchmarks, the resulting system achieves competitive classification performance, reaching a macro-F1-Score of 0.792 on a binary version of LIAR and 0.731 on PolitiFact, while ensuring outputs are well-calibrated and auditable. We conclude that proactively designing for interpretability enables systems that are both highly accurate and trustworthy by design, establishing a new standard for collaborative AI in the fight against disinformation.
| Język oryginału | angielski |
|---|---|
| Strony (od–do) | 168-182 |
| Liczba stron | 15 |
| Czasopismo | CEUR Workshop Proceedings |
| Tom | 4141 |
| Status publikacji | Opublikowano - 2025 |
| Wydarzenie | 1st Workshop on Advanced AI in Explainability and Ethics for the Sustainable Development Goals, ExplAI-2025 - Khmelnytskyi, Ukraina Czas trwania: 7 lis 2025 → 7 lis 2025 |
Obszary tematyczne ASJC Scopus
- Informatyka ogólna
Fingerprint
Zanurz się w tematy badawcze publikacji „Verifiable by construction: evidence-anchored LLMs for explainable fake news detection”. Razem tworzą niepowtarzalny odcisk palca.Cytowanie
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver