ClickGuard: Detecting and Spoiling Clickbait News with Informativeness Measures and Large Language Models

작성자

카테고리:

← 피드로
arXiv cs.AI · Wojciech Michaluk, Tymoteusz Urban, Mateusz Kubita, Soveatin Kuntur, Anna Wr'oblewska · 2026-07-24 AI

[Submitted on 18 May 2026]

View PDF HTML (experimental)

Abstract:This paper presents an AI-driven browser extension that identifies clickbait to help users avoid misleading Internet articles. Moving beyond traditional detection, the application employs a hybrid machine learning architecture that combines transformer-based embeddings with linguistically motivated features and a custom “baitness” score. After evaluating various natural language processing techniques — from classic vectorizers to large language model (LLM) embeddings — an XGBoost-based model was developed that achieves an F1-score of 91% on the open combined dataset. Most importantly, the tool can warn users before and after they access a clickbait article. After opening an article, the user receives a percentage score indicating the likelihood that it is clickbait. The prediction is explained based on the analyzed metrics, including those specifically developed within the proposed system. The browser extension also provides a clickbait spoiler — a one- to two-sentence summary of the entire article. Demo video:this https URL}{this https URL

Submission history

From: Soveatin Kuntur [view email]
[v1] Mon, 18 May 2026 10:12:47 UTC (511 KB)

원문에서 계속 ↗

추출 본문 · 출처: arxiv.org · https://arxiv.org/abs/2607.20463

코멘트

답글 남기기

이메일 주소는 공개되지 않습니다. 필수 필드는 *로 표시됩니다