NL2SHACL-Bench: A Benchmark Suite for Natural Language to SHACL Translation

작성자

카테고리:

← 피드로
arXiv cs.AI · Yuchen Zhou, Niels Bobet, Maribel Acosta · 2026-08-11 AI

[Submitted on 24 Jul 2026]

View PDF HTML (experimental)

Abstract:SHACL is a core technology for validating the conformance of RDF knowledge graphs (KGs). Yet, authoring SHACL shapes requires technical expertise that most domain experts lack. Translating natural language requirements into SHACL (NL2SHACL) would lower this barrier. However, there is no dedicated benchmark for NL2SHACL, and evaluating generated shapes requires methods beyond string comparison, as semantically equivalent shapes can differ in serialisation and structure. To tackle these challenges, we present NL2SHACL-Bench, a benchmark suite for natural language to SHACL translation. Using NL2SHACL-Bench, we evaluate four state-of-the-art large language models (LLMs) for this task. Our results show that current LLMs are highly capable of generating syntactically valid SHACL, but still struggle to produce semantically equivalent constraints for complex logical and structural patterns. This indicates that NL2SHACL-Bench provides a meaningful basis for measuring advances in the NL2SHACL state of the art.

Submission history

From: Yuchen Zhou [view email]
[v1] Fri, 24 Jul 2026 10:25:35 UTC (916 KB)

원문에서 계속 ↗

추출 본문 · 출처: arxiv.org · https://arxiv.org/abs/2608.07530

코멘트

답글 남기기