Tech Meridian ← ENTITY INDEX
PROMY MERIDIAN RU

TOPIC · ENTITY #12390

retrieval-augmented evaluation

Related event timeline, sources and context from the news index.

EVENT TIMELINE

1

RESEARCH · 1 SOURCE · arXiv cs.AI

Build2SPARQL: large-scale text-to-SPARQL dataset for building knowledge graphs

The paper introduces Build2SPARQL, a benchmark produced by a KG-grounded pipeline that auto-generates 6,136 executable SPARQL queries and 30,680 natural-language questions over 201 building KGs (180 Brick, 21 ASHRAE 223P). Queries are produced and validated by graph-traversal code while LLMs only generate the natural-language paraphrases; human validation of 300 questions reports 98.8% semantic fidelity and naturalness and 84.0% operational plausibility, and retrieval-augmented evaluation boosted exact-match accuracy on three open-weight models from 0.2–20% (zero-shot) to 56–65% (three-shot retrieved).

6.0