Wikidata
PulseAugur coverage of Wikidata — every cluster mentioning Wikidata across labs, papers, and developer communities, ranked by signal.
6 day(s) with sentiment data
-
New system WiCleanData enhances Wikidata's consistency and accuracy
Researchers have developed WiCleanData, a system designed to improve the consistency and accuracy of Wikidata. This automated pipeline addresses issues like redundant classes, instance-vs-class ambiguity, incorrect taxo…
-
New SWORD benchmark reveals cross-lingual factual inconsistencies in LLMs
A new benchmark called SWORD has been developed to evaluate Large Language Models' (LLMs) ability to reject factual errors across different languages. SWORD uses distortions derived from Wikidata to create factually inc…
-
New framework improves multilingual entity linking for rare entities
Researchers have developed a new framework to improve multilingual entity linking, particularly for rare entities. The system uses a reasoning-capable vision-language model that dynamically searches and reasons over Wik…
-
New research proposes improved evaluation for continual knowledge updating in LLMs
A new research paper on arXiv proposes a more robust method for evaluating continual knowledge updating in language models. The study highlights that traditional evaluations, which often rely on a single final checkpoin…
-
LLM overconfidence linked to knowledge popularity, new study finds
A new research paper explores the phenomenon of large language models (LLMs) exhibiting high confidence in incorrect answers, a problem termed overconfidence. The study, focusing on knowledge popularity, found that hall…
-
New AI tool detects temporal disinformation in political news
Researchers have developed a Temporal Coherence Score (TCS) to identify temporal inconsistencies in political news, a form of disinformation that bypasses traditional fake news detectors. The TCS system uses a four-stag…
-
Keenable AI open-sources NEEDLE, a dynamic web search benchmark
Keenable AI has open-sourced NEEDLE, a new benchmark designed to evaluate web search APIs by dynamically generating query sets hourly and daily. This approach prevents agents from accessing pre-existing answers, ensurin…
-
Wikidata powers AI reasoning beyond LLM limitations
Structured knowledge graphs, exemplified by Wikidata, are essential for advancing AI beyond probabilistic guessing towards true machine reasoning. These graphs provide a layer of factual grounding that pure Large Langua…
-
LLM pipeline transforms heritage records into FAIR knowledge graphs
Researchers have developed a pipeline to transform flat cultural heritage records into a FAIR Digital Object (FDO) compliant knowledge graph using CIDOC-CRM. This system employs a large language model to distinguish bet…
-
New protocol generates plausible unknown names for LLM evaluation
Researchers have developed a protocol called PUN (Plausible Unknown Names) to create and validate person names that are not easily identifiable online. This method combines components from Wikidata, web-based LLM screen…
-
Vision-Language Models Show Temporal Knowledge in Art Dating, But Biases Remain
Researchers have developed a method for dating artworks using vision-language models (VLMs), addressing the issue of "temporal entanglement" where models appear to encode historical time but actually reflect institution…
-
LLM outputs transformed into interactive historical maps
A developer has created a client-side WebGL application that transforms unstructured LLM outputs into spatial narrative visualizations, effectively turning any large language model into a historical map generator. This …
-
Scholarly networks traced back 900 years using Fields Medalists
Researchers have developed a method to reconstruct scholarly mentor-student networks spanning approximately nine centuries, utilizing data from Wikidata that aggregates the Mathematics Genealogy Project and the MacTutor…
-
New framework infuses semantic knowledge into traffic forecasting models
Researchers have developed a new framework for spatio-temporal traffic forecasting that enhances Graph Neural Networks (GNNs) by integrating external semantic knowledge. This approach uses general-purpose knowledge grap…
-
New AdaPop method improves LLM unlearning by prioritizing popular facts
Researchers have developed a new method called AdaPop (Adaptive Popularity) to improve the unlearning process in large language models (LLMs). Unlike existing methods that apply uniform pressure to remove data, AdaPop c…
-
New AdaPop method improves LLM unlearning by adapting to fact popularity
Researchers have developed a new method called AdaPop to improve the process of unlearning information from large language models (LLMs). Unlike previous methods that applied uniform pressure to remove data, AdaPop adju…
-
New benchmark ENTLORE tests latent organizational reasoning in enterprise QA · 6 sources tracked
Researchers have introduced ENTLORE, a new benchmark designed to evaluate latent organizational reasoning in enterprise question answering systems. This framework reconstructs enterprise structures from documents and or…
-
Canonical IDs are crucial for knowledge graphs to prevent data forks
This article discusses the importance of canonical IDs for knowledge graphs to prevent data forks and inconsistencies. It proposes using prefixed ULIDs or UUIDv7 for identifiers, which are time-sortable and self-describ…
-
Machine learning models detect user deaths on social media
A new dissertation details the development of machine learning classifiers capable of automatically detecting deceased users on social networking sites. The research utilized a new dataset compiled from Wikidata and X (…
-
New benchmark UNLINK-VL evaluates cross-modal knowledge unlearning in VLMs
Researchers have introduced UNLINK-VL, a new benchmark designed to evaluate how effectively knowledge can be removed from vision-language models (VLMs). The benchmark focuses on the transfer of unlearning across differe…