Python notebook of Princeton GraphMERT Paper – a better knowledge graph

https://github.com/creativeautomaton/graphMERT-python

2•7jewve5rws•2h ago

Comments

7jewve5rws•2h ago

For those interested in a new knowledge graph model, I implemented the Princeton work on GraphMERT into a python notebook for experimentation.

abstract: The notebook uses a Romeo and Juliet corpus text that is embedded with a sentencetransformers model then trained with to build the GraphMERT model which is used to build the knowledge graph in a GraphRAG inference setup.

For a detailed review of the GraphMERT paper watch this great european youtuber: discover ai - https://www.youtube.com/watch?v=xh6R2WR49yM&t=1s

About GraphMERT:

GraphMERT: Efficient and Scalable Distillation of Reliable Knowledge Graphs from Unstructured Data Margarita Belova, Jiaxin Xiao, Shikhar Tuli, Niraj K. Jha

    Researchers have pursued neurosymbolic artificial intelligence (AI) applications for nearly three decades because symbolic components provide abstraction while neural components provide generalization. Thus, a marriage of the two components can lead to rapid advancements in AI. Yet, the field has not realized this promise since most neurosymbolic AI frameworks fail to scale. In addition, the implicit representations and approximate reasoning of neural approaches limit interpretability and trust. Knowledge graphs (KGs), a gold-standard representation of explicit semantic knowledge, can address the symbolic side. However, automatically deriving reliable KGs from text corpora has remained an open problem. We address these challenges by introducing GraphMERT, a tiny graphical encoder-only model that distills high-quality KGs from unstructured text corpora and its own internal representations. GraphMERT and its equivalent KG form a modular neurosymbolic stack: neural learning of abstractions; symbolic KGs for verifiable reasoning. GraphMERT + KG is the first efficient and scalable neurosymbolic model to achieve state-of-the-art benchmark accuracy along with superior symbolic representations relative to baselines.
    Concretely, we target reliable domain-specific KGs that are both (1) factual (with provenance) and (2) valid (ontology-consistent relations with domain-appropriate semantics). When a large language model (LLM), e.g., Qwen3-32B, generates domain-specific KGs, it falls short on reliability due to prompt sensitivity, shallow domain expertise, and hallucinated relations. On text obtained from PubMed papers on diabetes, our 80M-parameter GraphMERT yields a KG with a 69.8% FActScore; a 32B-parameter baseline LLM yields a KG that achieves only 40.2% FActScore. The GraphMERT KG also attains a higher ValidityScore of 68.8%, versus 43.0% for the LLM baseline.

Nuclear fusion, the 'holy grail' of power

Will the explainer post go extinct?

Where are we on XChat security?

iOS 26.1 beta transparency toggle changes liquid glass

The GUI S-curve is peaking

The Cost of Cloud, a Trillion Dollar Paradox (2021)

Kohler's Dekoda Toilet Camera

China Went from Clean Energy Copycat to Global Innovator

NORAD's Cheyenne Mountain Combat Center, C.1966

Start an AI PhD Now

US NSA alleged to have launched a cyber attack on Chinese timekeeping agency

We Built WebSocket Servers for Vercel Functions

BlackRock Says Insurers Expect to Keep Ramping Up Private Bets

It was a weather balloon, not space debris, that struck a United Airlines plane

It was DNS

The breach that broke the internet: The untold story of Log4Shell

I Could Have Lived Without AI

U.S. Banks Are Hunting for Collateral to Back $20B Argentina Bailout

Free Seedream 4.0 – No Login Required

Sam Altman got Silicon Valley's giants to tether their fates to his company

IKEA Phone Bed

Ask HN: What books are you reading now?

Ask HN: What software dev tasks have you found LLMs to be good at versus bad at?

Proposed DNS RFC 8767: Serving Stale Data to Improve DNS Resiliency (2020)

Show HN: WatchDoggo – simple open-source service status monitor

Why 'Functor' Doesn't Matter (2019)

Sonoluminescence

Fundraiser with Safe Using Stripe Atlas

NobelBiz – Erlang/OTP and Elassandra/Cassandra|Full-Time|Remote|80K-100K USD

The Great Crown Caper – Two crowns, one crime, one unsolved mystery