Extracting knowledge graphs from language models as a diagnostic benchmark of model performance.

Last update: Oct 25, 2022

Overview

Interpreting Language Models Through Knowledge Graph Extraction

Idea: How do we interpret what a language model learns at various stages of training? Language models have been recently described as open knowledge bases. We can generate knowledge graphs by extracting relation triples from masked language models at sequential epochs or architecture variants to examine the knowledge acquisition process.

Dataset: Squad, Google-RE (3 flavors)

Models: BERT, RoBeRTa, DistilBert, training RoBERTa from scratch

Authors: Vinitra Swamy, Angelika Romanou, Martin Jaggi

This repository is the official implementation of the NeurIPS 2021 XAI4Debugging paper titled "Interpreting Language Models Through Knowledge Graph Extraction". Found this work useful? Please cite our paper.

Quick Start Guide

Pretrained Model (BERT, DistilBERT, RoBERTa) -> Knowlege Graph

Install requirements and clone repository

git clone https://github.com/epfml/interpret-lm-knowledge.git
pip install git+https://github.com/huggingface/transformers   
pip install textacy
cd interpret-lm-knowledge/scripts

Generate knowledge graphs and dataframes python run_knowledge_graph_experiments.py <dataset> <model> <use_spacy>
e.g. squad Bert spacy
e.g. re-place-birth Roberta

options:

dataset=squad - "squad", "re-place-birth", "re-date-birth", "re-place-death"  
model=Roberta - "Bert", "Roberta", "DistilBert"  
extractor=spacy - "spacy", "textacy", "custom"

See run_lm_experiments notebook for examples.

Train LM model from scratch -> Knowledge Graph

Install requirements and clone repository

!pip install git+https://github.com/huggingface/transformers
!pip list | grep -E 'transformers|tokenizers'
!pip install textacy

Run wikipedia_train_from_scratch_lm.ipynb.
As included in the last cell of the notebook, you can run the KG generation experiments by:

from run_training_kg_experiments import *
run_experiments(tokenizer, model, unmasker, "Roberta3e")

Citations

@inproceedings{swamy2021interpreting,
 author = {Swamy, Vinitra and Romanou, Angelika and Jaggi, Martin},
 booktitle = {Advances in Neural Information Processing Systems, Workshop on eXplainable AI Approaches for Debugging and Diagnosis},
 title = {Interpreting Language Models Through Knowledge Graph Extraction},
 volume = {35},
 year = {2021}
}

Extracting knowledge graphs from language models as a diagnostic benchmark of model performance.

Related tags

Overview

Interpreting Language Models Through Knowledge Graph Extraction

Quick Start Guide

Pretrained Model (BERT, DistilBERT, RoBERTa) -> Knowlege Graph

Train LM model from scratch -> Knowledge Graph

Citations

Owner

EPFL Machine Learning and Optimization Laboratory

(NeurIPS 2021) Realistic Evaluation of Transductive Few-Shot Learning

A distributed deep learning framework that supports flexible parallelization strategies.

SASM - simple crossplatform IDE for NASM, MASM, GAS and FASM assembly languages

Cancer-and-Tumor-Detection-Using-Inception-model - In this repo i am gonna show you how i did cancer/tumor detection in lungs using deep neural networks, specifically here the Inception model by google.

A simple library that implements CLIP guided loss in PyTorch.

The codebase for our paper "Generative Occupancy Fields for 3D Surface-Aware Image Synthesis" (NeurIPS 2021)

Here is the diagnostic tool for BMVC 2021 paper Diagnosing Errors in Video Relation Detectors.

FaceQgen: Semi-Supervised Deep Learning for Face Image Quality Assessment

🍅🍅🍅YOLOv5-Lite: lighter, faster and easier to deploy. Evolved from yolov5 and the size of model is only 1.7M (int8) and 3.3M (fp16). It can reach 10+ FPS on the Raspberry Pi 4B when the input size is 320×320~

Implementation of parameterized soft-exponential activation function.

HuSpaCy: industrial-strength Hungarian natural language processing

LIMEcraft: Handcrafted superpixel selectionand inspection for Visual eXplanations

TART - A PyTorch implementation for Transition Matrix Representation of Trees with Transposed Convolutions

PaSST: Efficient Training of Audio Transformers with Patchout

The pytorch implementation of the paper "text-guided neural image inpainting" at MM'2020

Using knowledge-informed machine learning on the PRONOSTIA (FEMTO) and IMS bearing data sets. Predict remaining-useful-life (RUL).

An implementation of the BADGE batch active learning algorithm.

Code for the paper "TadGAN: Time Series Anomaly Detection Using Generative Adversarial Networks"

TransGAN: Two Transformers Can Make One Strong GAN

Data labels and scripts for fastMRI.org