Cross-Document Coreference Resolution

Last update: Nov 28, 2022

Related tags

Overview

Cross-Document Coreference Resolution

This repository contains code and models for end-to-end cross-document coreference resolution, as decribed in our papers:

Cross-document Coreference Resolution over Predicted Mentions (Findings of ACL 2021)
Realistic Evaluation Principles for Cross-document Coreference Resolution (*SEM 2021)

The models are trained on ECB+, but they can be used for any setting of multiple documents.

Getting started

Install python3 requirements pip install -r requirements.txt

Extract mentions and raw text from ECB+

Run the following script in order to extract the data from ECB+ dataset and build the gold conll files. The ECB+ corpus can be downloaded here.

python get_ecb_data.py --data_path path_to_data

Training Instructions

The core of our model is the pairwise scorer between two spans, which indicates how likely two spans belong to the same cluster.

Training method

We present 3 ways to train this pairwise scorer:

Pipeline: first train a span scorer, then train the pairwise scorer using the same spans at each epoch.
Continue: pre-train the span scorer, then train the pairwise scorer while keep training the span scorer.
End-to-end: train together both models from scratch.

In order to choose the training method, you need to set the value of the training_method in the config_pairwise.json to pipeline, continue or e2e. In our paper, we found the continue method to perform the best for event coreference and we apply it for entity and ALL as well.

What are the labels ?

In ECB+, the entity and event coreference clusters are annotated separately, making it possible to train a model only on event or entity coreference. Therefore, our model also allows to be trained on events, entity, or both. You need to set the value of the mention_type in the config_pairwise.json (and config_span_scorer.json) to events, entities or mixed (corresponding to ALL in the paper).

Running the model

In both pipeline and continue methods, you need to first run the span scorer model

python train_span_scorer --config configs/config_span_scorer.json

For the pairwise scorer, run the following script

python train_pairwise_scorer --config configs/config_pairwise.json

Some important parameters in config_pairwise.json:

max_mention_span
top_k: pruning coefficient
training_method: (pipeline, continue, e2e)
subtopic: (true, false) whether to train at the topic or subtopic level (ECB+ notions).

Tuning threshold for agglomerative clustering

The training above will save 10 models (one for each epoch) in the specified directory, while each model is composed of a span_repr, a span scorer and a pairwise scorer. In order to find the best model and the best threshold for the agglomerative clustering, you need to do an hyperparameter search on the 10 models + several values for threshold, evaluated on the dev set. To do that, please set the config_clustering.json (split: dev) and run the two following scripts:

python tuned_threshold.py --config configs/config_clustering.json

python run_scorer.py [path_of_directory_of_conll_files] [mention_type]

Prediction

Given the trained pairwise scorer, the best model_num and the threshold from the above training and tuning, set the config_clustering.json (split: test) and run the following script.

python predict.py --config configs/config_clustering

(model_path corresponds to the directory in which you've stored the trained models)

An important configuration in the config_clustering is the topic_level. If you set false , you need to provide the path to the predicted topics in predicted_topics_path to produce conll files at the corpus level.

Evaluation

The output of the predict.py script is a file in the standard conll format. Then, it's straightforward to evaluate it with its corresponding gold conll file (created in the first step), using the official conll coreference scorer that you can find here or the coval system (python implementation).

Make sure to use the gold files of the same evaluation level (topic or corpus) as the predictions.

Notes

If you chose to train the pairwise with the end-to-end method, you don't need to provide a span_repr_path or a span_scorer_path in the config_pairwise.json.
If you use this model with gold mentions, the span scorer is not relevant, you should ignore the training method.
If you're interested in a newer but heavier model, check out our cross-encoder model

Cross-Document Coreference Resolution

Related tags

Overview

Cross-Document Coreference Resolution

Getting started

Extract mentions and raw text from ECB+

Training Instructions

Training method

What are the labels ?

Running the model

Tuning threshold for agglomerative clustering

Prediction

Evaluation

Notes

Team

Owner

Arie Cattan

A copy of Ares that costs 30 fucking dollars.

A universal memory dumper using Frida

A containerized REST API around OpenAI's CLIP model.

PyTea: PyTorch Tensor shape error analyzer

Digital Twin Mobility Profiling: A Spatio-Temporal Graph Learning Approach

Code for SyncTwin: Treatment Effect Estimation with Longitudinal Outcomes (NeurIPS 2021)

This repository contains the data and code for the paper "Diverse Text Generation via Variational Encoder-Decoder Models with Gaussian Process Priors" ([email protected])

Tooling for converting STAC metadata to ODC data model

Semantic Segmentation in Pytorch. Network include: FCN、FCN_ResNet、SegNet、UNet、BiSeNet、BiSeNetV2、PSPNet、DeepLabv3_plus、 HRNet、DDRNet

Implementation of GGB color space

AugLiChem - The augmentation library for chemical systems.

A large-scale video dataset for the training and evaluation of 3D human pose estimation models

BYOL for Audio: Self-Supervised Learning for General-Purpose Audio Representation

Real-ESRGAN aims at developing Practical Algorithms for General Image Restoration.

Repo for "Benchmarking Robustness of 3D Point Cloud Recognition against Common Corruptions" https://arxiv.org/abs/2201.12296

Аналитика доходности инвестиционного портфеля в Тинькофф брокере

Algorithmic encoding of protected characteristics and its implications on disparities across subgroups

FEMDA: Robust classification with Flexible Discriminant Analysis in heterogeneous data

SlotRefine: A Fast Non-Autoregressive Model forJoint Intent Detection and Slot Filling

Reimplementation of NeurIPS'19: "Meta-Weight-Net: Learning an Explicit Mapping For Sample Weighting" by Shu et al.