Official Implementation of "DialogLM: Pre-trained Model for Long Dialogue Understanding and Summarization."

Last update: Dec 19, 2022

Related tags

Deep Learning DialogLM

Overview

DialogLM

Code for AAAI 2022 paper: DialogLM: Pre-trained Model for Long Dialogue Understanding and Summarization.

Pre-trained Models

We release two versions of pre-trained models.

DialogLM is based on UniLMv2. According to whether sparse attention is introduced, it can be divided into two different versions to process dialogs of different lengths.
DialogLED builds on Longformer-Encoder-Decoder (LED) architecture and uses window-based denoising as the pre-training task on a large amount of long dialogue data for further training. You can use its base version and large version directly through HuggingFace.

Datasets

Please download the five datasets we used in our paper here (AMI, ICSI, QMSum, ForeverDreaming, TVMegaSite).

Finetuning for Downstream Tasks

Please go to specific folders to apply them to downstream tasks related to long dialogues.

Contributing

This project welcomes contributions and suggestions. Most contributions require you to agree to a Contributor License Agreement (CLA) declaring that you have the right to, and actually do, grant us the rights to use your contribution. For details, visit https://cla.opensource.microsoft.com.

When you submit a pull request, a CLA bot will automatically determine whether you need to provide a CLA and decorate the PR appropriately (e.g., status check, comment). Simply follow the instructions provided by the bot. You will only need to do this once across all repos using our CLA.

This project has adopted the Microsoft Open Source Code of Conduct. For more information see the Code of Conduct FAQ or contact [email protected] with any additional questions or comments.

Trademarks

This project may contain trademarks or logos for projects, products, or services. Authorized use of Microsoft trademarks or logos is subject to and must follow Microsoft's Trademark & Brand Guidelines. Use of Microsoft trademarks or logos in modified versions of this project must not cause confusion or imply Microsoft sponsorship. Any use of third-party trademarks or logos are subject to those third-party's policies.

Official Implementation of "DialogLM: Pre-trained Model for Long Dialogue Understanding and Summarization."

Related tags

Overview

DialogLM

Pre-trained Models

Datasets

Finetuning for Downstream Tasks

Contributing

Trademarks

Owner

Microsoft

This project uses Template Matching technique for object detecting by detection of template image over base image.

EsViT: Efficient self-supervised Vision Transformers

MediaPipe is a an open-source framework from Google for building multimodal

Notspot robot simulation - Python version

pybaum provides tools to work with pytrees which is a concept burrowed from JAX.

Implementation of "Learning to Match Features with Seeded Graph Matching Network" ICCV2021

⚖️🔁🔮🕵️‍♂️🦹🖼️ Code for Measuring the Contribution of Multiple Model Representations in Detecting Adversarial Instances paper.

Exploring Visual Engagement Signals for Representation Learning

A curated list of Machine Learning and Deep Learning tutorials in Jupyter Notebook format ready to run in Google Colaboratory

Understanding the Properties of Minimum Bayes Risk Decoding in Neural Machine Translation.

Official Python implementation of the 'Sparse deconvolution'-v0.3.0

Housing Price Prediction

Provided is code that demonstrates the training and evaluation of the work presented in the paper: "On the Detection of Digital Face Manipulation" published in CVPR 2020.

EfficientNetV2 implementation using PyTorch

MAME is a multi-purpose emulation framework.

Supplemental learning materials for "Fourier Feature Networks and Neural Volume Rendering"

Source code for Fathony, Sahu, Willmott, & Kolter, "Multiplicative Filter Networks", ICLR 2021.

Do Neural Networks for Segmentation Understand Insideness?

Code for the paper "Benchmarking and Analyzing Point Cloud Classification under Corruptions"

🔮 A refreshing functional take on deep learning, compatible with your favorite libraries

Official Implementation of "DialogLM: Pre-trained Model for Long Dialogue Understanding and Summarization."

Related tags

Overview

DialogLM

Pre-trained Models

Datasets

Finetuning for Downstream Tasks

Contributing

Trademarks

Owner

Microsoft

This project uses Template Matching technique for object detecting by detection of template image over base image.

EsViT: Efficient self-supervised Vision Transformers

MediaPipe is a an open-source framework from Google for building multimodal

Notspot robot simulation - Python version

pybaum provides tools to work with pytrees which is a concept burrowed from JAX.

Implementation of "Learning to Match Features with Seeded Graph Matching Network" ICCV2021

⚖️🔁🔮🕵️‍♂️🦹🖼️ Code for *Measuring the Contribution of Multiple Model Representations in Detecting Adversarial Instances* paper.

Exploring Visual Engagement Signals for Representation Learning

A curated list of Machine Learning and Deep Learning tutorials in Jupyter Notebook format ready to run in Google Colaboratory

Understanding the Properties of Minimum Bayes Risk Decoding in Neural Machine Translation.

Official Python implementation of the 'Sparse deconvolution'-v0.3.0

Housing Price Prediction

Provided is code that demonstrates the training and evaluation of the work presented in the paper: "On the Detection of Digital Face Manipulation" published in CVPR 2020.

EfficientNetV2 implementation using PyTorch

MAME is a multi-purpose emulation framework.

Supplemental learning materials for "Fourier Feature Networks and Neural Volume Rendering"

Source code for Fathony, Sahu, Willmott, & Kolter, "Multiplicative Filter Networks", ICLR 2021.

Do Neural Networks for Segmentation Understand Insideness?

Code for the paper "Benchmarking and Analyzing Point Cloud Classification under Corruptions"

🔮 A refreshing functional take on deep learning, compatible with your favorite libraries

⚖️🔁🔮🕵️‍♂️🦹🖼️ Code for Measuring the Contribution of Multiple Model Representations in Detecting Adversarial Instances paper.