BART-SLED (SLiding-Encoder and Decoder, base-sized model)
SLED models use pretrained, short-range encoder-decoder models, and apply them over
long-text inputs by splitting the input into multiple overlapping chunks, encoding each independently and perform fusion-in-decoder
Model description
This SLED model is based on the BART model, which is described in its
model card
.
BART is particularly effective when fine-tuned for text generation (e.g. summarization, translation) but also works
well for comprehension tasks (e.g. text classification, question answering). When used as a BART-SLED model, it can be applied on long text tasks.
Once installed, SLED is fully compatible with HuggingFace's AutoClasses (AutoTokenizer, AutoConfig, AutoModel
and AutoModelForCausalLM) and can be loaded using the from_pretrained methods
import sled # *** required so that SledModels will be registered for the AutoClasses ***
model = AutoModel.from_pretrained('tau/bart-base-sled')
Here is how to use this model in PyTorch:
from sled import SledTokenizer, SledModel
tokenizer = SledTokenizer.from_pretrained('tau/bart-base-sled')
model = SledModel.from_pretrained('tau/bart-base-sled')
inputs = tokenizer("Hello, my dog is cute", return_tensors="pt")
outputs = model(**inputs)
last_hidden_states = outputs.last_hidden_state
You can also replace SledModel by SledModelForConditionalGeneration for Seq2Seq generation
model = SledModelForConditionalGeneration.from_pretrained('tau/bart-base-sled')
In case you wish to apply SLED on a task containing a prefix (e.g. question) which should be given as a context to
every chunk, you can pass the
prefix_length
tensor input as well (A LongTensor in the length of the batch size).
import torch
import sled # *** required so that SledModels will be registered for the AutoClasses ***
tokenizer = AutoTokenizer.from_pretrained('tau/bart-base-sled')
model = AutoModel.from_pretrained('tau/bart-base-sled')
document_input_ids = tokenizer("Dogs are great for you.", return_tensors="pt").input_ids
prefix_input_ids = tokenizer("Are dogs good for you?", return_tensors="pt").input_ids
input_ids = torch.cat((prefix_input_ids, document_input_ids), dim=-1)
attention_mask = torch.ones_like(input_ids)
prefix_length = torch.LongTensor([[prefix_input_ids.size(1)]])
outputs = model(input_ids=input_ids, attention_mask=attention_mask, prefix_length=prefix_length)
last_hidden_states = outputs.last_hidden_state
BibTeX entry and citation info
Please cite both the SLED
paper
and the BART
paper
by Lewis et al as well as SummScreenFD by Chen et. al.
@inproceedings{Ivgi2022EfficientLU,
title={Efficient Long-Text Understanding with Short-Text Models},
author={Maor Ivgi and Uri Shaham and Jonathan Berant},
year={2022}
}
@article{DBLP:journals/corr/abs-1910-13461,
author = {Mike Lewis and
Yinhan Liu and
Naman Goyal and
Marjan Ghazvininejad and
Abdelrahman Mohamed and
Omer Levy and
Veselin Stoyanov and
Luke Zettlemoyer},
title = {{BART:} Denoising Sequence-to-Sequence Pre-training for Natural Language
Generation, Translation, and Comprehension},
journal = {CoRR},
volume = {abs/1910.13461},
year = {2019},
url = {http://arxiv.org/abs/1910.13461},
eprinttype = {arXiv},
eprint = {1910.13461},
timestamp = {Thu, 31 Oct 2019 14:02:26 +0100},
biburl = {https://dblp.org/rec/journals/corr/abs-1910-13461.bib},
bibsource = {dblp computer science bibliography, https://dblp.org}
}
@inproceedings{Chen2022SummScreenAD,
title={SummScreen: A Dataset for Abstractive Screenplay Summarization},
author={Mingda Chen and Zewei Chu and Sam Wiseman and Kevin Gimpel},
booktitle={ACL},
year={2022}
}
Runs of tau bart-base-sled-summscreenfd on huggingface.co
11
Total runs
1
24-hour runs
1
3-day runs
1
7-day runs
6
30-day runs
More Information About bart-base-sled-summscreenfd huggingface.co Model
More bart-base-sled-summscreenfd license Visit here:
bart-base-sled-summscreenfd huggingface.co is an AI model on huggingface.co that provides bart-base-sled-summscreenfd's model effect (), which can be used instantly with this tau bart-base-sled-summscreenfd model. huggingface.co supports a free trial of the bart-base-sled-summscreenfd model, and also provides paid use of the bart-base-sled-summscreenfd. Support call bart-base-sled-summscreenfd model through api, including Node.js, Python, http.
bart-base-sled-summscreenfd huggingface.co is an online trial and call api platform, which integrates bart-base-sled-summscreenfd's modeling effects, including api services, and provides a free online trial of bart-base-sled-summscreenfd, you can try bart-base-sled-summscreenfd online for free by clicking the link below.
tau bart-base-sled-summscreenfd online free url in huggingface.co:
bart-base-sled-summscreenfd is an open source model from GitHub that offers a free installation service, and any user can find bart-base-sled-summscreenfd on GitHub to install. At the same time, huggingface.co provides the effect of bart-base-sled-summscreenfd install, users can directly use bart-base-sled-summscreenfd installed effect in huggingface.co for debugging and trial. It also supports api for free installation.
bart-base-sled-summscreenfd install url in huggingface.co: