BookCorpus

E1312410 UNEXPLORED

BookCorpus is a large collection of freely available books commonly used as a pretraining dataset for natural language processing models.

All labels observed (3)

Label Occurrences
BookCorpus canonical 1
BookCorpus (for some variants) 1
BooksCorpus 1

How this entity was disambiguated

Referenced by (3)

Full triples — surface form annotated when it differs from this entity's canonical label.

RoBERTa trainingDataSource BookCorpus
DeBERTa usesPretrainingData BookCorpus (for some variants)
linked to: BookCorpus
BERT trainingCorpus BooksCorpus
linked to: BookCorpus