Références :
Utilisez la commande suivante pour charger cet ensemble de données dans TFDS :
ds = tfds.load('huggingface:multi_nli')
- Description :
The Multi-Genre Natural Language Inference (MultiNLI) corpus is a
crowd-sourced collection of 433k sentence pairs annotated with textual
entailment information. The corpus is modeled on the SNLI corpus, but differs in
that covers a range of genres of spoken and written text, and supports a
distinctive cross-genre generalization evaluation. The corpus served as the
basis for the shared task of the RepEval 2017 Workshop at EMNLP in Copenhagen.
- Licence : Aucune licence connue
- Version : 0.0.0
- Divisions :
Diviser | Exemples |
---|---|
'train' | 392702 |
'validation_matched' | 9815 |
'validation_mismatched' | 9832 |
- Caractéristiques :
{
"promptID": {
"dtype": "int32",
"id": null,
"_type": "Value"
},
"pairID": {
"dtype": "string",
"id": null,
"_type": "Value"
},
"premise": {
"dtype": "string",
"id": null,
"_type": "Value"
},
"premise_binary_parse": {
"dtype": "string",
"id": null,
"_type": "Value"
},
"premise_parse": {
"dtype": "string",
"id": null,
"_type": "Value"
},
"hypothesis": {
"dtype": "string",
"id": null,
"_type": "Value"
},
"hypothesis_binary_parse": {
"dtype": "string",
"id": null,
"_type": "Value"
},
"hypothesis_parse": {
"dtype": "string",
"id": null,
"_type": "Value"
},
"genre": {
"dtype": "string",
"id": null,
"_type": "Value"
},
"label": {
"num_classes": 3,
"names": [
"entailment",
"neutral",
"contradiction"
],
"names_file": null,
"id": null,
"_type": "ClassLabel"
}
}