सन्दर्भ:
इस डेटासेट को TFDS में लोड करने के लिए निम्नलिखित कमांड का उपयोग करें:
ds = tfds.load('huggingface:arxiv_dataset')
- विवरण :
A dataset of 1.7 million arXiv articles for applications like trend analysis, paper recommender engines, category prediction, co-citation networks, knowledge graph construction and semantic search interfaces.
- लाइसेंस : कोई ज्ञात लाइसेंस नहीं
- संस्करण : 1.1.0
- विभाजन :
विभाजित करना | उदाहरण |
---|---|
'train' | 1796911 |
- विशेषताएँ :
{
"id": {
"dtype": "string",
"id": null,
"_type": "Value"
},
"submitter": {
"dtype": "string",
"id": null,
"_type": "Value"
},
"authors": {
"dtype": "string",
"id": null,
"_type": "Value"
},
"title": {
"dtype": "string",
"id": null,
"_type": "Value"
},
"comments": {
"dtype": "string",
"id": null,
"_type": "Value"
},
"journal-ref": {
"dtype": "string",
"id": null,
"_type": "Value"
},
"doi": {
"dtype": "string",
"id": null,
"_type": "Value"
},
"report-no": {
"dtype": "string",
"id": null,
"_type": "Value"
},
"categories": {
"dtype": "string",
"id": null,
"_type": "Value"
},
"license": {
"dtype": "string",
"id": null,
"_type": "Value"
},
"abstract": {
"dtype": "string",
"id": null,
"_type": "Value"
},
"update_date": {
"dtype": "string",
"id": null,
"_type": "Value"
}
}