مراجع:
en
برای بارگذاری این مجموعه داده در TFDS از دستور زیر استفاده کنید:
ds = tfds.load('huggingface:medical_dialog/en')
- توضیحات :
The MedDialog dataset (English) contains conversations (in English) between doctors and patients.It has 0.26 million dialogues. The data is continuously growing and more dialogues will be added. The raw dialogues are from healthcaremagic.com and icliniq.com.
All copyrights of the data belong to healthcaremagic.com and icliniq.com.
- مجوز : مجوز شناخته شده ای وجود ندارد
- نسخه : 1.0.0
- تقسیمات :
تقسیم کنید | نمونه ها |
---|---|
'train' | 229674 |
- ویژگی ها :
{
"file_name": {
"dtype": "string",
"id": null,
"_type": "Value"
},
"dialogue_id": {
"dtype": "int32",
"id": null,
"_type": "Value"
},
"dialogue_url": {
"dtype": "string",
"id": null,
"_type": "Value"
},
"dialogue_turns": {
"feature": {
"speaker": {
"num_classes": 2,
"names": [
"Patient",
"Doctor"
],
"id": null,
"_type": "ClassLabel"
},
"utterance": {
"dtype": "string",
"id": null,
"_type": "Value"
}
},
"length": -1,
"id": null,
"_type": "Sequence"
}
}
zh
برای بارگذاری این مجموعه داده در TFDS از دستور زیر استفاده کنید:
ds = tfds.load('huggingface:medical_dialog/zh')
- توضیحات :
The MedDialog dataset (English) contains conversations (in English) between doctors and patients.It has 0.26 million dialogues. The data is continuously growing and more dialogues will be added. The raw dialogues are from healthcaremagic.com and icliniq.com.
All copyrights of the data belong to healthcaremagic.com and icliniq.com.
- مجوز : مجوز شناخته شده ای وجود ندارد
- نسخه : 1.0.0
- تقسیمات :
تقسیم کنید | نمونه ها |
---|---|
'train' | 1921127 |
- ویژگی ها :
{
"file_name": {
"dtype": "string",
"id": null,
"_type": "Value"
},
"dialogue_id": {
"dtype": "int32",
"id": null,
"_type": "Value"
},
"dialogue_url": {
"dtype": "string",
"id": null,
"_type": "Value"
},
"dialogue_turns": {
"feature": {
"speaker": {
"num_classes": 2,
"names": [
"\u75c5\u4eba",
"\u533b\u751f"
],
"id": null,
"_type": "ClassLabel"
},
"utterance": {
"dtype": "string",
"id": null,
"_type": "Value"
}
},
"length": -1,
"id": null,
"_type": "Sequence"
}
}
پردازش شده.en
برای بارگذاری این مجموعه داده در TFDS از دستور زیر استفاده کنید:
ds = tfds.load('huggingface:medical_dialog/processed.en')
- توضیحات :
The MedDialog dataset (English) contains conversations (in English) between doctors and patients.It has 0.26 million dialogues. The data is continuously growing and more dialogues will be added. The raw dialogues are from healthcaremagic.com and icliniq.com.
All copyrights of the data belong to healthcaremagic.com and icliniq.com.
- مجوز : حق چاپ
- نسخه : 2.0.0
- تقسیمات :
تقسیم کنید | نمونه ها |
---|---|
'test' | 61 |
'train' | 482 |
'validation' | 60 |
- ویژگی ها :
{
"description": {
"dtype": "string",
"id": null,
"_type": "Value"
},
"utterances": {
"feature": {
"dtype": "string",
"id": null,
"_type": "Value"
},
"length": -1,
"id": null,
"_type": "Sequence"
}
}
پردازش شده.zh
برای بارگذاری این مجموعه داده در TFDS از دستور زیر استفاده کنید:
ds = tfds.load('huggingface:medical_dialog/processed.zh')
- توضیحات :
The MedDialog dataset (English) contains conversations (in English) between doctors and patients.It has 0.26 million dialogues. The data is continuously growing and more dialogues will be added. The raw dialogues are from healthcaremagic.com and icliniq.com.
All copyrights of the data belong to healthcaremagic.com and icliniq.com.
- مجوز : حق چاپ
- نسخه : 2.0.0
- تقسیمات :
تقسیم کنید | نمونه ها |
---|---|
'test' | 340754 |
'train' | 2725989 |
'validation' | 340748 |
- ویژگی ها :
{
"utterances": {
"feature": {
"dtype": "string",
"id": null,
"_type": "Value"
},
"length": -1,
"id": null,
"_type": "Sequence"
}
}