নরওয়েজিয়ান_নার

তথ্যসূত্র:

বোকমাল

TFDS এ এই ডেটাসেট লোড করতে নিম্নলিখিত কমান্ডটি ব্যবহার করুন:

ds = tfds.load('huggingface:norwegian_ner/bokmaal')

বর্ণনা :

Named entities Recognition dataset for Norwegian.

লাইসেন্স : কোনো পরিচিত লাইসেন্স নেই
সংস্করণ : 1.0.0
বিভাজন :

বিভক্ত	উদাহরণ
`'test'`	1939
`'train'`	15696
`'validation'`	2410

বৈশিষ্ট্য :

{
    "idx": {
        "dtype": "string",
        "id": null,
        "_type": "Value"
    },
    "text": {
        "dtype": "string",
        "id": null,
        "_type": "Value"
    },
    "tokens": {
        "feature": {
            "dtype": "string",
            "id": null,
            "_type": "Value"
        },
        "length": -1,
        "id": null,
        "_type": "Sequence"
    },
    "lemmas": {
        "feature": {
            "dtype": "string",
            "id": null,
            "_type": "Value"
        },
        "length": -1,
        "id": null,
        "_type": "Sequence"
    },
    "pos_tags": {
        "feature": {
            "num_classes": 17,
            "names": [
                "NOUN",
                "PUNCT",
                "ADP",
                "NUM",
                "SYM",
                "SCONJ",
                "ADJ",
                "PART",
                "DET",
                "CCONJ",
                "PROPN",
                "PRON",
                "X",
                "ADV",
                "INTJ",
                "VERB",
                "AUX"
            ],
            "names_file": null,
            "id": null,
            "_type": "ClassLabel"
        },
        "length": -1,
        "id": null,
        "_type": "Sequence"
    },
    "ner_tags": {
        "feature": {
            "num_classes": 17,
            "names": [
                "O",
                "B-OTH",
                "I-OTH",
                "E-OTH",
                "S-OTH",
                "B-ORG",
                "I-ORG",
                "E-ORG",
                "S-ORG",
                "B-PRS",
                "I-PRS",
                "E-PRS",
                "S-PRS",
                "B-GEO",
                "I-GEO",
                "E-GEO",
                "S-GEO"
            ],
            "names_file": null,
            "id": null,
            "_type": "ClassLabel"
        },
        "length": -1,
        "id": null,
        "_type": "Sequence"
    }
}

নাইনরস্ক

TFDS এ এই ডেটাসেট লোড করতে নিম্নলিখিত কমান্ডটি ব্যবহার করুন:

ds = tfds.load('huggingface:norwegian_ner/nynorsk')

বর্ণনা :

Named entities Recognition dataset for Norwegian.

লাইসেন্স : কোনো পরিচিত লাইসেন্স নেই
সংস্করণ : 1.0.0
বিভাজন :

বিভক্ত	উদাহরণ
`'test'`	1511
`'train'`	14174
`'validation'`	1890

বৈশিষ্ট্য :

{
    "idx": {
        "dtype": "string",
        "id": null,
        "_type": "Value"
    },
    "text": {
        "dtype": "string",
        "id": null,
        "_type": "Value"
    },
    "tokens": {
        "feature": {
            "dtype": "string",
            "id": null,
            "_type": "Value"
        },
        "length": -1,
        "id": null,
        "_type": "Sequence"
    },
    "lemmas": {
        "feature": {
            "dtype": "string",
            "id": null,
            "_type": "Value"
        },
        "length": -1,
        "id": null,
        "_type": "Sequence"
    },
    "pos_tags": {
        "feature": {
            "num_classes": 17,
            "names": [
                "NOUN",
                "PUNCT",
                "ADP",
                "NUM",
                "SYM",
                "SCONJ",
                "ADJ",
                "PART",
                "DET",
                "CCONJ",
                "PROPN",
                "PRON",
                "X",
                "ADV",
                "INTJ",
                "VERB",
                "AUX"
            ],
            "names_file": null,
            "id": null,
            "_type": "ClassLabel"
        },
        "length": -1,
        "id": null,
        "_type": "Sequence"
    },
    "ner_tags": {
        "feature": {
            "num_classes": 17,
            "names": [
                "O",
                "B-OTH",
                "I-OTH",
                "E-OTH",
                "S-OTH",
                "B-ORG",
                "I-ORG",
                "E-ORG",
                "S-ORG",
                "B-PRS",
                "I-PRS",
                "E-PRS",
                "S-PRS",
                "B-GEO",
                "I-GEO",
                "E-GEO",
                "S-GEO"
            ],
            "names_file": null,
            "id": null,
            "_type": "ClassLabel"
        },
        "length": -1,
        "id": null,
        "_type": "Sequence"
    }
}

samnorsk

TFDS এ এই ডেটাসেট লোড করতে নিম্নলিখিত কমান্ডটি ব্যবহার করুন:

ds = tfds.load('huggingface:norwegian_ner/samnorsk')

বর্ণনা :

Named entities Recognition dataset for Norwegian.

লাইসেন্স : কোনো পরিচিত লাইসেন্স নেই
সংস্করণ : 1.0.0
বিভাজন :

বিভক্ত	উদাহরণ
`'test'`	3450
`'train'`	34170
`'validation'`	4300

বৈশিষ্ট্য :

{
    "idx": {
        "dtype": "string",
        "id": null,
        "_type": "Value"
    },
    "text": {
        "dtype": "string",
        "id": null,
        "_type": "Value"
    },
    "tokens": {
        "feature": {
            "dtype": "string",
            "id": null,
            "_type": "Value"
        },
        "length": -1,
        "id": null,
        "_type": "Sequence"
    },
    "lemmas": {
        "feature": {
            "dtype": "string",
            "id": null,
            "_type": "Value"
        },
        "length": -1,
        "id": null,
        "_type": "Sequence"
    },
    "pos_tags": {
        "feature": {
            "num_classes": 17,
            "names": [
                "NOUN",
                "PUNCT",
                "ADP",
                "NUM",
                "SYM",
                "SCONJ",
                "ADJ",
                "PART",
                "DET",
                "CCONJ",
                "PROPN",
                "PRON",
                "X",
                "ADV",
                "INTJ",
                "VERB",
                "AUX"
            ],
            "names_file": null,
            "id": null,
            "_type": "ClassLabel"
        },
        "length": -1,
        "id": null,
        "_type": "Sequence"
    },
    "ner_tags": {
        "feature": {
            "num_classes": 17,
            "names": [
                "O",
                "B-OTH",
                "I-OTH",
                "E-OTH",
                "S-OTH",
                "B-ORG",
                "I-ORG",
                "E-ORG",
                "S-ORG",
                "B-PRS",
                "I-PRS",
                "E-PRS",
                "S-PRS",
                "B-GEO",
                "I-GEO",
                "E-GEO",
                "S-GEO"
            ],
            "names_file": null,
            "id": null,
            "_type": "ClassLabel"
        },
        "length": -1,
        "id": null,
        "_type": "Sequence"
    }
}