TFDS รองรับ รูปแบบ Croissant 🥐 แล้ว! อ่าน เอกสาร เพื่อทราบข้อมูลเพิ่มเติม

หน้านี้ได้รับการแปลโดย Cloud Translation API

opus_dgt

อ้างอิง:

บีจี-กา

ใช้คำสั่งต่อไปนี้เพื่อโหลดชุดข้อมูลนี้ใน TFDS:

ds = tfds.load('huggingface:opus_dgt/bg-ga')

คำอธิบาย :

A collection of translation memories provided by the JRC. Source: https://ec.europa.eu/jrc/en/language-technologies/dgt-translation-memory
25 languages, 299 bitexts
total number of files: 817,410
total number of tokens: 2.13G
total number of sentence fragments: 113.52M

ใบอนุญาต : ไม่มีใบอนุญาตที่รู้จัก
เวอร์ชัน : 1.0.0
แยก :

แยก	ตัวอย่าง
`'train'`	179142

คุณสมบัติ :

{
    "id": {
        "dtype": "string",
        "id": null,
        "_type": "Value"
    },
    "translation": {
        "languages": [
            "bg",
            "ga"
        ],
        "id": null,
        "_type": "Translation"
    }
}

บีจี-ชม

ใช้คำสั่งต่อไปนี้เพื่อโหลดชุดข้อมูลนี้ใน TFDS:

ds = tfds.load('huggingface:opus_dgt/bg-hr')

คำอธิบาย :

A collection of translation memories provided by the JRC. Source: https://ec.europa.eu/jrc/en/language-technologies/dgt-translation-memory
25 languages, 299 bitexts
total number of files: 817,410
total number of tokens: 2.13G
total number of sentence fragments: 113.52M

ใบอนุญาต : ไม่มีใบอนุญาตที่รู้จัก
เวอร์ชัน : 1.0.0
แยก :

แยก	ตัวอย่าง
`'train'`	701572

คุณสมบัติ :

{
    "id": {
        "dtype": "string",
        "id": null,
        "_type": "Value"
    },
    "translation": {
        "languages": [
            "bg",
            "hr"
        ],
        "id": null,
        "_type": "Translation"
    }
}

บีจี-ช

ใช้คำสั่งต่อไปนี้เพื่อโหลดชุดข้อมูลนี้ใน TFDS:

ds = tfds.load('huggingface:opus_dgt/bg-sh')

คำอธิบาย :

A collection of translation memories provided by the JRC. Source: https://ec.europa.eu/jrc/en/language-technologies/dgt-translation-memory
25 languages, 299 bitexts
total number of files: 817,410
total number of tokens: 2.13G
total number of sentence fragments: 113.52M

ใบอนุญาต : ไม่มีใบอนุญาตที่รู้จัก
เวอร์ชัน : 1.0.0
แยก :

แยก	ตัวอย่าง
`'train'`	1488507

คุณสมบัติ :

{
    "id": {
        "dtype": "string",
        "id": null,
        "_type": "Value"
    },
    "translation": {
        "languages": [
            "bg",
            "sh"
        ],
        "id": null,
        "_type": "Translation"
    }
}

ฟิ-กา

ใช้คำสั่งต่อไปนี้เพื่อโหลดชุดข้อมูลนี้ใน TFDS:

ds = tfds.load('huggingface:opus_dgt/fi-ga')

คำอธิบาย :

A collection of translation memories provided by the JRC. Source: https://ec.europa.eu/jrc/en/language-technologies/dgt-translation-memory
25 languages, 299 bitexts
total number of files: 817,410
total number of tokens: 2.13G
total number of sentence fragments: 113.52M

ใบอนุญาต : ไม่มีใบอนุญาตที่รู้จัก
เวอร์ชัน : 1.0.0
แยก :

แยก	ตัวอย่าง
`'train'`	178619

คุณสมบัติ :

{
    "id": {
        "dtype": "string",
        "id": null,
        "_type": "Value"
    },
    "translation": {
        "languages": [
            "fi",
            "ga"
        ],
        "id": null,
        "_type": "Translation"
    }
}

es-ga

ใช้คำสั่งต่อไปนี้เพื่อโหลดชุดข้อมูลนี้ใน TFDS:

ds = tfds.load('huggingface:opus_dgt/es-ga')

คำอธิบาย :

A collection of translation memories provided by the JRC. Source: https://ec.europa.eu/jrc/en/language-technologies/dgt-translation-memory
25 languages, 299 bitexts
total number of files: 817,410
total number of tokens: 2.13G
total number of sentence fragments: 113.52M

ใบอนุญาต : ไม่มีใบอนุญาตที่รู้จัก
เวอร์ชัน : 1.0.0
แยก :

แยก	ตัวอย่าง
`'train'`	178696

คุณสมบัติ :

{
    "id": {
        "dtype": "string",
        "id": null,
        "_type": "Value"
    },
    "translation": {
        "languages": [
            "es",
            "ga"
        ],
        "id": null,
        "_type": "Translation"
    }
}

กาช

ใช้คำสั่งต่อไปนี้เพื่อโหลดชุดข้อมูลนี้ใน TFDS:

ds = tfds.load('huggingface:opus_dgt/ga-sh')

คำอธิบาย :

A collection of translation memories provided by the JRC. Source: https://ec.europa.eu/jrc/en/language-technologies/dgt-translation-memory
25 languages, 299 bitexts
total number of files: 817,410
total number of tokens: 2.13G
total number of sentence fragments: 113.52M

ใบอนุญาต : ไม่มีใบอนุญาตที่รู้จัก
เวอร์ชัน : 1.0.0
แยก :

แยก	ตัวอย่าง
`'train'`	91613

คุณสมบัติ :

{
    "id": {
        "dtype": "string",
        "id": null,
        "_type": "Value"
    },
    "translation": {
        "languages": [
            "ga",
            "sh"
        ],
        "id": null,
        "_type": "Translation"
    }
}

ชม.-sk

ใช้คำสั่งต่อไปนี้เพื่อโหลดชุดข้อมูลนี้ใน TFDS:

ds = tfds.load('huggingface:opus_dgt/hr-sk')

คำอธิบาย :

A collection of translation memories provided by the JRC. Source: https://ec.europa.eu/jrc/en/language-technologies/dgt-translation-memory
25 languages, 299 bitexts
total number of files: 817,410
total number of tokens: 2.13G
total number of sentence fragments: 113.52M

ใบอนุญาต : ไม่มีใบอนุญาตที่รู้จัก
เวอร์ชัน : 1.0.0
แยก :

แยก	ตัวอย่าง
`'train'`	689263

คุณสมบัติ :

{
    "id": {
        "dtype": "string",
        "id": null,
        "_type": "Value"
    },
    "translation": {
        "languages": [
            "hr",
            "sk"
        ],
        "id": null,
        "_type": "Translation"
    }
}

MT-ช

ใช้คำสั่งต่อไปนี้เพื่อโหลดชุดข้อมูลนี้ใน TFDS:

ds = tfds.load('huggingface:opus_dgt/mt-sh')

คำอธิบาย :

A collection of translation memories provided by the JRC. Source: https://ec.europa.eu/jrc/en/language-technologies/dgt-translation-memory
25 languages, 299 bitexts
total number of files: 817,410
total number of tokens: 2.13G
total number of sentence fragments: 113.52M

ใบอนุญาต : ไม่มีใบอนุญาตที่รู้จัก
เวอร์ชัน : 1.0.0
แยก :

แยก	ตัวอย่าง
`'train'`	1450424

คุณสมบัติ :

{
    "id": {
        "dtype": "string",
        "id": null,
        "_type": "Value"
    },
    "translation": {
        "languages": [
            "mt",
            "sh"
        ],
        "id": null,
        "_type": "Translation"
    }
}

ชม.-sv

ใช้คำสั่งต่อไปนี้เพื่อโหลดชุดข้อมูลนี้ใน TFDS:

ds = tfds.load('huggingface:opus_dgt/hr-sv')

คำอธิบาย :

A collection of translation memories provided by the JRC. Source: https://ec.europa.eu/jrc/en/language-technologies/dgt-translation-memory
25 languages, 299 bitexts
total number of files: 817,410
total number of tokens: 2.13G
total number of sentence fragments: 113.52M

ใบอนุญาต : ไม่มีใบอนุญาตที่รู้จัก
เวอร์ชัน : 1.0.0
แยก :

แยก	ตัวอย่าง
`'train'`	696334

คุณสมบัติ :

{
    "id": {
        "dtype": "string",
        "id": null,
        "_type": "Value"
    },
    "translation": {
        "languages": [
            "hr",
            "sv"
        ],
        "id": null,
        "_type": "Translation"
    }
}

กา-nl

ใช้คำสั่งต่อไปนี้เพื่อโหลดชุดข้อมูลนี้ใน TFDS:

ds = tfds.load('huggingface:opus_dgt/ga-nl')

คำอธิบาย :

A collection of translation memories provided by the JRC. Source: https://ec.europa.eu/jrc/en/language-technologies/dgt-translation-memory
25 languages, 299 bitexts
total number of files: 817,410
total number of tokens: 2.13G
total number of sentence fragments: 113.52M

ใบอนุญาต : ไม่มีใบอนุญาตที่รู้จัก
เวอร์ชัน : 1.0.0
แยก :

แยก	ตัวอย่าง
`'train'`	170644

คุณสมบัติ :

{
    "id": {
        "dtype": "string",
        "id": null,
        "_type": "Value"
    },
    "translation": {
        "languages": [
            "ga",
            "nl"
        ],
        "id": null,
        "_type": "Translation"
    }
}