reddit_tifu

مراجع:

کوتاه

برای بارگذاری این مجموعه داده در TFDS از دستور زیر استفاده کنید:

ds = tfds.load('huggingface:reddit_tifu/short')
  • توضیحات :
Reddit dataset, where TIFU denotes the name of subbreddit /r/tifu.
As defined in the publication, styel "short" uses title as summary and
"long" uses tldr as summary.

Features includes:
  - document: post text without tldr.
  - tldr: tldr line.
  - title: trimmed title without tldr.
  - ups: upvotes.
  - score: score.
  - num_comments: number of comments.
  - upvote_ratio: upvote ratio.
  • مجوز : مجوز شناخته شده ای وجود ندارد
  • نسخه : 1.1.0
  • تقسیم ها :
تقسیم کنید نمونه ها
'train' 79740
  • ویژگی ها :
{
    "ups": {
        "dtype": "float32",
        "id": null,
        "_type": "Value"
    },
    "num_comments": {
        "dtype": "float32",
        "id": null,
        "_type": "Value"
    },
    "upvote_ratio": {
        "dtype": "float32",
        "id": null,
        "_type": "Value"
    },
    "score": {
        "dtype": "float32",
        "id": null,
        "_type": "Value"
    },
    "documents": {
        "dtype": "string",
        "id": null,
        "_type": "Value"
    },
    "tldr": {
        "dtype": "string",
        "id": null,
        "_type": "Value"
    },
    "title": {
        "dtype": "string",
        "id": null,
        "_type": "Value"
    }
}

طولانی

برای بارگذاری این مجموعه داده در TFDS از دستور زیر استفاده کنید:

ds = tfds.load('huggingface:reddit_tifu/long')
  • توضیحات :
Reddit dataset, where TIFU denotes the name of subbreddit /r/tifu.
As defined in the publication, styel "short" uses title as summary and
"long" uses tldr as summary.

Features includes:
  - document: post text without tldr.
  - tldr: tldr line.
  - title: trimmed title without tldr.
  - ups: upvotes.
  - score: score.
  - num_comments: number of comments.
  - upvote_ratio: upvote ratio.
  • مجوز : مجوز شناخته شده ای وجود ندارد
  • نسخه : 1.1.0
  • تقسیم ها :
تقسیم کنید نمونه ها
'train' 42139
  • ویژگی ها :
{
    "ups": {
        "dtype": "float32",
        "id": null,
        "_type": "Value"
    },
    "num_comments": {
        "dtype": "float32",
        "id": null,
        "_type": "Value"
    },
    "upvote_ratio": {
        "dtype": "float32",
        "id": null,
        "_type": "Value"
    },
    "score": {
        "dtype": "float32",
        "id": null,
        "_type": "Value"
    },
    "documents": {
        "dtype": "string",
        "id": null,
        "_type": "Value"
    },
    "tldr": {
        "dtype": "string",
        "id": null,
        "_type": "Value"
    },
    "title": {
        "dtype": "string",
        "id": null,
        "_type": "Value"
    }
}