{
  "id": 188901,
  "title": "Paper and more datasets by the organizer Riiid",
  "url": "/competitions/riiid-test-answer-prediction/discussion/188901",
  "author_name": "Sirish Somanchi",
  "post_date": "2020-10-05T19:39:24.837000",
  "votes": 33,
  "comment_count": 8,
  "views": 0,
  "content": "<p>[<a href=\"https://arxiv.org/pdf/1912.03072.pdf\" target=\"_blank\">Paper</a>]: <a href=\"https://arxiv.org/abs/1912.03072\" target=\"_blank\">EdNet</a>: A Large-Scale Hierarchical Dataset in Education</p>\n<p>See <a href=\"https://github.com/riiid/ednet\" target=\"_blank\">Github</a> for additional datasets by the organizer:</p>\n<ul>\n<li>EdNet-KT1 : bit.ly/ednet_kt1</li>\n<li>EdNet-KT2 : bit.ly/ednet-kt2</li>\n<li>EdNet-KT3 : bit.ly/ednet-kt3</li>\n<li>EdNet-KT4 : bit.ly/ednet-kt4</li>\n<li>Contents : bit.ly/ednet-content</li>\n</ul>\n<p>See EDA <a href=\"https://github.com/premonish/EdNet/blob/master/notebooks/\" target=\"_blank\">jupyter notebooks</a></p>",
  "messages": [
    {
      "id": 1038443,
      "postDate": "2020-10-05T19:39:24.837Z",
      "content": "<p>[<a href=\"https://arxiv.org/pdf/1912.03072.pdf\" target=\"_blank\">Paper</a>]: <a href=\"https://arxiv.org/abs/1912.03072\" target=\"_blank\">EdNet</a>: A Large-Scale Hierarchical Dataset in Education</p>\n<p>See <a href=\"https://github.com/riiid/ednet\" target=\"_blank\">Github</a> for additional datasets by the organizer:</p>\n<ul>\n<li>EdNet-KT1 : bit.ly/ednet_kt1</li>\n<li>EdNet-KT2 : bit.ly/ednet-kt2</li>\n<li>EdNet-KT3 : bit.ly/ednet-kt3</li>\n<li>EdNet-KT4 : bit.ly/ednet-kt4</li>\n<li>Contents : bit.ly/ednet-content</li>\n</ul>\n<p>See EDA <a href=\"https://github.com/premonish/EdNet/blob/master/notebooks/\" target=\"_blank\">jupyter notebooks</a></p>",
      "rawMarkdown": "[[Paper](https://arxiv.org/pdf/1912.03072.pdf)]: [EdNet](https://arxiv.org/abs/1912.03072): A Large-Scale Hierarchical Dataset in Education\n\nSee [Github](https://github.com/riiid/ednet) for additional datasets by the organizer:\n- EdNet-KT1 : bit.ly/ednet_kt1\n- EdNet-KT2 : bit.ly/ednet-kt2\n- EdNet-KT3 : bit.ly/ednet-kt3\n- EdNet-KT4 : bit.ly/ednet-kt4\n- Contents : bit.ly/ednet-content\n\nSee EDA [jupyter notebooks](https://github.com/premonish/EdNet/blob/master/notebooks/)\n",
      "votes": 33
    },
    {
      "id": 1045560,
      "postDate": "2020-10-10T18:42:27.020Z",
      "content": "<p>Among the above for datasets, which is the competition dataset provided?</p>",
      "rawMarkdown": "Among the above for datasets, which is the competition dataset provided?",
      "votes": 1,
      "replies": [
        {
          "id": 1045574,
          "postDate": "2020-10-10T19:09:41.633Z",
          "content": "<p>Yes, are we sure these are not contained within the competition dataset?</p>",
          "rawMarkdown": "Yes, are we sure these are not contained within the competition dataset?"
        },
        {
          "id": 1045888,
          "postDate": "2020-10-11T06:10:33.860Z",
          "content": "<p>After my inspection, I think KT4 is provided to us. Since EDNet is a Hierarchical dataset, as we go from KT1 to KT4 the granularity of the data increases. So Since we are provided with the one with the highest granularity it is possible to extract any other lower Dataset by aggregation. The attached screenshot is taken from EDNet paper which explains my findings</p>",
          "rawMarkdown": "After my inspection, I think KT4 is provided to us. Since EDNet is a Hierarchical dataset, as we go from KT1 to KT4 the granularity of the data increases. So Since we are provided with the one with the highest granularity it is possible to extract any other lower Dataset by aggregation. The attached screenshot is taken from EDNet paper which explains my findings\n\n",
          "votes": 1
        },
        {
          "id": 1047503,
          "postDate": "2020-10-12T17:06:26.167Z",
          "content": "<p>The dataset provided to us doesn't strictly fall into any category in KT1-4 hierarchy. </p>\n<p>We are provided (q,a) pairs like in KT1.<br>\nWe are <strong>not</strong> provided the info on a student alternating between choices like in KT2<br>\nWe are provided the provided info on lecture and explanation viewing like in KT3<br>\nWe are <strong>not</strong> provided the fine details like payment, audio plays etc. like in KT4</p>",
          "rawMarkdown": "The dataset provided to us doesn't strictly fall into any category in KT1-4 hierarchy. \n\nWe are provided (q,a) pairs like in KT1.\nWe are **not** provided the info on a student alternating between choices like in KT2\nWe are provided the provided info on lecture and explanation viewing like in KT3\nWe are **not** provided the fine details like payment, audio plays etc. like in KT4",
          "votes": 3
        },
        {
          "id": 1105752,
          "postDate": "2020-12-08T06:59:35.927Z",
          "content": "<p><a href=\"https://www.kaggle.com/gautham11\" target=\"_blank\">@gautham11</a>  is there any more data also there apart from Train.csv  i see people in SAKT public kernel talking about training with more no of Users </p>",
          "rawMarkdown": "@gautham11  is there any more data also there apart from Train.csv  i see people in SAKT public kernel talking about training with more no of Users "
        }
      ]
    },
    {
      "id": 1042770,
      "postDate": "2020-10-08T13:04:37.783Z",
      "rawMarkdown": "",
      "votes": 1,
      "isDeleted": true
    },
    {
      "id": 1045566,
      "postDate": "2020-10-10T18:57:58.820Z",
      "content": "<p>Amazing, thanks for sharing :)</p>",
      "rawMarkdown": "Amazing, thanks for sharing :)",
      "votes": 2
    },
    {
      "id": 1070935,
      "postDate": "2020-11-06T10:27:34.843Z",
      "content": "<p>Thanks, very useful indeed!</p>",
      "rawMarkdown": "Thanks, very useful indeed!"
    }
  ],
  "comments": [
    {
      "id": 1045560,
      "author_name": "Aravind P",
      "author_url": "",
      "post_date": "2020-10-10T18:42:27.020000",
      "content": "<p>Among the above for datasets, which is the competition dataset provided?</p>",
      "votes": 1,
      "replies": [
        {
          "id": 1045574,
          "author_name": "Alex",
          "author_url": "",
          "post_date": "2020-10-10T19:09:41.633000",
          "content": "<p>Yes, are we sure these are not contained within the competition dataset?</p>",
          "votes": 0,
          "replies": []
        },
        {
          "id": 1045888,
          "author_name": "Aravind P",
          "author_url": "",
          "post_date": "2020-10-11T06:10:33.860000",
          "content": "<p>After my inspection, I think KT4 is provided to us. Since EDNet is a Hierarchical dataset, as we go from KT1 to KT4 the granularity of the data increases. So Since we are provided with the one with the highest granularity it is possible to extract any other lower Dataset by aggregation. The attached screenshot is taken from EDNet paper which explains my findings</p>",
          "votes": 1,
          "replies": []
        },
        {
          "id": 1047503,
          "author_name": "GauthamKumaran",
          "author_url": "",
          "post_date": "2020-10-12T17:06:26.167000",
          "content": "<p>The dataset provided to us doesn't strictly fall into any category in KT1-4 hierarchy. </p>\n<p>We are provided (q,a) pairs like in KT1.<br>\nWe are <strong>not</strong> provided the info on a student alternating between choices like in KT2<br>\nWe are provided the provided info on lecture and explanation viewing like in KT3<br>\nWe are <strong>not</strong> provided the fine details like payment, audio plays etc. like in KT4</p>",
          "votes": 3,
          "replies": []
        },
        {
          "id": 1105752,
          "author_name": "Jaideep",
          "author_url": "",
          "post_date": "2020-12-08T06:59:35.927000",
          "content": "<p><a href=\"https://www.kaggle.com/gautham11\" target=\"_blank\">@gautham11</a>  is there any more data also there apart from Train.csv  i see people in SAKT public kernel talking about training with more no of Users </p>",
          "votes": 0,
          "replies": []
        }
      ]
    },
    {
      "id": 1042770,
      "author_name": "",
      "author_url": "",
      "post_date": "2020-10-08T13:04:37.783000",
      "content": "",
      "votes": 1,
      "replies": []
    },
    {
      "id": 1045566,
      "author_name": "domiziano.stingi95",
      "author_url": "",
      "post_date": "2020-10-10T18:57:58.820000",
      "content": "<p>Amazing, thanks for sharing :)</p>",
      "votes": 2,
      "replies": []
    },
    {
      "id": 1070935,
      "author_name": "Jorge Marcher",
      "author_url": "",
      "post_date": "2020-11-06T10:27:34.843000",
      "content": "<p>Thanks, very useful indeed!</p>",
      "votes": 0,
      "replies": []
    }
  ],
  "raw_markdown_by_id": {
    "1038443": "[[Paper](https://arxiv.org/pdf/1912.03072.pdf)]: [EdNet](https://arxiv.org/abs/1912.03072): A Large-Scale Hierarchical Dataset in Education\n\nSee [Github](https://github.com/riiid/ednet) for additional datasets by the organizer:\n- EdNet-KT1 : bit.ly/ednet_kt1\n- EdNet-KT2 : bit.ly/ednet-kt2\n- EdNet-KT3 : bit.ly/ednet-kt3\n- EdNet-KT4 : bit.ly/ednet-kt4\n- Contents : bit.ly/ednet-content\n\nSee EDA [jupyter notebooks](https://github.com/premonish/EdNet/blob/master/notebooks/)\n",
    "1045560": "Among the above for datasets, which is the competition dataset provided?",
    "1042770": "",
    "1045566": "Amazing, thanks for sharing :)",
    "1070935": "Thanks, very useful indeed!"
  }
}