{
  "id": 36470,
  "title": "rows in submission file",
  "url": "/competitions/expedia-hotel-recommendations/discussion/36470",
  "author_name": "rbaral",
  "post_date": "2017-07-16T16:52:05.505000",
  "votes": 0,
  "comment_count": 0,
  "views": 0,
  "content": "<p>Hello everyone,</p>\n\n<p>I was just exploring this topic for fun even though it is already completed. In the submission file, the expected number of rows is 2528243. The sample submission file shows that every userid in the submission is unique. My submission was generated for the test data which also contains 2528243 entries, but the userid are repeated. The submission does not allow duplicate user rows. I tested with the unique userids from train, test, and both, but still do not get that many userids.\nSo, how to get 2528243 rows without having duplicate userid?</p>\n\n<p>Thanks.</p>",
  "messages": [
    {
      "id": 203846,
      "postDate": "2017-07-16T16:52:05.507Z",
      "content": "<p>Hello everyone,</p>\n\n<p>I was just exploring this topic for fun even though it is already completed. In the submission file, the expected number of rows is 2528243. The sample submission file shows that every userid in the submission is unique. My submission was generated for the test data which also contains 2528243 entries, but the userid are repeated. The submission does not allow duplicate user rows. I tested with the unique userids from train, test, and both, but still do not get that many userids.\nSo, how to get 2528243 rows without having duplicate userid?</p>\n\n<p>Thanks.</p>",
      "rawMarkdown": "Hello everyone,\n\nI was just exploring this topic for fun even though it is already completed. In the submission file, the expected number of rows is 2528243. The sample submission file shows that every userid in the submission is unique. My submission was generated for the test data which also contains 2528243 entries, but the userid are repeated. The submission does not allow duplicate user rows. I tested with the unique userids from train, test, and both, but still do not get that many userids.\nSo, how to get 2528243 rows without having duplicate userid?\n\nThanks.\n"
    }
  ],
  "comments": [],
  "raw_markdown_by_id": {
    "203846": "Hello everyone,\n\nI was just exploring this topic for fun even though it is already completed. In the submission file, the expected number of rows is 2528243. The sample submission file shows that every userid in the submission is unique. My submission was generated for the test data which also contains 2528243 entries, but the userid are repeated. The submission does not allow duplicate user rows. I tested with the unique userids from train, test, and both, but still do not get that many userids.\nSo, how to get 2528243 rows without having duplicate userid?\n\nThanks.\n"
  }
}