{
  "id": 172907,
  "title": "Public and Private training data",
  "url": "/competitions/landmark-recognition-2020/discussion/172907",
  "author_name": "",
  "post_date": "2020-08-06T23:02:05.769663200Z",
  "votes": 1,
  "comment_count": 2,
  "views": 0,
  "content": "<p>It's not clear to me what the private training data (~100k examples) referred to in the description is. The train.csv is over a million rows and all the image files for those appear to be in the dataset. </p>\n\n<p>Apologies if it's obvious! First time participating in one of these landmark competitions.</p>",
  "messages": [
    {
      "id": "961102",
      "postDate": "08/06/2020 23:02:05",
      "content": "<p>It's not clear to me what the private training data (~100k examples) referred to in the description is. The train.csv is over a million rows and all the image files for those appear to be in the dataset. </p>\n\n<p>Apologies if it's obvious! First time participating in one of these landmark competitions.</p>",
      "rawMarkdown": "It's not clear to me what the private training data (~100k examples) referred to in the description is. The train.csv is over a million rows and all the image files for those appear to be in the dataset. \n\nApologies if it's obvious! First time participating in one of these landmark competitions.",
      "votes": null
    },
    {
      "id": "962170",
      "postDate": "08/07/2020 21:43:16",
      "content": "<p><a href=\"https://www.kaggle.com/c/landmark-recognition-2020/discussion/173111\">https://www.kaggle.com/c/landmark-recognition-2020/discussion/173111</a></p>",
      "rawMarkdown": "https://www.kaggle.com/c/landmark-recognition-2020/discussion/173111",
      "votes": null
    },
    {
      "id": "985065",
      "postDate": "08/25/2020 13:33:47",
      "content": "<p>This is my first time participating in a Kaggle competition and I was asking myself the same question. Based on what is written in the Data section of the competition \"When you submit your notebook, Kaggle will rerun your code on the private dataset.\" What I understood about the private dataset is that, if not mistaken, it is a subset of the public training data that will be used to run our submitted notebooks, and ranking will be determined based on the outcome of it.</p>",
      "rawMarkdown": "This is my first time participating in a Kaggle competition and I was asking myself the same question. Based on what is written in the Data section of the competition \"When you submit your notebook, Kaggle will rerun your code on the private dataset.\" What I understood about the private dataset is that, if not mistaken, it is a subset of the public training data that will be used to run our submitted notebooks, and ranking will be determined based on the outcome of it.",
      "votes": null
    }
  ],
  "comments": [
    {
      "id": 985065,
      "author_name": "nghweigeok",
      "author_url": "",
      "post_date": "08/25/2020 13:33:47",
      "content": "<p>This is my first time participating in a Kaggle competition and I was asking myself the same question. Based on what is written in the Data section of the competition \"When you submit your notebook, Kaggle will rerun your code on the private dataset.\" What I understood about the private dataset is that, if not mistaken, it is a subset of the public training data that will be used to run our submitted notebooks, and ranking will be determined based on the outcome of it.</p>",
      "votes": null,
      "replies": []
    },
    {
      "id": 962170,
      "author_name": "usmannkhan",
      "author_url": "",
      "post_date": "08/07/2020 21:43:16",
      "content": "<p><a href=\"https://www.kaggle.com/c/landmark-recognition-2020/discussion/173111\">https://www.kaggle.com/c/landmark-recognition-2020/discussion/173111</a></p>",
      "votes": null,
      "replies": []
    }
  ],
  "raw_markdown_by_id": {
    "961102": "It's not clear to me what the private training data (~100k examples) referred to in the description is. The train.csv is over a million rows and all the image files for those appear to be in the dataset. \n\nApologies if it's obvious! First time participating in one of these landmark competitions.",
    "962170": "https://www.kaggle.com/c/landmark-recognition-2020/discussion/173111",
    "985065": "This is my first time participating in a Kaggle competition and I was asking myself the same question. Based on what is written in the Data section of the competition \"When you submit your notebook, Kaggle will rerun your code on the private dataset.\" What I understood about the private dataset is that, if not mistaken, it is a subset of the public training data that will be used to run our submitted notebooks, and ranking will be determined based on the outcome of it."
  },
  "source": "meta"
}