{
  "id": 307254,
  "title": "Are there any plan for publishing test data set after competition ended?",
  "url": "/competitions/tensorflow-great-barrier-reef/discussion/307254",
  "author_name": "",
  "post_date": "2022-02-13T12:43:47.641015400Z",
  "votes": 21,
  "comment_count": 5,
  "views": 0,
  "content": "<p><a href=\"https://www.kaggle.com/addisonhoward\" target=\"_blank\">@addisonhoward</a> <br>\n<a href=\"https://www.kaggle.com/sohier\" target=\"_blank\">@sohier</a> </p>\n<p>Dear Kaggle staff,</p>\n<p>I want to ask you if it is possible to publish test data set when competition is ended.</p>\n<p>I guess most of the participants (including me) have been struggled to find out why infer with high resolution of input images produces high score on LB.</p>\n<p>If possible, I want to investigate it after the competition.</p>\n<p>Thanks.</p>",
  "messages": [
    {
      "id": "1688172",
      "postDate": "02/13/2022 12:43:47",
      "content": "<p><a href=\"https://www.kaggle.com/addisonhoward\" target=\"_blank\">@addisonhoward</a> <br>\n<a href=\"https://www.kaggle.com/sohier\" target=\"_blank\">@sohier</a> </p>\n<p>Dear Kaggle staff,</p>\n<p>I want to ask you if it is possible to publish test data set when competition is ended.</p>\n<p>I guess most of the participants (including me) have been struggled to find out why infer with high resolution of input images produces high score on LB.</p>\n<p>If possible, I want to investigate it after the competition.</p>\n<p>Thanks.</p>",
      "rawMarkdown": "addisonhoward \n@sohier \n\nDear Kaggle staff,\n\nI want to ask you if it is possible to publish test data set when competition is ended.\n\nI guess most of the participants (including me) have been struggled to find out why infer with high resolution of input images produces high score on LB.\n\nIf possible, I want to investigate it after the competition.\n\nThanks.",
      "votes": null
    },
    {
      "id": "1690092",
      "postDate": "02/14/2022 17:25:07",
      "content": "<p>The Competition hosts ultimately have the final say if they'd like to release the final test data set, but we typically prefer it not be released - primarily to allow participants to continue to make late submissions to the competition without the ground truth being fully available to impact models that are being tested against an unseen test set.</p>\n<p>This is the <em>preferable</em> situation for us, but not a hard requirement we hold - however the hosts will need to make that decision.</p>",
      "rawMarkdown": "The Competition hosts ultimately have the final say if they'd like to release the final test data set, but we typically prefer it not be released - primarily to allow participants to continue to make late submissions to the competition without the ground truth being fully available to impact models that are being tested against an unseen test set.\n\nThis is the *preferable* situation for us, but not a hard requirement we hold - however the hosts will need to make that decision.",
      "votes": null
    },
    {
      "id": "1690695",
      "postDate": "02/15/2022 04:11:05",
      "content": "<p>Hi <a href=\"https://www.kaggle.com/tatamikenn\" target=\"_blank\">@tatamikenn</a> , I'm part of CSIRO's host team. Currently, the most likely case going forward is that we will keep the private test set private for the reason Addison mentioned, but we do plan to release more annotated videos/datasets which can be used as unseen test sets.</p>",
      "rawMarkdown": "Hi @tatamikenn , I'm part of CSIRO's host team. Currently, the most likely case going forward is that we will keep the private test set private for the reason Addison mentioned, but we do plan to release more annotated videos/datasets which can be used as unseen test sets.",
      "votes": null
    },
    {
      "id": "1691208",
      "postDate": "02/15/2022 09:47:42",
      "content": "<p>Wonderful. It’ll definitely help Kaggle community if it is released.<br>\nThank you for the reply.</p>",
      "rawMarkdown": "Wonderful. It’ll definitely help Kaggle community if it is released.\nThank you for the reply.",
      "votes": null
    },
    {
      "id": "1691217",
      "postDate": "02/15/2022 09:56:34",
      "content": "<p>By the way, I don’t know if CSIRO team noticed, we (some participants in the competition including me) found a lot of missing labels in the train dataset.<br>\nIf you are planning to release another dataset, I would recommend to label aided with more accurate model (e.g. this competition’s winner models) than the model you used to create the train labels. It will provide further accurate models if they are trained with it.</p>",
      "rawMarkdown": "By the way, I don’t know if CSIRO team noticed, we (some participants in the competition including me) found a lot of missing labels in the train dataset.\nIf you are planning to release another dataset, I would recommend to label aided with more accurate model (e.g. this competition’s winner models) than the model you used to create the train labels. It will provide further accurate models if they are trained with it.",
      "votes": null
    },
    {
      "id": "1691249",
      "postDate": "02/15/2022 10:06:41",
      "content": "<p>I somewhat understand the reasoning, but as to this competition dataset, I suspect the labels are not accurate enough. I, as well as some participants, noticed a lot of missing labels in the train dataset.<br>\nIf it is also true for the test labels, training to fit unseen inaccurate labels would only produce suboptimal experiment results. So I think it is more valuable to disclose entire dataset, enabling to check if the models prediction is truly wrong or not.</p>\n<p>However, this is the matter the host decides.</p>",
      "rawMarkdown": "I somewhat understand the reasoning, but as to this competition dataset, I suspect the labels are not accurate enough. I, as well as some participants, noticed a lot of missing labels in the train dataset.\nIf it is also true for the test labels, training to fit unseen inaccurate labels would only produce suboptimal experiment results. So I think it is more valuable to disclose entire dataset, enabling to check if the models prediction is truly wrong or not.\n\nHowever, this is the matter the host decides.",
      "votes": null
    }
  ],
  "comments": [
    {
      "id": 1690092,
      "author_name": "addisonhoward",
      "author_url": "",
      "post_date": "02/14/2022 17:25:07",
      "content": "<p>The Competition hosts ultimately have the final say if they'd like to release the final test data set, but we typically prefer it not be released - primarily to allow participants to continue to make late submissions to the competition without the ground truth being fully available to impact models that are being tested against an unseen test set.</p>\n<p>This is the <em>preferable</em> situation for us, but not a hard requirement we hold - however the hosts will need to make that decision.</p>",
      "votes": null,
      "replies": [
        {
          "id": 1691249,
          "author_name": "tatamikenn",
          "author_url": "",
          "post_date": "02/15/2022 10:06:41",
          "content": "<p>I somewhat understand the reasoning, but as to this competition dataset, I suspect the labels are not accurate enough. I, as well as some participants, noticed a lot of missing labels in the train dataset.<br>\nIf it is also true for the test labels, training to fit unseen inaccurate labels would only produce suboptimal experiment results. So I think it is more valuable to disclose entire dataset, enabling to check if the models prediction is truly wrong or not.</p>\n<p>However, this is the matter the host decides.</p>",
          "votes": null,
          "replies": []
        }
      ]
    },
    {
      "id": 1690695,
      "author_name": "ryanliu2021",
      "author_url": "",
      "post_date": "02/15/2022 04:11:05",
      "content": "<p>Hi <a href=\"https://www.kaggle.com/tatamikenn\" target=\"_blank\">@tatamikenn</a> , I'm part of CSIRO's host team. Currently, the most likely case going forward is that we will keep the private test set private for the reason Addison mentioned, but we do plan to release more annotated videos/datasets which can be used as unseen test sets.</p>",
      "votes": null,
      "replies": [
        {
          "id": 1691208,
          "author_name": "tatamikenn",
          "author_url": "",
          "post_date": "02/15/2022 09:47:42",
          "content": "<p>Wonderful. It’ll definitely help Kaggle community if it is released.<br>\nThank you for the reply.</p>",
          "votes": null,
          "replies": []
        },
        {
          "id": 1691217,
          "author_name": "tatamikenn",
          "author_url": "",
          "post_date": "02/15/2022 09:56:34",
          "content": "<p>By the way, I don’t know if CSIRO team noticed, we (some participants in the competition including me) found a lot of missing labels in the train dataset.<br>\nIf you are planning to release another dataset, I would recommend to label aided with more accurate model (e.g. this competition’s winner models) than the model you used to create the train labels. It will provide further accurate models if they are trained with it.</p>",
          "votes": null,
          "replies": []
        }
      ]
    }
  ],
  "raw_markdown_by_id": {
    "1688172": "addisonhoward \n@sohier \n\nDear Kaggle staff,\n\nI want to ask you if it is possible to publish test data set when competition is ended.\n\nI guess most of the participants (including me) have been struggled to find out why infer with high resolution of input images produces high score on LB.\n\nIf possible, I want to investigate it after the competition.\n\nThanks.",
    "1690092": "The Competition hosts ultimately have the final say if they'd like to release the final test data set, but we typically prefer it not be released - primarily to allow participants to continue to make late submissions to the competition without the ground truth being fully available to impact models that are being tested against an unseen test set.\n\nThis is the *preferable* situation for us, but not a hard requirement we hold - however the hosts will need to make that decision.",
    "1690695": "Hi @tatamikenn , I'm part of CSIRO's host team. Currently, the most likely case going forward is that we will keep the private test set private for the reason Addison mentioned, but we do plan to release more annotated videos/datasets which can be used as unseen test sets.",
    "1691208": "Wonderful. It’ll definitely help Kaggle community if it is released.\nThank you for the reply.",
    "1691217": "By the way, I don’t know if CSIRO team noticed, we (some participants in the competition including me) found a lot of missing labels in the train dataset.\nIf you are planning to release another dataset, I would recommend to label aided with more accurate model (e.g. this competition’s winner models) than the model you used to create the train labels. It will provide further accurate models if they are trained with it.",
    "1691249": "I somewhat understand the reasoning, but as to this competition dataset, I suspect the labels are not accurate enough. I, as well as some participants, noticed a lot of missing labels in the train dataset.\nIf it is also true for the test labels, training to fit unseen inaccurate labels would only produce suboptimal experiment results. So I think it is more valuable to disclose entire dataset, enabling to check if the models prediction is truly wrong or not.\n\nHowever, this is the matter the host decides."
  },
  "source": "meta"
}