{
  "id": 127968,
  "title": "The Size of Private Test Set",
  "url": "/competitions/deepfake-detection-challenge/discussion/127968",
  "author_name": "",
  "post_date": "2020-01-28T05:27:37.907121900Z",
  "votes": 2,
  "comment_count": 2,
  "views": 0,
  "content": "<p>To better simulating LB, I create my personal validation set with 4000 videos (P:N=1:1) with no actors and scenes overlapped. However, when I monitor the Val loss history, I found that the logloss always fluctuates (see above plot).</p>\n\n<p><img src=\"https://www.googleapis.com/download/storage/v1/b/kaggle-user-content/o/inbox%2F1497137%2Fde357e72e4f224db36b18f69c85ffdcc%2Fval_loss%20(1\" alt=\"\">.svg?generation=1580188608858698&amp;alt=media)</p>\n\n<p>My conclusion is that probably the size of a private test set is too small, and made the LB  lose its power.  Did anyone have better suggestions about building a reliable local validation set?</p>",
  "messages": [
    {
      "id": "730910",
      "postDate": "01/28/2020 05:27:37",
      "content": "<p>To better simulating LB, I create my personal validation set with 4000 videos (P:N=1:1) with no actors and scenes overlapped. However, when I monitor the Val loss history, I found that the logloss always fluctuates (see above plot).</p>\n\n<p><img src=\"https://www.googleapis.com/download/storage/v1/b/kaggle-user-content/o/inbox%2F1497137%2Fde357e72e4f224db36b18f69c85ffdcc%2Fval_loss%20(1\" alt=\"\">.svg?generation=1580188608858698&amp;alt=media)</p>\n\n<p>My conclusion is that probably the size of a private test set is too small, and made the LB  lose its power.  Did anyone have better suggestions about building a reliable local validation set?</p>",
      "rawMarkdown": "To better simulating LB, I create my personal validation set with 4000 videos (P:N=1:1) with no actors and scenes overlapped. However, when I monitor the Val loss history, I found that the logloss always fluctuates (see above plot).\n\n![](https://www.googleapis.com/download/storage/v1/b/kaggle-user-content/o/inbox%2F1497137%2Fde357e72e4f224db36b18f69c85ffdcc%2Fval_loss%20(1).svg?generation=1580188608858698&amp;alt=media)\n\nMy conclusion is that probably the size of a private test set is too small, and made the LB  lose its power.  Did anyone have better suggestions about building a reliable local validation set?",
      "votes": null
    },
    {
      "id": "731049",
      "postDate": "01/28/2020 09:42:52",
      "content": "<p>Did you build this validation set from the provided training data or from some external dataset? (If from the provided training data, I don't think you can have 4000 videos with no actors overlapping, since there are only 100 or so different actors in the dataset.)</p>",
      "rawMarkdown": "Did you build this validation set from the provided training data or from some external dataset? (If from the provided training data, I don't think you can have 4000 videos with no actors overlapping, since there are only 100 or so different actors in the dataset.)",
      "votes": null
    },
    {
      "id": "731514",
      "postDate": "01/28/2020 18:24:07",
      "content": "<p>Sorry to confuse you, what I mean no overlapping actors and scenes are based on the combination of them. And also due to my clustering methods, some scenes and actors might be divided into different groups.</p>",
      "rawMarkdown": "Sorry to confuse you, what I mean no overlapping actors and scenes are based on the combination of them. And also due to my clustering methods, some scenes and actors might be divided into different groups.",
      "votes": null
    }
  ],
  "comments": [
    {
      "id": 731049,
      "author_name": "humananalog",
      "author_url": "",
      "post_date": "01/28/2020 09:42:52",
      "content": "<p>Did you build this validation set from the provided training data or from some external dataset? (If from the provided training data, I don't think you can have 4000 videos with no actors overlapping, since there are only 100 or so different actors in the dataset.)</p>",
      "votes": null,
      "replies": [
        {
          "id": 731514,
          "author_name": "wufanyou",
          "author_url": "",
          "post_date": "01/28/2020 18:24:07",
          "content": "<p>Sorry to confuse you, what I mean no overlapping actors and scenes are based on the combination of them. And also due to my clustering methods, some scenes and actors might be divided into different groups.</p>",
          "votes": null,
          "replies": []
        }
      ]
    }
  ],
  "raw_markdown_by_id": {
    "730910": "To better simulating LB, I create my personal validation set with 4000 videos (P:N=1:1) with no actors and scenes overlapped. However, when I monitor the Val loss history, I found that the logloss always fluctuates (see above plot).\n\n![](https://www.googleapis.com/download/storage/v1/b/kaggle-user-content/o/inbox%2F1497137%2Fde357e72e4f224db36b18f69c85ffdcc%2Fval_loss%20(1).svg?generation=1580188608858698&amp;alt=media)\n\nMy conclusion is that probably the size of a private test set is too small, and made the LB  lose its power.  Did anyone have better suggestions about building a reliable local validation set?",
    "731049": "Did you build this validation set from the provided training data or from some external dataset? (If from the provided training data, I don't think you can have 4000 videos with no actors overlapping, since there are only 100 or so different actors in the dataset.)",
    "731514": "Sorry to confuse you, what I mean no overlapping actors and scenes are based on the combination of them. And also due to my clustering methods, some scenes and actors might be divided into different groups."
  },
  "source": "meta"
}