{
  "id": 61750,
  "title": "Performance gap between the offline validation and the online test",
  "url": "/competitions/google-ai-open-images-object-detection-track/discussion/61750",
  "author_name": "ChristianFans",
  "post_date": "2018-07-23T04:14:54.414000",
  "votes": 3,
  "comment_count": 13,
  "views": 0,
  "content": "<p>I followed&nbsp;<a href=\"https://github.com/tensorflow/models/blob/master/research/object_detection/g3doc/challenge_evaluation.md#object-detection-track%C2%A0to\">https://github.com/tensorflow/models/blob/master/research/object_detection/g3doc/challenge_evaluation.md#object-detection-track&nbsp;to</a> evaluate the validation mAP and found the validation mAP Score is much higher than the online test OpenImagesObjectDetectionAP(about 0.10 mAP). </p>\n\n<p>And I am sure the training set and validation set that I used is excatly the same with the offical recommeded split. So the gap between the validation mAP and the online test mAP really bothers me. </p>\n\n<p>Has anyone else encountered the same place?</p>",
  "messages": [
    {
      "id": 360706,
      "postDate": "2018-07-23T04:14:54.413Z",
      "content": "<p>I followed&nbsp;<a href=\"https://github.com/tensorflow/models/blob/master/research/object_detection/g3doc/challenge_evaluation.md#object-detection-track%C2%A0to\">https://github.com/tensorflow/models/blob/master/research/object_detection/g3doc/challenge_evaluation.md#object-detection-track&nbsp;to</a> evaluate the validation mAP and found the validation mAP Score is much higher than the online test OpenImagesObjectDetectionAP(about 0.10 mAP). </p>\n\n<p>And I am sure the training set and validation set that I used is excatly the same with the offical recommeded split. So the gap between the validation mAP and the online test mAP really bothers me. </p>\n\n<p>Has anyone else encountered the same place?</p>",
      "rawMarkdown": "I followed&nbsp;https://github.com/tensorflow/models/blob/master/research/object_detection/g3doc/challenge_evaluation.md#object-detection-track&nbsp;to evaluate the validation mAP and found the validation mAP Score is much higher than the online test OpenImagesObjectDetectionAP(about 0.10 mAP). \n\nAnd I am sure the training set and validation set that I used is excatly the same with the offical recommeded split. So the gap between the validation mAP and the online test mAP really bothers me. \n\nHas anyone else encountered the same place?",
      "votes": 3
    },
    {
      "id": 367061,
      "postDate": "2018-08-07T02:14:35.583Z",
      "content": "<p>same problem~\nHave you solved it ?\nlooking for your reply~</p>",
      "rawMarkdown": "same problem~\nHave you solved it ?\nlooking for your reply~",
      "replies": [
        {
          "id": 367604,
          "postDate": "2018-08-08T04:41:05.233Z",
          "content": "<p>The problem still exists, I think it is possibly a inherent property of the dataset.</p>",
          "rawMarkdown": "The problem still exists, I think it is possibly a inherent property of the dataset."
        },
        {
          "id": 369465,
          "postDate": "2018-08-13T05:40:48.830Z",
          "content": "<p>How much is your mAP gap between the two sets?</p>",
          "rawMarkdown": "How much is your mAP gap between the two sets?"
        },
        {
          "id": 369549,
          "postDate": "2018-08-13T10:26:09.360Z",
          "content": "<p>0.1map</p>",
          "rawMarkdown": "0.1map"
        }
      ]
    },
    {
      "id": 364301,
      "postDate": "2018-07-31T08:57:44.477Z",
      "content": "<p>Hi, where did you download your validation dataset? \nMy validation set was followed by the official recommended and have 100000 images, and I get nearly correct result~</p>",
      "rawMarkdown": "Hi, where did you download your validation dataset? \nMy validation set was followed by the official recommended and have 100000 images, and I get nearly correct result~",
      "replies": [
        {
          "id": 364363,
          "postDate": "2018-07-31T11:34:28.343Z",
          "content": "<p>I exactly use the official recommended validation set, and I follow the tutorial at <a href=\"https://github.com/tensorflow/models/blob/master/research/object_detection/g3doc/challenge_evaluation.md\">https://github.com/tensorflow/models/blob/master/research/object_detection/g3doc/challenge_evaluation.md</a>, I expand the bounding box annotations and image label annotations as illustrated. And my prediction list each bounding box in a new line with (ImageID,LabelName,Score,XMin,YMin,XMax,YMax), then I run the latest TF ObjectDetection API to evaluate it,  concretely, we run the  oid_od_challenge_evaluation.py script as advised, then I get a csv file and find that my result of validation is 0.10 higher compared to the test result on leaderboard. Any detailed difference with your evaluation process?</p>",
          "rawMarkdown": "I exactly use the official recommended validation set, and I follow the tutorial at https://github.com/tensorflow/models/blob/master/research/object_detection/g3doc/challenge_evaluation.md, I expand the bounding box annotations and image label annotations as illustrated. And my prediction list each bounding box in a new line with (ImageID,LabelName,Score,XMin,YMin,XMax,YMax), then I run the latest TF ObjectDetection API to evaluate it,  concretely, we run the  oid_od_challenge_evaluation.py script as advised, then I get a csv file and find that my result of validation is 0.10 higher compared to the test result on leaderboard. Any detailed difference with your evaluation process?",
          "votes": 1
        },
        {
          "id": 366780,
          "postDate": "2018-08-06T13:45:50.933Z",
          "content": "<p>Do you use the script for expanding classes for both local validation and submission?</p>",
          "rawMarkdown": "Do you use the script for expanding classes for both local validation and submission?"
        },
        {
          "id": 366788,
          "postDate": "2018-08-06T14:11:38.150Z",
          "content": "<p>Yes, I have expanded the bounding boxes label and images label for both validation and test set with the offered python scripts</p>",
          "rawMarkdown": "Yes, I have expanded the bounding boxes label and images label for both validation and test set with the offered python scripts",
          "votes": 1
        },
        {
          "id": 367063,
          "postDate": "2018-08-07T02:19:06.790Z",
          "content": "<p>I got the same problem...\nsomeone help?</p>",
          "rawMarkdown": "I got the same problem...\nsomeone help?"
        }
      ]
    },
    {
      "id": 362418,
      "postDate": "2018-07-26T10:07:29.617Z",
      "content": "<p>I found 10 points diff in Score metric, not mAP.</p>",
      "rawMarkdown": "I found 10 points diff in Score metric, not mAP."
    },
    {
      "id": 362416,
      "postDate": "2018-07-26T10:05:53.473Z",
      "content": "<p>same problem.  any progress?</p>",
      "rawMarkdown": "same problem.  any progress?"
    },
    {
      "id": 369463,
      "postDate": "2018-08-13T05:36:47.943Z",
      "rawMarkdown": "",
      "isDeleted": true,
      "replies": [
        {
          "id": 374851,
          "postDate": "2018-08-24T01:48:49.660Z",
          "content": "<p>你是中国人嘛</p>",
          "rawMarkdown": "你是中国人嘛"
        }
      ]
    }
  ],
  "comments": [
    {
      "id": 367061,
      "author_name": "MagicWolf",
      "author_url": "",
      "post_date": "2018-08-07T02:14:35.583000",
      "content": "<p>same problem~\nHave you solved it ?\nlooking for your reply~</p>",
      "votes": 0,
      "replies": [
        {
          "id": 367604,
          "author_name": "ChristianFans",
          "author_url": "",
          "post_date": "2018-08-08T04:41:05.233000",
          "content": "<p>The problem still exists, I think it is possibly a inherent property of the dataset.</p>",
          "votes": 0,
          "replies": []
        },
        {
          "id": 369465,
          "author_name": "Louis",
          "author_url": "",
          "post_date": "2018-08-13T05:40:48.830000",
          "content": "<p>How much is your mAP gap between the two sets?</p>",
          "votes": 0,
          "replies": []
        },
        {
          "id": 369549,
          "author_name": "MagicWolf",
          "author_url": "",
          "post_date": "2018-08-13T10:26:09.360000",
          "content": "<p>0.1map</p>",
          "votes": 0,
          "replies": []
        }
      ]
    },
    {
      "id": 364301,
      "author_name": "MagicWolf",
      "author_url": "",
      "post_date": "2018-07-31T08:57:44.477000",
      "content": "<p>Hi, where did you download your validation dataset? \nMy validation set was followed by the official recommended and have 100000 images, and I get nearly correct result~</p>",
      "votes": 0,
      "replies": [
        {
          "id": 364363,
          "author_name": "ChristianFans",
          "author_url": "",
          "post_date": "2018-07-31T11:34:28.343000",
          "content": "<p>I exactly use the official recommended validation set, and I follow the tutorial at <a href=\"https://github.com/tensorflow/models/blob/master/research/object_detection/g3doc/challenge_evaluation.md\">https://github.com/tensorflow/models/blob/master/research/object_detection/g3doc/challenge_evaluation.md</a>, I expand the bounding box annotations and image label annotations as illustrated. And my prediction list each bounding box in a new line with (ImageID,LabelName,Score,XMin,YMin,XMax,YMax), then I run the latest TF ObjectDetection API to evaluate it,  concretely, we run the  oid_od_challenge_evaluation.py script as advised, then I get a csv file and find that my result of validation is 0.10 higher compared to the test result on leaderboard. Any detailed difference with your evaluation process?</p>",
          "votes": 1,
          "replies": []
        },
        {
          "id": 366780,
          "author_name": "taraspiotr",
          "author_url": "",
          "post_date": "2018-08-06T13:45:50.933000",
          "content": "<p>Do you use the script for expanding classes for both local validation and submission?</p>",
          "votes": 0,
          "replies": []
        },
        {
          "id": 366788,
          "author_name": "ChristianFans",
          "author_url": "",
          "post_date": "2018-08-06T14:11:38.150000",
          "content": "<p>Yes, I have expanded the bounding boxes label and images label for both validation and test set with the offered python scripts</p>",
          "votes": 1,
          "replies": []
        },
        {
          "id": 367063,
          "author_name": "MagicWolf",
          "author_url": "",
          "post_date": "2018-08-07T02:19:06.790000",
          "content": "<p>I got the same problem...\nsomeone help?</p>",
          "votes": 0,
          "replies": []
        }
      ]
    },
    {
      "id": 362418,
      "author_name": "fromsea",
      "author_url": "",
      "post_date": "2018-07-26T10:07:29.617000",
      "content": "<p>I found 10 points diff in Score metric, not mAP.</p>",
      "votes": 0,
      "replies": []
    },
    {
      "id": 362416,
      "author_name": "fromsea",
      "author_url": "",
      "post_date": "2018-07-26T10:05:53.473000",
      "content": "<p>same problem.  any progress?</p>",
      "votes": 0,
      "replies": []
    },
    {
      "id": 369463,
      "author_name": "",
      "author_url": "",
      "post_date": "2018-08-13T05:36:47.943000",
      "content": "",
      "votes": 0,
      "replies": [
        {
          "id": 374851,
          "author_name": "MagicWolf",
          "author_url": "",
          "post_date": "2018-08-24T01:48:49.660000",
          "content": "<p>你是中国人嘛</p>",
          "votes": 0,
          "replies": []
        }
      ]
    }
  ],
  "raw_markdown_by_id": {
    "360706": "I followed&nbsp;https://github.com/tensorflow/models/blob/master/research/object_detection/g3doc/challenge_evaluation.md#object-detection-track&nbsp;to evaluate the validation mAP and found the validation mAP Score is much higher than the online test OpenImagesObjectDetectionAP(about 0.10 mAP). \n\nAnd I am sure the training set and validation set that I used is excatly the same with the offical recommeded split. So the gap between the validation mAP and the online test mAP really bothers me. \n\nHas anyone else encountered the same place?",
    "367061": "same problem~\nHave you solved it ?\nlooking for your reply~",
    "364301": "Hi, where did you download your validation dataset? \nMy validation set was followed by the official recommended and have 100000 images, and I get nearly correct result~",
    "362418": "I found 10 points diff in Score metric, not mAP.",
    "362416": "same problem.  any progress?",
    "369463": ""
  }
}