{
  "id": 238584,
  "title": "A decimal point finish?",
  "url": "/competitions/seti-breakthrough-listen/discussion/238584",
  "author_name": "Kamal Das",
  "post_date": "2021-05-12T17:11:52.544000",
  "votes": 3,
  "comment_count": 8,
  "views": 0,
  "content": "<p>A few days in, the Public LB is at 0.97 and public code is for 0.95<br>\nchasing a perfect score of 1.00 ; <br>\nlooks like doing well in this competition will be about grinding out small metric score differences?</p>\n<p>…. and obviously how solid your CV is!! </p>",
  "messages": [
    {
      "id": 1304525,
      "postDate": "2021-05-12T17:11:52.543Z",
      "content": "<p>A few days in, the Public LB is at 0.97 and public code is for 0.95<br>\nchasing a perfect score of 1.00 ; <br>\nlooks like doing well in this competition will be about grinding out small metric score differences?</p>\n<p>…. and obviously how solid your CV is!! </p>",
      "rawMarkdown": "A few days in, the Public LB is at 0.97 and public code is for 0.95\nchasing a perfect score of 1.00 ; \nlooks like doing well in this competition will be about grinding out small metric score differences?\n\n.... and obviously how solid your CV is!! \n\n",
      "votes": 3
    },
    {
      "id": 1304532,
      "postDate": "2021-05-12T17:19:32.673Z",
      "content": "<p><a href=\"https://www.kaggle.com/kmldas\" target=\"_blank\">@kmldas</a> <a href=\"https://www.kaggle.com/maggiemd\" target=\"_blank\">@maggiemd</a> - maybe we could persuade Kaggle to show more than 2dp on the LB for this one?</p>",
      "rawMarkdown": "@kmldas @maggiemd - maybe we could persuade Kaggle to show more than 2dp on the LB for this one?",
      "votes": 2,
      "replies": [
        {
          "id": 1304536,
          "postDate": "2021-05-12T17:24:21.257Z",
          "content": "<p>Sure, to enable LB probing. ;)</p>",
          "rawMarkdown": "Sure, to enable LB probing. ;)",
          "votes": 1
        },
        {
          "id": 1304607,
          "postDate": "2021-05-12T18:23:29.020Z",
          "content": "<p>hi <a href=\"https://www.kaggle.com/jbomitchell\" target=\"_blank\">@jbomitchell</a> , <br>\n<a href=\"https://www.kaggle.com/cpmpml\" target=\"_blank\">@cpmpml</a> has highlighted the reasons for limiting scores decimal points in another thread \"Limiting the display accuracy is meant to prevent LB probing. I doubt it will be changed. If you want to blame someone for this low digit display then blame me as I am the one who suggested limited display accuracy to Kaggle…\"<br>\n<a href=\"https://www.kaggle.com/c/birdclef-2021/discussion/232913\" target=\"_blank\">https://www.kaggle.com/c/birdclef-2021/discussion/232913</a></p>\n<p>Also, the results of the last few competitions show that public score is not a great indicator! many including me have been hit with differences in image size and not debugging properly</p>\n<p>We should trust our CV…. and Theo Viel's suggestion in the below thread are very helpful: <a href=\"https://www.kaggle.com/c/hubmap-kidney-segmentation/discussion/238150\" target=\"_blank\">https://www.kaggle.com/c/hubmap-kidney-segmentation/discussion/238150</a> </p>",
          "rawMarkdown": "hi @jbomitchell , \n@cpmpml has highlighted the reasons for limiting scores decimal points in another thread \"Limiting the display accuracy is meant to prevent LB probing. I doubt it will be changed. If you want to blame someone for this low digit display then blame me as I am the one who suggested limited display accuracy to Kaggle…\"\nhttps://www.kaggle.com/c/birdclef-2021/discussion/232913\n\nAlso, the results of the last few competitions show that public score is not a great indicator! many including me have been hit with differences in image size and not debugging properly\n\nWe should trust our CV.... and Theo Viel's suggestion in the below thread are very helpful: https://www.kaggle.com/c/hubmap-kidney-segmentation/discussion/238150 ",
          "votes": 4
        },
        {
          "id": 1304630,
          "postDate": "2021-05-12T18:42:03.950Z",
          "content": "<p>I believe that internal validation vs LB correlation, or lack thereof, has been discussed in a number of other places, probably including shopping arcades in China. When 20 people are on 1.00 and the next 150 on 0.99 (which may be soon given the rocket powered start to this competition), we will all be flying blind - apart from the \"Did that submission move me up the queue of Kagglers on identical scores\" metric (and <a href=\"https://www.kaggle.com/kmldas\" target=\"_blank\">@kmldas</a> on that measure you are currently two places further up the line than me).</p>\n<p>\"Your submission scored 1.00, which is not an improvement of your best score. Keep trying!\"</p>",
          "rawMarkdown": "I believe that internal validation vs LB correlation, or lack thereof, has been discussed in a number of other places, probably including shopping arcades in China. When 20 people are on 1.00 and the next 150 on 0.99 (which may be soon given the rocket powered start to this competition), we will all be flying blind - apart from the \"Did that submission move me up the queue of Kagglers on identical scores\" metric (and @kmldas on that measure you are currently two places further up the line than me).\n\n\"Your submission scored 1.00, which is not an improvement of your best score. Keep trying!\"",
          "votes": 1
        },
        {
          "id": 1304639,
          "postDate": "2021-05-12T18:48:05.230Z",
          "content": "<blockquote>\n  <p>\"Your submission scored 1.00, which is not an improvement of your best score. Keep trying!\" </p>\n</blockquote>\n<p>😄 I can definitely see that.. and hoping to see you at 1.00 <a href=\"https://www.kaggle.com/jbomitchell\" target=\"_blank\">@jbomitchell</a> !! My best wishes!! </p>\n<p>agree!  the close competition and lack of Public LB validation with limited decimal points is a challenge. To be honest, I am surprised I got this far 2 days into a competition.  And am unsure what more I can do … </p>\n<p>However, unlikely we will get more decimal points so no use wondering about it. <br>\nBest we may do is focus on CV and not be lured by ensembling which is already a top 10 contender in Public LB…</p>",
          "rawMarkdown": "> \"Your submission scored 1.00, which is not an improvement of your best score. Keep trying!\" \n\n😄 I can definitely see that.. and hoping to see you at 1.00 @jbomitchell !! My best wishes!! \n\nagree!  the close competition and lack of Public LB validation with limited decimal points is a challenge. To be honest, I am surprised I got this far 2 days into a competition.  And am unsure what more I can do ... \n\nHowever, unlikely we will get more decimal points so no use wondering about it. \nBest we may do is focus on CV and not be lured by ensembling which is already a top 10 contender in Public LB...\n",
          "votes": 1
        },
        {
          "id": 1304704,
          "postDate": "2021-05-12T19:31:28.117Z",
          "content": "<p>Also, a similar request here: <a href=\"https://www.kaggle.com/c/seti-breakthrough-listen/discussion/238270\" target=\"_blank\">https://www.kaggle.com/c/seti-breakthrough-listen/discussion/238270</a></p>",
          "rawMarkdown": "Also, a similar request here: https://www.kaggle.com/c/seti-breakthrough-listen/discussion/238270",
          "votes": 2
        },
        {
          "id": 1304739,
          "postDate": "2021-05-12T20:31:27.403Z",
          "content": "<blockquote>\n  <p>When 20 people are on 1.00 and the next 150 on 0.99 (which may be soon given the rocket powered start to this competition), we will all be flying blind </p>\n</blockquote>\n<p>Why would this be a problem?</p>\n<p>Using feedback from test set is one of the worst practice in real life machine learning.  It is a good thing if it is useless here.</p>",
          "rawMarkdown": "> When 20 people are on 1.00 and the next 150 on 0.99 (which may be soon given the rocket powered start to this competition), we will all be flying blind \n\nWhy would this be a problem?\n\nUsing feedback from test set is one of the worst practice in real life machine learning.  It is a good thing if it is useless here.",
          "votes": 3
        },
        {
          "id": 1304767,
          "postDate": "2021-05-12T21:13:42.863Z",
          "content": "<p>[Completely rewritten] What we are being asked to do is essentially to assign every true instance a higher score than any false instance. While AUC is a decent measure for weak predictors, at the level of accuracy we're seeing here, the AUC to 2dp is a very coarse measure of the number of misranked instances. If I were refereeing  a paper or thesis that used this metric and precision for predictors like these, I would insist that it was changed. </p>",
          "rawMarkdown": "[Completely rewritten] What we are being asked to do is essentially to assign every true instance a higher score than any false instance. While AUC is a decent measure for weak predictors, at the level of accuracy we're seeing here, the AUC to 2dp is a very coarse measure of the number of misranked instances. If I were refereeing  a paper or thesis that used this metric and precision for predictors like these, I would insist that it was changed. ",
          "votes": 2
        }
      ]
    }
  ],
  "comments": [
    {
      "id": 1304532,
      "author_name": "John Mitchell",
      "author_url": "",
      "post_date": "2021-05-12T17:19:32.673000",
      "content": "<p><a href=\"https://www.kaggle.com/kmldas\" target=\"_blank\">@kmldas</a> <a href=\"https://www.kaggle.com/maggiemd\" target=\"_blank\">@maggiemd</a> - maybe we could persuade Kaggle to show more than 2dp on the LB for this one?</p>",
      "votes": 2,
      "replies": [
        {
          "id": 1304536,
          "author_name": "CPMP",
          "author_url": "",
          "post_date": "2021-05-12T17:24:21.257000",
          "content": "<p>Sure, to enable LB probing. ;)</p>",
          "votes": 1,
          "replies": []
        },
        {
          "id": 1304607,
          "author_name": "Kamal Das",
          "author_url": "",
          "post_date": "2021-05-12T18:23:29.020000",
          "content": "<p>hi <a href=\"https://www.kaggle.com/jbomitchell\" target=\"_blank\">@jbomitchell</a> , <br>\n<a href=\"https://www.kaggle.com/cpmpml\" target=\"_blank\">@cpmpml</a> has highlighted the reasons for limiting scores decimal points in another thread \"Limiting the display accuracy is meant to prevent LB probing. I doubt it will be changed. If you want to blame someone for this low digit display then blame me as I am the one who suggested limited display accuracy to Kaggle…\"<br>\n<a href=\"https://www.kaggle.com/c/birdclef-2021/discussion/232913\" target=\"_blank\">https://www.kaggle.com/c/birdclef-2021/discussion/232913</a></p>\n<p>Also, the results of the last few competitions show that public score is not a great indicator! many including me have been hit with differences in image size and not debugging properly</p>\n<p>We should trust our CV…. and Theo Viel's suggestion in the below thread are very helpful: <a href=\"https://www.kaggle.com/c/hubmap-kidney-segmentation/discussion/238150\" target=\"_blank\">https://www.kaggle.com/c/hubmap-kidney-segmentation/discussion/238150</a> </p>",
          "votes": 4,
          "replies": []
        },
        {
          "id": 1304630,
          "author_name": "John Mitchell",
          "author_url": "",
          "post_date": "2021-05-12T18:42:03.950000",
          "content": "<p>I believe that internal validation vs LB correlation, or lack thereof, has been discussed in a number of other places, probably including shopping arcades in China. When 20 people are on 1.00 and the next 150 on 0.99 (which may be soon given the rocket powered start to this competition), we will all be flying blind - apart from the \"Did that submission move me up the queue of Kagglers on identical scores\" metric (and <a href=\"https://www.kaggle.com/kmldas\" target=\"_blank\">@kmldas</a> on that measure you are currently two places further up the line than me).</p>\n<p>\"Your submission scored 1.00, which is not an improvement of your best score. Keep trying!\"</p>",
          "votes": 1,
          "replies": []
        },
        {
          "id": 1304639,
          "author_name": "Kamal Das",
          "author_url": "",
          "post_date": "2021-05-12T18:48:05.230000",
          "content": "<blockquote>\n  <p>\"Your submission scored 1.00, which is not an improvement of your best score. Keep trying!\" </p>\n</blockquote>\n<p>😄 I can definitely see that.. and hoping to see you at 1.00 <a href=\"https://www.kaggle.com/jbomitchell\" target=\"_blank\">@jbomitchell</a> !! My best wishes!! </p>\n<p>agree!  the close competition and lack of Public LB validation with limited decimal points is a challenge. To be honest, I am surprised I got this far 2 days into a competition.  And am unsure what more I can do … </p>\n<p>However, unlikely we will get more decimal points so no use wondering about it. <br>\nBest we may do is focus on CV and not be lured by ensembling which is already a top 10 contender in Public LB…</p>",
          "votes": 1,
          "replies": []
        },
        {
          "id": 1304704,
          "author_name": "Kamal Das",
          "author_url": "",
          "post_date": "2021-05-12T19:31:28.117000",
          "content": "<p>Also, a similar request here: <a href=\"https://www.kaggle.com/c/seti-breakthrough-listen/discussion/238270\" target=\"_blank\">https://www.kaggle.com/c/seti-breakthrough-listen/discussion/238270</a></p>",
          "votes": 2,
          "replies": []
        },
        {
          "id": 1304739,
          "author_name": "CPMP",
          "author_url": "",
          "post_date": "2021-05-12T20:31:27.403000",
          "content": "<blockquote>\n  <p>When 20 people are on 1.00 and the next 150 on 0.99 (which may be soon given the rocket powered start to this competition), we will all be flying blind </p>\n</blockquote>\n<p>Why would this be a problem?</p>\n<p>Using feedback from test set is one of the worst practice in real life machine learning.  It is a good thing if it is useless here.</p>",
          "votes": 3,
          "replies": []
        },
        {
          "id": 1304767,
          "author_name": "John Mitchell",
          "author_url": "",
          "post_date": "2021-05-12T21:13:42.863000",
          "content": "<p>[Completely rewritten] What we are being asked to do is essentially to assign every true instance a higher score than any false instance. While AUC is a decent measure for weak predictors, at the level of accuracy we're seeing here, the AUC to 2dp is a very coarse measure of the number of misranked instances. If I were refereeing  a paper or thesis that used this metric and precision for predictors like these, I would insist that it was changed. </p>",
          "votes": 2,
          "replies": []
        }
      ]
    }
  ],
  "raw_markdown_by_id": {
    "1304525": "A few days in, the Public LB is at 0.97 and public code is for 0.95\nchasing a perfect score of 1.00 ; \nlooks like doing well in this competition will be about grinding out small metric score differences?\n\n.... and obviously how solid your CV is!! \n\n",
    "1304532": "@kmldas @maggiemd - maybe we could persuade Kaggle to show more than 2dp on the LB for this one?"
  }
}