{
  "id": 174515,
  "title": "Another 0.99+ team...",
  "url": "/competitions/siim-isic-melanoma-classification/discussion/174515",
  "author_name": "Dieter",
  "post_date": "2020-08-13T21:13:51.305000",
  "votes": 38,
  "comment_count": 51,
  "views": 0,
  "content": "<p>I can only speculate, but telling from the number of submissions, they just intensively probed image per image on the LB to label the test images which are part of the public test set… Whats the point? While you can use the probed labels as additional training data and probably get a better score, how does this help anyone but your own ego?</p>",
  "messages": [
    {
      "id": 970111,
      "postDate": "2020-08-14T08:03:07.497Z",
      "content": "<p>Hi, please let me clarify as much as I can now. That result is based on LB probing, it comprises many probing submissions, each of the probing submission scores is in range 0.4-0.6 that is why it was not visible before. This submission will be very bad for the private LB, it is basically very overfited to the public. I plan to publish more details on the procedure when the competition ends. I submitted it now to showcase what is possible to achieve with the probing. We didn't submit it earlier in order not to disrupt the leaderboard.</p>\n<p>My motivation: I like solving tasks with mixed integer programming, I worked with it at IBM many years ago and I practise it when appropriate since then. I try to distill as much information as possible from each submission, it is challenging and interesting.</p>\n<p>I understand that the area is a gray area, so I try to maintain full transparency, specifically on how we use it in our final solutions, if at all. Will be glad to share more when the competition ends.</p>\n<p>Good luck to you in the last few days!</p>",
      "rawMarkdown": "Hi, please let me clarify as much as I can now. That result is based on LB probing, it comprises many probing submissions, each of the probing submission scores is in range 0.4-0.6 that is why it was not visible before. This submission will be very bad for the private LB, it is basically very overfited to the public. I plan to publish more details on the procedure when the competition ends. I submitted it now to showcase what is possible to achieve with the probing. We didn't submit it earlier in order not to disrupt the leaderboard.\n\nMy motivation: I like solving tasks with mixed integer programming, I worked with it at IBM many years ago and I practise it when appropriate since then. I try to distill as much information as possible from each submission, it is challenging and interesting.\n\nI understand that the area is a gray area, so I try to maintain full transparency, specifically on how we use it in our final solutions, if at all. Will be glad to share more when the competition ends.\n\nGood luck to you in the last few days!",
      "votes": 33,
      "replies": [
        {
          "id": 970163,
          "postDate": "2020-08-14T09:00:23.573Z",
          "content": "<p>Thanks for clarifying.</p>",
          "rawMarkdown": "Thanks for clarifying."
        },
        {
          "id": 970217,
          "postDate": "2020-08-14T09:31:13.457Z",
          "content": "<blockquote>\n  <p>I like solving tasks with mixed integer programming,</p>\n</blockquote>\n<p>I love it.  Thanks for clarifying.</p>\n<p>Did you use CPLEX while at IBM?  That's what I was working on as you probably know.</p>",
          "rawMarkdown": ">  I like solving tasks with mixed integer programming,\n\nI love it.  Thanks for clarifying.\n\nDid you use CPLEX while at IBM?  That's what I was working on as you probably know.",
          "votes": 2
        },
        {
          "id": 970300,
          "postDate": "2020-08-14T10:46:29.843Z",
          "content": "<p>Yes, I remember you also participated in the latest Santa competition. At IBM we used CPLEX, of course, but during the Santa competition it became clear to us that Gurobi worked better there, as reported also by many other people. So I switched to Gurobi, and here also I used Gurobi.</p>",
          "rawMarkdown": "Yes, I remember you also participated in the latest Santa competition. At IBM we used CPLEX, of course, but during the Santa competition it became clear to us that Gurobi worked better there, as reported also by many other people. So I switched to Gurobi, and here also I used Gurobi.",
          "votes": 1
        },
        {
          "id": 970358,
          "postDate": "2020-08-14T11:36:50.890Z",
          "rawMarkdown": "",
          "isDeleted": true,
          "replies": [
            {
              "id": 970359,
              "postDate": "2020-08-14T11:37:27.910Z",
              "content": "<p><a href=\"https://www.kaggle.com/cpmpml/number-of-public-melanoma-is-78-or-77\" target=\"_blank\">That's already known</a>. 77 or 78 positive images.</p>",
              "rawMarkdown": "[That's already known](https://www.kaggle.com/cpmpml/number-of-public-melanoma-is-78-or-77). 77 or 78 positive images.",
              "votes": 3
            },
            {
              "id": 970374,
              "postDate": "2020-08-14T11:46:48.433Z",
              "rawMarkdown": "",
              "isDeleted": true
            },
            {
              "id": 970440,
              "postDate": "2020-08-14T13:06:02.497Z",
              "content": "<p>It is the best notebook ever.</p>",
              "rawMarkdown": "It is the best notebook ever.",
              "votes": 8
            }
          ]
        },
        {
          "id": 973382,
          "postDate": "2020-08-17T09:20:57.800Z",
          "content": "<p><a href=\"https://www.kaggle.com/zaharch\" target=\"_blank\">@zaharch</a> was also the one who had an impressive public LB score in LANL Earthquake Prediction not long ago.  He did some smart optimization to know public rows of submissions. </p>",
          "rawMarkdown": "@zaharch was also the one who had an impressive public LB score in LANL Earthquake Prediction not long ago.  He did some smart optimization to know public rows of submissions. "
        }
      ]
    },
    {
      "id": 969690,
      "postDate": "2020-08-13T21:13:51.307Z",
      "content": "<p>I can only speculate, but telling from the number of submissions, they just intensively probed image per image on the LB to label the test images which are part of the public test set… Whats the point? While you can use the probed labels as additional training data and probably get a better score, how does this help anyone but your own ego?</p>",
      "rawMarkdown": "I can only speculate, but telling from the number of submissions, they just intensively probed image per image on the LB to label the test images which are part of the public test set... Whats the point? While you can use the probed labels as additional training data and probably get a better score, how does this help anyone but your own ego?",
      "votes": 37
    },
    {
      "id": 969722,
      "postDate": "2020-08-13T22:09:23.657Z",
      "content": "<p>I don't think we should make any kind of allegations or comments without knowing their side of the story.<br>\nIf it's probing it has been done already by a few this competition. <br>\nAs long as the private solution is clean, who cares. </p>",
      "rawMarkdown": "I don't think we should make any kind of allegations or comments without knowing their side of the story.\nIf it's probing it has been done already by a few this competition. \nAs long as the private solution is clean, who cares. \n",
      "votes": 29,
      "replies": [
        {
          "id": 970408,
          "postDate": "2020-08-14T12:38:01.443Z",
          "content": "<p>Totally agreeing with you</p>",
          "rawMarkdown": "Totally agreeing with you"
        }
      ]
    },
    {
      "id": 969719,
      "postDate": "2020-08-13T22:02:13.193Z",
      "content": "<p>Kaggle should make this one a kernel competition.</p>",
      "rawMarkdown": "Kaggle should make this one a kernel competition.",
      "votes": 13,
      "replies": [
        {
          "id": 969731,
          "postDate": "2020-08-13T22:22:27.067Z",
          "content": "<blockquote>\n  <p>Kaggle should make this one a kernel competition.</p>\n</blockquote>\n<p>Or 2 stages competition with 2nd stage test set available just a week before the end. That would still allow internet use (for TPU for instance)<br>\nThere used to be many 2 stages CV competitions before on kaggle( in order to prevent handlabelling among others).  I don't why that's not the case anymore.</p>",
          "rawMarkdown": "> Kaggle should make this one a kernel competition.\n\nOr 2 stages competition with 2nd stage test set available just a week before the end. That would still allow internet use (for TPU for instance)\nThere used to be many 2 stages CV competitions before on kaggle( in order to prevent handlabelling among others).  I don't why that's not the case anymore.",
          "votes": 4
        }
      ]
    },
    {
      "id": 969729,
      "postDate": "2020-08-13T22:18:46.813Z",
      "content": "<p>Two of these things are not like the others 🤔</p>\n<p><img src=\"https://www.googleapis.com/download/storage/v1/b/kaggle-forum-message-attachments/o/inbox%2F644036%2Ff7fc749a373c3ba12da61593d5af3697%2FScreenshot%20from%202020-08-13%2018-17-40.png?generation=1597357117685509&amp;alt=media\" alt=\"\"></p>",
      "rawMarkdown": "Two of these things are not like the others 🤔\n\n![](https://www.googleapis.com/download/storage/v1/b/kaggle-forum-message-attachments/o/inbox%2F644036%2Ff7fc749a373c3ba12da61593d5af3697%2FScreenshot%20from%202020-08-13%2018-17-40.png?generation=1597357117685509&alt=media)",
      "votes": 12,
      "replies": [
        {
          "id": 969747,
          "postDate": "2020-08-13T22:44:56.793Z",
          "content": "<p>Great plot. In the first month of the comp, Yuval was in 1st place using a single model that scored LB 0.960 explained <a href=\"https://www.kaggle.com/c/siim-isic-melanoma-classification/discussion/168152#939957\" target=\"_blank\">here</a>. After Sirish took first, Sirish explained what he did to climb LB <a href=\"https://www.kaggle.com/c/siim-isic-melanoma-classification/discussion/161497#901185\" target=\"_blank\">here</a>. Then NoSound joined Yuval and their LB did not increase for 22 days until today when it jumped up to match Sirish's.</p>",
          "rawMarkdown": "Great plot. In the first month of the comp, Yuval was in 1st place using a single model that scored LB 0.960 explained [here][2]. After Sirish took first, Sirish explained what he did to climb LB [here][1]. Then NoSound joined Yuval and their LB did not increase for 22 days until today when it jumped up to match Sirish's.\n\n[1]: https://www.kaggle.com/c/siim-isic-melanoma-classification/discussion/161497#901185\n[2]: https://www.kaggle.com/c/siim-isic-melanoma-classification/discussion/168152#939957",
          "votes": 7
        },
        {
          "id": 969760,
          "postDate": "2020-08-13T23:03:35.020Z",
          "content": "<blockquote>\n  <p>their LB did not increase for 22 days until today</p>\n</blockquote>\n<p>It shows LB is not required to measure progress.</p>",
          "rawMarkdown": "> their LB did not increase for 22 days until today\n\nIt shows LB is not required to measure progress.",
          "votes": 7
        },
        {
          "id": 970024,
          "postDate": "2020-08-14T06:20:43.227Z",
          "content": "<p>How many subs during the no-progress period?</p>",
          "rawMarkdown": "How many subs during the no-progress period?"
        },
        {
          "id": 970031,
          "postDate": "2020-08-14T06:25:21.370Z",
          "content": "<p>I forget. Can we check that in meta Kaggle dataset? If I remember correctly, i think they did about 5 a day, so about 100 subs.</p>",
          "rawMarkdown": "I forget. Can we check that in meta Kaggle dataset? If I remember correctly, i think they did about 5 a day, so about 100 subs.",
          "votes": 1
        },
        {
          "id": 970038,
          "postDate": "2020-08-14T06:32:28.003Z",
          "content": "<p>That is probably enough to increase 0.960 to 0.990. I think you could check your model's top 3000 predictions by making 100 subs where you set the top <code>pred[30*k,30*(k+1)]=1</code> for <code>k in range(100)</code> (and other predictions to zero). Then you order all those groups by their LB score (keeping them above the bottom 7000). I could run a simulation on OOF doing this procedure to see how much it increases AUC.</p>",
          "rawMarkdown": "That is probably enough to increase 0.960 to 0.990. I think you could check your model's top 3000 predictions by making 100 subs where you set the top `pred[30*k,30*(k+1)]=1` for `k in range(100)` (and other predictions to zero). Then you order all those groups by their LB score (keeping them above the bottom 7000). I could run a simulation on OOF doing this procedure to see how much it increases AUC.",
          "votes": 6,
          "replies": [
            {
              "id": 970159,
              "postDate": "2020-08-14T08:57:49.247Z",
              "content": "<p><a href=\"https://www.kaggle.com/cdeotte\" target=\"_blank\">@cdeotte</a> Could you explain how this is done or could you create a notebook about this one??</p>",
              "rawMarkdown": "@cdeotte Could you explain how this is done or could you create a notebook about this one??\n"
            },
            {
              "id": 970510,
              "postDate": "2020-08-14T14:15:45.710Z",
              "content": "<p>I did a quick simulation of what I describe above using OOF. It's not as effective as I thought. </p>\n<p>If you do 100 probes (size 32 each) your LB would increase from 0.9600 to 0.9650. <br>\nIf you do 200 probes (size 16 each) then your LB would increase from 0.9600 to 0.9675. <br>\nIf you do 400 probes (size 8 each) then your LB would increase from 0.9600 to 0.9715. <br>\nIf you do 800 probes (size 4 each) then your LB would increase from 0.9600 to 0.9800.</p>\n<p>Furthermore, simulation suggests that this procedure would increase public LB but it would simultaneously decrease private LB.</p>",
              "rawMarkdown": "I did a quick simulation of what I describe above using OOF. It's not as effective as I thought. \n\nIf you do 100 probes (size 32 each) your LB would increase from 0.9600 to 0.9650. \nIf you do 200 probes (size 16 each) then your LB would increase from 0.9600 to 0.9675. \nIf you do 400 probes (size 8 each) then your LB would increase from 0.9600 to 0.9715. \nIf you do 800 probes (size 4 each) then your LB would increase from 0.9600 to 0.9800.\n\nFurthermore, simulation suggests that this procedure would increase public LB but it would simultaneously decrease private LB.",
              "votes": 3
            }
          ]
        },
        {
          "id": 970075,
          "postDate": "2020-08-14T07:15:46.937Z",
          "content": "<p>Sounds reasonable <a href=\"https://www.kaggle.com/cdeotte\" target=\"_blank\">@cdeotte</a>. Let's see your simulation results.</p>",
          "rawMarkdown": "Sounds reasonable @cdeotte. Let's see your simulation results.",
          "votes": 1
        }
      ]
    },
    {
      "id": 969694,
      "postDate": "2020-08-13T21:26:06.180Z",
      "content": "<p>I think they jumped from 962 directly to 99x which is somehow weird :)</p>",
      "rawMarkdown": "I think they jumped from 962 directly to 99x which is somehow weird :)",
      "votes": 4,
      "replies": [
        {
          "id": 969713,
          "postDate": "2020-08-13T21:54:05.047Z",
          "content": "<p>I agree, more weird than the 1st.<br>\nPlus, seeing many GMs like you not in the front part makes me feel a shake will happen.</p>",
          "rawMarkdown": "I agree, more weird than the 1st.\nPlus, seeing many GMs like you not in the front part makes me feel a shake will happen.",
          "votes": 4
        },
        {
          "id": 969718,
          "postDate": "2020-08-13T22:00:39.697Z",
          "content": "<p>So big jump is very strange for probing … maybe they found some leak, doctor who labeled image by hand (joke). At least they have good PL and can train model on test dataset which is a big advantage</p>",
          "rawMarkdown": "So big jump is very strange for probing ... maybe they found some leak, doctor who labeled image by hand (joke). At least they have good PL and can train model on test dataset which is a big advantage"
        },
        {
          "id": 969721,
          "postDate": "2020-08-13T22:07:47.447Z",
          "content": "<p>Or they probed images separately and then combined the probe.</p>",
          "rawMarkdown": "Or they probed images separately and then combined the probe.",
          "votes": 5
        },
        {
          "id": 969725,
          "postDate": "2020-08-13T22:13:04.863Z",
          "content": "<p>💯.         </p>",
          "rawMarkdown": "💯.         ",
          "votes": 1
        }
      ]
    },
    {
      "id": 969739,
      "postDate": "2020-08-13T22:31:28.750Z",
      "content": "<blockquote>\n  <p>While you can use the probed labels as additional training data </p>\n</blockquote>\n<p>No you can't.  Hand labeling of test data is prohibited.</p>",
      "rawMarkdown": "> While you can use the probed labels as additional training data \n\nNo you can't.  Hand labeling of test data is prohibited.",
      "votes": 1,
      "replies": [
        {
          "id": 969743,
          "postDate": "2020-08-13T22:38:01.613Z",
          "content": "<p>its not hand labeling, its Public LB labeling :D</p>",
          "rawMarkdown": "its not hand labeling, its Public LB labeling :D",
          "votes": 2
        },
        {
          "id": 969750,
          "postDate": "2020-08-13T22:47:36.483Z",
          "content": "<p>You can completely automate it with a script. A script (without any human intervention) can make a submission to LB and store the result. (A human never sees the result). Then the same script can impart it's knowledge into your model.</p>",
          "rawMarkdown": "You can completely automate it with a script. A script (without any human intervention) can make a submission to LB and store the result. (A human never sees the result). Then the same script can impart it's knowledge into your model.",
          "votes": 3
        },
        {
          "id": 969755,
          "postDate": "2020-08-13T22:58:15.960Z",
          "content": "<p>Well, if this is how they did it then fine.  But I doubt it very much.  And I'm curious to see how host interprets it.</p>",
          "rawMarkdown": "Well, if this is how they did it then fine.  But I doubt it very much.  And I'm curious to see how host interprets it."
        },
        {
          "id": 969758,
          "postDate": "2020-08-13T23:02:08.807Z",
          "content": "<p>There is a lot of gray area here. What if you cluster the images unsupervised into groups of 30. Then you make 333 LB submissions. Afterward is it hand labeling or a decision tree? </p>\n<pre><code> If test image belongs to cluster X then classify as Y\n</code></pre>",
          "rawMarkdown": "There is a lot of gray area here. What if you cluster the images unsupervised into groups of 30. Then you make 333 LB submissions. Afterward is it hand labeling or a decision tree? \n\n     If test image belongs to cluster X then classify as Y",
          "votes": 1
        },
        {
          "id": 969852,
          "postDate": "2020-08-14T02:09:04.773Z",
          "content": "<p>You said it : it is a gray area.</p>\n<p>I interpreted rules to forbid it, but I get that another interpretation is possible.</p>\n<p>Best would be to have Kaggle or host weigh in.</p>",
          "rawMarkdown": "You said it : it is a gray area.\n\nI interpreted rules to forbid it, but I get that another interpretation is possible.\n\nBest would be to have Kaggle or host weigh in.\n",
          "votes": 1
        }
      ]
    },
    {
      "id": 970133,
      "postDate": "2020-08-14T08:30:10.750Z",
      "content": "<p>I think LB probing is fine if someone is curious.  Personally I just could never pass the thought of using up all those submissions, especially toward the end.  It would take alot of submissions to probe to get to .99, and they have two in the team, but still, 385 submissions = 77 days of submissions, so burn then however you like.  The only issue I can see is if someone has made dummy accounts to do the probing, although Kaggle has done a good job of slow playing people like that and then dealing with it when competitions end.  </p>",
      "rawMarkdown": "I think LB probing is fine if someone is curious.  Personally I just could never pass the thought of using up all those submissions, especially toward the end.  It would take alot of submissions to probe to get to .99, and they have two in the team, but still, 385 submissions = 77 days of submissions, so burn then however you like.  The only issue I can see is if someone has made dummy accounts to do the probing, although Kaggle has done a good job of slow playing people like that and then dealing with it when competitions end.  ",
      "votes": 2
    },
    {
      "id": 970041,
      "postDate": "2020-08-14T06:35:03.357Z",
      "content": "<p>Well I’m not assuming that they used probing to get there, but here is a simple set up you can use to probe (or simply check your progress) as much as you want in this competition while staying under the radars:</p>\n<p>Simply take your submission and do 1 - submission before submitting it, then your AUC will be 1-real_AUC. No one will know about how good your models are performing. Once you are coming close to the end of the competition and want to have a real submission to compete on the private LB then submit the real submission file. IMHO this is one possible explanation for such a sudden jump!</p>",
      "rawMarkdown": "Well I’m not assuming that they used probing to get there, but here is a simple set up you can use to probe (or simply check your progress) as much as you want in this competition while staying under the radars:\n\nSimply take your submission and do 1 - submission before submitting it, then your AUC will be 1-real_AUC. No one will know about how good your models are performing. Once you are coming close to the end of the competition and want to have a real submission to compete on the private LB then submit the real submission file. IMHO this is one possible explanation for such a sudden jump!",
      "votes": 2,
      "replies": [
        {
          "id": 970048,
          "postDate": "2020-08-14T06:41:14.570Z",
          "content": "<p>To avoid this, in AUC metric competitions, Kaggle should probably show LB as max(AUC, 1-AUC)</p>",
          "rawMarkdown": "To avoid this, in AUC metric competitions, Kaggle should probably show LB as max(AUC, 1-AUC)"
        },
        {
          "id": 970084,
          "postDate": "2020-08-14T07:23:27.613Z",
          "content": "<p>It could be an explanation but perhaps not for this one. <br>\nThey DID get a score about 0.96 before their jump. Which means they need to submit 300X scores lower than 0.04 seemed really impossible. </p>",
          "rawMarkdown": "It could be an explanation but perhaps not for this one. \nThey DID get a score about 0.96 before their jump. Which means they need to submit 300X scores lower than 0.04 seemed really impossible. "
        },
        {
          "id": 970119,
          "postDate": "2020-08-14T08:09:59.493Z",
          "content": "<p>not sure to understand your comment. They had 0.96, wanted to try something new without being visible, created a submission file, took 1 - subfile and submited: score is 0.03. Then they know they can get to 0.97 but you don't. That's all I'm saying, there is no need for 300 submissions.</p>",
          "rawMarkdown": "not sure to understand your comment. They had 0.96, wanted to try something new without being visible, created a submission file, took 1 - subfile and submited: score is 0.03. Then they know they can get to 0.97 but you don't. That's all I'm saying, there is no need for 300 submissions.",
          "votes": 1
        },
        {
          "id": 970371,
          "postDate": "2020-08-14T11:46:04.610Z",
          "content": "<p>I read somewhere some time ago that kaggle already does this.</p>",
          "rawMarkdown": "I read somewhere some time ago that kaggle already does this.",
          "votes": 1
        },
        {
          "id": 970441,
          "postDate": "2020-08-14T13:06:53.717Z",
          "content": "<blockquote>\n  <p>Kaggle should probably show LB as max(AUC, 1-AUC)</p>\n</blockquote>\n<p>Kaggle does it.</p>",
          "rawMarkdown": "> Kaggle should probably show LB as max(AUC, 1-AUC)\n\nKaggle does it."
        },
        {
          "id": 970513,
          "postDate": "2020-08-14T14:17:36.773Z",
          "content": "<p>I think Kaggle only does it if your AUC is very low. For example if your AUC is 0.49 then Kaggle does not report 0.51. But if your AUC is 0.10 then perhaps Kaggle reports 0.90.</p>",
          "rawMarkdown": "I think Kaggle only does it if your AUC is very low. For example if your AUC is 0.49 then Kaggle does not report 0.51. But if your AUC is 0.10 then perhaps Kaggle reports 0.90.",
          "votes": 1
        },
        {
          "id": 970527,
          "postDate": "2020-08-14T14:34:27.167Z",
          "content": "<p>nope. Had a small bug and got LB AUC of 0.06 instead of 0.94</p>",
          "rawMarkdown": "nope. Had a small bug and got LB AUC of 0.06 instead of 0.94",
          "votes": 4,
          "replies": [
            {
              "id": 970698,
              "postDate": "2020-08-14T17:26:57.243Z",
              "content": "<p>ok, Kaggle did it in some competitions, I've seen it ;)</p>",
              "rawMarkdown": "ok, Kaggle did it in some competitions, I've seen it ;)",
              "votes": 1
            }
          ]
        }
      ]
    },
    {
      "id": 969753,
      "postDate": "2020-08-13T22:52:00.617Z",
      "content": "<p>I never did LB probing but heard many times during many competitions. By its intuition, personally we didn't feel any excitement on this technique, so never wanted to try. But I also read that this technique is not prohibited in competition. So, just wondering why it looks like many participants get bothered about it!! But I must agree with Guanshuo Xu, it would be better if it's a kernel competition. </p>",
      "rawMarkdown": "I never did LB probing but heard many times during many competitions. By its intuition, personally we didn't feel any excitement on this technique, so never wanted to try. But I also read that this technique is not prohibited in competition. So, just wondering why it looks like many participants get bothered about it!! But I must agree with Guanshuo Xu, it would be better if it's a kernel competition. ",
      "votes": 2,
      "replies": [
        {
          "id": 969754,
          "postDate": "2020-08-13T22:54:02.933Z",
          "content": "<p>In real life your \"employer\" tells you where the test images come from. In Kaggle \"LB probing\" tells you where the test images come from. It's important to know if the test images come from the same distribution as train or if the features are shifted (explained <a href=\"https://towardsdatascience.com/understanding-dataset-shift-f2a5a262a766\" target=\"_blank\">here</a>).</p>",
          "rawMarkdown": "In real life your \"employer\" tells you where the test images come from. In Kaggle \"LB probing\" tells you where the test images come from. It's important to know if the test images come from the same distribution as train or if the features are shifted (explained [here] [1]).\n\n[1]: https://towardsdatascience.com/understanding-dataset-shift-f2a5a262a766",
          "votes": 5
        },
        {
          "id": 970364,
          "postDate": "2020-08-14T11:40:16.130Z",
          "rawMarkdown": "",
          "votes": 1,
          "isDeleted": true
        }
      ]
    },
    {
      "id": 970507,
      "postDate": "2020-08-14T14:14:09.007Z",
      "content": "<p>Agree. Bad practice! Kaggle should have a simple script to check this and deny such mal practice.</p>",
      "rawMarkdown": "Agree. Bad practice! Kaggle should have a simple script to check this and deny such mal practice.",
      "votes": -4
    },
    {
      "id": 969740,
      "postDate": "2020-08-13T22:32:56.693Z",
      "content": "<p>Is this the start of a leak hunt?</p>",
      "rawMarkdown": "Is this the start of a leak hunt?",
      "replies": [
        {
          "id": 969751,
          "postDate": "2020-08-13T22:51:31.990Z",
          "content": "<p>Quote from Sirish <a href=\"https://www.kaggle.com/c/siim-isic-melanoma-classification/discussion/161497#900907\" target=\"_blank\">here</a></p>\n<blockquote>\n  <p>I did not find/exploit any LEAK. I can understand your concern, especially because recently top teams lost their ranking (and prize money), for example during Liverpool competition link1 link2</p>\n</blockquote>",
          "rawMarkdown": "Quote from Sirish [here][1]\n> I did not find/exploit any LEAK. I can understand your concern, especially because recently top teams lost their ranking (and prize money), for example during Liverpool competition link1 link2\n\n[1]: https://www.kaggle.com/c/siim-isic-melanoma-classification/discussion/161497#900907",
          "votes": 3
        },
        {
          "id": 969759,
          "postDate": "2020-08-13T23:02:10.283Z",
          "content": "<p>I once wrote I was not using a leak then was accused of misleading people on purpose.  What is a leak and what is not  a leak is not well defined.</p>",
          "rawMarkdown": "I once wrote I was not using a leak then was accused of misleading people on purpose.  What is a leak and what is not  a leak is not well defined.",
          "votes": 3
        }
      ]
    },
    {
      "id": 969724,
      "postDate": "2020-08-13T22:13:02.347Z",
      "rawMarkdown": "",
      "isDeleted": true
    }
  ],
  "comments": [
    {
      "id": 970111,
      "author_name": "nosound",
      "author_url": "",
      "post_date": "2020-08-14T08:03:07.497000",
      "content": "<p>Hi, please let me clarify as much as I can now. That result is based on LB probing, it comprises many probing submissions, each of the probing submission scores is in range 0.4-0.6 that is why it was not visible before. This submission will be very bad for the private LB, it is basically very overfited to the public. I plan to publish more details on the procedure when the competition ends. I submitted it now to showcase what is possible to achieve with the probing. We didn't submit it earlier in order not to disrupt the leaderboard.</p>\n<p>My motivation: I like solving tasks with mixed integer programming, I worked with it at IBM many years ago and I practise it when appropriate since then. I try to distill as much information as possible from each submission, it is challenging and interesting.</p>\n<p>I understand that the area is a gray area, so I try to maintain full transparency, specifically on how we use it in our final solutions, if at all. Will be glad to share more when the competition ends.</p>\n<p>Good luck to you in the last few days!</p>",
      "votes": 33,
      "replies": [
        {
          "id": 970163,
          "author_name": "jsyphil",
          "author_url": "",
          "post_date": "2020-08-14T09:00:23.573000",
          "content": "<p>Thanks for clarifying.</p>",
          "votes": 0,
          "replies": []
        },
        {
          "id": 970217,
          "author_name": "CPMP",
          "author_url": "",
          "post_date": "2020-08-14T09:31:13.457000",
          "content": "<blockquote>\n  <p>I like solving tasks with mixed integer programming,</p>\n</blockquote>\n<p>I love it.  Thanks for clarifying.</p>\n<p>Did you use CPLEX while at IBM?  That's what I was working on as you probably know.</p>",
          "votes": 2,
          "replies": []
        },
        {
          "id": 970300,
          "author_name": "nosound",
          "author_url": "",
          "post_date": "2020-08-14T10:46:29.843000",
          "content": "<p>Yes, I remember you also participated in the latest Santa competition. At IBM we used CPLEX, of course, but during the Santa competition it became clear to us that Gurobi worked better there, as reported also by many other people. So I switched to Gurobi, and here also I used Gurobi.</p>",
          "votes": 1,
          "replies": []
        },
        {
          "id": 970358,
          "author_name": "",
          "author_url": "",
          "post_date": "2020-08-14T11:36:50.890000",
          "content": "",
          "votes": 0,
          "replies": [
            {
              "id": 970359,
              "author_name": "Gilles Vandewiele",
              "author_url": "",
              "post_date": "2020-08-14T11:37:27.910000",
              "content": "<p><a href=\"https://www.kaggle.com/cpmpml/number-of-public-melanoma-is-78-or-77\" target=\"_blank\">That's already known</a>. 77 or 78 positive images.</p>",
              "votes": 3,
              "replies": []
            },
            {
              "id": 970374,
              "author_name": "",
              "author_url": "",
              "post_date": "2020-08-14T11:46:48.433000",
              "content": "",
              "votes": 0,
              "replies": []
            },
            {
              "id": 970440,
              "author_name": "CPMP",
              "author_url": "",
              "post_date": "2020-08-14T13:06:02.497000",
              "content": "<p>It is the best notebook ever.</p>",
              "votes": 8,
              "replies": []
            }
          ]
        },
        {
          "id": 973382,
          "author_name": "Kha Vo",
          "author_url": "",
          "post_date": "2020-08-17T09:20:57.800000",
          "content": "<p><a href=\"https://www.kaggle.com/zaharch\" target=\"_blank\">@zaharch</a> was also the one who had an impressive public LB score in LANL Earthquake Prediction not long ago.  He did some smart optimization to know public rows of submissions. </p>",
          "votes": 0,
          "replies": []
        }
      ]
    },
    {
      "id": 969722,
      "author_name": "Abhishek Thakur",
      "author_url": "",
      "post_date": "2020-08-13T22:09:23.657000",
      "content": "<p>I don't think we should make any kind of allegations or comments without knowing their side of the story.<br>\nIf it's probing it has been done already by a few this competition. <br>\nAs long as the private solution is clean, who cares. </p>",
      "votes": 29,
      "replies": [
        {
          "id": 970408,
          "author_name": "Ali Abdin",
          "author_url": "",
          "post_date": "2020-08-14T12:38:01.443000",
          "content": "<p>Totally agreeing with you</p>",
          "votes": 0,
          "replies": []
        }
      ]
    },
    {
      "id": 969719,
      "author_name": "Guanshuo Xu",
      "author_url": "",
      "post_date": "2020-08-13T22:02:13.193000",
      "content": "<p>Kaggle should make this one a kernel competition.</p>",
      "votes": 13,
      "replies": [
        {
          "id": 969731,
          "author_name": "Serigne ",
          "author_url": "",
          "post_date": "2020-08-13T22:22:27.067000",
          "content": "<blockquote>\n  <p>Kaggle should make this one a kernel competition.</p>\n</blockquote>\n<p>Or 2 stages competition with 2nd stage test set available just a week before the end. That would still allow internet use (for TPU for instance)<br>\nThere used to be many 2 stages CV competitions before on kaggle( in order to prevent handlabelling among others).  I don't why that's not the case anymore.</p>",
          "votes": 4,
          "replies": []
        }
      ]
    },
    {
      "id": 969729,
      "author_name": "Rob Mulla",
      "author_url": "",
      "post_date": "2020-08-13T22:18:46.813000",
      "content": "<p>Two of these things are not like the others 🤔</p>\n<p><img src=\"https://www.googleapis.com/download/storage/v1/b/kaggle-forum-message-attachments/o/inbox%2F644036%2Ff7fc749a373c3ba12da61593d5af3697%2FScreenshot%20from%202020-08-13%2018-17-40.png?generation=1597357117685509&amp;alt=media\" alt=\"\"></p>",
      "votes": 12,
      "replies": [
        {
          "id": 969747,
          "author_name": "Chris Deotte",
          "author_url": "",
          "post_date": "2020-08-13T22:44:56.793000",
          "content": "<p>Great plot. In the first month of the comp, Yuval was in 1st place using a single model that scored LB 0.960 explained <a href=\"https://www.kaggle.com/c/siim-isic-melanoma-classification/discussion/168152#939957\" target=\"_blank\">here</a>. After Sirish took first, Sirish explained what he did to climb LB <a href=\"https://www.kaggle.com/c/siim-isic-melanoma-classification/discussion/161497#901185\" target=\"_blank\">here</a>. Then NoSound joined Yuval and their LB did not increase for 22 days until today when it jumped up to match Sirish's.</p>",
          "votes": 7,
          "replies": []
        },
        {
          "id": 969760,
          "author_name": "CPMP",
          "author_url": "",
          "post_date": "2020-08-13T23:03:35.020000",
          "content": "<blockquote>\n  <p>their LB did not increase for 22 days until today</p>\n</blockquote>\n<p>It shows LB is not required to measure progress.</p>",
          "votes": 7,
          "replies": []
        },
        {
          "id": 970024,
          "author_name": "Psi",
          "author_url": "",
          "post_date": "2020-08-14T06:20:43.227000",
          "content": "<p>How many subs during the no-progress period?</p>",
          "votes": 0,
          "replies": []
        },
        {
          "id": 970031,
          "author_name": "Chris Deotte",
          "author_url": "",
          "post_date": "2020-08-14T06:25:21.370000",
          "content": "<p>I forget. Can we check that in meta Kaggle dataset? If I remember correctly, i think they did about 5 a day, so about 100 subs.</p>",
          "votes": 1,
          "replies": []
        },
        {
          "id": 970038,
          "author_name": "Chris Deotte",
          "author_url": "",
          "post_date": "2020-08-14T06:32:28.003000",
          "content": "<p>That is probably enough to increase 0.960 to 0.990. I think you could check your model's top 3000 predictions by making 100 subs where you set the top <code>pred[30*k,30*(k+1)]=1</code> for <code>k in range(100)</code> (and other predictions to zero). Then you order all those groups by their LB score (keeping them above the bottom 7000). I could run a simulation on OOF doing this procedure to see how much it increases AUC.</p>",
          "votes": 6,
          "replies": [
            {
              "id": 970159,
              "author_name": "MhdSharuk",
              "author_url": "",
              "post_date": "2020-08-14T08:57:49.247000",
              "content": "<p><a href=\"https://www.kaggle.com/cdeotte\" target=\"_blank\">@cdeotte</a> Could you explain how this is done or could you create a notebook about this one??</p>",
              "votes": 0,
              "replies": []
            },
            {
              "id": 970510,
              "author_name": "Chris Deotte",
              "author_url": "",
              "post_date": "2020-08-14T14:15:45.710000",
              "content": "<p>I did a quick simulation of what I describe above using OOF. It's not as effective as I thought. </p>\n<p>If you do 100 probes (size 32 each) your LB would increase from 0.9600 to 0.9650. <br>\nIf you do 200 probes (size 16 each) then your LB would increase from 0.9600 to 0.9675. <br>\nIf you do 400 probes (size 8 each) then your LB would increase from 0.9600 to 0.9715. <br>\nIf you do 800 probes (size 4 each) then your LB would increase from 0.9600 to 0.9800.</p>\n<p>Furthermore, simulation suggests that this procedure would increase public LB but it would simultaneously decrease private LB.</p>",
              "votes": 3,
              "replies": []
            }
          ]
        },
        {
          "id": 970075,
          "author_name": "Changyi",
          "author_url": "",
          "post_date": "2020-08-14T07:15:46.937000",
          "content": "<p>Sounds reasonable <a href=\"https://www.kaggle.com/cdeotte\" target=\"_blank\">@cdeotte</a>. Let's see your simulation results.</p>",
          "votes": 1,
          "replies": []
        }
      ]
    },
    {
      "id": 969694,
      "author_name": "Psi",
      "author_url": "",
      "post_date": "2020-08-13T21:26:06.180000",
      "content": "<p>I think they jumped from 962 directly to 99x which is somehow weird :)</p>",
      "votes": 4,
      "replies": [
        {
          "id": 969713,
          "author_name": "Changyi",
          "author_url": "",
          "post_date": "2020-08-13T21:54:05.047000",
          "content": "<p>I agree, more weird than the 1st.<br>\nPlus, seeing many GMs like you not in the front part makes me feel a shake will happen.</p>",
          "votes": 4,
          "replies": []
        },
        {
          "id": 969718,
          "author_name": "Sxwat",
          "author_url": "",
          "post_date": "2020-08-13T22:00:39.697000",
          "content": "<p>So big jump is very strange for probing … maybe they found some leak, doctor who labeled image by hand (joke). At least they have good PL and can train model on test dataset which is a big advantage</p>",
          "votes": 0,
          "replies": []
        },
        {
          "id": 969721,
          "author_name": "Psi",
          "author_url": "",
          "post_date": "2020-08-13T22:07:47.447000",
          "content": "<p>Or they probed images separately and then combined the probe.</p>",
          "votes": 5,
          "replies": []
        },
        {
          "id": 969725,
          "author_name": "Dieter",
          "author_url": "",
          "post_date": "2020-08-13T22:13:04.863000",
          "content": "<p>💯.         </p>",
          "votes": 1,
          "replies": []
        }
      ]
    },
    {
      "id": 969739,
      "author_name": "CPMP",
      "author_url": "",
      "post_date": "2020-08-13T22:31:28.750000",
      "content": "<blockquote>\n  <p>While you can use the probed labels as additional training data </p>\n</blockquote>\n<p>No you can't.  Hand labeling of test data is prohibited.</p>",
      "votes": 1,
      "replies": [
        {
          "id": 969743,
          "author_name": "Dieter",
          "author_url": "",
          "post_date": "2020-08-13T22:38:01.613000",
          "content": "<p>its not hand labeling, its Public LB labeling :D</p>",
          "votes": 2,
          "replies": []
        },
        {
          "id": 969750,
          "author_name": "Chris Deotte",
          "author_url": "",
          "post_date": "2020-08-13T22:47:36.483000",
          "content": "<p>You can completely automate it with a script. A script (without any human intervention) can make a submission to LB and store the result. (A human never sees the result). Then the same script can impart it's knowledge into your model.</p>",
          "votes": 3,
          "replies": []
        },
        {
          "id": 969755,
          "author_name": "CPMP",
          "author_url": "",
          "post_date": "2020-08-13T22:58:15.960000",
          "content": "<p>Well, if this is how they did it then fine.  But I doubt it very much.  And I'm curious to see how host interprets it.</p>",
          "votes": 0,
          "replies": []
        },
        {
          "id": 969758,
          "author_name": "Chris Deotte",
          "author_url": "",
          "post_date": "2020-08-13T23:02:08.807000",
          "content": "<p>There is a lot of gray area here. What if you cluster the images unsupervised into groups of 30. Then you make 333 LB submissions. Afterward is it hand labeling or a decision tree? </p>\n<pre><code> If test image belongs to cluster X then classify as Y\n</code></pre>",
          "votes": 1,
          "replies": []
        },
        {
          "id": 969852,
          "author_name": "CPMP",
          "author_url": "",
          "post_date": "2020-08-14T02:09:04.773000",
          "content": "<p>You said it : it is a gray area.</p>\n<p>I interpreted rules to forbid it, but I get that another interpretation is possible.</p>\n<p>Best would be to have Kaggle or host weigh in.</p>",
          "votes": 1,
          "replies": []
        }
      ]
    },
    {
      "id": 970133,
      "author_name": "Signal",
      "author_url": "",
      "post_date": "2020-08-14T08:30:10.750000",
      "content": "<p>I think LB probing is fine if someone is curious.  Personally I just could never pass the thought of using up all those submissions, especially toward the end.  It would take alot of submissions to probe to get to .99, and they have two in the team, but still, 385 submissions = 77 days of submissions, so burn then however you like.  The only issue I can see is if someone has made dummy accounts to do the probing, although Kaggle has done a good job of slow playing people like that and then dealing with it when competitions end.  </p>",
      "votes": 2,
      "replies": []
    },
    {
      "id": 970041,
      "author_name": "Optimo",
      "author_url": "",
      "post_date": "2020-08-14T06:35:03.357000",
      "content": "<p>Well I’m not assuming that they used probing to get there, but here is a simple set up you can use to probe (or simply check your progress) as much as you want in this competition while staying under the radars:</p>\n<p>Simply take your submission and do 1 - submission before submitting it, then your AUC will be 1-real_AUC. No one will know about how good your models are performing. Once you are coming close to the end of the competition and want to have a real submission to compete on the private LB then submit the real submission file. IMHO this is one possible explanation for such a sudden jump!</p>",
      "votes": 2,
      "replies": [
        {
          "id": 970048,
          "author_name": "Optimo",
          "author_url": "",
          "post_date": "2020-08-14T06:41:14.570000",
          "content": "<p>To avoid this, in AUC metric competitions, Kaggle should probably show LB as max(AUC, 1-AUC)</p>",
          "votes": 0,
          "replies": []
        },
        {
          "id": 970084,
          "author_name": "Changyi",
          "author_url": "",
          "post_date": "2020-08-14T07:23:27.613000",
          "content": "<p>It could be an explanation but perhaps not for this one. <br>\nThey DID get a score about 0.96 before their jump. Which means they need to submit 300X scores lower than 0.04 seemed really impossible. </p>",
          "votes": 0,
          "replies": []
        },
        {
          "id": 970119,
          "author_name": "Optimo",
          "author_url": "",
          "post_date": "2020-08-14T08:09:59.493000",
          "content": "<p>not sure to understand your comment. They had 0.96, wanted to try something new without being visible, created a submission file, took 1 - subfile and submited: score is 0.03. Then they know they can get to 0.97 but you don't. That's all I'm saying, there is no need for 300 submissions.</p>",
          "votes": 1,
          "replies": []
        },
        {
          "id": 970371,
          "author_name": "Rohit Agarwal",
          "author_url": "",
          "post_date": "2020-08-14T11:46:04.610000",
          "content": "<p>I read somewhere some time ago that kaggle already does this.</p>",
          "votes": 1,
          "replies": []
        },
        {
          "id": 970441,
          "author_name": "CPMP",
          "author_url": "",
          "post_date": "2020-08-14T13:06:53.717000",
          "content": "<blockquote>\n  <p>Kaggle should probably show LB as max(AUC, 1-AUC)</p>\n</blockquote>\n<p>Kaggle does it.</p>",
          "votes": 0,
          "replies": []
        },
        {
          "id": 970513,
          "author_name": "Chris Deotte",
          "author_url": "",
          "post_date": "2020-08-14T14:17:36.773000",
          "content": "<p>I think Kaggle only does it if your AUC is very low. For example if your AUC is 0.49 then Kaggle does not report 0.51. But if your AUC is 0.10 then perhaps Kaggle reports 0.90.</p>",
          "votes": 1,
          "replies": []
        },
        {
          "id": 970527,
          "author_name": "Dieter",
          "author_url": "",
          "post_date": "2020-08-14T14:34:27.167000",
          "content": "<p>nope. Had a small bug and got LB AUC of 0.06 instead of 0.94</p>",
          "votes": 4,
          "replies": [
            {
              "id": 970698,
              "author_name": "CPMP",
              "author_url": "",
              "post_date": "2020-08-14T17:26:57.243000",
              "content": "<p>ok, Kaggle did it in some competitions, I've seen it ;)</p>",
              "votes": 1,
              "replies": []
            }
          ]
        }
      ]
    },
    {
      "id": 969753,
      "author_name": "Innat",
      "author_url": "",
      "post_date": "2020-08-13T22:52:00.617000",
      "content": "<p>I never did LB probing but heard many times during many competitions. By its intuition, personally we didn't feel any excitement on this technique, so never wanted to try. But I also read that this technique is not prohibited in competition. So, just wondering why it looks like many participants get bothered about it!! But I must agree with Guanshuo Xu, it would be better if it's a kernel competition. </p>",
      "votes": 2,
      "replies": [
        {
          "id": 969754,
          "author_name": "Chris Deotte",
          "author_url": "",
          "post_date": "2020-08-13T22:54:02.933000",
          "content": "<p>In real life your \"employer\" tells you where the test images come from. In Kaggle \"LB probing\" tells you where the test images come from. It's important to know if the test images come from the same distribution as train or if the features are shifted (explained <a href=\"https://towardsdatascience.com/understanding-dataset-shift-f2a5a262a766\" target=\"_blank\">here</a>).</p>",
          "votes": 5,
          "replies": []
        },
        {
          "id": 970364,
          "author_name": "",
          "author_url": "",
          "post_date": "2020-08-14T11:40:16.130000",
          "content": "",
          "votes": 1,
          "replies": []
        }
      ]
    },
    {
      "id": 970507,
      "author_name": "Patrick Chan",
      "author_url": "",
      "post_date": "2020-08-14T14:14:09.007000",
      "content": "<p>Agree. Bad practice! Kaggle should have a simple script to check this and deny such mal practice.</p>",
      "votes": -4,
      "replies": []
    },
    {
      "id": 969740,
      "author_name": "CPMP",
      "author_url": "",
      "post_date": "2020-08-13T22:32:56.693000",
      "content": "<p>Is this the start of a leak hunt?</p>",
      "votes": 0,
      "replies": [
        {
          "id": 969751,
          "author_name": "Chris Deotte",
          "author_url": "",
          "post_date": "2020-08-13T22:51:31.990000",
          "content": "<p>Quote from Sirish <a href=\"https://www.kaggle.com/c/siim-isic-melanoma-classification/discussion/161497#900907\" target=\"_blank\">here</a></p>\n<blockquote>\n  <p>I did not find/exploit any LEAK. I can understand your concern, especially because recently top teams lost their ranking (and prize money), for example during Liverpool competition link1 link2</p>\n</blockquote>",
          "votes": 3,
          "replies": []
        },
        {
          "id": 969759,
          "author_name": "CPMP",
          "author_url": "",
          "post_date": "2020-08-13T23:02:10.283000",
          "content": "<p>I once wrote I was not using a leak then was accused of misleading people on purpose.  What is a leak and what is not  a leak is not well defined.</p>",
          "votes": 3,
          "replies": []
        }
      ]
    },
    {
      "id": 969724,
      "author_name": "",
      "author_url": "",
      "post_date": "2020-08-13T22:13:02.347000",
      "content": "",
      "votes": 0,
      "replies": []
    }
  ],
  "raw_markdown_by_id": {
    "970111": "Hi, please let me clarify as much as I can now. That result is based on LB probing, it comprises many probing submissions, each of the probing submission scores is in range 0.4-0.6 that is why it was not visible before. This submission will be very bad for the private LB, it is basically very overfited to the public. I plan to publish more details on the procedure when the competition ends. I submitted it now to showcase what is possible to achieve with the probing. We didn't submit it earlier in order not to disrupt the leaderboard.\n\nMy motivation: I like solving tasks with mixed integer programming, I worked with it at IBM many years ago and I practise it when appropriate since then. I try to distill as much information as possible from each submission, it is challenging and interesting.\n\nI understand that the area is a gray area, so I try to maintain full transparency, specifically on how we use it in our final solutions, if at all. Will be glad to share more when the competition ends.\n\nGood luck to you in the last few days!",
    "969690": "I can only speculate, but telling from the number of submissions, they just intensively probed image per image on the LB to label the test images which are part of the public test set... Whats the point? While you can use the probed labels as additional training data and probably get a better score, how does this help anyone but your own ego?",
    "969722": "I don't think we should make any kind of allegations or comments without knowing their side of the story.\nIf it's probing it has been done already by a few this competition. \nAs long as the private solution is clean, who cares. \n",
    "969719": "Kaggle should make this one a kernel competition.",
    "969729": "Two of these things are not like the others 🤔\n\n![](https://www.googleapis.com/download/storage/v1/b/kaggle-forum-message-attachments/o/inbox%2F644036%2Ff7fc749a373c3ba12da61593d5af3697%2FScreenshot%20from%202020-08-13%2018-17-40.png?generation=1597357117685509&alt=media)",
    "969694": "I think they jumped from 962 directly to 99x which is somehow weird :)",
    "969739": "> While you can use the probed labels as additional training data \n\nNo you can't.  Hand labeling of test data is prohibited.",
    "970133": "I think LB probing is fine if someone is curious.  Personally I just could never pass the thought of using up all those submissions, especially toward the end.  It would take alot of submissions to probe to get to .99, and they have two in the team, but still, 385 submissions = 77 days of submissions, so burn then however you like.  The only issue I can see is if someone has made dummy accounts to do the probing, although Kaggle has done a good job of slow playing people like that and then dealing with it when competitions end.  ",
    "970041": "Well I’m not assuming that they used probing to get there, but here is a simple set up you can use to probe (or simply check your progress) as much as you want in this competition while staying under the radars:\n\nSimply take your submission and do 1 - submission before submitting it, then your AUC will be 1-real_AUC. No one will know about how good your models are performing. Once you are coming close to the end of the competition and want to have a real submission to compete on the private LB then submit the real submission file. IMHO this is one possible explanation for such a sudden jump!",
    "969753": "I never did LB probing but heard many times during many competitions. By its intuition, personally we didn't feel any excitement on this technique, so never wanted to try. But I also read that this technique is not prohibited in competition. So, just wondering why it looks like many participants get bothered about it!! But I must agree with Guanshuo Xu, it would be better if it's a kernel competition. ",
    "970507": "Agree. Bad practice! Kaggle should have a simple script to check this and deny such mal practice.",
    "969740": "Is this the start of a leak hunt?",
    "969724": ""
  }
}