{
  "id": 546228,
  "title": "Updating the Hidden Test Set - COMPLETED",
  "url": "/competitions/czii-cryo-et-object-identification/discussion/546228",
  "author_name": "",
  "post_date": "2024-11-14T14:24:11.784477900Z",
  "votes": 26,
  "comment_count": 11,
  "views": 0,
  "content": "<p>The hidden test set is being updated to address a data issue. Once that is complete, all previous submission notebook will be re-run against the updated dataset. This process will take some time, during which, the leaderboard will be in a state of weirdness. This thread will be updated once the process is complete.</p>",
  "messages": [
    {
      "id": "3045508",
      "postDate": "11/14/2024 14:24:11",
      "content": "<p>The hidden test set is being updated to address a data issue. Once that is complete, all previous submission notebook will be re-run against the updated dataset. This process will take some time, during which, the leaderboard will be in a state of weirdness. This thread will be updated once the process is complete.</p>",
      "rawMarkdown": "The hidden test set is being updated to address a data issue. Once that is complete, all previous submission notebook will be re-run against the updated dataset. This process will take some time, during which, the leaderboard will be in a state of weirdness. This thread will be updated once the process is complete.",
      "votes": null
    },
    {
      "id": "3045538",
      "postDate": "11/14/2024 15:03:22",
      "content": "<p>Thank you for updating the test set. Can you also confirm that this script is used for scoring? (With beta=4 and distance_multiplier=0.5)</p>\n<p><a href=\"https://www.kaggle.com/code/metric/czi-cryoet-84969\" target=\"_blank\">https://www.kaggle.com/code/metric/czi-cryoet-84969</a></p>",
      "rawMarkdown": "Thank you for updating the test set. Can you also confirm that this script is used for scoring? (With beta=4 and distance_multiplier=0.5)\n\nhttps://www.kaggle.com/code/metric/czi-cryoet-84969",
      "votes": null
    },
    {
      "id": "3045869",
      "postDate": "11/14/2024 21:38:53",
      "content": "<p>Yes, correct.</p>",
      "rawMarkdown": "Yes, correct.",
      "votes": null
    },
    {
      "id": "3047356",
      "postDate": "11/16/2024 15:33:28",
      "content": "<p>thanks for the update!</p>",
      "rawMarkdown": "thanks for the update!",
      "votes": null
    },
    {
      "id": "3047520",
      "postDate": "11/16/2024 19:15:21",
      "content": "<p>Even if I am in the top 2 at this moment, I must admit that there is a problem with evaluation algorithm 😑. The results are too high !</p>\n<p>Dieter said the code is this one :<br>\n<a href=\"https://www.kaggle.com/code/metric/czi-cryoet-84969\" target=\"_blank\">https://www.kaggle.com/code/metric/czi-cryoet-84969</a></p>\n<p>I went and gave a look to it. I think there is a mistake in this line :<br>\n<code>raw_matches = ref_tree.query_ball_tree(candidate_tree, r=reference_radius)</code></p>\n<p>It should be the inverse to have true positives in ref_tree instead of candidate tree :<br>\n<code>raw_matches = candidate_tree.query_ball_tree(ref_tree, r=reference_radius)</code></p>\n<p>The result is that submitting multiple candidate points in the radius give false 'true positives'.</p>\n<p>Please let me know if I am wrong.<br>\nI hope this helps !</p>",
      "rawMarkdown": "Even if I am in the top 2 at this moment, I must admit that there is a problem with evaluation algorithm 😑. The results are too high !\n\nDieter said the code is this one :\nhttps://www.kaggle.com/code/metric/czi-cryoet-84969\n\nI went and gave a look to it. I think there is a mistake in this line :\n`raw_matches = ref_tree.query_ball_tree(candidate_tree, r=reference_radius)`\n\nIt should be the inverse to have true positives in ref_tree instead of candidate tree :\n`raw_matches = candidate_tree.query_ball_tree(ref_tree, r=reference_radius)`\n\nThe result is that submitting multiple candidate points in the radius give false 'true positives'.\n\nPlease let me know if I am wrong.\nI hope this helps !",
      "votes": null
    },
    {
      "id": "3047535",
      "postDate": "11/16/2024 19:44:18",
      "content": "<p>it would be good if you could illustrate the problem with dummy data.</p>\n<p>i am suggesting using linear_sum_assignment for  cross checking the eval code<br>\n<a href=\"https://docs.scipy.org/doc/scipy/reference/generated/scipy.optimize.linear_sum_assignment.html\" target=\"_blank\">https://docs.scipy.org/doc/scipy/reference/generated/scipy.optimize.linear_sum_assignment.html</a></p>",
      "rawMarkdown": "it would be good if you could illustrate the problem with dummy data.\n\ni am suggesting using linear_sum_assignment for  cross checking the eval code\nhttps://docs.scipy.org/doc/scipy/reference/generated/scipy.optimize.linear_sum_assignment.html",
      "votes": null
    },
    {
      "id": "3047545",
      "postDate": "11/16/2024 20:14:01",
      "content": "<p>No problem, I made a quick notebook. You can see that before correction we have too much \"true positives\", and even negative (-1) \"false negatives\".<br>\nThe score is also 1,36 &gt; 1, which should not be realistic 😅<br>\n<a href=\"https://www.kaggle.com/code/mjz1977/czii-metric-is-it-a-mistake\" target=\"_blank\">https://www.kaggle.com/code/mjz1977/czii-metric-is-it-a-mistake</a></p>\n<p>After correction it is OK</p>",
      "rawMarkdown": "No problem, I made a quick notebook. You can see that before correction we have too much \"true positives\", and even negative (-1) \"false negatives\".\nThe score is also 1,36 > 1, which should not be realistic 😅\nhttps://www.kaggle.com/code/mjz1977/czii-metric-is-it-a-mistake\n\nAfter correction it is OK",
      "votes": null
    },
    {
      "id": "3047601",
      "postDate": "11/16/2024 23:35:46",
      "content": "<p><a href=\"https://www.kaggle.com/inversion\" target=\"_blank\">@inversion</a> <a href=\"https://www.kaggle.com/kharrington\" target=\"_blank\">@kharrington</a>  yet another evaluation bug discovered by <a href=\"https://www.kaggle.com/mjz1977\" target=\"_blank\">@mjz1977</a> <br>\n🐞 🐞 🐞</p>\n<p><img src=\"https://www.googleapis.com/download/storage/v1/b/kaggle-forum-message-attachments/o/inbox%2F113660%2F7e8bb892dd7ab2ea73bc6ee4f1307192%2FSelection_692.png?generation=1731800092073125&amp;alt=media\" alt=\"\"></p>",
      "rawMarkdown": "inversion @kharrington  yet another evaluation bug discovered by @mjz1977 \n🐞 🐞 🐞\n\n![](https://www.googleapis.com/download/storage/v1/b/kaggle-forum-message-attachments/o/inbox%2F113660%2F7e8bb892dd7ab2ea73bc6ee4f1307192%2FSelection_692.png?generation=1731800092073125&alt=media)",
      "votes": null
    },
    {
      "id": "3047621",
      "postDate": "11/17/2024 00:46:12",
      "content": "<p><a href=\"https://www.kaggle.com/mjz1977\" target=\"_blank\">@mjz1977</a> good 👀. Thank you for the notebook. We've been checking it over and will followup/update on Monday.</p>",
      "rawMarkdown": "mjz1977 good 👀. Thank you for the notebook. We've been checking it over and will followup/update on Monday.",
      "votes": null
    },
    {
      "id": "3047844",
      "postDate": "11/17/2024 08:57:21",
      "content": "<p>Probably has an easy and fast fix. But since the score counts hits as coordinates inside a ground truth radius why not to give han explicit spherical mask and do evalute with CELoss directly? The statistics of ML are still my Achilles heel. Sorry if is a terrible suggestion.</p>",
      "rawMarkdown": "Probably has an easy and fast fix. But since the score counts hits as coordinates inside a ground truth radius why not to give han explicit spherical mask and do evalute with CELoss directly? The statistics of ML are still my Achilles heel. Sorry if is a terrible suggestion.",
      "votes": null
    },
    {
      "id": "3048815",
      "postDate": "11/18/2024 11:32:51",
      "content": "<p>I think that evaluation should be OK after the correction of wrong line. If staff go to another method, they will have to validate it before, and perhaps some other issues will appear 😉</p>",
      "rawMarkdown": "I think that evaluation should be OK after the correction of wrong line. If staff go to another method, they will have to validate it before, and perhaps some other issues will appear 😉",
      "votes": null
    },
    {
      "id": "3049179",
      "postDate": "11/18/2024 19:50:36",
      "content": "<p><a href=\"https://www.kaggle.com/mjz1977\" target=\"_blank\">@mjz1977</a> Would you mind reaching out to me via email or Kaggle contact form?</p>",
      "rawMarkdown": "mjz1977 Would you mind reaching out to me via email or Kaggle contact form?",
      "votes": null
    }
  ],
  "comments": [
    {
      "id": 3045538,
      "author_name": "christofhenkel",
      "author_url": "",
      "post_date": "11/14/2024 15:03:22",
      "content": "<p>Thank you for updating the test set. Can you also confirm that this script is used for scoring? (With beta=4 and distance_multiplier=0.5)</p>\n<p><a href=\"https://www.kaggle.com/code/metric/czi-cryoet-84969\" target=\"_blank\">https://www.kaggle.com/code/metric/czi-cryoet-84969</a></p>",
      "votes": null,
      "replies": [
        {
          "id": 3045869,
          "author_name": "inversion",
          "author_url": "",
          "post_date": "11/14/2024 21:38:53",
          "content": "<p>Yes, correct.</p>",
          "votes": null,
          "replies": []
        }
      ]
    },
    {
      "id": 3047356,
      "author_name": "hengck23",
      "author_url": "",
      "post_date": "11/16/2024 15:33:28",
      "content": "<p>thanks for the update!</p>",
      "votes": null,
      "replies": []
    },
    {
      "id": 3047520,
      "author_name": "mjz1977",
      "author_url": "",
      "post_date": "11/16/2024 19:15:21",
      "content": "<p>Even if I am in the top 2 at this moment, I must admit that there is a problem with evaluation algorithm 😑. The results are too high !</p>\n<p>Dieter said the code is this one :<br>\n<a href=\"https://www.kaggle.com/code/metric/czi-cryoet-84969\" target=\"_blank\">https://www.kaggle.com/code/metric/czi-cryoet-84969</a></p>\n<p>I went and gave a look to it. I think there is a mistake in this line :<br>\n<code>raw_matches = ref_tree.query_ball_tree(candidate_tree, r=reference_radius)</code></p>\n<p>It should be the inverse to have true positives in ref_tree instead of candidate tree :<br>\n<code>raw_matches = candidate_tree.query_ball_tree(ref_tree, r=reference_radius)</code></p>\n<p>The result is that submitting multiple candidate points in the radius give false 'true positives'.</p>\n<p>Please let me know if I am wrong.<br>\nI hope this helps !</p>",
      "votes": null,
      "replies": [
        {
          "id": 3047535,
          "author_name": "hengck23",
          "author_url": "",
          "post_date": "11/16/2024 19:44:18",
          "content": "<p>it would be good if you could illustrate the problem with dummy data.</p>\n<p>i am suggesting using linear_sum_assignment for  cross checking the eval code<br>\n<a href=\"https://docs.scipy.org/doc/scipy/reference/generated/scipy.optimize.linear_sum_assignment.html\" target=\"_blank\">https://docs.scipy.org/doc/scipy/reference/generated/scipy.optimize.linear_sum_assignment.html</a></p>",
          "votes": null,
          "replies": [
            {
              "id": 3047545,
              "author_name": "mjz1977",
              "author_url": "",
              "post_date": "11/16/2024 20:14:01",
              "content": "<p>No problem, I made a quick notebook. You can see that before correction we have too much \"true positives\", and even negative (-1) \"false negatives\".<br>\nThe score is also 1,36 &gt; 1, which should not be realistic 😅<br>\n<a href=\"https://www.kaggle.com/code/mjz1977/czii-metric-is-it-a-mistake\" target=\"_blank\">https://www.kaggle.com/code/mjz1977/czii-metric-is-it-a-mistake</a></p>\n<p>After correction it is OK</p>",
              "votes": null,
              "replies": [
                {
                  "id": 3047621,
                  "author_name": "kharrington",
                  "author_url": "",
                  "post_date": "11/17/2024 00:46:12",
                  "content": "<p><a href=\"https://www.kaggle.com/mjz1977\" target=\"_blank\">@mjz1977</a> good 👀. Thank you for the notebook. We've been checking it over and will followup/update on Monday.</p>",
                  "votes": null,
                  "replies": []
                }
              ]
            }
          ]
        },
        {
          "id": 3047601,
          "author_name": "hengck23",
          "author_url": "",
          "post_date": "11/16/2024 23:35:46",
          "content": "<p><a href=\"https://www.kaggle.com/inversion\" target=\"_blank\">@inversion</a> <a href=\"https://www.kaggle.com/kharrington\" target=\"_blank\">@kharrington</a>  yet another evaluation bug discovered by <a href=\"https://www.kaggle.com/mjz1977\" target=\"_blank\">@mjz1977</a> <br>\n🐞 🐞 🐞</p>\n<p><img src=\"https://www.googleapis.com/download/storage/v1/b/kaggle-forum-message-attachments/o/inbox%2F113660%2F7e8bb892dd7ab2ea73bc6ee4f1307192%2FSelection_692.png?generation=1731800092073125&amp;alt=media\" alt=\"\"></p>",
          "votes": null,
          "replies": [
            {
              "id": 3047844,
              "author_name": "sacuscreed",
              "author_url": "",
              "post_date": "11/17/2024 08:57:21",
              "content": "<p>Probably has an easy and fast fix. But since the score counts hits as coordinates inside a ground truth radius why not to give han explicit spherical mask and do evalute with CELoss directly? The statistics of ML are still my Achilles heel. Sorry if is a terrible suggestion.</p>",
              "votes": null,
              "replies": [
                {
                  "id": 3048815,
                  "author_name": "mjz1977",
                  "author_url": "",
                  "post_date": "11/18/2024 11:32:51",
                  "content": "<p>I think that evaluation should be OK after the correction of wrong line. If staff go to another method, they will have to validate it before, and perhaps some other issues will appear 😉</p>",
                  "votes": null,
                  "replies": []
                }
              ]
            }
          ]
        },
        {
          "id": 3049179,
          "author_name": "inversion",
          "author_url": "",
          "post_date": "11/18/2024 19:50:36",
          "content": "<p><a href=\"https://www.kaggle.com/mjz1977\" target=\"_blank\">@mjz1977</a> Would you mind reaching out to me via email or Kaggle contact form?</p>",
          "votes": null,
          "replies": []
        }
      ]
    }
  ],
  "raw_markdown_by_id": {
    "3045508": "The hidden test set is being updated to address a data issue. Once that is complete, all previous submission notebook will be re-run against the updated dataset. This process will take some time, during which, the leaderboard will be in a state of weirdness. This thread will be updated once the process is complete.",
    "3045538": "Thank you for updating the test set. Can you also confirm that this script is used for scoring? (With beta=4 and distance_multiplier=0.5)\n\nhttps://www.kaggle.com/code/metric/czi-cryoet-84969",
    "3045869": "Yes, correct.",
    "3047356": "thanks for the update!",
    "3047520": "Even if I am in the top 2 at this moment, I must admit that there is a problem with evaluation algorithm 😑. The results are too high !\n\nDieter said the code is this one :\nhttps://www.kaggle.com/code/metric/czi-cryoet-84969\n\nI went and gave a look to it. I think there is a mistake in this line :\n`raw_matches = ref_tree.query_ball_tree(candidate_tree, r=reference_radius)`\n\nIt should be the inverse to have true positives in ref_tree instead of candidate tree :\n`raw_matches = candidate_tree.query_ball_tree(ref_tree, r=reference_radius)`\n\nThe result is that submitting multiple candidate points in the radius give false 'true positives'.\n\nPlease let me know if I am wrong.\nI hope this helps !",
    "3047535": "it would be good if you could illustrate the problem with dummy data.\n\ni am suggesting using linear_sum_assignment for  cross checking the eval code\nhttps://docs.scipy.org/doc/scipy/reference/generated/scipy.optimize.linear_sum_assignment.html",
    "3047545": "No problem, I made a quick notebook. You can see that before correction we have too much \"true positives\", and even negative (-1) \"false negatives\".\nThe score is also 1,36 > 1, which should not be realistic 😅\nhttps://www.kaggle.com/code/mjz1977/czii-metric-is-it-a-mistake\n\nAfter correction it is OK",
    "3047601": "inversion @kharrington  yet another evaluation bug discovered by @mjz1977 \n🐞 🐞 🐞\n\n![](https://www.googleapis.com/download/storage/v1/b/kaggle-forum-message-attachments/o/inbox%2F113660%2F7e8bb892dd7ab2ea73bc6ee4f1307192%2FSelection_692.png?generation=1731800092073125&alt=media)",
    "3047621": "mjz1977 good 👀. Thank you for the notebook. We've been checking it over and will followup/update on Monday.",
    "3047844": "Probably has an easy and fast fix. But since the score counts hits as coordinates inside a ground truth radius why not to give han explicit spherical mask and do evalute with CELoss directly? The statistics of ML are still my Achilles heel. Sorry if is a terrible suggestion.",
    "3048815": "I think that evaluation should be OK after the correction of wrong line. If staff go to another method, they will have to validate it before, and perhaps some other issues will appear 😉",
    "3049179": "mjz1977 Would you mind reaching out to me via email or Kaggle contact form?"
  },
  "source": "meta"
}