{
  "id": 468705,
  "title": "Which transformation is used for the label during official evaluation? label/sum(label)?softmax(label)?or?",
  "url": "/competitions/hms-harmful-brain-activity-classification/discussion/468705",
  "author_name": "Rib~",
  "post_date": "2024-01-17T17:21:29.729000",
  "votes": 22,
  "comment_count": 12,
  "views": 0,
  "content": "<p>Which transformation is used for the label during official evaluation? label/sum(label)?softmax(label)?or?</p>",
  "messages": [
    {
      "id": 2606497,
      "postDate": "2024-01-17T17:21:29.730Z",
      "content": "<p>Which transformation is used for the label during official evaluation? label/sum(label)?softmax(label)?or?</p>",
      "rawMarkdown": "Which transformation is used for the label during official evaluation? label/sum(label)?softmax(label)?or?",
      "votes": 21
    },
    {
      "id": 2606507,
      "postDate": "2024-01-17T17:27:21.093Z",
      "content": "<p>No transformation. If you rows don't add up to 1, then your submission will throw an error</p>",
      "rawMarkdown": "No transformation. If you rows don't add up to 1, then your submission will throw an error",
      "votes": 3,
      "replies": [
        {
          "id": 2606512,
          "postDate": "2024-01-17T17:32:49.473Z",
          "content": "<p>What I mean is label, not pred.</p>",
          "rawMarkdown": "What I mean is label, not pred.",
          "votes": 1
        },
        {
          "id": 2607129,
          "postDate": "2024-01-18T04:20:09.880Z",
          "content": "<p>Yep that explains a lot hahaha I have been losing my mind on what has caused these thrown exception errors</p>",
          "rawMarkdown": "Yep that explains a lot hahaha I have been losing my mind on what has caused these thrown exception errors",
          "votes": 1,
          "replies": [
            {
              "id": 2607182,
              "postDate": "2024-01-18T05:20:24.133Z",
              "rawMarkdown": "",
              "isDeleted": true
            }
          ]
        }
      ]
    },
    {
      "id": 2606531,
      "postDate": "2024-01-17T17:54:45.123Z",
      "content": "<p>The metric code is posted here: <a href=\"https://www.kaggle.com/code/metric/kullback-leibler-divergence/notebook\" target=\"_blank\">https://www.kaggle.com/code/metric/kullback-leibler-divergence/notebook</a></p>",
      "rawMarkdown": "The metric code is posted here: https://www.kaggle.com/code/metric/kullback-leibler-divergence/notebook",
      "votes": 1,
      "replies": [
        {
          "id": 2606565,
          "postDate": "2024-01-17T18:10:49.273Z",
          "content": "<p>I know about metrics, but still can't solve my question. We only know that the sum of labels is 1, but I have ten thousand ways to make the sum of labels be 1.</p>",
          "rawMarkdown": "I know about metrics, but still can't solve my question. We only know that the sum of labels is 1, but I have ten thousand ways to make the sum of labels be 1.",
          "votes": 7,
          "replies": [
            {
              "id": 2606567,
              "postDate": "2024-01-17T18:12:34.753Z",
              "content": "<p>Good point Rib. <a href=\"https://www.kaggle.com/sohier\" target=\"_blank\">@sohier</a> I think the question is how does Kaggle make the ground truth probabilities from the ground truth vote counts (during LB evaluation)?</p>",
              "rawMarkdown": "Good point Rib. @sohier I think the question is how does Kaggle make the ground truth probabilities from the ground truth vote counts (during LB evaluation)?",
              "votes": 3
            },
            {
              "id": 2606605,
              "postDate": "2024-01-17T18:38:08.323Z",
              "content": "<p>If there were 7 votes, 4 for seizure and 3 for other, the target values would be 4/7 and 3/7 respectively.</p>",
              "rawMarkdown": "If there were 7 votes, 4 for seizure and 3 for other, the target values would be 4/7 and 3/7 respectively.",
              "votes": 11
            },
            {
              "id": 2606636,
              "postDate": "2024-01-17T19:00:06.943Z",
              "content": "<p><a href=\"https://www.kaggle.com/sohier\" target=\"_blank\">@sohier</a> for egg_id = 11127485 as discussed <a href=\"https://www.kaggle.com/competitions/hms-harmful-brain-activity-classification/discussion/468208\" target=\"_blank\">here</a> </p>\n<p>seizure_vote = <strong>28/29</strong> and gpd_vote = <strong>1/29</strong> ( as we are aggregating all votes for each eeg_id from all its eeg_sub_ids) is this assumption right ? since test set have only <em>eeg_id, spectrogram_id, patient_id</em></p>\n<p><img src=\"https://www.googleapis.com/download/storage/v1/b/kaggle-forum-message-attachments/o/inbox%2F761268%2Fad8d8c0b230547a214cc7399f483845e%2FScreenshot%202024-01-16%20at%203.25.43AM.png?generation=1705355775031469&amp;alt=media\"></p>",
              "rawMarkdown": "@sohier for egg_id = 11127485 as discussed [here](https://www.kaggle.com/competitions/hms-harmful-brain-activity-classification/discussion/468208) \n\nseizure_vote = **28/29** and gpd_vote = **1/29** ( as we are aggregating all votes for each eeg_id from all its eeg_sub_ids) is this assumption right ? since test set have only *eeg_id, spectrogram_id, patient_id*\n\n![](https://www.googleapis.com/download/storage/v1/b/kaggle-forum-message-attachments/o/inbox%2F761268%2Fad8d8c0b230547a214cc7399f483845e%2FScreenshot%202024-01-16%20at%203.25.43AM.png?generation=1705355775031469&alt=media)"
            },
            {
              "id": 2606645,
              "postDate": "2024-01-17T19:07:00.133Z",
              "content": "<p>The test data only has 1 of each <code>eeg_id</code> in <code>test.csv</code>. So your question does not apply to the test data. Nor Kaggle's computation of our LB score.</p>",
              "rawMarkdown": "The test data only has 1 of each `eeg_id` in `test.csv`. So your question does not apply to the test data. Nor Kaggle's computation of our LB score.",
              "votes": 4
            },
            {
              "id": 2606655,
              "postDate": "2024-01-17T19:11:52.240Z",
              "content": "<p><a href=\"https://www.kaggle.com/cdeotte\" target=\"_blank\">@cdeotte</a> Understood, i miss these words <em>test_eegs</em>: <strong>Exactly 50 seconds of EEG data.</strong>. Thanks, my question is not relevant. <a href=\"https://www.kaggle.com/sohier\" target=\"_blank\">@sohier</a> ignore my question please.</p>",
              "rawMarkdown": "@cdeotte Understood, i miss these words *test_eegs*: **Exactly 50 seconds of EEG data.**. Thanks, my question is not relevant. @sohier ignore my question please.",
              "votes": 3
            }
          ]
        },
        {
          "id": 2611253,
          "postDate": "2024-01-20T16:34:41.233Z",
          "content": "<p><a href=\"https://www.kaggle.com/sohier\" target=\"_blank\">@sohier</a>, as <a href=\"https://en.wikipedia.org/wiki/Kullback%E2%80%93Leibler_divergence\" target=\"_blank\">KL</a> is not symmetric, the P will be expert probability and Q will be model predicted probability. Is that correct? Thank you,</p>",
          "rawMarkdown": "@sohier, as [KL](https://en.wikipedia.org/wiki/Kullback%E2%80%93Leibler_divergence) is not symmetric, the P will be expert probability and Q will be model predicted probability. Is that correct? Thank you,"
        }
      ]
    }
  ],
  "comments": [
    {
      "id": 2606507,
      "author_name": "Chris Deotte",
      "author_url": "",
      "post_date": "2024-01-17T17:27:21.093000",
      "content": "<p>No transformation. If you rows don't add up to 1, then your submission will throw an error</p>",
      "votes": 3,
      "replies": [
        {
          "id": 2606512,
          "author_name": "Rib~",
          "author_url": "",
          "post_date": "2024-01-17T17:32:49.473000",
          "content": "<p>What I mean is label, not pred.</p>",
          "votes": 1,
          "replies": []
        },
        {
          "id": 2607129,
          "author_name": "Cody_Null",
          "author_url": "",
          "post_date": "2024-01-18T04:20:09.880000",
          "content": "<p>Yep that explains a lot hahaha I have been losing my mind on what has caused these thrown exception errors</p>",
          "votes": 1,
          "replies": [
            {
              "id": 2607182,
              "author_name": "",
              "author_url": "",
              "post_date": "2024-01-18T05:20:24.133000",
              "content": "",
              "votes": 0,
              "replies": []
            }
          ]
        }
      ]
    },
    {
      "id": 2606531,
      "author_name": "Sohier Dane",
      "author_url": "",
      "post_date": "2024-01-17T17:54:45.123000",
      "content": "<p>The metric code is posted here: <a href=\"https://www.kaggle.com/code/metric/kullback-leibler-divergence/notebook\" target=\"_blank\">https://www.kaggle.com/code/metric/kullback-leibler-divergence/notebook</a></p>",
      "votes": 1,
      "replies": [
        {
          "id": 2606565,
          "author_name": "Rib~",
          "author_url": "",
          "post_date": "2024-01-17T18:10:49.273000",
          "content": "<p>I know about metrics, but still can't solve my question. We only know that the sum of labels is 1, but I have ten thousand ways to make the sum of labels be 1.</p>",
          "votes": 7,
          "replies": [
            {
              "id": 2606567,
              "author_name": "Chris Deotte",
              "author_url": "",
              "post_date": "2024-01-17T18:12:34.753000",
              "content": "<p>Good point Rib. <a href=\"https://www.kaggle.com/sohier\" target=\"_blank\">@sohier</a> I think the question is how does Kaggle make the ground truth probabilities from the ground truth vote counts (during LB evaluation)?</p>",
              "votes": 3,
              "replies": []
            },
            {
              "id": 2606605,
              "author_name": "Sohier Dane",
              "author_url": "",
              "post_date": "2024-01-17T18:38:08.323000",
              "content": "<p>If there were 7 votes, 4 for seizure and 3 for other, the target values would be 4/7 and 3/7 respectively.</p>",
              "votes": 11,
              "replies": []
            },
            {
              "id": 2606636,
              "author_name": "SeshuRaju 🧘‍♂️",
              "author_url": "",
              "post_date": "2024-01-17T19:00:06.943000",
              "content": "<p><a href=\"https://www.kaggle.com/sohier\" target=\"_blank\">@sohier</a> for egg_id = 11127485 as discussed <a href=\"https://www.kaggle.com/competitions/hms-harmful-brain-activity-classification/discussion/468208\" target=\"_blank\">here</a> </p>\n<p>seizure_vote = <strong>28/29</strong> and gpd_vote = <strong>1/29</strong> ( as we are aggregating all votes for each eeg_id from all its eeg_sub_ids) is this assumption right ? since test set have only <em>eeg_id, spectrogram_id, patient_id</em></p>\n<p><img src=\"https://www.googleapis.com/download/storage/v1/b/kaggle-forum-message-attachments/o/inbox%2F761268%2Fad8d8c0b230547a214cc7399f483845e%2FScreenshot%202024-01-16%20at%203.25.43AM.png?generation=1705355775031469&amp;alt=media\"></p>",
              "votes": 0,
              "replies": []
            },
            {
              "id": 2606645,
              "author_name": "Chris Deotte",
              "author_url": "",
              "post_date": "2024-01-17T19:07:00.133000",
              "content": "<p>The test data only has 1 of each <code>eeg_id</code> in <code>test.csv</code>. So your question does not apply to the test data. Nor Kaggle's computation of our LB score.</p>",
              "votes": 4,
              "replies": []
            },
            {
              "id": 2606655,
              "author_name": "SeshuRaju 🧘‍♂️",
              "author_url": "",
              "post_date": "2024-01-17T19:11:52.240000",
              "content": "<p><a href=\"https://www.kaggle.com/cdeotte\" target=\"_blank\">@cdeotte</a> Understood, i miss these words <em>test_eegs</em>: <strong>Exactly 50 seconds of EEG data.</strong>. Thanks, my question is not relevant. <a href=\"https://www.kaggle.com/sohier\" target=\"_blank\">@sohier</a> ignore my question please.</p>",
              "votes": 3,
              "replies": []
            }
          ]
        },
        {
          "id": 2611253,
          "author_name": "Zhenlan",
          "author_url": "",
          "post_date": "2024-01-20T16:34:41.233000",
          "content": "<p><a href=\"https://www.kaggle.com/sohier\" target=\"_blank\">@sohier</a>, as <a href=\"https://en.wikipedia.org/wiki/Kullback%E2%80%93Leibler_divergence\" target=\"_blank\">KL</a> is not symmetric, the P will be expert probability and Q will be model predicted probability. Is that correct? Thank you,</p>",
          "votes": 0,
          "replies": []
        }
      ]
    }
  ],
  "raw_markdown_by_id": {
    "2606497": "Which transformation is used for the label during official evaluation? label/sum(label)?softmax(label)?or?",
    "2606507": "No transformation. If you rows don't add up to 1, then your submission will throw an error",
    "2606531": "The metric code is posted here: https://www.kaggle.com/code/metric/kullback-leibler-divergence/notebook"
  }
}