{
  "id": 501659,
  "title": "What's the use of 'rating' column in the metadata?",
  "url": "/competitions/birdclef-2024/discussion/501659",
  "author_name": "tanxxx",
  "post_date": "2024-05-10T08:30:44.143000",
  "votes": 1,
  "comment_count": 6,
  "views": 0,
  "content": "<p>Does this column reflect the audio quality? Will using this column as an objective in multi-task learning make the network more robust?🧐</p>",
  "messages": [
    {
      "id": 2804846,
      "postDate": "2024-05-10T08:30:44.143Z",
      "content": "<p>Does this column reflect the audio quality? Will using this column as an objective in multi-task learning make the network more robust?🧐</p>",
      "rawMarkdown": "Does this column reflect the audio quality? Will using this column as an objective in multi-task learning make the network more robust?🧐",
      "votes": 1
    },
    {
      "id": 2806069,
      "postDate": "2024-05-10T21:06:38.123Z",
      "content": "<p>I tried to weight based off this rating but ultimately found no success with it. The CV score increases but the LB score falls off massively. I think it is just because of the quality comparison between the 2. As the competition leaders pointed out at one point. The CV is like you are listening for a bird in a room with only you and the bird. The test set is like standing in a garden and trying to listen for the bird. </p>",
      "rawMarkdown": "I tried to weight based off this rating but ultimately found no success with it. The CV score increases but the LB score falls off massively. I think it is just because of the quality comparison between the 2. As the competition leaders pointed out at one point. The CV is like you are listening for a bird in a room with only you and the bird. The test set is like standing in a garden and trying to listen for the bird. ",
      "votes": 2,
      "replies": [
        {
          "id": 2806291,
          "postDate": "2024-05-11T02:41:59.767Z",
          "content": "<p>Thanks Cody, but what's the meaning of 'weight based off the rating' ? does it mean the CV split strategy or mixup like weight? Sorry i am still a bit confused.</p>",
          "rawMarkdown": "Thanks Cody, but what's the meaning of 'weight based off the rating' ? does it mean the CV split strategy or mixup like weight? Sorry i am still a bit confused.",
          "votes": 1,
          "replies": [
            {
              "id": 2806352,
              "postDate": "2024-05-11T03:56:55.240Z",
              "rawMarkdown": "",
              "isDeleted": true
            },
            {
              "id": 2806704,
              "postDate": "2024-05-11T08:22:44.590Z",
              "content": "<p>Thanks Cody, I understand what you mean. Is the rating value marked by many different people? Or is it obtained by some objective rules? So if it is marked by different people, will it be less objective?</p>",
              "rawMarkdown": "Thanks Cody, I understand what you mean. Is the rating value marked by many different people? Or is it obtained by some objective rules? So if it is marked by different people, will it be less objective?"
            },
            {
              "id": 2807024,
              "postDate": "2024-05-11T13:04:14.567Z",
              "content": "<p>Since the rating is given by the Xeno-Canto community, it appears to be given by more than one name.<br>\n<a href=\"https://www.kaggle.com/competitions/birdclef-2024/discussion/491145\" target=\"_blank\">https://www.kaggle.com/competitions/birdclef-2024/discussion/491145</a></p>\n<p>Also, the rating itself seems to indicate whether or not there is noise other than the primary label bird, not the sound quality.<br>\n<a href=\"https://www.kaggle.com/competitions/birdclef-2024/discussion/493605\" target=\"_blank\">https://www.kaggle.com/competitions/birdclef-2024/discussion/493605</a></p>\n<p>In BirdCLEF2021, there is a top solution that uses mixup with rating as weight, so it may be a metadata that can increase the score if used well.<br>\n<a href=\"https://www.kaggle.com/competitions/birdclef-2021/discussion/243463\" target=\"_blank\">https://www.kaggle.com/competitions/birdclef-2021/discussion/243463</a></p>",
              "rawMarkdown": "Since the rating is given by the Xeno-Canto community, it appears to be given by more than one name.\nhttps://www.kaggle.com/competitions/birdclef-2024/discussion/491145\n\nAlso, the rating itself seems to indicate whether or not there is noise other than the primary label bird, not the sound quality.\nhttps://www.kaggle.com/competitions/birdclef-2024/discussion/493605\n\nIn BirdCLEF2021, there is a top solution that uses mixup with rating as weight, so it may be a metadata that can increase the score if used well.\nhttps://www.kaggle.com/competitions/birdclef-2021/discussion/243463"
            },
            {
              "id": 2809769,
              "postDate": "2024-05-13T01:37:42.290Z",
              "content": "<p>Thanks tc for your detailed reply!</p>",
              "rawMarkdown": "Thanks tc for your detailed reply!"
            }
          ]
        }
      ]
    }
  ],
  "comments": [
    {
      "id": 2806069,
      "author_name": "Cody_Null",
      "author_url": "",
      "post_date": "2024-05-10T21:06:38.123000",
      "content": "<p>I tried to weight based off this rating but ultimately found no success with it. The CV score increases but the LB score falls off massively. I think it is just because of the quality comparison between the 2. As the competition leaders pointed out at one point. The CV is like you are listening for a bird in a room with only you and the bird. The test set is like standing in a garden and trying to listen for the bird. </p>",
      "votes": 2,
      "replies": [
        {
          "id": 2806291,
          "author_name": "tanxxx",
          "author_url": "",
          "post_date": "2024-05-11T02:41:59.767000",
          "content": "<p>Thanks Cody, but what's the meaning of 'weight based off the rating' ? does it mean the CV split strategy or mixup like weight? Sorry i am still a bit confused.</p>",
          "votes": 1,
          "replies": [
            {
              "id": 2806352,
              "author_name": "",
              "author_url": "",
              "post_date": "2024-05-11T03:56:55.240000",
              "content": "",
              "votes": 0,
              "replies": []
            },
            {
              "id": 2806704,
              "author_name": "tanxxx",
              "author_url": "",
              "post_date": "2024-05-11T08:22:44.590000",
              "content": "<p>Thanks Cody, I understand what you mean. Is the rating value marked by many different people? Or is it obtained by some objective rules? So if it is marked by different people, will it be less objective?</p>",
              "votes": 0,
              "replies": []
            },
            {
              "id": 2807024,
              "author_name": "tc0000",
              "author_url": "",
              "post_date": "2024-05-11T13:04:14.567000",
              "content": "<p>Since the rating is given by the Xeno-Canto community, it appears to be given by more than one name.<br>\n<a href=\"https://www.kaggle.com/competitions/birdclef-2024/discussion/491145\" target=\"_blank\">https://www.kaggle.com/competitions/birdclef-2024/discussion/491145</a></p>\n<p>Also, the rating itself seems to indicate whether or not there is noise other than the primary label bird, not the sound quality.<br>\n<a href=\"https://www.kaggle.com/competitions/birdclef-2024/discussion/493605\" target=\"_blank\">https://www.kaggle.com/competitions/birdclef-2024/discussion/493605</a></p>\n<p>In BirdCLEF2021, there is a top solution that uses mixup with rating as weight, so it may be a metadata that can increase the score if used well.<br>\n<a href=\"https://www.kaggle.com/competitions/birdclef-2021/discussion/243463\" target=\"_blank\">https://www.kaggle.com/competitions/birdclef-2021/discussion/243463</a></p>",
              "votes": 0,
              "replies": []
            },
            {
              "id": 2809769,
              "author_name": "tanxxx",
              "author_url": "",
              "post_date": "2024-05-13T01:37:42.290000",
              "content": "<p>Thanks tc for your detailed reply!</p>",
              "votes": 0,
              "replies": []
            }
          ]
        }
      ]
    }
  ],
  "raw_markdown_by_id": {
    "2804846": "Does this column reflect the audio quality? Will using this column as an objective in multi-task learning make the network more robust?🧐",
    "2806069": "I tried to weight based off this rating but ultimately found no success with it. The CV score increases but the LB score falls off massively. I think it is just because of the quality comparison between the 2. As the competition leaders pointed out at one point. The CV is like you are listening for a bird in a room with only you and the bird. The test set is like standing in a garden and trying to listen for the bird. "
  }
}