{
  "id": 574461,
  "title": "Questions Regarding Data Labels",
  "url": "/competitions/beyond-visible-spectrum-ai-for-agriculture-2025/discussion/574461",
  "author_name": "",
  "post_date": "2025-04-22T03:56:12.807600300Z",
  "votes": 2,
  "comment_count": 2,
  "views": 0,
  "content": "<p>I noticed that the data description mentions that the labels are divided into three categories, but when I checked the train.csv file, the labels range from 0 to 100. Could you please clarify the following:</p>\n<p>1.Does this mean that the labels in the dataset represent continuous values rather than discrete categories?</p>\n<p>2.If the labels are indeed continuous, how should we handle them for classification tasks, as they are described as belonging to three categories in the documentation?</p>\n<p>3.Is there any specific preprocessing or transformation needed for the labels to align with the three categories mentioned in the data description?</p>\n<p>I would appreciate any clarification on this, as it will greatly help in understanding the task and proceeding with the model development.</p>",
  "messages": [
    {
      "id": "3184407",
      "postDate": "04/22/2025 03:56:12",
      "content": "<p>I noticed that the data description mentions that the labels are divided into three categories, but when I checked the train.csv file, the labels range from 0 to 100. Could you please clarify the following:</p>\n<p>1.Does this mean that the labels in the dataset represent continuous values rather than discrete categories?</p>\n<p>2.If the labels are indeed continuous, how should we handle them for classification tasks, as they are described as belonging to three categories in the documentation?</p>\n<p>3.Is there any specific preprocessing or transformation needed for the labels to align with the three categories mentioned in the data description?</p>\n<p>I would appreciate any clarification on this, as it will greatly help in understanding the task and proceeding with the model development.</p>",
      "rawMarkdown": "I noticed that the data description mentions that the labels are divided into three categories, but when I checked the train.csv file, the labels range from 0 to 100. Could you please clarify the following:\n\n1.Does this mean that the labels in the dataset represent continuous values rather than discrete categories?\n\n2.If the labels are indeed continuous, how should we handle them for classification tasks, as they are described as belonging to three categories in the documentation?\n\n3.Is there any specific preprocessing or transformation needed for the labels to align with the three categories mentioned in the data description?\n\nI would appreciate any clarification on this, as it will greatly help in understanding the task and proceeding with the model development.",
      "votes": null
    },
    {
      "id": "3184632",
      "postDate": "04/22/2025 09:56:50",
      "content": "<p>I have modified the description of the dataset. The task is to predict the percentage of crop disease in patched image from 1 to 100(int)</p>",
      "rawMarkdown": "I have modified the description of the dataset. The task is to predict the percentage of crop disease in patched image from 1 to 100(int)",
      "votes": null
    },
    {
      "id": "3213429",
      "postDate": "05/30/2025 02:14:04",
      "content": "<p>Specifically, could you please confirm which matrix (such as NDVI) is used as the label for the training data? </p>",
      "rawMarkdown": "Specifically, could you please confirm which matrix (such as NDVI) is used as the label for the training data?",
      "votes": null
    }
  ],
  "comments": [
    {
      "id": 3184632,
      "author_name": "robeson",
      "author_url": "",
      "post_date": "04/22/2025 09:56:50",
      "content": "<p>I have modified the description of the dataset. The task is to predict the percentage of crop disease in patched image from 1 to 100(int)</p>",
      "votes": null,
      "replies": [
        {
          "id": 3213429,
          "author_name": "masudrana71",
          "author_url": "",
          "post_date": "05/30/2025 02:14:04",
          "content": "<p>Specifically, could you please confirm which matrix (such as NDVI) is used as the label for the training data? </p>",
          "votes": null,
          "replies": []
        }
      ]
    }
  ],
  "raw_markdown_by_id": {
    "3184407": "I noticed that the data description mentions that the labels are divided into three categories, but when I checked the train.csv file, the labels range from 0 to 100. Could you please clarify the following:\n\n1.Does this mean that the labels in the dataset represent continuous values rather than discrete categories?\n\n2.If the labels are indeed continuous, how should we handle them for classification tasks, as they are described as belonging to three categories in the documentation?\n\n3.Is there any specific preprocessing or transformation needed for the labels to align with the three categories mentioned in the data description?\n\nI would appreciate any clarification on this, as it will greatly help in understanding the task and proceeding with the model development.",
    "3184632": "I have modified the description of the dataset. The task is to predict the percentage of crop disease in patched image from 1 to 100(int)",
    "3213429": "Specifically, could you please confirm which matrix (such as NDVI) is used as the label for the training data?"
  },
  "source": "meta"
}