{
  "id": 641906,
  "title": "Question about mapping the five scrolls from the website to train.csv IDs",
  "url": "/competitions/vesuvius-challenge-surface-detection/discussion/641906",
  "author_name": "",
  "post_date": "2025-11-27T12:42:38.561364Z",
  "votes": 1,
  "comment_count": 1,
  "views": 0,
  "content": "<p>Hello everyone,</p>\n<p>On the <a href=\"https://dl.ash2txt.org/full-scrolls/\" target=\"_blank\">website</a> there are five scrolls listed in the data section. Are these the same scrolls included in the competition dataset? If so, how do they correspond to the scroll IDs found in <code>train.csv</code>?</p>\n<p>Best regards,\nBenjamin</p>",
  "messages": [
    {
      "id": "3350235",
      "postDate": "11/27/2025 12:42:38",
      "content": "<p>Hello everyone,</p>\n<p>On the <a href=\"https://dl.ash2txt.org/full-scrolls/\" target=\"_blank\">website</a> there are five scrolls listed in the data section. Are these the same scrolls included in the competition dataset? If so, how do they correspond to the scroll IDs found in <code>train.csv</code>?</p>\n<p>Best regards,\nBenjamin</p>",
      "rawMarkdown": "Hello everyone,\n\nOn the [website](https://dl.ash2txt.org/full-scrolls/) there are five scrolls listed in the data section. Are these the same scrolls included in the competition dataset? If so, how do they correspond to the scroll IDs found in `train.csv`?\n\nBest regards,\nBenjamin",
      "votes": null
    },
    {
      "id": "3351129",
      "postDate": "11/28/2025 07:05:44",
      "content": "<p>Hi,</p>\n<p>Some of the released scrolls are included in the competition dataset, however the Kaggle team prefers to keep the mapping secret.\nFor your information, we scanned more than 30 scrolls, and most of them are unreleased, and data from these unreleased scrolls is part of the test set.</p>\n<p>We also have additional data <a href=\"https://data.aws.ash2txt.org/samples/\" target=\"_blank\">in this other repository</a>. I know it's a bit confusing, but our project is a continuous work in progress. You can access to these scans in several ways, but one way is following the instructions in <a href=\"https://www.kaggle.com/code/giorgioangelotti/access-to-additional-unlabeled-data\" target=\"_blank\">this notebook</a></p>",
      "rawMarkdown": "Hi,\n\nSome of the released scrolls are included in the competition dataset, however the Kaggle team prefers to keep the mapping secret.\nFor your information, we scanned more than 30 scrolls, and most of them are unreleased, and data from these unreleased scrolls is part of the test set.\n\nWe also have additional data [in this other repository](https://data.aws.ash2txt.org/samples/). I know it's a bit confusing, but our project is a continuous work in progress. You can access to these scans in several ways, but one way is following the instructions in [this notebook](https://www.kaggle.com/code/giorgioangelotti/access-to-additional-unlabeled-data)",
      "votes": null
    }
  ],
  "comments": [
    {
      "id": 3351129,
      "author_name": "giorgioangelotti",
      "author_url": "",
      "post_date": "11/28/2025 07:05:44",
      "content": "<p>Hi,</p>\n<p>Some of the released scrolls are included in the competition dataset, however the Kaggle team prefers to keep the mapping secret.\nFor your information, we scanned more than 30 scrolls, and most of them are unreleased, and data from these unreleased scrolls is part of the test set.</p>\n<p>We also have additional data <a href=\"https://data.aws.ash2txt.org/samples/\" target=\"_blank\">in this other repository</a>. I know it's a bit confusing, but our project is a continuous work in progress. You can access to these scans in several ways, but one way is following the instructions in <a href=\"https://www.kaggle.com/code/giorgioangelotti/access-to-additional-unlabeled-data\" target=\"_blank\">this notebook</a></p>",
      "votes": null,
      "replies": []
    }
  ],
  "raw_markdown_by_id": {
    "3350235": "Hello everyone,\n\nOn the [website](https://dl.ash2txt.org/full-scrolls/) there are five scrolls listed in the data section. Are these the same scrolls included in the competition dataset? If so, how do they correspond to the scroll IDs found in `train.csv`?\n\nBest regards,\nBenjamin",
    "3351129": "Hi,\n\nSome of the released scrolls are included in the competition dataset, however the Kaggle team prefers to keep the mapping secret.\nFor your information, we scanned more than 30 scrolls, and most of them are unreleased, and data from these unreleased scrolls is part of the test set.\n\nWe also have additional data [in this other repository](https://data.aws.ash2txt.org/samples/). I know it's a bit confusing, but our project is a continuous work in progress. You can access to these scans in several ways, but one way is following the instructions in [this notebook](https://www.kaggle.com/code/giorgioangelotti/access-to-additional-unlabeled-data)"
  },
  "source": "meta"
}