{
  "id": 42589,
  "title": "1GB Training Data",
  "url": "/competitions/cdiscount-image-classification-challenge/discussion/42589",
  "author_name": "",
  "post_date": "2017-11-01T22:27:58.602145Z",
  "votes": null,
  "comment_count": 2,
  "views": 0,
  "content": "<p>Hi All, </p>\n\n<p>I was wondering if the competition organizers could create a larger training sample that has about 1GB of training data. And similarly, a test dataset. Those could be used to build our code before we push it to the server machines like: AWS or Google Cloud.</p>\n\n<p>Is there chance that you could provide that?</p>\n\n<p>Regards,\nAmro</p>",
  "messages": [
    {
      "id": "238677",
      "postDate": "11/01/2017 22:27:58",
      "content": "<p>Hi All, </p>\n\n<p>I was wondering if the competition organizers could create a larger training sample that has about 1GB of training data. And similarly, a test dataset. Those could be used to build our code before we push it to the server machines like: AWS or Google Cloud.</p>\n\n<p>Is there chance that you could provide that?</p>\n\n<p>Regards,\nAmro</p>",
      "rawMarkdown": "Hi All, \n\nI was wondering if the competition organizers could create a larger training sample that has about 1GB of training data. And similarly, a test dataset. Those could be used to build our code before we push it to the server machines like: AWS or Google Cloud.\n\nIs there chance that you could provide that?\n\nRegards,\nAmro",
      "votes": null
    },
    {
      "id": "239336",
      "postDate": "11/03/2017 08:03:12",
      "content": "<p>You can make it by yourself.<br>\nThis kernel can stop anytime after extracting the some part of data.<br>\n<a href=\"https://www.kaggle.com/inversion/processing-bson-files\">https://www.kaggle.com/inversion/processing-bson-files</a></p>",
      "rawMarkdown": "You can make it by yourself.<br>\nThis kernel can stop anytime after extracting the some part of data.<br>\nhttps://www.kaggle.com/inversion/processing-bson-files",
      "votes": null
    },
    {
      "id": "239535",
      "postDate": "11/03/2017 18:05:54",
      "content": "<p>Hi Toshi, I have kernel that does that. The problem is that the total size limit is less than 512 MB. And the file generated is not downloadable.</p>",
      "rawMarkdown": "Hi Toshi, I have kernel that does that. The problem is that the total size limit is less than 512 MB. And the file generated is not downloadable.",
      "votes": null
    }
  ],
  "comments": [
    {
      "id": 239336,
      "author_name": "toshik",
      "author_url": "",
      "post_date": "11/03/2017 08:03:12",
      "content": "<p>You can make it by yourself.<br>\nThis kernel can stop anytime after extracting the some part of data.<br>\n<a href=\"https://www.kaggle.com/inversion/processing-bson-files\">https://www.kaggle.com/inversion/processing-bson-files</a></p>",
      "votes": null,
      "replies": [
        {
          "id": 239535,
          "author_name": "am1to2",
          "author_url": "",
          "post_date": "11/03/2017 18:05:54",
          "content": "<p>Hi Toshi, I have kernel that does that. The problem is that the total size limit is less than 512 MB. And the file generated is not downloadable.</p>",
          "votes": null,
          "replies": []
        }
      ]
    }
  ],
  "raw_markdown_by_id": {
    "238677": "Hi All, \n\nI was wondering if the competition organizers could create a larger training sample that has about 1GB of training data. And similarly, a test dataset. Those could be used to build our code before we push it to the server machines like: AWS or Google Cloud.\n\nIs there chance that you could provide that?\n\nRegards,\nAmro",
    "239336": "You can make it by yourself.<br>\nThis kernel can stop anytime after extracting the some part of data.<br>\nhttps://www.kaggle.com/inversion/processing-bson-files",
    "239535": "Hi Toshi, I have kernel that does that. The problem is that the total size limit is less than 512 MB. And the file generated is not downloadable."
  },
  "source": "meta"
}