{
  "id": 97774,
  "title": "Previous competition data unusable",
  "url": "/competitions/aptos2019-blindness-detection/discussion/97774",
  "author_name": "xhlulu",
  "post_date": "2019-06-28T19:43:23.664000",
  "votes": 6,
  "comment_count": 23,
  "views": 0,
  "content": "<p>I've taken a look at this previous competition: <a href=\"https://www.kaggle.com/c/diabetic-retinopathy-detection/data\">https://www.kaggle.com/c/diabetic-retinopathy-detection/data</a></p>\n\n<p>Apart from the 1000 sample images, It seems that most of the data (~20,000 images) is unusable. Although the zip file was broken down into multiple parts, I don't seem to be able to recombine them when I load them inside a kernel.</p>\n\n<p>Has anyone been able to use it?</p>",
  "messages": [
    {
      "id": 565134,
      "postDate": "2019-06-30T13:38:37.853Z",
      "content": "<p>I uploaded 896x896 resized <a href=\"https://www.kaggle.com/donkeys/retinopathy-train-2015\">dataset</a> with combined training data and the labels from the previous one.</p>\n\n<p>I tried the 1024x1024 but it was about 23GB so a bit large for the 20GB size limit. Also I made a <a href=\"https://www.kaggle.com/donkeys/looking-in-the-eyes-of-past-and-present\">kernel</a> with a few visualizations of it. Maybe a different image format or some other processing could keep it a bit larger but it lets me play with it. </p>",
      "rawMarkdown": "I uploaded 896x896 resized [dataset](https://www.kaggle.com/donkeys/retinopathy-train-2015) with combined training data and the labels from the previous one.\n\nI tried the 1024x1024 but it was about 23GB so a bit large for the 20GB size limit. Also I made a [kernel](https://www.kaggle.com/donkeys/looking-in-the-eyes-of-past-and-present) with a few visualizations of it. Maybe a different image format or some other processing could keep it a bit larger but it lets me play with it. ",
      "votes": 10,
      "replies": [
        {
          "id": 565475,
          "postDate": "2019-07-01T02:02:54.240Z",
          "content": "<p>Thank you for your good work! I will downsample them to 224 for imagenet-like models.</p>\n\n<p>Your public quota is 20 GB; are you worried about not being able to upload datasets in the future?</p>",
          "rawMarkdown": "Thank you for your good work! I will downsample them to 224 for imagenet-like models.\n\n Your public quota is 20 GB; are you worried about not being able to upload datasets in the future?",
          "votes": 1
        },
        {
          "id": 565745,
          "postDate": "2019-07-01T10:24:37.503Z",
          "content": "<p>Glad you like it! </p>\n\n<p>As far as I understand the private quota is 20GB but there is no quota for public datasets. So I hope it has no impact :)</p>",
          "rawMarkdown": "Glad you like it! \n\nAs far as I understand the private quota is 20GB but there is no quota for public datasets. So I hope it has no impact :)",
          "votes": 1
        },
        {
          "id": 565886,
          "postDate": "2019-07-01T13:50:48.293Z",
          "content": "<p>You are totally right! I forgot about that!</p>",
          "rawMarkdown": "You are totally right! I forgot about that!"
        }
      ]
    },
    {
      "id": 565498,
      "postDate": "2019-07-01T03:04:47.110Z",
      "content": "<p>I made the dataset available here:\n<a href=\"https://www.kaggle.com/c/aptos2019-blindness-detection/discussion/98059\">https://www.kaggle.com/c/aptos2019-blindness-detection/discussion/98059</a></p>\n\n<p>I have two versions. One that is simply resized, and another that is cropped first by finding the mask of the fundus. This cropping will hopefully improve the score, as I have noticed it has helped in my own experiments with the previous competition dataset.</p>\n\n<p>Sorry it was later than I promised. Unfortunately, I also didn't have time to work on training a model for this competition using the previous competition data. But that should be coming out in the next couple of days.</p>",
      "rawMarkdown": "I made the dataset available here:\nhttps://www.kaggle.com/c/aptos2019-blindness-detection/discussion/98059\n\nI have two versions. One that is simply resized, and another that is cropped first by finding the mask of the fundus. This cropping will hopefully improve the score, as I have noticed it has helped in my own experiments with the previous competition dataset.\n\nSorry it was later than I promised. Unfortunately, I also didn't have time to work on training a model for this competition using the previous competition data. But that should be coming out in the next couple of days.",
      "votes": 8
    },
    {
      "id": 563939,
      "postDate": "2019-06-28T19:43:23.663Z",
      "content": "<p>I've taken a look at this previous competition: <a href=\"https://www.kaggle.com/c/diabetic-retinopathy-detection/data\">https://www.kaggle.com/c/diabetic-retinopathy-detection/data</a></p>\n\n<p>Apart from the 1000 sample images, It seems that most of the data (~20,000 images) is unusable. Although the zip file was broken down into multiple parts, I don't seem to be able to recombine them when I load them inside a kernel.</p>\n\n<p>Has anyone been able to use it?</p>",
      "rawMarkdown": "I've taken a look at this previous competition: https://www.kaggle.com/c/diabetic-retinopathy-detection/data\n\nApart from the 1000 sample images, It seems that most of the data (~20,000 images) is unusable. Although the zip file was broken down into multiple parts, I don't seem to be able to recombine them when I load them inside a kernel.\n\nHas anyone been able to use it?",
      "votes": 6
    },
    {
      "id": 564049,
      "postDate": "2019-06-28T23:38:44.467Z",
      "content": "<p>I plan to make preprocessed versions of the previous competition data by tomorrow.</p>",
      "rawMarkdown": "I plan to make preprocessed versions of the previous competition data by tomorrow.",
      "votes": 5,
      "replies": [
        {
          "id": 564264,
          "postDate": "2019-06-29T08:00:13.910Z",
          "content": "<p>Thank you ! </p>",
          "rawMarkdown": "Thank you ! "
        },
        {
          "id": 564265,
          "postDate": "2019-06-29T08:00:33.710Z",
          "content": "<p><a href=\"/tanlikesmath\">@tanlikesmath</a> thanks for that. Can you please make it public? \nThis competition data will hit plateau as because of small size. Then we need to use that dataset.</p>",
          "rawMarkdown": "@tanlikesmath thanks for that. Can you please make it public? \nThis competition data will hit plateau as because of small size. Then we need to use that dataset."
        },
        {
          "id": 564392,
          "postDate": "2019-06-29T11:04:16.943Z",
          "content": "<p>Yes I will i will first make my submission with it and write a kernel on how to use it and then make it public. I have the dataset already created from previous work and is currently private but once I get it to work with this competition I will make it public! 😃</p>",
          "rawMarkdown": "Yes I will i will first make my submission with it and write a kernel on how to use it and then make it public. I have the dataset already created from previous work and is currently private but once I get it to work with this competition I will make it public! 😃",
          "votes": 1
        },
        {
          "id": 564716,
          "postDate": "2019-06-29T21:16:27.137Z",
          "content": "<p><a href=\"/tanlikesmath\">@tanlikesmath</a> thank you for your good work. What size will it be (e.g. 224x224)?</p>",
          "rawMarkdown": "@tanlikesmath thank you for your good work. What size will it be (e.g. 224x224)?"
        },
        {
          "id": 564740,
          "postDate": "2019-06-29T22:27:35.317Z",
          "content": "<p><a href=\"/xhlulu\">@xhlulu</a> it is 1024x1024. You can further resize the pictures later if you want. </p>",
          "rawMarkdown": "@xhlulu it is 1024x1024. You can further resize the pictures later if you want. ",
          "votes": 2
        }
      ]
    },
    {
      "id": 564883,
      "postDate": "2019-06-30T05:52:47.657Z",
      "content": "<p>I can't use it in my kernel, but I can use it in local.\n<a href=\"https://www.kaggle.com/c/aptos2019-blindness-detection/discussion/97947#latest-564872\">https://www.kaggle.com/c/aptos2019-blindness-detection/discussion/97947#latest-564872</a></p>\n\n<p>In this competition we don't need to train in kernel. So it is better not to load this data in your kernel because it takes a lot of time.</p>",
      "rawMarkdown": "I can't use it in my kernel, but I can use it in local.\nhttps://www.kaggle.com/c/aptos2019-blindness-detection/discussion/97947#latest-564872\n\nIn this competition we don't need to train in kernel. So it is better not to load this data in your kernel because it takes a lot of time.",
      "votes": 2,
      "replies": [
        {
          "id": 564890,
          "postDate": "2019-06-30T06:08:04.900Z",
          "content": "<p>Well worked ! \nThank you !</p>",
          "rawMarkdown": "Well worked ! \nThank you !",
          "votes": 1
        },
        {
          "id": 564900,
          "postDate": "2019-06-30T06:21:27.490Z",
          "content": "<p><a href=\"/bluexleoxgreen\">@bluexleoxgreen</a> a fix for this coming soon ;)</p>",
          "rawMarkdown": "@bluexleoxgreen a fix for this coming soon ;)",
          "votes": 1
        },
        {
          "id": 565215,
          "postDate": "2019-06-30T15:40:32.913Z",
          "content": "<p>&gt; a fix for this coming soon ;)</p>\n\n<p><a href=\"/tanlikesmath\">@tanlikesmath</a> waiting for that.. \nHope we can do all the things here on kernels. Many people don't have GPUs\nOnly labeled training data (no test images) also would be fine</p>",
          "rawMarkdown": "&gt; a fix for this coming soon ;)\n\n@tanlikesmath waiting for that.. \nHope we can do all the things here on kernels. Many people don't have GPUs\nOnly labeled training data (no test images) also would be fine",
          "votes": 2
        },
        {
          "id": 565278,
          "postDate": "2019-06-30T17:37:48.073Z",
          "content": "<p><a href=\"/prashantkikani\">@prashantkikani</a> \nI think the labels of test data are also published in the previous competition. <br>\nRetionopathy_solution csv file includes both labels. <br>\nFor more detail, see my above post.  </p>",
          "rawMarkdown": "@prashantkikani \nI think the labels of test data are also published in the previous competition.  \nRetionopathy_solution csv file includes both labels.   \nFor more detail, see my above post.  ",
          "votes": 1
        }
      ]
    },
    {
      "id": 564854,
      "postDate": "2019-06-30T04:49:16.263Z",
      "content": "<p>Updated: <br>\nAs <a href=\"/bluexleoxgreen\">@bluexleoxgreen</a> adviced me, these files are just divided files. <br>\nNow I became to extract files. <br>\nThanks <a href=\"/bluexleoxgreen\">@bluexleoxgreen</a> !</p>\n\n<p><strong>I deleted my post in external dataset thread not to confuse.</strong></p>\n\n<hr>\n\n<p><a href=\"/xhlulu\">@xhlulu</a> \nThank you for positing this issue. I have also faced this problem(I could not unzip).  </p>\n\n<p>Just to be sure, I have reported this problem in external dataset thread.\n<a href=\"https://www.kaggle.com/c/aptos2019-blindness-detection/discussion/97605#latest-564847\">https://www.kaggle.com/c/aptos2019-blindness-detection/discussion/97605#latest-564847</a></p>",
      "rawMarkdown": "Updated:  \nAs @bluexleoxgreen adviced me, these files are just divided files.  \nNow I became to extract files.  \nThanks @bluexleoxgreen !\n  \n**I deleted my post in external dataset thread not to confuse.**\n\n---\n\n@xhlulu \nThank you for positing this issue. I have also faced this problem(I could not unzip).  \n\nJust to be sure, I have reported this problem in external dataset thread.\nhttps://www.kaggle.com/c/aptos2019-blindness-detection/discussion/97605#latest-564847",
      "votes": 2
    },
    {
      "id": 566171,
      "postDate": "2019-07-01T21:34:33.643Z",
      "content": "<p>You'll need 7zip or Keka (if you're on a Mac) to be able to extract all the volumes in the old dataset. Due to size limitations, the data is broken into parts/volumes named train.zip.001 - train.zip.005.</p>",
      "rawMarkdown": "You'll need 7zip or Keka (if you're on a Mac) to be able to extract all the volumes in the old dataset. Due to size limitations, the data is broken into parts/volumes named train.zip.001 - train.zip.005."
    },
    {
      "id": 564874,
      "postDate": "2019-06-30T05:32:19.743Z",
      "content": "<p>You need 7-Zip to unzip those files. </p>",
      "rawMarkdown": "You need 7-Zip to unzip those files. ",
      "replies": [
        {
          "id": 564886,
          "postDate": "2019-06-30T05:55:38.500Z",
          "rawMarkdown": "",
          "isDeleted": true
        }
      ]
    },
    {
      "id": 564735,
      "postDate": "2019-06-29T22:09:28.593Z",
      "content": "<p>What do you mean by unusable? Also, are you doing an comparison of images from the two diff datasets? If so, how are you doing it?</p>",
      "rawMarkdown": "What do you mean by unusable? Also, are you doing an comparison of images from the two diff datasets? If so, how are you doing it?",
      "replies": [
        {
          "id": 565473,
          "postDate": "2019-07-01T02:00:24.427Z",
          "content": "<p>I'm not sure what you mean by comparison of images. They are unusable because when you open a kernel with the previous competition's dataset, you are unable to even locate the files using <code>os.listdir</code>.</p>",
          "rawMarkdown": "I'm not sure what you mean by comparison of images. They are unusable because when you open a kernel with the previous competition's dataset, you are unable to even locate the files using `os.listdir`."
        }
      ]
    },
    {
      "id": 565125,
      "postDate": "2019-06-30T13:23:38.963Z",
      "rawMarkdown": "",
      "isDeleted": true
    }
  ],
  "comments": [
    {
      "id": 565134,
      "author_name": "averagemn",
      "author_url": "",
      "post_date": "2019-06-30T13:38:37.853000",
      "content": "<p>I uploaded 896x896 resized <a href=\"https://www.kaggle.com/donkeys/retinopathy-train-2015\">dataset</a> with combined training data and the labels from the previous one.</p>\n\n<p>I tried the 1024x1024 but it was about 23GB so a bit large for the 20GB size limit. Also I made a <a href=\"https://www.kaggle.com/donkeys/looking-in-the-eyes-of-past-and-present\">kernel</a> with a few visualizations of it. Maybe a different image format or some other processing could keep it a bit larger but it lets me play with it. </p>",
      "votes": 10,
      "replies": [
        {
          "id": 565475,
          "author_name": "xhlulu",
          "author_url": "",
          "post_date": "2019-07-01T02:02:54.240000",
          "content": "<p>Thank you for your good work! I will downsample them to 224 for imagenet-like models.</p>\n\n<p>Your public quota is 20 GB; are you worried about not being able to upload datasets in the future?</p>",
          "votes": 1,
          "replies": []
        },
        {
          "id": 565745,
          "author_name": "averagemn",
          "author_url": "",
          "post_date": "2019-07-01T10:24:37.503000",
          "content": "<p>Glad you like it! </p>\n\n<p>As far as I understand the private quota is 20GB but there is no quota for public datasets. So I hope it has no impact :)</p>",
          "votes": 1,
          "replies": []
        },
        {
          "id": 565886,
          "author_name": "xhlulu",
          "author_url": "",
          "post_date": "2019-07-01T13:50:48.293000",
          "content": "<p>You are totally right! I forgot about that!</p>",
          "votes": 0,
          "replies": []
        }
      ]
    },
    {
      "id": 565498,
      "author_name": "ilovescience",
      "author_url": "",
      "post_date": "2019-07-01T03:04:47.110000",
      "content": "<p>I made the dataset available here:\n<a href=\"https://www.kaggle.com/c/aptos2019-blindness-detection/discussion/98059\">https://www.kaggle.com/c/aptos2019-blindness-detection/discussion/98059</a></p>\n\n<p>I have two versions. One that is simply resized, and another that is cropped first by finding the mask of the fundus. This cropping will hopefully improve the score, as I have noticed it has helped in my own experiments with the previous competition dataset.</p>\n\n<p>Sorry it was later than I promised. Unfortunately, I also didn't have time to work on training a model for this competition using the previous competition data. But that should be coming out in the next couple of days.</p>",
      "votes": 8,
      "replies": []
    },
    {
      "id": 564049,
      "author_name": "ilovescience",
      "author_url": "",
      "post_date": "2019-06-28T23:38:44.467000",
      "content": "<p>I plan to make preprocessed versions of the previous competition data by tomorrow.</p>",
      "votes": 5,
      "replies": [
        {
          "id": 564264,
          "author_name": "CVxTz",
          "author_url": "",
          "post_date": "2019-06-29T08:00:13.910000",
          "content": "<p>Thank you ! </p>",
          "votes": 0,
          "replies": []
        },
        {
          "id": 564265,
          "author_name": "Prashant Kikani",
          "author_url": "",
          "post_date": "2019-06-29T08:00:33.710000",
          "content": "<p><a href=\"/tanlikesmath\">@tanlikesmath</a> thanks for that. Can you please make it public? \nThis competition data will hit plateau as because of small size. Then we need to use that dataset.</p>",
          "votes": 0,
          "replies": []
        },
        {
          "id": 564392,
          "author_name": "ilovescience",
          "author_url": "",
          "post_date": "2019-06-29T11:04:16.943000",
          "content": "<p>Yes I will i will first make my submission with it and write a kernel on how to use it and then make it public. I have the dataset already created from previous work and is currently private but once I get it to work with this competition I will make it public! 😃</p>",
          "votes": 1,
          "replies": []
        },
        {
          "id": 564716,
          "author_name": "xhlulu",
          "author_url": "",
          "post_date": "2019-06-29T21:16:27.137000",
          "content": "<p><a href=\"/tanlikesmath\">@tanlikesmath</a> thank you for your good work. What size will it be (e.g. 224x224)?</p>",
          "votes": 0,
          "replies": []
        },
        {
          "id": 564740,
          "author_name": "ilovescience",
          "author_url": "",
          "post_date": "2019-06-29T22:27:35.317000",
          "content": "<p><a href=\"/xhlulu\">@xhlulu</a> it is 1024x1024. You can further resize the pictures later if you want. </p>",
          "votes": 2,
          "replies": []
        }
      ]
    },
    {
      "id": 564883,
      "author_name": "leo",
      "author_url": "",
      "post_date": "2019-06-30T05:52:47.657000",
      "content": "<p>I can't use it in my kernel, but I can use it in local.\n<a href=\"https://www.kaggle.com/c/aptos2019-blindness-detection/discussion/97947#latest-564872\">https://www.kaggle.com/c/aptos2019-blindness-detection/discussion/97947#latest-564872</a></p>\n\n<p>In this competition we don't need to train in kernel. So it is better not to load this data in your kernel because it takes a lot of time.</p>",
      "votes": 2,
      "replies": [
        {
          "id": 564890,
          "author_name": "Maxwell",
          "author_url": "",
          "post_date": "2019-06-30T06:08:04.900000",
          "content": "<p>Well worked ! \nThank you !</p>",
          "votes": 1,
          "replies": []
        },
        {
          "id": 564900,
          "author_name": "ilovescience",
          "author_url": "",
          "post_date": "2019-06-30T06:21:27.490000",
          "content": "<p><a href=\"/bluexleoxgreen\">@bluexleoxgreen</a> a fix for this coming soon ;)</p>",
          "votes": 1,
          "replies": []
        },
        {
          "id": 565215,
          "author_name": "Prashant Kikani",
          "author_url": "",
          "post_date": "2019-06-30T15:40:32.913000",
          "content": "<p>&gt; a fix for this coming soon ;)</p>\n\n<p><a href=\"/tanlikesmath\">@tanlikesmath</a> waiting for that.. \nHope we can do all the things here on kernels. Many people don't have GPUs\nOnly labeled training data (no test images) also would be fine</p>",
          "votes": 2,
          "replies": []
        },
        {
          "id": 565278,
          "author_name": "Maxwell",
          "author_url": "",
          "post_date": "2019-06-30T17:37:48.073000",
          "content": "<p><a href=\"/prashantkikani\">@prashantkikani</a> \nI think the labels of test data are also published in the previous competition. <br>\nRetionopathy_solution csv file includes both labels. <br>\nFor more detail, see my above post.  </p>",
          "votes": 1,
          "replies": []
        }
      ]
    },
    {
      "id": 564854,
      "author_name": "Maxwell",
      "author_url": "",
      "post_date": "2019-06-30T04:49:16.263000",
      "content": "<p>Updated: <br>\nAs <a href=\"/bluexleoxgreen\">@bluexleoxgreen</a> adviced me, these files are just divided files. <br>\nNow I became to extract files. <br>\nThanks <a href=\"/bluexleoxgreen\">@bluexleoxgreen</a> !</p>\n\n<p><strong>I deleted my post in external dataset thread not to confuse.</strong></p>\n\n<hr>\n\n<p><a href=\"/xhlulu\">@xhlulu</a> \nThank you for positing this issue. I have also faced this problem(I could not unzip).  </p>\n\n<p>Just to be sure, I have reported this problem in external dataset thread.\n<a href=\"https://www.kaggle.com/c/aptos2019-blindness-detection/discussion/97605#latest-564847\">https://www.kaggle.com/c/aptos2019-blindness-detection/discussion/97605#latest-564847</a></p>",
      "votes": 2,
      "replies": []
    },
    {
      "id": 566171,
      "author_name": "EdwardPie",
      "author_url": "",
      "post_date": "2019-07-01T21:34:33.643000",
      "content": "<p>You'll need 7zip or Keka (if you're on a Mac) to be able to extract all the volumes in the old dataset. Due to size limitations, the data is broken into parts/volumes named train.zip.001 - train.zip.005.</p>",
      "votes": 0,
      "replies": []
    },
    {
      "id": 564874,
      "author_name": "Xuan Cao",
      "author_url": "",
      "post_date": "2019-06-30T05:32:19.743000",
      "content": "<p>You need 7-Zip to unzip those files. </p>",
      "votes": 0,
      "replies": [
        {
          "id": 564886,
          "author_name": "",
          "author_url": "",
          "post_date": "2019-06-30T05:55:38.500000",
          "content": "",
          "votes": 0,
          "replies": []
        }
      ]
    },
    {
      "id": 564735,
      "author_name": "Tim Yee",
      "author_url": "",
      "post_date": "2019-06-29T22:09:28.593000",
      "content": "<p>What do you mean by unusable? Also, are you doing an comparison of images from the two diff datasets? If so, how are you doing it?</p>",
      "votes": 0,
      "replies": [
        {
          "id": 565473,
          "author_name": "xhlulu",
          "author_url": "",
          "post_date": "2019-07-01T02:00:24.427000",
          "content": "<p>I'm not sure what you mean by comparison of images. They are unusable because when you open a kernel with the previous competition's dataset, you are unable to even locate the files using <code>os.listdir</code>.</p>",
          "votes": 0,
          "replies": []
        }
      ]
    },
    {
      "id": 565125,
      "author_name": "",
      "author_url": "",
      "post_date": "2019-06-30T13:23:38.963000",
      "content": "",
      "votes": 0,
      "replies": []
    }
  ],
  "raw_markdown_by_id": {
    "565134": "I uploaded 896x896 resized [dataset](https://www.kaggle.com/donkeys/retinopathy-train-2015) with combined training data and the labels from the previous one.\n\nI tried the 1024x1024 but it was about 23GB so a bit large for the 20GB size limit. Also I made a [kernel](https://www.kaggle.com/donkeys/looking-in-the-eyes-of-past-and-present) with a few visualizations of it. Maybe a different image format or some other processing could keep it a bit larger but it lets me play with it. ",
    "565498": "I made the dataset available here:\nhttps://www.kaggle.com/c/aptos2019-blindness-detection/discussion/98059\n\nI have two versions. One that is simply resized, and another that is cropped first by finding the mask of the fundus. This cropping will hopefully improve the score, as I have noticed it has helped in my own experiments with the previous competition dataset.\n\nSorry it was later than I promised. Unfortunately, I also didn't have time to work on training a model for this competition using the previous competition data. But that should be coming out in the next couple of days.",
    "563939": "I've taken a look at this previous competition: https://www.kaggle.com/c/diabetic-retinopathy-detection/data\n\nApart from the 1000 sample images, It seems that most of the data (~20,000 images) is unusable. Although the zip file was broken down into multiple parts, I don't seem to be able to recombine them when I load them inside a kernel.\n\nHas anyone been able to use it?",
    "564049": "I plan to make preprocessed versions of the previous competition data by tomorrow.",
    "564883": "I can't use it in my kernel, but I can use it in local.\nhttps://www.kaggle.com/c/aptos2019-blindness-detection/discussion/97947#latest-564872\n\nIn this competition we don't need to train in kernel. So it is better not to load this data in your kernel because it takes a lot of time.",
    "564854": "Updated:  \nAs @bluexleoxgreen adviced me, these files are just divided files.  \nNow I became to extract files.  \nThanks @bluexleoxgreen !\n  \n**I deleted my post in external dataset thread not to confuse.**\n\n---\n\n@xhlulu \nThank you for positing this issue. I have also faced this problem(I could not unzip).  \n\nJust to be sure, I have reported this problem in external dataset thread.\nhttps://www.kaggle.com/c/aptos2019-blindness-detection/discussion/97605#latest-564847",
    "566171": "You'll need 7zip or Keka (if you're on a Mac) to be able to extract all the volumes in the old dataset. Due to size limitations, the data is broken into parts/volumes named train.zip.001 - train.zip.005.",
    "564874": "You need 7-Zip to unzip those files. ",
    "564735": "What do you mean by unusable? Also, are you doing an comparison of images from the two diff datasets? If so, how are you doing it?",
    "565125": ""
  }
}