{
  "id": 614588,
  "title": "Help: download ~85 GB Kaggle data (MemoryError)",
  "url": "/competitions/physionet-ecg-image-digitization/discussion/614588",
  "author_name": "",
  "post_date": "2025-11-04T22:45:42.023100200Z",
  "votes": null,
  "comment_count": 3,
  "views": 0,
  "content": "<p>Command:\n  <code>kaggle competitions download -c physionet-ecg-image-digitization -p .\\ECG</code></p>\n<p>Error: <code>MemoryError</code> (requests/iter_content)\nRules accepted, CLI up to date, disk OK.</p>\n<p>Best way for huge files? per-file + real resume ? Mirror ? Any working script to share?</p>\n<p>Thanks </p>",
  "messages": [
    {
      "id": "3311395",
      "postDate": "11/04/2025 22:45:42",
      "content": "<p>Command:\n  <code>kaggle competitions download -c physionet-ecg-image-digitization -p .\\ECG</code></p>\n<p>Error: <code>MemoryError</code> (requests/iter_content)\nRules accepted, CLI up to date, disk OK.</p>\n<p>Best way for huge files? per-file + real resume ? Mirror ? Any working script to share?</p>\n<p>Thanks </p>",
      "rawMarkdown": "Command:\n  `kaggle competitions download -c physionet-ecg-image-digitization -p .\\ECG`\n\nError: `MemoryError` (requests/iter_content)\nRules accepted, CLI up to date, disk OK.\n\nBest way for huge files? per-file + real resume ? Mirror ? Any working script to share?\n\nThanks",
      "votes": null
    },
    {
      "id": "3311439",
      "postDate": "11/05/2025 03:02:06",
      "content": "<p>Download from browser was the only thing that worked. Got rate limited 429 with individual files and this <code>kaggle competitions download -c physionet-ecg-image-digitization</code> also didn't work. </p>",
      "rawMarkdown": "Download from browser was the only thing that worked. Got rate limited 429 with individual files and this `kaggle competitions download -c physionet-ecg-image-digitization` also didn't work.",
      "votes": null
    },
    {
      "id": "3311767",
      "postDate": "11/05/2025 18:20:22",
      "content": "<p>Thanks—you were right. I kept getting 429 rate limits with my custom script, and both individual-file downloads and kaggle competitions download -c physionet-ecg-image-digitization failed. You are right , the only thing that worked was using the browser with resume right now.</p>\n<p>python .\\download.py\n[Info] Downloading root file: train.csv\n[Info] Downloading root file: test.csv\n[Info] 100 train IDs.\nDownloading files:  52%|████████████████████████████████████▍                                 | 104/200 [21:13&lt;2:52:27, 107.79s/it][Warn] train/4182272849/4182272849.csv: failed after 8 retries (429 Client Error: Too Many Requests for url: <a href=\"https://www.kaggle.com/api/v1/competitions/data/download/physionet-ecg-image-digitization/train%2F4182272849%2F4182272849.csv\" target=\"_blank\">https://www.kaggle.com/api/v1/competitions/data/download/physionet-ecg-image-digitization/train%2F4182272849%2F4182272849.csv</a>)\nDownloading files:  56%|███████████████████████████████████████▌                              | 113/200 [49:34&lt;4:50:32, 200.38s/it][Warn] train/2897718844/2897718844-0003.png: failed after 8 retries (429 Client Error: Too Many Requests for url: <a href=\"https://www.kaggle.com/api/v1/competitions/data/download/physionet-ecg-image-digitization/train%2F2897718844%2F2897718844-0003.png\" target=\"_blank\">https://www.kaggle.com/api/v1/competitions/data/download/physionet-ecg-image-digitization/train%2F2897718844%2F2897718844-0003.png</a>)\nDownloading files:  58%|████████████████████████████████████████▌                             | 116/200 [55:35&lt;2:51:40, 122.63s/it]</p>",
      "rawMarkdown": "Thanks—you were right. I kept getting 429 rate limits with my custom script, and both individual-file downloads and kaggle competitions download -c physionet-ecg-image-digitization failed. You are right , the only thing that worked was using the browser with resume right now.\n\npython .\\download.py\n[Info] Downloading root file: train.csv\n[Info] Downloading root file: test.csv\n[Info] 100 train IDs.\nDownloading files:  52%|████████████████████████████████████▍                                 | 104/200 [21:13<2:52:27, 107.79s/it][Warn] train/4182272849/4182272849.csv: failed after 8 retries (429 Client Error: Too Many Requests for url: https://www.kaggle.com/api/v1/competitions/data/download/physionet-ecg-image-digitization/train%2F4182272849%2F4182272849.csv)\nDownloading files:  56%|███████████████████████████████████████▌                              | 113/200 [49:34<4:50:32, 200.38s/it][Warn] train/2897718844/2897718844-0003.png: failed after 8 retries (429 Client Error: Too Many Requests for url: https://www.kaggle.com/api/v1/competitions/data/download/physionet-ecg-image-digitization/train%2F2897718844%2F2897718844-0003.png)\nDownloading files:  58%|████████████████████████████████████████▌                             | 116/200 [55:35<2:51:40, 122.63s/it]",
      "votes": null
    },
    {
      "id": "3312721",
      "postDate": "11/07/2025 19:18:24",
      "content": "<p>Thanks for reporting. I'll file an internal bug report.</p>",
      "rawMarkdown": "Thanks for reporting. I'll file an internal bug report.",
      "votes": null
    }
  ],
  "comments": [
    {
      "id": 3311439,
      "author_name": "rashmibanthia",
      "author_url": "",
      "post_date": "11/05/2025 03:02:06",
      "content": "<p>Download from browser was the only thing that worked. Got rate limited 429 with individual files and this <code>kaggle competitions download -c physionet-ecg-image-digitization</code> also didn't work. </p>",
      "votes": null,
      "replies": [
        {
          "id": 3311767,
          "author_name": "tonylica",
          "author_url": "",
          "post_date": "11/05/2025 18:20:22",
          "content": "<p>Thanks—you were right. I kept getting 429 rate limits with my custom script, and both individual-file downloads and kaggle competitions download -c physionet-ecg-image-digitization failed. You are right , the only thing that worked was using the browser with resume right now.</p>\n<p>python .\\download.py\n[Info] Downloading root file: train.csv\n[Info] Downloading root file: test.csv\n[Info] 100 train IDs.\nDownloading files:  52%|████████████████████████████████████▍                                 | 104/200 [21:13&lt;2:52:27, 107.79s/it][Warn] train/4182272849/4182272849.csv: failed after 8 retries (429 Client Error: Too Many Requests for url: <a href=\"https://www.kaggle.com/api/v1/competitions/data/download/physionet-ecg-image-digitization/train%2F4182272849%2F4182272849.csv\" target=\"_blank\">https://www.kaggle.com/api/v1/competitions/data/download/physionet-ecg-image-digitization/train%2F4182272849%2F4182272849.csv</a>)\nDownloading files:  56%|███████████████████████████████████████▌                              | 113/200 [49:34&lt;4:50:32, 200.38s/it][Warn] train/2897718844/2897718844-0003.png: failed after 8 retries (429 Client Error: Too Many Requests for url: <a href=\"https://www.kaggle.com/api/v1/competitions/data/download/physionet-ecg-image-digitization/train%2F2897718844%2F2897718844-0003.png\" target=\"_blank\">https://www.kaggle.com/api/v1/competitions/data/download/physionet-ecg-image-digitization/train%2F2897718844%2F2897718844-0003.png</a>)\nDownloading files:  58%|████████████████████████████████████████▌                             | 116/200 [55:35&lt;2:51:40, 122.63s/it]</p>",
          "votes": null,
          "replies": []
        }
      ]
    },
    {
      "id": 3312721,
      "author_name": "sohier",
      "author_url": "",
      "post_date": "11/07/2025 19:18:24",
      "content": "<p>Thanks for reporting. I'll file an internal bug report.</p>",
      "votes": null,
      "replies": []
    }
  ],
  "raw_markdown_by_id": {
    "3311395": "Command:\n  `kaggle competitions download -c physionet-ecg-image-digitization -p .\\ECG`\n\nError: `MemoryError` (requests/iter_content)\nRules accepted, CLI up to date, disk OK.\n\nBest way for huge files? per-file + real resume ? Mirror ? Any working script to share?\n\nThanks",
    "3311439": "Download from browser was the only thing that worked. Got rate limited 429 with individual files and this `kaggle competitions download -c physionet-ecg-image-digitization` also didn't work.",
    "3311767": "Thanks—you were right. I kept getting 429 rate limits with my custom script, and both individual-file downloads and kaggle competitions download -c physionet-ecg-image-digitization failed. You are right , the only thing that worked was using the browser with resume right now.\n\npython .\\download.py\n[Info] Downloading root file: train.csv\n[Info] Downloading root file: test.csv\n[Info] 100 train IDs.\nDownloading files:  52%|████████████████████████████████████▍                                 | 104/200 [21:13<2:52:27, 107.79s/it][Warn] train/4182272849/4182272849.csv: failed after 8 retries (429 Client Error: Too Many Requests for url: https://www.kaggle.com/api/v1/competitions/data/download/physionet-ecg-image-digitization/train%2F4182272849%2F4182272849.csv)\nDownloading files:  56%|███████████████████████████████████████▌                              | 113/200 [49:34<4:50:32, 200.38s/it][Warn] train/2897718844/2897718844-0003.png: failed after 8 retries (429 Client Error: Too Many Requests for url: https://www.kaggle.com/api/v1/competitions/data/download/physionet-ecg-image-digitization/train%2F2897718844%2F2897718844-0003.png)\nDownloading files:  58%|████████████████████████████████████████▌                             | 116/200 [55:35<2:51:40, 122.63s/it]",
    "3312721": "Thanks for reporting. I'll file an internal bug report."
  },
  "source": "meta"
}