{
  "id": 309037,
  "title": "Google Cloud AI Notebook, can not download dataset",
  "url": "/competitions/happy-whale-and-dolphin/discussion/309037",
  "author_name": "",
  "post_date": "2022-02-21T14:24:27.852816500Z",
  "votes": 4,
  "comment_count": 5,
  "views": 0,
  "content": "<p>I have created a GC AI notebook linked to my Kaggle kernel, following the link 'Upgrade to GC AI Notebooks'. After multiple attempts, I still have not been able to download the dataset using the first cell of the Jupyter notebook in GC AI. </p>\n<p>Normally, I get several warnings about Jupyter message output being temporary stopped (the download progress bar), before the kernel shuts down. By then, the progress is about 35 GB downloaded.</p>\n<p>Upon load failure, the notebook presents a link to the dataset. I could download this to my PC, then upload to the GC notebook. However, with the zip file being 57 GB, I`d rather not. Anyone else having this problem, or used some other method to get this dataset to a GC notebook?</p>",
  "messages": [
    {
      "id": "1699921",
      "postDate": "02/21/2022 14:24:27",
      "content": "<p>I have created a GC AI notebook linked to my Kaggle kernel, following the link 'Upgrade to GC AI Notebooks'. After multiple attempts, I still have not been able to download the dataset using the first cell of the Jupyter notebook in GC AI. </p>\n<p>Normally, I get several warnings about Jupyter message output being temporary stopped (the download progress bar), before the kernel shuts down. By then, the progress is about 35 GB downloaded.</p>\n<p>Upon load failure, the notebook presents a link to the dataset. I could download this to my PC, then upload to the GC notebook. However, with the zip file being 57 GB, I`d rather not. Anyone else having this problem, or used some other method to get this dataset to a GC notebook?</p>",
      "rawMarkdown": "I have created a GC AI notebook linked to my Kaggle kernel, following the link 'Upgrade to GC AI Notebooks'. After multiple attempts, I still have not been able to download the dataset using the first cell of the Jupyter notebook in GC AI. \n\nNormally, I get several warnings about Jupyter message output being temporary stopped (the download progress bar), before the kernel shuts down. By then, the progress is about 35 GB downloaded.\n\nUpon load failure, the notebook presents a link to the dataset. I could download this to my PC, then upload to the GC notebook. However, with the zip file being 57 GB, I`d rather not. Anyone else having this problem, or used some other method to get this dataset to a GC notebook?",
      "votes": null
    },
    {
      "id": "1699931",
      "postDate": "02/21/2022 14:30:30",
      "content": "<p><a href=\"https://www.kaggle.com/c/tensorflow-great-barrier-reef/discussion/300601\" target=\"_blank\">https://www.kaggle.com/c/tensorflow-great-barrier-reef/discussion/300601</a></p>\n<p>please upvote if it helps!</p>",
      "rawMarkdown": "https://www.kaggle.com/c/tensorflow-great-barrier-reef/discussion/300601\n\n\nplease upvote if it helps!",
      "votes": null
    },
    {
      "id": "1700826",
      "postDate": "02/22/2022 09:30:11",
      "content": "<p>After starting a clean Google AI notebook, the \"built-in\" download procedure seemed to work, until approximately 32 GB was completed (to a 100 GB disk):</p>\n<blockquote>\n  <p>Unexpected error while saving file: imported/NOTEBOOK-2e194c5e-aa4f-4452-85f6-6de807677f4d.ipynb HTTP 500: Internal Server Error (Unexpected error while saving file: importedNOTEBOOK-2e194c5e-aa4f-4452-85f6-6de807677f4d.ipynb database or disk is full)</p>\n</blockquote>\n<p>I will try the procedure from <a href=\"https://www.kaggle.com/dragonzhang\" target=\"_blank\">@dragonzhang</a>, still curious about other experiences regarding this dataset and GC AI notebooks.</p>",
      "rawMarkdown": "After starting a clean Google AI notebook, the \"built-in\" download procedure seemed to work, until approximately 32 GB was completed (to a 100 GB disk):\n\n> Unexpected error while saving file: imported/NOTEBOOK-2e194c5e-aa4f-4452-85f6-6de807677f4d.ipynb HTTP 500: Internal Server Error (Unexpected error while saving file: importedNOTEBOOK-2e194c5e-aa4f-4452-85f6-6de807677f4d.ipynb database or disk is full)\n\nI will try the procedure from @dragonzhang, still curious about other experiences regarding this dataset and GC AI notebooks.",
      "votes": null
    },
    {
      "id": "1700984",
      "postDate": "02/22/2022 12:31:05",
      "content": "<p>About to give up using Google Cloud notebooks on this data. I have tried</p>\n<ol>\n<li>original dataset fetch method</li>\n<li>direct file upload from PC</li>\n<li>wget and curl, using the Jupyter terminal</li>\n</ol>\n<p>All methods fail after approximately 50% progress.</p>",
      "rawMarkdown": "About to give up using Google Cloud notebooks on this data. I have tried\n\n1. original dataset fetch method\n2. direct file upload from PC\n3. wget and curl, using the Jupyter terminal\n\nAll methods fail after approximately 50% progress.",
      "votes": null
    },
    {
      "id": "1700985",
      "postDate": "02/22/2022 12:32:24",
      "content": "<p>It seems your method is not compatible with the GC Jupyter notebooks, I got several errors about incompatible pip libraries. Thanks anyway :) </p>",
      "rawMarkdown": "It seems your method is not compatible with the GC Jupyter notebooks, I got several errors about incompatible pip libraries. Thanks anyway :)",
      "votes": null
    },
    {
      "id": "1703702",
      "postDate": "02/24/2022 18:45:15",
      "content": "<p>It is about 4G/hr to copy.  But I did not copy  too big dataset before.<br>\nperhaps you need upgrade your Colab account.</p>",
      "rawMarkdown": "It is about 4G/hr to copy.  But I did not copy  too big dataset before.\nperhaps you need upgrade your Colab account.",
      "votes": null
    }
  ],
  "comments": [
    {
      "id": 1699931,
      "author_name": "dragonzhang",
      "author_url": "",
      "post_date": "02/21/2022 14:30:30",
      "content": "<p><a href=\"https://www.kaggle.com/c/tensorflow-great-barrier-reef/discussion/300601\" target=\"_blank\">https://www.kaggle.com/c/tensorflow-great-barrier-reef/discussion/300601</a></p>\n<p>please upvote if it helps!</p>",
      "votes": null,
      "replies": [
        {
          "id": 1700985,
          "author_name": "frontsideflip",
          "author_url": "",
          "post_date": "02/22/2022 12:32:24",
          "content": "<p>It seems your method is not compatible with the GC Jupyter notebooks, I got several errors about incompatible pip libraries. Thanks anyway :) </p>",
          "votes": null,
          "replies": []
        }
      ]
    },
    {
      "id": 1700826,
      "author_name": "frontsideflip",
      "author_url": "",
      "post_date": "02/22/2022 09:30:11",
      "content": "<p>After starting a clean Google AI notebook, the \"built-in\" download procedure seemed to work, until approximately 32 GB was completed (to a 100 GB disk):</p>\n<blockquote>\n  <p>Unexpected error while saving file: imported/NOTEBOOK-2e194c5e-aa4f-4452-85f6-6de807677f4d.ipynb HTTP 500: Internal Server Error (Unexpected error while saving file: importedNOTEBOOK-2e194c5e-aa4f-4452-85f6-6de807677f4d.ipynb database or disk is full)</p>\n</blockquote>\n<p>I will try the procedure from <a href=\"https://www.kaggle.com/dragonzhang\" target=\"_blank\">@dragonzhang</a>, still curious about other experiences regarding this dataset and GC AI notebooks.</p>",
      "votes": null,
      "replies": [
        {
          "id": 1703702,
          "author_name": "dragonzhang",
          "author_url": "",
          "post_date": "02/24/2022 18:45:15",
          "content": "<p>It is about 4G/hr to copy.  But I did not copy  too big dataset before.<br>\nperhaps you need upgrade your Colab account.</p>",
          "votes": null,
          "replies": []
        }
      ]
    },
    {
      "id": 1700984,
      "author_name": "frontsideflip",
      "author_url": "",
      "post_date": "02/22/2022 12:31:05",
      "content": "<p>About to give up using Google Cloud notebooks on this data. I have tried</p>\n<ol>\n<li>original dataset fetch method</li>\n<li>direct file upload from PC</li>\n<li>wget and curl, using the Jupyter terminal</li>\n</ol>\n<p>All methods fail after approximately 50% progress.</p>",
      "votes": null,
      "replies": []
    }
  ],
  "raw_markdown_by_id": {
    "1699921": "I have created a GC AI notebook linked to my Kaggle kernel, following the link 'Upgrade to GC AI Notebooks'. After multiple attempts, I still have not been able to download the dataset using the first cell of the Jupyter notebook in GC AI. \n\nNormally, I get several warnings about Jupyter message output being temporary stopped (the download progress bar), before the kernel shuts down. By then, the progress is about 35 GB downloaded.\n\nUpon load failure, the notebook presents a link to the dataset. I could download this to my PC, then upload to the GC notebook. However, with the zip file being 57 GB, I`d rather not. Anyone else having this problem, or used some other method to get this dataset to a GC notebook?",
    "1699931": "https://www.kaggle.com/c/tensorflow-great-barrier-reef/discussion/300601\n\n\nplease upvote if it helps!",
    "1700826": "After starting a clean Google AI notebook, the \"built-in\" download procedure seemed to work, until approximately 32 GB was completed (to a 100 GB disk):\n\n> Unexpected error while saving file: imported/NOTEBOOK-2e194c5e-aa4f-4452-85f6-6de807677f4d.ipynb HTTP 500: Internal Server Error (Unexpected error while saving file: importedNOTEBOOK-2e194c5e-aa4f-4452-85f6-6de807677f4d.ipynb database or disk is full)\n\nI will try the procedure from @dragonzhang, still curious about other experiences regarding this dataset and GC AI notebooks.",
    "1700984": "About to give up using Google Cloud notebooks on this data. I have tried\n\n1. original dataset fetch method\n2. direct file upload from PC\n3. wget and curl, using the Jupyter terminal\n\nAll methods fail after approximately 50% progress.",
    "1700985": "It seems your method is not compatible with the GC Jupyter notebooks, I got several errors about incompatible pip libraries. Thanks anyway :)",
    "1703702": "It is about 4G/hr to copy.  But I did not copy  too big dataset before.\nperhaps you need upgrade your Colab account."
  },
  "source": "meta"
}