{
  "id": 209954,
  "title": "Full data not getting download kaggle",
  "url": "/competitions/rfcx-species-audio-detection/discussion/209954",
  "author_name": "",
  "post_date": "2021-01-09T07:35:54.656071Z",
  "votes": 3,
  "comment_count": 3,
  "views": 0,
  "content": "<p>I try to download(in colab) complete dataset from kaggle  which says just <code>kaggle competitions download -c rfcx-species-audio-detection</code> , then in order to get tfrecords separately I write this</p>\n<pre><code>import os, zipfile\n\n\ndir_name = os.getcwd()\next_loc = os.path.join(tf, 'train') \n#all train tfrecords files are ending with 148\nextension = \"148.tfrec.zip\"\n\nfor item in os.listdir(dir_name): # loop through items in dir\n    if item.endswith(extension): # check for \"148.tfrec.zip\" extension\n        file_name = os.path.abspath(item) # get full path of files\n        zip_ref = zipfile.ZipFile(file_name) # create zipfile object\n        zip_ref.extractall(ext_loc) # extract file to dir\n        zip_ref.close() # close file\n        os.remove(file_name) # delete zipped file\n</code></pre>\n<p>After doing this I am getting 20 files but in the actual train dataset 31 files are present ,this means complete data is not getting downloaded. How to download complete data?Same case apply for flac files </p>",
  "messages": [
    {
      "id": "1145502",
      "postDate": "01/09/2021 07:35:54",
      "content": "<p>I try to download(in colab) complete dataset from kaggle  which says just <code>kaggle competitions download -c rfcx-species-audio-detection</code> , then in order to get tfrecords separately I write this</p>\n<pre><code>import os, zipfile\n\n\ndir_name = os.getcwd()\next_loc = os.path.join(tf, 'train') \n#all train tfrecords files are ending with 148\nextension = \"148.tfrec.zip\"\n\nfor item in os.listdir(dir_name): # loop through items in dir\n    if item.endswith(extension): # check for \"148.tfrec.zip\" extension\n        file_name = os.path.abspath(item) # get full path of files\n        zip_ref = zipfile.ZipFile(file_name) # create zipfile object\n        zip_ref.extractall(ext_loc) # extract file to dir\n        zip_ref.close() # close file\n        os.remove(file_name) # delete zipped file\n</code></pre>\n<p>After doing this I am getting 20 files but in the actual train dataset 31 files are present ,this means complete data is not getting downloaded. How to download complete data?Same case apply for flac files </p>",
      "rawMarkdown": "I try to download(in colab) complete dataset from kaggle  which says just `kaggle competitions download -c rfcx-species-audio-detection` , then in order to get tfrecords separately I write this\n\n```\nimport os, zipfile\n\n\ndir_name = os.getcwd()\next_loc = os.path.join(tf, 'train') \n#all train tfrecords files are ending with 148\nextension = \"148.tfrec.zip\"\n\nfor item in os.listdir(dir_name): # loop through items in dir\n    if item.endswith(extension): # check for \"148.tfrec.zip\" extension\n        file_name = os.path.abspath(item) # get full path of files\n        zip_ref = zipfile.ZipFile(file_name) # create zipfile object\n        zip_ref.extractall(ext_loc) # extract file to dir\n        zip_ref.close() # close file\n        os.remove(file_name) # delete zipped file\n```\nAfter doing this I am getting 20 files but in the actual train dataset 31 files are present ,this means complete data is not getting downloaded. How to download complete data?Same case apply for flac files",
      "votes": null
    },
    {
      "id": "1151225",
      "postDate": "01/13/2021 07:42:54",
      "content": "<p>I'm experiencing the same problem. kaggle cli results</p>\n<pre><code>$ kaggle competitions files -c rfcx-species-audio-detection\ntfrecords/train/15-148.tfrec  813MB  2020-11-17 20:12:16\ntfrecords/test/04-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/12-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/17-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/16-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/08-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/19-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/01-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/06-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/11-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/13-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/10-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/18-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/09-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/05-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/15-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/00-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/02-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/14-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/07-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/03-63.tfrec    346MB  2020-11-17 20:12:16\ntest/019db5220.flac             3MB  2020-11-17 20:12:16\n</code></pre>",
      "rawMarkdown": "I'm experiencing the same problem. kaggle cli results\n\n```\n$ kaggle competitions files -c rfcx-species-audio-detection\ntfrecords/train/15-148.tfrec  813MB  2020-11-17 20:12:16\ntfrecords/test/04-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/12-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/17-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/16-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/08-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/19-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/01-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/06-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/11-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/13-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/10-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/18-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/09-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/05-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/15-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/00-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/02-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/14-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/07-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/03-63.tfrec    346MB  2020-11-17 20:12:16\ntest/019db5220.flac             3MB  2020-11-17 20:12:16\n```",
      "votes": null
    },
    {
      "id": "1191812",
      "postDate": "02/08/2021 17:38:20",
      "content": "<p>Did you manage to solve this?</p>",
      "rawMarkdown": "Did you manage to solve this?",
      "votes": null
    },
    {
      "id": "1192570",
      "postDate": "02/09/2021 07:32:05",
      "content": "<p>I had to download it manually.</p>",
      "rawMarkdown": "I had to download it manually.",
      "votes": null
    }
  ],
  "comments": [
    {
      "id": 1151225,
      "author_name": "takadaat",
      "author_url": "",
      "post_date": "01/13/2021 07:42:54",
      "content": "<p>I'm experiencing the same problem. kaggle cli results</p>\n<pre><code>$ kaggle competitions files -c rfcx-species-audio-detection\ntfrecords/train/15-148.tfrec  813MB  2020-11-17 20:12:16\ntfrecords/test/04-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/12-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/17-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/16-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/08-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/19-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/01-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/06-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/11-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/13-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/10-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/18-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/09-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/05-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/15-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/00-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/02-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/14-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/07-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/03-63.tfrec    346MB  2020-11-17 20:12:16\ntest/019db5220.flac             3MB  2020-11-17 20:12:16\n</code></pre>",
      "votes": null,
      "replies": [
        {
          "id": 1191812,
          "author_name": "framoni",
          "author_url": "",
          "post_date": "02/08/2021 17:38:20",
          "content": "<p>Did you manage to solve this?</p>",
          "votes": null,
          "replies": []
        },
        {
          "id": 1192570,
          "author_name": "takadaat",
          "author_url": "",
          "post_date": "02/09/2021 07:32:05",
          "content": "<p>I had to download it manually.</p>",
          "votes": null,
          "replies": []
        }
      ]
    }
  ],
  "raw_markdown_by_id": {
    "1145502": "I try to download(in colab) complete dataset from kaggle  which says just `kaggle competitions download -c rfcx-species-audio-detection` , then in order to get tfrecords separately I write this\n\n```\nimport os, zipfile\n\n\ndir_name = os.getcwd()\next_loc = os.path.join(tf, 'train') \n#all train tfrecords files are ending with 148\nextension = \"148.tfrec.zip\"\n\nfor item in os.listdir(dir_name): # loop through items in dir\n    if item.endswith(extension): # check for \"148.tfrec.zip\" extension\n        file_name = os.path.abspath(item) # get full path of files\n        zip_ref = zipfile.ZipFile(file_name) # create zipfile object\n        zip_ref.extractall(ext_loc) # extract file to dir\n        zip_ref.close() # close file\n        os.remove(file_name) # delete zipped file\n```\nAfter doing this I am getting 20 files but in the actual train dataset 31 files are present ,this means complete data is not getting downloaded. How to download complete data?Same case apply for flac files",
    "1151225": "I'm experiencing the same problem. kaggle cli results\n\n```\n$ kaggle competitions files -c rfcx-species-audio-detection\ntfrecords/train/15-148.tfrec  813MB  2020-11-17 20:12:16\ntfrecords/test/04-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/12-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/17-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/16-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/08-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/19-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/01-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/06-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/11-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/13-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/10-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/18-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/09-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/05-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/15-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/00-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/02-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/14-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/07-63.tfrec    346MB  2020-11-17 20:12:16\ntfrecords/test/03-63.tfrec    346MB  2020-11-17 20:12:16\ntest/019db5220.flac             3MB  2020-11-17 20:12:16\n```",
    "1191812": "Did you manage to solve this?",
    "1192570": "I had to download it manually."
  },
  "source": "meta"
}