{
  "id": 28600,
  "title": "How to download these satellite images to an EC2?",
  "url": "/competitions/dstl-satellite-imagery-feature-detection/discussion/28600",
  "author_name": "ZhuangfangYi",
  "post_date": "2017-02-08T20:43:33.644000",
  "votes": 0,
  "comment_count": 4,
  "views": 0,
  "content": "<p>Hi guys, \nI have a quick question (hope it's a quick one):\nHow can I download these image data to AWS EC2? I currently have an EC2 instance with GPU, and try to download these image data to the EC2. But when I use \"wget ... \"data, it would not unzip. I guess because the data is securely locked by kaggle (don't know the right words for this)? and I also tried to use kaggle CLI to configure and download the data, but my kaggle account was signed up and logged in with my Gmail, and basically I don't remember the password correctly. What is the best way to download the data to EC2 in this case? \nRegards, \nnana</p>",
  "messages": [
    {
      "id": 160727,
      "postDate": "2017-02-09T07:53:08.843Z",
      "content": "<ol>\n<li><p>In chrome Use Ctrl + Shift + C (or Cmd + Shift + C on Mac) to open the DevTools</p></li>\n<li><p>Go to the data site then click on download and pause</p></li>\n<li><p>copy the cUrl in network tab of developer tools, by right clicking on the link-&gt;copy-&gt;copy as cUrl</p></li>\n<li><p>$ wget cUrl (remove --compressed at the end) -O file_name</p></li>\n</ol>",
      "rawMarkdown": "1. In chrome Use Ctrl + Shift + C (or Cmd + Shift + C on Mac) to open the DevTools\n\n2. Go to the data site then click on download and pause\n\n3. copy the cUrl in network tab of developer tools, by right clicking on the link->copy->copy as cUrl\n\n4. $ wget cUrl (remove --compressed at the end) -O file_name\n",
      "votes": 1
    },
    {
      "id": 161675,
      "postDate": "2017-02-15T05:41:27.433Z",
      "content": "<p>I also used my google account when signing up. To solve the problem of not having a password, I first clicked forgotten password, and then followed the instructions to create a password for my account. Then I used the kaggle CLI to download the data.</p>",
      "rawMarkdown": "I also used my google account when signing up. To solve the problem of not having a password, I first clicked forgotten password, and then followed the instructions to create a password for my account. Then I used the kaggle CLI to download the data."
    },
    {
      "id": 161398,
      "postDate": "2017-02-13T15:30:29.637Z",
      "content": "<p>Alternately, you could use a chrome extension such as <a href=\"https://chrome.google.com/webstore/detail/cookietxt-export/lopabhfecdfhgogdbojmaicoicjekelh\">cookie.txt export</a>.</p>\n\n<ol>\n<li>Log into kaggle, go to the data download page</li>\n<li>Click on the extension, and save the entire text as cookie.txt</li>\n<li><p>On your instance, use the --load-cookies.txt option for wget:</p>\n\n<p>wget --load-cookies path/to/saved/cookie.txt <a href=\"http://url.zip/\">http://url.zip</a></p></li>\n</ol>",
      "rawMarkdown": "Alternately, you could use a chrome extension such as [cookie.txt export][1].\n\n1. Log into kaggle, go to the data download page\n2. Click on the extension, and save the entire text as cookie.txt\n3. On your instance, use the --load-cookies.txt option for wget:\n\n    wget --load-cookies path/to/saved/cookie.txt http://url.zip\n\n\n  [1]: https://chrome.google.com/webstore/detail/cookietxt-export/lopabhfecdfhgogdbojmaicoicjekelh"
    },
    {
      "id": 160623,
      "postDate": "2017-02-08T20:43:33.643Z",
      "content": "<p>Hi guys, \nI have a quick question (hope it's a quick one):\nHow can I download these image data to AWS EC2? I currently have an EC2 instance with GPU, and try to download these image data to the EC2. But when I use \"wget ... \"data, it would not unzip. I guess because the data is securely locked by kaggle (don't know the right words for this)? and I also tried to use kaggle CLI to configure and download the data, but my kaggle account was signed up and logged in with my Gmail, and basically I don't remember the password correctly. What is the best way to download the data to EC2 in this case? \nRegards, \nnana</p>",
      "rawMarkdown": "Hi guys, \nI have a quick question (hope it's a quick one):\nHow can I download these image data to AWS EC2? I currently have an EC2 instance with GPU, and try to download these image data to the EC2. But when I use \"wget ... \"data, it would not unzip. I guess because the data is securely locked by kaggle (don't know the right words for this)? and I also tried to use kaggle CLI to configure and download the data, but my kaggle account was signed up and logged in with my Gmail, and basically I don't remember the password correctly. What is the best way to download the data to EC2 in this case? \nRegards, \nnana"
    },
    {
      "id": 163660,
      "postDate": "2017-02-25T00:00:55.403Z",
      "rawMarkdown": "",
      "isDeleted": true
    }
  ],
  "comments": [
    {
      "id": 160727,
      "author_name": "__MR__",
      "author_url": "",
      "post_date": "2017-02-09T07:53:08.843000",
      "content": "<ol>\n<li><p>In chrome Use Ctrl + Shift + C (or Cmd + Shift + C on Mac) to open the DevTools</p></li>\n<li><p>Go to the data site then click on download and pause</p></li>\n<li><p>copy the cUrl in network tab of developer tools, by right clicking on the link-&gt;copy-&gt;copy as cUrl</p></li>\n<li><p>$ wget cUrl (remove --compressed at the end) -O file_name</p></li>\n</ol>",
      "votes": 1,
      "replies": []
    },
    {
      "id": 161675,
      "author_name": "Hugo",
      "author_url": "",
      "post_date": "2017-02-15T05:41:27.433000",
      "content": "<p>I also used my google account when signing up. To solve the problem of not having a password, I first clicked forgotten password, and then followed the instructions to create a password for my account. Then I used the kaggle CLI to download the data.</p>",
      "votes": 0,
      "replies": []
    },
    {
      "id": 161398,
      "author_name": "grantbey",
      "author_url": "",
      "post_date": "2017-02-13T15:30:29.637000",
      "content": "<p>Alternately, you could use a chrome extension such as <a href=\"https://chrome.google.com/webstore/detail/cookietxt-export/lopabhfecdfhgogdbojmaicoicjekelh\">cookie.txt export</a>.</p>\n\n<ol>\n<li>Log into kaggle, go to the data download page</li>\n<li>Click on the extension, and save the entire text as cookie.txt</li>\n<li><p>On your instance, use the --load-cookies.txt option for wget:</p>\n\n<p>wget --load-cookies path/to/saved/cookie.txt <a href=\"http://url.zip/\">http://url.zip</a></p></li>\n</ol>",
      "votes": 0,
      "replies": []
    },
    {
      "id": 163660,
      "author_name": "",
      "author_url": "",
      "post_date": "2017-02-25T00:00:55.403000",
      "content": "",
      "votes": 0,
      "replies": []
    }
  ],
  "raw_markdown_by_id": {
    "160727": "1. In chrome Use Ctrl + Shift + C (or Cmd + Shift + C on Mac) to open the DevTools\n\n2. Go to the data site then click on download and pause\n\n3. copy the cUrl in network tab of developer tools, by right clicking on the link->copy->copy as cUrl\n\n4. $ wget cUrl (remove --compressed at the end) -O file_name\n",
    "161675": "I also used my google account when signing up. To solve the problem of not having a password, I first clicked forgotten password, and then followed the instructions to create a password for my account. Then I used the kaggle CLI to download the data.",
    "161398": "Alternately, you could use a chrome extension such as [cookie.txt export][1].\n\n1. Log into kaggle, go to the data download page\n2. Click on the extension, and save the entire text as cookie.txt\n3. On your instance, use the --load-cookies.txt option for wget:\n\n    wget --load-cookies path/to/saved/cookie.txt http://url.zip\n\n\n  [1]: https://chrome.google.com/webstore/detail/cookietxt-export/lopabhfecdfhgogdbojmaicoicjekelh",
    "160623": "Hi guys, \nI have a quick question (hope it's a quick one):\nHow can I download these image data to AWS EC2? I currently have an EC2 instance with GPU, and try to download these image data to the EC2. But when I use \"wget ... \"data, it would not unzip. I guess because the data is securely locked by kaggle (don't know the right words for this)? and I also tried to use kaggle CLI to configure and download the data, but my kaggle account was signed up and logged in with my Gmail, and basically I don't remember the password correctly. What is the best way to download the data to EC2 in this case? \nRegards, \nnana",
    "163660": ""
  }
}