{
  "id": 125242,
  "title": "Original Files",
  "url": "/competitions/deepfake-detection-challenge/discussion/125242",
  "author_name": "",
  "post_date": "2020-01-09T13:27:13.979035500Z",
  "votes": null,
  "comment_count": 5,
  "views": 0,
  "content": "<p>Where do I find the original files before the deepfake was applied? I downloaded \"deepfake-detection-challenge.zip\" and I don't see them in there.</p>",
  "messages": [
    {
      "id": "714476",
      "postDate": "01/09/2020 13:27:13",
      "content": "<p>Where do I find the original files before the deepfake was applied? I downloaded \"deepfake-detection-challenge.zip\" and I don't see them in there.</p>",
      "rawMarkdown": "Where do I find the original files before the deepfake was applied? I downloaded \"deepfake-detection-challenge.zip\" and I don't see them in there.",
      "votes": null
    },
    {
      "id": "714770",
      "postDate": "01/09/2020 18:20:25",
      "content": "<p>I believe the original files are identified in the json metadata files.</p>",
      "rawMarkdown": "I believe the original files are identified in the json metadata files.",
      "votes": null
    },
    {
      "id": "714789",
      "postDate": "01/09/2020 18:45:26",
      "content": "<p>They are identified in the json metadata files.  If you only downloaded the sample train file than many of the originals will not be in that batch of 400.  You need to grab the entire download if you want to find all of the originals.  </p>",
      "rawMarkdown": "They are identified in the json metadata files.  If you only downloaded the sample train file than many of the originals will not be in that batch of 400.  You need to grab the entire download if you want to find all of the originals.",
      "votes": null
    },
    {
      "id": "715253",
      "postDate": "01/10/2020 10:00:03",
      "content": "<p>is there an easy way to download the full 500GB+ dataset via CLI or API? i tried with kaggle command, and it's only downloading the 4GB sample</p>",
      "rawMarkdown": "is there an easy way to download the full 500GB+ dataset via CLI or API? i tried with kaggle command, and it's only downloading the 4GB sample",
      "votes": null
    },
    {
      "id": "715279",
      "postDate": "01/10/2020 10:47:45",
      "content": "<p>Not sure what you mean but if you are looking for a way to download it on a headless machine (like cloud server), use cookies from browser and wget. Just search for wget in the discussions of this competition (on phone so awkward to look it up for you) \nThe links are in the data tab of the competition. </p>",
      "rawMarkdown": "Not sure what you mean but if you are looking for a way to download it on a headless machine (like cloud server), use cookies from browser and wget. Just search for wget in the discussions of this competition (on phone so awkward to look it up for you) \nThe links are in the data tab of the competition.",
      "votes": null
    },
    {
      "id": "715282",
      "postDate": "01/10/2020 10:54:07",
      "content": "<p>thanks, that was it ;)</p>",
      "rawMarkdown": "thanks, that was it ;)",
      "votes": null
    }
  ],
  "comments": [
    {
      "id": 714770,
      "author_name": "petewills",
      "author_url": "",
      "post_date": "01/09/2020 18:20:25",
      "content": "<p>I believe the original files are identified in the json metadata files.</p>",
      "votes": null,
      "replies": []
    },
    {
      "id": 714789,
      "author_name": "pcjimmmy",
      "author_url": "",
      "post_date": "01/09/2020 18:45:26",
      "content": "<p>They are identified in the json metadata files.  If you only downloaded the sample train file than many of the originals will not be in that batch of 400.  You need to grab the entire download if you want to find all of the originals.  </p>",
      "votes": null,
      "replies": []
    },
    {
      "id": 715253,
      "author_name": "antonio1979",
      "author_url": "",
      "post_date": "01/10/2020 10:00:03",
      "content": "<p>is there an easy way to download the full 500GB+ dataset via CLI or API? i tried with kaggle command, and it's only downloading the 4GB sample</p>",
      "votes": null,
      "replies": [
        {
          "id": 715279,
          "author_name": "moshel",
          "author_url": "",
          "post_date": "01/10/2020 10:47:45",
          "content": "<p>Not sure what you mean but if you are looking for a way to download it on a headless machine (like cloud server), use cookies from browser and wget. Just search for wget in the discussions of this competition (on phone so awkward to look it up for you) \nThe links are in the data tab of the competition. </p>",
          "votes": null,
          "replies": []
        },
        {
          "id": 715282,
          "author_name": "antonio1979",
          "author_url": "",
          "post_date": "01/10/2020 10:54:07",
          "content": "<p>thanks, that was it ;)</p>",
          "votes": null,
          "replies": []
        }
      ]
    }
  ],
  "raw_markdown_by_id": {
    "714476": "Where do I find the original files before the deepfake was applied? I downloaded \"deepfake-detection-challenge.zip\" and I don't see them in there.",
    "714770": "I believe the original files are identified in the json metadata files.",
    "714789": "They are identified in the json metadata files.  If you only downloaded the sample train file than many of the originals will not be in that batch of 400.  You need to grab the entire download if you want to find all of the originals.",
    "715253": "is there an easy way to download the full 500GB+ dataset via CLI or API? i tried with kaggle command, and it's only downloading the 4GB sample",
    "715279": "Not sure what you mean but if you are looking for a way to download it on a headless machine (like cloud server), use cookies from browser and wget. Just search for wget in the discussions of this competition (on phone so awkward to look it up for you) \nThe links are in the data tab of the competition.",
    "715282": "thanks, that was it ;)"
  },
  "source": "meta"
}