{
  "id": 287955,
  "title": "Test images does not downloaded",
  "url": "/competitions/wikipedia-image-caption/discussion/287955",
  "author_name": "",
  "post_date": "2021-11-16T05:05:11.016039100Z",
  "votes": 5,
  "comment_count": 6,
  "views": 0,
  "content": "<p>Here, in test.tsv, some test images are not downloaded using image_url. Does anyone face this issue?</p>",
  "messages": [
    {
      "id": "1583823",
      "postDate": "11/16/2021 05:05:11",
      "content": "<p>Here, in test.tsv, some test images are not downloaded using image_url. Does anyone face this issue?</p>",
      "rawMarkdown": "Here, in test.tsv, some test images are not downloaded using image_url. Does anyone face this issue?",
      "votes": null
    },
    {
      "id": "1584074",
      "postDate": "11/16/2021 08:51:57",
      "content": "<p>All test images are included in folder image_data_test/image_pixels. u can download this folder directly</p>",
      "rawMarkdown": "All test images are included in folder image_data_test/image_pixels. u can download this folder directly",
      "votes": null
    },
    {
      "id": "1584355",
      "postDate": "11/16/2021 13:34:55",
      "content": "<p>Those are csv files, and those have <br>\n<em>b64_bytes: base64 encoded bytes of the image file at a 300px resolution</em></p>\n<p>but I need actual resolutions.</p>",
      "rawMarkdown": "Those are csv files, and those have \n*b64_bytes: base64 encoded bytes of the image file at a 300px resolution*\n\nbut I need actual resolutions.",
      "votes": null
    },
    {
      "id": "1585644",
      "postDate": "11/17/2021 12:06:18",
      "content": "<p>Some of them are .svg or .tif etc.<br>\nAre those the images that does not work for you?</p>",
      "rawMarkdown": "Some of them are .svg or .tif etc.\nAre those the images that does not work for you?",
      "votes": null
    },
    {
      "id": "1585853",
      "postDate": "11/17/2021 15:53:08",
      "content": "<p>No, not only those. For an example, if I have found out an image does not parse by urlib, then I try by manually copying the URL and pasting it in the browser. Then I do not see the images. So, I assume some URLs are broken. I am now trying to use base64 as I need to optimize my model also</p>",
      "rawMarkdown": "No, not only those. For an example, if I have found out an image does not parse by urlib, then I try by manually copying the URL and pasting it in the browser. Then I do not see the images. So, I assume some URLs are broken. I am now trying to use base64 as I need to optimize my model also",
      "votes": null
    },
    {
      "id": "1597407",
      "postDate": "11/27/2021 13:34:38",
      "content": "<blockquote>\n  <p>All test images are included in folder image_data_test/image_pixels. u can download this folder directly</p>\n</blockquote>\n<p>Hello,  I find out that there are only 44762 test images in folder image_data_test/image_pixels. But there are 90'000+ image urls in file test.tsv. Dose this mean we lost part of test images?</p>",
      "rawMarkdown": "> All test images are included in folder image_data_test/image_pixels. u can download this folder directly\n\nHello,  I find out that there are only 44762 test images in folder image_data_test/image_pixels. But there are 90'000+ image urls in file test.tsv. Dose this mean we lost part of test images?",
      "votes": null
    },
    {
      "id": "1600179",
      "postDate": "11/30/2021 07:31:58",
      "content": "<p>Some URLs in <code>test.tsv</code> are repeated. There are only 44762 unique URLs, and the corresponding images are all provided.</p>",
      "rawMarkdown": "Some URLs in `test.tsv` are repeated. There are only 44762 unique URLs, and the corresponding images are all provided.",
      "votes": null
    }
  ],
  "comments": [
    {
      "id": 1584074,
      "author_name": "penglu2097",
      "author_url": "",
      "post_date": "11/16/2021 08:51:57",
      "content": "<p>All test images are included in folder image_data_test/image_pixels. u can download this folder directly</p>",
      "votes": null,
      "replies": [
        {
          "id": 1584355,
          "author_name": "faisaltfc",
          "author_url": "",
          "post_date": "11/16/2021 13:34:55",
          "content": "<p>Those are csv files, and those have <br>\n<em>b64_bytes: base64 encoded bytes of the image file at a 300px resolution</em></p>\n<p>but I need actual resolutions.</p>",
          "votes": null,
          "replies": []
        },
        {
          "id": 1597407,
          "author_name": "hfzhong",
          "author_url": "",
          "post_date": "11/27/2021 13:34:38",
          "content": "<blockquote>\n  <p>All test images are included in folder image_data_test/image_pixels. u can download this folder directly</p>\n</blockquote>\n<p>Hello,  I find out that there are only 44762 test images in folder image_data_test/image_pixels. But there are 90'000+ image urls in file test.tsv. Dose this mean we lost part of test images?</p>",
          "votes": null,
          "replies": []
        },
        {
          "id": 1600179,
          "author_name": "penglu2097",
          "author_url": "",
          "post_date": "11/30/2021 07:31:58",
          "content": "<p>Some URLs in <code>test.tsv</code> are repeated. There are only 44762 unique URLs, and the corresponding images are all provided.</p>",
          "votes": null,
          "replies": []
        }
      ]
    },
    {
      "id": 1585644,
      "author_name": "alexanderbader",
      "author_url": "",
      "post_date": "11/17/2021 12:06:18",
      "content": "<p>Some of them are .svg or .tif etc.<br>\nAre those the images that does not work for you?</p>",
      "votes": null,
      "replies": [
        {
          "id": 1585853,
          "author_name": "faisaltfc",
          "author_url": "",
          "post_date": "11/17/2021 15:53:08",
          "content": "<p>No, not only those. For an example, if I have found out an image does not parse by urlib, then I try by manually copying the URL and pasting it in the browser. Then I do not see the images. So, I assume some URLs are broken. I am now trying to use base64 as I need to optimize my model also</p>",
          "votes": null,
          "replies": []
        }
      ]
    }
  ],
  "raw_markdown_by_id": {
    "1583823": "Here, in test.tsv, some test images are not downloaded using image_url. Does anyone face this issue?",
    "1584074": "All test images are included in folder image_data_test/image_pixels. u can download this folder directly",
    "1584355": "Those are csv files, and those have \n*b64_bytes: base64 encoded bytes of the image file at a 300px resolution*\n\nbut I need actual resolutions.",
    "1585644": "Some of them are .svg or .tif etc.\nAre those the images that does not work for you?",
    "1585853": "No, not only those. For an example, if I have found out an image does not parse by urlib, then I try by manually copying the URL and pasting it in the browser. Then I do not see the images. So, I assume some URLs are broken. I am now trying to use base64 as I need to optimize my model also",
    "1597407": "> All test images are included in folder image_data_test/image_pixels. u can download this folder directly\n\nHello,  I find out that there are only 44762 test images in folder image_data_test/image_pixels. But there are 90'000+ image urls in file test.tsv. Dose this mean we lost part of test images?",
    "1600179": "Some URLs in `test.tsv` are repeated. There are only 44762 unique URLs, and the corresponding images are all provided."
  },
  "source": "meta"
}