{
  "id": 122623,
  "title": "What is a parquet file?",
  "url": "/competitions/bengaliai-cv19/discussion/122623",
  "author_name": "",
  "post_date": "2019-12-21T14:18:41.659850700Z",
  "votes": 6,
  "comment_count": 6,
  "views": 0,
  "content": "<p>Is it a latest image file format?</p>",
  "messages": [
    {
      "id": "700140",
      "postDate": "12/21/2019 14:18:41",
      "content": "<p>Is it a latest image file format?</p>",
      "rawMarkdown": "Is it a latest image file format?",
      "votes": null
    },
    {
      "id": "700327",
      "postDate": "12/21/2019 19:47:11",
      "content": "<p><a href=\"/shrutimechlearn\">@shrutimechlearn</a> : Nope. Its not a image file format. Parquet format is open source format for Hadoop. You can read about Parquet format here. <a href=\"https://acadgild.com/blog/parquet-file-format-hadoop\">https://acadgild.com/blog/parquet-file-format-hadoop</a></p>",
      "rawMarkdown": "shrutimechlearn : Nope. Its not a image file format. Parquet format is open source format for Hadoop. You can read about Parquet format here. https://acadgild.com/blog/parquet-file-format-hadoop",
      "votes": null
    },
    {
      "id": "700339",
      "postDate": "12/21/2019 19:56:00",
      "content": "<p>It's a really magical file format. It is super small size, it is super fast to read and write. way better than csv or other data file formats</p>",
      "rawMarkdown": "It's a really magical file format. It is super small size, it is super fast to read and write. way better than csv or other data file formats",
      "votes": null
    },
    {
      "id": "700533",
      "postDate": "12/22/2019 06:55:04",
      "content": "<p>How do I use it efficiently when I have access data row by row?</p>",
      "rawMarkdown": "How do I use it efficiently when I have access data row by row?",
      "votes": null
    },
    {
      "id": "700717",
      "postDate": "12/22/2019 13:49:58",
      "content": "<p>It's just efficient when you are reading/writing. Once you have it loaded into RAM via <code>pd.read_parquet</code> it works just like any other DataFrame.</p>",
      "rawMarkdown": "It's just efficient when you are reading/writing. Once you have it loaded into RAM via `pd.read_parquet` it works just like any other DataFrame.",
      "votes": null
    },
    {
      "id": "700830",
      "postDate": "12/22/2019 17:10:08",
      "content": "<p>Parquet is just like a csv file but more compressed. It can shrink a csv from gigabytes to a few hundred megabytes. Commonly used in cloud storage.</p>",
      "rawMarkdown": "Parquet is just like a csv file but more compressed. It can shrink a csv from gigabytes to a few hundred megabytes. Commonly used in cloud storage.",
      "votes": null
    },
    {
      "id": "1201198",
      "postDate": "02/15/2021 08:12:23",
      "content": "<p>How can I create an image dataset in parquet file format?</p>",
      "rawMarkdown": "How can I create an image dataset in parquet file format?",
      "votes": null
    }
  ],
  "comments": [
    {
      "id": 1201198,
      "author_name": "vietviet",
      "author_url": "",
      "post_date": "02/15/2021 08:12:23",
      "content": "<p>How can I create an image dataset in parquet file format?</p>",
      "votes": null,
      "replies": []
    },
    {
      "id": 700327,
      "author_name": "manojprabhaakr",
      "author_url": "",
      "post_date": "12/21/2019 19:47:11",
      "content": "<p><a href=\"/shrutimechlearn\">@shrutimechlearn</a> : Nope. Its not a image file format. Parquet format is open source format for Hadoop. You can read about Parquet format here. <a href=\"https://acadgild.com/blog/parquet-file-format-hadoop\">https://acadgild.com/blog/parquet-file-format-hadoop</a></p>",
      "votes": null,
      "replies": []
    },
    {
      "id": 700339,
      "author_name": "returnofsputnik",
      "author_url": "",
      "post_date": "12/21/2019 19:56:00",
      "content": "<p>It's a really magical file format. It is super small size, it is super fast to read and write. way better than csv or other data file formats</p>",
      "votes": null,
      "replies": [
        {
          "id": 700533,
          "author_name": "ibraheemmoosa",
          "author_url": "",
          "post_date": "12/22/2019 06:55:04",
          "content": "<p>How do I use it efficiently when I have access data row by row?</p>",
          "votes": null,
          "replies": []
        },
        {
          "id": 700717,
          "author_name": "returnofsputnik",
          "author_url": "",
          "post_date": "12/22/2019 13:49:58",
          "content": "<p>It's just efficient when you are reading/writing. Once you have it loaded into RAM via <code>pd.read_parquet</code> it works just like any other DataFrame.</p>",
          "votes": null,
          "replies": []
        }
      ]
    },
    {
      "id": 700830,
      "author_name": "notjohnsmith",
      "author_url": "",
      "post_date": "12/22/2019 17:10:08",
      "content": "<p>Parquet is just like a csv file but more compressed. It can shrink a csv from gigabytes to a few hundred megabytes. Commonly used in cloud storage.</p>",
      "votes": null,
      "replies": []
    }
  ],
  "raw_markdown_by_id": {
    "700140": "Is it a latest image file format?",
    "700327": "shrutimechlearn : Nope. Its not a image file format. Parquet format is open source format for Hadoop. You can read about Parquet format here. https://acadgild.com/blog/parquet-file-format-hadoop",
    "700339": "It's a really magical file format. It is super small size, it is super fast to read and write. way better than csv or other data file formats",
    "700533": "How do I use it efficiently when I have access data row by row?",
    "700717": "It's just efficient when you are reading/writing. Once you have it loaded into RAM via `pd.read_parquet` it works just like any other DataFrame.",
    "700830": "Parquet is just like a csv file but more compressed. It can shrink a csv from gigabytes to a few hundred megabytes. Commonly used in cloud storage.",
    "1201198": "How can I create an image dataset in parquet file format?"
  },
  "source": "meta"
}