{
  "id": 55432,
  "title": "File Structures/ Data types and #nulls",
  "url": "/competitions/avito-demand-prediction/discussion/55432",
  "author_name": "",
  "post_date": "2018-04-26T13:18:56.348507600Z",
  "votes": 28,
  "comment_count": 5,
  "views": 0,
  "content": "<p>I tried to pack basic information into 1 page. Enjoy! </p>",
  "messages": [
    {
      "id": "319629",
      "postDate": "04/26/2018 13:18:56",
      "content": "<p>I tried to pack basic information into 1 page. Enjoy! </p>",
      "rawMarkdown": "I tried to pack basic information into 1 page. Enjoy!",
      "votes": null
    },
    {
      "id": "319678",
      "postDate": "04/26/2018 15:26:11",
      "content": "<p>This is super cool ! Thanks you </p>",
      "rawMarkdown": "This is super cool ! Thanks you",
      "votes": null
    },
    {
      "id": "324067",
      "postDate": "05/07/2018 04:24:13",
      "content": "<p>I guess that you used <code>wc -l</code> to count how many rows in the <code>train_active.csv</code> and <code>test_active.csv</code> because you couldn't load them. </p>\n\n<p>Number of rows in <code>train_active.csv</code>:  14,129,821</p>\n\n<p>Number of rows in <code>test_active.csv</code>:  12,824,068</p>",
      "rawMarkdown": "I guess that you used `wc -l` to count how many rows in the `train_active.csv` and `test_active.csv` because you couldn't load them. \n\nNumber of rows in `train_active.csv`:  14,129,821\n\nNumber of rows in `test_active.csv`:  12,824,068",
      "votes": null
    },
    {
      "id": "325300",
      "postDate": "05/08/2018 09:35:17",
      "content": "<p>Yes, that's correct. These files are huge, I don't have unique values also for them. Reading the file with chunksize might help. </p>",
      "rawMarkdown": "Yes, that's correct. These files are huge, I don't have unique values also for them. Reading the file with chunksize might help.",
      "votes": null
    },
    {
      "id": "325726",
      "postDate": "05/08/2018 19:43:25",
      "content": "<p>Its Helping, Thanks.</p>",
      "rawMarkdown": "Its Helping, Thanks.",
      "votes": null
    },
    {
      "id": "326965",
      "postDate": "05/10/2018 15:21:31",
      "content": "<p>Nice summary!</p>",
      "rawMarkdown": "Nice summary!",
      "votes": null
    }
  ],
  "comments": [
    {
      "id": 319678,
      "author_name": "arthurllau",
      "author_url": "",
      "post_date": "04/26/2018 15:26:11",
      "content": "<p>This is super cool ! Thanks you </p>",
      "votes": null,
      "replies": []
    },
    {
      "id": 324067,
      "author_name": "hqphat",
      "author_url": "",
      "post_date": "05/07/2018 04:24:13",
      "content": "<p>I guess that you used <code>wc -l</code> to count how many rows in the <code>train_active.csv</code> and <code>test_active.csv</code> because you couldn't load them. </p>\n\n<p>Number of rows in <code>train_active.csv</code>:  14,129,821</p>\n\n<p>Number of rows in <code>test_active.csv</code>:  12,824,068</p>",
      "votes": null,
      "replies": [
        {
          "id": 325300,
          "author_name": "rashmibanthia",
          "author_url": "",
          "post_date": "05/08/2018 09:35:17",
          "content": "<p>Yes, that's correct. These files are huge, I don't have unique values also for them. Reading the file with chunksize might help. </p>",
          "votes": null,
          "replies": []
        }
      ]
    },
    {
      "id": 325726,
      "author_name": "rishabhgarg1023",
      "author_url": "",
      "post_date": "05/08/2018 19:43:25",
      "content": "<p>Its Helping, Thanks.</p>",
      "votes": null,
      "replies": []
    },
    {
      "id": 326965,
      "author_name": "zehaiwang",
      "author_url": "",
      "post_date": "05/10/2018 15:21:31",
      "content": "<p>Nice summary!</p>",
      "votes": null,
      "replies": []
    }
  ],
  "raw_markdown_by_id": {
    "319629": "I tried to pack basic information into 1 page. Enjoy!",
    "319678": "This is super cool ! Thanks you",
    "324067": "I guess that you used `wc -l` to count how many rows in the `train_active.csv` and `test_active.csv` because you couldn't load them. \n\nNumber of rows in `train_active.csv`:  14,129,821\n\nNumber of rows in `test_active.csv`:  12,824,068",
    "325300": "Yes, that's correct. These files are huge, I don't have unique values also for them. Reading the file with chunksize might help.",
    "325726": "Its Helping, Thanks.",
    "326965": "Nice summary!"
  },
  "source": "meta"
}