{
  "id": 327104,
  "title": "The tabular data is soooooo huge that reading is a problem! Sad but exciting.",
  "url": "/competitions/amex-default-prediction/discussion/327104",
  "author_name": "",
  "post_date": "2022-05-25T15:47:29.619577800Z",
  "votes": 6,
  "comment_count": 5,
  "views": 0,
  "content": "<p>Any good suggestions for reading this data? :)</p>",
  "messages": [
    {
      "id": "1801300",
      "postDate": "05/25/2022 15:47:29",
      "content": "<p>Any good suggestions for reading this data? :)</p>",
      "rawMarkdown": "Any good suggestions for reading this data? :)",
      "votes": null
    },
    {
      "id": "1801326",
      "postDate": "05/25/2022 16:03:47",
      "content": "<p>A number of years ago, back in the days when memory was often a limiter, there were some cool techniques to read and process data line-by-line, e.g., <a href=\"https://www.kaggle.com/competitions/avazu-ctr-prediction/discussion/10927\" target=\"_blank\">FTRL</a>. </p>",
      "rawMarkdown": "A number of years ago, back in the days when memory was often a limiter, there were some cool techniques to read and process data line-by-line, e.g., [FTRL](https://www.kaggle.com/competitions/avazu-ctr-prediction/discussion/10927).",
      "votes": null
    },
    {
      "id": "1801328",
      "postDate": "05/25/2022 16:04:06",
      "content": "<p>How do you plan to import this data to any other ecosystem? Is there a ZIP link?</p>",
      "rawMarkdown": "How do you plan to import this data to any other ecosystem? Is there a ZIP link?",
      "votes": null
    },
    {
      "id": "1801334",
      "postDate": "05/25/2022 16:10:47",
      "content": "<p>I recommend <a href=\"https://www.kaggle.com/rohanrao\" target=\"_blank\">@rohanrao</a> 's post <br>\n<a href=\"https://www.kaggle.com/competitions/riiid-test-answer-prediction/discussion/191751\" target=\"_blank\">https://www.kaggle.com/competitions/riiid-test-answer-prediction/discussion/191751</a></p>\n<p>found it very useful when I attempted Riiid Answer Correctness Prediction last year</p>",
      "rawMarkdown": "I recommend @rohanrao 's post \nhttps://www.kaggle.com/competitions/riiid-test-answer-prediction/discussion/191751\n\nfound it very useful when I attempted Riiid Answer Correctness Prediction last year",
      "votes": null
    },
    {
      "id": "1801536",
      "postDate": "05/25/2022 20:57:59",
      "content": "<p>You can explore the data using the first 1 million lines, for example (hoping your conclusions don't get biased) and then when you train and test your model by reading the CSVs in <a href=\"https://stackoverflow.com/questions/25962114/how-do-i-read-a-large-csv-file-with-pandas\" target=\"_blank\">chunks</a></p>",
      "rawMarkdown": "You can explore the data using the first 1 million lines, for example (hoping your conclusions don't get biased) and then when you train and test your model by reading the CSVs in [chunks](https://stackoverflow.com/questions/25962114/how-do-i-read-a-large-csv-file-with-pandas)",
      "votes": null
    },
    {
      "id": "1801660",
      "postDate": "05/26/2022 02:38:36",
      "content": "<p>Thank u guys. These are very helpful to me -v-</p>",
      "rawMarkdown": "Thank u guys. These are very helpful to me -v-",
      "votes": null
    }
  ],
  "comments": [
    {
      "id": 1801326,
      "author_name": "inversion",
      "author_url": "",
      "post_date": "05/25/2022 16:03:47",
      "content": "<p>A number of years ago, back in the days when memory was often a limiter, there were some cool techniques to read and process data line-by-line, e.g., <a href=\"https://www.kaggle.com/competitions/avazu-ctr-prediction/discussion/10927\" target=\"_blank\">FTRL</a>. </p>",
      "votes": null,
      "replies": [
        {
          "id": 1801334,
          "author_name": "kmldas",
          "author_url": "",
          "post_date": "05/25/2022 16:10:47",
          "content": "<p>I recommend <a href=\"https://www.kaggle.com/rohanrao\" target=\"_blank\">@rohanrao</a> 's post <br>\n<a href=\"https://www.kaggle.com/competitions/riiid-test-answer-prediction/discussion/191751\" target=\"_blank\">https://www.kaggle.com/competitions/riiid-test-answer-prediction/discussion/191751</a></p>\n<p>found it very useful when I attempted Riiid Answer Correctness Prediction last year</p>",
          "votes": null,
          "replies": []
        },
        {
          "id": 1801660,
          "author_name": "carlmaxace",
          "author_url": "",
          "post_date": "05/26/2022 02:38:36",
          "content": "<p>Thank u guys. These are very helpful to me -v-</p>",
          "votes": null,
          "replies": []
        }
      ]
    },
    {
      "id": 1801328,
      "author_name": "pratyush",
      "author_url": "",
      "post_date": "05/25/2022 16:04:06",
      "content": "<p>How do you plan to import this data to any other ecosystem? Is there a ZIP link?</p>",
      "votes": null,
      "replies": []
    },
    {
      "id": 1801536,
      "author_name": "bgmello",
      "author_url": "",
      "post_date": "05/25/2022 20:57:59",
      "content": "<p>You can explore the data using the first 1 million lines, for example (hoping your conclusions don't get biased) and then when you train and test your model by reading the CSVs in <a href=\"https://stackoverflow.com/questions/25962114/how-do-i-read-a-large-csv-file-with-pandas\" target=\"_blank\">chunks</a></p>",
      "votes": null,
      "replies": []
    }
  ],
  "raw_markdown_by_id": {
    "1801300": "Any good suggestions for reading this data? :)",
    "1801326": "A number of years ago, back in the days when memory was often a limiter, there were some cool techniques to read and process data line-by-line, e.g., [FTRL](https://www.kaggle.com/competitions/avazu-ctr-prediction/discussion/10927).",
    "1801328": "How do you plan to import this data to any other ecosystem? Is there a ZIP link?",
    "1801334": "I recommend @rohanrao 's post \nhttps://www.kaggle.com/competitions/riiid-test-answer-prediction/discussion/191751\n\nfound it very useful when I attempted Riiid Answer Correctness Prediction last year",
    "1801536": "You can explore the data using the first 1 million lines, for example (hoping your conclusions don't get biased) and then when you train and test your model by reading the CSVs in [chunks](https://stackoverflow.com/questions/25962114/how-do-i-read-a-large-csv-file-with-pandas)",
    "1801660": "Thank u guys. These are very helpful to me -v-"
  },
  "source": "meta"
}