{
  "id": 469645,
  "title": "Relating the Train.CSV Data set to data from the Train_Spectrogram parguet files?",
  "url": "/competitions/hms-harmful-brain-activity-classification/discussion/469645",
  "author_name": "Michael Wayne Treasure",
  "post_date": "2024-01-21T13:28:03.532000",
  "votes": 1,
  "comment_count": 0,
  "views": 0,
  "content": "<p>From the Train_Spectrogram parquet files, there seems to be 4,279,506 records, of which 102,722 have null values. These loaded parquet files contain data readings at different frequencies for LL, RP, RL and RP along with a \"time\" labeled column.  </p>\n<p><strong>What I am trying to figure out is - how do I tie back this set of parquet data to the train.csv data?</strong></p>\n<p>In the first reference image below, I show an example of 25 records from the Train_Spectrum parquet files. In the second reference image I show the schema for the train.csv file.</p>\n<p><img src=\"https://www.googleapis.com/download/storage/v1/b/kaggle-forum-message-attachments/o/inbox%2F16083849%2F11513c533726e8434c7b2e32ca624b4c%2Fparquet_train_spectrogram_example.PNG?generation=1705843437531953&amp;alt=media\"></p>\n<p><img src=\"https://www.googleapis.com/download/storage/v1/b/kaggle-forum-message-attachments/o/inbox%2F16083849%2F84e111eb70424381570209d8e00603a8%2Fschema_train_csv.PNG?generation=1705843615002841&amp;alt=media\"></p>",
  "messages": [
    {
      "id": 2612571,
      "postDate": "2024-01-21T13:28:03.533Z",
      "content": "<p>From the Train_Spectrogram parquet files, there seems to be 4,279,506 records, of which 102,722 have null values. These loaded parquet files contain data readings at different frequencies for LL, RP, RL and RP along with a \"time\" labeled column.  </p>\n<p><strong>What I am trying to figure out is - how do I tie back this set of parquet data to the train.csv data?</strong></p>\n<p>In the first reference image below, I show an example of 25 records from the Train_Spectrum parquet files. In the second reference image I show the schema for the train.csv file.</p>\n<p><img src=\"https://www.googleapis.com/download/storage/v1/b/kaggle-forum-message-attachments/o/inbox%2F16083849%2F11513c533726e8434c7b2e32ca624b4c%2Fparquet_train_spectrogram_example.PNG?generation=1705843437531953&amp;alt=media\"></p>\n<p><img src=\"https://www.googleapis.com/download/storage/v1/b/kaggle-forum-message-attachments/o/inbox%2F16083849%2F84e111eb70424381570209d8e00603a8%2Fschema_train_csv.PNG?generation=1705843615002841&amp;alt=media\"></p>",
      "rawMarkdown": "From the Train_Spectrogram parquet files, there seems to be 4,279,506 records, of which 102,722 have null values. These loaded parquet files contain data readings at different frequencies for LL, RP, RL and RP along with a \"time\" labeled column.  \n\n**What I am trying to figure out is - how do I tie back this set of parquet data to the train.csv data?**\n\nIn the first reference image below, I show an example of 25 records from the Train_Spectrum parquet files. In the second reference image I show the schema for the train.csv file.\n\n![](https://www.googleapis.com/download/storage/v1/b/kaggle-forum-message-attachments/o/inbox%2F16083849%2F11513c533726e8434c7b2e32ca624b4c%2Fparquet_train_spectrogram_example.PNG?generation=1705843437531953&alt=media)\n\n![](https://www.googleapis.com/download/storage/v1/b/kaggle-forum-message-attachments/o/inbox%2F16083849%2F84e111eb70424381570209d8e00603a8%2Fschema_train_csv.PNG?generation=1705843615002841&alt=media)",
      "votes": 1
    }
  ],
  "comments": [],
  "raw_markdown_by_id": {
    "2612571": "From the Train_Spectrogram parquet files, there seems to be 4,279,506 records, of which 102,722 have null values. These loaded parquet files contain data readings at different frequencies for LL, RP, RL and RP along with a \"time\" labeled column.  \n\n**What I am trying to figure out is - how do I tie back this set of parquet data to the train.csv data?**\n\nIn the first reference image below, I show an example of 25 records from the Train_Spectrum parquet files. In the second reference image I show the schema for the train.csv file.\n\n![](https://www.googleapis.com/download/storage/v1/b/kaggle-forum-message-attachments/o/inbox%2F16083849%2F11513c533726e8434c7b2e32ca624b4c%2Fparquet_train_spectrogram_example.PNG?generation=1705843437531953&alt=media)\n\n![](https://www.googleapis.com/download/storage/v1/b/kaggle-forum-message-attachments/o/inbox%2F16083849%2F84e111eb70424381570209d8e00603a8%2Fschema_train_csv.PNG?generation=1705843615002841&alt=media)"
  }
}