{
  "id": 246424,
  "title": "It seems some data of \"***_derived.csv\" are missed, is it correct?",
  "url": "/competitions/google-smartphone-decimeter-challenge/discussion/246424",
  "author_name": "",
  "post_date": "2021-06-15T09:42:37.143403200Z",
  "votes": 3,
  "comment_count": 3,
  "views": 0,
  "content": "<p>Hi! I am a beginner of Kaggle so if there is something strange its sorry.<br>\nNow, I am trying to load data from ***_derived.csv as mentioned in <a href=\"https://www.kaggle.com/c/google-smartphone-decimeter-challenge/discussion/246226\" target=\"_blank\">https://www.kaggle.com/c/google-smartphone-decimeter-challenge/discussion/246226</a></p>\n<p>I found some data, which shown in baseline_locations_train.csv , are missed in ***_derived.csv.<br>\nThe example of the missed data are here.(Too many of missed data so they are some example)</p>\n<table>\n<thead>\n<tr>\n<th>collectionName</th>\n<th>phoneName</th>\n<th>millisSinceGpsEpoch</th>\n</tr>\n</thead>\n<tbody>\n<tr>\n<td>2020-05-14-US-MTV-1</td>\n<td>Pixel4</td>\n<td>1273529463442</td>\n</tr>\n<tr>\n<td>2020-05-14-US-MTV-1</td>\n<td>Pixel4XLModded</td>\n<td>1273529466449</td>\n</tr>\n<tr>\n<td>2020-05-14-US-MTV-1</td>\n<td>Pixel4XLModded</td>\n<td>1273529467449</td>\n</tr>\n<tr>\n<td>2020-05-14-US-MTV-1</td>\n<td>Pixel4XLModded</td>\n<td>1273529468449</td>\n</tr>\n<tr>\n<td>2020-05-14-US-MTV-1</td>\n<td>Pixel4XLModded</td>\n<td>1273529469449</td>\n</tr>\n<tr>\n<td>2020-05-14-US-MTV-1</td>\n<td>Pixel4XLModded</td>\n<td>1273529470449</td>\n</tr>\n</tbody>\n</table>\n<p>2020-05-14-US-MTV-1 Pixel4 1273529463442 is the first row of the baseline_locations_train.csv and 2020-05-14-US-MTV-1/Pixel4/Pixel4_derived.csv so it can easily check it.</p>\n<p>Is it correct for this competition? If there is my misunderstanding, sorry.</p>",
  "messages": [
    {
      "id": "1350135",
      "postDate": "06/15/2021 09:42:37",
      "content": "<p>Hi! I am a beginner of Kaggle so if there is something strange its sorry.<br>\nNow, I am trying to load data from ***_derived.csv as mentioned in <a href=\"https://www.kaggle.com/c/google-smartphone-decimeter-challenge/discussion/246226\" target=\"_blank\">https://www.kaggle.com/c/google-smartphone-decimeter-challenge/discussion/246226</a></p>\n<p>I found some data, which shown in baseline_locations_train.csv , are missed in ***_derived.csv.<br>\nThe example of the missed data are here.(Too many of missed data so they are some example)</p>\n<table>\n<thead>\n<tr>\n<th>collectionName</th>\n<th>phoneName</th>\n<th>millisSinceGpsEpoch</th>\n</tr>\n</thead>\n<tbody>\n<tr>\n<td>2020-05-14-US-MTV-1</td>\n<td>Pixel4</td>\n<td>1273529463442</td>\n</tr>\n<tr>\n<td>2020-05-14-US-MTV-1</td>\n<td>Pixel4XLModded</td>\n<td>1273529466449</td>\n</tr>\n<tr>\n<td>2020-05-14-US-MTV-1</td>\n<td>Pixel4XLModded</td>\n<td>1273529467449</td>\n</tr>\n<tr>\n<td>2020-05-14-US-MTV-1</td>\n<td>Pixel4XLModded</td>\n<td>1273529468449</td>\n</tr>\n<tr>\n<td>2020-05-14-US-MTV-1</td>\n<td>Pixel4XLModded</td>\n<td>1273529469449</td>\n</tr>\n<tr>\n<td>2020-05-14-US-MTV-1</td>\n<td>Pixel4XLModded</td>\n<td>1273529470449</td>\n</tr>\n</tbody>\n</table>\n<p>2020-05-14-US-MTV-1 Pixel4 1273529463442 is the first row of the baseline_locations_train.csv and 2020-05-14-US-MTV-1/Pixel4/Pixel4_derived.csv so it can easily check it.</p>\n<p>Is it correct for this competition? If there is my misunderstanding, sorry.</p>",
      "rawMarkdown": "Hi! I am a beginner of Kaggle so if there is something strange its sorry.\nNow, I am trying to load data from ***_derived.csv as mentioned in https://www.kaggle.com/c/google-smartphone-decimeter-challenge/discussion/246226\n\nI found some data, which shown in baseline_locations_train.csv , are missed in ***_derived.csv.\nThe example of the missed data are here.(Too many of missed data so they are some example)\n\n| collectionName | phoneName | millisSinceGpsEpoch  |\n| --- | --- | --- |\n|  2020-05-14-US-MTV-1 | Pixel4 | 1273529463442 |\n| 2020-05-14-US-MTV-1 | Pixel4XLModded | 1273529466449 |\n| 2020-05-14-US-MTV-1 | Pixel4XLModded | 1273529467449 |\n| 2020-05-14-US-MTV-1 | Pixel4XLModded | 1273529468449 |\n| 2020-05-14-US-MTV-1 | Pixel4XLModded | 1273529469449 |\n| 2020-05-14-US-MTV-1 | Pixel4XLModded | 1273529470449 |\n\n2020-05-14-US-MTV-1 Pixel4 1273529463442 is the first row of the baseline_locations_train.csv and 2020-05-14-US-MTV-1/Pixel4/Pixel4_derived.csv so it can easily check it.\n\nIs it correct for this competition? If there is my misunderstanding, sorry.",
      "votes": null
    },
    {
      "id": "1353308",
      "postDate": "06/17/2021 03:37:44",
      "content": "<p>This is a reference information.<br>\n<a href=\"https://www.kaggle.com/gymf123/tips-notes-from-the-competition-hosts\" target=\"_blank\">https://www.kaggle.com/gymf123/tips-notes-from-the-competition-hosts</a></p>\n<blockquote>\n  <p>Tip 1: millisSinceGpsEpoch in _derived.csv refers to the timestamp of the next epoch not the current epoch.<br>\n  Important Note: this may affect your final results if you are only using _derived.csv in your positioning algorithm.</p>\n</blockquote>",
      "rawMarkdown": "This is a reference information.\nhttps://www.kaggle.com/gymf123/tips-notes-from-the-competition-hosts\n\n> Tip 1: millisSinceGpsEpoch in _derived.csv refers to the timestamp of the next epoch not the current epoch.\nImportant Note: this may affect your final results if you are only using _derived.csv in your positioning algorithm.",
      "votes": null
    },
    {
      "id": "1359109",
      "postDate": "06/21/2021 04:53:55",
      "content": "<p>Hi I think this is Ok, the data is not missing per se, it was not recorded (to my understanding), you can check the notebook I mentioned in your previous post, there the author used merge_asof to overcome this problem.</p>",
      "rawMarkdown": "Hi I think this is Ok, the data is not missing per se, it was not recorded (to my understanding), you can check the notebook I mentioned in your previous post, there the author used merge_asof to overcome this problem.",
      "votes": null
    },
    {
      "id": "1360245",
      "postDate": "06/22/2021 00:10:04",
      "content": "<p>Thank you! I tried merge_asof and some data, I thought missed, are found. But the problem seems not solved. baseline_locations_train.csv has 131342 lines but the data I created, applied merge_asof, has 130339.</p>\n<p>--&gt; sry. something miss understand, I applied other functinos, not merge_asof XD and it solved. </p>\n<p>I can't much understand about \"millisSinceGpsEpoch\"…<br>\n--&gt; still I don't understand about it… merge_asof seems tackling this problem but… what is \"millisSinceGpsEpoch\" haha</p>",
      "rawMarkdown": "Thank you! I tried merge_asof and some data, I thought missed, are found. But the problem seems not solved. baseline_locations_train.csv has 131342 lines but the data I created, applied merge_asof, has 130339.\n\n--> sry. something miss understand, I applied other functinos, not merge_asof XD and it solved. \n\nI can't much understand about \"millisSinceGpsEpoch\"...\n--> still I don't understand about it... merge_asof seems tackling this problem but... what is \"millisSinceGpsEpoch\" haha",
      "votes": null
    }
  ],
  "comments": [
    {
      "id": 1353308,
      "author_name": "asobod11138",
      "author_url": "",
      "post_date": "06/17/2021 03:37:44",
      "content": "<p>This is a reference information.<br>\n<a href=\"https://www.kaggle.com/gymf123/tips-notes-from-the-competition-hosts\" target=\"_blank\">https://www.kaggle.com/gymf123/tips-notes-from-the-competition-hosts</a></p>\n<blockquote>\n  <p>Tip 1: millisSinceGpsEpoch in _derived.csv refers to the timestamp of the next epoch not the current epoch.<br>\n  Important Note: this may affect your final results if you are only using _derived.csv in your positioning algorithm.</p>\n</blockquote>",
      "votes": null,
      "replies": []
    },
    {
      "id": 1359109,
      "author_name": "avivlevi815",
      "author_url": "",
      "post_date": "06/21/2021 04:53:55",
      "content": "<p>Hi I think this is Ok, the data is not missing per se, it was not recorded (to my understanding), you can check the notebook I mentioned in your previous post, there the author used merge_asof to overcome this problem.</p>",
      "votes": null,
      "replies": [
        {
          "id": 1360245,
          "author_name": "asobod11138",
          "author_url": "",
          "post_date": "06/22/2021 00:10:04",
          "content": "<p>Thank you! I tried merge_asof and some data, I thought missed, are found. But the problem seems not solved. baseline_locations_train.csv has 131342 lines but the data I created, applied merge_asof, has 130339.</p>\n<p>--&gt; sry. something miss understand, I applied other functinos, not merge_asof XD and it solved. </p>\n<p>I can't much understand about \"millisSinceGpsEpoch\"…<br>\n--&gt; still I don't understand about it… merge_asof seems tackling this problem but… what is \"millisSinceGpsEpoch\" haha</p>",
          "votes": null,
          "replies": []
        }
      ]
    }
  ],
  "raw_markdown_by_id": {
    "1350135": "Hi! I am a beginner of Kaggle so if there is something strange its sorry.\nNow, I am trying to load data from ***_derived.csv as mentioned in https://www.kaggle.com/c/google-smartphone-decimeter-challenge/discussion/246226\n\nI found some data, which shown in baseline_locations_train.csv , are missed in ***_derived.csv.\nThe example of the missed data are here.(Too many of missed data so they are some example)\n\n| collectionName | phoneName | millisSinceGpsEpoch  |\n| --- | --- | --- |\n|  2020-05-14-US-MTV-1 | Pixel4 | 1273529463442 |\n| 2020-05-14-US-MTV-1 | Pixel4XLModded | 1273529466449 |\n| 2020-05-14-US-MTV-1 | Pixel4XLModded | 1273529467449 |\n| 2020-05-14-US-MTV-1 | Pixel4XLModded | 1273529468449 |\n| 2020-05-14-US-MTV-1 | Pixel4XLModded | 1273529469449 |\n| 2020-05-14-US-MTV-1 | Pixel4XLModded | 1273529470449 |\n\n2020-05-14-US-MTV-1 Pixel4 1273529463442 is the first row of the baseline_locations_train.csv and 2020-05-14-US-MTV-1/Pixel4/Pixel4_derived.csv so it can easily check it.\n\nIs it correct for this competition? If there is my misunderstanding, sorry.",
    "1353308": "This is a reference information.\nhttps://www.kaggle.com/gymf123/tips-notes-from-the-competition-hosts\n\n> Tip 1: millisSinceGpsEpoch in _derived.csv refers to the timestamp of the next epoch not the current epoch.\nImportant Note: this may affect your final results if you are only using _derived.csv in your positioning algorithm.",
    "1359109": "Hi I think this is Ok, the data is not missing per se, it was not recorded (to my understanding), you can check the notebook I mentioned in your previous post, there the author used merge_asof to overcome this problem.",
    "1360245": "Thank you! I tried merge_asof and some data, I thought missed, are found. But the problem seems not solved. baseline_locations_train.csv has 131342 lines but the data I created, applied merge_asof, has 130339.\n\n--> sry. something miss understand, I applied other functinos, not merge_asof XD and it solved. \n\nI can't much understand about \"millisSinceGpsEpoch\"...\n--> still I don't understand about it... merge_asof seems tackling this problem but... what is \"millisSinceGpsEpoch\" haha"
  },
  "source": "meta"
}