{
  "id": 244696,
  "title": "Follow-up on training-evaluation notebook not working",
  "url": "/competitions/birdclef-2021/discussion/244696",
  "author_name": "",
  "post_date": "2021-06-07T20:32:48.135396700Z",
  "votes": 2,
  "comment_count": 4,
  "views": 0,
  "content": "<p>Okay, now that the competition is over, I implore you to help me understand, why this notebook will not submit for evaluation, i.e. why it cannot find the submission.csv file. It it a mix of Stefan Kahls two notebooks, the one for training and the one for submitting to the competition. </p>\n<p>If you have access to the versions, I will tell you that version 4 succeded in showing that a submission.csv file was created, only I had called it submission1.csv and submission2.csv, and thus it could not be used for anything. BUT IT COULD TELL THAT IT WAS THERE. However, all other versions fail. </p>\n<p>I asked Kaggle for the difference, between e.g. version 6 and version 4, and this is what it shows:<br>\n<img src=\"https://i.imgur.com/HJ82SvS.png\" alt=\"\"></p>\n<p>The ONLY difference was the name. And the fact that I used a maximum of 5 files in version 4 and 1500 in all others.</p>\n<p><img src=\"https://i.imgur.com/15sQ99k.png\" alt=\"\"></p>\n<p><img src=\"https://i.imgur.com/a4Q7r9h.png\" alt=\"\"></p>\n<p>I know that I'm saving the submissions.csv file two times, to different paths, that resolve to the same place. But as far as I can tell, it makes no difference, whether I do both of them or only one of them.</p>\n<p>If you make a commit and run all version of the latest version and try to submit it to the competition, what then? Does the submissions file appear for you? If not, can you spot why it wouldn't!?</p>\n<p><a href=\"https://www.kaggle.com/tobiasbonnesen/birdclef2021-model-training-thesis-baseline\" target=\"_blank\">https://www.kaggle.com/tobiasbonnesen/birdclef2021-model-training-thesis-baseline</a></p>\n<p>I'm training my real model on my own computer, so I'm not super dependent on this for further scoring. It just annoys me so much, that it doesn't work, when as far as I can tell, it should.</p>\n<p>And no, it's not because of training time. This takes no time to train. You can even change the max audio files settings to allow for super-duper fast training time. 🤓</p>\n<p>Thank you 🙏🙏🙏🙏🙏🙏🙏</p>",
  "messages": [
    {
      "id": "1340378",
      "postDate": "06/07/2021 20:32:48",
      "content": "<p>Okay, now that the competition is over, I implore you to help me understand, why this notebook will not submit for evaluation, i.e. why it cannot find the submission.csv file. It it a mix of Stefan Kahls two notebooks, the one for training and the one for submitting to the competition. </p>\n<p>If you have access to the versions, I will tell you that version 4 succeded in showing that a submission.csv file was created, only I had called it submission1.csv and submission2.csv, and thus it could not be used for anything. BUT IT COULD TELL THAT IT WAS THERE. However, all other versions fail. </p>\n<p>I asked Kaggle for the difference, between e.g. version 6 and version 4, and this is what it shows:<br>\n<img src=\"https://i.imgur.com/HJ82SvS.png\" alt=\"\"></p>\n<p>The ONLY difference was the name. And the fact that I used a maximum of 5 files in version 4 and 1500 in all others.</p>\n<p><img src=\"https://i.imgur.com/15sQ99k.png\" alt=\"\"></p>\n<p><img src=\"https://i.imgur.com/a4Q7r9h.png\" alt=\"\"></p>\n<p>I know that I'm saving the submissions.csv file two times, to different paths, that resolve to the same place. But as far as I can tell, it makes no difference, whether I do both of them or only one of them.</p>\n<p>If you make a commit and run all version of the latest version and try to submit it to the competition, what then? Does the submissions file appear for you? If not, can you spot why it wouldn't!?</p>\n<p><a href=\"https://www.kaggle.com/tobiasbonnesen/birdclef2021-model-training-thesis-baseline\" target=\"_blank\">https://www.kaggle.com/tobiasbonnesen/birdclef2021-model-training-thesis-baseline</a></p>\n<p>I'm training my real model on my own computer, so I'm not super dependent on this for further scoring. It just annoys me so much, that it doesn't work, when as far as I can tell, it should.</p>\n<p>And no, it's not because of training time. This takes no time to train. You can even change the max audio files settings to allow for super-duper fast training time. 🤓</p>\n<p>Thank you 🙏🙏🙏🙏🙏🙏🙏</p>",
      "rawMarkdown": "Okay, now that the competition is over, I implore you to help me understand, why this notebook will not submit for evaluation, i.e. why it cannot find the submission.csv file. It it a mix of Stefan Kahls two notebooks, the one for training and the one for submitting to the competition. \n\nIf you have access to the versions, I will tell you that version 4 succeded in showing that a submission.csv file was created, only I had called it submission1.csv and submission2.csv, and thus it could not be used for anything. BUT IT COULD TELL THAT IT WAS THERE. However, all other versions fail. \n\nI asked Kaggle for the difference, between e.g. version 6 and version 4, and this is what it shows:\n![](https://i.imgur.com/HJ82SvS.png)\n\nThe ONLY difference was the name. And the fact that I used a maximum of 5 files in version 4 and 1500 in all others.\n\n![](https://i.imgur.com/15sQ99k.png)\n\n![](https://i.imgur.com/a4Q7r9h.png)\n\nI know that I'm saving the submissions.csv file two times, to different paths, that resolve to the same place. But as far as I can tell, it makes no difference, whether I do both of them or only one of them.\n\nIf you make a commit and run all version of the latest version and try to submit it to the competition, what then? Does the submissions file appear for you? If not, can you spot why it wouldn't!?\n\nhttps://www.kaggle.com/tobiasbonnesen/birdclef2021-model-training-thesis-baseline\n\nI'm training my real model on my own computer, so I'm not super dependent on this for further scoring. It just annoys me so much, that it doesn't work, when as far as I can tell, it should.\n\nAnd no, it's not because of training time. This takes no time to train. You can even change the max audio files settings to allow for super-duper fast training time. 🤓\n\nThank you 🙏🙏🙏🙏🙏🙏🙏",
      "votes": null
    },
    {
      "id": "1341132",
      "postDate": "06/08/2021 12:55:51",
      "content": "<p>Delete all unneccesary output except submission.csv. That might help.</p>",
      "rawMarkdown": "Delete all unneccesary output except submission.csv. That might help.",
      "votes": null
    },
    {
      "id": "1341144",
      "postDate": "06/08/2021 13:02:04",
      "content": "<p>Thanks for the reply :-) </p>\n<p>I tried that, it didn't do any diffeference. Interestingly enough, for version 4, where it actually detects some files as outputted, it is the case that several submissions files were outputted. </p>\n<p>If you go to my latest version where I'm only outputting one csv file<br>\n<a href=\"https://www.kaggle.com/tobiasbonnesen/birdclef2021-model-training-thesis-baseline\" target=\"_blank\">https://www.kaggle.com/tobiasbonnesen/birdclef2021-model-training-thesis-baseline</a><br>\nAnd try to submit it it for evaluation, you'll see it doesn't work. </p>\n<p>I'm pretty sure this is a kaggle bug. &lt;.&lt; Though what triggers it I don't know. But maybe I'm wrong. </p>",
      "rawMarkdown": "Thanks for the reply :-) \n\nI tried that, it didn't do any diffeference. Interestingly enough, for version 4, where it actually detects some files as outputted, it is the case that several submissions files were outputted. \n\nIf you go to my latest version where I'm only outputting one csv file\nhttps://www.kaggle.com/tobiasbonnesen/birdclef2021-model-training-thesis-baseline\nAnd try to submit it it for evaluation, you'll see it doesn't work. \n\nI'm pretty sure this is a kaggle bug. <.< Though what triggers it I don't know. But maybe I'm wrong.",
      "votes": null
    },
    {
      "id": "1341177",
      "postDate": "06/08/2021 13:33:58",
      "content": "<p>Output Size  20.4 MB (on version 9/9)</p>\n<p>That is not \"one csv file\". Delete all your spectrograms that you generated during training.</p>",
      "rawMarkdown": "Output Size  20.4 MB (on version 9/9)\n\nThat is not \"one csv file\". Delete all your spectrograms that you generated during training.",
      "votes": null
    },
    {
      "id": "1341361",
      "postDate": "06/08/2021 15:54:05",
      "content": "<p>There is another difference.  In the left code you output submission2.csv, and you don't output it in the right code. Anyway, the submission file must be named submission.csv.</p>\n<p>Also why are you using a path for saving your sub?  Here is what I used:</p>\n<pre><code>df[['row_id', 'birds']].to_csv('submission.csv', index=False)\n</code></pre>",
      "rawMarkdown": "There is another difference.  In the left code you output submission2.csv, and you don't output it in the right code. Anyway, the submission file must be named submission.csv.\n\nAlso why are you using a path for saving your sub?  Here is what I used:\n\n```\ndf[['row_id', 'birds']].to_csv('submission.csv', index=False)\n\n```",
      "votes": null
    }
  ],
  "comments": [
    {
      "id": 1341132,
      "author_name": "fffrrt",
      "author_url": "",
      "post_date": "06/08/2021 12:55:51",
      "content": "<p>Delete all unneccesary output except submission.csv. That might help.</p>",
      "votes": null,
      "replies": [
        {
          "id": 1341144,
          "author_name": "tobiasbonnesen",
          "author_url": "",
          "post_date": "06/08/2021 13:02:04",
          "content": "<p>Thanks for the reply :-) </p>\n<p>I tried that, it didn't do any diffeference. Interestingly enough, for version 4, where it actually detects some files as outputted, it is the case that several submissions files were outputted. </p>\n<p>If you go to my latest version where I'm only outputting one csv file<br>\n<a href=\"https://www.kaggle.com/tobiasbonnesen/birdclef2021-model-training-thesis-baseline\" target=\"_blank\">https://www.kaggle.com/tobiasbonnesen/birdclef2021-model-training-thesis-baseline</a><br>\nAnd try to submit it it for evaluation, you'll see it doesn't work. </p>\n<p>I'm pretty sure this is a kaggle bug. &lt;.&lt; Though what triggers it I don't know. But maybe I'm wrong. </p>",
          "votes": null,
          "replies": []
        },
        {
          "id": 1341177,
          "author_name": "fffrrt",
          "author_url": "",
          "post_date": "06/08/2021 13:33:58",
          "content": "<p>Output Size  20.4 MB (on version 9/9)</p>\n<p>That is not \"one csv file\". Delete all your spectrograms that you generated during training.</p>",
          "votes": null,
          "replies": []
        }
      ]
    },
    {
      "id": 1341361,
      "author_name": "cpmpml",
      "author_url": "",
      "post_date": "06/08/2021 15:54:05",
      "content": "<p>There is another difference.  In the left code you output submission2.csv, and you don't output it in the right code. Anyway, the submission file must be named submission.csv.</p>\n<p>Also why are you using a path for saving your sub?  Here is what I used:</p>\n<pre><code>df[['row_id', 'birds']].to_csv('submission.csv', index=False)\n</code></pre>",
      "votes": null,
      "replies": []
    }
  ],
  "raw_markdown_by_id": {
    "1340378": "Okay, now that the competition is over, I implore you to help me understand, why this notebook will not submit for evaluation, i.e. why it cannot find the submission.csv file. It it a mix of Stefan Kahls two notebooks, the one for training and the one for submitting to the competition. \n\nIf you have access to the versions, I will tell you that version 4 succeded in showing that a submission.csv file was created, only I had called it submission1.csv and submission2.csv, and thus it could not be used for anything. BUT IT COULD TELL THAT IT WAS THERE. However, all other versions fail. \n\nI asked Kaggle for the difference, between e.g. version 6 and version 4, and this is what it shows:\n![](https://i.imgur.com/HJ82SvS.png)\n\nThe ONLY difference was the name. And the fact that I used a maximum of 5 files in version 4 and 1500 in all others.\n\n![](https://i.imgur.com/15sQ99k.png)\n\n![](https://i.imgur.com/a4Q7r9h.png)\n\nI know that I'm saving the submissions.csv file two times, to different paths, that resolve to the same place. But as far as I can tell, it makes no difference, whether I do both of them or only one of them.\n\nIf you make a commit and run all version of the latest version and try to submit it to the competition, what then? Does the submissions file appear for you? If not, can you spot why it wouldn't!?\n\nhttps://www.kaggle.com/tobiasbonnesen/birdclef2021-model-training-thesis-baseline\n\nI'm training my real model on my own computer, so I'm not super dependent on this for further scoring. It just annoys me so much, that it doesn't work, when as far as I can tell, it should.\n\nAnd no, it's not because of training time. This takes no time to train. You can even change the max audio files settings to allow for super-duper fast training time. 🤓\n\nThank you 🙏🙏🙏🙏🙏🙏🙏",
    "1341132": "Delete all unneccesary output except submission.csv. That might help.",
    "1341144": "Thanks for the reply :-) \n\nI tried that, it didn't do any diffeference. Interestingly enough, for version 4, where it actually detects some files as outputted, it is the case that several submissions files were outputted. \n\nIf you go to my latest version where I'm only outputting one csv file\nhttps://www.kaggle.com/tobiasbonnesen/birdclef2021-model-training-thesis-baseline\nAnd try to submit it it for evaluation, you'll see it doesn't work. \n\nI'm pretty sure this is a kaggle bug. <.< Though what triggers it I don't know. But maybe I'm wrong.",
    "1341177": "Output Size  20.4 MB (on version 9/9)\n\nThat is not \"one csv file\". Delete all your spectrograms that you generated during training.",
    "1341361": "There is another difference.  In the left code you output submission2.csv, and you don't output it in the right code. Anyway, the submission file must be named submission.csv.\n\nAlso why are you using a path for saving your sub?  Here is what I used:\n\n```\ndf[['row_id', 'birds']].to_csv('submission.csv', index=False)\n\n```"
  },
  "source": "meta"
}