{
  "id": 498731,
  "title": "Submission CSV NOT Found error  pops after running a few hours",
  "url": "/competitions/image-matching-challenge-2024/discussion/498731",
  "author_name": "",
  "post_date": "2024-04-29T11:43:39.954080900Z",
  "votes": 1,
  "comment_count": 7,
  "views": 0,
  "content": "<p>We used an end-to-end model for feature matching and successfully ran it on the 41 samples from test set after submission. <br>\nHowever, after running on the actual test set for a few hours, we encountered the error \"Submission CSV NOT Found\". <br>\nWe have successfully obtained scores in the following attempts(we use the original  create_submission function as in post <a href=\"https://www.kaggle.com/competitions/image-matching-challenge-2024/discussion/491321#2778205):\" target=\"_blank\">https://www.kaggle.com/competitions/image-matching-challenge-2024/discussion/491321#2778205):</a></p>\n<ul>\n<li>run our pipeline only in the first three datasets,</li>\n<li>traverse all scenes without any processing</li>\n</ul>\n<p>The scores were as low as expected, which proves that the function create_submission works well.</p>\n<p>We also used try-except to handle exceptions in image reading and processing, but the error still exists. After local testing, the submission file can always be obtained, even if the h5 files for colmap (keypoints, matches) do not exist.</p>\n<p>Can anyone share some experience of submission or give some advice?</p>",
  "messages": [
    {
      "id": "2782613",
      "postDate": "04/29/2024 11:43:39",
      "content": "<p>We used an end-to-end model for feature matching and successfully ran it on the 41 samples from test set after submission. <br>\nHowever, after running on the actual test set for a few hours, we encountered the error \"Submission CSV NOT Found\". <br>\nWe have successfully obtained scores in the following attempts(we use the original  create_submission function as in post <a href=\"https://www.kaggle.com/competitions/image-matching-challenge-2024/discussion/491321#2778205):\" target=\"_blank\">https://www.kaggle.com/competitions/image-matching-challenge-2024/discussion/491321#2778205):</a></p>\n<ul>\n<li>run our pipeline only in the first three datasets,</li>\n<li>traverse all scenes without any processing</li>\n</ul>\n<p>The scores were as low as expected, which proves that the function create_submission works well.</p>\n<p>We also used try-except to handle exceptions in image reading and processing, but the error still exists. After local testing, the submission file can always be obtained, even if the h5 files for colmap (keypoints, matches) do not exist.</p>\n<p>Can anyone share some experience of submission or give some advice?</p>",
      "rawMarkdown": "We used an end-to-end model for feature matching and successfully ran it on the 41 samples from test set after submission. \nHowever, after running on the actual test set for a few hours, we encountered the error \"Submission CSV NOT Found\". \nWe have successfully obtained scores in the following attempts(we use the original  create_submission function as in post https://www.kaggle.com/competitions/image-matching-challenge-2024/discussion/491321#2778205):\n- run our pipeline only in the first three datasets,\n- traverse all scenes without any processing\n\nThe scores were as low as expected, which proves that the function create_submission works well.\n\nWe also used try-except to handle exceptions in image reading and processing, but the error still exists. After local testing, the submission file can always be obtained, even if the h5 files for colmap (keypoints, matches) do not exist.\n\nCan anyone share some experience of submission or give some advice?",
      "votes": null
    },
    {
      "id": "2783472",
      "postDate": "04/29/2024 18:58:21",
      "content": "<ul>\n<li><p>When you make a submission, a version of the notebook will be saved to run on the sample test data. You can check if there is a <code>submission.csv</code> or any typo of the name from the output tab. <br>\n<img src=\"https://www.googleapis.com/download/storage/v1/b/kaggle-forum-message-attachments/o/inbox%2F3964695%2F324e1da61629da261e86e436c53fea1e%2FScreenshot%20from%202024-04-29%2014-35-39.png?generation=1714415756585089&amp;alt=media\"></p></li>\n<li><p>I have seen errors when the <code>submission.csv</code> is not the first (alphabetical order) file in the output (<code>/kaggle/working/</code>). I think the problem was solved by moving all other intermediate files to a <code>/kaggle/tmp/</code> directory. </p></li>\n</ul>",
      "rawMarkdown": "When you make a submission, a version of the notebook will be saved to run on the sample test data. You can check if there is a `submission.csv` or any typo of the name from the output tab. \n![](https://www.googleapis.com/download/storage/v1/b/kaggle-forum-message-attachments/o/inbox%2F3964695%2F324e1da61629da261e86e436c53fea1e%2FScreenshot%20from%202024-04-29%2014-35-39.png?generation=1714415756585089&alt=media)\n\n- I have seen errors when the `submission.csv` is not the first (alphabetical order) file in the output (`/kaggle/working/`). I think the problem was solved by moving all other intermediate files to a `/kaggle/tmp/` directory.",
      "votes": null
    },
    {
      "id": "2783799",
      "postDate": "04/30/2024 00:40:29",
      "content": "<blockquote>\n  <p>We have successfully obtained scores in the following attempts(we use the original  create_submission function as in post <a href=\"https://www.kaggle.com/competitions/image-matching-challenge-2024/discussion/491321#2778205):\" target=\"_blank\">https://www.kaggle.com/competitions/image-matching-challenge-2024/discussion/491321#2778205):</a></p>\n  <ul>\n  <li>run our pipeline only in the first three datasets,</li>\n  <li>traverse all scenes without any processing</li>\n  </ul>\n  <p>Can anyone share some experience of submission or give some advice?</p>\n</blockquote>\n<p>Since your pipeline works work the first three dataset, would that be possible that the intermediate files generated by your pipeline jammed the disk space so that your \"submissions.csv\" cannot be saved properly due to full disk space?</p>",
      "rawMarkdown": "> We have successfully obtained scores in the following attempts(we use the original  create_submission function as in post https://www.kaggle.com/competitions/image-matching-challenge-2024/discussion/491321#2778205):\n> - run our pipeline only in the first three datasets,\n> - traverse all scenes without any processing\n\n> Can anyone share some experience of submission or give some advice?\n\nSince your pipeline works work the first three dataset, would that be possible that the intermediate files generated by your pipeline jammed the disk space so that your \"submissions.csv\" cannot be saved properly due to full disk space?",
      "votes": null
    },
    {
      "id": "2784553",
      "postDate": "04/30/2024 10:51:02",
      "content": "<p>It's strange. Our pipeline does not produce much intermediate data. <br>\nThanks for your comment.</p>",
      "rawMarkdown": "It's strange. Our pipeline does not produce much intermediate data. \nThanks for your comment.",
      "votes": null
    },
    {
      "id": "2784558",
      "postDate": "04/30/2024 10:54:15",
      "content": "<p>Thanks. We will try your second suggestion. Our pipeline makes a folder in the <code>/kaggle/working/</code>. </p>",
      "rawMarkdown": "Thanks. We will try your second suggestion. Our pipeline makes a folder in the `/kaggle/working/`.",
      "votes": null
    },
    {
      "id": "2825411",
      "postDate": "05/20/2024 11:37:31",
      "content": "<p>Hello. I would like to ask you about same problem.<br>\nHow was the result of 2nd suggestion trying? Did it work?</p>",
      "rawMarkdown": "Hello. I would like to ask you about same problem.\nHow was the result of 2nd suggestion trying? Did it work?",
      "votes": null
    },
    {
      "id": "2825448",
      "postDate": "05/20/2024 12:07:25",
      "content": "<p>That problem may not be the key one. We added more exceptions handling, such as input check before colmap, image reading check, and we solved it.</p>",
      "rawMarkdown": "That problem may not be the key one. We added more exceptions handling, such as input check before colmap, image reading check, and we solved it.",
      "votes": null
    },
    {
      "id": "2825455",
      "postDate": "05/20/2024 12:13:04",
      "content": "<p>I see that it was not important thing.<br>\nI will consider your comment. Thank you so much!!!</p>",
      "rawMarkdown": "I see that it was not important thing.\nI will consider your comment. Thank you so much!!!",
      "votes": null
    }
  ],
  "comments": [
    {
      "id": 2783472,
      "author_name": "maxchen303",
      "author_url": "",
      "post_date": "04/29/2024 18:58:21",
      "content": "<ul>\n<li><p>When you make a submission, a version of the notebook will be saved to run on the sample test data. You can check if there is a <code>submission.csv</code> or any typo of the name from the output tab. <br>\n<img src=\"https://www.googleapis.com/download/storage/v1/b/kaggle-forum-message-attachments/o/inbox%2F3964695%2F324e1da61629da261e86e436c53fea1e%2FScreenshot%20from%202024-04-29%2014-35-39.png?generation=1714415756585089&amp;alt=media\"></p></li>\n<li><p>I have seen errors when the <code>submission.csv</code> is not the first (alphabetical order) file in the output (<code>/kaggle/working/</code>). I think the problem was solved by moving all other intermediate files to a <code>/kaggle/tmp/</code> directory. </p></li>\n</ul>",
      "votes": null,
      "replies": [
        {
          "id": 2784558,
          "author_name": "xikegy",
          "author_url": "",
          "post_date": "04/30/2024 10:54:15",
          "content": "<p>Thanks. We will try your second suggestion. Our pipeline makes a folder in the <code>/kaggle/working/</code>. </p>",
          "votes": null,
          "replies": [
            {
              "id": 2825411,
              "author_name": "katsuhisamasuda7",
              "author_url": "",
              "post_date": "05/20/2024 11:37:31",
              "content": "<p>Hello. I would like to ask you about same problem.<br>\nHow was the result of 2nd suggestion trying? Did it work?</p>",
              "votes": null,
              "replies": [
                {
                  "id": 2825448,
                  "author_name": "xikegy",
                  "author_url": "",
                  "post_date": "05/20/2024 12:07:25",
                  "content": "<p>That problem may not be the key one. We added more exceptions handling, such as input check before colmap, image reading check, and we solved it.</p>",
                  "votes": null,
                  "replies": [
                    {
                      "id": 2825455,
                      "author_name": "katsuhisamasuda7",
                      "author_url": "",
                      "post_date": "05/20/2024 12:13:04",
                      "content": "<p>I see that it was not important thing.<br>\nI will consider your comment. Thank you so much!!!</p>",
                      "votes": null,
                      "replies": []
                    }
                  ]
                }
              ]
            }
          ]
        }
      ]
    },
    {
      "id": 2783799,
      "author_name": "photunix",
      "author_url": "",
      "post_date": "04/30/2024 00:40:29",
      "content": "<blockquote>\n  <p>We have successfully obtained scores in the following attempts(we use the original  create_submission function as in post <a href=\"https://www.kaggle.com/competitions/image-matching-challenge-2024/discussion/491321#2778205):\" target=\"_blank\">https://www.kaggle.com/competitions/image-matching-challenge-2024/discussion/491321#2778205):</a></p>\n  <ul>\n  <li>run our pipeline only in the first three datasets,</li>\n  <li>traverse all scenes without any processing</li>\n  </ul>\n  <p>Can anyone share some experience of submission or give some advice?</p>\n</blockquote>\n<p>Since your pipeline works work the first three dataset, would that be possible that the intermediate files generated by your pipeline jammed the disk space so that your \"submissions.csv\" cannot be saved properly due to full disk space?</p>",
      "votes": null,
      "replies": [
        {
          "id": 2784553,
          "author_name": "xikegy",
          "author_url": "",
          "post_date": "04/30/2024 10:51:02",
          "content": "<p>It's strange. Our pipeline does not produce much intermediate data. <br>\nThanks for your comment.</p>",
          "votes": null,
          "replies": []
        }
      ]
    }
  ],
  "raw_markdown_by_id": {
    "2782613": "We used an end-to-end model for feature matching and successfully ran it on the 41 samples from test set after submission. \nHowever, after running on the actual test set for a few hours, we encountered the error \"Submission CSV NOT Found\". \nWe have successfully obtained scores in the following attempts(we use the original  create_submission function as in post https://www.kaggle.com/competitions/image-matching-challenge-2024/discussion/491321#2778205):\n- run our pipeline only in the first three datasets,\n- traverse all scenes without any processing\n\nThe scores were as low as expected, which proves that the function create_submission works well.\n\nWe also used try-except to handle exceptions in image reading and processing, but the error still exists. After local testing, the submission file can always be obtained, even if the h5 files for colmap (keypoints, matches) do not exist.\n\nCan anyone share some experience of submission or give some advice?",
    "2783472": "When you make a submission, a version of the notebook will be saved to run on the sample test data. You can check if there is a `submission.csv` or any typo of the name from the output tab. \n![](https://www.googleapis.com/download/storage/v1/b/kaggle-forum-message-attachments/o/inbox%2F3964695%2F324e1da61629da261e86e436c53fea1e%2FScreenshot%20from%202024-04-29%2014-35-39.png?generation=1714415756585089&alt=media)\n\n- I have seen errors when the `submission.csv` is not the first (alphabetical order) file in the output (`/kaggle/working/`). I think the problem was solved by moving all other intermediate files to a `/kaggle/tmp/` directory.",
    "2783799": "> We have successfully obtained scores in the following attempts(we use the original  create_submission function as in post https://www.kaggle.com/competitions/image-matching-challenge-2024/discussion/491321#2778205):\n> - run our pipeline only in the first three datasets,\n> - traverse all scenes without any processing\n\n> Can anyone share some experience of submission or give some advice?\n\nSince your pipeline works work the first three dataset, would that be possible that the intermediate files generated by your pipeline jammed the disk space so that your \"submissions.csv\" cannot be saved properly due to full disk space?",
    "2784553": "It's strange. Our pipeline does not produce much intermediate data. \nThanks for your comment.",
    "2784558": "Thanks. We will try your second suggestion. Our pipeline makes a folder in the `/kaggle/working/`.",
    "2825411": "Hello. I would like to ask you about same problem.\nHow was the result of 2nd suggestion trying? Did it work?",
    "2825448": "That problem may not be the key one. We added more exceptions handling, such as input check before colmap, image reading check, and we solved it.",
    "2825455": "I see that it was not important thing.\nI will consider your comment. Thank you so much!!!"
  },
  "source": "meta"
}