{
  "id": 488145,
  "title": "Submission scoring error?",
  "url": "/competitions/home-credit-credit-risk-model-stability/discussion/488145",
  "author_name": "",
  "post_date": "2024-04-01T08:49:51.994702800Z",
  "votes": 1,
  "comment_count": 8,
  "views": 0,
  "content": "<p>Hello, all.</p>\n<p>I've made my own notebook and run everything OK with the notebook itself.<br>\nAnd I've submitted the notebook but I got the error of submission scoring error.<br>\nI checked the log and everything was fine, and I can even see the predicted / created file of sample test set.</p>\n<p>I have no idea why this error occurs…</p>\n<p>Does any of you experiences like me?</p>",
  "messages": [
    {
      "id": "2726530",
      "postDate": "04/01/2024 08:49:51",
      "content": "<p>Hello, all.</p>\n<p>I've made my own notebook and run everything OK with the notebook itself.<br>\nAnd I've submitted the notebook but I got the error of submission scoring error.<br>\nI checked the log and everything was fine, and I can even see the predicted / created file of sample test set.</p>\n<p>I have no idea why this error occurs…</p>\n<p>Does any of you experiences like me?</p>",
      "rawMarkdown": "Hello, all.\n\nI've made my own notebook and run everything OK with the notebook itself.\nAnd I've submitted the notebook but I got the error of submission scoring error.\nI checked the log and everything was fine, and I can even see the predicted / created file of sample test set.\n\nI have no idea why this error occurs...\n\nDoes any of you experiences like me?",
      "votes": null
    },
    {
      "id": "2726969",
      "postDate": "04/01/2024 14:06:42",
      "content": "<p>Hey I had a similar bug driving me crazy a couple days ago. Basically I was generating multiple case id predictions because of bad aggregations. So make sure you are only outputting 10 rows of predictions (case id is always without duplicated).<br>\nGood luck!</p>",
      "rawMarkdown": "Hey I had a similar bug driving me crazy a couple days ago. Basically I was generating multiple case id predictions because of bad aggregations. So make sure you are only outputting 10 rows of predictions (case id is always without duplicated).\nGood luck!",
      "votes": null
    },
    {
      "id": "2726985",
      "postDate": "04/01/2024 14:19:29",
      "content": "<p>In submissions section check the error written under the version number of failed scoring, it could be -  notebook out of memory </p>",
      "rawMarkdown": "In submissions section check the error written under the version number of failed scoring, it could be -  notebook out of memory",
      "votes": null
    },
    {
      "id": "2727818",
      "postDate": "04/02/2024 00:20:52",
      "content": "<p><img src=\"https://www.googleapis.com/download/storage/v1/b/kaggle-forum-message-attachments/o/inbox%2F568032%2Ffe991159effba42ec133cbfa0c74b4b5%2F2024-04-02%20%209.20.13.png?generation=1712017224124757&amp;alt=media\" alt=\"Image\"><br>\nLog says it ran suceeded but scoring is failed</p>",
      "rawMarkdown": "![Image](https://www.googleapis.com/download/storage/v1/b/kaggle-forum-message-attachments/o/inbox%2F568032%2Ffe991159effba42ec133cbfa0c74b4b5%2F2024-04-02%20%209.20.13.png?generation=1712017224124757&alt=media)\nLog says it ran suceeded but scoring is failed",
      "votes": null
    },
    {
      "id": "2727821",
      "postDate": "04/02/2024 00:22:34",
      "content": "<p><img src=\"https://www.googleapis.com/download/storage/v1/b/kaggle-forum-message-attachments/o/inbox%2F568032%2F34928b9cffefeceaa57b961592029719%2F2024-04-02%20%209.21.18.png?generation=1712017297304300&amp;alt=media\" alt=\"image\"><br>\nOutput of the sample test set is exactly 10. You meant, there could be a problem in real test set (that is hidden)?</p>",
      "rawMarkdown": "![image](https://www.googleapis.com/download/storage/v1/b/kaggle-forum-message-attachments/o/inbox%2F568032%2F34928b9cffefeceaa57b961592029719%2F2024-04-02%20%209.21.18.png?generation=1712017297304300&alt=media)\nOutput of the sample test set is exactly 10. You meant, there could be a problem in real test set (that is hidden)?",
      "votes": null
    },
    {
      "id": "2728249",
      "postDate": "04/02/2024 06:38:12",
      "content": "<p>You are RIGHT. Thank you so much. I figured out that I didn't do the aggregate with 'other_1' file. I thought it has only 1 row for case_id, as it is in the training data set. It might have one or more rows for each unique case_id in actual test dataset! Thanks!</p>",
      "rawMarkdown": "You are RIGHT. Thank you so much. I figured out that I didn't do the aggregate with 'other_1' file. I thought it has only 1 row for case_id, as it is in the training data set. It might have one or more rows for each unique case_id in actual test dataset! Thanks!",
      "votes": null
    },
    {
      "id": "2728260",
      "postDate": "04/02/2024 06:44:26",
      "content": "<p>Try: submission['score'] = submission['score'].fillna(0)</p>",
      "rawMarkdown": "Try: submission['score'] = submission['score'].fillna(0)",
      "votes": null
    },
    {
      "id": "2729017",
      "postDate": "04/02/2024 14:11:56",
      "content": "<p>Glad to help! Can you give me some tip to improve score, or wether you have set up CV with correlation on LB? 🤝🤝🤝</p>",
      "rawMarkdown": "Glad to help! Can you give me some tip to improve score, or wether you have set up CV with correlation on LB? 🤝🤝🤝",
      "votes": null
    },
    {
      "id": "2731962",
      "postDate": "04/03/2024 00:11:14",
      "content": "<p>Sorry, I don't get the decent score right now with the ones I made on my own. .. around 0.374 only.</p>",
      "rawMarkdown": "Sorry, I don't get the decent score right now with the ones I made on my own. .. around 0.374 only.",
      "votes": null
    }
  ],
  "comments": [
    {
      "id": 2726969,
      "author_name": "davidcanorosillo",
      "author_url": "",
      "post_date": "04/01/2024 14:06:42",
      "content": "<p>Hey I had a similar bug driving me crazy a couple days ago. Basically I was generating multiple case id predictions because of bad aggregations. So make sure you are only outputting 10 rows of predictions (case id is always without duplicated).<br>\nGood luck!</p>",
      "votes": null,
      "replies": [
        {
          "id": 2727821,
          "author_name": "gilgarad",
          "author_url": "",
          "post_date": "04/02/2024 00:22:34",
          "content": "<p><img src=\"https://www.googleapis.com/download/storage/v1/b/kaggle-forum-message-attachments/o/inbox%2F568032%2F34928b9cffefeceaa57b961592029719%2F2024-04-02%20%209.21.18.png?generation=1712017297304300&amp;alt=media\" alt=\"image\"><br>\nOutput of the sample test set is exactly 10. You meant, there could be a problem in real test set (that is hidden)?</p>",
          "votes": null,
          "replies": []
        },
        {
          "id": 2728249,
          "author_name": "gilgarad",
          "author_url": "",
          "post_date": "04/02/2024 06:38:12",
          "content": "<p>You are RIGHT. Thank you so much. I figured out that I didn't do the aggregate with 'other_1' file. I thought it has only 1 row for case_id, as it is in the training data set. It might have one or more rows for each unique case_id in actual test dataset! Thanks!</p>",
          "votes": null,
          "replies": [
            {
              "id": 2729017,
              "author_name": "davidcanorosillo",
              "author_url": "",
              "post_date": "04/02/2024 14:11:56",
              "content": "<p>Glad to help! Can you give me some tip to improve score, or wether you have set up CV with correlation on LB? 🤝🤝🤝</p>",
              "votes": null,
              "replies": [
                {
                  "id": 2731962,
                  "author_name": "gilgarad",
                  "author_url": "",
                  "post_date": "04/03/2024 00:11:14",
                  "content": "<p>Sorry, I don't get the decent score right now with the ones I made on my own. .. around 0.374 only.</p>",
                  "votes": null,
                  "replies": []
                }
              ]
            }
          ]
        }
      ]
    },
    {
      "id": 2726985,
      "author_name": "eu1234",
      "author_url": "",
      "post_date": "04/01/2024 14:19:29",
      "content": "<p>In submissions section check the error written under the version number of failed scoring, it could be -  notebook out of memory </p>",
      "votes": null,
      "replies": [
        {
          "id": 2727818,
          "author_name": "gilgarad",
          "author_url": "",
          "post_date": "04/02/2024 00:20:52",
          "content": "<p><img src=\"https://www.googleapis.com/download/storage/v1/b/kaggle-forum-message-attachments/o/inbox%2F568032%2Ffe991159effba42ec133cbfa0c74b4b5%2F2024-04-02%20%209.20.13.png?generation=1712017224124757&amp;alt=media\" alt=\"Image\"><br>\nLog says it ran suceeded but scoring is failed</p>",
          "votes": null,
          "replies": []
        }
      ]
    },
    {
      "id": 2728260,
      "author_name": "fengpan23",
      "author_url": "",
      "post_date": "04/02/2024 06:44:26",
      "content": "<p>Try: submission['score'] = submission['score'].fillna(0)</p>",
      "votes": null,
      "replies": []
    }
  ],
  "raw_markdown_by_id": {
    "2726530": "Hello, all.\n\nI've made my own notebook and run everything OK with the notebook itself.\nAnd I've submitted the notebook but I got the error of submission scoring error.\nI checked the log and everything was fine, and I can even see the predicted / created file of sample test set.\n\nI have no idea why this error occurs...\n\nDoes any of you experiences like me?",
    "2726969": "Hey I had a similar bug driving me crazy a couple days ago. Basically I was generating multiple case id predictions because of bad aggregations. So make sure you are only outputting 10 rows of predictions (case id is always without duplicated).\nGood luck!",
    "2726985": "In submissions section check the error written under the version number of failed scoring, it could be -  notebook out of memory",
    "2727818": "![Image](https://www.googleapis.com/download/storage/v1/b/kaggle-forum-message-attachments/o/inbox%2F568032%2Ffe991159effba42ec133cbfa0c74b4b5%2F2024-04-02%20%209.20.13.png?generation=1712017224124757&alt=media)\nLog says it ran suceeded but scoring is failed",
    "2727821": "![image](https://www.googleapis.com/download/storage/v1/b/kaggle-forum-message-attachments/o/inbox%2F568032%2F34928b9cffefeceaa57b961592029719%2F2024-04-02%20%209.21.18.png?generation=1712017297304300&alt=media)\nOutput of the sample test set is exactly 10. You meant, there could be a problem in real test set (that is hidden)?",
    "2728249": "You are RIGHT. Thank you so much. I figured out that I didn't do the aggregate with 'other_1' file. I thought it has only 1 row for case_id, as it is in the training data set. It might have one or more rows for each unique case_id in actual test dataset! Thanks!",
    "2728260": "Try: submission['score'] = submission['score'].fillna(0)",
    "2729017": "Glad to help! Can you give me some tip to improve score, or wether you have set up CV with correlation on LB? 🤝🤝🤝",
    "2731962": "Sorry, I don't get the decent score right now with the ones I made on my own. .. around 0.374 only."
  },
  "source": "meta"
}