{
  "id": 540608,
  "title": "Be careful and test your code",
  "url": "/competitions/jane-street-real-time-market-data-forecasting/discussion/540608",
  "author_name": "",
  "post_date": "2024-10-15T11:16:58.728311400Z",
  "votes": 18,
  "comment_count": 3,
  "views": 0,
  "content": "<p>As this is another code competition where the model will need to handle new data after the competition ends, I want to remind everyone to be mindful when writing code. I've touched on this in a previous post:</p>\n<p><a href=\"https://www.kaggle.com/competitions/optiver-trading-at-the-close/discussion/457610\" target=\"_blank\">https://www.kaggle.com/competitions/optiver-trading-at-the-close/discussion/457610</a></p>\n<p>Being careless can lead to losing months of effort, and unfortunately, it's something that happened often in the past. Take the time to write solid, test-driven code.</p>",
  "messages": [
    {
      "id": "3017925",
      "postDate": "10/15/2024 11:16:58",
      "content": "<p>As this is another code competition where the model will need to handle new data after the competition ends, I want to remind everyone to be mindful when writing code. I've touched on this in a previous post:</p>\n<p><a href=\"https://www.kaggle.com/competitions/optiver-trading-at-the-close/discussion/457610\" target=\"_blank\">https://www.kaggle.com/competitions/optiver-trading-at-the-close/discussion/457610</a></p>\n<p>Being careless can lead to losing months of effort, and unfortunately, it's something that happened often in the past. Take the time to write solid, test-driven code.</p>",
      "rawMarkdown": "As this is another code competition where the model will need to handle new data after the competition ends, I want to remind everyone to be mindful when writing code. I've touched on this in a previous post:\n\nhttps://www.kaggle.com/competitions/optiver-trading-at-the-close/discussion/457610\n\nBeing careless can lead to losing months of effort, and unfortunately, it's something that happened often in the past. Take the time to write solid, test-driven code.",
      "votes": null
    },
    {
      "id": "3017930",
      "postDate": "10/15/2024 11:24:38",
      "content": "<p><a href=\"https://www.kaggle.com/fritzcremer\" target=\"_blank\">@fritzcremer</a> chance of error is low here as we have float32 columns here. Errors are mostly caused with improper null treatments with non-tree models. I think this is a big factor here as well.<br>\nAlso, chance of OOM issues are lesser here due to the API structure, but the chance of 10-minute inference could be an issue for large and complex pipelines. </p>",
      "rawMarkdown": "fritzcremer chance of error is low here as we have float32 columns here. Errors are mostly caused with improper null treatments with non-tree models. I think this is a big factor here as well.\nAlso, chance of OOM issues are lesser here due to the API structure, but the chance of 10-minute inference could be an issue for large and complex pipelines.",
      "votes": null
    },
    {
      "id": "3018011",
      "postDate": "10/15/2024 12:56:31",
      "content": "<p>Right, code-failiure is probably not as likely as with some other competitions. But still, people may implement complex pipelines and there is always to find a way to break things. People should be especially careful when they are used to comfort of normal code-competitions, which already run your code on the private test data without showing the score and tell you if it crashed or not.</p>",
      "rawMarkdown": "Right, code-failiure is probably not as likely as with some other competitions. But still, people may implement complex pipelines and there is always to find a way to break things. People should be especially careful when they are used to comfort of normal code-competitions, which already run your code on the private test data without showing the score and tell you if it crashed or not.",
      "votes": null
    },
    {
      "id": "3018153",
      "postDate": "10/15/2024 14:42:04",
      "content": "<p>That is a big problem in code competitions. One gets submission errors but no error statement leading to all sorts of confusion <a href=\"https://www.kaggle.com/fritzcremer\" target=\"_blank\">@fritzcremer</a> </p>",
      "rawMarkdown": "That is a big problem in code competitions. One gets submission errors but no error statement leading to all sorts of confusion @fritzcremer",
      "votes": null
    }
  ],
  "comments": [
    {
      "id": 3017930,
      "author_name": "ravi20076",
      "author_url": "",
      "post_date": "10/15/2024 11:24:38",
      "content": "<p><a href=\"https://www.kaggle.com/fritzcremer\" target=\"_blank\">@fritzcremer</a> chance of error is low here as we have float32 columns here. Errors are mostly caused with improper null treatments with non-tree models. I think this is a big factor here as well.<br>\nAlso, chance of OOM issues are lesser here due to the API structure, but the chance of 10-minute inference could be an issue for large and complex pipelines. </p>",
      "votes": null,
      "replies": [
        {
          "id": 3018011,
          "author_name": "fritzcremer",
          "author_url": "",
          "post_date": "10/15/2024 12:56:31",
          "content": "<p>Right, code-failiure is probably not as likely as with some other competitions. But still, people may implement complex pipelines and there is always to find a way to break things. People should be especially careful when they are used to comfort of normal code-competitions, which already run your code on the private test data without showing the score and tell you if it crashed or not.</p>",
          "votes": null,
          "replies": [
            {
              "id": 3018153,
              "author_name": "ravi20076",
              "author_url": "",
              "post_date": "10/15/2024 14:42:04",
              "content": "<p>That is a big problem in code competitions. One gets submission errors but no error statement leading to all sorts of confusion <a href=\"https://www.kaggle.com/fritzcremer\" target=\"_blank\">@fritzcremer</a> </p>",
              "votes": null,
              "replies": []
            }
          ]
        }
      ]
    }
  ],
  "raw_markdown_by_id": {
    "3017925": "As this is another code competition where the model will need to handle new data after the competition ends, I want to remind everyone to be mindful when writing code. I've touched on this in a previous post:\n\nhttps://www.kaggle.com/competitions/optiver-trading-at-the-close/discussion/457610\n\nBeing careless can lead to losing months of effort, and unfortunately, it's something that happened often in the past. Take the time to write solid, test-driven code.",
    "3017930": "fritzcremer chance of error is low here as we have float32 columns here. Errors are mostly caused with improper null treatments with non-tree models. I think this is a big factor here as well.\nAlso, chance of OOM issues are lesser here due to the API structure, but the chance of 10-minute inference could be an issue for large and complex pipelines.",
    "3018011": "Right, code-failiure is probably not as likely as with some other competitions. But still, people may implement complex pipelines and there is always to find a way to break things. People should be especially careful when they are used to comfort of normal code-competitions, which already run your code on the private test data without showing the score and tell you if it crashed or not.",
    "3018153": "That is a big problem in code competitions. One gets submission errors but no error statement leading to all sorts of confusion @fritzcremer"
  },
  "source": "meta"
}