{
  "id": 438498,
  "title": "How to deal with layout nlp data?",
  "url": "/competitions/predict-ai-model-runtime/discussion/438498",
  "author_name": "ljjsxx",
  "post_date": "2023-09-11T13:15:11.116000",
  "votes": 10,
  "comment_count": 11,
  "views": 0,
  "content": "<p>My model works very well on layout xla dataset, but has no effect at all on layout nlp. What should I do? Can you tell me some models that handle layout nlp dataset? Thank you so much！</p>",
  "messages": [
    {
      "id": 2433284,
      "postDate": "2023-09-11T13:15:11.117Z",
      "content": "<p>My model works very well on layout xla dataset, but has no effect at all on layout nlp. What should I do? Can you tell me some models that handle layout nlp dataset? Thank you so much！</p>",
      "rawMarkdown": "My model works very well on layout xla dataset, but has no effect at all on layout nlp. What should I do? Can you tell me some models that handle layout nlp dataset? Thank you so much！",
      "votes": 8
    },
    {
      "id": 2454517,
      "postDate": "2023-09-24T21:48:00.920Z",
      "content": "<p>I also have very poor performance across training, validation and testing on the NLP data, which is surprising since very similar results between NLP and XLA are reported in the paper. <a href=\"https://www.kaggle.com/mangpophothilimthana\" target=\"_blank\">@mangpophothilimthana</a> is there any chance that the data might be corrupted?</p>",
      "rawMarkdown": "I also have very poor performance across training, validation and testing on the NLP data, which is surprising since very similar results between NLP and XLA are reported in the paper. @mangpophothilimthana is there any chance that the data might be corrupted?",
      "votes": 5,
      "replies": [
        {
          "id": 2454541,
          "postDate": "2023-09-24T22:47:11.750Z",
          "content": "<p>I was thinking the same thing… In the paper they show similar results between nlp and xla default and even suggest that xla random is harder to train than nlp default which in reality is not the case. </p>",
          "rawMarkdown": "I was thinking the same thing... In the paper they show similar results between nlp and xla default and even suggest that xla random is harder to train than nlp default which in reality is not the case. ",
          "replies": [
            {
              "id": 2454547,
              "postDate": "2023-09-24T22:54:47.850Z",
              "rawMarkdown": "",
              "isDeleted": true
            }
          ]
        },
        {
          "id": 2458961,
          "postDate": "2023-09-27T23:43:34.910Z",
          "content": "<p>Thank you so much for raising this concern! We checked the data, and some parts of the data were indeed corrupted. We have regenerated the data, and tried our best to verified that the new data is correct. Please see this <a href=\"https://www.kaggle.com/competitions/predict-ai-model-runtime/discussion/443581\" target=\"_blank\">post</a> about the new data. Sorry for the inconvenient this may have caused.</p>",
          "rawMarkdown": "Thank you so much for raising this concern! We checked the data, and some parts of the data were indeed corrupted. We have regenerated the data, and tried our best to verified that the new data is correct. Please see this [post](https://www.kaggle.com/competitions/predict-ai-model-runtime/discussion/443581) about the new data. Sorry for the inconvenient this may have caused.",
          "votes": 3
        }
      ]
    },
    {
      "id": 2438162,
      "postDate": "2023-09-14T05:06:10.550Z",
      "content": "<p>If you are using one model for both NLP and XLA datasets, then this is a good reason to use different models for them. </p>",
      "rawMarkdown": "If you are using one model for both NLP and XLA datasets, then this is a good reason to use different models for them. ",
      "replies": [
        {
          "id": 2440430,
          "postDate": "2023-09-15T14:00:36.363Z",
          "content": "<p>Thanks I will try</p>",
          "rawMarkdown": "Thanks I will try"
        }
      ]
    },
    {
      "id": 2433928,
      "postDate": "2023-09-12T01:19:25.600Z",
      "content": "<p>Same problem. I used SAGE and GCN, but they didn't work well in NLP. I also tried to train the layout files together but the effect was worse</p>",
      "rawMarkdown": "Same problem. I used SAGE and GCN, but they didn't work well in NLP. I also tried to train the layout files together but the effect was worse\n\n\n",
      "replies": [
        {
          "id": 2433949,
          "postDate": "2023-09-12T02:09:30.387Z",
          "content": "<p>Hello, have you ever tried the github code provided by the organizer? I tried it but it was difficult to run on kaggle.</p>",
          "rawMarkdown": "Hello, have you ever tried the github code provided by the organizer? I tried it but it was difficult to run on kaggle.",
          "replies": [
            {
              "id": 2433959,
              "postDate": "2023-09-12T02:18:55.257Z",
              "content": "<p>No, I refer to the kaggle code and improve it. I'll try the github code</p>",
              "rawMarkdown": "No, I refer to the kaggle code and improve it. I'll try the github code"
            }
          ]
        },
        {
          "id": 2440032,
          "postDate": "2023-09-15T08:22:07.950Z",
          "content": "<p>Did you implement GST in your xla/nlp model?</p>",
          "rawMarkdown": "Did you implement GST in your xla/nlp model?",
          "replies": [
            {
              "id": 2440432,
              "postDate": "2023-09-15T14:01:28.193Z",
              "content": "<p>I haven't use GST. I will try later</p>",
              "rawMarkdown": "I haven't use GST. I will try later",
              "votes": 1
            }
          ]
        }
      ]
    }
  ],
  "comments": [
    {
      "id": 2454517,
      "author_name": "Peiyuan Liao",
      "author_url": "",
      "post_date": "2023-09-24T21:48:00.920000",
      "content": "<p>I also have very poor performance across training, validation and testing on the NLP data, which is surprising since very similar results between NLP and XLA are reported in the paper. <a href=\"https://www.kaggle.com/mangpophothilimthana\" target=\"_blank\">@mangpophothilimthana</a> is there any chance that the data might be corrupted?</p>",
      "votes": 5,
      "replies": [
        {
          "id": 2454541,
          "author_name": "Amit Aharoni",
          "author_url": "",
          "post_date": "2023-09-24T22:47:11.750000",
          "content": "<p>I was thinking the same thing… In the paper they show similar results between nlp and xla default and even suggest that xla random is harder to train than nlp default which in reality is not the case. </p>",
          "votes": 0,
          "replies": [
            {
              "id": 2454547,
              "author_name": "",
              "author_url": "",
              "post_date": "2023-09-24T22:54:47.850000",
              "content": "",
              "votes": 0,
              "replies": []
            }
          ]
        },
        {
          "id": 2458961,
          "author_name": "Mangpo Phothilimthana",
          "author_url": "",
          "post_date": "2023-09-27T23:43:34.910000",
          "content": "<p>Thank you so much for raising this concern! We checked the data, and some parts of the data were indeed corrupted. We have regenerated the data, and tried our best to verified that the new data is correct. Please see this <a href=\"https://www.kaggle.com/competitions/predict-ai-model-runtime/discussion/443581\" target=\"_blank\">post</a> about the new data. Sorry for the inconvenient this may have caused.</p>",
          "votes": 3,
          "replies": []
        }
      ]
    },
    {
      "id": 2438162,
      "author_name": "PassengerC07",
      "author_url": "",
      "post_date": "2023-09-14T05:06:10.550000",
      "content": "<p>If you are using one model for both NLP and XLA datasets, then this is a good reason to use different models for them. </p>",
      "votes": 0,
      "replies": [
        {
          "id": 2440430,
          "author_name": "ljjsxx",
          "author_url": "",
          "post_date": "2023-09-15T14:00:36.363000",
          "content": "<p>Thanks I will try</p>",
          "votes": 0,
          "replies": []
        }
      ]
    },
    {
      "id": 2433928,
      "author_name": "MiHu",
      "author_url": "",
      "post_date": "2023-09-12T01:19:25.600000",
      "content": "<p>Same problem. I used SAGE and GCN, but they didn't work well in NLP. I also tried to train the layout files together but the effect was worse</p>",
      "votes": 0,
      "replies": [
        {
          "id": 2433949,
          "author_name": "ljjsxx",
          "author_url": "",
          "post_date": "2023-09-12T02:09:30.387000",
          "content": "<p>Hello, have you ever tried the github code provided by the organizer? I tried it but it was difficult to run on kaggle.</p>",
          "votes": 0,
          "replies": [
            {
              "id": 2433959,
              "author_name": "MiHu",
              "author_url": "",
              "post_date": "2023-09-12T02:18:55.257000",
              "content": "<p>No, I refer to the kaggle code and improve it. I'll try the github code</p>",
              "votes": 0,
              "replies": []
            }
          ]
        },
        {
          "id": 2440032,
          "author_name": "Amit Aharoni",
          "author_url": "",
          "post_date": "2023-09-15T08:22:07.950000",
          "content": "<p>Did you implement GST in your xla/nlp model?</p>",
          "votes": 0,
          "replies": [
            {
              "id": 2440432,
              "author_name": "ljjsxx",
              "author_url": "",
              "post_date": "2023-09-15T14:01:28.193000",
              "content": "<p>I haven't use GST. I will try later</p>",
              "votes": 1,
              "replies": []
            }
          ]
        }
      ]
    }
  ],
  "raw_markdown_by_id": {
    "2433284": "My model works very well on layout xla dataset, but has no effect at all on layout nlp. What should I do? Can you tell me some models that handle layout nlp dataset? Thank you so much！",
    "2454517": "I also have very poor performance across training, validation and testing on the NLP data, which is surprising since very similar results between NLP and XLA are reported in the paper. @mangpophothilimthana is there any chance that the data might be corrupted?",
    "2438162": "If you are using one model for both NLP and XLA datasets, then this is a good reason to use different models for them. ",
    "2433928": "Same problem. I used SAGE and GCN, but they didn't work well in NLP. I also tried to train the layout files together but the effect was worse\n\n\n"
  }
}