{
  "id": 151084,
  "title": "What are the people at the top doing differently?",
  "url": "/competitions/jigsaw-multilingual-toxic-comment-classification/discussion/151084",
  "author_name": "Mr_KnowNothing",
  "post_date": "2020-05-14T09:24:52.692000",
  "votes": 10,
  "comment_count": 23,
  "views": 0,
  "content": "<p>I would request people at the top to please share some insights of their approaches to motivate others to be in the competition. I have studied Alex's kernels and tried to apply most of the things that he has taught but I am stuck at 0.9464-0.9463 . I see no improvements with anything else\nHere is what all I have tried \n- Use translated dataset for training and validation\n- Balanced the dataset\n- Different Type of Pre-processing (I am not sure if it works )\n- Different architectures on top of XLM-Roberta</p>\n\n<p>I Have the following questions and what has worked differently for the table toppers\n- Will training on more data will help? If so how many data points? I am currently training on 200,000 examples as limited by memory usage \n- Does changing hyperparams like max_len and batch size lead to iimprovement in model ? I have tried different max_len and bs but didn't improve anything\n- Will it be a good idea to train different classifiers for different languages\n- I know a lot of people have been doing ensembling but using ensembling as brute force can be fatal rigght? What would be a good number of models to use to ensemble?</p>",
  "messages": [
    {
      "id": 849499,
      "postDate": "2020-05-15T20:06:25.460Z",
      "content": "<p>You got some questions, that's a good starting point. Next step is designing some experiments to answer those questions. </p>\n\n<p>In my personal experience, the most valuable things I learn from a kaggle competition usually come from:\n- The ideas (both successful and failed ones) I came up with and implemented myself during the competition. \n- The interesting/unique ideas in winning solutions that I integrated with my solution and verified their effectiveness after the competition ends. </p>\n\n<p>The specific ideas are important to win a competition, but the more valuable thing is how we come up with such ideas and how to verify their performances. </p>",
      "rawMarkdown": "You got some questions, that's a good starting point. Next step is designing some experiments to answer those questions. \n\nIn my personal experience, the most valuable things I learn from a kaggle competition usually come from:\n- The ideas (both successful and failed ones) I came up with and implemented myself during the competition. \n- The interesting/unique ideas in winning solutions that I integrated with my solution and verified their effectiveness after the competition ends. \n\nThe specific ideas are important to win a competition, but the more valuable thing is how we come up with such ideas and how to verify their performances. ",
      "votes": 17,
      "replies": [
        {
          "id": 849665,
          "postDate": "2020-05-16T00:41:51.463Z",
          "content": "<p>Wise words <a href=\"/naivelamb\">@naivelamb</a> , for my experience, for the last year I basically just get into competitions that I don't previously know how to solve, this way I can learn a lot, and in some way, I always win 😄 </p>",
          "rawMarkdown": "Wise words @naivelamb , for my experience, for the last year I basically just get into competitions that I don't previously know how to solve, this way I can learn a lot, and in some way, I always win 😄 ",
          "votes": 4
        },
        {
          "id": 849770,
          "postDate": "2020-05-16T03:55:16.313Z",
          "content": "<p>Thanks a lot for the lovely advices GM <a href=\"/naivelamb\">@naivelamb</a>\nI will keep these in mind always</p>",
          "rawMarkdown": "Thanks a lot for the lovely advices GM @naivelamb\nI will keep these in mind always",
          "votes": 1
        }
      ]
    },
    {
      "id": 848890,
      "postDate": "2020-05-15T10:22:11.433Z",
      "content": "<p>In my opinion, this is not in the spirit of competitions to beg for high ranked participants to share their solutions (even hints), and I will even encourage them not to \"overshare\". Especially the #1 team.</p>\n\n<p>Some people currently have an idea (or more) that can win them the competition, why give it away ?\nSharing these ideas would really not benefit anyone's ranking since they will be made public, it will only result in everybody using it and the leaderboard getting even more crowded. People who worked hard to reach high scores may suffer from this, as an approach better than theirs was made available to everybody, regardless of the time invested on the competition.</p>\n\n<p>The good time for sharing winning ideas is when the competition ends.</p>",
      "rawMarkdown": "In my opinion, this is not in the spirit of competitions to beg for high ranked participants to share their solutions (even hints), and I will even encourage them not to \"overshare\". Especially the #1 team.\n\nSome people currently have an idea (or more) that can win them the competition, why give it away ?\nSharing these ideas would really not benefit anyone's ranking since they will be made public, it will only result in everybody using it and the leaderboard getting even more crowded. People who worked hard to reach high scores may suffer from this, as an approach better than theirs was made available to everybody, regardless of the time invested on the competition.\n\nThe good time for sharing winning ideas is when the competition ends.",
      "votes": 14,
      "replies": [
        {
          "id": 848895,
          "postDate": "2020-05-15T10:32:24.527Z",
          "content": "<p>I agree that the winning solutions should only be shared only after the competition and I get your point ,However , I am not asking to share their winning Idea or Unique Idea , only a hint that too on the questions I have asked . I am sorry if I have conveyed the wrong Idea I also don't want the lb to be polluted .</p>",
          "rawMarkdown": "I agree that the winning solutions should only be shared only after the competition and I get your point ,However , I am not asking to share their winning Idea or Unique Idea , only a hint that too on the questions I have asked . I am sorry if I have conveyed the wrong Idea I also don't want the lb to be polluted .",
          "votes": -3
        },
        {
          "id": 848901,
          "postDate": "2020-05-15T10:37:18.617Z",
          "content": "<p>I totally agree with this</p>\n\n<p>Of course I was ironic (smiley) with my message about the 1st place team, as I know they won't, fortunately, share their approach before the end </p>\n\n<p>General Ideas may still be shared but spoon feeding is definitely not fair for all those who are hardworking so far to get at the top. </p>",
          "rawMarkdown": "I totally agree with this\n\nOf course I was ironic (smiley) with my message about the 1st place team, as I know they won't, fortunately, share their approach before the end \n\nGeneral Ideas may still be shared but spoon feeding is definitely not fair for all those who are hardworking so far to get at the top. ",
          "votes": 2
        }
      ]
    },
    {
      "id": 847291,
      "postDate": "2020-05-14T09:24:52.693Z",
      "content": "<p>I would request people at the top to please share some insights of their approaches to motivate others to be in the competition. I have studied Alex's kernels and tried to apply most of the things that he has taught but I am stuck at 0.9464-0.9463 . I see no improvements with anything else\nHere is what all I have tried \n- Use translated dataset for training and validation\n- Balanced the dataset\n- Different Type of Pre-processing (I am not sure if it works )\n- Different architectures on top of XLM-Roberta</p>\n\n<p>I Have the following questions and what has worked differently for the table toppers\n- Will training on more data will help? If so how many data points? I am currently training on 200,000 examples as limited by memory usage \n- Does changing hyperparams like max_len and batch size lead to iimprovement in model ? I have tried different max_len and bs but didn't improve anything\n- Will it be a good idea to train different classifiers for different languages\n- I know a lot of people have been doing ensembling but using ensembling as brute force can be fatal rigght? What would be a good number of models to use to ensemble?</p>",
      "rawMarkdown": "I would request people at the top to please share some insights of their approaches to motivate others to be in the competition. I have studied Alex's kernels and tried to apply most of the things that he has taught but I am stuck at 0.9464-0.9463 . I see no improvements with anything else\nHere is what all I have tried \n- Use translated dataset for training and validation\n- Balanced the dataset\n- Different Type of Pre-processing (I am not sure if it works )\n- Different architectures on top of XLM-Roberta\n\nI Have the following questions and what has worked differently for the table toppers\n- Will training on more data will help? If so how many data points? I am currently training on 200,000 examples as limited by memory usage \n- Does changing hyperparams like max_len and batch size lead to iimprovement in model ? I have tried different max_len and bs but didn't improve anything\n- Will it be a good idea to train different classifiers for different languages\n- I know a lot of people have been doing ensembling but using ensembling as brute force can be fatal rigght? What would be a good number of models to use to ensemble?",
      "votes": 10
    },
    {
      "id": 850286,
      "postDate": "2020-05-16T13:33:58.923Z",
      "content": "<p>Top kernel have 256 entries that shows that they are trying lot of things and very few of them works, so I think solution is try  as many things as you can and see if anything works.</p>",
      "rawMarkdown": "Top kernel have 256 entries that shows that they are trying lot of things and very few of them works, so I think solution is try  as many things as you can and see if anything works.",
      "votes": 3,
      "replies": [
        {
          "id": 850307,
          "postDate": "2020-05-16T14:00:27.763Z",
          "content": "<p>Good Point , I am also trying a lot of stuffs and their are still many days left , I hope to do better</p>",
          "rawMarkdown": "Good Point , I am also trying a lot of stuffs and their are still many days left , I hope to do better"
        }
      ]
    },
    {
      "id": 847610,
      "postDate": "2020-05-14T14:17:11.513Z",
      "content": "<p>I'm also curious to know what the 1st place team is doing ;) <br>\nThe margin is huge . </p>\n\n<p>But one of them is the co-winner of the first Jigsaw competition. </p>",
      "rawMarkdown": "I'm also curious to know what the 1st place team is doing ;)  \nThe margin is huge . \n\nBut one of them is the co-winner of the first Jigsaw competition. \n\n",
      "votes": 4,
      "replies": [
        {
          "id": 847654,
          "postDate": "2020-05-14T14:40:02.317Z",
          "content": "<p>Meanwhile it would be helpful if you share some insights of your approach . Best single model score?\nHow many models have used for ensemble? And if you can answer my above questions I will be glad</p>",
          "rawMarkdown": "Meanwhile it would be helpful if you share some insights of your approach . Best single model score?\nHow many models have used for ensemble? And if you can answer my above questions I will be glad"
        },
        {
          "id": 847768,
          "postDate": "2020-05-14T15:44:18.803Z",
          "content": "<p>I've reported my best single model score 0.9465 in <a href=\"https://www.kaggle.com/c/jigsaw-multilingual-toxic-comment-classification/discussion/141825\">this topic</a></p>\n\n<p>I've tried almost all the appoaches you listed here without really meaningful success. </p>\n\n<p>The dataset is huge...That gives the opportunity to chose carefully the subset for training. And I use about 150 000- 250 000 datapoints in my best model. </p>\n\n<p>Beware of overfitting too, given the difference in languages between the train and test set (Even if you use data translation and some languages like French and Russian have very sophisticated grammar, so difficult to translate correctly by these API ).</p>\n\n<p>Good training strategy is the key here, othewise your model will quickly overfit the training set. </p>\n\n<p>And of course ensemble WORKS.  I see no reason why it wouldn't work.</p>",
          "rawMarkdown": "I've reported my best single model score 0.9465 in [this topic](https://www.kaggle.com/c/jigsaw-multilingual-toxic-comment-classification/discussion/141825)\n\nI've tried almost all the appoaches you listed here without really meaningful success. \n\nThe dataset is huge...That gives the opportunity to chose carefully the subset for training. And I use about 150 000- 250 000 datapoints in my best model. \n\nBeware of overfitting too, given the difference in languages between the train and test set (Even if you use data translation and some languages like French and Russian have very sophisticated grammar, so difficult to translate correctly by these API ).\n\n Good training strategy is the key here, othewise your model will quickly overfit the training set. \n\n\nAnd of course ensemble WORKS.  I see no reason why it wouldn't work.\n",
          "votes": 8
        },
        {
          "id": 847904,
          "postDate": "2020-05-14T16:47:12.667Z",
          "content": "<p>Thanks for sharing\n<a href=\"/rftexas\">@rftexas</a> You need to see this</p>",
          "rawMarkdown": "Thanks for sharing\n@rftexas You need to see this",
          "votes": 1
        }
      ]
    },
    {
      "id": 857688,
      "postDate": "2020-05-22T20:30:31.277Z",
      "content": "<p><a href=\"/tanulsingh077\">@tanulsingh077</a> you can try make pseudo-labeling for test set such as <a href=\"https://www.kaggle.com/c/jigsaw-toxic-comment-classification-challenge/discussion/52557\">here</a>, I believe it can give boost </p>\n\n<p><a href=\"https://storage.googleapis.com/kaggle-forum-message-attachments/302467/8860/pl_capture.PNG\">https://storage.googleapis.com/kaggle-forum-message-attachments/302467/8860/pl_capture.PNG</a></p>",
      "rawMarkdown": "@tanulsingh077 you can try make pseudo-labeling for test set such as [here](https://www.kaggle.com/c/jigsaw-toxic-comment-classification-challenge/discussion/52557), I believe it can give boost \n\nhttps://storage.googleapis.com/kaggle-forum-message-attachments/302467/8860/pl_capture.PNG",
      "votes": 1,
      "replies": [
        {
          "id": 858239,
          "postDate": "2020-05-23T10:19:27.717Z",
          "content": "<p>Thank you <a href=\"/shonenkov\">@shonenkov</a> I will surely try it out</p>",
          "rawMarkdown": "Thank you @shonenkov I will surely try it out",
          "votes": 1
        },
        {
          "id": 876615,
          "postDate": "2020-06-06T20:37:27.153Z",
          "content": "<p><a href=\"/shonenkov\">@shonenkov</a> ,  Pseudo labelling is leaky and it will only gives huge CV boost. <a href=\"https://www.kaggle.com/c/tweet-sentiment-extraction/discussion/156556#876422\">ref here</a>. I tried it on Tweet Sentiment and my CV improved from ~0.71 to ~0.75 but no improvement on the LB. </p>",
          "rawMarkdown": "@shonenkov ,  Pseudo labelling is leaky and it will only gives huge CV boost. [ref here](https://www.kaggle.com/c/tweet-sentiment-extraction/discussion/156556#876422). I tried it on Tweet Sentiment and my CV improved from ~0.71 to ~0.75 but no improvement on the LB. "
        }
      ]
    },
    {
      "id": 847737,
      "postDate": "2020-05-14T15:20:49.783Z",
      "content": "<p>I think that those on top are very experienced persons . That's all.</p>",
      "rawMarkdown": "I think that those on top are very experienced persons . That's all.",
      "votes": 1
    },
    {
      "id": 848863,
      "postDate": "2020-05-15T09:50:45.410Z",
      "content": "<p>Maybe they manually improve their predictions :) Just need some language skills and the time to go through some thousand sentences.</p>",
      "rawMarkdown": "Maybe they manually improve their predictions :) Just need some language skills and the time to go through some thousand sentences.",
      "votes": -6,
      "replies": [
        {
          "id": 848891,
          "postDate": "2020-05-15T10:25:52.733Z",
          "content": "<p>Hand labelling is against the competition rules .  I think all top teams are aware of that</p>\n\n<p>May be if you downplay hardworking and ethics to compete fairly,  you may still try it and see if you will be at the top on both public and private LB ^^</p>",
          "rawMarkdown": "\nHand labelling is against the competition rules .  I think all top teams are aware of that\n\n\nMay be if you downplay hardworking and ethics to compete fairly,  you may still try it and see if you will be at the top on both public and private LB ^^",
          "votes": 2
        },
        {
          "id": 872390,
          "postDate": "2020-06-03T06:30:45.270Z",
          "rawMarkdown": "",
          "isDeleted": true
        }
      ]
    },
    {
      "id": 849904,
      "postDate": "2020-05-16T06:56:51.870Z",
      "rawMarkdown": "",
      "isDeleted": true,
      "replies": [
        {
          "id": 853708,
          "postDate": "2020-05-19T12:27:23.870Z",
          "content": "<p>I don't think that's a good idea.The model can easily learn the comments with curse/swear words as toxic.But there can be false positives like\n<code>Example:“Oh, I feel like such an asshole now.Sorry, bud.”</code>\nI think the model should actually learn to correctly classify this sort of comments.</p>",
          "rawMarkdown": "I don't think that's a good idea.The model can easily learn the comments with curse/swear words as toxic.But there can be false positives like\n`Example:“Oh, I feel like such an asshole now.Sorry, bud.”`\nI think the model should actually learn to correctly classify this sort of comments."
        },
        {
          "id": 890335,
          "postDate": "2020-06-17T13:04:52.310Z",
          "content": "<p>exactly, as we're in the world of multi-directional language models which take into account each word together with other word in a given sentence, e.g. in the attention mechanism.</p>",
          "rawMarkdown": "exactly, as we're in the world of multi-directional language models which take into account each word together with other word in a given sentence, e.g. in the attention mechanism."
        }
      ]
    },
    {
      "id": 847903,
      "postDate": "2020-05-14T16:47:01.357Z",
      "rawMarkdown": "",
      "votes": -1,
      "isDeleted": true
    }
  ],
  "comments": [
    {
      "id": 849499,
      "author_name": "Xuan Cao",
      "author_url": "",
      "post_date": "2020-05-15T20:06:25.460000",
      "content": "<p>You got some questions, that's a good starting point. Next step is designing some experiments to answer those questions. </p>\n\n<p>In my personal experience, the most valuable things I learn from a kaggle competition usually come from:\n- The ideas (both successful and failed ones) I came up with and implemented myself during the competition. \n- The interesting/unique ideas in winning solutions that I integrated with my solution and verified their effectiveness after the competition ends. </p>\n\n<p>The specific ideas are important to win a competition, but the more valuable thing is how we come up with such ideas and how to verify their performances. </p>",
      "votes": 17,
      "replies": [
        {
          "id": 849665,
          "author_name": "DimitreOliveira",
          "author_url": "",
          "post_date": "2020-05-16T00:41:51.463000",
          "content": "<p>Wise words <a href=\"/naivelamb\">@naivelamb</a> , for my experience, for the last year I basically just get into competitions that I don't previously know how to solve, this way I can learn a lot, and in some way, I always win 😄 </p>",
          "votes": 4,
          "replies": []
        },
        {
          "id": 849770,
          "author_name": "Mr_KnowNothing",
          "author_url": "",
          "post_date": "2020-05-16T03:55:16.313000",
          "content": "<p>Thanks a lot for the lovely advices GM <a href=\"/naivelamb\">@naivelamb</a>\nI will keep these in mind always</p>",
          "votes": 1,
          "replies": []
        }
      ]
    },
    {
      "id": 848890,
      "author_name": "Theo Viel",
      "author_url": "",
      "post_date": "2020-05-15T10:22:11.433000",
      "content": "<p>In my opinion, this is not in the spirit of competitions to beg for high ranked participants to share their solutions (even hints), and I will even encourage them not to \"overshare\". Especially the #1 team.</p>\n\n<p>Some people currently have an idea (or more) that can win them the competition, why give it away ?\nSharing these ideas would really not benefit anyone's ranking since they will be made public, it will only result in everybody using it and the leaderboard getting even more crowded. People who worked hard to reach high scores may suffer from this, as an approach better than theirs was made available to everybody, regardless of the time invested on the competition.</p>\n\n<p>The good time for sharing winning ideas is when the competition ends.</p>",
      "votes": 14,
      "replies": [
        {
          "id": 848895,
          "author_name": "Mr_KnowNothing",
          "author_url": "",
          "post_date": "2020-05-15T10:32:24.527000",
          "content": "<p>I agree that the winning solutions should only be shared only after the competition and I get your point ,However , I am not asking to share their winning Idea or Unique Idea , only a hint that too on the questions I have asked . I am sorry if I have conveyed the wrong Idea I also don't want the lb to be polluted .</p>",
          "votes": -3,
          "replies": []
        },
        {
          "id": 848901,
          "author_name": "Serigne ",
          "author_url": "",
          "post_date": "2020-05-15T10:37:18.617000",
          "content": "<p>I totally agree with this</p>\n\n<p>Of course I was ironic (smiley) with my message about the 1st place team, as I know they won't, fortunately, share their approach before the end </p>\n\n<p>General Ideas may still be shared but spoon feeding is definitely not fair for all those who are hardworking so far to get at the top. </p>",
          "votes": 2,
          "replies": []
        }
      ]
    },
    {
      "id": 850286,
      "author_name": "Maunish dave",
      "author_url": "",
      "post_date": "2020-05-16T13:33:58.923000",
      "content": "<p>Top kernel have 256 entries that shows that they are trying lot of things and very few of them works, so I think solution is try  as many things as you can and see if anything works.</p>",
      "votes": 3,
      "replies": [
        {
          "id": 850307,
          "author_name": "Mr_KnowNothing",
          "author_url": "",
          "post_date": "2020-05-16T14:00:27.763000",
          "content": "<p>Good Point , I am also trying a lot of stuffs and their are still many days left , I hope to do better</p>",
          "votes": 0,
          "replies": []
        }
      ]
    },
    {
      "id": 847610,
      "author_name": "Serigne ",
      "author_url": "",
      "post_date": "2020-05-14T14:17:11.513000",
      "content": "<p>I'm also curious to know what the 1st place team is doing ;) <br>\nThe margin is huge . </p>\n\n<p>But one of them is the co-winner of the first Jigsaw competition. </p>",
      "votes": 4,
      "replies": [
        {
          "id": 847654,
          "author_name": "Mr_KnowNothing",
          "author_url": "",
          "post_date": "2020-05-14T14:40:02.317000",
          "content": "<p>Meanwhile it would be helpful if you share some insights of your approach . Best single model score?\nHow many models have used for ensemble? And if you can answer my above questions I will be glad</p>",
          "votes": 0,
          "replies": []
        },
        {
          "id": 847768,
          "author_name": "Serigne ",
          "author_url": "",
          "post_date": "2020-05-14T15:44:18.803000",
          "content": "<p>I've reported my best single model score 0.9465 in <a href=\"https://www.kaggle.com/c/jigsaw-multilingual-toxic-comment-classification/discussion/141825\">this topic</a></p>\n\n<p>I've tried almost all the appoaches you listed here without really meaningful success. </p>\n\n<p>The dataset is huge...That gives the opportunity to chose carefully the subset for training. And I use about 150 000- 250 000 datapoints in my best model. </p>\n\n<p>Beware of overfitting too, given the difference in languages between the train and test set (Even if you use data translation and some languages like French and Russian have very sophisticated grammar, so difficult to translate correctly by these API ).</p>\n\n<p>Good training strategy is the key here, othewise your model will quickly overfit the training set. </p>\n\n<p>And of course ensemble WORKS.  I see no reason why it wouldn't work.</p>",
          "votes": 8,
          "replies": []
        },
        {
          "id": 847904,
          "author_name": "Mr_KnowNothing",
          "author_url": "",
          "post_date": "2020-05-14T16:47:12.667000",
          "content": "<p>Thanks for sharing\n<a href=\"/rftexas\">@rftexas</a> You need to see this</p>",
          "votes": 1,
          "replies": []
        }
      ]
    },
    {
      "id": 857688,
      "author_name": "Alex Shonenkov",
      "author_url": "",
      "post_date": "2020-05-22T20:30:31.277000",
      "content": "<p><a href=\"/tanulsingh077\">@tanulsingh077</a> you can try make pseudo-labeling for test set such as <a href=\"https://www.kaggle.com/c/jigsaw-toxic-comment-classification-challenge/discussion/52557\">here</a>, I believe it can give boost </p>\n\n<p><a href=\"https://storage.googleapis.com/kaggle-forum-message-attachments/302467/8860/pl_capture.PNG\">https://storage.googleapis.com/kaggle-forum-message-attachments/302467/8860/pl_capture.PNG</a></p>",
      "votes": 1,
      "replies": [
        {
          "id": 858239,
          "author_name": "Mr_KnowNothing",
          "author_url": "",
          "post_date": "2020-05-23T10:19:27.717000",
          "content": "<p>Thank you <a href=\"/shonenkov\">@shonenkov</a> I will surely try it out</p>",
          "votes": 1,
          "replies": []
        },
        {
          "id": 876615,
          "author_name": "Tushar",
          "author_url": "",
          "post_date": "2020-06-06T20:37:27.153000",
          "content": "<p><a href=\"/shonenkov\">@shonenkov</a> ,  Pseudo labelling is leaky and it will only gives huge CV boost. <a href=\"https://www.kaggle.com/c/tweet-sentiment-extraction/discussion/156556#876422\">ref here</a>. I tried it on Tweet Sentiment and my CV improved from ~0.71 to ~0.75 but no improvement on the LB. </p>",
          "votes": 0,
          "replies": []
        }
      ]
    },
    {
      "id": 847737,
      "author_name": "",
      "author_url": "",
      "post_date": "2020-05-14T15:20:49.783000",
      "content": "<p>I think that those on top are very experienced persons . That's all.</p>",
      "votes": 1,
      "replies": []
    },
    {
      "id": 848863,
      "author_name": "Damian Berger",
      "author_url": "",
      "post_date": "2020-05-15T09:50:45.410000",
      "content": "<p>Maybe they manually improve their predictions :) Just need some language skills and the time to go through some thousand sentences.</p>",
      "votes": -6,
      "replies": [
        {
          "id": 848891,
          "author_name": "Serigne ",
          "author_url": "",
          "post_date": "2020-05-15T10:25:52.733000",
          "content": "<p>Hand labelling is against the competition rules .  I think all top teams are aware of that</p>\n\n<p>May be if you downplay hardworking and ethics to compete fairly,  you may still try it and see if you will be at the top on both public and private LB ^^</p>",
          "votes": 2,
          "replies": []
        },
        {
          "id": 872390,
          "author_name": "",
          "author_url": "",
          "post_date": "2020-06-03T06:30:45.270000",
          "content": "",
          "votes": 0,
          "replies": []
        }
      ]
    },
    {
      "id": 849904,
      "author_name": "",
      "author_url": "",
      "post_date": "2020-05-16T06:56:51.870000",
      "content": "",
      "votes": 0,
      "replies": [
        {
          "id": 853708,
          "author_name": "Shahules",
          "author_url": "",
          "post_date": "2020-05-19T12:27:23.870000",
          "content": "<p>I don't think that's a good idea.The model can easily learn the comments with curse/swear words as toxic.But there can be false positives like\n<code>Example:“Oh, I feel like such an asshole now.Sorry, bud.”</code>\nI think the model should actually learn to correctly classify this sort of comments.</p>",
          "votes": 0,
          "replies": []
        },
        {
          "id": 890335,
          "author_name": "Dron Dronych",
          "author_url": "",
          "post_date": "2020-06-17T13:04:52.310000",
          "content": "<p>exactly, as we're in the world of multi-directional language models which take into account each word together with other word in a given sentence, e.g. in the attention mechanism.</p>",
          "votes": 0,
          "replies": []
        }
      ]
    },
    {
      "id": 847903,
      "author_name": "",
      "author_url": "",
      "post_date": "2020-05-14T16:47:01.357000",
      "content": "",
      "votes": -1,
      "replies": []
    }
  ],
  "raw_markdown_by_id": {
    "849499": "You got some questions, that's a good starting point. Next step is designing some experiments to answer those questions. \n\nIn my personal experience, the most valuable things I learn from a kaggle competition usually come from:\n- The ideas (both successful and failed ones) I came up with and implemented myself during the competition. \n- The interesting/unique ideas in winning solutions that I integrated with my solution and verified their effectiveness after the competition ends. \n\nThe specific ideas are important to win a competition, but the more valuable thing is how we come up with such ideas and how to verify their performances. ",
    "848890": "In my opinion, this is not in the spirit of competitions to beg for high ranked participants to share their solutions (even hints), and I will even encourage them not to \"overshare\". Especially the #1 team.\n\nSome people currently have an idea (or more) that can win them the competition, why give it away ?\nSharing these ideas would really not benefit anyone's ranking since they will be made public, it will only result in everybody using it and the leaderboard getting even more crowded. People who worked hard to reach high scores may suffer from this, as an approach better than theirs was made available to everybody, regardless of the time invested on the competition.\n\nThe good time for sharing winning ideas is when the competition ends.",
    "847291": "I would request people at the top to please share some insights of their approaches to motivate others to be in the competition. I have studied Alex's kernels and tried to apply most of the things that he has taught but I am stuck at 0.9464-0.9463 . I see no improvements with anything else\nHere is what all I have tried \n- Use translated dataset for training and validation\n- Balanced the dataset\n- Different Type of Pre-processing (I am not sure if it works )\n- Different architectures on top of XLM-Roberta\n\nI Have the following questions and what has worked differently for the table toppers\n- Will training on more data will help? If so how many data points? I am currently training on 200,000 examples as limited by memory usage \n- Does changing hyperparams like max_len and batch size lead to iimprovement in model ? I have tried different max_len and bs but didn't improve anything\n- Will it be a good idea to train different classifiers for different languages\n- I know a lot of people have been doing ensembling but using ensembling as brute force can be fatal rigght? What would be a good number of models to use to ensemble?",
    "850286": "Top kernel have 256 entries that shows that they are trying lot of things and very few of them works, so I think solution is try  as many things as you can and see if anything works.",
    "847610": "I'm also curious to know what the 1st place team is doing ;)  \nThe margin is huge . \n\nBut one of them is the co-winner of the first Jigsaw competition. \n\n",
    "857688": "@tanulsingh077 you can try make pseudo-labeling for test set such as [here](https://www.kaggle.com/c/jigsaw-toxic-comment-classification-challenge/discussion/52557), I believe it can give boost \n\nhttps://storage.googleapis.com/kaggle-forum-message-attachments/302467/8860/pl_capture.PNG",
    "847737": "I think that those on top are very experienced persons . That's all.",
    "848863": "Maybe they manually improve their predictions :) Just need some language skills and the time to go through some thousand sentences.",
    "849904": "",
    "847903": ""
  }
}