{
  "id": 1169,
  "title": "General Questions",
  "url": "/competitions/GestureChallenge/discussion/1169",
  "author_name": "",
  "post_date": "2011-12-16T06:54:13.870Z",
  "votes": null,
  "comment_count": 1,
  "views": 1819,
  "content": "<p><strong>General Questions:</strong></p>\r\n<p>So since their are 2 types either &quot;devel&quot; or &quot;valid&quot; and 20 choies for each of those, there would be a grand total of 40 batches with each batch having 100 gestures?</p>\r\n<p>So by saying that their are &quot;There are instances of N unique gestures from a vocabulary of 8 to 15 gestures&quot; you are saying that each batch contains a unique subset combination of the 8-15 gestures? Actually I think it would be permutation since order does\r\n matter(what order the gestures are done in)?</p>\r\n<p>Just wanted to confirm that I do not have to publish any papers on this topic correct seeing as to how the deadlines for submission are before final evaluations?</p>\r\n<p><strong>Training Data/Test Data</strong></p>\r\n<p>Am I allowed to use cross validation or do I have to keep data as training data and testing data seperate?</p>\r\n<p>Along the same lines am I allowed to use methods such as boosting/bagging?</p>\r\n<p>I don't really understand how to read the training data/test data:</p>\r\n<p>Ex: In &quot;devel01_train.csv&quot;, cell A1 specifies following:&nbsp;</p>\r\n<table border=\"0\" cellspacing=\"0\" cellpadding=\"0\" style=\"width:64px\">\r\n<tbody>\r\n<tr>\r\n<td width=\"64\" height=\"20\">\r\n<p>devel01_1,10</p>\r\n</td>\r\n</tr>\r\n</tbody>\r\n</table>\r\n<p>Should I go ahead and split the data on the commas, because in excel all the data is shown in one column?</p>\r\n<p>So devel01_1 would be the row id which wold be unique and 10 would be a label? What is this label? Is that label the classification?</p>\r\n<p>Thanks!</p>",
  "messages": [
    {
      "id": "7228",
      "postDate": "12/16/2011 06:54:13",
      "content": "<p><strong>General Questions:</strong></p>\r\n<p>So since their are 2 types either &quot;devel&quot; or &quot;valid&quot; and 20 choies for each of those, there would be a grand total of 40 batches with each batch having 100 gestures?</p>\r\n<p>So by saying that their are &quot;There are instances of N unique gestures from a vocabulary of 8 to 15 gestures&quot; you are saying that each batch contains a unique subset combination of the 8-15 gestures? Actually I think it would be permutation since order does\r\n matter(what order the gestures are done in)?</p>\r\n<p>Just wanted to confirm that I do not have to publish any papers on this topic correct seeing as to how the deadlines for submission are before final evaluations?</p>\r\n<p><strong>Training Data/Test Data</strong></p>\r\n<p>Am I allowed to use cross validation or do I have to keep data as training data and testing data seperate?</p>\r\n<p>Along the same lines am I allowed to use methods such as boosting/bagging?</p>\r\n<p>I don't really understand how to read the training data/test data:</p>\r\n<p>Ex: In &quot;devel01_train.csv&quot;, cell A1 specifies following:&nbsp;</p>\r\n<table border=\"0\" cellspacing=\"0\" cellpadding=\"0\" style=\"width:64px\">\r\n<tbody>\r\n<tr>\r\n<td width=\"64\" height=\"20\">\r\n<p>devel01_1,10</p>\r\n</td>\r\n</tr>\r\n</tbody>\r\n</table>\r\n<p>Should I go ahead and split the data on the commas, because in excel all the data is shown in one column?</p>\r\n<p>So devel01_1 would be the row id which wold be unique and 10 would be a label? What is this label? Is that label the classification?</p>\r\n<p>Thanks!</p>",
      "rawMarkdown": "",
      "votes": null
    },
    {
      "id": "7248",
      "postDate": "12/16/2011 19:16:00",
      "content": "<p><strong>Data:</strong></p>\r\n<p>You are now seeing a subset of the data, we will be releasing shortly more develxx batches for development. For final testing we will release finalxx batches.</p>\r\n<p>All the batches are organized in the same way. The all contain:</p>\r\n<p>- 47 M files (RGB videos) and 47 corresponding K files (depth videos)</p>\r\n<p>- The video contain sequences of 1 to 5 recorded gestures. In total 100 gestures are recorded.</p>\r\n<p>- The valid and final batches contain labels for the first N videos used as training examples. The devel batches also contain the labels of the other videos (the test examples).</p>\r\n<p>The number N &nbsp;maybe different for each batch. It is between 8 and 15. It corresponds to the number of gesture tokens (unique gestures) in the lexicon associated to that batch (different gestures are played in different batches, e.g., one has diving signals,\r\n another has referee signals, etc.). The gestures are played in the various videos following a script that prescribed to the users in which order to play them. The script was obtained by drawing at random labels from the N possible values.</p>\r\n<p><strong>Publication:</strong></p>\r\n<p>Publishing methods is optional, we just provide this opportunity ro the participants.</p>\r\n<p><strong>Development data:</strong></p>\r\n<p>You are allowed to use the develxx files in any way you want. In particular, you can do cross-validation. We provide a split between training and test data to make the devel batches look like the valid and final batches. But this split does not need to be\r\n kept for your purposes. I am not sure how you want to use boosting/bagging, but feel free to try. In the end, the task will be:</p>\r\n<p>Given new batches from new lexicons of gestures you have not seen before, train on the N first examples and make predictions on the remainder.</p>\r\n<p><strong>CSV file format:</strong></p>\r\n<p>- First column =&gt; the ID of the example. devel01_1 means the 1st example of the batch devel01. valid04_12 means the 12th example of the batch valid04.</p>\r\n<p>- Second column =&gt; the list of gestures that were played (space separated)</p>\r\n<p>&nbsp;</p>",
      "rawMarkdown": "",
      "votes": null
    }
  ],
  "comments": [
    {
      "id": 7248,
      "author_name": "challengeadmin",
      "author_url": "",
      "post_date": "12/16/2011 19:16:00",
      "content": "<p><strong>Data:</strong></p>\r\n<p>You are now seeing a subset of the data, we will be releasing shortly more develxx batches for development. For final testing we will release finalxx batches.</p>\r\n<p>All the batches are organized in the same way. The all contain:</p>\r\n<p>- 47 M files (RGB videos) and 47 corresponding K files (depth videos)</p>\r\n<p>- The video contain sequences of 1 to 5 recorded gestures. In total 100 gestures are recorded.</p>\r\n<p>- The valid and final batches contain labels for the first N videos used as training examples. The devel batches also contain the labels of the other videos (the test examples).</p>\r\n<p>The number N &nbsp;maybe different for each batch. It is between 8 and 15. It corresponds to the number of gesture tokens (unique gestures) in the lexicon associated to that batch (different gestures are played in different batches, e.g., one has diving signals,\r\n another has referee signals, etc.). The gestures are played in the various videos following a script that prescribed to the users in which order to play them. The script was obtained by drawing at random labels from the N possible values.</p>\r\n<p><strong>Publication:</strong></p>\r\n<p>Publishing methods is optional, we just provide this opportunity ro the participants.</p>\r\n<p><strong>Development data:</strong></p>\r\n<p>You are allowed to use the develxx files in any way you want. In particular, you can do cross-validation. We provide a split between training and test data to make the devel batches look like the valid and final batches. But this split does not need to be\r\n kept for your purposes. I am not sure how you want to use boosting/bagging, but feel free to try. In the end, the task will be:</p>\r\n<p>Given new batches from new lexicons of gestures you have not seen before, train on the N first examples and make predictions on the remainder.</p>\r\n<p><strong>CSV file format:</strong></p>\r\n<p>- First column =&gt; the ID of the example. devel01_1 means the 1st example of the batch devel01. valid04_12 means the 12th example of the batch valid04.</p>\r\n<p>- Second column =&gt; the list of gestures that were played (space separated)</p>\r\n<p>&nbsp;</p>",
      "votes": null,
      "replies": []
    }
  ],
  "raw_markdown_by_id": {
    "7228": "",
    "7248": ""
  },
  "source": "meta"
}