{
  "id": 176660,
  "title": "Is my understanding of problem statement correct?",
  "url": "/competitions/osic-pulmonary-fibrosis-progression/discussion/176660",
  "author_name": "",
  "post_date": "2020-08-22T20:06:11.602894Z",
  "votes": 1,
  "comment_count": 2,
  "views": 0,
  "content": "<p>For each patient, we are given his info and for each week he goes for test the CT scan (Dicom images ) are given and for training label we are given the FCV and percent (confidence)<br>\nIn the test data, we are given the patients info and Dicom images and we are asked to predict the FCV and percent(confidence level)<br>\nIs my understanding correct?</p>\n<p><strong>Another doubt: How is confidence level calculated?</strong><br>\nfrom my understanding, the confidence level (say 95%)tells that if I take the N samples from the population and build a 95% confidence interval on each sample then 95% percent of the samples will have the population mean between the confidence interval. <br>\n Is this definition anyhow related to the calculation of confidence here?</p>",
  "messages": [
    {
      "id": "981895",
      "postDate": "08/22/2020 20:06:11",
      "content": "<p>For each patient, we are given his info and for each week he goes for test the CT scan (Dicom images ) are given and for training label we are given the FCV and percent (confidence)<br>\nIn the test data, we are given the patients info and Dicom images and we are asked to predict the FCV and percent(confidence level)<br>\nIs my understanding correct?</p>\n<p><strong>Another doubt: How is confidence level calculated?</strong><br>\nfrom my understanding, the confidence level (say 95%)tells that if I take the N samples from the population and build a 95% confidence interval on each sample then 95% percent of the samples will have the population mean between the confidence interval. <br>\n Is this definition anyhow related to the calculation of confidence here?</p>",
      "rawMarkdown": "For each patient, we are given his info and for each week he goes for test the CT scan (Dicom images ) are given and for training label we are given the FCV and percent (confidence)\nIn the test data, we are given the patients info and Dicom images and we are asked to predict the FCV and percent(confidence level)\nIs my understanding correct?\n\n**Another doubt: How is confidence level calculated?**\nfrom my understanding, the confidence level (say 95%)tells that if I take the N samples from the population and build a 95% confidence interval on each sample then 95% percent of the samples will have the population mean between the confidence interval. \n Is this definition anyhow related to the calculation of confidence here?",
      "votes": null
    },
    {
      "id": "981925",
      "postDate": "08/22/2020 21:10:10",
      "content": "<p>Hi. We are given baseline information, a single baseline CT (week 0), and subsequent FVC measurements are taken (other weeks). Although asked to predict FVCs for all weeks (-12 to 133) in reality only \"the last 3 visits\" (which are unknown to us) are going to be evaluated.</p>\n<p>I think the term confidence is unfortunate and is an overloading of the term. It's confusing because a lower \"confidence\" value actually means that the model is \"more confident\" about its prediction.</p>\n<p>Given the metric, other users have <a href=\"https://www.kaggle.com/c/osic-pulmonary-fibrosis-progression/discussion/173636\" target=\"_blank\">shown the relationship between \"confidence\" and predicted FVC</a> that will give the optimal score. </p>\n<p>Users in their notebooks, particularly those using Quantile Regression, often calculate the \"confidence\" as the difference between the upper value and the lower value.</p>",
      "rawMarkdown": "Hi. We are given baseline information, a single baseline CT (week 0), and subsequent FVC measurements are taken (other weeks). Although asked to predict FVCs for all weeks (-12 to 133) in reality only \"the last 3 visits\" (which are unknown to us) are going to be evaluated.\n\nI think the term confidence is unfortunate and is an overloading of the term. It's confusing because a lower \"confidence\" value actually means that the model is \"more confident\" about its prediction.\n\nGiven the metric, other users have [shown the relationship between \"confidence\" and predicted FVC](https://www.kaggle.com/c/osic-pulmonary-fibrosis-progression/discussion/173636) that will give the optimal score. \n\nUsers in their notebooks, particularly those using Quantile Regression, often calculate the \"confidence\" as the difference between the upper value and the lower value.",
      "votes": null
    },
    {
      "id": "981962",
      "postDate": "08/22/2020 22:18:46",
      "content": "<p>No, there is only one CT scan per patient. The CT scan is considered the week 0, weeks before are negatives. But the patients keep revising the doctor and doing FVC tests, which shows how many air their lungs are capable to blow (the disease keeps damaging the lungs).  In the test, we need to use both image and the FVC tests to predict the FVC and confidence interval of our predictions.</p>",
      "rawMarkdown": "No, there is only one CT scan per patient. The CT scan is considered the week 0, weeks before are negatives. But the patients keep revising the doctor and doing FVC tests, which shows how many air their lungs are capable to blow (the disease keeps damaging the lungs).  In the test, we need to use both image and the FVC tests to predict the FVC and confidence interval of our predictions.",
      "votes": null
    }
  ],
  "comments": [
    {
      "id": 981925,
      "author_name": "jjinho",
      "author_url": "",
      "post_date": "08/22/2020 21:10:10",
      "content": "<p>Hi. We are given baseline information, a single baseline CT (week 0), and subsequent FVC measurements are taken (other weeks). Although asked to predict FVCs for all weeks (-12 to 133) in reality only \"the last 3 visits\" (which are unknown to us) are going to be evaluated.</p>\n<p>I think the term confidence is unfortunate and is an overloading of the term. It's confusing because a lower \"confidence\" value actually means that the model is \"more confident\" about its prediction.</p>\n<p>Given the metric, other users have <a href=\"https://www.kaggle.com/c/osic-pulmonary-fibrosis-progression/discussion/173636\" target=\"_blank\">shown the relationship between \"confidence\" and predicted FVC</a> that will give the optimal score. </p>\n<p>Users in their notebooks, particularly those using Quantile Regression, often calculate the \"confidence\" as the difference between the upper value and the lower value.</p>",
      "votes": null,
      "replies": []
    },
    {
      "id": 981962,
      "author_name": "cafalchio",
      "author_url": "",
      "post_date": "08/22/2020 22:18:46",
      "content": "<p>No, there is only one CT scan per patient. The CT scan is considered the week 0, weeks before are negatives. But the patients keep revising the doctor and doing FVC tests, which shows how many air their lungs are capable to blow (the disease keeps damaging the lungs).  In the test, we need to use both image and the FVC tests to predict the FVC and confidence interval of our predictions.</p>",
      "votes": null,
      "replies": []
    }
  ],
  "raw_markdown_by_id": {
    "981895": "For each patient, we are given his info and for each week he goes for test the CT scan (Dicom images ) are given and for training label we are given the FCV and percent (confidence)\nIn the test data, we are given the patients info and Dicom images and we are asked to predict the FCV and percent(confidence level)\nIs my understanding correct?\n\n**Another doubt: How is confidence level calculated?**\nfrom my understanding, the confidence level (say 95%)tells that if I take the N samples from the population and build a 95% confidence interval on each sample then 95% percent of the samples will have the population mean between the confidence interval. \n Is this definition anyhow related to the calculation of confidence here?",
    "981925": "Hi. We are given baseline information, a single baseline CT (week 0), and subsequent FVC measurements are taken (other weeks). Although asked to predict FVCs for all weeks (-12 to 133) in reality only \"the last 3 visits\" (which are unknown to us) are going to be evaluated.\n\nI think the term confidence is unfortunate and is an overloading of the term. It's confusing because a lower \"confidence\" value actually means that the model is \"more confident\" about its prediction.\n\nGiven the metric, other users have [shown the relationship between \"confidence\" and predicted FVC](https://www.kaggle.com/c/osic-pulmonary-fibrosis-progression/discussion/173636) that will give the optimal score. \n\nUsers in their notebooks, particularly those using Quantile Regression, often calculate the \"confidence\" as the difference between the upper value and the lower value.",
    "981962": "No, there is only one CT scan per patient. The CT scan is considered the week 0, weeks before are negatives. But the patients keep revising the doctor and doing FVC tests, which shows how many air their lungs are capable to blow (the disease keeps damaging the lungs).  In the test, we need to use both image and the FVC tests to predict the FVC and confidence interval of our predictions."
  },
  "source": "meta"
}