{
  "id": 397289,
  "title": "Audio Segmentation",
  "url": "/competitions/ml-olympiad-dialectrecognition/discussion/397289",
  "author_name": "",
  "post_date": "2023-03-24T21:56:43.063067Z",
  "votes": 2,
  "comment_count": 2,
  "views": 0,
  "content": "<p>Hello Community,</p>\n<p>I hope you are all doing well, and enjoying your projects in Ramadan. I wanted to share a resource that might be helpful to you especially if you are approaching the dialect recognition problem as an audio problem. It is a segmented audio dataset for all dialects in the SADA corpus, each segment labeled with the corresponding <code>SegmentID.</code>+<code>.wav</code> aligned with their dialect and  transcripts. </p>\n<p>The segmentation makes it easier and faster to process and infer from the audio data. You can find the dataset here: <a href=\"https://www.kaggle.com/code/amjadkhatabi/segmented-audio-data-for-arabic-dialects-sada/notebook\" target=\"_blank\">https://www.kaggle.com/code/amjadkhatabi/segmented-audio-data-for-arabic-dialects-sada/notebook</a></p>\n<p>I hope you find this resource useful and interesting. Please feel free to share your feedback or questions with me. I would love to hear from you.</p>\n<p>Best wishes</p>",
  "messages": [
    {
      "id": "2195758",
      "postDate": "03/24/2023 21:56:43",
      "content": "<p>Hello Community,</p>\n<p>I hope you are all doing well, and enjoying your projects in Ramadan. I wanted to share a resource that might be helpful to you especially if you are approaching the dialect recognition problem as an audio problem. It is a segmented audio dataset for all dialects in the SADA corpus, each segment labeled with the corresponding <code>SegmentID.</code>+<code>.wav</code> aligned with their dialect and  transcripts. </p>\n<p>The segmentation makes it easier and faster to process and infer from the audio data. You can find the dataset here: <a href=\"https://www.kaggle.com/code/amjadkhatabi/segmented-audio-data-for-arabic-dialects-sada/notebook\" target=\"_blank\">https://www.kaggle.com/code/amjadkhatabi/segmented-audio-data-for-arabic-dialects-sada/notebook</a></p>\n<p>I hope you find this resource useful and interesting. Please feel free to share your feedback or questions with me. I would love to hear from you.</p>\n<p>Best wishes</p>",
      "rawMarkdown": "Hello Community,\n\nI hope you are all doing well, and enjoying your projects in Ramadan. I wanted to share a resource that might be helpful to you especially if you are approaching the dialect recognition problem as an audio problem. It is a segmented audio dataset for all dialects in the SADA corpus, each segment labeled with the corresponding `SegmentID.`+`.wav` aligned with their dialect and  transcripts. \n\nThe segmentation makes it easier and faster to process and infer from the audio data. You can find the dataset here: https://www.kaggle.com/code/amjadkhatabi/segmented-audio-data-for-arabic-dialects-sada/notebook\n\nI hope you find this resource useful and interesting. Please feel free to share your feedback or questions with me. I would love to hear from you.\n\nBest wishes",
      "votes": null
    },
    {
      "id": "2195764",
      "postDate": "03/24/2023 22:05:05",
      "content": "<p>Thank you so much for sharing</p>",
      "rawMarkdown": "Thank you so much for sharing",
      "votes": null
    },
    {
      "id": "2195809",
      "postDate": "03/24/2023 22:43:19",
      "content": "<p>My pleasure! Many thanks to you and all community for this exceptional learning experience.</p>",
      "rawMarkdown": "My pleasure! Many thanks to you and all community for this exceptional learning experience.",
      "votes": null
    }
  ],
  "comments": [
    {
      "id": 2195764,
      "author_name": "ruqiyas",
      "author_url": "",
      "post_date": "03/24/2023 22:05:05",
      "content": "<p>Thank you so much for sharing</p>",
      "votes": null,
      "replies": [
        {
          "id": 2195809,
          "author_name": "amjadkhatabi",
          "author_url": "",
          "post_date": "03/24/2023 22:43:19",
          "content": "<p>My pleasure! Many thanks to you and all community for this exceptional learning experience.</p>",
          "votes": null,
          "replies": []
        }
      ]
    }
  ],
  "raw_markdown_by_id": {
    "2195758": "Hello Community,\n\nI hope you are all doing well, and enjoying your projects in Ramadan. I wanted to share a resource that might be helpful to you especially if you are approaching the dialect recognition problem as an audio problem. It is a segmented audio dataset for all dialects in the SADA corpus, each segment labeled with the corresponding `SegmentID.`+`.wav` aligned with their dialect and  transcripts. \n\nThe segmentation makes it easier and faster to process and infer from the audio data. You can find the dataset here: https://www.kaggle.com/code/amjadkhatabi/segmented-audio-data-for-arabic-dialects-sada/notebook\n\nI hope you find this resource useful and interesting. Please feel free to share your feedback or questions with me. I would love to hear from you.\n\nBest wishes",
    "2195764": "Thank you so much for sharing",
    "2195809": "My pleasure! Many thanks to you and all community for this exceptional learning experience."
  },
  "source": "meta"
}