{
  "id": 667554,
  "title": "Why the Next Competition Needs a Fixed Dataset ?",
  "url": "/competitions/duality-ai-lunate-ai-geospatial-object-detection/discussion/667554",
  "author_name": "",
  "post_date": "2026-01-13T05:22:40.208416900Z",
  "votes": null,
  "comment_count": 2,
  "views": 0,
  "content": "<p>In the next competition, if the authority provides a fixed dataset, competitors will be able to focus more on code and algorithms rather than spending most of their time creating data. With Falcon being a generative model, the outputs vary every time, so results change unpredictably. Because of this, success depends less on skill and more on luck. A fixed dataset would make the competition fairer and more skill-based.</p>",
  "messages": [
    {
      "id": "3390408",
      "postDate": "01/13/2026 05:22:40",
      "content": "<p>In the next competition, if the authority provides a fixed dataset, competitors will be able to focus more on code and algorithms rather than spending most of their time creating data. With Falcon being a generative model, the outputs vary every time, so results change unpredictably. Because of this, success depends less on skill and more on luck. A fixed dataset would make the competition fairer and more skill-based.</p>",
      "rawMarkdown": "In the next competition, if the authority provides a fixed dataset, competitors will be able to focus more on code and algorithms rather than spending most of their time creating data. With Falcon being a generative model, the outputs vary every time, so results change unpredictably. Because of this, success depends less on skill and more on luck. A fixed dataset would make the competition fairer and more skill-based.",
      "votes": null
    },
    {
      "id": "3390566",
      "postDate": "01/13/2026 11:58:40",
      "content": "<p>I agree, but the main idea of the competition is to generate your own data - it's part of the competition. It's cool for me because in the future you can use falcon for your other purposes.</p>",
      "rawMarkdown": "I agree, but the main idea of the competition is to generate your own data - it's part of the competition. It's cool for me because in the future you can use falcon for your other purposes.",
      "votes": null
    },
    {
      "id": "3390703",
      "postDate": "01/13/2026 19:34:46",
      "content": "<p>Part of the competition is understanding data needs for model improvement, which is why users are asked to create their own data. This is a very common skill for synthetic data use. Please feel free to reach out for advice on how to use the tools available to create effective data if you're needing support using the set parameters!\nThis element does differentiate our competitions from other ones available on the platform, but we see it as an important part of working with synthetic data, since the ability to create specific data elements is a major reason companies would employ synthetic data pipelines.</p>",
      "rawMarkdown": "Part of the competition is understanding data needs for model improvement, which is why users are asked to create their own data. This is a very common skill for synthetic data use. Please feel free to reach out for advice on how to use the tools available to create effective data if you're needing support using the set parameters!\nThis element does differentiate our competitions from other ones available on the platform, but we see it as an important part of working with synthetic data, since the ability to create specific data elements is a major reason companies would employ synthetic data pipelines.",
      "votes": null
    }
  ],
  "comments": [
    {
      "id": 3390566,
      "author_name": "antonoof",
      "author_url": "",
      "post_date": "01/13/2026 11:58:40",
      "content": "<p>I agree, but the main idea of the competition is to generate your own data - it's part of the competition. It's cool for me because in the future you can use falcon for your other purposes.</p>",
      "votes": null,
      "replies": []
    },
    {
      "id": 3390703,
      "author_name": "rebekahduality",
      "author_url": "",
      "post_date": "01/13/2026 19:34:46",
      "content": "<p>Part of the competition is understanding data needs for model improvement, which is why users are asked to create their own data. This is a very common skill for synthetic data use. Please feel free to reach out for advice on how to use the tools available to create effective data if you're needing support using the set parameters!\nThis element does differentiate our competitions from other ones available on the platform, but we see it as an important part of working with synthetic data, since the ability to create specific data elements is a major reason companies would employ synthetic data pipelines.</p>",
      "votes": null,
      "replies": []
    }
  ],
  "raw_markdown_by_id": {
    "3390408": "In the next competition, if the authority provides a fixed dataset, competitors will be able to focus more on code and algorithms rather than spending most of their time creating data. With Falcon being a generative model, the outputs vary every time, so results change unpredictably. Because of this, success depends less on skill and more on luck. A fixed dataset would make the competition fairer and more skill-based.",
    "3390566": "I agree, but the main idea of the competition is to generate your own data - it's part of the competition. It's cool for me because in the future you can use falcon for your other purposes.",
    "3390703": "Part of the competition is understanding data needs for model improvement, which is why users are asked to create their own data. This is a very common skill for synthetic data use. Please feel free to reach out for advice on how to use the tools available to create effective data if you're needing support using the set parameters!\nThis element does differentiate our competitions from other ones available on the platform, but we see it as an important part of working with synthetic data, since the ability to create specific data elements is a major reason companies would employ synthetic data pipelines."
  },
  "source": "meta"
}