{
  "id": 600105,
  "title": "Dataset Downloading Problem",
  "url": "/competitions/rsna-intracranial-aneurysm-detection/discussion/600105",
  "author_name": "wallace",
  "post_date": "2025-08-21T00:15:26.978000",
  "votes": 3,
  "comment_count": 9,
  "views": 0,
  "content": "<p>Hello, I have a couple of questions about the dataset:</p>\n<ol>\n<li><p>I tried downloading the dataset using the Kaggle CLI (<code>kaggle competitions download -c rsna-intracranial-aneurysm-detection</code>), but it failed. I also tried downloading it from the website; however, after I clicked the link, the option disappeared. Could anyone advise what might be going on?</p></li>\n<li><p>It appears the data are organized by <code>SeriesInstanceUID</code>within the <code>series</code> folder. The number of series doesn’t necessarily equal the number of patients, correct? To determine the patient count, I would need to retrieve the PatientID for each series and then count the unique values, is that right?</p></li>\n</ol>\n<p>Thank you.</p>",
  "messages": [
    {
      "id": 3272439,
      "postDate": "2025-08-21T00:15:26.977Z",
      "content": "<p>Hello, I have a couple of questions about the dataset:</p>\n<ol>\n<li><p>I tried downloading the dataset using the Kaggle CLI (<code>kaggle competitions download -c rsna-intracranial-aneurysm-detection</code>), but it failed. I also tried downloading it from the website; however, after I clicked the link, the option disappeared. Could anyone advise what might be going on?</p></li>\n<li><p>It appears the data are organized by <code>SeriesInstanceUID</code>within the <code>series</code> folder. The number of series doesn’t necessarily equal the number of patients, correct? To determine the patient count, I would need to retrieve the PatientID for each series and then count the unique values, is that right?</p></li>\n</ol>\n<p>Thank you.</p>",
      "rawMarkdown": "Hello, I have a couple of questions about the dataset:\n\n1. I tried downloading the dataset using the Kaggle CLI (`kaggle competitions download -c rsna-intracranial-aneurysm-detection`), but it failed. I also tried downloading it from the website; however, after I clicked the link, the option disappeared. Could anyone advise what might be going on?\n\n2. It appears the data are organized by `SeriesInstanceUID `within the `series` folder. The number of series doesn’t necessarily equal the number of patients, correct? To determine the patient count, I would need to retrieve the PatientID for each series and then count the unique values, is that right?\n\nThank you.",
      "votes": 3
    },
    {
      "id": 3272751,
      "postDate": "2025-08-21T14:37:27.183Z",
      "content": "<p>The dataset was recently updated. Can you try downloading again?</p>\n<p>Please note that PatientID is completely randomized on a per-series basis. We recommend that you treat each series as independent.</p>",
      "rawMarkdown": "The dataset was recently updated. Can you try downloading again?\n\nPlease note that PatientID is completely randomized on a per-series basis. We recommend that you treat each series as independent.",
      "votes": 1,
      "replies": [
        {
          "id": 3272758,
          "postDate": "2025-08-21T14:45:17.563Z",
          "content": "<p>OK I confirm that it is showing ok. <br>\nThanks so much </p>",
          "rawMarkdown": "OK I confirm that it is showing ok. \nThanks so much "
        }
      ]
    },
    {
      "id": 3272905,
      "postDate": "2025-08-21T18:07:09.700Z",
      "content": "<p>Same here, I'm trying to download on a VM but it eats up all the RAM and nothing happens</p>",
      "rawMarkdown": "Same here, I'm trying to download on a VM but it eats up all the RAM and nothing happens",
      "replies": [
        {
          "id": 3272913,
          "postDate": "2025-08-21T18:32:36.653Z",
          "content": "<p>At first it does something weird yes.. I waited and it downloaded.. <br>\nNumber of series changed.. now &gt; 5k from previous ~4k</p>",
          "rawMarkdown": "At first it does something weird yes.. I waited and it downloaded.. \nNumber of series changed.. now > 5k from previous ~4k\n\n",
          "replies": [
            {
              "id": 3272928,
              "postDate": "2025-08-21T19:11:53.723Z",
              "content": "<p>How long did it take for you?</p>",
              "rawMarkdown": "How long did it take for you?"
            }
          ]
        }
      ]
    },
    {
      "id": 3272548,
      "postDate": "2025-08-21T07:49:24.090Z",
      "content": "<p>me too.. the dataset is not available. </p>",
      "rawMarkdown": "me too.. the dataset is not available. "
    },
    {
      "id": 3272464,
      "postDate": "2025-08-21T02:25:42.197Z",
      "content": "<p>Same issue</p>",
      "rawMarkdown": "Same issue"
    },
    {
      "id": 3272462,
      "postDate": "2025-08-21T02:22:34.050Z",
      "content": "<p>My dataset status is unavailable, does any one have same problem?</p>",
      "rawMarkdown": "My dataset status is unavailable, does any one have same problem?"
    },
    {
      "id": 3272443,
      "postDate": "2025-08-21T00:44:54.943Z",
      "content": "<p>I faced the same issue. Hope they will update soon. </p>",
      "rawMarkdown": "I faced the same issue. Hope they will update soon. \n"
    }
  ],
  "comments": [
    {
      "id": 3272751,
      "author_name": "Evan Calabrese",
      "author_url": "",
      "post_date": "2025-08-21T14:37:27.183000",
      "content": "<p>The dataset was recently updated. Can you try downloading again?</p>\n<p>Please note that PatientID is completely randomized on a per-series basis. We recommend that you treat each series as independent.</p>",
      "votes": 1,
      "replies": [
        {
          "id": 3272758,
          "author_name": "NikolAiD153",
          "author_url": "",
          "post_date": "2025-08-21T14:45:17.563000",
          "content": "<p>OK I confirm that it is showing ok. <br>\nThanks so much </p>",
          "votes": 0,
          "replies": []
        }
      ]
    },
    {
      "id": 3272905,
      "author_name": "ericmit25",
      "author_url": "",
      "post_date": "2025-08-21T18:07:09.700000",
      "content": "<p>Same here, I'm trying to download on a VM but it eats up all the RAM and nothing happens</p>",
      "votes": 0,
      "replies": [
        {
          "id": 3272913,
          "author_name": "NikolAiD153",
          "author_url": "",
          "post_date": "2025-08-21T18:32:36.653000",
          "content": "<p>At first it does something weird yes.. I waited and it downloaded.. <br>\nNumber of series changed.. now &gt; 5k from previous ~4k</p>",
          "votes": 0,
          "replies": [
            {
              "id": 3272928,
              "author_name": "ericmit25",
              "author_url": "",
              "post_date": "2025-08-21T19:11:53.723000",
              "content": "<p>How long did it take for you?</p>",
              "votes": 0,
              "replies": []
            }
          ]
        }
      ]
    },
    {
      "id": 3272548,
      "author_name": "NikolAiD153",
      "author_url": "",
      "post_date": "2025-08-21T07:49:24.090000",
      "content": "<p>me too.. the dataset is not available. </p>",
      "votes": 0,
      "replies": []
    },
    {
      "id": 3272464,
      "author_name": "Arunodhayan",
      "author_url": "",
      "post_date": "2025-08-21T02:25:42.197000",
      "content": "<p>Same issue</p>",
      "votes": 0,
      "replies": []
    },
    {
      "id": 3272462,
      "author_name": "Tom",
      "author_url": "",
      "post_date": "2025-08-21T02:22:34.050000",
      "content": "<p>My dataset status is unavailable, does any one have same problem?</p>",
      "votes": 0,
      "replies": []
    },
    {
      "id": 3272443,
      "author_name": "jucious",
      "author_url": "",
      "post_date": "2025-08-21T00:44:54.943000",
      "content": "<p>I faced the same issue. Hope they will update soon. </p>",
      "votes": 0,
      "replies": []
    }
  ],
  "raw_markdown_by_id": {
    "3272439": "Hello, I have a couple of questions about the dataset:\n\n1. I tried downloading the dataset using the Kaggle CLI (`kaggle competitions download -c rsna-intracranial-aneurysm-detection`), but it failed. I also tried downloading it from the website; however, after I clicked the link, the option disappeared. Could anyone advise what might be going on?\n\n2. It appears the data are organized by `SeriesInstanceUID `within the `series` folder. The number of series doesn’t necessarily equal the number of patients, correct? To determine the patient count, I would need to retrieve the PatientID for each series and then count the unique values, is that right?\n\nThank you.",
    "3272751": "The dataset was recently updated. Can you try downloading again?\n\nPlease note that PatientID is completely randomized on a per-series basis. We recommend that you treat each series as independent.",
    "3272905": "Same here, I'm trying to download on a VM but it eats up all the RAM and nothing happens",
    "3272548": "me too.. the dataset is not available. ",
    "3272464": "Same issue",
    "3272462": "My dataset status is unavailable, does any one have same problem?",
    "3272443": "I faced the same issue. Hope they will update soon. \n"
  }
}