{
  "id": 369125,
  "title": "New to Kaggle or Machine Learning? Check this out ~",
  "url": "/competitions/rsna-breast-cancer-detection/discussion/369125",
  "author_name": "Maggie",
  "post_date": "2022-11-29T04:05:13.718000",
  "votes": 15,
  "comment_count": 25,
  "views": 0,
  "content": "<p>New to machine learning and data science? No question is too basic or too simple. Feel free to start your own thread, or use this thread as a place to post any first-timer clarifying questions for the Kaggle community to help you with!</p>\n<p>If you would consider yourself a beginner but don't know where to get started, let other Kagglers help you take your first steps here!</p>\n<p>New to Kaggle? Take a look at a few videos Dr. Rachael Tatman has put together to learn a bit more about <a href=\"https://www.youtube.com/watch?v=aIus8si_Et0\" target=\"_blank\">site etiquette</a>, or <a href=\"https://www.youtube.com/watch?v=sEJHyuWKd-s\" target=\"_blank\">Kaggle lingo</a>.</p>\n<p>Remember: Kaggle is for everyone. Whether you're teaming up or sharing tips in the competition forum, we expect everyone to follow our <a href=\"https://www.kaggle.com/community-guidelines\" target=\"_blank\">Kaggle community guidelines</a>.</p>\n<p>A tip on sharing content - Kaggle is a collaborative community, whereby sharing techniques, starter notebooks, and ideas in the discussion forums are highly encouraged throughout the competition. However, as the competition draws closer to the final deadline it's customary to keep high-scoring notebooks withheld until after the competition has concluded. This maintains the spirit of the competition, while also allowing individuals to submit their own creative work without jeopardy of a higher-scoring notebook being available for an automatic higher rank (through copy/submit). We disable publishing of public notebooks within the final week of the competition, but encourage you to use your best judgment prior to that the deadline.</p>\n<p>Happy Modeling!</p>",
  "messages": [
    {
      "id": 2047812,
      "postDate": "2022-11-29T04:05:13.717Z",
      "content": "<p>New to machine learning and data science? No question is too basic or too simple. Feel free to start your own thread, or use this thread as a place to post any first-timer clarifying questions for the Kaggle community to help you with!</p>\n<p>If you would consider yourself a beginner but don't know where to get started, let other Kagglers help you take your first steps here!</p>\n<p>New to Kaggle? Take a look at a few videos Dr. Rachael Tatman has put together to learn a bit more about <a href=\"https://www.youtube.com/watch?v=aIus8si_Et0\" target=\"_blank\">site etiquette</a>, or <a href=\"https://www.youtube.com/watch?v=sEJHyuWKd-s\" target=\"_blank\">Kaggle lingo</a>.</p>\n<p>Remember: Kaggle is for everyone. Whether you're teaming up or sharing tips in the competition forum, we expect everyone to follow our <a href=\"https://www.kaggle.com/community-guidelines\" target=\"_blank\">Kaggle community guidelines</a>.</p>\n<p>A tip on sharing content - Kaggle is a collaborative community, whereby sharing techniques, starter notebooks, and ideas in the discussion forums are highly encouraged throughout the competition. However, as the competition draws closer to the final deadline it's customary to keep high-scoring notebooks withheld until after the competition has concluded. This maintains the spirit of the competition, while also allowing individuals to submit their own creative work without jeopardy of a higher-scoring notebook being available for an automatic higher rank (through copy/submit). We disable publishing of public notebooks within the final week of the competition, but encourage you to use your best judgment prior to that the deadline.</p>\n<p>Happy Modeling!</p>",
      "rawMarkdown": "New to machine learning and data science? No question is too basic or too simple. Feel free to start your own thread, or use this thread as a place to post any first-timer clarifying questions for the Kaggle community to help you with!\n\nIf you would consider yourself a beginner but don't know where to get started, let other Kagglers help you take your first steps here!\n\nNew to Kaggle? Take a look at a few videos Dr. Rachael Tatman has put together to learn a bit more about [site etiquette](https://www.youtube.com/watch?v=aIus8si_Et0), or [Kaggle lingo](https://www.youtube.com/watch?v=sEJHyuWKd-s).\n\nRemember: Kaggle is for everyone. Whether you're teaming up or sharing tips in the competition forum, we expect everyone to follow our [Kaggle community guidelines](https://www.kaggle.com/community-guidelines).\n\nA tip on sharing content - Kaggle is a collaborative community, whereby sharing techniques, starter notebooks, and ideas in the discussion forums are highly encouraged throughout the competition. However, as the competition draws closer to the final deadline it's customary to keep high-scoring notebooks withheld until after the competition has concluded. This maintains the spirit of the competition, while also allowing individuals to submit their own creative work without jeopardy of a higher-scoring notebook being available for an automatic higher rank (through copy/submit). We disable publishing of public notebooks within the final week of the competition, but encourage you to use your best judgment prior to that the deadline.\n\n  \n\nHappy Modeling!\n",
      "votes": 15
    },
    {
      "id": 2085792,
      "postDate": "2023-01-04T11:57:52.937Z",
      "content": "<p>Hi all :) <br>\nI want to ask a question about downloading data.<br>\nWhen clicking  **Data ** under the ribbon that include :<br>\n<em>Overview Data Code Discussion Leaderboard Rules Team</em>  there is option/ tab  that called <br>\n** Download All**  which include 324.72 GB. <br>\nI do not have so much free space on the laptop I work on. <br>\nWanted to ask if there is an option to download part of the data? </p>\n<p>I'm not sure if related, there is a link to Kaggle  API  installation and documentation, However I got lost there with unfamiliar terminology, and was sure if it's related to data downloading. <br>\nKindly your advice or guidance <br>\nTal </p>",
      "rawMarkdown": "Hi all :) \nI want to ask a question about downloading data.\nWhen clicking  **Data ** under the ribbon that include :\n*Overview Data Code Discussion Leaderboard Rules Team*  there is option/ tab  that called \n** Download All**  which include 324.72 GB. \nI do not have so much free space on the laptop I work on. \nWanted to ask if there is an option to download part of the data? \n\nI'm not sure if related, there is a link to Kaggle  API  installation and documentation, However I got lost there with unfamiliar terminology, and was sure if it's related to data downloading. \nKindly your advice or guidance \nTal ",
      "votes": 2,
      "replies": [
        {
          "id": 2087845,
          "postDate": "2023-01-05T20:59:59.337Z",
          "rawMarkdown": "",
          "votes": 3,
          "isDeleted": true,
          "replies": [
            {
              "id": 2088868,
              "postDate": "2023-01-06T17:59:45.500Z",
              "content": "<p>Hi Mohamed, <br>\nThank you for the detailed answer, much appeiciated!</p>\n<p>I followed your steps and when I run the first block of your code in my Jupyter notebook  I receive this error: <br>\n<code>FileNotFoundError: [Errno 2] No such file or directory: 'train.csv'</code></p>\n<p>Do I need to download this file 'train.csv' from somewhere?<br>\n(because the folder where I run my Jupyter notebook - \"C:\\Users.kaggle\\\" where kaggle.json is) </p>\n<p>kindly your advice or guidance<br>\nTal</p>",
              "rawMarkdown": "Hi Mohamed, \nThank you for the detailed answer, much appeiciated!\n\nI followed your steps and when I run the first block of your code in my Jupyter notebook  I receive this error: \n`FileNotFoundError: [Errno 2] No such file or directory: 'train.csv'`\n\nDo I need to download this file 'train.csv' from somewhere?\n(because the folder where I run my Jupyter notebook - \"C:\\Users\\.kaggle\\\" where kaggle.json is) \n\nkindly your advice or guidance\nTal"
            },
            {
              "id": 2088985,
              "postDate": "2023-01-06T19:45:31.877Z",
              "rawMarkdown": "",
              "isDeleted": true
            },
            {
              "id": 2093654,
              "postDate": "2023-01-10T07:34:42.923Z",
              "content": "<p>thank you very much its much clearer now :) </p>",
              "rawMarkdown": "thank you very much its much clearer now :) "
            },
            {
              "id": 2116413,
              "postDate": "2023-01-26T13:24:51.863Z",
              "content": "<p>Hi Mohamed, <br>\nI have another question with your permission.<br>\nthe code you wrote in part 3 <br>\n<code>(3) Iterate over the subset and **download every** image separately. My code:</code></p>\n<p>however the first line in the code <strong>(which I rum from the jupyter note book, is this correct?)</strong>  <br>\n<code>downloaded_images = [filename for filename in os.listdir() if '.dcm' in filename]</code> <br>\nseems to assume I already have .dcm files in my folder (which I do not). </p>\n<p>Can you please elaborate - how do I get .dcm files in my folder? <br>\nor is there something else that I should do? <br>\nThank you in advance </p>",
              "rawMarkdown": "Hi Mohamed, \nI have another question with your permission.\nthe code you wrote in part 3 \n`(3) Iterate over the subset and **download every** image separately. My code:`\n\nhowever the first line in the code **(which I rum from the jupyter note book, is this correct?)**  \n`downloaded_images = [filename for filename in os.listdir() if '.dcm' in filename]` \nseems to assume I already have .dcm files in my folder (which I do not). \n\nCan you please elaborate - how do I get .dcm files in my folder? \nor is there something else that I should do? \nThank you in advance \n"
            },
            {
              "id": 2116776,
              "postDate": "2023-01-26T18:03:56.260Z",
              "rawMarkdown": "",
              "isDeleted": true
            },
            {
              "id": 2131775,
              "postDate": "2023-02-06T10:54:59.193Z",
              "content": "<p>thank you Mohamed 🙏</p>",
              "rawMarkdown": "thank you Mohamed 🙏"
            }
          ]
        }
      ]
    },
    {
      "id": 2534655,
      "postDate": "2023-11-22T19:20:05.307Z",
      "content": "<p>hello everyone I want to work on this competition though its already over but  data size if 372 gb i  don't have much of space on my laptop is there any option to download part of the data?</p>",
      "rawMarkdown": "hello everyone I want to work on this competition though its already over but  data size if 372 gb i  don't have much of space on my laptop is there any option to download part of the data?\n"
    },
    {
      "id": 2165020,
      "postDate": "2023-03-01T23:48:47.210Z",
      "content": "<p>Any idea on how to incorporate Kaggle newly launched models in while training your model? <br>\nPreferably using fast ai. Thank you.</p>",
      "rawMarkdown": "Any idea on how to incorporate Kaggle newly launched models in while training your model? \nPreferably using fast ai. Thank you."
    },
    {
      "id": 2132209,
      "postDate": "2023-02-06T16:58:26.100Z",
      "content": "<p>Hi, <br>\nMy question is about the submission and how the evaluation is done.</p>\n<p>I'm quit lost in the workflow. </p>\n<p>The test data has only 4  files. </p>\n<p>So I do not understand where and at what stage all other 8000 patients (from the hidden test set) being added to the model for evaluation?<br>\nHow do I add all other 8000 patients with their images to the submission.csv file?</p>\n<p>(I know I'm missing something not sure what)</p>\n<p>kindly your help \\ guidance </p>",
      "rawMarkdown": "Hi, \nMy question is about the submission and how the evaluation is done.\n\nI'm quit lost in the workflow. \n\nThe test data has only 4  files. \n\nSo I do not understand where and at what stage all other 8000 patients (from the hidden test set) being added to the model for evaluation?\nHow do I add all other 8000 patients with their images to the submission.csv file?\n\n(I know I'm missing something not sure what)\n\nkindly your help \\ guidance \n \n",
      "replies": [
        {
          "id": 2132226,
          "postDate": "2023-02-06T17:12:55.417Z",
          "content": "<p>found the answer here <br>\n<a href=\"https://www.kaggle.com/competitions/rsna-breast-cancer-detection/discussion/382039\" target=\"_blank\">https://www.kaggle.com/competitions/rsna-breast-cancer-detection/discussion/382039</a></p>",
          "rawMarkdown": "found the answer here \nhttps://www.kaggle.com/competitions/rsna-breast-cancer-detection/discussion/382039"
        }
      ]
    },
    {
      "id": 2111117,
      "postDate": "2023-01-22T16:09:04.417Z",
      "content": "<p>Hi there, </p>\n<p>This is a newbie kaggle question.</p>\n<p>I'm running into issues due to not having internet access when making a competition submission.</p>\n<p>To this end, I'm looking for some tips on how to load a pretrained resnet50 model in the form of a .pth file.</p>\n<p>I've copied it over to the appropriate checkpoint folder from /kaggle/input using the code below.</p>\n<p>if not os.path.exists('/root/.cache/torch/hub/checkpoints/'):<br>\n        os.makedirs('/root/.cache/torch/hub/checkpoints/')<br>\n!cp '/kaggle/input/resnet50-with-pytorch/resnet50.pth' '/root/.cache/torch/hub/checkpoints/resnet50.pth'</p>\n<p>I've also defined a model class, however,  I get the following error when I try to instantiate my model class. </p>\n<p>NameError: name 'resnet50' is not defined</p>\n<p>Any help or pointers would be greatly appreciated. </p>\n<p>Thank you,<br>\nBernie</p>",
      "rawMarkdown": "Hi there, \n\nThis is a newbie kaggle question.\n\nI'm running into issues due to not having internet access when making a competition submission.\n\nTo this end, I'm looking for some tips on how to load a pretrained resnet50 model in the form of a .pth file.\n\nI've copied it over to the appropriate checkpoint folder from /kaggle/input using the code below.\n\nif not os.path.exists('/root/.cache/torch/hub/checkpoints/'):\n        os.makedirs('/root/.cache/torch/hub/checkpoints/')\n!cp '/kaggle/input/resnet50-with-pytorch/resnet50.pth' '/root/.cache/torch/hub/checkpoints/resnet50.pth'\n\nI've also defined a model class, however,  I get the following error when I try to instantiate my model class. \n\nNameError: name 'resnet50' is not defined\n\nAny help or pointers would be greatly appreciated. \n\nThank you,\nBernie",
      "replies": [
        {
          "id": 2111167,
          "postDate": "2023-01-22T16:42:51.547Z",
          "content": "<p>could you post the first lines of your model class and the way you try to load your weights?</p>",
          "rawMarkdown": "could you post the first lines of your model class and the way you try to load your weights?\n"
        }
      ]
    },
    {
      "id": 2087605,
      "postDate": "2023-01-05T17:53:54.597Z",
      "content": "<p>Hi All. I've used kaggle before, but years ago, and it seems that the set-up has changed. Please could you help asnwer my list of possibly stupid questions:</p>\n<ol>\n<li><p>How do we use open source software available on pip/conda? I've seen some notebooks installing packages from pip, when I try to install something I get : [Errno -3] Temporary failure in name resolution')', is there something else I am supposed to do? - i think I found a way to do this, using something called a wheel file, will update if it works</p></li>\n<li><p>Can we train a model, save it, and then upload the model and use that to predict on the test data or do we need to submit a notebook that trains the model and then predicts on the test data as the submission?</p></li>\n<li><p>I saw some mention of a time limit in the discussions, but there's nothing in the rules about this. Is there a time limit for training models or predicting on an image? </p></li>\n</ol>\n<p>Thanks in advance for any help here.</p>",
      "rawMarkdown": "Hi All. I've used kaggle before, but years ago, and it seems that the set-up has changed. Please could you help asnwer my list of possibly stupid questions:\n\n1. How do we use open source software available on pip/conda? I've seen some notebooks installing packages from pip, when I try to install something I get : [Errno -3] Temporary failure in name resolution')', is there something else I am supposed to do? - i think I found a way to do this, using something called a wheel file, will update if it works\n\n2. Can we train a model, save it, and then upload the model and use that to predict on the test data or do we need to submit a notebook that trains the model and then predicts on the test data as the submission?\n\n3. I saw some mention of a time limit in the discussions, but there's nothing in the rules about this. Is there a time limit for training models or predicting on an image? \n\nThanks in advance for any help here.",
      "replies": [
        {
          "id": 2087638,
          "postDate": "2023-01-05T18:21:36.670Z",
          "content": "<p><a href=\"https://www.kaggle.com/competitions/rsna-breast-cancer-detection/overview/code-requirements\" target=\"_blank\">https://www.kaggle.com/competitions/rsna-breast-cancer-detection/overview/code-requirements</a></p>",
          "rawMarkdown": "https://www.kaggle.com/competitions/rsna-breast-cancer-detection/overview/code-requirements",
          "replies": [
            {
              "id": 2088660,
              "postDate": "2023-01-06T14:26:26.967Z",
              "content": "<p>Ah, thanks, I had missed this somehow :)</p>",
              "rawMarkdown": "Ah, thanks, I had missed this somehow :)"
            }
          ]
        },
        {
          "id": 2087642,
          "postDate": "2023-01-05T18:23:27.833Z",
          "content": "<p>2  - train anywhere you want - upload model and predict - this is very common in kaggle challenges.  But often a pain to get it to work :)</p>",
          "rawMarkdown": "2  - train anywhere you want - upload model and predict - this is very common in kaggle challenges.  But often a pain to get it to work :)",
          "votes": 1,
          "replies": [
            {
              "id": 2088659,
              "postDate": "2023-01-06T14:26:12.180Z",
              "content": "<p>thanks for all your answers PC Jimmy, that's very helpful! :)</p>",
              "rawMarkdown": "thanks for all your answers PC Jimmy, that's very helpful! :)"
            }
          ]
        },
        {
          "id": 2087646,
          "postDate": "2023-01-05T18:26:51Z",
          "content": "<p>1 - to install packages using pip - the package needs to be in a data set that you have added.  (the wheel file)   It's a bit tricky to do this the first time so start simple with a single wheel in a data set you created.  </p>",
          "rawMarkdown": "1 - to install packages using pip - the package needs to be in a data set that you have added.  (the wheel file)   It's a bit tricky to do this the first time so start simple with a single wheel in a data set you created.  "
        }
      ]
    },
    {
      "id": 2072764,
      "postDate": "2022-12-22T12:21:25.100Z",
      "content": "<p>Hi, beginner question here. </p>\n<p>I see others publishing datasets of converted PNGs in this competition. Is it ok to use these datasets in submissions? Does it invalidate the entry? </p>\n<p>I think the alternative would be to include all image processing from DICOM to PNG in the submission code itself.</p>\n<p>Thanks.</p>",
      "rawMarkdown": "Hi, beginner question here. \n\nI see others publishing datasets of converted PNGs in this competition. Is it ok to use these datasets in submissions? Does it invalidate the entry? \n\nI think the alternative would be to include all image processing from DICOM to PNG in the submission code itself.\n\nThanks.",
      "replies": [
        {
          "id": 2087652,
          "postDate": "2023-01-05T18:28:07.773Z",
          "content": "<p>converted PNGs will NOT work in submission - because we do not have access to the test dicom files before hand to create them.  In the submission you will need to read and process the dicom files as part of the script.</p>",
          "rawMarkdown": "converted PNGs will NOT work in submission - because we do not have access to the test dicom files before hand to create them.  In the submission you will need to read and process the dicom files as part of the script."
        }
      ]
    },
    {
      "id": 2062738,
      "postDate": "2022-12-12T10:58:42.567Z",
      "content": "<p>Dear all<br>\nQuite new in Kaggle and Deep learning.<br>\nI have a question: what is the aim of this competition ?<br>\nIs it to detect a cancer from imagery or is it do diagnose (False Postive or False Negative or True Positve or True Negative) ?<br>\nAs it is mentioned in the headline \"It is important to identify cases of cancer for obvious reasons, but false positives also have downsides for patients.\"</p>",
      "rawMarkdown": "Dear all\nQuite new in Kaggle and Deep learning.\nI have a question: what is the aim of this competition ?\nIs it to detect a cancer from imagery or is it do diagnose (False Postive or False Negative or True Positve or True Negative) ?\nAs it is mentioned in the headline \"It is important to identify cases of cancer for obvious reasons, but false positives also have downsides for patients.\"",
      "replies": [
        {
          "id": 2074718,
          "postDate": "2022-12-24T14:38:08.693Z",
          "rawMarkdown": "",
          "votes": 3,
          "isDeleted": true
        }
      ]
    },
    {
      "id": 2157807,
      "postDate": "2023-02-24T10:47:38.163Z",
      "content": "<p>Thanks for the tips!</p>",
      "rawMarkdown": "Thanks for the tips!"
    }
  ],
  "comments": [
    {
      "id": 2085792,
      "author_name": "Tal Wac",
      "author_url": "",
      "post_date": "2023-01-04T11:57:52.937000",
      "content": "<p>Hi all :) <br>\nI want to ask a question about downloading data.<br>\nWhen clicking  **Data ** under the ribbon that include :<br>\n<em>Overview Data Code Discussion Leaderboard Rules Team</em>  there is option/ tab  that called <br>\n** Download All**  which include 324.72 GB. <br>\nI do not have so much free space on the laptop I work on. <br>\nWanted to ask if there is an option to download part of the data? </p>\n<p>I'm not sure if related, there is a link to Kaggle  API  installation and documentation, However I got lost there with unfamiliar terminology, and was sure if it's related to data downloading. <br>\nKindly your advice or guidance <br>\nTal </p>",
      "votes": 2,
      "replies": [
        {
          "id": 2087845,
          "author_name": "",
          "author_url": "",
          "post_date": "2023-01-05T20:59:59.337000",
          "content": "",
          "votes": 3,
          "replies": [
            {
              "id": 2088868,
              "author_name": "Tal Wac",
              "author_url": "",
              "post_date": "2023-01-06T17:59:45.500000",
              "content": "<p>Hi Mohamed, <br>\nThank you for the detailed answer, much appeiciated!</p>\n<p>I followed your steps and when I run the first block of your code in my Jupyter notebook  I receive this error: <br>\n<code>FileNotFoundError: [Errno 2] No such file or directory: 'train.csv'</code></p>\n<p>Do I need to download this file 'train.csv' from somewhere?<br>\n(because the folder where I run my Jupyter notebook - \"C:\\Users.kaggle\\\" where kaggle.json is) </p>\n<p>kindly your advice or guidance<br>\nTal</p>",
              "votes": 0,
              "replies": []
            },
            {
              "id": 2088985,
              "author_name": "",
              "author_url": "",
              "post_date": "2023-01-06T19:45:31.877000",
              "content": "",
              "votes": 0,
              "replies": []
            },
            {
              "id": 2093654,
              "author_name": "Tal Wac",
              "author_url": "",
              "post_date": "2023-01-10T07:34:42.923000",
              "content": "<p>thank you very much its much clearer now :) </p>",
              "votes": 0,
              "replies": []
            },
            {
              "id": 2116413,
              "author_name": "Tal Wac",
              "author_url": "",
              "post_date": "2023-01-26T13:24:51.863000",
              "content": "<p>Hi Mohamed, <br>\nI have another question with your permission.<br>\nthe code you wrote in part 3 <br>\n<code>(3) Iterate over the subset and **download every** image separately. My code:</code></p>\n<p>however the first line in the code <strong>(which I rum from the jupyter note book, is this correct?)</strong>  <br>\n<code>downloaded_images = [filename for filename in os.listdir() if '.dcm' in filename]</code> <br>\nseems to assume I already have .dcm files in my folder (which I do not). </p>\n<p>Can you please elaborate - how do I get .dcm files in my folder? <br>\nor is there something else that I should do? <br>\nThank you in advance </p>",
              "votes": 0,
              "replies": []
            },
            {
              "id": 2116776,
              "author_name": "",
              "author_url": "",
              "post_date": "2023-01-26T18:03:56.260000",
              "content": "",
              "votes": 0,
              "replies": []
            },
            {
              "id": 2131775,
              "author_name": "Tal Wac",
              "author_url": "",
              "post_date": "2023-02-06T10:54:59.193000",
              "content": "<p>thank you Mohamed 🙏</p>",
              "votes": 0,
              "replies": []
            }
          ]
        }
      ]
    },
    {
      "id": 2534655,
      "author_name": "KUNAL SOOD",
      "author_url": "",
      "post_date": "2023-11-22T19:20:05.307000",
      "content": "<p>hello everyone I want to work on this competition though its already over but  data size if 372 gb i  don't have much of space on my laptop is there any option to download part of the data?</p>",
      "votes": 0,
      "replies": []
    },
    {
      "id": 2165020,
      "author_name": "Ifeanyichukwu Nwobodo",
      "author_url": "",
      "post_date": "2023-03-01T23:48:47.210000",
      "content": "<p>Any idea on how to incorporate Kaggle newly launched models in while training your model? <br>\nPreferably using fast ai. Thank you.</p>",
      "votes": 0,
      "replies": []
    },
    {
      "id": 2132209,
      "author_name": "Tal Wac",
      "author_url": "",
      "post_date": "2023-02-06T16:58:26.100000",
      "content": "<p>Hi, <br>\nMy question is about the submission and how the evaluation is done.</p>\n<p>I'm quit lost in the workflow. </p>\n<p>The test data has only 4  files. </p>\n<p>So I do not understand where and at what stage all other 8000 patients (from the hidden test set) being added to the model for evaluation?<br>\nHow do I add all other 8000 patients with their images to the submission.csv file?</p>\n<p>(I know I'm missing something not sure what)</p>\n<p>kindly your help \\ guidance </p>",
      "votes": 0,
      "replies": [
        {
          "id": 2132226,
          "author_name": "Tal Wac",
          "author_url": "",
          "post_date": "2023-02-06T17:12:55.417000",
          "content": "<p>found the answer here <br>\n<a href=\"https://www.kaggle.com/competitions/rsna-breast-cancer-detection/discussion/382039\" target=\"_blank\">https://www.kaggle.com/competitions/rsna-breast-cancer-detection/discussion/382039</a></p>",
          "votes": 0,
          "replies": []
        }
      ]
    },
    {
      "id": 2111117,
      "author_name": "bmjdoherty",
      "author_url": "",
      "post_date": "2023-01-22T16:09:04.417000",
      "content": "<p>Hi there, </p>\n<p>This is a newbie kaggle question.</p>\n<p>I'm running into issues due to not having internet access when making a competition submission.</p>\n<p>To this end, I'm looking for some tips on how to load a pretrained resnet50 model in the form of a .pth file.</p>\n<p>I've copied it over to the appropriate checkpoint folder from /kaggle/input using the code below.</p>\n<p>if not os.path.exists('/root/.cache/torch/hub/checkpoints/'):<br>\n        os.makedirs('/root/.cache/torch/hub/checkpoints/')<br>\n!cp '/kaggle/input/resnet50-with-pytorch/resnet50.pth' '/root/.cache/torch/hub/checkpoints/resnet50.pth'</p>\n<p>I've also defined a model class, however,  I get the following error when I try to instantiate my model class. </p>\n<p>NameError: name 'resnet50' is not defined</p>\n<p>Any help or pointers would be greatly appreciated. </p>\n<p>Thank you,<br>\nBernie</p>",
      "votes": 0,
      "replies": [
        {
          "id": 2111167,
          "author_name": "Eleftherios Fanioudakis",
          "author_url": "",
          "post_date": "2023-01-22T16:42:51.547000",
          "content": "<p>could you post the first lines of your model class and the way you try to load your weights?</p>",
          "votes": 0,
          "replies": []
        }
      ]
    },
    {
      "id": 2087605,
      "author_name": "ellagale",
      "author_url": "",
      "post_date": "2023-01-05T17:53:54.597000",
      "content": "<p>Hi All. I've used kaggle before, but years ago, and it seems that the set-up has changed. Please could you help asnwer my list of possibly stupid questions:</p>\n<ol>\n<li><p>How do we use open source software available on pip/conda? I've seen some notebooks installing packages from pip, when I try to install something I get : [Errno -3] Temporary failure in name resolution')', is there something else I am supposed to do? - i think I found a way to do this, using something called a wheel file, will update if it works</p></li>\n<li><p>Can we train a model, save it, and then upload the model and use that to predict on the test data or do we need to submit a notebook that trains the model and then predicts on the test data as the submission?</p></li>\n<li><p>I saw some mention of a time limit in the discussions, but there's nothing in the rules about this. Is there a time limit for training models or predicting on an image? </p></li>\n</ol>\n<p>Thanks in advance for any help here.</p>",
      "votes": 0,
      "replies": [
        {
          "id": 2087638,
          "author_name": "PC Jimmmy",
          "author_url": "",
          "post_date": "2023-01-05T18:21:36.670000",
          "content": "<p><a href=\"https://www.kaggle.com/competitions/rsna-breast-cancer-detection/overview/code-requirements\" target=\"_blank\">https://www.kaggle.com/competitions/rsna-breast-cancer-detection/overview/code-requirements</a></p>",
          "votes": 0,
          "replies": [
            {
              "id": 2088660,
              "author_name": "ellagale",
              "author_url": "",
              "post_date": "2023-01-06T14:26:26.967000",
              "content": "<p>Ah, thanks, I had missed this somehow :)</p>",
              "votes": 0,
              "replies": []
            }
          ]
        },
        {
          "id": 2087642,
          "author_name": "PC Jimmmy",
          "author_url": "",
          "post_date": "2023-01-05T18:23:27.833000",
          "content": "<p>2  - train anywhere you want - upload model and predict - this is very common in kaggle challenges.  But often a pain to get it to work :)</p>",
          "votes": 1,
          "replies": [
            {
              "id": 2088659,
              "author_name": "ellagale",
              "author_url": "",
              "post_date": "2023-01-06T14:26:12.180000",
              "content": "<p>thanks for all your answers PC Jimmy, that's very helpful! :)</p>",
              "votes": 0,
              "replies": []
            }
          ]
        },
        {
          "id": 2087646,
          "author_name": "PC Jimmmy",
          "author_url": "",
          "post_date": "2023-01-05T18:26:51",
          "content": "<p>1 - to install packages using pip - the package needs to be in a data set that you have added.  (the wheel file)   It's a bit tricky to do this the first time so start simple with a single wheel in a data set you created.  </p>",
          "votes": 0,
          "replies": []
        }
      ]
    },
    {
      "id": 2072764,
      "author_name": "frnkc",
      "author_url": "",
      "post_date": "2022-12-22T12:21:25.100000",
      "content": "<p>Hi, beginner question here. </p>\n<p>I see others publishing datasets of converted PNGs in this competition. Is it ok to use these datasets in submissions? Does it invalidate the entry? </p>\n<p>I think the alternative would be to include all image processing from DICOM to PNG in the submission code itself.</p>\n<p>Thanks.</p>",
      "votes": 0,
      "replies": [
        {
          "id": 2087652,
          "author_name": "PC Jimmmy",
          "author_url": "",
          "post_date": "2023-01-05T18:28:07.773000",
          "content": "<p>converted PNGs will NOT work in submission - because we do not have access to the test dicom files before hand to create them.  In the submission you will need to read and process the dicom files as part of the script.</p>",
          "votes": 0,
          "replies": []
        }
      ]
    },
    {
      "id": 2062738,
      "author_name": "Thibaut Jacquin",
      "author_url": "",
      "post_date": "2022-12-12T10:58:42.567000",
      "content": "<p>Dear all<br>\nQuite new in Kaggle and Deep learning.<br>\nI have a question: what is the aim of this competition ?<br>\nIs it to detect a cancer from imagery or is it do diagnose (False Postive or False Negative or True Positve or True Negative) ?<br>\nAs it is mentioned in the headline \"It is important to identify cases of cancer for obvious reasons, but false positives also have downsides for patients.\"</p>",
      "votes": 0,
      "replies": [
        {
          "id": 2074718,
          "author_name": "",
          "author_url": "",
          "post_date": "2022-12-24T14:38:08.693000",
          "content": "",
          "votes": 3,
          "replies": []
        }
      ]
    },
    {
      "id": 2157807,
      "author_name": "Drobnitchi Daniel",
      "author_url": "",
      "post_date": "2023-02-24T10:47:38.163000",
      "content": "<p>Thanks for the tips!</p>",
      "votes": 0,
      "replies": []
    }
  ],
  "raw_markdown_by_id": {
    "2047812": "New to machine learning and data science? No question is too basic or too simple. Feel free to start your own thread, or use this thread as a place to post any first-timer clarifying questions for the Kaggle community to help you with!\n\nIf you would consider yourself a beginner but don't know where to get started, let other Kagglers help you take your first steps here!\n\nNew to Kaggle? Take a look at a few videos Dr. Rachael Tatman has put together to learn a bit more about [site etiquette](https://www.youtube.com/watch?v=aIus8si_Et0), or [Kaggle lingo](https://www.youtube.com/watch?v=sEJHyuWKd-s).\n\nRemember: Kaggle is for everyone. Whether you're teaming up or sharing tips in the competition forum, we expect everyone to follow our [Kaggle community guidelines](https://www.kaggle.com/community-guidelines).\n\nA tip on sharing content - Kaggle is a collaborative community, whereby sharing techniques, starter notebooks, and ideas in the discussion forums are highly encouraged throughout the competition. However, as the competition draws closer to the final deadline it's customary to keep high-scoring notebooks withheld until after the competition has concluded. This maintains the spirit of the competition, while also allowing individuals to submit their own creative work without jeopardy of a higher-scoring notebook being available for an automatic higher rank (through copy/submit). We disable publishing of public notebooks within the final week of the competition, but encourage you to use your best judgment prior to that the deadline.\n\n  \n\nHappy Modeling!\n",
    "2085792": "Hi all :) \nI want to ask a question about downloading data.\nWhen clicking  **Data ** under the ribbon that include :\n*Overview Data Code Discussion Leaderboard Rules Team*  there is option/ tab  that called \n** Download All**  which include 324.72 GB. \nI do not have so much free space on the laptop I work on. \nWanted to ask if there is an option to download part of the data? \n\nI'm not sure if related, there is a link to Kaggle  API  installation and documentation, However I got lost there with unfamiliar terminology, and was sure if it's related to data downloading. \nKindly your advice or guidance \nTal ",
    "2534655": "hello everyone I want to work on this competition though its already over but  data size if 372 gb i  don't have much of space on my laptop is there any option to download part of the data?\n",
    "2165020": "Any idea on how to incorporate Kaggle newly launched models in while training your model? \nPreferably using fast ai. Thank you.",
    "2132209": "Hi, \nMy question is about the submission and how the evaluation is done.\n\nI'm quit lost in the workflow. \n\nThe test data has only 4  files. \n\nSo I do not understand where and at what stage all other 8000 patients (from the hidden test set) being added to the model for evaluation?\nHow do I add all other 8000 patients with their images to the submission.csv file?\n\n(I know I'm missing something not sure what)\n\nkindly your help \\ guidance \n \n",
    "2111117": "Hi there, \n\nThis is a newbie kaggle question.\n\nI'm running into issues due to not having internet access when making a competition submission.\n\nTo this end, I'm looking for some tips on how to load a pretrained resnet50 model in the form of a .pth file.\n\nI've copied it over to the appropriate checkpoint folder from /kaggle/input using the code below.\n\nif not os.path.exists('/root/.cache/torch/hub/checkpoints/'):\n        os.makedirs('/root/.cache/torch/hub/checkpoints/')\n!cp '/kaggle/input/resnet50-with-pytorch/resnet50.pth' '/root/.cache/torch/hub/checkpoints/resnet50.pth'\n\nI've also defined a model class, however,  I get the following error when I try to instantiate my model class. \n\nNameError: name 'resnet50' is not defined\n\nAny help or pointers would be greatly appreciated. \n\nThank you,\nBernie",
    "2087605": "Hi All. I've used kaggle before, but years ago, and it seems that the set-up has changed. Please could you help asnwer my list of possibly stupid questions:\n\n1. How do we use open source software available on pip/conda? I've seen some notebooks installing packages from pip, when I try to install something I get : [Errno -3] Temporary failure in name resolution')', is there something else I am supposed to do? - i think I found a way to do this, using something called a wheel file, will update if it works\n\n2. Can we train a model, save it, and then upload the model and use that to predict on the test data or do we need to submit a notebook that trains the model and then predicts on the test data as the submission?\n\n3. I saw some mention of a time limit in the discussions, but there's nothing in the rules about this. Is there a time limit for training models or predicting on an image? \n\nThanks in advance for any help here.",
    "2072764": "Hi, beginner question here. \n\nI see others publishing datasets of converted PNGs in this competition. Is it ok to use these datasets in submissions? Does it invalidate the entry? \n\nI think the alternative would be to include all image processing from DICOM to PNG in the submission code itself.\n\nThanks.",
    "2062738": "Dear all\nQuite new in Kaggle and Deep learning.\nI have a question: what is the aim of this competition ?\nIs it to detect a cancer from imagery or is it do diagnose (False Postive or False Negative or True Positve or True Negative) ?\nAs it is mentioned in the headline \"It is important to identify cases of cancer for obvious reasons, but false positives also have downsides for patients.\"",
    "2157807": "Thanks for the tips!"
  }
}