{
  "id": 115778,
  "title": "stage-2 duplicates",
  "url": "/competitions/rsna-intracranial-hemorrhage-detection/discussion/115778",
  "author_name": "Mobassir",
  "post_date": "2019-11-05T06:39:13.516000",
  "votes": 9,
  "comment_count": 4,
  "views": 0,
  "content": "<p>duplicates_to_remove = [\n        56346,56347,56348,56349,\n        56350,56351,1171830,1171831,\n        1171832,1171833,1171834,1171835,\n        3705312,3705313,3705314,3705315,\n        3705316,3705317,3842478,3842479,\n        3842480,3842481,3842482,3842483\n    ]\n    df = df.drop(index = duplicates_to_remove)</p>",
  "messages": [
    {
      "id": 665592,
      "postDate": "2019-11-05T06:39:13.517Z",
      "content": "<p>duplicates_to_remove = [\n        56346,56347,56348,56349,\n        56350,56351,1171830,1171831,\n        1171832,1171833,1171834,1171835,\n        3705312,3705313,3705314,3705315,\n        3705316,3705317,3842478,3842479,\n        3842480,3842481,3842482,3842483\n    ]\n    df = df.drop(index = duplicates_to_remove)</p>",
      "rawMarkdown": "duplicates_to_remove = [\n        56346,56347,56348,56349,\n        56350,56351,1171830,1171831,\n        1171832,1171833,1171834,1171835,\n        3705312,3705313,3705314,3705315,\n        3705316,3705317,3842478,3842479,\n        3842480,3842481,3842482,3842483\n    ]\n    df = df.drop(index = duplicates_to_remove)",
      "votes": 9
    },
    {
      "id": 666014,
      "postDate": "2019-11-05T16:33:44.997Z",
      "content": "<p>But how do you use this code?\nWe cannot modify any lines of code in stage2.</p>",
      "rawMarkdown": "But how do you use this code?\nWe cannot modify any lines of code in stage2.",
      "votes": 1,
      "replies": [
        {
          "id": 666065,
          "postDate": "2019-11-05T17:28:36.990Z",
          "content": "<p>you can do it mate,because in stage-1 we also removed those same duplicate pictures,as stage 2 dataset becomes huge after merging train and test set of stage 1 so those duplicate images indices have changed,so you can definitely use this code in stage-2</p>",
          "rawMarkdown": "you can do it mate,because in stage-1 we also removed those same duplicate pictures,as stage 2 dataset becomes huge after merging train and test set of stage 1 so those duplicate images indices have changed,so you can definitely use this code in stage-2"
        }
      ]
    },
    {
      "id": 665917,
      "postDate": "2019-11-05T14:29:55.380Z",
      "content": "<p><a href=\"/mobassir\">@mobassir</a> - You have started with the preparation for stage 2 :). Awesome mate. Thanks for sharing the list of duplicate images list. </p>",
      "rawMarkdown": "@mobassir - You have started with the preparation for stage 2 :). Awesome mate. Thanks for sharing the list of duplicate images list. ",
      "votes": 2,
      "replies": [
        {
          "id": 665922,
          "postDate": "2019-11-05T14:33:41.883Z",
          "content": "<p>yeah,the model was training when i was chatting with you,it is still training ha ha ha :D</p>",
          "rawMarkdown": "yeah,the model was training when i was chatting with you,it is still training ha ha ha :D",
          "votes": 1
        }
      ]
    }
  ],
  "comments": [
    {
      "id": 666014,
      "author_name": "shirakia",
      "author_url": "",
      "post_date": "2019-11-05T16:33:44.997000",
      "content": "<p>But how do you use this code?\nWe cannot modify any lines of code in stage2.</p>",
      "votes": 1,
      "replies": [
        {
          "id": 666065,
          "author_name": "Mobassir",
          "author_url": "",
          "post_date": "2019-11-05T17:28:36.990000",
          "content": "<p>you can do it mate,because in stage-1 we also removed those same duplicate pictures,as stage 2 dataset becomes huge after merging train and test set of stage 1 so those duplicate images indices have changed,so you can definitely use this code in stage-2</p>",
          "votes": 0,
          "replies": []
        }
      ]
    },
    {
      "id": 665917,
      "author_name": "Manoj Prabhakar",
      "author_url": "",
      "post_date": "2019-11-05T14:29:55.380000",
      "content": "<p><a href=\"/mobassir\">@mobassir</a> - You have started with the preparation for stage 2 :). Awesome mate. Thanks for sharing the list of duplicate images list. </p>",
      "votes": 2,
      "replies": [
        {
          "id": 665922,
          "author_name": "Mobassir",
          "author_url": "",
          "post_date": "2019-11-05T14:33:41.883000",
          "content": "<p>yeah,the model was training when i was chatting with you,it is still training ha ha ha :D</p>",
          "votes": 1,
          "replies": []
        }
      ]
    }
  ],
  "raw_markdown_by_id": {
    "665592": "duplicates_to_remove = [\n        56346,56347,56348,56349,\n        56350,56351,1171830,1171831,\n        1171832,1171833,1171834,1171835,\n        3705312,3705313,3705314,3705315,\n        3705316,3705317,3842478,3842479,\n        3842480,3842481,3842482,3842483\n    ]\n    df = df.drop(index = duplicates_to_remove)",
    "666014": "But how do you use this code?\nWe cannot modify any lines of code in stage2.",
    "665917": "@mobassir - You have started with the preparation for stage 2 :). Awesome mate. Thanks for sharing the list of duplicate images list. "
  }
}