{
  "id": 428506,
  "title": "Size of hidden test dataset?",
  "url": "/competitions/google-research-identify-contrails-reduce-global-warming/discussion/428506",
  "author_name": "Phaedrus",
  "post_date": "2023-08-01T16:43:27.644000",
  "votes": 2,
  "comment_count": 3,
  "views": 0,
  "content": "<p>Hi Folks,</p>\n<p>Do we know the number of images/npy files in hidden test set? I searched but couldn't find info about it on the discussion threads. Even a ballpark is fine.</p>\n<p>Thanks, </p>",
  "messages": [
    {
      "id": 2369348,
      "postDate": "2023-08-01T17:23:15.857Z",
      "content": "<p>test/ - the test set; your objective is to identify contrails found in these records. Note: Since this is a Code competition, you do not have access to the actual test set that your notebook is rerun against. The records shown here are copies of the first two records of the validation data (without the labels). <strong>The hidden test set is approximately the same size (± 5%) as the validation set</strong>. IMPORTANT: Submissions should use run-length encoding with empty predictions (e.g., no contrails) should be marked by '-' in the submission. (See this notebook for details.)</p>\n<p>Same size as validation set +- 5% (validation folder)</p>",
      "rawMarkdown": "test/ - the test set; your objective is to identify contrails found in these records. Note: Since this is a Code competition, you do not have access to the actual test set that your notebook is rerun against. The records shown here are copies of the first two records of the validation data (without the labels). **The hidden test set is approximately the same size (± 5%) as the validation set**. IMPORTANT: Submissions should use run-length encoding with empty predictions (e.g., no contrails) should be marked by '-' in the submission. (See this notebook for details.)\n\nSame size as validation set +- 5% (validation folder)",
      "votes": 3,
      "replies": [
        {
          "id": 2369724,
          "postDate": "2023-08-02T01:50:04.827Z",
          "content": "<p>thankyou! </p>",
          "rawMarkdown": "thankyou! "
        },
        {
          "id": 2370284,
          "postDate": "2023-08-02T09:48:41.007Z",
          "content": "<p>In addition to that, from the Leaderboard section:</p>\n<blockquote>\n  <p>This leaderboard is calculated with approximately 15% of the test data. The final results will be based on the other 85%, so the final standings may be different.</p>\n</blockquote>\n<p>So the private hidden test set is around 85% the size of the validation set, which would be <strong>1575 aprox.</strong>, as the validation set contains 1856 samples. The public part of the hidden test set then would be around 290 images, which as usual, will probably cause some LB shake-up.</p>",
          "rawMarkdown": "In addition to that, from the Leaderboard section:\n\n>This leaderboard is calculated with approximately 15% of the test data. The final results will be based on the other 85%, so the final standings may be different.\n\nSo the private hidden test set is around 85% the size of the validation set, which would be **1575 aprox.**, as the validation set contains 1856 samples. The public part of the hidden test set then would be around 290 images, which as usual, will probably cause some LB shake-up."
        }
      ]
    },
    {
      "id": 2369307,
      "postDate": "2023-08-01T16:43:27.643Z",
      "content": "<p>Hi Folks,</p>\n<p>Do we know the number of images/npy files in hidden test set? I searched but couldn't find info about it on the discussion threads. Even a ballpark is fine.</p>\n<p>Thanks, </p>",
      "rawMarkdown": "Hi Folks,\n\nDo we know the number of images/npy files in hidden test set? I searched but couldn't find info about it on the discussion threads. Even a ballpark is fine.\n\nThanks, ",
      "votes": 1
    }
  ],
  "comments": [
    {
      "id": 2369348,
      "author_name": "Martin Kovacevic Buvinic",
      "author_url": "",
      "post_date": "2023-08-01T17:23:15.857000",
      "content": "<p>test/ - the test set; your objective is to identify contrails found in these records. Note: Since this is a Code competition, you do not have access to the actual test set that your notebook is rerun against. The records shown here are copies of the first two records of the validation data (without the labels). <strong>The hidden test set is approximately the same size (± 5%) as the validation set</strong>. IMPORTANT: Submissions should use run-length encoding with empty predictions (e.g., no contrails) should be marked by '-' in the submission. (See this notebook for details.)</p>\n<p>Same size as validation set +- 5% (validation folder)</p>",
      "votes": 3,
      "replies": [
        {
          "id": 2369724,
          "author_name": "Phaedrus",
          "author_url": "",
          "post_date": "2023-08-02T01:50:04.827000",
          "content": "<p>thankyou! </p>",
          "votes": 0,
          "replies": []
        },
        {
          "id": 2370284,
          "author_name": "Enric Domingo",
          "author_url": "",
          "post_date": "2023-08-02T09:48:41.007000",
          "content": "<p>In addition to that, from the Leaderboard section:</p>\n<blockquote>\n  <p>This leaderboard is calculated with approximately 15% of the test data. The final results will be based on the other 85%, so the final standings may be different.</p>\n</blockquote>\n<p>So the private hidden test set is around 85% the size of the validation set, which would be <strong>1575 aprox.</strong>, as the validation set contains 1856 samples. The public part of the hidden test set then would be around 290 images, which as usual, will probably cause some LB shake-up.</p>",
          "votes": 0,
          "replies": []
        }
      ]
    }
  ],
  "raw_markdown_by_id": {
    "2369348": "test/ - the test set; your objective is to identify contrails found in these records. Note: Since this is a Code competition, you do not have access to the actual test set that your notebook is rerun against. The records shown here are copies of the first two records of the validation data (without the labels). **The hidden test set is approximately the same size (± 5%) as the validation set**. IMPORTANT: Submissions should use run-length encoding with empty predictions (e.g., no contrails) should be marked by '-' in the submission. (See this notebook for details.)\n\nSame size as validation set +- 5% (validation folder)",
    "2369307": "Hi Folks,\n\nDo we know the number of images/npy files in hidden test set? I searched but couldn't find info about it on the discussion threads. Even a ballpark is fine.\n\nThanks, "
  }
}