Row 2775

Row ID: 2775 | Dataset Entry | Axioma AXP Content Repository

Content Data

This page contains data entry 2775 from the Axioma AXP content repository. The structured data below represents the complete record for this entry.

Which version are you referring to? In my article, I analyze Gemini Ultra, which indeed isn't public yet, but I'm basing my analysis on the results reported on their website and in the technical report that they released. The technical report gives some insight into how the models (all 3 versions) work (and that's the first part of the article) and how they compare to other models such as GPT-4 (and that's what I base my "controversy" claims on - their own reported results for their best model). I didn't run the benchmark myself on smaller models, I analysed results they obtained internally for their best model and published in technical reports.

FieldValue
text Which version are you referring to? In my article, I analyze Gemini Ultra, which indeed isn't public yet, but I'm basing my analysis on the results reported on their website and in the technical report that they released. The technical report gives some insight into how the models (all 3 versions) work (and that's the first part of the article) and how they compare to other models such as GPT-4 (and that's what I base my "controversy" claims on - their own reported results for their best model).…
label r/deepmind
dataType comment
communityName r/deepmind
datetime 2023-12-08
username_encoded Z0FBQUFBQm5LakwwTW41NVpoeGNjZmwzZHlJSjdIbnF1SlJUSHNDdkJHc0xFYVdjay1ZbzM2amxNUHkxZnNZS1k1VVpBd0c4MXUzcXg5RXFDX1NLNjBLbUFoOEp0VXpVRUE9PQ==
url_encoded Z0FBQUFBQm5Lak9FTmo0Q1dmcGlGdXdlb2tlcHNSVnF3ZVVKa1NZNDgyZzREMU1sZHF6U3hDZngyNW13TEVOQTRDN2hTampZb1c1OWc1WXVrbmt6WjBFRVpGTmxIU0FBSm95ZXQ5cDNnS1B0VHVXeG9walNrRDgtQmZTbUZNODk1S0Z5eGw4YWZQQlUwZkFVSWNBZXJ0MlJFTE9BeVgwb1U1N1VIeDIyeVhJeXlPMDFaZkdxOXdfWFB4RHE1T3lDczcySHZzR3I5WWhIOVZ4VEFVaTVkN0d2Y0s3Z21TbU9PZz09

Raw Record

{
  "text": "Which version are you referring to? In my article, I analyze Gemini Ultra, which indeed isn't public yet, but I'm basing my analysis on the results reported on their website and in the technical report that they released. The technical report gives some insight into how the models (all 3 versions) work (and that's the first part of the article) and how they compare to other models such as GPT-4 (and that's what I base my \"controversy\" claims on - their own reported results for their best model). I didn't run the benchmark myself on smaller models, I analysed results they obtained internally for their best model and published in technical reports.",
  "label": "r/deepmind",
  "dataType": "comment",
  "communityName": "r/deepmind",
  "datetime": "2023-12-08",
  "username_encoded": "Z0FBQUFBQm5LakwwTW41NVpoeGNjZmwzZHlJSjdIbnF1SlJUSHNDdkJHc0xFYVdjay1ZbzM2amxNUHkxZnNZS1k1VVpBd0c4MXUzcXg5RXFDX1NLNjBLbUFoOEp0VXpVRUE9PQ==",
  "url_encoded": "Z0FBQUFBQm5Lak9FTmo0Q1dmcGlGdXdlb2tlcHNSVnF3ZVVKa1NZNDgyZzREMU1sZHF6U3hDZngyNW13TEVOQTRDN2hTampZb1c1OWc1WXVrbmt6WjBFRVpGTmxIU0FBSm95ZXQ5cDNnS1B0VHVXeG9walNrRDgtQmZTbUZNODk1S0Z5eGw4YWZQQlUwZkFVSWNBZXJ0MlJFTE9BeVgwb1U1N1VIeDIyeVhJeXlPMDFaZkdxOXdfWFB4RHE1T3lDczcySHZzR3I5WWhIOVZ4VEFVaTVkN0d2Y0s3Z21TbU9PZz09"
}

Entry Information