Row 56372

Row ID: 56372 | Dataset Entry | Axioma AXP Content Repository

Content Data

This page contains data entry 56372 from the Axioma AXP content repository. The structured data below represents the complete record for this entry.

Google’s new Gemini model was tested on expert questions in domains like biology, physics, and chemistry. the questions were designed by experts in the fields after Gemini was trained. Human experts scored 39%, then they were given access to Google search while answering and scored 49%. Gemini without internet access scored 60%. These weren’t memory tests, they required real problem solving like calculating genetic frequencies and enthalpy changes.

There’s a lot of logic going on when processing a prompt through billions of interconnected parameters that allow it to accurately “predict the next word.” Humans also “predict the next word” when answering but that doesn’t mean humans aren’t using logic under the hood.

FieldValue
text Google’s new Gemini model was tested on expert questions in domains like biology, physics, and chemistry. the questions were designed by experts in the fields after Gemini was trained. Human experts scored 39%, then they were given access to Google search while answering and scored 49%. Gemini without internet access scored 60%. These weren’t memory tests, they required real problem solving like calculating genetic frequencies and enthalpy changes. There’s a lot of logic going on when processin…
label r/technology
dataType comment
communityName r/technology
datetime 2024-05-23
username_encoded Z0FBQUFBQm5Lak1XRDJfT1VKMDdVY09tNjlHUVFmN2VqMVNrQkZEUVVvR25kRE04YlQ1VkRYQU83bnN3bDZDYXR6aXg5UzdhNVlmTzNNQi1MajVBLU1HMDkxU2pZMGw0Rmc9PQ==
url_encoded Z0FBQUFBQm5Lak9tbFhlbDFWNWFFMGtnbHZydElMSWxKMFd1Y21ySTRUOUFtRkZmVG1CaVAtN0Y5N3ZGbm9OTnJ6UDV2RUxhb2V0aHVOZDczVDFQVHMyZWR6RE9RMjNPOC1uNUpRRG9mVWxZTGNJdEpsNzgxZVdzRmFZanNFQUNKWUxyeFlxY2FNa3Zyd3FrRDY5WDgzRWdnTTFDSE0wdzJSVWNOb2Z1Z1NZZTNyXzN2VVVReHFENUhUR19SMC1fMkdqWkhNd2EtQmdtRmVUSDJRVWE5cGd1M0hRS24tVlk0Zz09

Raw Record

{
  "text": "Google’s new Gemini model was tested on expert questions in domains like biology, physics, and chemistry. the questions were designed by experts in the fields after Gemini was trained. Human experts scored 39%, then they were given access to Google search while answering and scored 49%. Gemini without internet access scored 60%. These weren’t memory tests, they required real problem solving like calculating genetic frequencies and enthalpy changes.\n\nThere’s a lot of logic going on when processing a prompt through billions of interconnected parameters that allow it to accurately “predict the next word.” Humans also “predict the next word” when answering but that doesn’t mean humans aren’t using logic under the hood.",
  "label": "r/technology",
  "dataType": "comment",
  "communityName": "r/technology",
  "datetime": "2024-05-23",
  "username_encoded": "Z0FBQUFBQm5Lak1XRDJfT1VKMDdVY09tNjlHUVFmN2VqMVNrQkZEUVVvR25kRE04YlQ1VkRYQU83bnN3bDZDYXR6aXg5UzdhNVlmTzNNQi1MajVBLU1HMDkxU2pZMGw0Rmc9PQ==",
  "url_encoded": "Z0FBQUFBQm5Lak9tbFhlbDFWNWFFMGtnbHZydElMSWxKMFd1Y21ySTRUOUFtRkZmVG1CaVAtN0Y5N3ZGbm9OTnJ6UDV2RUxhb2V0aHVOZDczVDFQVHMyZWR6RE9RMjNPOC1uNUpRRG9mVWxZTGNJdEpsNzgxZVdzRmFZanNFQUNKWUxyeFlxY2FNa3Zyd3FrRDY5WDgzRWdnTTFDSE0wdzJSVWNOb2Z1Z1NZZTNyXzN2VVVReHFENUhUR19SMC1fMkdqWkhNd2EtQmdtRmVUSDJRVWE5cGd1M0hRS24tVlk0Zz09"
}

Entry Information