Row 46422

Row ID: 46422 | Dataset Entry | Axioma AXP Content Repository

Content Data

This page contains data entry 46422 from the Axioma AXP content repository. The structured data below represents the complete record for this entry.

As the models near “perfect”, it’s going to be much harder to feel the differences between generations just by having it perform casual tasks or conversations. You’re going to need to run much more specific & focused tasks in order to notice any meaningful differences, like with modern computing benchmarks.

Right now, we’re still nowhere near “perfect”, so the differences are still very noticeable. Although it might be hard to tell a difference between GPT-4 and 3.5 based on conversation alone, it’s very noticeable when it comes to any sort of problem solving.

Eventually, the only way to tell a difference would probably be to ask ridiculously complex questions that no average user would ever ask. The focus would probably shift to power/cost efficiency long before this point though.

FieldValue
text As the models near “perfect”, it’s going to be much harder to feel the differences between generations just by having it perform casual tasks or conversations. You’re going to need to run much more specific & focused tasks in order to notice any meaningful differences, like with modern computing benchmarks. Right now, we’re still nowhere near “perfect”, so the differences are still very noticeable. Although it might be hard to tell a difference between GPT-4 and 3.5 based on conversation alone,…
label r/openai
dataType comment
communityName r/OpenAI
datetime 2024-05-22
username_encoded Z0FBQUFBQm5Lak1RWVhpdmN1bWRzUGVrMDZWbzZSY2FWektIWnhYUmJOd1V6UGg5VGNqV0JQUWxLRHE4Mnc0T1VxSTZnTVdsUmI1anEyZmpJZTBBZzBqUUdxRVNGU3p4Z3dXQmFjZ2YzQ2lZd2FEaXZiQk1KR3c9
url_encoded Z0FBQUFBQm5Lak9mWmg4NXVSVF9Na1lYNGlNdkhxUVBTMklUSnoyVjZVTllLR3l6b3hsbE5Nc3JtXzFtMXk5VWtSX0dWbzlaZFIydWFaQ0hVTTRfcnhoLWZ2M0JZOWNhc0t6UWdBWk96em9yWjBySEZwTzRnU3FkREticzcxeXJQTUtyLXp3NzVGWHpjTUhVTHNvV1IzcU9NeUIyUWp3dVd6SDJjMmlMVncwSXI3cTJMWDlQWk5OWEZia3dmcFR5OTVJZDZWckVKUlhBNkxhNnVScnN5MnRnclRVeEtWZWNHZz09

Raw Record

{
  "text": "As the models near “perfect”, it’s going to be much harder to feel the differences between generations just by having it perform casual tasks or conversations. You’re going to need to run much more specific & focused tasks in order to notice any meaningful differences, like with modern computing benchmarks.\n\nRight now, we’re still nowhere near “perfect”, so the differences are still very noticeable. Although it might be hard to tell a difference between GPT-4 and 3.5 based on conversation alone, it’s very noticeable when it comes to any sort of problem solving. \n\nEventually, the only way to tell a difference would probably be to ask ridiculously complex questions that no average user would ever ask. The focus would probably shift to power/cost efficiency long before this point though.",
  "label": "r/openai",
  "dataType": "comment",
  "communityName": "r/OpenAI",
  "datetime": "2024-05-22",
  "username_encoded": "Z0FBQUFBQm5Lak1RWVhpdmN1bWRzUGVrMDZWbzZSY2FWektIWnhYUmJOd1V6UGg5VGNqV0JQUWxLRHE4Mnc0T1VxSTZnTVdsUmI1anEyZmpJZTBBZzBqUUdxRVNGU3p4Z3dXQmFjZ2YzQ2lZd2FEaXZiQk1KR3c9",
  "url_encoded": "Z0FBQUFBQm5Lak9mWmg4NXVSVF9Na1lYNGlNdkhxUVBTMklUSnoyVjZVTllLR3l6b3hsbE5Nc3JtXzFtMXk5VWtSX0dWbzlaZFIydWFaQ0hVTTRfcnhoLWZ2M0JZOWNhc0t6UWdBWk96em9yWjBySEZwTzRnU3FkREticzcxeXJQTUtyLXp3NzVGWHpjTUhVTHNvV1IzcU9NeUIyUWp3dVd6SDJjMmlMVncwSXI3cTJMWDlQWk5OWEZia3dmcFR5OTVJZDZWckVKUlhBNkxhNnVScnN5MnRnclRVeEtWZWNHZz09"
}

Entry Information