Row 88820
Content Data
This page contains data entry 88820 from the Axioma AXP content repository. The structured data below represents the complete record for this entry.
I don't really think that's the problem. I use the internet to research topics but I won't tell people to put glue on pizza. That's because I at least have some understanding of how to discern good and bad data sources and can reason through the data I read.
Current LLMs do not really have the capability to do that. It's not that the source data is bad, it's that they don't have the ability to synthesize new knowledge accurately.
For example, [Ars Technica's article](https://arstechnica.com/information-technology/2024/05/googles-ai-overview-can-give-false-misleading-and-dangerous-answers/) gives an example of asking "southernmost point in mainland Alaska" leading to an innocuously wrong answer involving "Amatignak Island" (which is not part of mainland Alaska). Inaccurate answers like this isn't due to bad input data. It's due to the Artificial Intelligence having no actual intelligence at all and can't understand proper semantics.
| Field | Value |
|---|---|
| text | I don't really think that's the problem. I use the internet to research topics but I won't tell people to put glue on pizza. That's because I at least have some understanding of how to discern good and bad data sources and can reason through the data I read. Current LLMs do not really have the capability to do that. It's not that the source data is bad, it's that they don't have the ability to synthesize new knowledge accurately. For example, [Ars Technica's article](https://arstechnica.com/in… |
| label | r/technology |
| dataType | comment |
| communityName | r/technology |
| datetime | 2024-05-25 |
| username_encoded | Z0FBQUFBQm5Lak1xVFFWenlSZktaTXRtSnNBWUpVbXBYZGEtcUZsMHZ3RDF4aW85RXZ3bVM2bWpKWXYybk1wVXBGdElubm4zeG1FZWN6OWpETXJEd0NGVEtJdmtSQThrSXc9PQ== |
| url_encoded | Z0FBQUFBQm5Lak83OFJ1Vnd2VklBRkEyeUpmeHFPbjRXVHg1RElkNVd6WEtabGZ5RXpJTHA2ejFjVXJTWlduNV9JLTZER2I3bEVUMWhIUlZuU2ljczN6T0dwZm9hRS1Od3o3aHR3YWU4WEJuLXNPeVlBeWJTXy1VYjFiSVV3d3ZydGY4YU5KOWJGM19xLWVDRG9FOHNLWklpLVZENnpJSllKM254Z3BtZUxLdjJhZm5TRHpVR0x0cFY2NWt0VEdGVEY5c0pLS0ZQODJ3TmhFM1paNFk1Q2VocUVLX3dONVRQZz09 |
Raw Record
{
"text": "I don't really think that's the problem. I use the internet to research topics but I won't tell people to put glue on pizza. That's because I at least have some understanding of how to discern good and bad data sources and can reason through the data I read.\n\nCurrent LLMs do not really have the capability to do that. It's not that the source data is bad, it's that they don't have the ability to synthesize new knowledge accurately.\n\nFor example, [Ars Technica's article](https://arstechnica.com/information-technology/2024/05/googles-ai-overview-can-give-false-misleading-and-dangerous-answers/) gives an example of asking \"southernmost point in mainland Alaska\" leading to an innocuously wrong answer involving \"Amatignak Island\" (which is not part of mainland Alaska). Inaccurate answers like this isn't due to bad input data. It's due to the Artificial Intelligence having no actual intelligence at all and can't understand proper semantics.",
"label": "r/technology",
"dataType": "comment",
"communityName": "r/technology",
"datetime": "2024-05-25",
"username_encoded": "Z0FBQUFBQm5Lak1xVFFWenlSZktaTXRtSnNBWUpVbXBYZGEtcUZsMHZ3RDF4aW85RXZ3bVM2bWpKWXYybk1wVXBGdElubm4zeG1FZWN6OWpETXJEd0NGVEtJdmtSQThrSXc9PQ==",
"url_encoded": "Z0FBQUFBQm5Lak83OFJ1Vnd2VklBRkEyeUpmeHFPbjRXVHg1RElkNVd6WEtabGZ5RXpJTHA2ejFjVXJTWlduNV9JLTZER2I3bEVUMWhIUlZuU2ljczN6T0dwZm9hRS1Od3o3aHR3YWU4WEJuLXNPeVlBeWJTXy1VYjFiSVV3d3ZydGY4YU5KOWJGM19xLWVDRG9FOHNLWklpLVZENnpJSllKM254Z3BtZUxLdjJhZm5TRHpVR0x0cFY2NWt0VEdGVEY5c0pLS0ZQODJ3TmhFM1paNFk1Q2VocUVLX3dONVRQZz09"
}
Entry Information
- Entry ID: 88820
- Repository: Axioma AXP
- Dataset: arrmlet/reddit_dataset_36
- Total Entries: 100,000