Row 90805
Content Data
This page contains data entry 90805 from the Axioma AXP content repository. The structured data below represents the complete record for this entry.
I think it's also important to remember that the goal for many of these companies is making an AI that sounds and seems human, because that's what consumers want.
So it has to "talk" like a regular person when asked a question, even if the answer is scientific or complicated. If you just train any LLM on scientific journals and encyclopedias, sure it's probably going to get the correct answer, but it's also going to sound like a scientific journal which can be beyond the grasp of most people. Ergo it has to train on regular peoples conversation and comments, but that "taints" its knowledge.
| Field | Value |
|---|---|
| text | I think it's also important to remember that the goal for many of these companies is making an AI that sounds and seems human, because that's what consumers want. So it has to "talk" like a regular person when asked a question, even if the answer is scientific or complicated. If you just train any LLM on scientific journals and encyclopedias, sure it's probably going to get the correct answer, but it's also going to sound like a scientific journal which can be beyond the grasp of most people. E… |
| label | r/technology |
| dataType | comment |
| communityName | r/technology |
| datetime | 2024-05-25 |
| username_encoded | Z0FBQUFBQm5Lak1yRk9qWHh0Z3NtZ0NKbVpCeXdfMk5wYmctekttMnI2MGIyN1VDSDZWVzE1eHZZbTNaOGFFZVI0RGlObmpFVl9DYy1TZW5zbEp3eDlSTktsZVM3VUMwY1E9PQ== |
| url_encoded | Z0FBQUFBQm5Lak85R0MzTzhqT1FKMDJId18tV2NLYzliaVdwb2pHU0htZTRuS1FXS25XeFI2RW1BOXFrLVJfNkNWa1hjZklUVnIyQkVEZ2Q2X3ZmOWt0Rm5PdXk5WmFoNW9yRmltTWlWbWpfTjY5Q2MyZ0R4UzV3dTR6eDA0X24tN1lrSmdSbWdUTHhmNHhhdEFMQ2w3ekpuNzdqVEowdmtkUkJhbDZKZkRSME9ZOVRiNWY3UDJXM0NpMHphckJUZENiV21TbkotblF1b1BFNHpKWkNvd1Z1NWNUc3dGOGJndz09 |
Raw Record
{
"text": "I think it's also important to remember that the goal for many of these companies is making an AI that sounds and seems human, because that's what consumers want.\n\nSo it has to \"talk\" like a regular person when asked a question, even if the answer is scientific or complicated. If you just train any LLM on scientific journals and encyclopedias, sure it's probably going to get the correct answer, but it's also going to sound like a scientific journal which can be beyond the grasp of most people. Ergo it has to train on regular peoples conversation and comments, but that \"taints\" its knowledge.",
"label": "r/technology",
"dataType": "comment",
"communityName": "r/technology",
"datetime": "2024-05-25",
"username_encoded": "Z0FBQUFBQm5Lak1yRk9qWHh0Z3NtZ0NKbVpCeXdfMk5wYmctekttMnI2MGIyN1VDSDZWVzE1eHZZbTNaOGFFZVI0RGlObmpFVl9DYy1TZW5zbEp3eDlSTktsZVM3VUMwY1E9PQ==",
"url_encoded": "Z0FBQUFBQm5Lak85R0MzTzhqT1FKMDJId18tV2NLYzliaVdwb2pHU0htZTRuS1FXS25XeFI2RW1BOXFrLVJfNkNWa1hjZklUVnIyQkVEZ2Q2X3ZmOWt0Rm5PdXk5WmFoNW9yRmltTWlWbWpfTjY5Q2MyZ0R4UzV3dTR6eDA0X24tN1lrSmdSbWdUTHhmNHhhdEFMQ2w3ekpuNzdqVEowdmtkUkJhbDZKZkRSME9ZOVRiNWY3UDJXM0NpMHphckJUZENiV21TbkotblF1b1BFNHpKWkNvd1Z1NWNUc3dGOGJndz09"
}
Entry Information
- Entry ID: 90805
- Repository: Axioma AXP
- Dataset: arrmlet/reddit_dataset_36
- Total Entries: 100,000