Row 69450
Content Data
This page contains data entry 69450 from the Axioma AXP content repository. The structured data below represents the complete record for this entry.
I get what you're arguing, but
>AI doesn't have tact or bias
is just flat out wrong. Its biased towards whatever is in the training data - if your training data is 50% mein kampf, its going to have some bias. And while the size of the datasets used helps normalize this, bias isn't evenly distributed (e.g. theres not an equal amount of counter-stereotypes that say "white people love watermelon")
If the data that was being used to train them was guaranteed to be factual, then maybe youre right - the stereotypes would just be a reflection of that dataset. But the dataset is just the internet
But all of these are technical issues that ARE being worked on by white people who don't want racist models - I dont like the insinuation either that because they're white, they don't care or are incapable of prioritizing them
I mean, remember when Google released their model that tried *too* hard not to be racist, leading to hilarity? That's because Google knows the model has bias and is trying to account for it with prompting
| Field | Value |
|---|---|
| text | I get what you're arguing, but >AI doesn't have tact or bias is just flat out wrong. Its biased towards whatever is in the training data - if your training data is 50% mein kampf, its going to have some bias. And while the size of the datasets used helps normalize this, bias isn't evenly distributed (e.g. theres not an equal amount of counter-stereotypes that say "white people love watermelon") If the data that was being used to train them was guaranteed to be factual, then maybe youre right… |
| label | r/technology |
| dataType | comment |
| communityName | r/technology |
| datetime | 2024-05-23 |
| username_encoded | Z0FBQUFBQm5Lak1lNVU5a1NKMG04TjY1elpBd0R0MVdhRzFUdE9QZUd2MXE3bnNOeHhwQUo0MzZzb2I5LVZlX0k5NVhKdUk1cjRGNDFrU0FZcFNPSU55Y2Nhd054REFEUi1mY2NVSkk2ZGd0MEROMHA2b3daalU9 |
| url_encoded | Z0FBQUFBQm5Lak92X0xNZ2oxWmlsakotRi1DcEpXMjhQOXVHM190NFFFb3VsaUZXV0pjZXNnTlNuYmtjb0dJLXA2WUk5WmVYZmxOdlNsdklOcnRpQ0JEMllvV3VrX2VSTzV3VmRsRGluczJQdGRrd3lOc1NHMTNkTldoelV1SUo4WjQ5bHZhbS1XTUpYNzV6LVRkUUkxa3ZRaU9zZzB1QVM2VWdlMjlqdDhTM2MxRFlTbUk3bzNLT3dDQXhVR1QtSmpqRmlDWWwtYVdqRlFnR1hGVFc0dHpVd0JhamE0TUt1Zz09 |
Raw Record
{
"text": "I get what you're arguing, but \n\n>AI doesn't have tact or bias\n\nis just flat out wrong. Its biased towards whatever is in the training data - if your training data is 50% mein kampf, its going to have some bias. And while the size of the datasets used helps normalize this, bias isn't evenly distributed (e.g. theres not an equal amount of counter-stereotypes that say \"white people love watermelon\")\n\nIf the data that was being used to train them was guaranteed to be factual, then maybe youre right - the stereotypes would just be a reflection of that dataset. But the dataset is just the internet\n\nBut all of these are technical issues that ARE being worked on by white people who don't want racist models - I dont like the insinuation either that because they're white, they don't care or are incapable of prioritizing them\n\nI mean, remember when Google released their model that tried *too* hard not to be racist, leading to hilarity? That's because Google knows the model has bias and is trying to account for it with prompting",
"label": "r/technology",
"dataType": "comment",
"communityName": "r/technology",
"datetime": "2024-05-23",
"username_encoded": "Z0FBQUFBQm5Lak1lNVU5a1NKMG04TjY1elpBd0R0MVdhRzFUdE9QZUd2MXE3bnNOeHhwQUo0MzZzb2I5LVZlX0k5NVhKdUk1cjRGNDFrU0FZcFNPSU55Y2Nhd054REFEUi1mY2NVSkk2ZGd0MEROMHA2b3daalU9",
"url_encoded": "Z0FBQUFBQm5Lak92X0xNZ2oxWmlsakotRi1DcEpXMjhQOXVHM190NFFFb3VsaUZXV0pjZXNnTlNuYmtjb0dJLXA2WUk5WmVYZmxOdlNsdklOcnRpQ0JEMllvV3VrX2VSTzV3VmRsRGluczJQdGRrd3lOc1NHMTNkTldoelV1SUo4WjQ5bHZhbS1XTUpYNzV6LVRkUUkxa3ZRaU9zZzB1QVM2VWdlMjlqdDhTM2MxRFlTbUk3bzNLT3dDQXhVR1QtSmpqRmlDWWwtYVdqRlFnR1hGVFc0dHpVd0JhamE0TUt1Zz09"
}
Entry Information
- Entry ID: 69450
- Repository: Axioma AXP
- Dataset: arrmlet/reddit_dataset_36
- Total Entries: 100,000