Row 4311
Content Data
This page contains data entry 4311 from the Axioma AXP content repository. The structured data below represents the complete record for this entry.
Last year, Stuart Russell's team devised a strategy to beat AlphaGo. A top amateur player won using a "loop" tactic that a human would have easily spot and dispatched. Because this tactic presumably did not occur in many of the human-played games in AlphaGo's training set, AlphaGo failed to recognize it and lost. [https://www.ft.com/content/175e5314-a7f7-4741-a786-273219f433a1](https://www.ft.com/content/175e5314-a7f7-4741-a786-273219f433a1)
My question: Since AlphaZero is not trained on human-played games, but rather learned through self-play, it should not be susceptible to this sort of "hack". True?
| Field | Value |
|---|---|
| text | Last year, Stuart Russell's team devised a strategy to beat AlphaGo. A top amateur player won using a "loop" tactic that a human would have easily spot and dispatched. Because this tactic presumably did not occur in many of the human-played games in AlphaGo's training set, AlphaGo failed to recognize it and lost. [https://www.ft.com/content/175e5314-a7f7-4741-a786-273219f433a1](https://www.ft.com/content/175e5314-a7f7-4741-a786-273219f433a1) My question: Since AlphaZero is not trained on human… |
| label | r/neuralnetworks |
| dataType | post |
| communityName | r/neuralnetworks |
| datetime | 2024-04-07 |
| username_encoded | Z0FBQUFBQm5LakwxSU5ham5fUXp0QWhtSkJRRk5RUFFpMzZBU2hjLWROVWZzaWljcDNKTEk1VWgwZTJJR3h4WUM1NmdmTjJmOHZEaExCTU8wbTU0R2J2QkpRRlVhY2ZxWVE9PQ== |
| url_encoded | Z0FBQUFBQm5Lak9GRE95X0pieHVuY0I1blVlV0NXdjNMVnZuNFZxTmlkbWEwYmFDVjBqbnBBZDBtS2pHenFjcnhKdjhiV0FfcEN3SGhpSVlrT01IaTZuSTdSYk1VV0pjMWNPSUdaRklJcXlNcEVsRTV2VWRKRENIcUk4OGlLVjN1c1ZVeU9uNk81Z2FvZ29EYlV5UUhrWDNlNHZNMXI5Rmlyd1hCTTk0cVNFeUZVSUdMU0N5MVJsWENHQkdCa1JLSHFDZ2NjbUxLRlg1YXV6LS1XQWhheTE2bXljdm9QZGZBdz09 |
Raw Record
{
"text": "Last year, Stuart Russell's team devised a strategy to beat AlphaGo. A top amateur player won using a \"loop\" tactic that a human would have easily spot and dispatched. Because this tactic presumably did not occur in many of the human-played games in AlphaGo's training set, AlphaGo failed to recognize it and lost. [https://www.ft.com/content/175e5314-a7f7-4741-a786-273219f433a1](https://www.ft.com/content/175e5314-a7f7-4741-a786-273219f433a1)\n\nMy question: Since AlphaZero is not trained on human-played games, but rather learned through self-play, it should not be susceptible to this sort of \"hack\". True? ",
"label": "r/neuralnetworks",
"dataType": "post",
"communityName": "r/neuralnetworks",
"datetime": "2024-04-07",
"username_encoded": "Z0FBQUFBQm5LakwxSU5ham5fUXp0QWhtSkJRRk5RUFFpMzZBU2hjLWROVWZzaWljcDNKTEk1VWgwZTJJR3h4WUM1NmdmTjJmOHZEaExCTU8wbTU0R2J2QkpRRlVhY2ZxWVE9PQ==",
"url_encoded": "Z0FBQUFBQm5Lak9GRE95X0pieHVuY0I1blVlV0NXdjNMVnZuNFZxTmlkbWEwYmFDVjBqbnBBZDBtS2pHenFjcnhKdjhiV0FfcEN3SGhpSVlrT01IaTZuSTdSYk1VV0pjMWNPSUdaRklJcXlNcEVsRTV2VWRKRENIcUk4OGlLVjN1c1ZVeU9uNk81Z2FvZ29EYlV5UUhrWDNlNHZNMXI5Rmlyd1hCTTk0cVNFeUZVSUdMU0N5MVJsWENHQkdCa1JLSHFDZ2NjbUxLRlg1YXV6LS1XQWhheTE2bXljdm9QZGZBdz09"
}
Entry Information
- Entry ID: 4311
- Repository: Axioma AXP
- Dataset: arrmlet/reddit_dataset_36
- Total Entries: 100,000