Row 94873

Row ID: 94873 | Dataset Entry | Axioma AXP Content Repository

Content Data

This page contains data entry 94873 from the Axioma AXP content repository. The structured data below represents the complete record for this entry.

[https://transformer-circuits.pub/2024/scaling-monosemanticity/index.html?s=09%2F/](https://transformer-circuits.pub/2024/scaling-monosemanticity/index.html?s=09%2F/) here's the whole paper, it's a good read front to back, very promising that we will make meaningful progress on mechanistic interpretability as capabilities increase. This work is on claude sonnet, which is a large enough proof of concept that the auto encoders should be able to scale to the current biggest models.

FieldValue
text [https://transformer-circuits.pub/2024/scaling-monosemanticity/index.html?s=09%2F/](https://transformer-circuits.pub/2024/scaling-monosemanticity/index.html?s=09%2F/) here's the whole paper, it's a good read front to back, very promising that we will make meaningful progress on mechanistic interpretability as capabilities increase. This work is on claude sonnet, which is a large enough proof of concept that the auto encoders should be able to scale to the current biggest models.
label r/openai
dataType comment
communityName r/OpenAI
datetime 2024-05-25
username_encoded Z0FBQUFBQm5Lak11QzIzamotR1NuS24wYXpCdzdvU1h0OHRaSHoteE5sXzUyblkzMk5Wc25fSW02S1owWlphVFJxNHVLSFhfNEVqUUR4SEl5T1FUdVNNMjBHSW1BYlIteHc9PQ==
url_encoded Z0FBQUFBQm5LalBBUTljcTZUSTBqbGFyVDRQeElaTUZBYkhoWHZtQzhpRVB2Sm5WM0g4VnB0Ui1RWnQ2UHdNOThBZE5xcDNWX1M3LW5BOWNoa0FESTc1MVlzbzNDYUxmdC1EQ3g2NTVWNEJtT1dUb0VYdTN3OVJtMlIzT2lDRlFYQkY3U0hYaV9yYXJhQjJXNVQyYUJnazRBV0d4R1VsVURUa05wZm9vRHdMTnFzTk5Vb3Z2cFlJRG9mbkxWTGZ4ZGNDYWZaUzdudFBvUTVpd002eE9zUHpjaUhzVHBEb0ktdz09

Raw Record

{
  "text": "[https://transformer-circuits.pub/2024/scaling-monosemanticity/index.html?s=09%2F/](https://transformer-circuits.pub/2024/scaling-monosemanticity/index.html?s=09%2F/)  \nhere's the whole paper, it's a good read front to back, very promising that we will make meaningful progress on mechanistic interpretability as capabilities increase. This work is on claude sonnet, which is a large enough proof of concept that the auto encoders should be able to scale to the current biggest models.",
  "label": "r/openai",
  "dataType": "comment",
  "communityName": "r/OpenAI",
  "datetime": "2024-05-25",
  "username_encoded": "Z0FBQUFBQm5Lak11QzIzamotR1NuS24wYXpCdzdvU1h0OHRaSHoteE5sXzUyblkzMk5Wc25fSW02S1owWlphVFJxNHVLSFhfNEVqUUR4SEl5T1FUdVNNMjBHSW1BYlIteHc9PQ==",
  "url_encoded": "Z0FBQUFBQm5LalBBUTljcTZUSTBqbGFyVDRQeElaTUZBYkhoWHZtQzhpRVB2Sm5WM0g4VnB0Ui1RWnQ2UHdNOThBZE5xcDNWX1M3LW5BOWNoa0FESTc1MVlzbzNDYUxmdC1EQ3g2NTVWNEJtT1dUb0VYdTN3OVJtMlIzT2lDRlFYQkY3U0hYaV9yYXJhQjJXNVQyYUJnazRBV0d4R1VsVURUa05wZm9vRHdMTnFzTk5Vb3Z2cFlJRG9mbkxWTGZ4ZGNDYWZaUzdudFBvUTVpd002eE9zUHpjaUhzVHBEb0ktdz09"
}

Entry Information