Row 94873
Content Data
This page contains data entry 94873 from the Axioma AXP content repository. The structured data below represents the complete record for this entry.
[https://transformer-circuits.pub/2024/scaling-monosemanticity/index.html?s=09%2F/](https://transformer-circuits.pub/2024/scaling-monosemanticity/index.html?s=09%2F/) here's the whole paper, it's a good read front to back, very promising that we will make meaningful progress on mechanistic interpretability as capabilities increase. This work is on claude sonnet, which is a large enough proof of concept that the auto encoders should be able to scale to the current biggest models.
| Field | Value |
|---|---|
| text | [https://transformer-circuits.pub/2024/scaling-monosemanticity/index.html?s=09%2F/](https://transformer-circuits.pub/2024/scaling-monosemanticity/index.html?s=09%2F/) here's the whole paper, it's a good read front to back, very promising that we will make meaningful progress on mechanistic interpretability as capabilities increase. This work is on claude sonnet, which is a large enough proof of concept that the auto encoders should be able to scale to the current biggest models. |
| label | r/openai |
| dataType | comment |
| communityName | r/OpenAI |
| datetime | 2024-05-25 |
| username_encoded | Z0FBQUFBQm5Lak11QzIzamotR1NuS24wYXpCdzdvU1h0OHRaSHoteE5sXzUyblkzMk5Wc25fSW02S1owWlphVFJxNHVLSFhfNEVqUUR4SEl5T1FUdVNNMjBHSW1BYlIteHc9PQ== |
| url_encoded | Z0FBQUFBQm5LalBBUTljcTZUSTBqbGFyVDRQeElaTUZBYkhoWHZtQzhpRVB2Sm5WM0g4VnB0Ui1RWnQ2UHdNOThBZE5xcDNWX1M3LW5BOWNoa0FESTc1MVlzbzNDYUxmdC1EQ3g2NTVWNEJtT1dUb0VYdTN3OVJtMlIzT2lDRlFYQkY3U0hYaV9yYXJhQjJXNVQyYUJnazRBV0d4R1VsVURUa05wZm9vRHdMTnFzTk5Vb3Z2cFlJRG9mbkxWTGZ4ZGNDYWZaUzdudFBvUTVpd002eE9zUHpjaUhzVHBEb0ktdz09 |
Raw Record
{
"text": "[https://transformer-circuits.pub/2024/scaling-monosemanticity/index.html?s=09%2F/](https://transformer-circuits.pub/2024/scaling-monosemanticity/index.html?s=09%2F/) \nhere's the whole paper, it's a good read front to back, very promising that we will make meaningful progress on mechanistic interpretability as capabilities increase. This work is on claude sonnet, which is a large enough proof of concept that the auto encoders should be able to scale to the current biggest models.",
"label": "r/openai",
"dataType": "comment",
"communityName": "r/OpenAI",
"datetime": "2024-05-25",
"username_encoded": "Z0FBQUFBQm5Lak11QzIzamotR1NuS24wYXpCdzdvU1h0OHRaSHoteE5sXzUyblkzMk5Wc25fSW02S1owWlphVFJxNHVLSFhfNEVqUUR4SEl5T1FUdVNNMjBHSW1BYlIteHc9PQ==",
"url_encoded": "Z0FBQUFBQm5LalBBUTljcTZUSTBqbGFyVDRQeElaTUZBYkhoWHZtQzhpRVB2Sm5WM0g4VnB0Ui1RWnQ2UHdNOThBZE5xcDNWX1M3LW5BOWNoa0FESTc1MVlzbzNDYUxmdC1EQ3g2NTVWNEJtT1dUb0VYdTN3OVJtMlIzT2lDRlFYQkY3U0hYaV9yYXJhQjJXNVQyYUJnazRBV0d4R1VsVURUa05wZm9vRHdMTnFzTk5Vb3Z2cFlJRG9mbkxWTGZ4ZGNDYWZaUzdudFBvUTVpd002eE9zUHpjaUhzVHBEb0ktdz09"
}
Entry Information
- Entry ID: 94873
- Repository: Axioma AXP
- Dataset: arrmlet/reddit_dataset_36
- Total Entries: 100,000