Row 90206
Content Data
This page contains data entry 90206 from the Axioma AXP content repository. The structured data below represents the complete record for this entry.
I guess I'm confused as to why we need two inputs, when it seems like we only need one list of retrieved items.
Each query would have different positive and negative pairs, so for each query we would have to record which passage is positive and which is negative along with their position in the original corpus (since it seems like FAISS returns the indices of documents). When we retrieve the documents we would check if there are positive samples there and calculate our ranking metrics accordingly. I don't know why we need a separate reference array.
| Field | Value |
|---|---|
| text | I guess I'm confused as to why we need two inputs, when it seems like we only need one list of retrieved items. Each query would have different positive and negative pairs, so for each query we would have to record which passage is positive and which is negative along with their position in the original corpus (since it seems like FAISS returns the indices of documents). When we retrieve the documents we would check if there are positive samples there and calculate our ranking metrics according… |
| label | r/machinelearning |
| dataType | comment |
| communityName | r/MachineLearning |
| datetime | 2024-05-25 |
| username_encoded | Z0FBQUFBQm5Lak1yRXY3ZFB0TldGX3A4QUtVUGx5ZnhWVkNWVHlYTlFDNDJCRUh4MTczbXhuejhMVWFvQ1JmckdQTU9oVHVrdm1lSTV5LUVlMnpLOW0wbGxqYWduQUFHaEE9PQ== |
| url_encoded | Z0FBQUFBQm5Lak84WW5maXlEbmVXLTRvOXZGU3kxYzJBQ2ZlVVMtUG1wNDhSc0JMRjNSZjh0SGsyRVRpZkZfMXE4N0tQTUZjREJlalRrS25sSzZERDctRFJ5NmdYREFWSjIwalpaUVhKUEExM0pZbEh2NklFY1EtNk1kUVRNcy05X2hwWUt2endNaUNzeGwya3R3MXZVRlRpeGhNbk9HeTM4aEdJcThEQ1VRTmpncmo4WEttSFBxREdUdXdLcFpIUkxhaXhlY1ZXVE9STUdhNExsOVZDMEVBVTBHZ21RVUYydz09 |
Raw Record
{
"text": "I guess I'm confused as to why we need two inputs, when it seems like we only need one list of retrieved items.\n\nEach query would have different positive and negative pairs, so for each query we would have to record which passage is positive and which is negative along with their position in the original corpus (since it seems like FAISS returns the indices of documents). When we retrieve the documents we would check if there are positive samples there and calculate our ranking metrics accordingly. I don't know why we need a separate reference array.",
"label": "r/machinelearning",
"dataType": "comment",
"communityName": "r/MachineLearning",
"datetime": "2024-05-25",
"username_encoded": "Z0FBQUFBQm5Lak1yRXY3ZFB0TldGX3A4QUtVUGx5ZnhWVkNWVHlYTlFDNDJCRUh4MTczbXhuejhMVWFvQ1JmckdQTU9oVHVrdm1lSTV5LUVlMnpLOW0wbGxqYWduQUFHaEE9PQ==",
"url_encoded": "Z0FBQUFBQm5Lak84WW5maXlEbmVXLTRvOXZGU3kxYzJBQ2ZlVVMtUG1wNDhSc0JMRjNSZjh0SGsyRVRpZkZfMXE4N0tQTUZjREJlalRrS25sSzZERDctRFJ5NmdYREFWSjIwalpaUVhKUEExM0pZbEh2NklFY1EtNk1kUVRNcy05X2hwWUt2endNaUNzeGwya3R3MXZVRlRpeGhNbk9HeTM4aEdJcThEQ1VRTmpncmo4WEttSFBxREdUdXdLcFpIUkxhaXhlY1ZXVE9STUdhNExsOVZDMEVBVTBHZ21RVUYydz09"
}
Entry Information
- Entry ID: 90206
- Repository: Axioma AXP
- Dataset: arrmlet/reddit_dataset_36
- Total Entries: 100,000