Row 90206

Row ID: 90206 | Dataset Entry | Axioma AXP Content Repository

Content Data

This page contains data entry 90206 from the Axioma AXP content repository. The structured data below represents the complete record for this entry.

I guess I'm confused as to why we need two inputs, when it seems like we only need one list of retrieved items.

Each query would have different positive and negative pairs, so for each query we would have to record which passage is positive and which is negative along with their position in the original corpus (since it seems like FAISS returns the indices of documents). When we retrieve the documents we would check if there are positive samples there and calculate our ranking metrics accordingly. I don't know why we need a separate reference array.

FieldValue
text I guess I'm confused as to why we need two inputs, when it seems like we only need one list of retrieved items. Each query would have different positive and negative pairs, so for each query we would have to record which passage is positive and which is negative along with their position in the original corpus (since it seems like FAISS returns the indices of documents). When we retrieve the documents we would check if there are positive samples there and calculate our ranking metrics according…
label r/machinelearning
dataType comment
communityName r/MachineLearning
datetime 2024-05-25
username_encoded Z0FBQUFBQm5Lak1yRXY3ZFB0TldGX3A4QUtVUGx5ZnhWVkNWVHlYTlFDNDJCRUh4MTczbXhuejhMVWFvQ1JmckdQTU9oVHVrdm1lSTV5LUVlMnpLOW0wbGxqYWduQUFHaEE9PQ==
url_encoded Z0FBQUFBQm5Lak84WW5maXlEbmVXLTRvOXZGU3kxYzJBQ2ZlVVMtUG1wNDhSc0JMRjNSZjh0SGsyRVRpZkZfMXE4N0tQTUZjREJlalRrS25sSzZERDctRFJ5NmdYREFWSjIwalpaUVhKUEExM0pZbEh2NklFY1EtNk1kUVRNcy05X2hwWUt2endNaUNzeGwya3R3MXZVRlRpeGhNbk9HeTM4aEdJcThEQ1VRTmpncmo4WEttSFBxREdUdXdLcFpIUkxhaXhlY1ZXVE9STUdhNExsOVZDMEVBVTBHZ21RVUYydz09

Raw Record

{
  "text": "I guess I'm confused as to why we need two inputs, when it seems like we only need one list of retrieved items.\n\nEach query would have different positive and negative pairs, so for each query we would have to record which passage is positive and which is negative along with their position in the original corpus (since it seems like FAISS returns the indices of documents). When we retrieve the documents we would check if there are positive samples there and calculate our ranking metrics accordingly. I don't know why we need a separate reference array.",
  "label": "r/machinelearning",
  "dataType": "comment",
  "communityName": "r/MachineLearning",
  "datetime": "2024-05-25",
  "username_encoded": "Z0FBQUFBQm5Lak1yRXY3ZFB0TldGX3A4QUtVUGx5ZnhWVkNWVHlYTlFDNDJCRUh4MTczbXhuejhMVWFvQ1JmckdQTU9oVHVrdm1lSTV5LUVlMnpLOW0wbGxqYWduQUFHaEE9PQ==",
  "url_encoded": "Z0FBQUFBQm5Lak84WW5maXlEbmVXLTRvOXZGU3kxYzJBQ2ZlVVMtUG1wNDhSc0JMRjNSZjh0SGsyRVRpZkZfMXE4N0tQTUZjREJlalRrS25sSzZERDctRFJ5NmdYREFWSjIwalpaUVhKUEExM0pZbEh2NklFY1EtNk1kUVRNcy05X2hwWUt2endNaUNzeGwya3R3MXZVRlRpeGhNbk9HeTM4aEdJcThEQ1VRTmpncmo4WEttSFBxREdUdXdLcFpIUkxhaXhlY1ZXVE9STUdhNExsOVZDMEVBVTBHZ21RVUYydz09"
}

Entry Information