Row 21805
Content Data
This page contains data entry 21805 from the Axioma AXP content repository. The structured data below represents the complete record for this entry.
In the context of multimodal search you can definitely tell the model the type of each vector by simply concatenating that information. You can then adjust your search metric so that similar images and texts remain close to each other by ignoring those concatenated dimensions in your calculation.
In a generative context you can just use special tokens to indicate where the image starts and ends.
| Field | Value |
|---|---|
| text | In the context of multimodal search you can definitely tell the model the type of each vector by simply concatenating that information. You can then adjust your search metric so that similar images and texts remain close to each other by ignoring those concatenated dimensions in your calculation. In a generative context you can just use special tokens to indicate where the image starts and ends. |
| label | r/machinelearning |
| dataType | comment |
| communityName | r/MachineLearning |
| datetime | 2024-05-21 |
| username_encoded | Z0FBQUFBQm5Lak1BUW1mSlpTdkpQUFAwWk5qOU53UXduV0labl9pYWZpbE53V2RmQmRiOThDaUh4eGk0OTJMMXNWWmdWYzNIWDE1Mm92S2xZcE5PUlp6ZXJZeDR0bGtxYmc9PQ== |
| url_encoded | Z0FBQUFBQm5Lak9QR2stZEVGc2RNRXJWVDhjMVhaa0huQjhRY2wySzBZOGR4akhyNFBuSFVYeE1CalJlMW9UQnlJNUoxSEpia19UMWk3dGgwenN1Y3RWS3A4TV9rMjR6QzczM0dVYXQ2aUdIRkR5TWY5NF9fdDZYTGJYaG1nQVRrMTVGWVp0VUdabHhuNWwxYXJva29JaUZmUUhWRG9PcWgwQ0NURERSak5XeUdpaldPMV9yVnBPcEZsY3l2alFUSDU2Z3RpWldXOHk1YzRrTVMxN1U1ZFRSSmZIMzQ3Mk8wNzd4VWtSazRGTU1UU1d0QUJtN2MxTT0= |
Raw Record
{
"text": "In the context of multimodal search you can definitely tell the model the type of each vector by simply concatenating that information. You can then adjust your search metric so that similar images and texts remain close to each other by ignoring those concatenated dimensions in your calculation.\n\nIn a generative context you can just use special tokens to indicate where the image starts and ends.",
"label": "r/machinelearning",
"dataType": "comment",
"communityName": "r/MachineLearning",
"datetime": "2024-05-21",
"username_encoded": "Z0FBQUFBQm5Lak1BUW1mSlpTdkpQUFAwWk5qOU53UXduV0labl9pYWZpbE53V2RmQmRiOThDaUh4eGk0OTJMMXNWWmdWYzNIWDE1Mm92S2xZcE5PUlp6ZXJZeDR0bGtxYmc9PQ==",
"url_encoded": "Z0FBQUFBQm5Lak9QR2stZEVGc2RNRXJWVDhjMVhaa0huQjhRY2wySzBZOGR4akhyNFBuSFVYeE1CalJlMW9UQnlJNUoxSEpia19UMWk3dGgwenN1Y3RWS3A4TV9rMjR6QzczM0dVYXQ2aUdIRkR5TWY5NF9fdDZYTGJYaG1nQVRrMTVGWVp0VUdabHhuNWwxYXJva29JaUZmUUhWRG9PcWgwQ0NURERSak5XeUdpaldPMV9yVnBPcEZsY3l2alFUSDU2Z3RpWldXOHk1YzRrTVMxN1U1ZFRSSmZIMzQ3Mk8wNzd4VWtSazRGTU1UU1d0QUJtN2MxTT0="
}
Entry Information
- Entry ID: 21805
- Repository: Axioma AXP
- Dataset: arrmlet/reddit_dataset_36
- Total Entries: 100,000