Row 66143
Content Data
This page contains data entry 66143 from the Axioma AXP content repository. The structured data below represents the complete record for this entry.
So, I used to have the same understanding as you wrote above.
Though, doesn't this paper \[1\] show the opposite? That for Sparse GP models, FITC (minimizes fKL) UNDERESTIMATES the predictive variance and VFE (minimizes rKL) OVERESTIMATES it (see section 3.1). This result is what started to confuse me really and what made me interested in the question. Might just be something specific to sparse GPs though...
\[1\] Bauer, M., Van der Wilk, M., & Rasmussen, C. E. (2016). Understanding probabilistic sparse Gaussian process approximations. *Advances in neural information processing systems*, *29*
| Field | Value |
|---|---|
| text | So, I used to have the same understanding as you wrote above. Though, doesn't this paper \[1\] show the opposite? That for Sparse GP models, FITC (minimizes fKL) UNDERESTIMATES the predictive variance and VFE (minimizes rKL) OVERESTIMATES it (see section 3.1). This result is what started to confuse me really and what made me interested in the question. Might just be something specific to sparse GPs though... \[1\] Bauer, M., Van der Wilk, M., & Rasmussen, C. E. (2016). Understanding probabi… |
| label | r/machinelearning |
| dataType | comment |
| communityName | r/MachineLearning |
| datetime | 2024-05-23 |
| username_encoded | Z0FBQUFBQm5Lak1jNU1NNHJVQ1VhX1BGazJ3UGwyRFdJaWpSQTljM3lOd0MxMkc1N19rTXpvWW1lSlJWN2pOaWIxY19DZlRKVlI2WktYUkpMQXlTZk5Gd0FTZXYtUGdUdmc9PQ== |
| url_encoded | Z0FBQUFBQm5Lak9zSWlJUndPOFVxcGFBd19RWXpVREE5a1VlTTF6OTJIcFQzVWpoOFM4U0NqTzJGRG5nSVdzQjVmUFkwYVlIWFMybENwNkJ5OVJWalo2R19VTURWMm1OU19NTnhrN20wWHlhSGNzV3M4d1FtWTFmYlZMenY1VEQ4S0dva3Q1VXpuY3NLZldkR1AwT1lscFJjZVEzTHFMV0dBYk9uNzZBTUpJMXVoYUtMTjVsUFZ2a3VISHp0ZzhDZHJPaE1hbC03U1g1dnhYbjNUYzNDQ1ZxaWRFaVJRZWpYZUdVRU9VeU9oX01zSnFFNVRfMG9Tbz0= |
Raw Record
{
"text": "So, I used to have the same understanding as you wrote above.\n\nThough, doesn't this paper \\[1\\] show the opposite? That for Sparse GP models, FITC (minimizes fKL) UNDERESTIMATES the predictive variance and VFE (minimizes rKL) OVERESTIMATES it (see section 3.1). This result is what started to confuse me really and what made me interested in the question. Might just be something specific to sparse GPs though...\n\n \n\\[1\\] Bauer, M., Van der Wilk, M., & Rasmussen, C. E. (2016). Understanding probabilistic sparse Gaussian process approximations. *Advances in neural information processing systems*, *29*",
"label": "r/machinelearning",
"dataType": "comment",
"communityName": "r/MachineLearning",
"datetime": "2024-05-23",
"username_encoded": "Z0FBQUFBQm5Lak1jNU1NNHJVQ1VhX1BGazJ3UGwyRFdJaWpSQTljM3lOd0MxMkc1N19rTXpvWW1lSlJWN2pOaWIxY19DZlRKVlI2WktYUkpMQXlTZk5Gd0FTZXYtUGdUdmc9PQ==",
"url_encoded": "Z0FBQUFBQm5Lak9zSWlJUndPOFVxcGFBd19RWXpVREE5a1VlTTF6OTJIcFQzVWpoOFM4U0NqTzJGRG5nSVdzQjVmUFkwYVlIWFMybENwNkJ5OVJWalo2R19VTURWMm1OU19NTnhrN20wWHlhSGNzV3M4d1FtWTFmYlZMenY1VEQ4S0dva3Q1VXpuY3NLZldkR1AwT1lscFJjZVEzTHFMV0dBYk9uNzZBTUpJMXVoYUtMTjVsUFZ2a3VISHp0ZzhDZHJPaE1hbC03U1g1dnhYbjNUYzNDQ1ZxaWRFaVJRZWpYZUdVRU9VeU9oX01zSnFFNVRfMG9Tbz0="
}
Entry Information
- Entry ID: 66143
- Repository: Axioma AXP
- Dataset: arrmlet/reddit_dataset_36
- Total Entries: 100,000