Row 39195
Content Data
This page contains data entry 39195 from the Axioma AXP content repository. The structured data below represents the complete record for this entry.
LLM model doesn't learn using online chat but using huge amount of data. what are you talking about is somewhat online Q-learning where agents of the model learn simultaneously using reward function given by the user. This type of learning is resources demanding, imagine milion of user training their agent, OpenAI servers would burn.
people often mistake between model learning (training) and interaction (conversation) during the conversation despite correcting the agent in one session, It will forget everything with new session. Because every session is running the standard GPT-model.
| Field | Value |
|---|---|
| text | LLM model doesn't learn using online chat but using huge amount of data. what are you talking about is somewhat online Q-learning where agents of the model learn simultaneously using reward function given by the user. This type of learning is resources demanding, imagine milion of user training their agent, OpenAI servers would burn. people often mistake between model learning (training) and interaction (conversation) during the conversation despite correcting the agent in one session, It will… |
| label | r/openai |
| dataType | comment |
| communityName | r/OpenAI |
| datetime | 2024-05-22 |
| username_encoded | Z0FBQUFBQm5Lak1MVlpKNDVTbklSdW9aVVdnek1KSTYtcVpIM3VfWExGRERCejF4NW5Ic1JPX1d2QTRkOVdpeExxcjFYQ2dILW93SDhobmVicGtVRmNLNWwteUpBWUFzX0E9PQ== |
| url_encoded | Z0FBQUFBQm5Lak9heGg5c3ZITjBOYXU5U2xMb0ppX2pCREtZTFFjalpvdGVIc0h4ejFEM2FfTHF1dGVLTFI4THp0bDUzZ0lXeWxrUDdJRm9hN2E4YnZ6RXNsUG1tbGZ2OTcwd0ZSTG50M3UtTGJKN3QxSENCdTc1SGNKUVNBcWIyV1JFek9fNUEzOFByMGlFUmFvbkpHN2NtYm9DdnlpWDhhMkFrSnZ5QTY3Wld2LVgtZ2tuZHloMHZyTWhiaEwxRlo0LUdQZ1cwSVZaa3lJbDVjRF9ZaFNLdTVMWlZJMDQ2Zz09 |
Raw Record
{
"text": "LLM model doesn't learn using online chat but using huge amount of data. what are you talking about is somewhat online Q-learning where agents of the model learn simultaneously using reward function given by the user. This type of learning is resources demanding, imagine milion of user training their agent, OpenAI servers would burn.\n\npeople often mistake between model learning (training) and interaction (conversation) during the conversation despite correcting the agent in one session, It will forget everything with new session. Because every session is running the standard GPT-model.",
"label": "r/openai",
"dataType": "comment",
"communityName": "r/OpenAI",
"datetime": "2024-05-22",
"username_encoded": "Z0FBQUFBQm5Lak1MVlpKNDVTbklSdW9aVVdnek1KSTYtcVpIM3VfWExGRERCejF4NW5Ic1JPX1d2QTRkOVdpeExxcjFYQ2dILW93SDhobmVicGtVRmNLNWwteUpBWUFzX0E9PQ==",
"url_encoded": "Z0FBQUFBQm5Lak9heGg5c3ZITjBOYXU5U2xMb0ppX2pCREtZTFFjalpvdGVIc0h4ejFEM2FfTHF1dGVLTFI4THp0bDUzZ0lXeWxrUDdJRm9hN2E4YnZ6RXNsUG1tbGZ2OTcwd0ZSTG50M3UtTGJKN3QxSENCdTc1SGNKUVNBcWIyV1JFek9fNUEzOFByMGlFUmFvbkpHN2NtYm9DdnlpWDhhMkFrSnZ5QTY3Wld2LVgtZ2tuZHloMHZyTWhiaEwxRlo0LUdQZ1cwSVZaa3lJbDVjRF9ZaFNLdTVMWlZJMDQ2Zz09"
}
Entry Information
- Entry ID: 39195
- Repository: Axioma AXP
- Dataset: arrmlet/reddit_dataset_36
- Total Entries: 100,000