Row 39195

Row ID: 39195 | Dataset Entry | Axioma AXP Content Repository

Content Data

This page contains data entry 39195 from the Axioma AXP content repository. The structured data below represents the complete record for this entry.

LLM model doesn't learn using online chat but using huge amount of data. what are you talking about is somewhat online Q-learning where agents of the model learn simultaneously using reward function given by the user. This type of learning is resources demanding, imagine milion of user training their agent, OpenAI servers would burn.

people often mistake between model learning (training) and interaction (conversation) during the conversation despite correcting the agent in one session, It will forget everything with new session. Because every session is running the standard GPT-model.

FieldValue
text LLM model doesn't learn using online chat but using huge amount of data. what are you talking about is somewhat online Q-learning where agents of the model learn simultaneously using reward function given by the user. This type of learning is resources demanding, imagine milion of user training their agent, OpenAI servers would burn. people often mistake between model learning (training) and interaction (conversation) during the conversation despite correcting the agent in one session, It will…
label r/openai
dataType comment
communityName r/OpenAI
datetime 2024-05-22
username_encoded Z0FBQUFBQm5Lak1MVlpKNDVTbklSdW9aVVdnek1KSTYtcVpIM3VfWExGRERCejF4NW5Ic1JPX1d2QTRkOVdpeExxcjFYQ2dILW93SDhobmVicGtVRmNLNWwteUpBWUFzX0E9PQ==
url_encoded Z0FBQUFBQm5Lak9heGg5c3ZITjBOYXU5U2xMb0ppX2pCREtZTFFjalpvdGVIc0h4ejFEM2FfTHF1dGVLTFI4THp0bDUzZ0lXeWxrUDdJRm9hN2E4YnZ6RXNsUG1tbGZ2OTcwd0ZSTG50M3UtTGJKN3QxSENCdTc1SGNKUVNBcWIyV1JFek9fNUEzOFByMGlFUmFvbkpHN2NtYm9DdnlpWDhhMkFrSnZ5QTY3Wld2LVgtZ2tuZHloMHZyTWhiaEwxRlo0LUdQZ1cwSVZaa3lJbDVjRF9ZaFNLdTVMWlZJMDQ2Zz09

Raw Record

{
  "text": "LLM model doesn't learn using online chat but using huge amount of data. what are you talking about is somewhat online Q-learning where agents of the model learn simultaneously  using reward function given by the user. This type of learning is resources demanding, imagine milion of user training their agent, OpenAI servers would burn.\n\npeople often mistake between model learning (training) and interaction (conversation) during the conversation despite correcting the agent in one session, It will forget everything with new session. Because every session is running the standard GPT-model.",
  "label": "r/openai",
  "dataType": "comment",
  "communityName": "r/OpenAI",
  "datetime": "2024-05-22",
  "username_encoded": "Z0FBQUFBQm5Lak1MVlpKNDVTbklSdW9aVVdnek1KSTYtcVpIM3VfWExGRERCejF4NW5Ic1JPX1d2QTRkOVdpeExxcjFYQ2dILW93SDhobmVicGtVRmNLNWwteUpBWUFzX0E9PQ==",
  "url_encoded": "Z0FBQUFBQm5Lak9heGg5c3ZITjBOYXU5U2xMb0ppX2pCREtZTFFjalpvdGVIc0h4ejFEM2FfTHF1dGVLTFI4THp0bDUzZ0lXeWxrUDdJRm9hN2E4YnZ6RXNsUG1tbGZ2OTcwd0ZSTG50M3UtTGJKN3QxSENCdTc1SGNKUVNBcWIyV1JFek9fNUEzOFByMGlFUmFvbkpHN2NtYm9DdnlpWDhhMkFrSnZ5QTY3Wld2LVgtZ2tuZHloMHZyTWhiaEwxRlo0LUdQZ1cwSVZaa3lJbDVjRF9ZaFNLdTVMWlZJMDQ2Zz09"
}

Entry Information