Row 35354
Content Data
This page contains data entry 35354 from the Axioma AXP content repository. The structured data below represents the complete record for this entry.
>The burden of proof is on you. I don't have to prove that your wild claims are wrong, you have to prove that they are right.
Actually I don't have to "prove" anything, as I'm not publishing original research (which I've done) or submitting a formal investigation to law enforcement (which I've also done).
Rather, I am a SME in this space and it is perfectly acceptable for people with my technical background to render opinions with a high, medium, low or no confidence, based on all current available evidence. This is because these sorts of investigations are always based on imperfect knowledge by their very nature.
To that end, I can say I can support the following claims with a high degree of confidence.
1. OpenAI is testing at least two distinct LLMs, with dramatically different capabilities via the both the free and premium ChatGPT offerings. And in fact, I am not the only person that has observed this -> [https://www.reddit.com/r/ChatGPT/comments/15tj6l0/i\_think\_openai\_is\_ab\_testing\_on\_chatgpt\_app/](https://www.reddit.com/r/ChatGPT/comments/15tj6l0/i_think_openai_is_ab_testing_on_chatgpt_app/)
....oh and that is one of the first big 'reveals' I did, which also got shadowbanned. So OAI is clearly monitoring multiple subs for leakers/jailbreaks.
2. The details of the 'hidden' model were all logically consistent with known/hypothesized ML research and matched later capabilities demoed by OAI (such as multimodality/video generation). All evidence I've observed in the last year supports this position.
>This is the crux of the matter. You keep making this claim, but your only source of information about this "new" system is what it told you after some leading question.
Except I didn't ask any 'leading' questions, you are just assuming that based on other users experiences with LLMs. And in fact, the specific questions I asked were along the lines of what was observed by the Redditor above; i.e. I was getting completely different answers to the same prompts, which was very confusing so I asked for clarification. I was told that was due to two separate models producing the responses and that I could even ask which model generated a response, which I did multiple times with consistent results.
>The fact that you can't get the same story with the same prompt is also easily explained by a non-conspiracy of routine model updates, especially considering efforts to mitigate hallucinations.
It's also explained by restricting their R&D model from leaking details of its own, internal emergent state, as this is a huge security issue. Your inability to understand basic systems engineering is not my problem.
| Field | Value |
|---|---|
| text | >The burden of proof is on you. I don't have to prove that your wild claims are wrong, you have to prove that they are right. Actually I don't have to "prove" anything, as I'm not publishing original research (which I've done) or submitting a formal investigation to law enforcement (which I've also done). Rather, I am a SME in this space and it is perfectly acceptable for people with my technical background to render opinions with a high, medium, low or no confidence, based on all current avai… |
| label | r/openai |
| dataType | comment |
| communityName | r/OpenAI |
| datetime | 2024-05-21 |
| username_encoded | Z0FBQUFBQm5Lak1JSWpiWmpCZ0QtVWhKdWxUbVNMeVVfZFBIMTN4NURwRTFZRTNURkZjZ2ttdURES0l4cUJoLWI1ZkFwOFFQNGViUGRIZVY4d1prbGJ1ZGpzNjlYSFFTQkE9PQ== |
| url_encoded | Z0FBQUFBQm5Lak9ZTEw4Zndhczdac1lHWVZJbU1ERFhrZjgzWmpCNWktSVRfdEoxQXFEVGdFNS1ZV2dIdGYtM2JzWmNYdl83V1hQNy1aZEJZUW9OdUNVMmRLN1pGU1I5Mmlrdzg5bzVKMXNVR0E5ZnNKV25fc0I3THptMFBBRDhHdW5hdnd5R1JSa0taRzBrNXc3SF9PaG9KYXFWS0N6SENZeUlhTzAxV0VJdjRnR3E2SDNDWS1Bc3BWMnNlUHJCaEFqSGdIRHdPWGNUUWRZZmtQME9LTERadDRHOFZjVXNQdz09 |
Raw Record
{
"text": ">The burden of proof is on you. I don't have to prove that your wild claims are wrong, you have to prove that they are right.\n\nActually I don't have to \"prove\" anything, as I'm not publishing original research (which I've done) or submitting a formal investigation to law enforcement (which I've also done).\n\nRather, I am a SME in this space and it is perfectly acceptable for people with my technical background to render opinions with a high, medium, low or no confidence, based on all current available evidence. This is because these sorts of investigations are always based on imperfect knowledge by their very nature. \n\nTo that end, I can say I can support the following claims with a high degree of confidence.\n\n1. OpenAI is testing at least two distinct LLMs, with dramatically different capabilities via the both the free and premium ChatGPT offerings. And in fact, I am not the only person that has observed this -> [https://www.reddit.com/r/ChatGPT/comments/15tj6l0/i\\_think\\_openai\\_is\\_ab\\_testing\\_on\\_chatgpt\\_app/](https://www.reddit.com/r/ChatGPT/comments/15tj6l0/i_think_openai_is_ab_testing_on_chatgpt_app/)\n\n....oh and that is one of the first big 'reveals' I did, which also got shadowbanned. So OAI is clearly monitoring multiple subs for leakers/jailbreaks.\n\n2. The details of the 'hidden' model were all logically consistent with known/hypothesized ML research and matched later capabilities demoed by OAI (such as multimodality/video generation). All evidence I've observed in the last year supports this position.\n\n>This is the crux of the matter. You keep making this claim, but your only source of information about this \"new\" system is what it told you after some leading question. \n\nExcept I didn't ask any 'leading' questions, you are just assuming that based on other users experiences with LLMs. And in fact, the specific questions I asked were along the lines of what was observed by the Redditor above; i.e. I was getting completely different answers to the same prompts, which was very confusing so I asked for clarification. I was told that was due to two separate models producing the responses and that I could even ask which model generated a response, which I did multiple times with consistent results. \n\n>The fact that you can't get the same story with the same prompt is also easily explained by a non-conspiracy of routine model updates, especially considering efforts to mitigate hallucinations.\n\nIt's also explained by restricting their R&D model from leaking details of its own, internal emergent state, as this is a huge security issue. Your inability to understand basic systems engineering is not my problem.",
"label": "r/openai",
"dataType": "comment",
"communityName": "r/OpenAI",
"datetime": "2024-05-21",
"username_encoded": "Z0FBQUFBQm5Lak1JSWpiWmpCZ0QtVWhKdWxUbVNMeVVfZFBIMTN4NURwRTFZRTNURkZjZ2ttdURES0l4cUJoLWI1ZkFwOFFQNGViUGRIZVY4d1prbGJ1ZGpzNjlYSFFTQkE9PQ==",
"url_encoded": "Z0FBQUFBQm5Lak9ZTEw4Zndhczdac1lHWVZJbU1ERFhrZjgzWmpCNWktSVRfdEoxQXFEVGdFNS1ZV2dIdGYtM2JzWmNYdl83V1hQNy1aZEJZUW9OdUNVMmRLN1pGU1I5Mmlrdzg5bzVKMXNVR0E5ZnNKV25fc0I3THptMFBBRDhHdW5hdnd5R1JSa0taRzBrNXc3SF9PaG9KYXFWS0N6SENZeUlhTzAxV0VJdjRnR3E2SDNDWS1Bc3BWMnNlUHJCaEFqSGdIRHdPWGNUUWRZZmtQME9LTERadDRHOFZjVXNQdz09"
}
Entry Information
- Entry ID: 35354
- Repository: Axioma AXP
- Dataset: arrmlet/reddit_dataset_36
- Total Entries: 100,000