Row 43435

Row ID: 43435 | Dataset Entry | Axioma AXP Content Repository

Content Data

This page contains data entry 43435 from the Axioma AXP content repository. The structured data below represents the complete record for this entry.

I could be wrong, but the reason that is in my understanding is because we are using the 4o model. The 4o LLM model, which is significantly faster, which is also faster in voice conversations. However, there is a voice model part of the 4o model, which is natively multimodal, and we're definitely not speaking to that one.

That is to say, the model we're speaking to in voice right now is the one that doesn't hear voice but translates voice into text first and then responds to the text. But the one that we will be getting is the one that hears voice and is trained to understand human inflections intonation and also respond with the same.

FieldValue
text I could be wrong, but the reason that is in my understanding is because we are using the 4o model. The 4o LLM model, which is significantly faster, which is also faster in voice conversations. However, there is a voice model part of the 4o model, which is natively multimodal, and we're definitely not speaking to that one. That is to say, the model we're speaking to in voice right now is the one that doesn't hear voice but translates voice into text first and then responds to the text. But the o…
label r/openai
dataType comment
communityName r/OpenAI
datetime 2024-05-22
username_encoded Z0FBQUFBQm5Lak1PMlFUTDN1VUhfd1dmWjY1aVZCVG9xRVhoMkpNQWNkeTcxY0VkRVViQWY2aWpfZHBySlZFUzVOQS1MRVNjUkMxNjBMeHFwQTlELUMzVjRxTzFtb2paa1E9PQ==
url_encoded Z0FBQUFBQm5Lak9kNy1fc0p3YVJxZE85WVVHZ3JKOUhudzlYZVJEc0dRdDNfcHYwRldJajZPQlNGNDRpRURxaHVjZnpJaUpUYjNkdHUzMkd1MVRmdkNVLUNYOWN4blY3VGRiMlM3V2I1TnpNalYydVZHanRENFlMNFduVzkxN1V3bWtTcENUSnZuUkd4eGdDbFpkZlk0SF9BQ3FmWWZHdnpDMlJibGpuT3l5VWdmeDlFcEVOVFJXOFRfZWJCS1ppWnRpUjUwQlA5Z0plYjhWakNYQTV5TU1Bdl9TVktaM1EyQT09

Raw Record

{
  "text": "I could be wrong, but the reason that is in my understanding is because we are using the 4o model. The 4o LLM model, which is significantly faster, which is also faster in voice conversations. However, there is a voice model part of the 4o model, which is natively multimodal, and we're definitely not speaking to that one.\n\nThat is to say, the model we're speaking to in voice right now is the one that doesn't hear voice but translates voice into text first and then responds to the text. But the one that we will be getting is the one that hears voice and is trained to understand human inflections intonation and also respond with the same.",
  "label": "r/openai",
  "dataType": "comment",
  "communityName": "r/OpenAI",
  "datetime": "2024-05-22",
  "username_encoded": "Z0FBQUFBQm5Lak1PMlFUTDN1VUhfd1dmWjY1aVZCVG9xRVhoMkpNQWNkeTcxY0VkRVViQWY2aWpfZHBySlZFUzVOQS1MRVNjUkMxNjBMeHFwQTlELUMzVjRxTzFtb2paa1E9PQ==",
  "url_encoded": "Z0FBQUFBQm5Lak9kNy1fc0p3YVJxZE85WVVHZ3JKOUhudzlYZVJEc0dRdDNfcHYwRldJajZPQlNGNDRpRURxaHVjZnpJaUpUYjNkdHUzMkd1MVRmdkNVLUNYOWN4blY3VGRiMlM3V2I1TnpNalYydVZHanRENFlMNFduVzkxN1V3bWtTcENUSnZuUkd4eGdDbFpkZlk0SF9BQ3FmWWZHdnpDMlJibGpuT3l5VWdmeDlFcEVOVFJXOFRfZWJCS1ppWnRpUjUwQlA5Z0plYjhWakNYQTV5TU1Bdl9TVktaM1EyQT09"
}

Entry Information