Row 4996

Row ID: 4996 | Dataset Entry | Axioma AXP Content Repository

Content Data

This page contains data entry 4996 from the Axioma AXP content repository. The structured data below represents the complete record for this entry.

I'm doing a talking face generation project In my problem statement is to create a tutor kinda of talking face, generated from a image (or else a video) and onemore thing is that to give a response atleast in a 5 sec.

I tried a SADTalker and im not happy with the results and its taking a long time to generate and also i checked out the other sota model like make it talk and Audio2 head but its very basic.

The backend part of ASR-MT-TTS are ready but the UI of this talking face need to be figure out, I need a open source models suggestion from you guys!!!

FieldValue
text I'm doing a talking face generation project In my problem statement is to create a tutor kinda of talking face, generated from a image (or else a video) and onemore thing is that to give a response atleast in a 5 sec. I tried a SADTalker and im not happy with the results and its taking a long time to generate and also i checked out the other sota model like make it talk and Audio2 head but its very basic. The backend part of ASR-MT-TTS are ready but the UI of this talking face need to be figu…
label r/deeplearning
dataType post
communityName r/deeplearning
datetime 2024-04-24
username_encoded Z0FBQUFBQm5LakwyMW9jRjZkbm5lRGYtMEFQMjFmV1dpU3RzeHZUOG9LZFhhdmNCOEh5b2NEZHl3MUt5NmFpMnZMZkNUZ0huQVRYRjB3bzRDbGZXYmoybUNCdmxZRGJXeEJybE1QZ3pHbEZQZGdONXZVNTZyaDg9
url_encoded Z0FBQUFBQm5Lak9Gc0sxME1ZaWRwcnl1aURWT3pGSE95cWtyZFJRYmtQczhiaXRuSlJwemRkLTFpVTVhbURQLW1YLUg4OFZjandHa3ItaTY5bVptUVNaZ0hRa2N3NkFleHphdkhiRDdTM1NsX1pESE5RUkdIX3pBa3FKMFZ4aWE5Z1U0NUEtdFMzVFF2UEZGaXU0eUpKVDdmVEJEWFlBLVM1elFYaFUyeGxsV1NBWDFuNXlYY3NFPQ==

Raw Record

{
  "text": "I'm doing a talking face generation project In my problem statement is to create a tutor kinda of talking face, generated from a image (or else a video) and onemore thing is that to give a response atleast in a 5 sec. \n\nI tried a SADTalker and im not happy with the results and its taking a long time to generate and also i checked out the other sota model like make it talk and Audio2 head but its very basic.\n\nThe backend part of ASR-MT-TTS are ready but the UI of this talking face need to be figure out, I need a open source models suggestion from you guys!!!",
  "label": "r/deeplearning",
  "dataType": "post",
  "communityName": "r/deeplearning",
  "datetime": "2024-04-24",
  "username_encoded": "Z0FBQUFBQm5LakwyMW9jRjZkbm5lRGYtMEFQMjFmV1dpU3RzeHZUOG9LZFhhdmNCOEh5b2NEZHl3MUt5NmFpMnZMZkNUZ0huQVRYRjB3bzRDbGZXYmoybUNCdmxZRGJXeEJybE1QZ3pHbEZQZGdONXZVNTZyaDg9",
  "url_encoded": "Z0FBQUFBQm5Lak9Gc0sxME1ZaWRwcnl1aURWT3pGSE95cWtyZFJRYmtQczhiaXRuSlJwemRkLTFpVTVhbURQLW1YLUg4OFZjandHa3ItaTY5bVptUVNaZ0hRa2N3NkFleHphdkhiRDdTM1NsX1pESE5RUkdIX3pBa3FKMFZ4aWE5Z1U0NUEtdFMzVFF2UEZGaXU0eUpKVDdmVEJEWFlBLVM1elFYaFUyeGxsV1NBWDFuNXlYY3NFPQ=="
}

Entry Information