Row 4996
Content Data
This page contains data entry 4996 from the Axioma AXP content repository. The structured data below represents the complete record for this entry.
I'm doing a talking face generation project In my problem statement is to create a tutor kinda of talking face, generated from a image (or else a video) and onemore thing is that to give a response atleast in a 5 sec.
I tried a SADTalker and im not happy with the results and its taking a long time to generate and also i checked out the other sota model like make it talk and Audio2 head but its very basic.
The backend part of ASR-MT-TTS are ready but the UI of this talking face need to be figure out, I need a open source models suggestion from you guys!!!
| Field | Value |
|---|---|
| text | I'm doing a talking face generation project In my problem statement is to create a tutor kinda of talking face, generated from a image (or else a video) and onemore thing is that to give a response atleast in a 5 sec. I tried a SADTalker and im not happy with the results and its taking a long time to generate and also i checked out the other sota model like make it talk and Audio2 head but its very basic. The backend part of ASR-MT-TTS are ready but the UI of this talking face need to be figu… |
| label | r/deeplearning |
| dataType | post |
| communityName | r/deeplearning |
| datetime | 2024-04-24 |
| username_encoded | Z0FBQUFBQm5LakwyMW9jRjZkbm5lRGYtMEFQMjFmV1dpU3RzeHZUOG9LZFhhdmNCOEh5b2NEZHl3MUt5NmFpMnZMZkNUZ0huQVRYRjB3bzRDbGZXYmoybUNCdmxZRGJXeEJybE1QZ3pHbEZQZGdONXZVNTZyaDg9 |
| url_encoded | Z0FBQUFBQm5Lak9Gc0sxME1ZaWRwcnl1aURWT3pGSE95cWtyZFJRYmtQczhiaXRuSlJwemRkLTFpVTVhbURQLW1YLUg4OFZjandHa3ItaTY5bVptUVNaZ0hRa2N3NkFleHphdkhiRDdTM1NsX1pESE5RUkdIX3pBa3FKMFZ4aWE5Z1U0NUEtdFMzVFF2UEZGaXU0eUpKVDdmVEJEWFlBLVM1elFYaFUyeGxsV1NBWDFuNXlYY3NFPQ== |
Raw Record
{
"text": "I'm doing a talking face generation project In my problem statement is to create a tutor kinda of talking face, generated from a image (or else a video) and onemore thing is that to give a response atleast in a 5 sec. \n\nI tried a SADTalker and im not happy with the results and its taking a long time to generate and also i checked out the other sota model like make it talk and Audio2 head but its very basic.\n\nThe backend part of ASR-MT-TTS are ready but the UI of this talking face need to be figure out, I need a open source models suggestion from you guys!!!",
"label": "r/deeplearning",
"dataType": "post",
"communityName": "r/deeplearning",
"datetime": "2024-04-24",
"username_encoded": "Z0FBQUFBQm5LakwyMW9jRjZkbm5lRGYtMEFQMjFmV1dpU3RzeHZUOG9LZFhhdmNCOEh5b2NEZHl3MUt5NmFpMnZMZkNUZ0huQVRYRjB3bzRDbGZXYmoybUNCdmxZRGJXeEJybE1QZ3pHbEZQZGdONXZVNTZyaDg9",
"url_encoded": "Z0FBQUFBQm5Lak9Gc0sxME1ZaWRwcnl1aURWT3pGSE95cWtyZFJRYmtQczhiaXRuSlJwemRkLTFpVTVhbURQLW1YLUg4OFZjandHa3ItaTY5bVptUVNaZ0hRa2N3NkFleHphdkhiRDdTM1NsX1pESE5RUkdIX3pBa3FKMFZ4aWE5Z1U0NUEtdFMzVFF2UEZGaXU0eUpKVDdmVEJEWFlBLVM1elFYaFUyeGxsV1NBWDFuNXlYY3NFPQ=="
}
Entry Information
- Entry ID: 4996
- Repository: Axioma AXP
- Dataset: arrmlet/reddit_dataset_36
- Total Entries: 100,000