Row 26558
Content Data
This page contains data entry 26558 from the Axioma AXP content repository. The structured data below represents the complete record for this entry.
Oh, that was well hidden. Awesome, I'll have a deep dive into Unstructured. At least this takes away the burden of having to serialize / load the dataset for each step. With my current llamaindex setup, storing local files (and especially different processed versions of the same documents) is really cumbersome.
After skimming the other links, semantic chunking and proposition extraction sounds pretty much like what I'm aiming for. I'll give this a go.
Thanks again!
| Field | Value |
|---|---|
| text | Oh, that was well hidden. Awesome, I'll have a deep dive into Unstructured. At least this takes away the burden of having to serialize / load the dataset for each step. With my current llamaindex setup, storing local files (and especially different processed versions of the same documents) is really cumbersome. After skimming the other links, semantic chunking and proposition extraction sounds pretty much like what I'm aiming for. I'll give this a go. Thanks again! |
| label | r/datascience |
| dataType | comment |
| communityName | r/datascience |
| datetime | 2024-05-21 |
| username_encoded | Z0FBQUFBQm5Lak1EWko3SnN5em5SZUNYUnZIUEYwWjBoWGRWWWhaWGI0ZGlxSF9rZW5jNWxFTVI2UXhGcDUxdWpxSDJfVi1TNWJUQ2VDb3NIZGZ0cWx5LTZ5a08tQnhJOEE9PQ== |
| url_encoded | Z0FBQUFBQm5Lak9TWjBRVjJFX2I5YmVhZDRYNkd2YnYzYlpUY1dBNzdqbUV6VnAxY1RDazNBekZWNTVWckNYakY4V3BVeEt1bFJ6cjY5Y1d1Zm4ydVhqbF9XVmhtTV9WRW5PR2Nqb2o4U1ZtejhwVkR4N3ZzeVV4M3BiR0tyNU1zVFdWNDRMS180eF9tTjJQT2E0cVZ0ZmwxT09rajdscmFMMkE5RGd0NTlaX0s3X2tDMFgxcXYxREM3empUa3k1WHRHN3NsekNldFYwZlJiN0FZT2NZX25uQVI3aXAxVll5Zz09 |
Raw Record
{
"text": "Oh, that was well hidden. Awesome, I'll have a deep dive into Unstructured. At least this takes away the burden of having to serialize / load the dataset for each step. With my current llamaindex setup, storing local files (and especially different processed versions of the same documents) is really cumbersome.\n\nAfter skimming the other links, semantic chunking and proposition extraction sounds pretty much like what I'm aiming for. I'll give this a go.\n\nThanks again!",
"label": "r/datascience",
"dataType": "comment",
"communityName": "r/datascience",
"datetime": "2024-05-21",
"username_encoded": "Z0FBQUFBQm5Lak1EWko3SnN5em5SZUNYUnZIUEYwWjBoWGRWWWhaWGI0ZGlxSF9rZW5jNWxFTVI2UXhGcDUxdWpxSDJfVi1TNWJUQ2VDb3NIZGZ0cWx5LTZ5a08tQnhJOEE9PQ==",
"url_encoded": "Z0FBQUFBQm5Lak9TWjBRVjJFX2I5YmVhZDRYNkd2YnYzYlpUY1dBNzdqbUV6VnAxY1RDazNBekZWNTVWckNYakY4V3BVeEt1bFJ6cjY5Y1d1Zm4ydVhqbF9XVmhtTV9WRW5PR2Nqb2o4U1ZtejhwVkR4N3ZzeVV4M3BiR0tyNU1zVFdWNDRMS180eF9tTjJQT2E0cVZ0ZmwxT09rajdscmFMMkE5RGd0NTlaX0s3X2tDMFgxcXYxREM3empUa3k1WHRHN3NsekNldFYwZlJiN0FZT2NZX25uQVI3aXAxVll5Zz09"
}
Entry Information
- Entry ID: 26558
- Repository: Axioma AXP
- Dataset: arrmlet/reddit_dataset_36
- Total Entries: 100,000