Row 26558

Row ID: 26558 | Dataset Entry | Axioma AXP Content Repository

Content Data

This page contains data entry 26558 from the Axioma AXP content repository. The structured data below represents the complete record for this entry.

Oh, that was well hidden. Awesome, I'll have a deep dive into Unstructured. At least this takes away the burden of having to serialize / load the dataset for each step. With my current llamaindex setup, storing local files (and especially different processed versions of the same documents) is really cumbersome.

After skimming the other links, semantic chunking and proposition extraction sounds pretty much like what I'm aiming for. I'll give this a go.

Thanks again!

FieldValue
text Oh, that was well hidden. Awesome, I'll have a deep dive into Unstructured. At least this takes away the burden of having to serialize / load the dataset for each step. With my current llamaindex setup, storing local files (and especially different processed versions of the same documents) is really cumbersome. After skimming the other links, semantic chunking and proposition extraction sounds pretty much like what I'm aiming for. I'll give this a go. Thanks again!
label r/datascience
dataType comment
communityName r/datascience
datetime 2024-05-21
username_encoded Z0FBQUFBQm5Lak1EWko3SnN5em5SZUNYUnZIUEYwWjBoWGRWWWhaWGI0ZGlxSF9rZW5jNWxFTVI2UXhGcDUxdWpxSDJfVi1TNWJUQ2VDb3NIZGZ0cWx5LTZ5a08tQnhJOEE9PQ==
url_encoded Z0FBQUFBQm5Lak9TWjBRVjJFX2I5YmVhZDRYNkd2YnYzYlpUY1dBNzdqbUV6VnAxY1RDazNBekZWNTVWckNYakY4V3BVeEt1bFJ6cjY5Y1d1Zm4ydVhqbF9XVmhtTV9WRW5PR2Nqb2o4U1ZtejhwVkR4N3ZzeVV4M3BiR0tyNU1zVFdWNDRMS180eF9tTjJQT2E0cVZ0ZmwxT09rajdscmFMMkE5RGd0NTlaX0s3X2tDMFgxcXYxREM3empUa3k1WHRHN3NsekNldFYwZlJiN0FZT2NZX25uQVI3aXAxVll5Zz09

Raw Record

{
  "text": "Oh, that was well hidden. Awesome, I'll have a deep dive into Unstructured. At least this takes away the burden of having to serialize / load the dataset for each step. With my current llamaindex setup, storing local files (and especially different processed versions of the same documents) is really cumbersome.\n\nAfter skimming the other links, semantic chunking and proposition extraction sounds pretty much like what I'm aiming for. I'll give this a go.\n\nThanks again!",
  "label": "r/datascience",
  "dataType": "comment",
  "communityName": "r/datascience",
  "datetime": "2024-05-21",
  "username_encoded": "Z0FBQUFBQm5Lak1EWko3SnN5em5SZUNYUnZIUEYwWjBoWGRWWWhaWGI0ZGlxSF9rZW5jNWxFTVI2UXhGcDUxdWpxSDJfVi1TNWJUQ2VDb3NIZGZ0cWx5LTZ5a08tQnhJOEE9PQ==",
  "url_encoded": "Z0FBQUFBQm5Lak9TWjBRVjJFX2I5YmVhZDRYNkd2YnYzYlpUY1dBNzdqbUV6VnAxY1RDazNBekZWNTVWckNYakY4V3BVeEt1bFJ6cjY5Y1d1Zm4ydVhqbF9XVmhtTV9WRW5PR2Nqb2o4U1ZtejhwVkR4N3ZzeVV4M3BiR0tyNU1zVFdWNDRMS180eF9tTjJQT2E0cVZ0ZmwxT09rajdscmFMMkE5RGd0NTlaX0s3X2tDMFgxcXYxREM3empUa3k1WHRHN3NsekNldFYwZlJiN0FZT2NZX25uQVI3aXAxVll5Zz09"
}

Entry Information