Row 6518
Content Data
This page contains data entry 6518 from the Axioma AXP content repository. The structured data below represents the complete record for this entry.
I have used Spark on Databricks quite long, without understanding it properly (my known language is Python, so I use pyspark but would like to dig deeper into spark/scala). I like the idea of Spark being open source so it can be relevant for understanding other tools such as Databricks better in-depth and my impression is that big data processing/ML in the academia/research is often done directly on Spark. I have one foot in research and could work in that context some day, however currently it is better to prioritize more industry-valid stuff. So if I deep-dive into Spark, will I get projects? (Projects where I can really use it?) I am located in Northern Europe.
| Field | Value |
|---|---|
| text | I have used Spark on Databricks quite long, without understanding it properly (my known language is Python, so I use pyspark but would like to dig deeper into spark/scala). I like the idea of Spark being open source so it can be relevant for understanding other tools such as Databricks better in-depth and my impression is that big data processing/ML in the academia/research is often done directly on Spark. I have one foot in research and could work in that context some day, however currently it … |
| label | r/datascience |
| dataType | post |
| communityName | r/datascience |
| datetime | 2024-05-11 |
| username_encoded | Z0FBQUFBQm5LakwyUkVzT3c2a055TXNzeklsLW5tX1d4RVpEa2ZqZkw5eFd4SkoxcTk0QVlDYzVSTk1ENml6Z2dscl82M1VVUW5OTzJQdTNuMUo0Z1ZxamlNc2tpTzdwc2hBbVVpd1dnNnNSd0laZGVjejhPa009 |
| url_encoded | Z0FBQUFBQm5Lak9HSEZ2d3JmSlpBdTZTY091UUFnNW9mTVppS2RPMkp3eTB0a0ZjVjhhWjl5Q1hNTzBhV3Noc01CMFlzZUtVQVRabndyZWlVTnRjQTJpc0FyQ3hpcXhwRGRoZ0hpUVZwNWhoSUxyQnMwNlpqOG5RbllpT0FNd0t0SWxUVy12dmVBc2NwY1JJNFUzSmhmRlJfcG5iNTQ1MEt4NXhicjlfRFJya0x4NTE0WVFpUGVKQjFnajhNdGJ2QS1SeFlMeXJmUzVq |
Raw Record
{
"text": "I have used Spark on Databricks quite long, without understanding it properly (my known language is Python, so I use pyspark but would like to dig deeper into spark/scala). I like the idea of Spark being open source so it can be relevant for understanding other tools such as Databricks better in-depth and my impression is that big data processing/ML in the academia/research is often done directly on Spark. I have one foot in research and could work in that context some day, however currently it is better to prioritize more industry-valid stuff. So if I deep-dive into Spark, will I get projects? (Projects where I can really use it?) I am located in Northern Europe.",
"label": "r/datascience",
"dataType": "post",
"communityName": "r/datascience",
"datetime": "2024-05-11",
"username_encoded": "Z0FBQUFBQm5LakwyUkVzT3c2a055TXNzeklsLW5tX1d4RVpEa2ZqZkw5eFd4SkoxcTk0QVlDYzVSTk1ENml6Z2dscl82M1VVUW5OTzJQdTNuMUo0Z1ZxamlNc2tpTzdwc2hBbVVpd1dnNnNSd0laZGVjejhPa009",
"url_encoded": "Z0FBQUFBQm5Lak9HSEZ2d3JmSlpBdTZTY091UUFnNW9mTVppS2RPMkp3eTB0a0ZjVjhhWjl5Q1hNTzBhV3Noc01CMFlzZUtVQVRabndyZWlVTnRjQTJpc0FyQ3hpcXhwRGRoZ0hpUVZwNWhoSUxyQnMwNlpqOG5RbllpT0FNd0t0SWxUVy12dmVBc2NwY1JJNFUzSmhmRlJfcG5iNTQ1MEt4NXhicjlfRFJya0x4NTE0WVFpUGVKQjFnajhNdGJ2QS1SeFlMeXJmUzVq"
}
Entry Information
- Entry ID: 6518
- Repository: Axioma AXP
- Dataset: arrmlet/reddit_dataset_36
- Total Entries: 100,000