Row 2270
Content Data
This page contains data entry 2270 from the Axioma AXP content repository. The structured data below represents the complete record for this entry.
From my perspective as someone who doesn't code, the current learning curve for using AWS to train a custom LLM on new data is too steep for me to bother with right now. Formatting the data isn't an issue, it's trying to figure out what the training steps are after that (and trying not to get dinged for accidentally using compute after one has given up on it and getting a bill). Any open source project which made the process more like using a word processor or spreadsheet (with non technical instructions) would be 1. opening up a huge Pandora's box, which perhaps should remain locked for now) 2. inundated with new users with ideas and a lack of coding expertise.
| Field | Value |
|---|---|
| text | From my perspective as someone who doesn't code, the current learning curve for using AWS to train a custom LLM on new data is too steep for me to bother with right now. Formatting the data isn't an issue, it's trying to figure out what the training steps are after that (and trying not to get dinged for accidentally using compute after one has given up on it and getting a bill). Any open source project which made the process more like using a word processor or spreadsheet (with non technical ins… |
| label | r/opensourceai |
| dataType | comment |
| communityName | r/OpenSourceAI |
| datetime | 2023-07-28 |
| username_encoded | Z0FBQUFBQm5LakwwUUFsZG82a3VkeXFkMFVYNzhaQW1ZN0pTazYtRFVEYmwxSEc5VzlJeDNMQUd5SEpQRHJEeHdHbGh3azE5SUJhT3BwU0FFNjJNWlJJRGZqcHlvNW81cnc9PQ== |
| url_encoded | Z0FBQUFBQm5Lak9FbnZwdVkydGdFMjJCbjkxaHhjeG0xbWhzbmZ3dFd2S1JQOTRtVE0yc3g1V1JzLXJ4U3RaY04xZVA5QVJKb1l0cFd5cU9PV2xTVUN5VllpVmJFVjhTdUxqUmRCck91UF9uMG9BbEpxZDFpQ3c5bXVsQU1ZendCTW1BdnJPS2FWQWhCZ1RSVG9GVEJqS0xDZTA3MC0wM3RjZmJPcG81eXdyYlE4eUd1cGpZYWVfV1ZVZHhNb2lfcjQ0eFRwUGpmTlBFUmo2dWItNVYyQ0wxMXE5dy02RFRFR1YxckNpUDlsMko0YXVDYnFQVDVPdz0= |
Raw Record
{
"text": "From my perspective as someone who doesn't code, the current learning curve for using AWS to train a custom LLM on new data is too steep for me to bother with right now. Formatting the data isn't an issue, it's trying to figure out what the training steps are after that (and trying not to get dinged for accidentally using compute after one has given up on it and getting a bill). Any open source project which made the process more like using a word processor or spreadsheet (with non technical instructions) would be 1. opening up a huge Pandora's box, which perhaps should remain locked for now) 2. inundated with new users with ideas and a lack of coding expertise.",
"label": "r/opensourceai",
"dataType": "comment",
"communityName": "r/OpenSourceAI",
"datetime": "2023-07-28",
"username_encoded": "Z0FBQUFBQm5LakwwUUFsZG82a3VkeXFkMFVYNzhaQW1ZN0pTazYtRFVEYmwxSEc5VzlJeDNMQUd5SEpQRHJEeHdHbGh3azE5SUJhT3BwU0FFNjJNWlJJRGZqcHlvNW81cnc9PQ==",
"url_encoded": "Z0FBQUFBQm5Lak9FbnZwdVkydGdFMjJCbjkxaHhjeG0xbWhzbmZ3dFd2S1JQOTRtVE0yc3g1V1JzLXJ4U3RaY04xZVA5QVJKb1l0cFd5cU9PV2xTVUN5VllpVmJFVjhTdUxqUmRCck91UF9uMG9BbEpxZDFpQ3c5bXVsQU1ZendCTW1BdnJPS2FWQWhCZ1RSVG9GVEJqS0xDZTA3MC0wM3RjZmJPcG81eXdyYlE4eUd1cGpZYWVfV1ZVZHhNb2lfcjQ0eFRwUGpmTlBFUmo2dWItNVYyQ0wxMXE5dy02RFRFR1YxckNpUDlsMko0YXVDYnFQVDVPdz0="
}
Entry Information
- Entry ID: 2270
- Repository: Axioma AXP
- Dataset: arrmlet/reddit_dataset_36
- Total Entries: 100,000