@
[email protected]
Actually since yesterday I'm pondering about the idea to build a #federated version of stackoverflow, nothing written yet, I'm reading, researching.
Also, right now I was checking this stack exchange sqlite db under CC BY-SA 4.0 to check how useful and doable would be import this data and using as a base for the federated version.
Also wondering if we could use this data somehow to train our own opensource AI to help the community, but I'm do not have knowledge on LLM/AI things. Please if there is any expert I would appreciate the opinion on that.
https://seqlite.puny.engineering
EDIT:
A better place to download the dump content, with more interesting tables, like the one with the Votes; the other link the dumped data only contains two tables Users and Posts. Right now downloading the whole data related with stack overflow, and will take some time due my humble home internet connection, so I didn't have the chance to take a look at the data, but I guess that's the interesting thing.
here:
https://archive.org/download/stackexchange
They have already access to SO’s CC content, why would they get it from the fediverse?
They already have it.
I said alternative to SO. As in, likely, a place to post new content (answers, comments). Nothing can really be done with the content OAI already got their hands on other than firing off a few well-placed EMP bombs.
Yes, but you mentioned importing old content is problematic, and I don’t see why?
Because to import old content, you have to respect the old license (or get every contributor of back-then to relicense). That would mean having a site with contents under differing licenses depending on date, which is something the corpos can use as an excuse to continue siphoning everything without consequence.
I’m fine with a mirror / archive of SO. But it shoudl very definitively be a different thing than an active SO alternative, and their users and data storages should be also different.