upvote
I can imagine this having negative consequences for open platforms, such as Wikipedia, Bluesky, Mastodon, Common Crawl, etc. It could have implications for Stack Exchange keeping their datasets downloadable, and general API availability of most platforms. How would this make provisions for keeping the things that are currently free to access still free? People are already buckling under the weight of subscription fatigue, and I can only see this making that worse. I'm willing to give up a lot for the sake of privacy, but the people who would be most affected by such a change could stand to lose a lot more.
reply
I am not sure what the point of this is.

If they are exporting data for sale, we can tax the value they are selling it at. If they are using data to generate ad revenue, we can tax that.

If we want to enforce privacy standards, we can make liability laws around sharing that data with third parties; we can make it harder for users to consent to the sharing of data by making users decide and approve on a case by case basis.

There is so much to do before exotic schemes, for instance way less intrusive things like making it so privacy policies legally survive bankruptcy.

reply
How does that work for non-profit web crawls like the Common Crawl Foundation?
reply
Right, because if you want a solid solution to a problem, you ask governments to tax it! Always works.
reply
I like the overall direction, but how to tax it without screwing over hobbyist archivists aka data hoarders? That is, how to assess the value of some datastore, somehow distinguishing between personal preference profiles, random LLM software slop and, say, vintage anime collections?
reply
> how to tax it without screwing over hobbyist archivists

If you're a corporation, you're taxed. If you're a human, you're not.

reply
Perhaps you are human that wants to isolate the expenses, activity and liability into a corporate organization, like Wikimedia?
reply