The Internet began when people started publishing data in electronic form: that is, not just storing it for themselves, but providing access to it via remote connection. To ensure this, a minimum set of components is required: 1) a data storage device; 2) a communication channel to which the storage device is connected; 3) a protocol for providing remote access to this data. There can be many different types of devices, channels, and protocols, as long as they can somehow interface with each other. This looks quite promising in terms of ensuring decentralized data exchange, doesn’t it? Of course. But there is also a vast open field here for deepening centralization.
It is cheaper to produce identical devices. Identical communication channels are cheaper to maintain. The use of a single protocol is an invaluable asset. But the main thing is the functional separation that occurs between different system components when, for example, it turns out that storing data is cheaper on a large stationary server with a thick, uninterrupted communication channel, while accessing it is more convenient from inexpensive personal clients. And just like that, a platform economy begins, in which those who need to store and consume data become completely dependent on the rules established by the storage and publishing platform.
Naturally, throughout history, this state of affairs has greatly irritated both data producers and consumers, who would prefer to retain control. However, the question is how to have this control without spending too much money, ensuring everything works fast and doesn’t require mega-skills from every user. The struggle for the user, accompanied by the involvement of the state as a conflict regulator and the state’s own initiatives to ensure the interests of politicians, is all well-documented and generally well-known. The struggle for digital autonomy is also well-documented and quite known. One could provide an overview of the dynamic equilibrium between centralization and decentralization processes at the time of writing this text, but I want to approach this from a different angle. Specifically: how do I personally see the ideal functioning of the Internet given the current level of technology?
Current snapshot of the technological level
- Data storage. Without investing in specialized devices, the average user easily secures volumes on the order of hundreds of gigabytes on a phone and several terabytes on bulkier personal devices.
- Communication channels. A more or less settled user can relatively easily afford a 24/7 unlimited channel with speeds ranging from units to tens of megabytes per second. An actively traveling user, without additional investment, may occasionally be offline, have traffic limits, and a more modest channel width.
- Clouds. For a price comparable to the cost of personal internet access, a user can rent cloud capacity for data storage and processing that slightly exceeds the power of their personal devices.
- Money. Thanks to cryptocurrencies, a user can technically pay for any services over the network directly to their providers in arbitrarily small fractions with an arbitrarily high frequency.
Now I will daydream
I produce only a few gigabytes of data per month, mostly crappy photos. With crappy videos, let’s say it would be tens of gigabytes (the highest subjective value for me, of course, is text, which is a mere few hundred kilobytes, and with all the chatter in chats, let’s say a few megabytes). I want to have unconditional access to all this content from any of my devices, as well as the ability to share access to individual pieces of content with both a limited and an unlimited circle of people. To achieve this, I need my texts to be fully synchronized between several of my personal devices and cloud storage, while photos and videos gradually settle into cheaper and higher-capacity storage, unobtrusively leaving, say, the phone—but with the ability to easily return any archived data back to local access.
Beyond this, I consume other people’s content. Here we are talking about hundreds of gigabytes per month. I need the ability to selectively save any data into personal storage. Ideally, while preserving metadata about where and when this content was obtained by me. It would also be good to substitute locally saved content when surfing the same network resource again to avoid re-downloading—and if the content on the network has updated, to have the ability to replace my version with a fresh one or save an archived version.
I also value a convenient way to donate directly for someone else’s content and receive donations for my own. For this, I need the ability to attach payment detail metadata to an arbitrary object.
So, effectively, I need an operating system for working online that would link locally stored materials with the network addresses where they should be available, provide flexible data synchronization between storage systems, allow for configuring access rights, contain tools for viewing and editing data, as well as a toolkit for managing money. Naturally, all of this must be sufficiently protected from unauthorized access.
It turns out that in order to provide me with the most decentralized workflow on the web, a quite comprehensive centralized solution is required. It could, of course, consist of modules from different providers, but these modules need to work together seamlessly, so centralization practically invites itself here.
And what do we have today?
There are separate services for organizing local data storage and for organizing cloud storage. There are services for synchronizing local folders with cloud ones. There are tools for various network activities: publishing photos, videos, texts, tracks, exchanging messages, money, plans, task lists, and a ton of other types of data. There is the ability to rent servers and host websites on them. There is the ability to rent a domain name. In short, there is a huge, flourishing variety of tools for working on the web, the development of which is quite decentralized. And none of this gives me the possibility of full control over my data.
Such is the paradox.
