Search arXiv⌕ Search

arXiv · 2609.38433

The Bureaucratization of the Internet: Analyzing the Diffusion of Governance Regimes on Reddit (2011--2023)

Abstract

As online platforms scale, distributed communities must develop complex governance structures to maintain order. Reddit represents a unique experiment in this process: millions of subreddits manage their own digital commons, yet together they form an interconnected ecosystem of platform governance. Using a historical dataset of subreddit rules from 2011 to 2023, we apply computational methods to characterize that ecosystem and identify mechanisms driving change. We identify seven distinct governance regimes and document a platform-wide drift toward irreversible bureaucratization operating through two reinforcing dynamics: a ratchet effect among existing communities and a cohort shift in which new communities increasingly launch with pre-packaged regulatory frameworks. Using Dyadic Event History Analysis, we find that community scale is the dominant and universal predictor of rule adoption, independent of network exposure. Diffusion flows through two complementary channels: normative transmission via shared moderator networks and mimetic transmission among topically similar communities. However, mimetic copying is selective: it drives adoption of content rules addressing shared topical challenges while communities resist copying operational and behavioral rules. Additionally, communities with overlapping user bases differentiate rather than converge, consistent with ecological accounts of niche competition, and prestige-based contagion is firmly rejected. Together, these findings reveal decentralized platform governance as a stratified ecosystem in which lateral coordination, scale pressure, structural inertia, cohort effects, and episodic platform coercion jointly produce irreversible formalization with direct implications for platform management.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Katherine Van Koevering, Yuanhao Liu, Jon Kleinberg. 2026-09-29. The Bureaucratization of the Internet: Analyzing the Diffusion of Governance Regimes on Reddit (2011--2023). https://arxiv.org/abs/2609.38433

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Beyond the Clique: Comparing Clique and Dowker Complexes for Co-occurrence Data in Learning Analytics

Learning analytics increasingly represents the relations in co-occurrence data---codes in a window of discourse, participants in a thread, tags on a post---as simplicial complexes and analyses them with persistent homology. The standard approach uses the clique complex. Because its simplices are determined by the pairwise network alone, it cannot distinguish "three elements co-occurred together" from "each of the three pairs co-occurred separately". The Dowker complex, in contrast, takes as simplices the sets of entities actually observed to co-occur. We compare analyses based on these two complexes. First, we show that, on the same 1-skeleton, the Dowker complex is a subcomplex of the clique complex, the map on first homology induced by the inclusion is surjective, and its kernel is generated by phantom triangles that never co-occurred; that is, the clique construction can only erase holes. We then test this on real data. On four Stack Exchange data sets, 46--91% of clique triangles are phantom and 522--1,970 holes are erased. On the example data of the learning analytics R packages tna/Nestimate, the first Betti number $β_1$ of the clique complex is 0 at every threshold, whereas the Dowker complex detects holes. Against a degree-preserving null model, the Dowker complex departs strongly in five of the six data sets. Moreover, this difference does not appear in the fixed-threshold analyses used in practice. We conclude that the construction should follow the data type and that, for observed groups, the Dowker complex is the appropriate choice; it fills a gap in current tooling. Code is available at https://github.com/igu-lab/beyond-the-clique.

cs.SI↗

Degree-Corrected Joint Matrix Factorization for Multilayer Community Detection

Multilayer networks allow the modeling of interactions between the same entities across different contexts, such as temporal observations, varying settings, or interactions of different types. The goal of community detection in multilayer networks is to identify groups of nodes exhibiting similar connectivity patterns, which may vary across layers. We propose a method based on a joint nonnegative symmetric matrix trifactorization for community detection in multilayer networks, where each graph is approximated by a nonnegative symmetric matrix trifactorization. Our approach enforces constraints on the factor matrices so that communities are disjoint and shared across layers, while allowing each layer to have its own connectivity patterns and node degrees. This flexibility enables the model to capture both local and global structural variations across layers. We also develop an algorithm to efficiently solve this problem. We evaluate multilayer community detection methods using the multilayer degree-corrected stochastic block model (MDCBM), a flexible framework for generating realistic multilayer graphs with heterogeneous degrees and varying connectivity patterns. Experiments show that our method reliably detects communities across diverse regimes, whereas existing state-of-the-art approaches are often limited by restrictive structural assumptions.

cs.SI↗

Out-of-Network Attention Dynamics on Bluesky

Personalized social media commonly relies on explicit follow graphs to shape what content users encounter; yet how much attention crosses ties they have not formed remains largely undocumented at scale. We study this question on Bluesky, a large decentralized microblogging platform whose default feed relies on a simple, reverse-chronological content recommender. We analyze 173 million user-author interactions (likes, reposts, replies, and quotes) collected from a near-complete platform dump between February and September 2023. We decompose each interaction by attention-path length (already followed, relayed by a followed account, reachable within two follow-hops, or beyond) and find that 74.5% of interactions reach the user through an account they already follow. Measured by distance in the follow graph rather than by route, 80.6% of interaction lands within two follow hops, far beyond the 22.4% an expected-degree null predicts. We then characterize how exploration varies across users and over tenure. A broad-reaching minority generates three quarters of all exploratory activity, while aggregate declines in exploration with tenure mask three distinct individual trajectories. Finally, attention reaching beyond two hops converts into new follow ties at less than one third the rate of two-hop-local exploratory attention. Together, these results depict a platform where out-of-network exploration is substantial in volume but strongly constrained by network proximity and unlikely to translate into new social ties.

cs.SI↗