Searcharxiv⌕ Search

arXiv · 2610.00305

Crude, Commercial, and Self-Referential: Chinese-Language Coordinated Activity in Japanese-Language X

Abstract

Malicious coordination has long been regarded as a principal source of information ecosystem pollution. Here, we focus on crude, text-repetition-based coordination. As the demand for mitigating its dissemination has grown, scholars have studied such coordination, focusing especially on bot detection. Few studies, however, have characterized malicious coordination per se or examined how it elicits reactions from general users. Leveraging a dataset of 734,173 Chinese-language coordinated accounts and around 495 million coordinated posts published between May 2024 and March 2026, this study analyzes the characteristics of coordinated behavior and how general users react to coordinated posts. We report three findings: (1) most coordinated accounts are crude and retain the classic marks of automation, and the same criterion applied to Japanese-language accounts over the same month yields a share six times lower; (2) their content is overwhelmingly non-political; (3) regarding their reach, most reactions within large observable cascades originate from coordinated accounts themselves, while posts classified as potentially harmful or illegal material receive a comparatively high proportion of reactions from outside the Chinese-dominant population. We provide a longitudinal quantitative map of crude Chinese-language coordination appearing in X's Japanese-classified stream.

Explore related subjects

Keep this discovery

Explore connections, maps & timelines

BibTeXRIS

Kei Ichikawa, Bruno T. Sugano, Genta Toya, Wu Qianyun, Yasuhiro Hashimoto, Masashi Toyoda, Naoki Yoshinaga, Kazutoshi Sasahara. 2026-09-28. Crude, Commercial, and Self-Referential: Chinese-Language Coordinated Activity in Japanese-Language X. https://arxiv.org/abs/2610.00305

Cite the original work for its findings. Save a collection to share your selection of sources.

KEEP EXPLORING

Related papers

Beyond the Clique: Comparing Clique and Dowker Complexes for Co-occurrence Data in Learning Analytics

Learning analytics increasingly represents the relations in co-occurrence data (codes in a window of discourse, participants in a thread, tags on a post) as simplicial complexes and analyses them with persistent homology. The standard approach uses the clique complex. Because its simplices are determined by the pairwise network alone, it cannot distinguish "three elements co-occurred together" from "each of the three pairs co-occurred separately". The Dowker complex, in contrast, takes as simplices the sets of entities actually observed to co-occur. We compare the two. First, we show that, on the same 1-skeleton, the Dowker complex is a subcomplex of the clique complex, the map on first homology induced by the inclusion is surjective, and its kernel is generated by phantom triangles that never co-occurred; that is, the clique construction can only erase holes. We also give a criterion for agreement that can be checked on triples of groups. We then test this on real data. On four Stack Exchange data sets, 46-91% of clique triangles are phantom and 522-1,970 holes are erased. On the example data of the learning-analytics R packages tna/Nestimate, the first Betti number $β_1$ of the clique complex is 0 at every threshold, whereas the Dowker complex detects holes. Against a degree-preserving null model, the Dowker complex departs strongly in five of the six data sets. This difference is invisible at the fixed thresholds used in practice. We conclude that the construction should follow the data type and that, for observed groups, the Dowker complex is the appropriate choice. Code is available at https://github.com/igu-lab/beyond-the-clique.

cs.SI↗

Degree-Corrected Joint Matrix Factorization for Multilayer Community Detection

Multilayer networks allow the modeling of interactions between the same entities across different contexts, such as temporal observations, varying settings, or interactions of different types. The goal of community detection in multilayer networks is to identify groups of nodes exhibiting similar connectivity patterns, which may vary across layers. We propose a method based on a joint nonnegative symmetric matrix trifactorization for community detection in multilayer networks, where each graph is approximated by a nonnegative symmetric matrix trifactorization. Our approach enforces constraints on the factor matrices so that communities are disjoint and shared across layers, while allowing each layer to have its own connectivity patterns and node degrees. This flexibility enables the model to capture both local and global structural variations across layers. We also develop an algorithm to efficiently solve this problem. We evaluate multilayer community detection methods using the multilayer degree-corrected stochastic block model (MDCBM), a flexible framework for generating realistic multilayer graphs with heterogeneous degrees and varying connectivity patterns. Experiments show that our method reliably detects communities across diverse regimes, whereas existing state-of-the-art approaches are often limited by restrictive structural assumptions.

cs.SI↗

Out-of-Network Attention Dynamics on Bluesky

Personalized social media commonly relies on explicit follow graphs to shape what content users encounter; yet how much attention crosses ties they have not formed remains largely undocumented at scale. We study this question on Bluesky, a large decentralized microblogging platform whose default feed relies on a simple, reverse-chronological content recommender. We analyze 173 million user-author interactions (likes, reposts, replies, and quotes) collected from a near-complete platform dump between February and September 2023. We decompose each interaction by attention-path length (already followed, relayed by a followed account, reachable within two follow-hops, or beyond) and find that 74.5% of interactions reach the user through an account they already follow. Measured by distance in the follow graph rather than by route, 80.6% of interaction lands within two follow hops, far beyond the 22.4% an expected-degree null predicts. We then characterize how exploration varies across users and over tenure. A broad-reaching minority generates three quarters of all exploratory activity, while aggregate declines in exploration with tenure mask three distinct individual trajectories. Finally, attention reaching beyond two hops converts into new follow ties at less than one third the rate of two-hop-local exploratory attention. Together, these results depict a platform where out-of-network exploration is substantial in volume but strongly constrained by network proximity and unlikely to translate into new social ties.

cs.SI↗