Behavioral Homophily in Social Media via Inverse Reinforcement Learning: A Reddit Case Study
Lanqin Yuan, Philipp J. Schneider, Marian-Andrei Rizoiu
Abstract
Online communities play a critical role in shaping societal discourse and influencing collective behavior in the real world. The tendency for people to connect with others who share similar characteristics and views, known as homophily, plays a key role in the formation of echo chambers which further amplify polarization and division. Existing works examining homophily in online communities traditionally infer it using content- or adjacency-based approaches, such as constructing explicit interaction networks or performing topic analysis. These methods fall short for platforms where interaction networks cannot be easily constructed and fail to capture the complex nature of user interactions across the platform. This work introduces a novel approach for quantifying user homophily. We first use an Inverse Reinforcement Learning (IRL) framework to infer users' policies, then use these policies as a measure of behavioral homophily. We apply our method to Reddit, conducting a case study across 5.9 million interactions over six years, demonstrating how this approach uncovers distinct behavioral patterns and user roles that vary across different communities. We further validate our behavioral homophily measure against traditional content-based homophily, offering a powerful method for analyzing social media dynamics and their broader societal implications. We find, among others, that users can behave very similarly (high behavioral homophily) when discussing entirely different topics like soccer vs e-sports (low topical homophily), and that there is an entire class of users on Reddit whose purpose seems to be to disagree with others.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext 832f4119-83f3-45df-a082-36ad5a90f52fCited by top-tier papers1
Ask how each one uses itBuilds on3
- DeBERTaV3: Improving DeBERTa using ELECTRA-Style Pre-Training with Gradient-Disentangled Embedding SharingPengcheng He, Jianfeng Gao, Weizhu ChenICLR 2023 · 394 citations
- Evidence of Demographic rather than Ideological Segregation in News Discussion on RedditCorrado Monti, Jacopo D'Ignazi, Michele Starnini, Gianmarco De Francisci MoralesWWW 2023 · 25 citations
- Analyzing the Strategy of Propaganda using Inverse Reinforcement Learning: Evidence from the 2022 Russian Invasion of UkraineDominique Geissler, Stefan FeuerriegelCSCW 2024 · 4 citations
Related papers
- Transformer-Based Quantification of the Echo Chamber Effect in Online CommunitiesVahid Ghafouri, Faisal Alatawi, Mansooreh Karami, Jose Such et al.CSCW 2024 · 4 citations
- Navigating Multidimensional Ideologies with Reddit's Political Compass: Economic Conflict and Social AffinityErnesto Colacrai, Federico Cinus, Gianmarco De Francisci Morales, Michele StarniniWWW 2024 · 7 citations
- Deliberate Exposure to Opposing Views and Its Association with Behavior and Rewards on Political CommunitiesAlexandros EfstratiouWWW 2024 · 2 citations
- Measuring User-Moderator Alignment on r/ChangeMyViewVinay Koshy, Tanvi Bajpai, Eshwar Chandrasekharan, Hari Sundaram et al.CSCW 2023 · 24 citations
- Online Platforms and the Fair Exposure Problem under HomophilyJakob Schoeffer, Alexander Ritchie, Keziah Naggita, Faidra Monachou et al.AAAI 2023 · 5 citations
