VendorLink: An NLP approach for Identifying & Linking Vendor Migrants & Potential Aliases on Darknet Markets
Vageesh Saxena, Nils Rethmeier, Gijs van Dijck, Gerasimos Spanakis
摘要
The anonymity on the Darknet allows vendors to stay undetected by using multiple vendor aliases or frequently migrating between markets. Consequently, illegal markets and their connections are challenging to uncover on the Darknet. To identify relationships between illegal markets and their vendors, we propose VendorLink, an NLP-based approach that examines writing patterns to verify, identify, and link unique vendor accounts across text advertisements (ads) on seven public Darknet markets. In contrast to existing literature, Ven-dorLink utilizes the strength of supervised pretraining to perform closed-set vendor verification, open-set vendor identification, and lowresource market adaption tasks. Through Ven-dorLink, we uncover (i) 15 migrants and 71 potential aliases in the Alphabay-Dreams-Silk dataset, (ii) 17 migrants and 3 potential aliases in the Valhalla-Berlusconi dataset, and (iii) 75 migrants and 10 potential aliases in the Traderoute-Agora dataset. Altogether, our approach can help Law Enforcement Agencies (LEA) make more informed decisions by verifying and identifying migrating vendors and their potential aliases on existing and Low-Resource (LR) emerging Darknet markets. 1
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了最后一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- IDTraffickers: An Authorship Attribution Dataset to link and connect Potential Human-Trafficking Operations on Text Escort AdvertisementsVageesh Saxena, Benjamin Bashpole, Gijs van Dijck, Gerasimos SpanakisEMNLP 2023 · 被引用 2 次
- Covering Cracks in Content Moderation: Delexicalized Distant Supervision for Illicit Drug Jargon DetectionMinkyoo Song, Eugene Jang, Jaehan Kim, Seungwon ShinKDD 2025
它引用的顶会 Paper10
- On the Sentence Embeddings from Pre-trained Language ModelsBohan Li, Hao Zhou, Junxian He, Mingxuan Wang 等EMNLP 2020 · 被引用 538 次
- PromptBERT: Improving BERT Sentence Embeddings with PromptsTing Jiang, Jian Jiao, Shaohan Huang, Zihan Zhang 等EMNLP 2022 · 被引用 148 次
- Improved Text Classification via Contrastive Adversarial TrainingLin Pan, Chung-Wei Hang, Avirup Sil, Saloni PotdarAAAI 2022 · 被引用 115 次
- Authorship Attribution for Neural Text GenerationAdaku Uchendu, Thai Le, Kai Shu, Dongwon LeeEMNLP 2020 · 被引用 110 次
- Plug and Prey? Measuring the Commoditization of Cybercrime via Online Anonymous MarketsRolf van Wegberg, Samaneh Tajalizadehkhoob, Kyle Soska, Ugur Akyazi 等USENIX Security 2018 · 被引用 97 次
相关 Paper
- eDarkFind: Unsupervised Multi-view Learning for Sybil Account DetectionRamnath Kumar, Shweta Yadav, Raminta Daniulaityte, Francois R. Lamy 等WWW 2020 · 被引用 28 次
- SYSML: StYlometry with Structure and Multitask Learning: Implications for Darknet Forum Migrant AnalysisPranav Maneriker, Yuntian He, Srinivasan ParthasarathyEMNLP 2021 · 被引用 6 次
- Go See a Specialist? Predicting Cybercrime Sales on Online Anonymous Markets from Vendor and Product CharacteristicsRolf van Wegberg, Fieke Miedema, Ugur Akyazi, Arman Noroozian 等WWW 2020 · 被引用 14 次
- DarkBERT: A Language Model for the Dark Side of the InternetYoungjin Jin, Eugene Jang, Jian Cui, Jin-Woo Chung 等ACL 2023 · 被引用 41 次
- Uncovering and Mitigating the Hidden Chasm: A Study on the Text-Text Domain Gap in Euphemism IdentificationYuxue Hu, Junsong Li, Mingmin Wu, Zhongqiang Huang 等AAAI 2024 · 被引用 1 次
