HW-Spy: Handwriting Inference by Tracing Pen-Tail Movements
Long Huang, Kang G. Shin
摘要
While keyboard typing has been the most common way of inputting texts, handwriting still plays an important role in generating, inputting, or recording information like filling out essential/private forms. Considerable research has been done to identify and demonstrate the risk of keystroke-inference attacks. However, little has been done on handwriting inference despite its high risk of leaking sensitive information. To assess this under-explored risk of information leakage, we present a novel handwriting-inference attack, called HW-Spy, by tracing the victim's pen-tail movements when both the pen tip and the writing surface are outside the view of the attacker's camera, which usually happens when the victim is multitasking during an online meeting, when the victim's writing scene (in a public space) is recorded by a remote camera, or when the victim's writing behaviors are captured by the surveillance camera in a bank/dealership/realty office.
In particular, we apply image segmentation to the recorded video frames of the victim's writing activities and extract the victim's pentail movements as a 2D coordinate sequence. We then identify the stroke-associated movements from the recorded pen's in-air video frames using a 1D U-Net model trained for stroke mask prediction and segment the characters based on the thus-derived motion features. The pen-tail's coordinate segments are then fed into a Long Short-Term Memory (LSTM) network to reconstruct the actual handwriting, which is processed further by a transformer-based model to infer the hand-written content. Our extensive experimentation shows HW-Spy to achieve an accuracy, up to 84.2%, of personalized handwriting inference and a comparable accuracy, up to 79.5%, of non-personalized handwriting inference.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
它引用的顶会 Paper14
- Learning Transferable Visual Models From Natural Language SupervisionAlec Radford, Jong Wook Kim, Chris Hallacy, Aditya Ramesh 等ICML 2021 · 被引用 47,906 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- Segment AnythingAlexander Kirillov, Eric Mintun, Nikhila Ravi, Hanzi Mao 等ICCV 2023 · 被引用 13,211 次
- Informer: Beyond Efficient Transformer for Long Sequence Time-Series ForecastingHaoyi Zhou, Shanghang Zhang, Jieqi Peng, Shuai Zhang 等AAAI 2021 · 被引用 7,289 次
- When CSI Meets Public WiFi: Inferring Your Mobile Phone Password via WiFi SignalsMengyuan Li, Yan Meng, Junyi Liu, Haojin Zhu 等CCS 2016 · 被引用 213 次
相关 Paper
- Towards a General Video-based Keystroke Inference AttackZhuolin Yang, Yuxin Chen, Zain Sarwar, Hadleigh Schwartz 等USENIX Security 2023
- ArmSpy: Video-assisted PIN Inference Leveraging Keystroke-induced Arm Posture ChangesYuefeng Chen, Yicong Du, Chunlong Xu, Yanghai Yu 等INFOCOM 2022 · 被引用 4 次
- VISIBLE: Video-Assisted Keystroke Inference from Tablet Backside MotionJingchao Sun, Xiaocong Jin, Yimin Chen, Jinxue Zhang 等NDSS 2016 · 被引用 72 次
- A Keylogging Inference Attack on Air-Tapping Keyboards in Virtual EnvironmentsÜlkü Meteriz-Yildiran, Necip Fazil Yildiran, Amro Awad, David MohaisenIEEE VR 2022 · 被引用 40 次
- EyeTell: Video-Assisted Touchscreen Keystroke Inference from Eye MovementsYimin Chen, Tao Li, Rui Zhang, Yanchao Zhang 等S&P 2018 · 被引用 60 次
