STEFANN: Scene Text Editor Using Font Adaptive Neural Network
Prasun Roy, Saumik Bhattacharya, Subhankar Ghosh, Umapada Pal
Abstract
Textual information in a captured scene plays an important role in scene interpretation and decision making. Though there exist methods that can successfully detect and interpret complex text regions present in a scene, to the best of our knowledge, there is no significant prior work that aims to modify the textual information in an image. The ability to edit text directly on images has several advantages including error correction, text restoration and image reusability. In this paper, we propose a method to modify text in an image at character-level. We approach the problem in two stages. At first, the unobserved character (target) is generated from an observed character (source) being modified. We propose two different neural network architectures -(a) FANnet to achieve structural consistency with source font and (b) Colornet to preserve source color. Next, we replace the source character with the generated character maintaining both geometric and visual consistency with neighboring characters. Our method works as a unified platform for modifying text in images. We present the effectiveness of our method on COCO-Text and ICDAR datasets both qualitatively and quantitatively.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Your agent calls
Luneget_paper_fulltext
Free to start. No credit card required.
Terminal
Install the CLIlune papers fulltext c1baa90c-b96f-4fd1-a13d-98204e55655aCited by top-tier papers22
- DiffUTE: Universal Text Editing Diffusion ModelHaoxing Chen, Zhuoer Xu, Zhangxuan Gu, Jun Lan et al.NeurIPS 2023 · 61 citations
- Brush Your Text: Synthesize Any Scene Text on Images via Diffusion ModelLingjun Zhang, Xinyuan Chen, Yaohui Wang, Yue Lu et al.AAAI 2024 · 54 citations
- Exploring Stroke-Level Modifications for Scene Text EditingYadong Qu, Qingfeng Tan, Hongtao Xie, Jianjun Xu et al.AAAI 2023 · 51 citations
- Wiffract: a new foundation for RF imaging via edge tracingAnurag Pallaprolu, Belal Korany, Yasamin MostofiMobiCom 2022 · 35 citations
- De-rendering Stylized TextsWataru Shimoda, Daichi Haraguchi, Seiichi Uchida, Kota YamaguchiICCV 2021 · 34 citations
Related papers
- SwapText: Image Based Texts Transfer in ScenesQiangpeng Yang, Jun Huang, Wei LinCVPR 2020
- Self-Supervised Cross-Language Scene Text EditingFuxiang Yang, Tonghua Su, Xiang Zhou, Donglin Di et al.ACM MM 2023 · 2 citations
- Show, Edit and Tell: A Framework for Editing Image CaptionsFawaz Sammani, Luke Melas-KyriaziCVPR 2020
- ManiTrans: Entity-Level Text-Guided Image Manipulation via Token-wise Semantic Alignment and GenerationJianan Wang, Guansong Lu, Hang Xu, Zhenguo Li et al.CVPR 2022 · 15 citations
- Text-Guided Neural Image InpaintingLisai Zhang, Qingcai Chen, Baotian Hu, Shuoran JiangACM MM 2020 · 53 citations
