TransforLearn: Interactive Visual Tutorial for the Transformer Model
Lin Gao, Zekai Shao, Ziqin Luo, Haibo Hu, Cagatay Turkay, Siming Chen
摘要
The widespread adoption of Transformers in deep learning, serving as the core framework for numerous large-scale language models, has sparked significant interest in understanding their underlying mechanisms. However, beginners face difficulties in comprehending and learning Transformers due to its complex structure and abstract data representation. We present TransforLearn, the first interactive visual tutorial designed for deep learning beginners and non-experts to comprehensively learn about Transformers. TransforLearn supports interactions for architecture-driven exploration and task-driven exploration, providing insight into different levels of model details and their working processes. It accommodates interactive views of each layer's operation and mathematical formula, helping users to understand the data flow of long text sequences. By altering the current decoder-based recursive prediction results and combining the downstream task abstractions, users can deeply explore model processes. Our user study revealed that the interactions of TransforLearn are positively received. We observe that TransforLearn facilitates users' accomplishment of study tasks and a grasp of key concepts in Transformer effectively.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper5
- Fine-Tuned Large Language Model for Visualization System: A Study on Self-Regulated Learning in EducationLin Gao, Jing Lu, Zekai Shao, Ziyue Lin 等IEEE VIS 2024 · 被引用 27 次
- Smartboard: Visual Exploration of Team Tactics with LLM AgentZiao Liu, Xiao Xie, Moqi He, Wenshuo Zhao 等IEEE VIS 2024 · 被引用 12 次
- Crowdsourced Think-Aloud StudiesZach Cutler, Lane Harrison, Carolina Nobre, Alexander LexCHI 2025 · 被引用 4 次
- ConceptViz: A Visual Analytics Approach for Exploring Concepts in Large Language ModelsHaoxuan Li, Zhen Wen, Qiqi Jiang, Chenxiao Li 等IEEE VIS 2025 · 被引用 3 次
- Transformer Explainer: Learning LLM Transformers with Interactive Visual Explanation and ExperimentationAeree Cho, Grace C. Kim, Alexander Karpekov, Seongmin Lee 等CHI 2026 · 被引用 2 次
它引用的顶会 Paper11
- Language Models are Few-Shot LearnersTom B. Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah 等NeurIPS 2020 · 被引用 64,255 次
- Training language models to follow instructions with human feedbackLong Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida 等NeurIPS 2022 · 被引用 24,707 次
- An Image is Worth 16x16 Words: Transformers for Image Recognition at ScaleAlexey Dosovitskiy, Lucas Beyer, Alexander Kolesnikov, Dirk Weissenborn 等ICLR 2021 · 被引用 21,477 次
- What is AI Literacy? Competencies and Design ConsiderationsDuri Long, Brian MagerkoCHI 2020 · 被引用 2,947 次
- CNN Explainer: Learning Convolutional Neural Networks with Interactive VisualizationZijie J. Wang, Robert Turko, Omar Shaikh, Haekyu Park 等IEEE VIS 2020 · 被引用 341 次
相关 Paper
- Transformers from an Optimization PerspectiveYongyi Yang, Zengfeng Huang, David P. WipfNeurIPS 2022 · 被引用 44 次
- AttentionViz: A Global View of Transformer AttentionCatherine Yeh, Yida Chen, Aoyu Wu, Cynthia Chen 等IEEE VIS 2023 · 被引用 78 次
- Tuformer: Data-driven Design of Transformers for Improved Generalization or EfficiencyXiaoyu Liu, Jiahao Su, Furong HuangICLR 2022 · 被引用 8 次
- Teaching Temporal Logics to Neural NetworksChristopher Hahn, Frederik Schmitt, Jens U. Kreber, Markus Norman Rabe 等ICLR 2021 · 被引用 78 次
- Visformer: The Vision-friendly TransformerZhengsu Chen, Lingxi Xie, Jianwei Niu, Xuefeng Liu 等ICCV 2021 · 被引用 293 次
