Drop Clause: Enhancing Performance, Robustness and Pattern Recognition Capabilities of the Tsetlin Machine
Jivitesh Sharma, Rohan Kumar Yadav, Ole-Christoffer Granmo, Lei Jiao
摘要
Logic-based machine learning has the crucial advantage of transparency. However, despite significant recent progress, further research is needed to close the accuracy gap between logic-based architectures and deep neural network ones. This paper introduces a novel variant of the Tsetlin machine (TM) that randomly drops clauses, the logical learning element of TMs. In effect, TM with Drop Clause ignores a random selection of the clauses in each epoch, selected according to a predefined probability. In this way, the TM learning phase becomes more diverse. To explore the effects that Drop Clause has on accuracy, training time and robustness, we conduct extensive experiments on nine benchmark datasets in natural language processing (IMDb, R8, R52, MR, and TREC) and image classification (MNIST, Fashion MNIST, CIFAR-10, and CIFAR-100). Our proposed model outperforms baseline machine learning algorithms by a wide margin and achieves competitive performance compared with recent deep learning models, such as BERT-Large and AlexNet-DFA. In brief, we observe up to 10% increase in accuracy and 2× to 4× faster in learning than those of the standard TM. We visualize the patterns learnt by Drop Clause TM in the form of heatmaps and show evidence of the ability of drop clause to learn more unique and discriminative patterns. We finally evaluate how Drop Clause affects learning robustness by introducing corruptions and alterations in the image/language test data, which exposes increased learning robustness.
问问这篇 Paper
智能体会读完全文。
Lune 把这篇 Paper 索引到了每一个公式,引用它的顶会 Paper 也一样。你提问,回答直接引用原文。
引用它的顶会 Paper2
- Generalized Convergence Analysis of Tsetlin Automaton Based Algorithms: A Probabilistic Approach to Concept LearningMohamed-Bachir Belaid, Jivitesh Sharma, Lei Jiao, Ole-Christoffer Granmo 等AAAI 2025
- Convergence Analysis of Tsetlin Machines under Noise-Free and Noisy Training Conditions: From 2 Bits to k BitsXuan Zhang, Lei Jiao, Ole-Christoffer GranmoICLR 2026
它引用的顶会 Paper5
- Simple Spectral Graph ConvolutionHao Zhu, Piotr KoniuszICLR 2021 · 被引用 352 次
- Human-Level Interpretable Learning for Aspect-Based Sentiment AnalysisRohan Kumar Yadav, Lei Jiao, Ole-Christoffer Granmo, Morten GoodwinAAAI 2021 · 被引用 96 次
- Efficient Exact Verification of Binarized Neural NetworksKai Jia, Martin C. RinardNeurIPS 2020 · 被引用 70 次
- Feature Projection for Improved Text ClassificationQi Qin, Wenpeng Hu, Bing LiuACL 2020 · 被引用 66 次
- Massively Parallel and Asynchronous Tsetlin Machine Architecture Supporting Almost Constant-Time ScalingKuruge Darshana Abeyrathna, Bimal Bhattarai, Morten Goodwin, Saeed Rahimi Gorji 等ICML 2021 · 被引用 45 次
相关 Paper
- MaxSAT-Based Compression for Tsetlin MachinesStefan SzeiderICML 2026
- Accelerating Training of Transformer-Based Language Models with Progressive Layer DroppingMinjia Zhang, Yuxiong HeNeurIPS 2020 · 被引用 126 次
- Tsetlin Machine for Solving Contextual Bandit ProblemsRaihan Seraj, Jivitesh Sharma, Ole-Christoffer GranmoNeurIPS 2022 · 被引用 19 次
- NeuroSelect: Learning to Select Clauses in SAT SolversHongduo Liu, Peng Xu, Yuan Pu, Lihao Yin 等DAC 2024 · 被引用 2 次
- The Surprising Effectiveness of Test-Time Training for Few-Shot LearningEkin Akyürek, Mehul Damani, Adam Zweiger, Linlu Qiu 等ICML 2025
