Towards Comprehensive Patent Approval Predictions: Beyond Traditional Document Classification
Xiaochen Gao, Zhaoyi Hou, Yifei Ning, Kewen Zhao, Beilei He, Jingbo Shang, Vish Krishnan
Abstract
Predicting the approval odds of a patent application is a challenging problem involving multiple factors. The most important factor is arguably the novelty -35 U.S. Code § 102 rejects applications that are not sufficiently differentiated from prior art. Novelty evaluation distinguishes the patent approval prediction from conventional document classification -toosimilar newer submissions are considered as not novel and would receive the opposite label, thus confusing standard document classifiers (e.g., BERT). To address this issue, we propose a novel framework AISeer that unifies the document classifier with handcrafted features, particularly time-dependent novelty scores. Specifically, we formulate the novelty scores by comparing each application with millions of prior art using a hybrid of efficient filters and a neural bi-encoder. Moreover, we impose a new regularization term into the classification objective to enforce the monotonic change of approval prediction w.r.t. novelty scores, From extensive experiments on a large-scale USPTO dataset, we find that standard BERT fine-tuning can partially learn the correct relationship between novelty and approvals from inconsistent data. However, our time-dependent novelty feature and other handcrafted features offer a significant boost on top of it. Also, our monotonic regularization, while shrinking the search space, can drive the optimizer to better local optima, yielding a further small performance gain.
Ask about this paper
Your agent reads all of it.
Lune indexed this paper to the last equation, along with the top-tier papers that cite it. Ask a question and the answer quotes them.
Cited by top-tier papers1
Ask how each one uses itBuilds on1
Related papers
- Is It Novel and Why? Fine-Grained Patent Novelty Prediction Based on Passage RetrievalValentin Knappich, Anna Hätty, Simon Razniewski, Annemarie FriedrichSIGIR 2026 · 1 citation
- 1+1>2: Integrating Deep Code Behaviors with Metadata Features for Malicious PyPI Package DetectionXiaobing Sun, Xingan Gao, Sicong Cao, Lili Bo et al.ASE 2024 · 3 citations
- RATE: Reviewer Profiling and Annotation-free Training for Expertise Ranking in Peer Review SystemsWeicong Liu, Zixuan Yang, Yibo Zhao, Xiang LiACL 2026
- A Generate-and-Rank Framework with Semantic Type Regularization for Biomedical Concept NormalizationDongfang Xu, Zeyu Zhang, Steven BethardACL 2020 · 43 citations
- SciMine: An Efficient Systematic Prioritization Model Based on Richer Semantic InformationFang Guo, Yun Luo, Linyi Yang, Yue ZhangSIGIR 2023 · 3 citations
