cs.IROct 1, 2026

From Rules to Neural Graphs: Scalable Structured Prediction for Patent Prior Art Search

Authors: Nikolai Zenovkin, Sebastian Björkqvist

Organizations: IPRally Technologies Oy, Helsinki, Finland

Abstract

Patent search requires processing documents routinely exceeding tens of thousands of tokens. Most neural retrieval approaches operate on truncated inputs, limiting their effectiveness. Graph-based retrieval addresses this by representing each patent as a structured invention graph, but constructing these graphs relies on brittle rule-based parsers. We present the neural parser, which adapts biaffine attention from dependency parsing to predict invention graphs directly from patent text. Our local biaffine attention restricts pairwise scoring to a sliding window, reducing complexity from O(n2)O(n^2) to O(n⋅w)O(n \cdot w). Since local and global scoring share the same weights, the model trains on short sequences and deploys on documents exceeding 40,000 tokens without retraining. Distilled from 1 million rule-parsed documents, it surpasses its teacher at 3×\times lower inference cost: neural graphs improve citation recall by 0.5% on short queries and 1.1% on full documents in a downstream Graph Transformer retrieval system.

Figures & tables

Explore similar work

CardsList
  1. Citation-Driven Multi-View Training for Patent Embeddings: QaECTER and Sophia-Bench

    Apr 24, 2026Younes Djemmal, You Zuo, Kim Gerdes +1PatentText Embeddings

  2. Heterogeneous Dependency Graph-Guided Attentionfor Patent Representation Learning

    May 11, 2026Yongmin Yoo, Qiongkai Xu, Zhangkai Wu +1PatentInterleaved Graph Attention

  3. Benchmarking Patent Embeddings: A Multi-Task Evaluation of 22 Models Across Retrieval, Classification, and Clustering

    May 22, 2026Amirhossein Yousefiramandi, Ciaran CooneyText EmbeddingsClustering