cs.CLSep 24, 2026

R-DEIM Net: An Efficient Rationale-Augmented Dual-Expert Interaction Model for Paraphrase Detection

Authors: Pushp, Vaibhav Prajapati, Himangshu Sarma

Organizations: Indian Institute of Information Technology (IIIT), Sri City, India · University of Technology Nuremberg (UTN), Germany

Abstract

Recent advances in paraphrase detection reveal a fundamental trade-off: large language models achieve high accuracy but require high computation, while efficient Siamese-BERT variants offer practical scalability with reduced transparency in rationale generation. We present R-DEIM Net, a 76M-parameter dual-expert architecture exploring whether moderate-scale models can achieve competitive accuracy on paraphrase detection while enabling human-readable rationale generation. The architecture combines two specialized components: an Interaction Expert that captures token-level similarity patterns through multi-scale 2D convolutions and attention head allowing variable input length, and a Reasoning Expert that uses a Flan-T5-small decoder to generate rationales as auxiliary supervision. Rather than re-encoding generated text, we extract and pool decoder hidden states as complementary features for classification. On the Quora Question Pairs dataset, R-DEIM Net achieves 90.07% accuracy and 90.16% F1-score via 10-fold cross-validation. This represents competitive performance with strong transformer-based baselines (e.g., MFAE BERT: 90.54% accuracy) and recent large language model based approaches (LLaMA-70B) while using a substantially smaller parameter budget. The model generates rationales alongside predictions, providing potential for auxiliary human-readable descriptions.

Figures & tables

Appendix figures & tables2 assets

Supplementary material from the paper’s appendix.

Appendix

Explore similar work

CardsList
  1. CAPS: Cascaded Adaptive Pairwise Selection for Efficient Parallel Reasoning

    May 15, 2026Fangzhou Lin, Shuo Xing, Peiran Li +6Reasoning Benchmark

  2. ThreadWeaver: Adaptive Threading for Efficient Parallel Reasoning in Language Models

    Nov 24, 2025Long Lian, Sida Wang, Felix Juefei-Xu +7LLM Reasoning StrategiesInference Latency

  3. Cut Your Losses! Learning to Prune Paths Early for Efficient Parallel Reasoning

    Apr 17, 2026Jiaxi Bi, Tongxu Luo, Wenyu Du +2Large Reasoning Models