cs.LGMar 23, 2026

Do Papers Tell the Whole Story? A Benchmark and Framework for Uncovering Hidden Implementation Gaps in Bioinformatics

Authors: Tianxiang Xu, Xiaoyan Zhu, Xin Lai, Xin Lian, Sizhe Dang, Hangyu Cheng, Jiayin Wang

Organizations: Xi’an Jiaotong University

Abstract

As bioinformatics software is increasingly applied across a broader range of scenarios and the rapid development of large language models (LLMs) further lowers the barriers to software use and development, the composition of the bioinformatics research community is undergoing substantial change. Consequently, a growing number of researchers require a deeper understanding of methodological details and software behavior. In this context, systematically analyzing the relationship between paper descriptions and code implementations is emerging as an important new challenge in the field. To address this challenge, we introduce paper-code consistency analysis as a new research perspective and construct BioCon, the first benchmark dataset for paper-code consistency analysis in bioinformatics. Furthermore, we develop a unified cross-modal analysis framework to systematically investigate this problem from three perspectives: sentence-level detection, cross-modal retrieval, and project-level assessment. Experimental results demonstrate that the proposed framework can effectively model the semantic relationships between scientific publications and software implementations. Further case studies reveal that paper-code inconsistency is not a single phenomenon but arises from multiple underlying causes, among which Author-Perceived Non-Essential Details represents the most prevalent category. These findings suggest that paper-code consistency analysis is not merely a technical problem but also raises broader discussions regarding knowledge dissemination, the boundaries of code disclosure, and community norms. We hope that this work will encourage the bioinformatics community to re-examine the relationship between scientific publications and software implementations while providing a foundation for future research in paper-code consistency analysis.

Figures & tables

Explore similar work

CardsList
  1. Supporting Workflow Reproducibility by Linking Bioinformatics Tools across Papers and Executable Code

    Mar 9, 2026Clémence Sebe, Olivier Ferret, Aurélie Névéol +3BioinformaticsAgentic Workflow Design

  2. Prompt-to-Paper: Agentic AI System for Bioinformatics

    Jul 5, 2026Ramsha Kamran, Maheera Amjad, Zartasha Mustansar +3Bioinformatics

  3. PRISMA-LLM: An Empirical Reporting Framework for AI-Assisted Systematic Reviews

    Sep 14, 2026Miguel Zabaleta, Baihan LinSystematic ReviewPeer Review