Do Papers Tell the Whole Story? A Benchmark and Framework for Uncovering Hidden Implementation Gaps in Bioinformatics
cs.LG, cs.SE
Submitted: 2026-03-23
Updated: 2026-09-28
DOI: 10.1093/bib/bbag509
Code: https://github.com/Ricardo1998-Xu/BioCon
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Terminology
Sources
- Retrieval-Augmented Code Generation: A Survey with Focus on Repository-Level Approaches
- A Survey on Large Language Models for Software Engineering
- PseudoBridge: Pseudo Code as the Bridge for Better Semantic and Logic Alignment in Code Retrieval
- Towards Realistic Project-Level Code Generation via Multi-Agent Collaboration and Semantic Architecture Modeling
- Code Representation Learning At Scale
- Investigating the Impact of Code Comment Inconsistency on Bug Introducing
- Code Comment Inconsistency Detection with BERT and Longformer
- SciCoQA: Quality Assurance for Scientific Paper--Code Alignment
- CodeGen: An Open Large Language Model for Code with Multi-Turn Program Synthesis
Related papers
- Polynomial-Augmented Neural Networks (PANNs) with Weak Orthogonality Constraints for Enhanced Function and PDE Approximation
- AIRL-S: Unifying Reinforcement Learning and Search-Based Test-Time Scaling via Adversarial Inverse Reinforcement Learning
- Transformers as Bayesian In-Context Experimenters: Smoothness-Adaptive Efficient ATE Estimation
- Convergence issues in Relational Concept Analysis based on AOC-posets
- Beliefs Beyond Posteriors: Local-Consistency Optimisation for Bayesian Neural Networks
- Understanding Diffusion Models via Ratio-Based Function Approximation with SignReLU Networks