BabelCoder: Agentic Code Translation with Specification Alignment
cs.SE, cs.AI
Submitted: 2025-12-07
Updated: 2026-09-24
Code: https://github.com/anonprox/babelcoder
Terminology
Sources
- AVATAR: A Parallel Corpus for Java-Python Program Translation
- Enhancing Code Translation in Language Models with Few-Shot Learning via Retrieval-Augmented Generation
- ExeCoder: Empowering Large Language Models with Executability Representation for Code Translation
- AlphaTrans: A Neuro-Symbolic Compositional Approach for Repository-Level Code Translation and Validation
- Self-Organized Agents: A LLM Multi-Agent Framework toward Ultra Large-Scale Code Generation and Optimization
- Unsupervised Translation of Programming Languages
- An Exploratory Study on Fine-Tuning Large Language Models for Secure Code Generation
- Secure-Instruct: An Automated Pipeline for Synthesizing Instruction-Tuning Datasets Using LLMs for Secure Code Generation
- RobuNFR: Evaluating the Robustness of Large Language Models on Non-Functional Requirements Aware Code Generation
- Beyond Functional Correctness: Exploring Hallucinations in LLM-Generated Code
- InterTrans: Leveraging Transitive Intermediate Translations to Enhance LLM-based Code Translation
- Evolving Triple Knowledge-Augmented LLMs for Code Translation in Repository Context
- CATCODER: Repository-Level Code Generation with Relevant Code and Type Context
- CodeNet: A Large-Scale AI for Code Dataset for Learning a Diversity of Coding Tasks
- A Multi-Language Perspective on the Robustness of LLM Code Generation
- Leveraging Automated Unit Tests for Unsupervised Code Translation
- Specification-Driven Code Translation Powered by Large Language Models: How Far Are We?
- Code Translation with Compiler Representations
- Teaching Code LLMs to Use Autocompletion Tools in Repository-Level Code Generation
- RepoTransBench: A Real-World Multilingual Benchmark for Repository-Level Code Translation
Related papers
- Falsification-Based Verification of LLM-Generated Optimization Models: Sound Test Batteries and Their Detection Limits
- GitSkills: A Dataset of Agent Skills on GitHub
- SABER: Benchmarking Operational Safety of LLM Coding Agents in Stateful Project Workspaces
- PackMonitor: Enabling Zero Package Hallucinations Through Decoding-Time Monitoring
- IntentCoding: Amplifying User Intent in Code Generation
- Incentives and Outcomes in Bug Bounties