Code Language Models

Momentum

10 papers in the last four weeks, up 150% on the four weeks before. 0.1% of all new papers.

Jul 13Week of Sep 28

Latest papers 85

All topics
CardsList
  1. LLM-Based Code Documentation Generation and Multi-Judge Evaluation

    May 11, 2026Ikbel Ghrab, Mohamed Dhieb, Ismail Khenissi +1LLM EvaluationLLM-as-a-Judge

  2. An Execution-Verified Multi-Language Benchmark for Code Semantic Reasoning

    May 10, 2026Yikun Li, Jinfeng Jiang, Ting Zhang +7LLM EvaluationCode Language Models

  3. Standing on the Shoulders of Giants: Stabilized Knowledge Distillation for Cross--Language Code Clone Detection

    May 4, 2026Mohamad Khajezade, Fatemeh H. Fard, Mohamed Sami ShehataLanguage Model DistillationKnowledge Distillation

  4. Code World Model Preparedness Report

    May 1, 2026Daniel Song, Peter Ney, Cristina Menghini +21AI Safety EvaluationLLM Safety Evaluation

  5. To Diff or Not to Diff? Structure-Aware and Adaptive Output Formats for Efficient LLM-based Code Editing

    Apr 30, 2026Wei Cheng, Yongchang Cao, Chen Shen +4Code GenerationLLM Inference Acceleration

  6. PLMGH: What Matters in PLM-GNN Hybrids for Code Classification and Vulnerability Detection

    Apr 28, 2026Mohamed Taoufik Kaouthar El Idrissi, Edward Zulkoski, Mohammad HamdaqaGraph Neural NetworksSoftware Vulnerability Detection

  7. Large Language Models for Multilingual Code Intelligence: A Survey

    Apr 27, 2026Chao Jiang, Dugang Liu, Cheng Wen +6Code TranslationCode Generation

  8. Query2Diagram: Answering Developer Queries with UML Diagrams

    Apr 26, 2026Oleg Baryshnikov, Anton M. Alekseev, Sergey I. NikolenkoSoftware EngineeringSoftware Reverse Engineering

  9. Parallel-SFT: Improving Zero-Shot Cross-Programming-Language Transfer for Code RL

    Apr 22, 2026Zhaofeng Wu, Shiqi Wang, Boya Peng +5Supervised Fine-TuningRL for Code Generation

  10. The Path Not Taken: Duality in Reasoning about Program Execution

    Apr 22, 2026Eshgin Hasanov, Md Mahadi Hassan Sibat, Santu Karmaker +1LLM EvaluationCode Language Models

  11. WebCompass: Towards Multimodal Web Coding Evaluation for Code Language Models

    Apr 20, 2026Xinping Lei, Xinyu Che, Junqi Xiong +16LLM Agent EvaluationCode Language Models

  12. CodePivot: Bootstrapping Multilingual Transpilation in LLMs via Reinforcement Learning without Parallel Corpora

    Apr 20, 2026Shangyu Li, Juyong Jiang, Meibo Ren +7Code TranslationRL for Language Models

  13. Understanding Secret Leakage Risks in Code LLMs: A Tokenization Perspective

    Apr 20, 2026Meifang Chen, Zhe Yang, Huang Nianchen +4Data LeakageMemorization in Language Models

  14. SynthFix: Adaptive Neuro-Symbolic Code Vulnerability Repair

    Apr 19, 2026Yifan Zhang, Jieyu Li, Kexin Pei +2Automated Program RepairNeuro-Symbolic AI

  15. Configuration Over Selection: Hyperparameter Sensitivity Exceeds Model Differences in Open-Source LLMs for RTL Generation

    Apr 18, 2026Minghao Shao, Zeng Wang, Weimin Fu +5LLM EvaluationLanguage Model Generation Evaluation

  16. Improving LLM Code Reasoning via Semantic Equivalence Self-Play with Formal Verification

    Apr 18, 2026Antonio Valerio Miceli Barone, Poon Tsz NokAdversarial TrainingSelf-Play RL

  17. SWE-Tester: Training Open-Source LLMs for Issue Reproduction in Real-World Repositories

    Jan 20, 2026Aditya Bharat Soni, Rajat Ghosh, Vaishnavi Bhargava +2LLM Post-TrainingAutomated Test Generation

  18. AgentPack: A Dataset of Code Changes, Co-Authored by Agents and Humans

    Sep 26, 2025Yangtian Zi, Zixuan Wu, Aleksander Boruch-Gruszecki +2Software Engineering AgentsHuman-AI Collaboration

  19. EyeMulator: Improving Code Language Models by Mimicking Human Visual Attention

    Aug 22, 2025Yifan Zhang, Chen Huang, Yueke Zhang +5Code TranslationLLM Fine-Tuning

  20. Are AI Coders Snitches? An Empirical Study of Pretraining Data Detection on Code Large Language Models

    Jul 23, 2025Tianlin Li, Yunxiang Wei, Zhiming Li +5LLM EvaluationCode Language Models

  21. RedCoder: Automated Multi-Turn Red Teaming for Code LLMs

    Jun 25, 2025Wenjie Jacky Mo, Qin Liu, Xiaofei Wen +5Adversarial Prompt GenerationLLM Red Teaming

  22. Is Your Benchmark Still Useful? Dynamic Benchmarking for Code Language Models

    Mar 9, 2025Batu Guan, Xiao Wu, Yuanyuan Yuan +1Benchmark DesignData Contamination in Language Models