arXiv.org
From Code Foundation Models to Agents and Applications: A Comprehensive Survey and Practical Guide to Code Intelligence
This survey comprehensively examines the lifecycle of code-focused large language models (LLMs), from data curation and pre-training to post-training, alignment, and deployment as autonomous agents. It analyzes both general-purpose LLMs (e.g., GPT-4, Claude, LLaMA) and code-specialized models (e.g., StarCoder, Code LLaMA, DeepSeek-Coder, QwenCoder),…
Jian Yang, Xianglong Liu, Weifeng Lv, Ken Deng, et al.- Published
- Nov 2025
- Upvotes
- 306
- Citations
- 12