Granite Code Models: A Family of Open Foundation Models for Code Intelligence

Mayank Mishra,Matt Stallone,Gaoyuan Zhang,Yikang Shen,Aditya Prasad,Adriana Meza Soria,Michele Merler,Parameswaran Selvam,Saptha Surendran,Shivdeep Singh,Manish Sethi,Xuan-Hong Dang,Pengyuan Li,Kun-Lung Wu,Syed Zawad,Andrew Coleman,Matthew White,Mark Lewis,Raju Pavuluri,Yan Koyfman,Boris Lublinsky,Maximilien de Bayser,Ibrahim Abdelaziz,Kinjal Basu,Mayank Agarwal,Yi Zhou,Chris Johnson,Aanchal Goyal,Hima Patel,Yousaf Shah,Petros Zerfos,Heiko Ludwig,Asim Munawar,Maxwell Crouse,Pavan Kapanipathi,Shweta Salaria,Bob Calio,Sophia Wen,Seetharami Seelam,Brian Belgodere,Carlos Fonseca,Amith Singhee,Nirmit Desai,David D. Cox,Ruchir Puri,Rameswar Panda

from arxiv, Corresponding Authors: Rameswar Panda, Ruchir Puri; Equal Contributors: Mayank Mishra, Matt Stallone, Gaoyuan Zhang

Large Language Models (LLMs) trained on code are revolutionizing the software development process. Increasingly, code LLMs are being integrated into software development environments to improve the productivity of human programmers, and LLM-based agents are beginning to show promise for handling complex tasks autonomously. Realizing the full potential of code LLMs requires a wide range of capabilities, including code generation, fixing bugs, explaining and documenting code, maintaining repositories, and more. In this work, we introduce the Granite series of decoder-only code models for code generative tasks, trained with code written in 116 programming languages. The Granite Code models family consists of models ranging in size from 3 to 34 billion parameters, suitable for applications ranging from complex application modernization tasks to on-device memory-constrained use cases. Evaluation on a comprehensive set of tasks demonstrates that Granite Code models consistently reaches state-of-the-art performance among available open-source code LLMs. The Granite Code model family was optimized for enterprise software development workflows and performs well across a range of coding tasks (e.g. code generation, fixing and explanation), making it a versatile all around code model. We release all our Granite Code models under an Apache 2.0 license for both research and commercial use.

翻译：暂无翻译

相关内容

MoDELS

关注 43

ACM/IEEE第23届模型驱动工程语言和系统国际会议，是模型驱动软件和系统工程的首要会议系列，由ACM-SIGSOFT和IEEE-TCSE支持组织。自1998年以来，模型涵盖了建模的各个方面，从语言和方法到工具和应用程序。模特的参加者来自不同的背景，包括研究人员、学者、工程师和工业专业人士。MODELS 2019是一个论坛，参与者可以围绕建模和模型驱动的软件和系统交流前沿研究成果和创新实践经验。今年的版本将为建模社区提供进一步推进建模基础的机会，并在网络物理系统、嵌入式系统、社会技术系统、云计算、大数据、机器学习、安全、开源等新兴领域提出建模的创新应用以及可持续性。官网链接：http://www.modelsconference.org/

《生成式模型: 变分自编码器与扩散模型》，75页ppt，Google DeepMind科学家Ruiqi Gao

专知会员服务

66+阅读 · 2023年6月10日

Linux导论，Introduction to Linux，96页ppt

专知会员服务

81+阅读 · 2020年7月26日

【CHI2020-微软】解释可解释性:理解数据科学家使用机器学习的可解释性工具，Interpreting Interpretability: Understanding Data Scientists’Use of Interpretability Tools for Machine Learning

专知会员服务

55+阅读 · 2020年3月8日

FlowQA: Grasping Flow in History for Conversational Machine Comprehension

专知会员服务

34+阅读 · 2019年10月18日