arrow
Return

Assessing Small Language Models for Code Generation: An Empirical Study with Benchmarks

delete2026-02-13
delete0
delete
OA
AI
M
Md Mahade Hasan
M
Muhammad Waseem
K
Kai‐Kristian Kemell
J
Jussi Rasku
J
Juha Ala-Rantala
P
Pekka Abrahamsson
DOI:10.1016/j.jss.2026.112815delete
deleteOriginal
deleteShare
deleteSave
View PDF
Abstract

Abstract

En 中文
• Evaluated 20 open-source Small Language Models (SLMs) on five code generation benchmarks, covering tasks such as code synthesis, summarization, and repair. • Identified that several compact SLMs (-<3B parameters) achieve competitive performance with lower resource requirements, making them suitable for deployment in memory-constrained environments. • Larger SLMs offer higher accuracy but demand up to 4 times more resources (VRAM) for a 10% improvement in pass@1, highlighting clear performance efficiency tradeoffs of SLMs. • No statistically significant performance differences across programming languages, suggesting SLMs, generalizability in multilingual code generation tasks within the programming languages tested in the study.
Keywords:
Small Language Models
Code Generation
Empirical Study
Benchmarks
Software Engineering
Generative AI
AI Summary

AI Summary

Key information extracted from the uploaded paper, including a brief overview, abstract, background, key highlights, visual analysis, and future outlook.

Journal

Journal of Systems and Software cover
Journal of Systems and Software
IF:
4.1
Papers:
5.5K
Citations:
8.4K

Organization

No organization information available
Cited Papers

Cited Papers

A Survey on Optimization Techniques for Edge Artificial Intelligence (AI)
errSENSORS
IF3.5
err2023-01-22
err21
errOAAI
errSurianarayanan, Chellammal; Lawrence, John Jeyasekaran; Chelliah, Pethuru Raj; Prakash, Edmond; Hewage, Chaminda
errShare
errSave
Competition-level code generation with AlphaCode
err2022-12-09
err0
errOAAI
errYujia Li; David Choi; Junyoung Chung; Nate Kushman; Julian Schrittwieser; Rémi Leblond; Tom Eccles; James Keeling; Felix Gimeno; Agustin Dal Lago; Thomas Hubert; Peter Choy; Cyprien de Masson d’Autume; Igor Babuschkin; Xinyun Chen; Po-Sen Huang; Johannes Welbl; Sven Gowal; Alexey Cherepanov; James Molloy; Daniel J. Mankowitz; Esme Sutherland Robson; Pushmeet Kohli; Nando de Freitas; Koray Kavukcuoglu; Oriol Vinyals
errShare
errSave
LLM-Based Test-Driven Interactive Code Generation: User Study and Empirical Evaluation
err2024-09-01
err1
errOAAI
errFakhoury, Sarah; Naik, Aaditya; Sakkas, Georgios; Chakraborty, Saikat; Lahiri, Shuvendu K.
errShare
errSave
A Survey of Machine Learning for Big Code and Naturalness
err2018-07-31
err477
errOAAI
errAllamanis, Miltiadis; Barr, Earl T.; Devanbu, Premkumar; Sutton, Charles
errShare
errSave
no more