arxiv: 2501.05465 · v2 · pith:2A3QF3NHnew · submitted 2025-01-03 · 💻 cs.CL

Small Language Models (SLMs) Can Still Pack a Punch: A survey (updated 2026)

Akanksha Gupta , Bijo Thomas , Harshita Asnani , Phanindra Reddy Madduru , Samia Feroze , Shreyas Subramanian , Vikram Elango , Mecit Gungor This is my paper

classification 💻 cs.CL

keywords modelsslmslanguagesmallsurveyagnosticarisesbalancing

0 comments p. Extension

Add this Pith Number to your LaTeX paper

What is a Pith Number?

\usepackage{pith}
\pithnumber{2A3QF3NH}

Prints a linked pith:2A3QF3NH badge after your title and writes the identifier into PDF metadata. Compiles on arXiv with no extra files. Learn more

read the original abstract

As foundation AI models continue to increase in size, an important question arises - is massive scale the only path forward? This survey of about 160 papers presents a family of Small Language Models (SLMs) in the 1 to 8 billion parameter range that demonstrate smaller models can perform as well, or even outperform large models. We explore task agnostic, general purpose SLMs, task-specific SLMs and techniques to create SLMs that can guide the community to build models while balancing performance, efficiency, scalability and cost. Furthermore we define and characterize SLMs' effective sizes, representing increased capability with respect to LLMs.

This paper has not been read by Pith yet.

discussion (0)

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

ShadowNPU: System and Algorithm Co-design for NPU-Centric On-Device LLM Inference
cs.PF 2025-08 unverdicted novelty 5.0

ShadowNPU presents shadowAttn, a co-designed sparse attention system that uses NPU pilot compute and techniques like graph bucketing and per-head sparsity to minimize CPU/GPU fallback during on-device LLM inference wh...
Small Language Models are the Future of Agentic AI
cs.AI 2025-06 unverdicted novelty 5.0

Small language models are sufficiently capable, more suitable, and far more economical than large models for the repetitive tasks that dominate agentic AI systems.
SLM Finetuning for Natural Language to Domain Specific Code Generation in Production
cs.LG 2026-04 unverdicted novelty 3.0

Fine-tuned small language models outperform larger models in natural language to domain-specific code generation with improved performance, latency, and the ability to adapt to customer-specific scenarios without losi...