Browse State-of-the-Art › Small Language Model
Small Language Model
41 papers with code · 0 benchmarks · 0 datasets archive 2025-07-28
Benchmarks archive 2025-07-28
No benchmark for this task in the archive.
Libraries
Not in the archive: the export carries no per-task library table, so there is nothing to show at snapshot 2025-07-28.
Datasets archive 2025-07-28
No dataset record in the archive lists this task.
Subtasks archive 2025-07-28
No subtask under this task in the archive's task tree.
Most implemented papers archive 2025-07-28
30 shown of 41 papers with code (109 tagged with this task in all), ordered by repositories listed in the archive, not by stars (the archive holds no stars, so PwC's “Social” and “Latest” sorts cannot be reproduced). Papers without a page here are shown as plain text.
-
27 Jan 2025 2 repositories listedWe introduce Atla Selene Mini, a state-of-the-art small language model-as-a-judge (SLMJ).
-
4 Jan 2024 2 repositories listed Syntology ran 8 of 9 samples · 1 unverifiedWe present TinyLlama, a compact 1.
-
17 Jun 2025 1 repository listedTo address this, a secondary model, known as a relevant grader, can be served to verify its relevance.
-
4 Jun 2025 1 repository listedRecently, Large Language Models (LLMs) have demonstrated significant potential for data annotation, markedly reducing the labor costs associated with downstream applications.
-
28 May 2025 1 repository listedVision encoders trained on the surrogate can then be directly transferred to the larger model, a process we call zero-shot grafting -- when plugged directly into the full-size target LLM, the grafted pair surpasses the…
-
21 May 2025 1 repository listedThis study explores the enhancement of medical knowledge in a small language model by leveraging accessible online data, including a crawled corpus from medical magazines and a dataset of real doctor-patient QA pairs.
-
21 Apr 2025 1 repository listedTo address this, we propose CRAVE, a Conflicting Reasoning Approach for explainable claim VErification, that verify the complex claims based on the conflicting rationales reasoned by large language models (LLMs).
-
7 Apr 2025 1 repository listed Syntology ran 0 of 1 samples · 1 unverifiedRetrieval-Augmented Generation (RAG) systems often struggle to handle multi-hop question-answering tasks accurately due to irrelevant context retrieval and limited complex reasoning capabilities.
-
27 Mar 2025 1 repository listedTo tackle these challenges, we propose Mobile-VideoGPT, an efficient multimodal framework designed to operate with fewer than a billion parameters.
-
20 Mar 2025 1 repository listedThis paper presents a survey on distributed solutions for various LMs, including large language models (LLMs), vision language models (VLMs), multimodal LLMs (MLLMs), and small language models (SLMs).
-
11 Feb 2025 1 repository listedHowever, the task of extracting longer entity spans (e.
-
19 Jan 2025 1 repository listedSpecifically, to efficiently query the LLM, we propose an adaptive selection strategy based on the uncertainty estimation of the SLM, where the LLM is invoked only when the SLM is uncertain.
-
16 Dec 2024 1 repository listed Syntology ran 0 of 5 samples · 5 unverifiedWe address the challenge of utilizing large language models (LLMs) for complex embodied tasks, in the environment where decision-making systems operate timely on capacity-limited, off-the-shelf devices.
-
24 Nov 2024 1 repository listedMinimal duplication positively impacted model accuracy (+0.
-
21 Nov 2024 1 repository listedThis progress is largely attributed to the development of generalizable source code representations that effectively capture the syntactic and semantic characteristics of code.
-
18 Nov 2024 1 repository listedInteractions with online Large Language Models raise privacy issues where providers can gather sensitive information about users and their companies from the prompts.
-
7 Nov 2024 1 repository listedThe interest in developing small language models (SLM) for on-device deployment is fast growing.
-
4 Nov 2024 1 repository listedThe telecommunications industry's rapid evolution demands intelligent systems capable of managing complex networks and adapting to emerging technologies.
-
29 Oct 2024 1 repository listed Syntology ran 3 of 4 samples · 1 unverified · 4 pointer-only (licence)Large language models (LLMs) have shown impressive capabilities across various tasks, but their performance on domain-specific tasks remains limited.
-
10 Oct 2024 1 repository listed Syntology ran 2 of 2 samples · 0 unverified · 2 pointer-only (licence)A mechanistic understanding of how MLPs do computation in deep neural networks remains elusive.
-
17 Sep 2024 1 repository listedIn this paper, we evaluate the creative fiction writing abilities of a fine-tuned small language model (SLM), BART-large, and compare its performance to human writers and two large language models (LLMs): GPT-3.
-
6 Sep 2024 1 repository listedFurthermore, our approach exhibits major cost benefits: the average prediction quality of AnyMatch is within 4.
-
1 Sep 2024 1 repository listed Syntology ran 1 of 1 samples · 0 unverifiedRecent large language models (LLMs) have enabled the development of advanced agentic systems that can integrate various tools and APIs to fulfill user queries through function calling.
-
30 Aug 2024 1 repository listedOur findings reveal that while BERT-based models generally outperform both the LLMs and SLM, the performance of the large generative models is still noteworthy.
-
24 Aug 2024 1 repository listed Syntology ran 10 of 14 samples · 4 unverifiedThe widespread adoption of cloud-based proprietary large language models (LLMs) has introduced significant challenges, including operational dependencies, privacy concerns, and the necessity of continuous internet…
-
22 Aug 2024 1 repository listedLarge language models (LLMs) are highly capable but face latency challenges in real-time applications, such as conducting online hallucination detection.
-
21 Aug 2024 1 repository listedRecent studies show that large language models (LLMs) struggle with technical standards in telecommunications.
-
28 Jul 2024 1 repository listed Syntology ran 0 of 1 samples · 1 unverified · 1 pointer-only (licence)However, existing methods are designed for specific models with fixed prompts, limiting their adaptability to the fast-evolving models and diverse practical scenarios.
-
21 Jun 2024 1 repository listedThe goal of text style transfer is to transform the style of texts while preserving their original meaning, often with only a few examples of the target style.
-
6 Jun 2024 1 repository listedRecent advancements in text-to-speech (TTS) powered by language models have showcased remarkable capabilities in achieving naturalness and zero-shot voice cloning.
Syntology lines on 8 of the papers shown; no Syntology record for the others (a paper without an arXiv id cannot be joined to the graph, and absence from the graph layer is not a recorded non-run). “Ran” means the sample executed on a synthesized fixture, not that the paper's result was reproduced. Read from the graph 2026-09-24.
Report a problem or propose a change · a person checks every report against the paper or source before anything changes; decisions are listed on /corrections