Deprecated: The each() function is deprecated. This message will be suppressed on further calls in /home/zhenxiangba/zhenxiangba.com/public_html/phproxy-improved-master/index.php on line 456
Paper page - Stable LM 2 1.6B Technical Report
[go: Go Back, main page]

Papers
arxiv:2402.17834

Stable LM 2 1.6B Technical Report

Published on Feb 27, 2024
Authors:
,
,
,
,
,
,
,
,
,
,
,
,
,
,

Abstract

StableLM 2 1.6B, a language model with 1.6B parameters, achieves state-of-the-art performance for small-sized models across various benchmarks and is optimized for edge devices.

We introduce StableLM 2 1.6B, the first in a new generation of our language model series. In this technical report, we present in detail the data and training procedure leading to the base and instruction-tuned versions of StableLM 2 1.6B. The weights for both models are available via Hugging Face for anyone to download and use. The report contains thorough evaluations of these models, including zero- and few-shot benchmarks, multilingual benchmarks, and the MT benchmark focusing on multi-turn dialogues. At the time of publishing this report, StableLM 2 1.6B was the state-of-the-art open model under 2B parameters by a significant margin. Given its appealing small size, we also provide throughput measurements on a number of edge devices. In addition, we open source several quantized checkpoints and provide their performance metrics compared to the original model.

Community

Sign up or log in to comment

Get this paper in your agent:

hf papers read 2402.17834
Don't have the latest CLI?
curl -LsSf https://hf.co/cli/install.sh | bash

Models citing this paper 11

Browse 11 models citing this paper

Datasets citing this paper 0

No dataset linking this paper

Cite arxiv.org/abs/2402.17834 in a dataset README.md to link it from this page.

Spaces citing this paper 149

Browse 149 spaces citing this paper

Collections including this paper 0

No Collection including this paper

Add this paper to a collection to link it from this page.