paper-0190 · paper · 2023
Baichuan Inc.
Primary technical report for a notable AI model (verified primary source).
Abstract
Large language models (LLMs) have demonstrated remarkable performance on a variety of natural language tasks based on just a few examples of natural language instructions, reducing the need for extensive feature engineering. However, most powerful LLMs are closed-source or limited in their capability for languages other than English. In this technical report, we present Baichuan 2, a series of large-scale multilingual language models containing 7 billion and 13 billion parameters, trained from scratch, on 2.6 trillion tokens. Baichuan 2 matches or outperforms other open-source models of similar size on public benchmarks like MMLU, CMMLU, GSM8K, and HumanEval. Furthermore, Baichuan 2 excels in vertical domains such as medicine and law. We will release all pre-training model checkpoints to benefit the research community in better understanding the training dynamics of Baichuan 2. [OpenAlex]
Academic, score -0.2140
| Metric | Status | Value | Norm. | Weight | Contribution | Source | Confidence | License | Provenance |
|---|---|---|---|---|---|---|---|---|---|
| citation_count | present | 125.0 | 0.000558 | 0.5 | 0.000279 | OpenAlex | high | OpenAlex, CC0 metadata | link |
| library_holdings | missing | recorded as missing, penalized by rule, never imputed | −0.1 | recorded as missing; penalized by rule, never imputed | |||||
| readership_persistence | present | 4.0 | 0.214286 | 0.05 | 0.010714 | OpenAlex | medium | OpenAlex, CC0 metadata | link |
| syllabus_adoptions | missing | recorded as missing, penalized by rule, never imputed | −0.125 | recorded as missing; penalized by rule, never imputed | |||||
Broad Influence, score -0.1142
| Metric | Status | Value | Norm. | Weight | Contribution | Source | Confidence | License | Provenance |
|---|---|---|---|---|---|---|---|---|---|
| citation_count | present | 125.0 | 0.000558 | 0.2 | 0.000112 | OpenAlex | high | OpenAlex, CC0 metadata | link |
| library_holdings | missing | recorded as missing, penalized by rule, never imputed | −0.125 | recorded as missing; penalized by rule, never imputed | |||||
| readership_persistence | present | 4.0 | 0.214286 | 0.4 | 0.085714 | OpenAlex | medium | OpenAlex, CC0 metadata | link |
| syllabus_adoptions | missing | recorded as missing, penalized by rule, never imputed | −0.075 | recorded as missing; penalized by rule, never imputed | |||||
Governance Practitioner, score -0.3034
| Metric | Status | Value | Norm. | Weight | Contribution | Source | Confidence | License | Provenance |
|---|---|---|---|---|---|---|---|---|---|
| citation_count | present | 125.0 | 0.000558 | 0.25 | 0.00014 | OpenAlex | high | OpenAlex, CC0 metadata | link |
| library_holdings | missing | recorded as missing, penalized by rule, never imputed | −0.15 | recorded as missing; penalized by rule, never imputed | |||||
| readership_persistence | present | 4.0 | 0.214286 | 0.1 | 0.021429 | OpenAlex | medium | OpenAlex, CC0 metadata | link |
| syllabus_adoptions | missing | recorded as missing, penalized by rule, never imputed | −0.175 | recorded as missing; penalized by rule, never imputed | |||||
A rank is not a verdict on intrinsic worth. It is a transparent output of declared evidence, weights, and missing-data rules at a specific release date.
Disagree with this rank or a number? Challenge it with your evidence. Every challenge gets a public identifier and a published resolution.