Pith. sign in
Pith Number

pith:UYV4HIS6

pith:2023:UYV4HIS62QZ3V2ZORIBLOX56TA
not attested not anchored not stored refs resolved

Mistral 7B

Albert Q. Jiang, Alexandre Sablayrolles, Arthur Mensch, Chris Bamford, Devendra Singh Chaplot, Diego de las Casas, Florian Bressand, Gianna Lengyel, Guillaume Lample, L\'elio Renard Lavaud, Lucile Saulnier, Marie-Anne Lachaux, Pierre Stock, Teven Le Scao, Thibaut Lavril, Thomas Wang, Timoth\'ee Lacroix, William El Sayed

A 7-billion-parameter model outperforms Llama 2 13B on every benchmark and Llama 1 34B on reasoning, mathematics, and code generation.

arxiv:2310.06825 v1 · 2023-10-10 · cs.CL · cs.AI · cs.LG

Add to your LaTeX paper
\usepackage{pith}
\pithnumber{UYV4HIS62QZ3V2ZORIBLOX56TA}

Prints a linked badge after your title and injects PDF metadata. Compiles on arXiv. Learn more · Embed verified badge

Record completeness

1 Bitcoin timestamp
2 Internet Archive
3 Author claim open · sign in to claim
4 Citations open
5 Replications open
Portable graph bundle live · download bundle · merged state
The bundle contains the canonical record plus signed events. A mirror can host it anywhere and recompute the same current state with the deterministic merge algorithm.

Claims

C1strongest claim

Mistral 7B outperforms Llama 2 13B across all evaluated benchmarks, and Llama 1 34B in reasoning, mathematics, and code generation.

C2weakest assumption

The chosen evaluation benchmarks and test sets are representative of downstream usefulness and contain no data leakage from the training corpus.

C3one line summary

Mistral 7B is a 7B-parameter LLM that outperforms Llama 2 13B across benchmarks via grouped-query attention and sliding-window attention while remaining efficient.

References

29 extracted · 29 resolved · 21 Pith anchors

[1] GQA: Training Generalized Multi-Query Transformer Models from Multi-Head Checkpoints 2023 · arXiv:2305.13245
[2] Program Synthesis with Large Language Models 2021 · arXiv:2108.07732
[3] Longformer: The Long-Document Transformer 2004 · arXiv:2004.05150
[4] Piqa: Reasoning about phys- ical commonsense in natural language 2020
[5] Evaluating Large Language Models Trained on Code 2021 · arXiv:2107.03374

Cited by

713 papers in Pith

Receipt and verification
First computed 2026-07-05T06:59:24.626993Z
Builder pith-number-builder-2026-05-17-v1
Signature Pith Ed25519 (pith-v1-2026-05) · public key
Schema pith-number/v1.0

Canonical hash

a62bc3a25ed433baeb2e8a02b75fbe9801ea22b68648adaaa87b7e9049ffd391

Aliases

arxiv: 2310.06825 · arxiv_version: 2310.06825v1 · doi: 10.48550/arxiv.2310.06825 · pith_short_12: UYV4HIS62QZ3 · pith_short_16: UYV4HIS62QZ3V2ZO · pith_short_8: UYV4HIS6
Agent API
Verify this Pith Number yourself
curl -sH 'Accept: application/ld+json' https://pith.science/pith/UYV4HIS62QZ3V2ZORIBLOX56TA \
  | jq -c '.canonical_record' \
  | python3 -c "import sys,json,hashlib; b=json.dumps(json.loads(sys.stdin.read()), sort_keys=True, separators=(',',':'), ensure_ascii=False).encode(); print(hashlib.sha256(b).hexdigest())"
# expect: a62bc3a25ed433baeb2e8a02b75fbe9801ea22b68648adaaa87b7e9049ffd391
Canonical record JSON
{
  "metadata": {
    "abstract_canon_sha256": "96e32cda302bd4f4c818c808146c0da97d0025b59edbb48fdfff8038778a488d",
    "cross_cats_sorted": [
      "cs.AI",
      "cs.LG"
    ],
    "license": "http://creativecommons.org/licenses/by/4.0/",
    "primary_cat": "cs.CL",
    "submitted_at": "2023-10-10T17:54:58Z",
    "title_canon_sha256": "7f8336369be8aaabe3568c438cf97f40b8f4f596e45196c70422f34d3b23fbb9"
  },
  "schema_version": "1.0",
  "source": {
    "id": "2310.06825",
    "kind": "arxiv",
    "version": 1
  }
}