All models

#10 Llama 3.1 70B Instruct

Meta · Llama 3.1 70B Instruct

Released: · Source

Text

Use cases
Text
Accepts
No verified data
Produces
No verified data

Access

Price USD

No verified data · Local: hardware costs apply

Evaluator & score

Independent run · default (effort unspecified by source) · 20.09.2026 · Higher is better

Snapshot: 2026-09-20

Conditions

GPQA Diamond; Epoch AI own run; best score across scorers

Epoch AI own run; source version: Llama-3.1-70B-Instruct; file: gpqa_diamond.csv; row: MEwMcmSV32JHrsNoAq4BFL.

Measurement date: 27.01.2025

Independent run · default (effort unspecified by source) · 20.09.2026 · Higher is better

Snapshot: 2026-09-20

Conditions

MATH level 5; Epoch AI own run; best score across scorers

Epoch AI own run; MATH level 5; version: Llama-3.1-70B-Instruct; mode: default (effort unspecified by source); source id: 6JiXiQ67iyiC7cFFRiogvs; file: math_level_5.csv. CC BY 4.0, Epoch AI. Snapshot 2026-09-20. Source column Best score (across scorers), not AIpedia's selection of the best effort.

Measurement date: 27.01.2025

Independent run · default (effort unspecified by source) · 20.09.2026 · Higher is better

Snapshot: 2026-09-20

Conditions

Mock AIME 2024/2025; Epoch AI own run; best score across scorers

Epoch AI own run; source version: Llama-3.1-70B-Instruct; file: otis_mock_aime_2024_2025.csv; row: X2tQXGELXtEbc2VX6gARd7.

Measurement date: 25.02.2025

SourceVerified 18.09.2026