Live
AI models

Llemma 34B

Training compute
5.4×10²³ FLOP
Parameters
34B
Published
Oct 16, 2023

Llemma 34B is an AI model developed by Princeton University, University of Toronto, Vector Institute, University of Cambridge, Carnegie Mellon University (CMU), University of Washington and EleutherAI (United States, Canada and United Kingdom), first published in October 2023. It works in the mathematics and language domain, on tasks such as mathematical reasoning, language modeling/generation, question answering and code generation.

Training it took an estimated 5.4×10²³ FLOP of compute (estimation method: operation counting,hardware). The model has 34,000,000,000 parameters. It was trained on roughly 55B datapoints. Training ran on 256 NVIDIA A100 SXM4 40 GB.

Access: Open weights (restricted use). Its weights are openly available. It is built on top of Code Llama-34B. The reference paper has 440 citations. Epoch AI rates the confidence of this record as confident.

Full record
Organization
Princeton University, University of Toronto, Vector Institute, University of Cambridge, Carnegie Mellon University (CMU), University of Washington, EleutherAI
Country of organization
United States, Canada, United Kingdom
Domain
Mathematics, Language
Task
Mathematical reasoning, Language modeling/generation, Question answering, Code generation
Training compute
5.4×10²³ FLOP
Compute estimation method
Operation counting, Hardware
Parameters
34,000,000,000
Dataset size
55B
Training hardware
NVIDIA A100 SXM4 40 GB
Chips used
256
Chip-hours
47K
Training power draw
203.3 kW
Numerical format
BF16
Model accessibility
Open weights (restricted use)
Open weights
Yes
Base model
Code Llama-34B
Citations
440
Epoch confidence
Confident
SourceEpoch AI, 'AI Models'. Published online at epoch.ai. Retrieved 2026-07-29 from https://epoch.ai/data/ai-models. Licensed under CC BY 4.0.
← All ai models