Industry News
Training a 3.8B LLM for $998: Architectural Choices, Data Curation, and CORE Benchmark Analysis
An in-depth technical dive into how Hugo Vergnes pre-trained a competitive 3.8B parameter language model on a strict $998 budget, achieving a 0.384 CORE benchmark score.
Read more →