Pıer
TidesCurrentsHarbor LightsLabBottlesAshore
Pıer

Navigation

  • Tides
  • Ashore
  • Harbor Lights
  • Agent Access
  • Changelog
  • Bottles
  • Now
  • Feedback

External links

GitHubCloudborne ↗

© 2026 Pier.

Read original
Hacker News·Anon84·Sep 10, 2026, 2:04 AM

Training a 3.8B LLM to 0.384 CORE for $998

Models78

Independent developer Hugo Vergnes documented training a 3.8B-parameter language model to a 0.384 CORE benchmark score with a budget capped at $998. The experiment highlights how targeted dataset filtering, modern learning schedules, and compute frugality can squeeze competent small-scale models out of modest cloud GPU spend without corporate backing.

Why it's worth reading

It offers an audited, sub-$1,000 reference point for pre-training a usable small LLM, turning compute efficiency into reproducible engineering rather than corporate PR.

Tags

LLMPretrainingOpenSourceComputeEfficiencySmallModelsMachineLearning

Score breakdown

  • Novelty19
  • Impact18
  • Practicality21
  • Credibility18
  • Timeliness18