#631 · Primary category: Education & Research

bigscience

machine-learning models nlp training

Central place for the engineering/scaling WG: documentation, SLURM scripts and logs, compute environment and data.

Project last updated:07/29/24

GitHub Stars

1.0K

Forks

102

Contributors

16

License

Other

Why we included this project

BigScience's repo is the working file cabinet of the research workshop behind BLOOM, the multilingual model trained by a large open collaboration. It doesn't ship a single installable tool; instead it collects the engineering group's SLURM cluster scripts, live training logs, experiment write-ups, dataset notes, and walkthroughs of the compute environment. Anyone gearing up to pretrain or fine-tune a model with billions of parameters will find practical details here: how to set up clusters, what scaling baselines looked like, and candid lessons-learned documents about what broke across multiple long runs. For ML educators and research groups, it also works as a primary-source case study of how collaborative open training actually happens. Expect to read shell scripts and cluster documentation rather than a tidy API, but that's exactly where the value is.

Articles for this project

No articles for this project yet.

To suggest a topic or contribute an article, contact us.

Related projects in this category