Skip to content
#

small-language-models

Here are 337 public repositories matching this topic...

This Repository provides a Jupyter Notebook for building a small language model from scratch using 'TinyStories' dataset. Covers data preprocessing, BPE tokenization, binary storage, GPU memory management, and training a Transformer in PyTorch. Generate sample stories to test your model. Ideal for learning NLP and PyTorch.

  • Updated Jun 7, 2025
  • Jupyter Notebook

A governed local AI build-and-memory system that trains small brains, compares them, protects the better one, archives the worse one, and preserves the evidence of why. v1.0.0/governed-v2.2.0+

  • Updated May 16, 2026
  • Python

Readable, composable PyTorch library for language models: build Llama, Qwen, Gemma, DeepSeek, Kimi and 20+ other architectures from plain nn.Modules, then train them on CPU, one GPU, or multi-GPU.

  • Updated Sep 27, 2026
  • Python

Add this topic to your repo

To associate your repository with the small-language-models topic, visit your repo's landing page and select "manage topics."

Learn more