mini-AGI alternatives
Curated alternatives to mini-AGI — and why you'd switch.
train-llm-from-scratch
The full LLM pipeline hand-written in plain PyTorch — tokens, transformer, pretraining, then SFT, reward model, PPO, DPO, GRPO. No trl, no peft: read every algorithm, train on one GPU.
Why switchBoth train your own small LM from scratch on a single GPU; train-llm-from-scratch teaches the standard pretrain→SFT→RL pipeline, mini-AGI experiments with never-ending continual learning instead.
Full comparison →