Implement a ChatGPT-like LLM in PyTorch from scratch, step by step
Fix encoding of multiple preceding spaces in BPE tokenizer. (#945)
* Fix encoding of multiple preceding spaces in BPE tokenizer. * Add test --------- Co-authored-by: rasbt <mail@sebastianraschka.com>
M
Maxwell De Jong committed
e0dbec33314cc5f4d1f1d70258f31bb8f6dd0588
Parent: 90e0f3c
Committed by GitHub <noreply@github.com>
on 1/10/2026, 4:27:23 PM