TimeCapsuleLLM is an experimental project training language models from scratch on historical datasets to eliminate modern bias. By avoiding fine-tuning, it aims to simulate the worldview and language of specific eras, currently focusing on 1800–1850 London.
Highlights
Trains from scratch to prevent modern concept leakage
Uses nanoGPT architecture
Focuses on 1800–1850 London data
Scaling dataset from 50 to 600 texts for better coherence
Contribute to angelos-p/llm-from-scratch development by creating an account on GitHub.
llm-chronicles.com
LLM Chronicles
A fast-paced whiteboard animation series unraveling Deep Learning and Large Language Models. Dive into core concepts, from Neural Networks basics to t...
github.com
GitHub - arman-bd/guppylm: A ~9M parameter LLM that talks like a small fish.
A ~9M parameter LLM that talks like a small fish. Contribute to arman-bd/guppylm development by creating an account on GitHub.