Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Compactbot
/
ldt-10m
Like
0
Text Generation
Safetensors
English
ldt
tiny
slm
small-language-model
from-scratch
10m
License:
apache-2.0
Model card
Files
Files and versions
xet
Community
15
Copy to bucket
new
main
ldt-10m
42 MB
Ctrl+K
Ctrl+K
1 contributor
History:
24 commits
Compactbot
Fix param count: 10,026,880 โ 10,284,480 (verified from safetensors header)
32a8b17
verified
about 21 hours ago
.gitattributes
Safe
1.52 kB
initial commit
9 days ago
README.md
2.37 kB
Fix param count: 10,026,880 โ 10,284,480 (verified from safetensors header)
about 21 hours ago
config.json
287 Bytes
Add config.json
1 day ago
model.py
Safe
5.62 kB
Self-contained LDT model definition (#2)
9 days ago
model.safetensors
Safe
41.1 MB
xet
v8: continued training to 2.6B tokens (from 1.51B). Improved coherence on narrative prompts. Same architecture (10,052,864 params, 5L d320 5-head SwiGLU RoPE). (#14)
5 days ago
tokenizer.json
Safe
830 kB
Byte-level BPE tokenizer (vocab 12288) (#3)
9 days ago
tokenizer_config.json
Safe
250 Bytes
Tokenizer config (#5)
9 days ago