lightning-ai/lit-llama
steadyImplementation of the LLaMA language model based on nanoGPT. Supports flash attention, Int8 and GPTQ 4bit quantization, LoRA and LLaMA-Adapter fine-tuning, pre-training. Apache 2.0-licensed.
Python
View on GitHub
Stars
6,083
Forks
519
Open issues
100
24h
0
0.0%
7d
+3
+0.0%
Refresh
2h
Star history (7 days)
Last checked
1h ago
Last pushed
01 Jul 2025
Next check
just now