rabbitllm
Run 70B+ LLMs on a single 4GB GPU — no quantization required. Layer-streaming inference for consumer hardware.
Activity
- Latest release
- 6mo ago
- Total releases
- 3
- Cadence
- ~2 days
- Last 12 months
- 3
Reach
- Stars
- —
Details
- License
- MIT
- First release
- Feb 22, 2026
Releases
| Version | Released | |
|---|---|---|
1.1.0
minor
| ||
1.0.1
patch
| ||
1.0.0
initial
|