Original caption
Full setup is in my profile. This is Colibri, and it can run absolutely massive AI models locally by using your SSD, RAM and VRAM together. We’re talking 744 BILLION to 2.8 TRILLION parameters on consumer hardware. It’s still slow right now. But if this gets faster, local AI gets very interesting very quickly. And NVIDIA might want to pay attention. 👀