{"name":"expert-streaming-engine","description":"Run oversized sparse MoE GGUF models with bounded NVMe-to-RAM-to-VRAM caching, native multi-GPU planning, and a Linux desktop Studio.","screenshots":["/contribute_ss.webp","/contribute_ss.webp"],"sites":["https://github.com/xero00000/expert-streaming-engine"],"archived":false}