Chat with quantized LLMs using OpenVINO CPU backend
Kroma style LoRA for Krea 2 Turbo text-to-image
MiniMax H3 video prompts via CPU GGUF