MiniMaxAI/MiniMax-H3 Image-Text-to-Video β’ 33B β’ Updated about 17 hours ago β’ 1.61M β’ β’ 3.81k
deepseek-ai/DeepSeek-V4-Flash-0731 Text Generation β’ 304B β’ Updated 13 days ago β’ 1.43M β’ β’ 3.31k
view post Post 4926 Weβre releasing Gemma 4 NVFP4 quants that run 1.5Γ faster on your GPU.Gemma-4-12B NVFP4 works on 11GB VRAM.26B-A4B hits 13K tok/s (B200).Unsloth NVFP4 enables faster, more accurate 4-bit Blackwell inference.Blog: https://unsloth.ai/docs/basics/nvfp4Gemma NVFP4: https://huggingface.co/collections/unsloth/nvfp4 See translation 3 replies Β· π₯ 12 12 π€ 3 3 π 2 2 π 2 2 + Reply