huihui-ai/Huihui-GLM-5.2-abliterated-GGUF Text Generation • 754B • Updated 2 days ago • 15.4k • 259
coder3101/gemma-4-31B-it-qat-q4_0-unquantized-heretic Image-Text-to-Text • 31B • Updated Jun 6 • 63 • • 5
nvidia/nemotron-3.5-asr-streaming-0.6b Automatic Speech Recognition • 0.6B • Updated 18 days ago • 750k • • 926
nvidia/NVIDIA-Nemotron-Nano-12B-v2-VL-FP8 Image-Text-to-Text • 13B • Updated Nov 13, 2025 • 24.9k • 51
ibm-granite/granite-docling-258M Image-Text-to-Text • 0.3B • Updated Sep 23, 2025 • 205k • 1.23k
AutoTriton: Automatic Triton Programming with Reinforcement Learning in LLMs Paper • 2507.05687 • Published Jul 8, 2025 • 31
nvidia/Llama-3.1-Nemotron-Nano-VL-8B-V1 Image-Text-to-Text • 9B • Updated Dec 4, 2025 • 1.29M • 181
nvidia/OpenCodeReasoning-Nemotron-32B-IOI Text Generation • 33B • Updated May 7, 2025 • 47 • • 26
nvidia/Llama-3_1-Nemotron-Ultra-253B-v1 Text Generation • 253B • Updated Oct 15, 2025 • 2.76k • 352
PocketDoc/Dans-PersonalityEngine-V1.2.0-24b Text Generation • 24B • Updated May 23, 2025 • 76 • • 183
Can 1B LLM Surpass 405B LLM? Rethinking Compute-Optimal Test-Time Scaling Paper • 2502.06703 • Published Feb 10, 2025 • 153
meta-llama/Llama-3.2-90B-Vision-Instruct Image-Text-to-Text • 89B • Updated Mar 4, 2025 • 46.9k • 359