Deepseek deepseek-ai/DeepSeek-V4-Pro Text Generation • 1.6T • Updated Jun 22 • 782k • • 5.49k deepseek-ai/DeepSeek-V3 Text Generation • 685B • Updated Mar 27, 2025 • 1.08M • • 4.18k
transformer KV Cache Steering for Inducing Reasoning in Small Language Models Paper • 2507.08799 • Published Jul 11, 2025 • 40
KV Cache Steering for Inducing Reasoning in Small Language Models Paper • 2507.08799 • Published Jul 11, 2025 • 40
glm nvidia/GLM-5.2-NVFP4 Text Generation • 381B • Updated 1 day ago • 1.06M • 320 huihui-ai/Huihui-GLM-5.2-abliterated-GGUF Text Generation • 754B • Updated about 1 month ago • 11.2k • 343
huihui-ai/Huihui-GLM-5.2-abliterated-GGUF Text Generation • 754B • Updated about 1 month ago • 11.2k • 343
video Vidi: Large Multimodal Models for Video Understanding and Editing Paper • 2504.15681 • Published Apr 22, 2025 • 14
Vidi: Large Multimodal Models for Video Understanding and Editing Paper • 2504.15681 • Published Apr 22, 2025 • 14
microsoft phi 4 microsoft/Phi-4-multimodal-instruct Automatic Speech Recognition • 6B • Updated Dec 10, 2025 • 259k • 1.61k
microsoft/Phi-4-multimodal-instruct Automatic Speech Recognition • 6B • Updated Dec 10, 2025 • 259k • 1.61k
foolingaround google/flan-t5-large 0.8B • Updated Jul 17, 2023 • 466k • 892 stepfun-ai/GOT-OCR-2.0-hf Image-Text-to-Text • 0.6B • Updated Jan 31, 2025 • 167k • 240
Deepseek deepseek-ai/DeepSeek-V4-Pro Text Generation • 1.6T • Updated Jun 22 • 782k • • 5.49k deepseek-ai/DeepSeek-V3 Text Generation • 685B • Updated Mar 27, 2025 • 1.08M • • 4.18k
glm nvidia/GLM-5.2-NVFP4 Text Generation • 381B • Updated 1 day ago • 1.06M • 320 huihui-ai/Huihui-GLM-5.2-abliterated-GGUF Text Generation • 754B • Updated about 1 month ago • 11.2k • 343
huihui-ai/Huihui-GLM-5.2-abliterated-GGUF Text Generation • 754B • Updated about 1 month ago • 11.2k • 343
transformer KV Cache Steering for Inducing Reasoning in Small Language Models Paper • 2507.08799 • Published Jul 11, 2025 • 40
KV Cache Steering for Inducing Reasoning in Small Language Models Paper • 2507.08799 • Published Jul 11, 2025 • 40
video Vidi: Large Multimodal Models for Video Understanding and Editing Paper • 2504.15681 • Published Apr 22, 2025 • 14
Vidi: Large Multimodal Models for Video Understanding and Editing Paper • 2504.15681 • Published Apr 22, 2025 • 14
microsoft phi 4 microsoft/Phi-4-multimodal-instruct Automatic Speech Recognition • 6B • Updated Dec 10, 2025 • 259k • 1.61k
microsoft/Phi-4-multimodal-instruct Automatic Speech Recognition • 6B • Updated Dec 10, 2025 • 259k • 1.61k
foolingaround google/flan-t5-large 0.8B • Updated Jul 17, 2023 • 466k • 892 stepfun-ai/GOT-OCR-2.0-hf Image-Text-to-Text • 0.6B • Updated Jan 31, 2025 • 167k • 240