Models 38
Latest releases first. Find a model for what you’re making.
0–7d New 8–30d Recent 31d+ Older · days since release
GLM 5.3 FlashX
2dGLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s.
Open weightsLing 3.0 Flash VL
10dLing 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding…
7.7K downloads / mo97 likesTaoMate-H3
13d635 downloads / mo94 likesMinimax-h3_Singularity
15d217.9K downloads / mo499 likesFastVideo-FastH3-8-Step-V2
16d1.2K downloads / mo75 likesQwen3.8 Max (0902)
17dQwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team.
Open weightsvdn-minimax-h3
18d664 downloads / mo312 likesViggle-Animate
20d0 downloads / mo240 likesGLM Flash Latest
24dThis model always redirects to the latest model in the GLM Flash family.
Open weightsQwen3.8 Flash
25dQwen3.8 Flash is a multimodal reasoning model from Alibaba.
724.1K downloads / mo5.4K likesGLM 5.3 Flash
25dGLM-5.3-Flash is a native multimodal model from Z.ai.
2.7M downloads / mo2.5K likesQwen3.8 27B
37dQwen3.8 27B is an open-weight dense vision-language model from Qwen.
7.4M downloads / mo15.7K likesMinimax-h3-Turbo
44d1.5M downloads / mo967 likesQwen3.7 Flash
55dQwen3.7 Flash is a vision-language reasoning model from Alibaba.
Open weightsLTX-2.5
59d1.6M downloads / mo4.3K likesKimi K3
66dKimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI.
2.2M downloads / mo11.4K likesMiniMax M3
112dMiniMax-M3 is a multimodal foundation model from MiniMax.
191.9K downloads / mo1.5K likesStep 3.7 Flash
115dStep 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model.
22.2K downloads / mo455 likesNemotron 3 Nano Omni (free)
145dNVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems.
272.4K downloads / mo427 likesKimi Latest
146dThis model always redirects to the latest model in the Kimi family.
Open weightsQwen3.6 27B
146dQwen3.6 27B is a dense 27-billion-parameter language model from the Qwen Team at Alibaba, released in April 2026.
3.7M downloads / mo2.3K likesQwen3.5 Plus 2026-04-20
146dQwen3.5 Plus (April 2026) is a large-scale multimodal language model from Alibaba.
Open weightsQwen3.6 Flash
146dQwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series.
Open weightsQwen3.6 35B A3B
146dQwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token.
3.4M downloads / mo2.8K likesMiMo-V2.5
151dMiMo-V2.5 is a native omnimodal model by Xiaomi.
296.7K downloads / mo424 likesGemma 4 26B A4B
170dGemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind.
9.8M downloads / mo1.5K likesQwen3.6 Plus
171dQwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and…
Open weightsGemma 4 31B
171dGemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output.
9M downloads / mo3.9K likesGLM 5V Turbo
172dGLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks.
Open weightsReka Edge
184dReka Edge is an extremely efficient 7B multimodal vision-language model that accepts image/video+text inputs and generates text outputs.
533 downloads / mo133 likesQwen3.5-9B
194dQwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient…
9.3M downloads / mo2K likesQwen3.5-122B-A10B
207dThe Qwen3.5 122B-A10B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse…
364.9K downloads / mo620 likesQwen3.5-35B-A3B
207dThe Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechanisms and a sparse…
1.9M downloads / mo1.5K likesQwen3.5-27B
207dThe Qwen3.5 27B native vision-language Dense model incorporates a linear attention mechanism, delivering fast response times while balancing inference speed…
1.9M downloads / mo1K likesQwen3.5-Flash
207dThe Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse…
Open weightsQwen3.5 Plus 2026-02-15
216dThe Qwen3.5 native vision-language series Plus models are built on a hybrid architecture that integrates linear attention mechanisms with sparse…
Open weightsQwen3.5 397B A17B
216dThe Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse…
200.6K downloads / mo1.6K likesGLM 4.6V
286dGLM-4.6V is a large multimodal model designed for high-fidelity visual understanding and long-context reasoning across images, documents, and mixed media.
4.6K downloads / mo396 likes
