Models 69
Latest releases first. Find a model for what you’re making.
0–7d New 8–30d Recent 31d+ Older · days since release
GLM 5.3 FlashX
2dGLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s.
Open weightsLing 3.0 Flash VL
10dLing 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding…
7.7K downloads / mo97 likesTaoMate-H3
13d635 downloads / mo94 likesMinimax-h3_Singularity
15d217.9K downloads / mo499 likesFastVideo-FastH3-8-Step-V2
16d1.2K downloads / mo75 likesQwen3.8 Max (0902)
17dQwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team.
Open weightsMuse Spark 1.3
18dMuse Spark 1.3 is a multimodal reasoning model from Meta for long-running agentic, multi-agent, and coding workflows.
API modelvdn-minimax-h3
18d664 downloads / mo312 likesGemini 3.8 Flash
18dGemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step…
API modelMuse Spark 1.3 Contributor
18dMuse Spark 1.3 Contributor is the cost-efficient contributor tier of Meta’s multimodal reasoning model for experimentation, learning, and early-stage agentic,…
API modelViggle-Animate
20d0 downloads / mo240 likesGLM Flash Latest
24dThis model always redirects to the latest model in the GLM Flash family.
Open weightsQwen3.8 Flash
25dQwen3.8 Flash is a multimodal reasoning model from Alibaba.
724.1K downloads / mo5.4K likesGLM 5.3 Flash
25dGLM-5.3-Flash is a native multimodal model from Z.ai.
2.7M downloads / mo2.5K likesMuse Spark 1.2 Contributor
30dMuse Spark 1.2 contributor tier is a reasoning model from Meta designed for developers who want to start building at an even lower cost.
API modelQwen3.8 27B
37dQwen3.8 27B is an open-weight dense vision-language model from Qwen.
7.4M downloads / mo15.7K likesGemini 3.7 Flash
38dGemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning.
API modelSeed 2.1 Turbo
39dSeed 2.1 Turbo is a multimodal model from ByteDance Seed for coding and long-horizon agent workflows.
API modelSeed-2.0-Code
39dSeed 2.0 Code is a model from ByteDance Seed optimized for agentic coding.
API modelMinimax-h3-Turbo
44d1.5M downloads / mo967 likesMuse Spark 1.2
46dMuse Spark 1.2 is a reasoning model from Meta, designed for complex agentic tasks.
API modelQwen3.7 Flash
55dQwen3.7 Flash is a vision-language reasoning model from Alibaba.
Open weightsLTX-2.5
59d1.6M downloads / mo4.3K likesGemini 3.5 Flash Lite
61dGemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities.
API modelGemini 3.6 Flash
61dGemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development.
API modelAuto Router (Beta)
65dThe experimental version of our Auto Router where we test new improvements.
API modelMuse Spark 1.1
66dMuse Spark 1.1 is a multimodal reasoning model from Meta, built for agentic tasks.
API modelKimi K3
66dKimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI.
2.2M downloads / mo11.4K likesMiniMax M3
112dMiniMax-M3 is a multimodal foundation model from MiniMax.
191.9K downloads / mo1.5K likesStep 3.7 Flash
115dStep 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model.
22.2K downloads / mo455 likesGemini 3.5 Flash
124dGemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed.
API modelPerceptron Mk1
131dPerceptron Mk1 (Mark One) is Perceptron's highest-quality vision-language model for video and embodied reasoning.** It accepts image and video inputs paired…
API modelGemini 3.1 Flash Lite
136dGemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads.
API modelNemotron 3 Nano Omni (free)
145dNVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems.
272.4K downloads / mo427 likesKimi Latest
146dThis model always redirects to the latest model in the Kimi family.
Open weightsQwen3.6 27B
146dQwen3.6 27B is a dense 27-billion-parameter language model from the Qwen Team at Alibaba, released in April 2026.
3.7M downloads / mo2.3K likesQwen3.5 Plus 2026-04-20
146dQwen3.5 Plus (April 2026) is a large-scale multimodal language model from Alibaba.
Open weightsQwen3.6 Flash
146dQwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series.
Open weightsGemini Flash Latest
146dThis model always redirects to the latest model in the Gemini Flash family.
API modelQwen3.6 35B A3B
146dQwen3.6-35B-A3B is an open-weight multimodal model from Alibaba Cloud with 35 billion total parameters and 3 billion active parameters per token.
3.4M downloads / mo2.8K likesGemini Pro Latest
146dThis model always redirects to the latest model in the Gemini Pro family.
API modelMiMo-V2.5
151dMiMo-V2.5 is a native omnimodal model by Xiaomi.
296.7K downloads / mo424 likesGemma 4 26B A4B
170dGemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind.
9.8M downloads / mo1.5K likesQwen3.6 Plus
171dQwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and…
Open weightsGemma 4 31B
171dGemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output.
9M downloads / mo3.9K likesGLM 5V Turbo
172dGLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks.
Open weightsReka Edge
184dReka Edge is an extremely efficient 7B multimodal vision-language model that accepts image/video+text inputs and generates text outputs.
533 downloads / mo133 likesSeed-2.0-Lite
194dSeed-2.0-Lite is a versatile, cost‑efficient enterprise workhorse that delivers strong multimodal and agent capabilities while offering noticeably lower…
API model
