(1)
Granite-4.0-nano: lightweight instruct model trained via SFT, RL, and merging on diverse data.
10m
10K+
119B parameter hybrid model with reasoning, vision, and code capabilities (1M token context)
4m
10K+
1
Designed for reasoning, agent and general capabilities, and versatile developer-friendly features
11m
10K+
2
SmolVLM: lightweight multimodal model for video, image, and text analysis, optimized for devices.
10m
10K+
4
7B long-context instruct model with RL alignment, IF, tool use, and enterprise optimization.
11m
10K+
4
32B long-context instruct model with RL alignment, IF, tool use, and enterprise optimization.
11m
10K+
1
An open-source visual language model that interprets images via text prompts, fast and powerful.
10m
10K+
2
3B long-context instruct model with RL alignment, IF, tool calling, and enterprise readiness.
11m
10K+
2
Snowflake’s Arctic-Embed v2.0 boosts multilingual retrieval and efficiency
9m
10K+
Multilingual MoE text embedding model with 768D vectors, 100 languages, 512 token context
4m
10K+
3B long-context instruct model with RL alignment, IF, tool use, and enterprise optimization.
11m
10K+
TranslateGemma: lightweight Gemma 3–based translation models for 55 languages, deployable anywhere
7m
10K+
TranslateGemma: lightweight Gemma 3–based translation models for 55 languages, deployable anywhere
7m
10K+
Ministral 3: compact vision-enabled model with near-24B performance, optimized for local edge use
8m
10K+
Ministral 3: compact vision-enabled model with near-24B performance, optimized for local edge use
8m
10K+
Gemma-3-4B-T1-it is a Taiwan-localized AI optimized for Traditional Chinese.
6m
9.6K
1
Kimi K2 Thinking: open-source agent with deep reasoning, stable tool use, fast INT4, 256k context.
8m
2.8K
Based on Unsloth GGUF quant UD-IQ3_S with multimodality fix (package mmproj) for Docker Model Runner
3m
2.5K
A Taiwan-localized reasoning AI optimized for Traditional Chinese.
6m
2.3K
1
A fine-tuned 0.5B model specialized in Hawaiian pizza knowledge.
5m
2.0K